Job Summary
We are looking for a skilled DevOps Engineer with 3–4 years of experience to manage and improve our cloud infrastructure, CI/CD pipelines, deployment processes, monitoring, and application reliability.
The ideal candidate should have strong hands-on experience with AWS, Docker, Kubernetes, Linux, CI/CD, Git, and infrastructure automation. The candidate will work closely with development, QA, and product teams to ensure reliable, secure, and scalable application deployments.
Key Responsibilities
1. Cloud & Infrastructure
- Manage and maintain AWS infrastructure across development, QA, staging, and production environments.
- Work with services such as:
-
EC2
-
S3
-
RDS
-
ECR
-
CloudFront
-
IAM
-
VPC
-
Route 53
-
CloudWatch
-
SES/SNS
- Configure and maintain secure networking, IAM roles, security groups, and access policies.
- Monitor infrastructure capacity, availability, and performance.
- Optimize AWS infrastructure and cloud costs.
2. CI/CD
- Design, implement, and maintain CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or similar tools.
- Automate:
Build
Unit/integration test execution
Docker image creation
Security scanning
Deployment
Rollback
- Implement deployment strategies such as rolling, blue-green, or canary deployments.
- Maintain separate deployment workflows for Dev, QA, Staging, and Production.
3. Docker & Kubernetes
- Build and maintain optimized Dockerfiles and Docker images.
- Manage containerized applications using Kubernetes.
- Work with:
-
Deployments
-
Services
-
ConfigMaps
-
Secrets
-
Ingress
-
Persistent Volumes
-
Liveness/Readiness probes
-
Resource limits and requests
- Troubleshoot pod failures, CrashLoopBackOff, image-pull issues, networking problems, and resource-related issues.
- Experience with Helm charts and Kubernetes configuration management is preferred.
4. Infrastructure as Code
Expected Experience
The candidate should be capable of independently handling:
Developer commits CI pipeline Build/Test Docker image Container registry Kubernetes deployment Monitoring Production support
and should be comfortable troubleshooting the complete deployment lifecycle.