Cloud / Site Reliability Engineer (SRE)
Location: Alpha-1, Greater Noida, Uttar Pradesh
Job Type: Full-time | Permanent
Openings: 1
Salary: ₹9 LPA – ₹15 LPA
Job Summary
We are looking for a Cloud/SRE professional to maintain platform reliability, performance, monitoring, scalability and incident response.
Key Responsibilities
- Define and monitor SLOs, error budgets and uptime targets.
- Build monitoring dashboards and alerts.
- Handle production incidents and root-cause analysis.
- Plan autoscaling and capacity requirements.
- Perform load testing and reliability improvements.
- Optimise cloud infrastructure costs.
- Improve system resilience and fault tolerance.
- Maintain incident runbooks.
- Work with backend teams on performance issues.
- Conduct disaster recovery drills.
Required Skills
- Experience in SRE, cloud operations or reliability engineering.
- Strong AWS/GCP/Azure knowledge.
- Monitoring and observability experience.
- Real-world incident management/on-call experience.
- Python, Go or Bash scripting.
- Understanding of distributed systems and caching.
- Linux troubleshooting and performance tuning.
Preferred
- Kubernetes.
- Terraform/Pulumi.
- k6, Locust or JMeter.
- Cloud architecture certification.
Pay: ₹900,000.00 - ₹1,500,000.00 per year
Application Question(s):
- Which cloud platform are you strongest in?
- Do you have production on-call experience?
- Which monitoring tools have you used?
- Have you worked with Kubernetes?
- Have you handled production incidents and root-cause analysis?
Experience:
- Site Reliability Engineer: 4 years (Required)
Location:
- Greater Noida, Uttar Pradesh (Required)
Work Location: In person