Track Lead - AWS IAC, Terraform, Python
Lucknow, Uttar Pradesh
Job Summary
We are seeking a highly skilled Cloud SME with strong expertise in AWS Cloud Engineering and Site Reliability Engineering (SRE). The ideal candidate will design, implement, automate, and manage highly available, scalable, secure, and cost-efficient cloud platforms. The role requires extensive experience across AWS services, cloud-native architectures, Infrastructure as Code (IaC), Kubernetes, observability, DevOps, and cloud operations, with secondary expertise in Google Cloud Platform (GCP) and multi-cloud environments.
Job Description : Key Performance Indicators (KPIs)\\\\r\\\\n\\\\r\\\\nCloud Platform Availability ( >99.9%)\\\\r\\\\nSLO/SLA Compliance\\\\r\\\\nMean Time to Recover (MTTR)\\\\r\\\\nChange Success Rate\\\\r\\\\nCloud Cost Optimization Savings\\\\r\\\\nSecurity Compliance Score\\\\r\\\\nAutomation Coverage\\\\r\\\\nIncident Reduction Trend
Key Responsibilities
Job Responsibilities : Cloud Architecture & Engineering Design and implement enterprise-grade cloud solutions on AWS. Develop scalable, resilient, and secure cloud architectures. Lead cloud transformation and modernization initiatives. Define cloud governance, standards, and best practices. Support hybrid and multi-cloud environments involving AWS and GCP. [SRE (AWS,GCP) | Word], [Job Descri...Consultant | Word], [AWS CLOUD Engineer | Word] Site Reliability Engineering (SRE) Establish and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs). Improve platform reliability, availability, scalability, and performance. Lead incident response, root cause analysis (RCA), and postmortem reviews. Implement proactive monitoring and automated remediation. Drive operational excellence through reliability engineering practices. [SRE (AWS,GCP) | Word], [HCL Detail...rity Group | PowerPoint] AWS Cloud Operations Manage AWS infrastructure including: EC2 S3 RDS EKS Lambda VPC Route53 CloudFront DynamoDB SNS/SQS Ensure optimal performance, security, and cost management. Implement high availability and disaster recovery solutions. [SRE (AWS,GCP) | Word], [AWS CLOUD Engineer | Word], [Job Descri...Consultant | Word] GCP & Multi-Cloud Support Support GCP environments including: Compute Engine GKE Cloud Storage Cloud SQL Cloud Functions Design multi-cloud architectures and governance models. Assist cloud migration and modernization projects. [SRE (AWS,GCP) | Word], [Job Descri...DevOps L3 | Word] Infrastructure as Code (IaC) Develop cloud infrastructure using: Terraform AWS CloudFormation GCP Deployment Manager Automate provisioning, configuration, and compliance processes. Implement GitOps and Infrastructure Automation practices. [SRE (AWS,GCP) | Word], [AWS CLOUD Engineer | Word], [JD- Devops Engineer | Word] DevOps & CI/CD Build and maintain CI/CD pipelines. Support tools such as: Jenkins GitHub Actions GitLab CI/CD AWS CodePipeline Automate application deployments and release management. Implement DevSecOps controls and compliance checks. [SRE (AWS,GCP) | Word], [AWS CLOUD Engineer | Word], [JD- Devops Engineer | Word] Observability & Monitoring Implement enterprise observability platforms. Configure and maintain: CloudWatch CloudTrail GCP Monitoring Prometheus Grafana OpenTelemetry ELK Stack Build dashboards, alerts, and operational reporting. [SRE (AWS,GCP) | Word], [AWS CLOUD Engineer | Word] Security & Compliance Implement cloud security frameworks and policies. Manage IAM, encryption, secrets management, and policy enforcement. Ensure compliance with organizational and regulatory requirements. Support vulnerability assessments and remediation activities. [SRE (AWS,GCP) | Word], [Job Descri...Consultant | Word] Disaster Recovery & Resiliency Design and test DR and Business Continuity solutions. Conduct failover, failback, and recovery testing. Ensure backup and restoration strategies
Skill Requirements
Technical Skills AWS (Primary) EC2 EKS Lambda S3 RDS DynamoDB VPC Route53 CloudFront SNS/SQS IAM KMS Secrets Manager CloudWatch CloudTrail GCP (Secondary) Compute Engine GKE Cloud SQL Cloud Storage Cloud Functions IAM Cloud Monitoring Cloud Logging DevOps & Automation Terraform CloudFormation Git Jenkins GitLab CI/CD GitHub Actions ArgoCD Containers & Platforms Kubernetes (EKS/GKE) Docker Helm Istio Programming Python Bash/Shell PowerShell YAML JSON Networking VPC VPN DNS Load Balancers Transit Gateway Private Endpoints Hybrid Connectivity Cloud Networking Architecture Monitoring & SRE Prometheus Grafana ELK OpenTelemetry Datadog (Preferred)
Other Requirements
Experience Senior Cloud Engineer 7–10 years of overall IT experience. Minimum 5 years in AWS cloud engineering. Cloud SME / Lead SRE 10–15+ years of experience. 5+ years designing enterprise AWS environments. 3+ years working with Kubernetes and SRE practices. Experience managing production-critical cloud platforms. [SRE (AWS,GCP) | Word], [Job Descri...DevOps L3 | Word], [Subash Raj...Profile 3 | PDF] Preferred Certifications AWS AWS Solutions Architect – Professional AWS DevOps Engineer – Professional AWS Solutions Architect – Associate AWS Advanced Networking Specialty Google Cloud Google Professional Cloud Architect Google Professional DevOps Engineer Other Certified Kubernetes Administrator (CKA) Terraform Associate ITIL v4 Foundation
#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-