You own production. Not a ticket queue — the real thing.
At Techdome, you'll keep Healthcare, FinTech, AI, and SaaS products running for real users who depend on uptime. You'll build the pipelines, handle the incidents, and ship releases without downtime. If Kubernetes, automation, and observability genuinely excite you, read on.
-
Keep production stable, fast, and scalable across every environment
-
Manage and optimize cloud infrastructure (AWS, Azure, or GCP)
-
Build zero-downtime CI/CD pipelines using Blue-Green, Rolling, and Canary deployments
-
Automate infrastructure with Terraform and Ansible
-
Set up monitoring and observability (Prometheus, Grafana, ELK, Datadog, OpenTelemetry)
-
Define and track SLIs, SLOs, and error budgets
-
Lead incident response, root-cause analysis, and post-incident reviews
-
Optimize cloud costs and plan for scale
-
Use AI tools to automate alert triage, incident summaries, and repetitive ops work
-
Join the on-call rotation
-
3+ years as an SRE, DevOps, Platform, or Cloud Engineer
-
Hands-on experience with AWS, Azure, or GCP
-
Strong Docker and Kubernetes experience
-
Infrastructure-as-Code skills (Terraform, Ansible, or similar)
-
Experience building CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI, etc.)
-
Solid Linux and networking fundamentals
-
Scripting ability in Python, Go, or Bash
-
Real experience running production systems at scale
-
Practical experience with Blue-Green, Canary, or Rolling deployments
-
Background in FinTech, Payments, Healthcare, or another high-availability industry
-
Comfortable using AI tools (Copilot, Claude, Cursor, ChatGPT) day-to-day
Nice to have: experience building AI-powered ops workflows (alert triage, incident summarization, automation) and strong SRE fundamentals — SLOs, error budgets, chaos engineering.
-
Work across AI, Healthcare, Payments, and SaaS — not one product for years
-
Own critical infrastructure from day one, no waiting period
-
Support systems used by thousands of real users
-
Work directly with founders and senior engineering leadership
-
Genuine AI-first tooling and culture, not just a mandate
-
Fast decisions and real growth, not a slow climb to a title