JOB DUTIES & RESPONSIBILITY
1. Infrastructure & Virtualization
-
Manage and optimize ProMax-based VM clusters, including provisioning, backup strategies, and resource allocation.
-
Develop reusable VM templates with preconfigured environments to standardize deployments.
-
Automate VM lifecycle tasks (creation, patching, cleanup) using Ansible or similar tools.
-
Implement user and access controls to ensure secure, auditable environments.
2. Infrastructure as Code & Automation
-
Implement IaC practices (Terraform, Pulumi, etc.) to standardize infrastructure provisioning.
-
Use Ansible/automation frameworks to script recurring workflows like backups, log rotation, and database initialization.
-
Maintain a library of reusable scripts and playbooks to improve operational efficiency
3. Containerization & Application Platform
-
Design, deploy, and maintain services using Docker and Docker Compose.
-
Manage containerized stacks (including Supabase-like platforms) for consistency and scalability.
-
Optimize resource utilization, lifecycle management, and monitoring across multi-node container deployments.
4. Database Administration (PostgreSQL)
-
Install, configure, and maintain PostgreSQL databases across multiple environments.
-
Plan and execute migrations, including splitting workloads into dedicated clusters.
-
Implement HA, replication, and disaster recovery strategies.
-
Apply best practices for database performance, tuning, and security.
-
Establish a structured migration/versioning system for schema management
5. CI/CD & Deployment
-
Design, implement, and maintain CI/CD pipelines to streamline builds, testing, and deployments.
-
Manage deployments across multiple environments, including cloud-native platforms and modern hosting services like Vercel.
-
Integrate infrastructure automation into pipelines to ensure consistency and reproducibility.
-
Collaborate with development teams to optimize deployment workflows and rollback strategies.
6. Monitoring, Observability & Reliability
-
Deploy and manage monitoring solutions (Prometheus, Grafana, ELK/EFK) for systems and applications.
-
Build dashboards, alerts, and usage reports to track performance, reliability, and API consumption.
-
Drive incident response processes, including root-cause analysis and post-mortems.
7. Scalability & Security
-
Implement scaling strategies for APIs and backend services (load balancing, caching, auto scaling).
-
Enforce strong security practices for infrastructure, databases, and deployments.
-
Continuously evaluate tools and frameworks to improve system efficiency and resilience.
QUALIFICATION & SKILL-SET
-
3 – 7 years of experience in DevOps, SRE, or Infrastructure roles.
-
Strong hands-on experience with PostgreSQL administration in production environments.
-
Proficiency with Docker/Docker Compose and containerized application stacks.
-
Experience with Supabase or similar database-driven PaaS.
-
Skilled in CI/CD pipeline design (GitHub Actions, GitLab CI, Jenkins, etc.).
-
Experience deploying to modern platforms such as Vercel, Netlify, or cloud-native PaaS.
-
Proficiency in IaC tools (Terraform, Pulumi) and configuration management (Ansible, Puppet, Chef).
-
Experience with virtualization platforms (Proxmox, VMware, KVM, or similar).
-
Solid scripting/programming skills (Bash, Python, or Go). Familiarity with observability stacks (Grafana, Prometheus, ELK).
-
Strong understanding of Linux system administration, networking, and security.
Nice-to-Have Skills
-
Cloud provider experience (AWS, GCP, Azure) and hybrid infra management.
-
Database optimization and query tuning experience.
-
Knowledge of Kubernetes or Nomad for container orchestration.
-
Experience with cost monitoring and optimization for infra and deployments.
Why Join Us
-
Shape the infrastructure for a next-generation data and geospatial analytics platform.
-
Direct ownership of automation, monitoring, and deployment pipelines across multiple products.
-
Collaborative and growth-oriented environment with opportunities to move into senior DevOps/SRE Architect roles