Platform Engineer – TechOps
Location: Bangalore
Experience: 1–2 Years
Employment Type: Full-time
Work Mode: On-site
Shift: 2:00 PM – 11:00 PM IST
Interview: Face-to-Face
Job Summary
We are looking for a Platform Engineer – TechOps to join our Platform SRE team and help maintain the reliability, availability, performance, and scalability of large-scale connected-device platforms.
The ideal candidate will have hands-on experience in Linux, AWS/Cloud, monitoring, incident management, networking, and production support, along with an automation-first mindset.
Key Responsibilities
- Monitor and maintain the reliability, availability, and performance of production platforms.
- Provide 24/7 operational support and participate in on-call rotations, including off-hours and weekends when required.
- Handle P1/P2 incidents, troubleshoot production issues, perform RCA, and contribute to post-incident reviews.
- Monitor system health, application performance, infrastructure, and connected devices using monitoring and observability tools.
- Perform Linux system troubleshooting, log analysis, service validation, and issue resolution.
- Work closely with Development, Operations, Support, and Engineering teams for incident resolution and platform improvements.
- Support deployment, commissioning, monitoring, maintenance, and decommissioning of connected devices.
- Assist with firmware changes and releases where required.
- Identify opportunities for automation and reduce repetitive manual operational tasks.
- Support capacity planning, performance tuning, resource optimization, and scalability initiatives.
- Follow security, compliance, documentation, and operational best practices.
- Troubleshoot network-related issues and connectivity problems across distributed environments.
Required Skills
- 1+ year of experience in SRE, DevOps, Platform Engineering, Production Support, or Technical Support.
- Strong knowledge of Linux/Unix administration.
- Hands-on experience with AWS or other cloud platforms.
- Experience with monitoring/observability tools such as:
- Grafana
- Prometheus
- Dynatrace
- Zabbix
- CloudWatch
- Good understanding of incident management, troubleshooting, RCA, and on-call operations.
- Knowledge of networking fundamentals including TCP/IP, DNS, HTTP/HTTPS, routing, firewalls, VLANs, ACLs, and subnetting.
- Scripting/automation experience using Bash, Shell, Python, or similar technologies.
- Good communication and stakeholder-management skills.
- Ability to work effectively in a fast-paced production environment.
Good to Have
- Experience working with IoT or connected-device platforms.
- Experience with EV charging infrastructure / EV charging platforms.
- Experience with device onboarding/offboarding and firmware management.
- Knowledge of AWS services and cloud infrastructure.
- Experience with Infrastructure as Code or DevOps tools.
- Experience working in a global production environment.
- Bachelor's degree in Computer Science, Information Technology, Electronics, Electrical Engineering, or a related field.
What We Are Looking For
- Strong troubleshooting and problem-solving skills.
- Ownership mindset with a focus on reliability and availability.
- Automation-first approach to eliminating operational toil.
- Willingness to work in a 2 PM – 11 PM shift and participate in on-call support.
- Strong written and verbal English communication.
- Candidates who can join immediately or within 30 days are preferred.
Pay: ₹500,000.00 - ₹1,000,000.00 per year
Application Question(s):
- Do you have experience handling P1/P2 production incidents, RCA and on-call support? Please mention briefly.
- Are you comfortable working from Bangalore in an on-site setup?
Are you comfortable working in the 2 PM–11 PM IST shift and participating in on-call/off-hours support when required?
- What is your current notice period and earliest possible joining date?
- What is your current CTC and expected CTC?
- How many years of experience do you have in SRE / DevOps / Platform Engineering / Production Support?
Do you have hands-on experience with AWS? If yes, which AWS services have you worked on?
Which monitoring/observability tools have you worked with?
Work Location: In person