About this Opportunity
We are looking for an Operation Assurance Engineer (L0) to support the operations and availability of critical IT infrastructure, applications, databases, and monitoring platforms. This role focuses on proactive monitoring, incident management, troubleshooting, preventive maintenance, and operational support to ensure service stability, SLA compliance, and customer satisfaction in a 24×7 environment.
________________________________________
What You Will Do
- Monitor infrastructure and applications using Zabbix and other monitoring tools to identify and respond to alarms and incidents.
- Perform first-level troubleshooting across applications, databases, Linux systems, servers, networks, and infrastructure components, restoring services and escalating where required.
- Manage incidents, changes, and service requests through ITSM/ITIL processes, ensuring SLA adherence and accurate documentation.
- Conduct health checks, capacity monitoring, backup verification, preventive maintenance, patching, upgrades, and lifecycle management activities.
- Support application operations, database monitoring, service restarts, and basic administration tasks.
- Collaborate with internal teams, vendors, and customer stakeholders to resolve issues and implement approved changes.
- Maintain CMDB accuracy, operational reports, and compliance with security and operational standards.
- Participate in a 24×7 support model, including shift and on-call responsibilities.
________________________________________
You Will Bring
- Bachelor's degree in Computer Science, IT, Computer Engineering, Telecommunications, or related field.
- 2–4 years of experience in IT Operations, Infrastructure Support, Application Support, Managed Services, or a similar environment.
- Strong knowledge of Linux, Java, Python, PostgreSQL, Oracle, application support, and infrastructure troubleshooting.
- Experience with Zabbix or similar monitoring tools, ITSM platforms, and ITIL-based Incident, Problem, Change, and Service Management processes.
- Ability to analyze logs, troubleshoot across application, database, OS, and infrastructure layers, and work effectively under pressure.
- Excellent communication, stakeholder management, ownership, and customer-focus skills.
- Preferred certifications: ITIL Foundation, Zabbix, Linux, Oracle/PostgreSQL, CCNA, Java, or Python.
Key Success Metrics:
Service Availability, Incident Response, Application & Database Support, Infrastructure Stability, SLA Compliance, Backup & Change Success, CMDB Accuracy, and Customer Satisfaction.
“All academic credentials must be from recognized and accredited institutions and are further subject to verification.”