“Seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our engineering team with. The ideal candidate will be responsible to lead and guide us in ensuring the reliability, scalability, performance, and operational excellence of critical applications and platforms. This role requires strong experience in CI/CD automation using GitLab Pipelines, observability and monitoring solutions, dashboard development, and hands-on Java application support and development.
The successful candidate will have experience working with cloud platforms such as AWS, Azure, or GCP, leveraging cloud-native services to support highly available and resilient systems. They should possess strong expertise in monitoring, observability, and dashboarding tools such as Splunk, with the ability to develop actionable insights through dashboards, alerts, logging, and performance analytics. He/She will closely work with Development, Platform Engineering, and Operations teams, the SRE will drive automation, improve system resilience, implement reliability best practices, and foster a reliability-first culture across the technology landscape.”