Job Description
- Ensure the stability, availability, and performance of production applications.
- Monitor, troubleshoot, and resolve incidents, problems, and performance issues.
- Participate in major incident management (P1/P2), root cause analysis (RCA), and continuous service improvement initiatives.
- Manage deployments, releases, and change requests following ITIL and DevOps best practices.
- Collaborate with Development, Infrastructure, Scrum teams, and external providers to deliver sustainable solutions.
- Implement and enhance monitoring and observability capabilities across production environments.
- Maintain technical documentation and promote knowledge sharing within global support teams.
- Participate in on-call rotations supporting business-critical applications.