Role Overview
As a key member of the SymphonyAI Financial Services team in Bangalore, you will be responsible for ensuring the reliability, performance, and availability of our SaaS solutions. You will work closely with engineering and customer success teams to provide world-class support, manage incidents, and drive continuous service improvements, playing a pivotal role in maintaining high service standards and enhancing client satisfaction.
Responsibilities
- Monitor and maintain the health of SaaS environments, ensuring high availability and optimal performance.
- Investigate and resolve incidents, problems, and service requests within agreed SLAs to maintain client service standards.
- Participate in root cause analysis and post-incident reviews for major outages, contributing to service improvement strategies.
- Implement and manage changes to cloud environments following ITIL best practices, ensuring compliance with industry standards.
- Utilize observability and monitoring tools like DataDog and PagerDuty to identify and respond to service degradations proactively.
Requirements
- Hands-on experience with Azure (preferred) and/or AWS cloud platforms.
- Hands-on experience in PostgreSQL and Oracle administration and optimization.
- Experience with Kubernetes and Azure Kubernetes Service (AKS).
- Proficiency in observability and incident response tools like DataDog and PagerDuty.
- Familiarity with ITIL processes, particularly Incident, Problem, Change, and Service Request Management.
- Experience working in Managed Services or SaaS Support environments.
Skills
- Azure
- AWS
- Kubernetes
- PostgreSQL
- DataDog