Source description
About the role
Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world’s leading brands, Fortune 500 companies and government institutions. For more information, please visit www.saviynt.com .
We are seeking an experienced Cloud Manager to lead the design, implementation, operation, and optimization of our cloud infrastructure. The ideal candidate will manage cloud platforms, drive cloud transformation initiatives, ensure security and compliance, and lead a team of cloud engineers to deliver scalable, secure, and cost-effective cloud solutions.
WHAT YOU WILL BE DOING
Oversee the monitoring of our SaaS application and underlying infrastructure (Kubernetes on AWS and Azure, VPN connections, customer applications, Elastic Search, MySQL) for alerts and performance issues.
Manage the full lifecycle of alerts, incidents, and service requests reported through FreshService, ensuring timely and accurate logging, prioritization, resolution, and escalation.
Develop, implement, and maintain operational procedures, runbooks, and knowledge base articles to standardize incident resolution and service request fulfilment
Drive continuous improvement initiatives to optimize operational efficiency, reduce incident rates, and improve service request turnaround times.
Collaborate with engineering, development, and SRE teams to troubleshoot complex issues, identify root causes, and implement preventative measures.
Ensure adherence to defined SLAs (Service Level Agreements) and KPIs (Key Performance Indicators) for operational performance.
Generate regular reports on operational metrics, incident trends, and service request performance for management review.
Participate in on-call escalations as needed and provide leadership during critical incidents. Foster a collaborative and high-performing team environment, providing coaching, mentoring, and performance feedback to team members
Manage and maintain operational documentation, including system diagrams, contact lists, and escalation paths.
Ensure compliance with relevant security and compliance policies within the operations centre.
Plan and coordinate scheduled maintenance activities with minimal impact to service availability.
WHAT YOU WILL BRING
Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
Relevant cloud certifications are highly desirable, such as: AWS Certified Solutions Architect – Associate/Professional Microsoft Certified: Azure Administrator or Azure Solutions Architect Expert
10–12 years of overall experience in IT infrastructure, systems administration, cloud engineering, or DevOps.
5+ years of hands-on experience managing enterprise cloud environments (AWS, Azure, or Google Cloud Platform).
Proven experience leading cloud migration, modernization, and infrastructure transformation projects.
Experience managing or mentoring cloud engineering or infrastructure teams.
Strong background in Infrastructure as Code (Terraform, CloudFormation, or ARM/Bicep) and CI/CD implementation.
Hands-on experience with containerization and orchestration technologies such as Docker and Kubernetes.
Experience in cloud security, governance, disaster recovery, and cost optimization (FinOps).
Familiarity with Agile/Scrum methodologies and cross-functional collaboration with development, security, and operations teams.
More at Saviynt
Related open roles
Principal Site Reliability Engineer, Google Cloud
United States · Hybrid
Principal Site Reliability Engineer
Vancouver · Hybrid
Staff platform Support Engineer
Bangalore · Hybrid
Technical Lead - Professional_Services
India · Hybrid
Platform Support Engineer
Atlanta · Hybrid
Senior / Staff Site Reliability, Platform Engineering
San Francisco Bay Area · Atlanta · Hybrid
