Lead enterprise AWS cloud operations, driving reliability, security, scalability, automation, cost optimization, and operational excellence.
09th September, 2026
Job Summary
We are seeking an experienced Cloud Operations Manager to lead and optimize cloud infrastructure operations within a large enterprise environment. The successful candidate will provide both technical leadership and strategic direction, overseeing teams across Cloud Engineering, Site Reliability Engineering (SRE), Networking, and Database Administration.
This role will be responsible for ensuring the reliability, scalability, security, performance, and cost efficiency of cloud infrastructure while driving automation and Infrastructure-as-Code (IaC) practices.
The Cloud Operations Manager will work closely with technology, project, program, and business stakeholders to define and implement effective cloud operating models and ensure that infrastructure and applications are designed for production readiness, maintainability, and ongoing support.
Key Responsibilities
Develop and execute cloud operations strategies aligned with business and technology objectives.
Lead and mentor teams of Cloud Engineers, SREs, Network Engineers, and Database Administrators.
Oversee the design, deployment, and ongoing operation of scalable and highly available cloud infrastructure.
Drive adoption of Infrastructure-as-Code (IaC) using tools such as Terraform and AWS CloudFormation.
Establish and continuously improve cloud operational processes, standards, and best practices.
Ensure high levels of system availability, reliability, performance, and resilience.
Implement effective monitoring, alerting, incident management, and root-cause analysis processes.
Lead proactive risk identification and remediation across cloud infrastructure and services.
Ensure appropriate security controls, data protection measures, and compliance requirements are implemented.
Drive cloud cost optimization initiatives while maintaining performance and reliability.
Partner with engineering and application teams to ensure new infrastructure is production-ready and supportable.
Contribute to cloud migration, modernization, scalability, and transformation initiatives.
Evaluate emerging cloud technologies and recommend solutions that improve operational efficiency and business value.
Establish and monitor operational KPIs, SLAs, SLOs, and service health metrics.
Collaborate with internal stakeholders and third-party technology partners to resolve service issues and improve cloud capabilities.
Support the development of disaster recovery, business continuity, and resilience strategies.
Ensure operational readiness for new software components, platforms, and infrastructure services.
Technical Requirements
Strong hands-on expertise with Amazon Web Services (AWS) and enterprise cloud environments.
Strong understanding of AWS architecture, infrastructure, networking, security, and core cloud services.
Experience with Infrastructure-as-Code, particularly Terraform and/or CloudFormation.
Strong knowledge of CI/CD pipelines and DevOps practices.
Experience with cloud monitoring, logging, observability, and incident management solutions.
Proficiency with scripting and automation using Python, Bash, or similar languages.
Strong understanding of high availability, scalability, disaster recovery, and cloud resilience.
Experience implementing cloud security and governance best practices.
Knowledge of cloud cost management and optimization.
Understanding of SRE principles, operational excellence, and service reliability practices.
Qualifications & Experience
Bachelor's or Master's degree in Computer Science, Information Technology, Engineering, or a related discipline, or equivalent professional experience.
14+ years of overall IT experience, preferably within large enterprise environments.
5+ years of experience managing AWS cloud environments and cloud operations teams.
Proven experience leading multidisciplinary technical teams.
Demonstrated experience in Cloud Operations, CloudOps, Program Management, or Delivery Management.
Strong experience working with enterprise-scale cloud infrastructure and operational models.
Experience working across technology, engineering, business, and third-party stakeholder groups.
Proven track record of delivering cloud infrastructure and operational transformation initiatives.
Preferred Certifications
AWS Certified Solutions Architect
AWS Certified SysOps Administrator
AWS Certified DevOps Engineer
Other relevant AWS or cloud certifications are advantageous.
Key Competencies
Strong technical leadership and people management skills.
Strategic and analytical mindset.
Excellent problem-solving and decision-making abilities.
Strong communication and stakeholder management skills.
Ability to manage multiple priorities within complex enterprise environments.
Strong focus on operational excellence, reliability, security, and continuous improvement.
Ability to translate business requirements into practical cloud operating strategies.