About the RoleWe are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for... Read more
About the Role
We are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessAll your saved jobs are no longer available or you've already applied.
for the following search criteria