About the RoleWe are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for... Read more
About the Role
We are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessAbout the RoleWe are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for... Read more
About the Role
We are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessAbout the RoleWe are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for... Read more
About the Role
We are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Preferred Qualifications
Experience implementing Site Reliability Engineering (SRE) practices.Experience with monitoring and observability platforms such as Prometheus, Grafana, APM tools, and centralized logging solutions.Experience with disaster recovery planning and multi-region infrastructure.Knowledge of compliance frameworks including SOC 2 and other security standards.Experience supporting enterprise-scale cloud platforms.Cloud certifications such as AWS Solutions Architect or Google Cloud Professional certifications.Experience with enterprise networking, telecommunications, wireless networking, or IoT environments.Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessAbout the RoleWe are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for... Read more
About the Role
We are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Preferred Qualifications
Experience implementing Site Reliability Engineering (SRE) practices.Experience with monitoring and observability platforms such as Prometheus, Grafana, APM tools, and centralized logging solutions.Experience with disaster recovery planning and multi-region infrastructure.Knowledge of compliance frameworks including SOC 2 and other security standards.Experience supporting enterprise-scale cloud platforms.Cloud certifications such as AWS Solutions Architect or Google Cloud Professional certifications.Experience with enterprise networking, telecommunications, wireless networking, or IoT environments.Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessAbout the RoleWe are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for... Read more
About the Role
We are seeking an experienced DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Preferred Qualifications
Experience implementing Site Reliability Engineering (SRE) practices.Experience with monitoring and observability platforms such as Prometheus, Grafana, APM tools, and centralized logging solutions.Experience with disaster recovery planning and multi-region infrastructure.Knowledge of compliance frameworks including SOC 2 and other security standards.Experience supporting enterprise-scale cloud platforms.Cloud certifications such as AWS Solutions Architect or Google Cloud Professional certifications.Experience with enterprise networking, telecommunications, wireless networking, or IoT environments.Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessAbout the RoleWe are seeking an experienced Senior DevOps Manager to lead a high-performing Technical Operations team responsible... Read more
About the Role
We are seeking an experienced Senior DevOps Manager to lead a high-performing Technical Operations team responsible for ensuring the reliability, scalability, and performance of a large-scale cloud infrastructure. This leadership role combines technical expertise with people management, driving operational excellence, infrastructure automation, and cross-functional collaboration across engineering, security, and product teams.
The ideal candidate is an experienced DevOps or Site Reliability Engineering (SRE) leader with a strong background in cloud infrastructure, incident management, infrastructure automation, and team development.
Key Responsibilities
Lead, mentor, and develop a high-performing Technical Operations/DevOps team.Ensure 24/7 operational stability, availability, and reliability of production infrastructure.Drive operational excellence through continuous improvement initiatives.Lead incident response, major incident management, and post-incident reviews (RCA).Establish and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.Oversee infrastructure automation using Infrastructure as Code (IaC) principles.Manage Kubernetes-based container platforms and cloud infrastructure.Improve deployment processes through CI/CD and automation.Collaborate with Product, Engineering, Security, and Infrastructure teams on strategic initiatives.Drive disaster recovery, resiliency, and business continuity planning.Develop monitoring, alerting, and observability strategies.Present operational metrics and project updates to senior leadership.Ensure compliance with security standards, audits, and regulatory requirements.Required Qualifications
8+ years of experience in DevOps, Technical Operations, Infrastructure Engineering, or Site Reliability Engineering.4+ years of experience managing technical operations, infrastructure, or SRE teams.Strong leadership, mentoring, and people management skills.Experience managing production cloud infrastructure in AWS and/or GCP.Hands-on experience with Kubernetes, Docker, and container orchestration.Strong experience with Terraform and Infrastructure as Code (IaC).Experience implementing CI/CD pipelines and deployment automation.Strong understanding of Linux systems administration.Experience with incident management, on-call operations, and production support.Knowledge of networking fundamentals, cloud security, and infrastructure best practices.Experience with Agile methodologies, Kanban, Jira, or similar project management tools.Excellent communication, stakeholder management, and executive presentation skills.Preferred Qualifications
Experience implementing Site Reliability Engineering (SRE) practices.Experience with monitoring and observability platforms such as Prometheus, Grafana, APM tools, and centralized logging solutions.Experience with disaster recovery planning and multi-region infrastructure.Knowledge of compliance frameworks including SOC 2 and other security standards.Experience supporting enterprise-scale cloud platforms.Cloud certifications such as AWS Solutions Architect or Google Cloud Professional certifications.Experience with enterprise networking, telecommunications, wireless networking, or IoT environments.Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field.Technical Environment
Google Cloud Platform (GCP)Amazon Web Services (AWS)KubernetesDockerTerraformLinuxCI/CD PipelinesGitPrometheusGrafanaApplication Performance Monitoring (APM)Infrastructure as Code (IaC)JiraAgile / KanbanNetworking & Cloud SecurityWhat You'll Gain
Opportunity to lead and mentor a high-performing DevOps and Technical Operations team.Work on large-scale, cloud-native infrastructure supporting enterprise applications and mission-critical services.Drive strategic initiatives in infrastructure automation, cloud modernization, and Site Reliability Engineering (SRE).Gain hands-on experience with cutting-edge technologies including Kubernetes, Terraform, AWS/GCP, CI/CD, and Infrastructure as Code (IaC).Influence technical direction and collaborate with cross-functional engineering, security, and product teams.GCS is acting as an Employment Business in relation to this vacancy.
Read lessSenior DevOps Manager Hybrid · 3 days/week onsite · Full-time About the role Were looking for a seasoned... Read more
Senior DevOps Manager
Hybrid · 3 days/week onsite · Full-time
About the role
Were looking for a seasoned Senior DevOps Manager to lead a global operations team delivering world-class infrastructure performance and reliability. This is a high-impact leadership role responsible for driving operational excellence, fostering cross-functional collaboration, and implementing scalable project-management practices to support enterprise customers.
As a key member of the engineering leadership team, youll ensure 24/7 operational stability across the infrastructure while continuously improving processes, systems, and team capabilities to meet evolving business needs.
What youll do
Lead and mentor Technical Operations engineers across multiple time zones
Build a collaborative team culture centred on knowledge sharing, innovation, and operational excellence
Own 24/7 operational stability - incident response, escalation procedures, and post-incident reviews
Drive incident management including alert management, outage response, and root cause analysis (RCA/CAR)
Implement and maintain SRE practices: SLOs, error budgets, and reliability engineering
Establish robust monitoring and alerting using APM tools and diagnostic dashboards
Lead technical project delivery with clear timelines, resource allocation, and stakeholder communication
Drive automation initiatives - runbook automation, deployment automation, and infrastructure-as-code
Manage infrastructure automation using Terraform, Kubernetes, and cloud platforms (GCP/AWS)
Provide executive reporting on operational metrics, project status, and team performance
Drive cost-optimisation and resource-planning initiatives
What youll bring
8+ years in technical operations, with 4+ years leading Technical Operations, SRE, or infrastructure teams
Proven track record developing and mentoring high-performing teams
Strong project management skills (Agile/Kanban, JIRA)
Excellent communication, including presenting to executive stakeholders
Deep SRE experience - incident management, monitoring, and reliability engineering
Infrastructure automation expertise (Terraform, Kubernetes, Docker, CI/CD)
Cloud platform proficiency (GCP/AWS), including networking, security, and cost management
Monitoring and observability experience (Prometheus, Grafana, APM, log aggregation)
24/7 operations experience, including on-call and global team coordination
Change management and compliance experience (SOC 2, security reviews, audits)
Linux, networking protocols, and security fundamentals
Nice to have
Cloud certifications (GCP Professional, AWS Solutions Architect) or SRE credentials
Enterprise networking, wireless infrastructure, IoT, or telecoms experience
Security and compliance expertise (vulnerability management, regulatory frameworks)
Degree in Computer Science, Engineering, or a related field
Multi-region operations experience across time zones
Whats on offer
Comprehensive medical, dental, and vision plans
Life and accidental death insurance
401(k) and company incentive plan
Paid holidays and vacation, plus additional leave options
Hybrid working (3 days/week onsite)
GCS is acting as an Employment Business in relation to this vacancy.
Read lessDevOps Manager Hybrid · 3 days/week onsite · Full-timeWe're looking for an experienced DevOps Manager to lead an... Read more
DevOps Manager
Hybrid · 3 days/week onsite · Full-time
We're looking for an experienced DevOps Manager to lead an operations team focused on infrastructure performance, reliability, and continuous improvement. In this role you'll guide and develop engineers, drive operational excellence, and help deliver stable, scalable, well-automated systems.
This is a leadership role that blends people management, technical direction, and hands-on problem solving - a strong fit for someone who enjoys building high-performing teams and modern operational practices.
What you'll do
Lead, mentor, and develop a team of operations/infrastructure engineersOwn operational stability, incident response, and continuous improvementDrive automation and modern DevOps/SRE practicesManage projects, timelines, and stakeholder communicationPartner across product, security, and development teamsWhat you'll bring
Experience leading operations, infrastructure, or SRE teamsStrong background in cloud, automation, and infrastructure practicesFamiliarity with tools like Kubernetes, Terraform, and CI/CDSolid communication and team-leadership skillsA track record of building reliable, scalable systemsNice to have
Cloud certifications or multi-cloud experienceExposure to security or compliance frameworksMonitoring/observability experienceBackground in networking or enterprise infrastructureWhat's on offer
Competitive compensation and benefitsHybrid workingThe opportunity to lead and shape a growing operations functionGCS is acting as an Employment Business in relation to this vacancy.
Read lessSenior DevOps Manager Hybrid · 3 days/week onsite · Full-timeAbout the roleWe're looking for a seasoned Senior DevOps... Read more
Senior DevOps Manager
Hybrid · 3 days/week onsite · Full-time
About the role
We're looking for a seasoned Senior DevOps Manager to lead a global operations team delivering world-class infrastructure performance and reliability. This is a high-impact leadership role responsible for driving operational excellence, fostering cross-functional collaboration, and implementing scalable project-management practices to support enterprise customers.
As a key member of the engineering leadership team, you'll ensure 24/7 operational stability across the infrastructure while continuously improving processes, systems, and team capabilities to meet evolving business needs.
What you'll do
Lead and mentor Technical Operations engineers across multiple time zonesBuild a collaborative team culture centred on knowledge sharing, innovation, and operational excellenceOwn 24/7 operational stability - incident response, escalation procedures, and post-incident reviewsDrive incident management including alert management, outage response, and root cause analysis (RCA/CAR)Implement and maintain SRE practices: SLOs, error budgets, and reliability engineeringEstablish robust monitoring and alerting using APM tools and diagnostic dashboardsLead technical project delivery with clear timelines, resource allocation, and stakeholder communicationDrive automation initiatives - runbook automation, deployment automation, and infrastructure-as-codeManage infrastructure automation using Terraform, Kubernetes, and cloud platforms (GCP/AWS)Provide executive reporting on operational metrics, project status, and team performanceDrive cost-optimisation and resource-planning initiativesWhat you'll bring
8+ years in technical operations, with 4+ years leading Technical Operations, SRE, or infrastructure teamsProven track record developing and mentoring high-performing teamsStrong project management skills (Agile/Kanban, JIRA)Excellent communication, including presenting to executive stakeholdersDeep SRE experience - incident management, monitoring, and reliability engineeringInfrastructure automation expertise (Terraform, Kubernetes, Docker, CI/CD)Cloud platform proficiency (GCP/AWS), including networking, security, and cost managementMonitoring and observability experience (Prometheus, Grafana, APM, log aggregation)24/7 operations experience, including on-call and global team coordinationChange management and compliance experience (SOC 2, security reviews, audits)Linux, networking protocols, and security fundamentalsNice to have
Cloud certifications (GCP Professional, AWS Solutions Architect) or SRE credentialsEnterprise networking, wireless infrastructure, IoT, or telecoms experienceSecurity and compliance expertise (vulnerability management, regulatory frameworks)Degree in Computer Science, Engineering, or a related fieldMulti-region operations experience across time zonesWhat's on offer
Comprehensive medical, dental, and vision plansLife and accidental death insurance401(k) and company incentive planPaid holidays and vacation, plus additional leave optionsHybrid working (3 days/week onsite)GCS is acting as an Employment Business in relation to this vacancy.
Read lessAll your saved jobs are no longer available or you've already applied.
for the following search criteria