MongoDB Site Reliability Engineer

Reference: MongoDB_SRE_1780401860

MongoDB SRE (AVP) - Knutsford (Hybrid)

Are you a MongoDB expert ready to step into a true engineering role? Join a global team modernising a large‑scale database estate and move beyond repetitive DBA work.

What You'll Do

  • Own MongoDB operations end‑to‑end (clusters, sharding, replica sets, backups).

  • Troubleshoot and resolve complex production issues across L1-L3.

  • Build automation using Python, Ansible, TDD, Agile.

  • Improve observability with better monitoring, alerting, and performance insights.

  • Reduce toil by engineering tools and automation that transform the platform.

Required Skills

  • Deep MongoDB administration expertise.

  • Strong experience with Ops Manager and backup tooling.

  • Solid troubleshooting and production support capability.

  • SRE fundamentals and an automation‑first mindset.

  • Hands‑on Python and Ansible experience.

  • Observability experience (monitoring, alerting, dashboards).

Why Apply

  • Perfect for Senior DBAs wanting to transition into SRE/Engineering.

  • ~25% of your time spent coding and automating, with more growth ahead.

  • High‑impact role shaping a global database platform.

GCS is acting as an Employment Agency in relation to this vacancy.

£45,000.00 - £60,000.00
Per annum
GBP45000 - GBP60000 per annum

Knutsford

Permanent

Added 02/06/2026
Reference: MongoDB_SRE_1780401860

MongoDB Site Reliability Engineer

Knutsford
Permanent

Other similar jobs

Site Reliability Engineer

Added 17/07/2026

Site Reliability Engineer (SRE) - Level 3Job SummaryWe are seeking a Site Reliability Engineer (SRE) to build, maintain, and support reliable, scalable, and secure cloud platforms. The ideal candidate will have experience with cloud infrastructure, Kubernetes, automation, monitoring, and production support while working closely with engineering teams to improve platform reliability and operational excellence.Key ResponsibilitiesManage and support cloud infrastructure and Kubernetes environments.Automate infrastructure using Infrastructure as Code (IaC).Monitor, troubleshoot, and optimize production systems.Support CI/CD pipelines and deployment automation.Participate in incident response, root cause analysis, and reliability improvements.Collaborate with engineering teams on platform enhancements.Develop automation using scripting and DevOps best practices.Participate...

Learn more

Site Reliability Engineer

Added 16/07/2026

Site Reliability Engineer III (AI Platform)Location: Mount Laurel, NJ (Onsite)Duration: ContractExperience: 4+ yearsAbout the RoleWe are seeking a Site Reliability Engineer (SRE) III to support a cutting-edge AI Platform Engineering team responsible for building and maintaining the infrastructure behind enterprise AI and machine learning applications. This is an exciting opportunity to work on large-scale distributed systems, Kubernetes environments, and cloud-native platforms that power next-generation AI solutions.The ideal candidate has a strong background in cloud infrastructure, Kubernetes, Infrastructure as Code, observability, and automation. Experience supporting production environments at scale is essential.ResponsibilitiesDesign, implement, and support highly available, scalable, and secure cloud infrastructure.Maintain...

Learn more

Site Reliability Engineer

Added 16/07/2026

Site Reliability Engineer III (AI Platform)Location: Mount Laurel, NJ (Onsite)Duration: ContractExperience: 4+ yearsAbout the RoleWe are seeking a Site Reliability Engineer (SRE) III to support a cutting-edge AI Platform Engineering team responsible for building and maintaining the infrastructure behind enterprise AI and machine learning applications. This is an exciting opportunity to work on large-scale distributed systems, Kubernetes environments, and cloud-native platforms that power next-generation AI solutions.The ideal candidate has a strong background in cloud infrastructure, Kubernetes, Infrastructure as Code, observability, and automation. Experience supporting production environments at scale is essential.ResponsibilitiesDesign, implement, and support highly available, scalable, and secure cloud infrastructure.Maintain...

Learn more

Microsoft SQL Database Site Reliability Engineer

Added 02/06/2026

Step into a high‑impact engineering role where you'll shape the future of Microsoft SQL operations at enterprise scale. As a Database SRE, you'll combine deep SQL Server expertise with modern SRE practices to build a more reliable, automated, and observable database platform for one of the world's largest financial institutions. What You'll DoLead SQL Engineering - Solve complex SQL Server 2016-2022 challenges across availability, tuning, performance, and architecture.Shape the MSSQL SRE practice - Influence standards, patterns, SLIs/SLOs, and operational models for the SQL estate.Act as the top technical escalation - Provide expert‑level guidance on incidents, root cause, and long‑term fixes.Drive...

Learn more

Azure Site Reliability Engineer

Added 29/05/2026

Azure Site Reliability Engineer (SRE)Location: Glasgow / Knutsford (Hybrid- 2 days a week in office)Team: 6 UK / 5 IndiaEnvironment: Part of a wider multi‑cloud engineering organisation (Azure, AWS, GCP)Growth: Significant technical development opportunities across cloud engineering, automation, and platform build Role OverviewWe are looking for a hands‑on Azure SRE who can design, build, and automate enterprise‑grade Azure Landing Zones and cloud governance frameworks. This is not an application development role - it is a platform engineering role focused on controls, policies, guardrails, IaC, and DevOps automation.You will work as part of a global SRE function, collaborating with engineers in...

Learn more

Site Reliability Specialist

Added 17/07/2026

Site Reliability Engineer (SRE) - Level 3Job SummaryWe are seeking a Site Reliability Engineer (SRE) to build, maintain, and support reliable, scalable, and secure cloud platforms. The ideal candidate will have experience with cloud infrastructure, Kubernetes, automation, monitoring, and production support while working closely with engineering teams to improve platform reliability and operational excellence.Key ResponsibilitiesManage and support cloud infrastructure and Kubernetes environments.Automate infrastructure using Infrastructure as Code (IaC).Monitor, troubleshoot, and optimize production systems.Support CI/CD pipelines and deployment automation.Participate in incident response, root cause analysis, and reliability improvements.Collaborate with engineering teams on platform enhancements.Develop automation using scripting and DevOps best practices.Participate...

Learn more

MongoDB SRE (AVP)

Added 17/04/2026

MongoDB SRE (AVP) - Knutsford (Hybrid)Are you a MongoDB expert ready to step into a true engineering role? Join a global team modernising a large‑scale database estate and move beyond repetitive DBA work.What You'll DoOwn MongoDB operations end‑to‑end (clusters, sharding, replica sets, backups).Troubleshoot and resolve complex production issues across L1-L3.Build automation using Python, Ansible, TDD, Agile.Improve observability with better monitoring, alerting, and performance insights.Reduce toil by engineering tools and automation that transform the platform.Required SkillsDeep MongoDB administration expertise.Strong experience with Ops Manager and backup tooling.Solid troubleshooting and production support capability.SRE fundamentals and an automation‑first mindset.Hands‑on Python and Ansible experience.Observability experience...

Learn more

Site Acquisition Specialist

Added 08/07/2026

Site Acquisition Specialist is responsible for managing wireless site acquisition activities from initial site identification through project completion. This includes obtaining properly leases, coordinating zoning and permitting, working with municipalities and landlords, and processing all project milestones through AT&T's LMPS system. The role supports wireless network deployment projects for AT&T and T-Mobile while partnering with internal teams, contractors, and customers to remain on schedule.Required experience:Site leasingZoning applicationsPermittingMunicipal approvalsLandlord negotiationsUnderstanding wireless cell site development processExperience coordinating with engineering, construction, and general contractorsDay To Day Duties:Manage the complete site acquisition process for assigned projectsSubmit and track projects through AT&T LMPS systemObtain and...

Learn more

Senior Software Engineer/Data Platform Engineer (Databricks, Graph, APIs)

Added 30/04/2026

Senior Software Engineer / Data Platform Engineer (Databricks, Graph, APIs)Location: Philadelphia, PA The team sits within the network technology organisation and is responsible for building advanced data platforms that support digital twin capabilities across the access network. The group combines network design data, telemetry, mapping technologies, and graph intelligence to improve troubleshooting, planning, operational efficiency, and market competitiveness.The team works on highly scalable engineering products including large data pipelines, graph databases, APIs, and mapping platforms. Their work enables smarter network decisions, faster fault resolution, and better use of operational resources.This is a technically strong team focused on solving complex real-world...

Learn more

Field Engineer

Added 20/07/2026

Client - TelecomDuration - 12+ monthsPosition - Hybrid (4 Days on site)Job Summary:We are seeking a highly Field Engineer III to join our remote network surveillance and troubleshooting team supporting the company's 5G RAN network. This on-site role focuses on monitoring, testing, maintaining, and restoring the 5G RAN network. The ideal candidate will lead Tier II field diagnostics, outage restoration, and mentor Tier I NOC staff while supporting new 5G site deployments and fronthaul transport initiatives.Key Responsibilities:Provide on-site operational support to detect incidents, troubleshoot problems, and implement resolutions for deployed 5G Core, RAN, Transport, and other network systems nationwide.Quickly analyse...

Learn more

Field Engineer

Added 20/07/2026

Client - TelecomDuration - 12+ monthsPosition - Hybrid (4 Days on site)Job Summary:We are seeking a highly Field Engineer III to join our remote network surveillance and troubleshooting team supporting the company's 5G RAN network. This on-site role focuses on monitoring, testing, maintaining, and restoring the 5G RAN network. The ideal candidate will lead Tier II field diagnostics, outage restoration, and mentor Tier I NOC staff while supporting new 5G site deployments and fronthaul transport initiatives.Key Responsibilities:Provide on-site operational support to detect incidents, troubleshoot problems, and implement resolutions for deployed 5G Core, RAN, Transport, and other network systems nationwide.Quickly analyse...

Learn more

Field Engineer

Added 20/07/2026

Client - Telecom Duration - 12+ monthsPosition - Hybrid (4 Days on site) Job Summary:We are seeking a highly Field Engineer III to join our remote network surveillance and troubleshooting team supporting the company's 5G RAN network. This on-site role focuses on monitoring, testing, maintaining, and restoring the 5G RAN network. The ideal candidate will lead Tier II field diagnostics, outage restoration, and mentor Tier I NOC staff while supporting new 5G site deployments and fronthaul transport initiatives.Key Responsibilities:Provide on-site operational support to detect incidents, troubleshoot problems, and implement resolutions for deployed 5G Core, RAN, Transport, and other network systems...

Learn more

Platform Engineer - Kubernetes focused

Added 17/07/2026

Senior Platform Engineer (Kubernetes) | ContractWe're partnering with a leading organisation seeking an experienced Senior Platform Engineer to help develop, automate and scale their cloud-native infrastructure environment.This is an excellent opportunity for an engineer who enjoys building robust Kubernetes platforms, driving automation, and working with modern cloud technologies in a highly collaborative environment.What You'll Be DoingDesign, build and support Kubernetes-based platform infrastructure.Drive automation across infrastructure and operational processes using Infrastructure as Code principles.Develop and maintain tooling using technologies such as Terraform, Ansible, Python and Go.Enhance platform reliability, scalability, security and performance.Implement and improve observability, monitoring and alerting solutions.Work closely with...

Learn more

Java Backend Engineer

Added 17/07/2026

Job Role - Java Backend EngineerLocation: Remote (Candidates must be based in Poland)Office Hub: Kraków, PolandContract: 12 Months (B2B/Freelance) We are seeking an experienced Senior Java Backend Engineer to join a high-performing engineering team on a 12-month B2B/Freelance contract, with a strong possibility of extension.This is an exciting opportunity to work on a leading technology platform, collaborating with top-tier engineers in a fast-paced, international environment. You'll play a key role in shaping scalable backend solutions, driving technical excellence, and supporting teams across multiple stakeholders.Key ResponsibilitiesCollaborate with highly skilled engineering teams to build and scale a best-in-class platform.Drive technical excellence and...

Learn more

Cloud Infrastructure Engineer

Added 17/07/2026

Job SummaryWe are seeking a Site Reliability Engineer (SRE) to build, maintain, and support reliable, scalable, and secure cloud platforms. The ideal candidate will have experience with cloud infrastructure, Kubernetes, automation, monitoring, and production support while working closely with engineering teams to improve platform reliability and operational excellence.Key ResponsibilitiesManage and support cloud infrastructure and Kubernetes environments.Automate infrastructure using Infrastructure as Code (IaC).Monitor, troubleshoot, and optimize production systems.Support CI/CD pipelines and deployment automation.Participate in incident response, root cause analysis, and reliability improvements.Collaborate with engineering teams on platform enhancements.Develop automation using scripting and DevOps best practices.Participate in production support and on-call rotations.Required...

Learn more
At least 8 characters, 1 uppercase, 1 lowercase and 1 special character or number
Your file must be a doc, docx or pdf. No larger than 5MB.