Senior Site Reliability Engineer

Reference: SSRE2508_1787657463

Role Title: Senior Site Reliability Engineer
Location: Knutsford or Glasgow - Hybrid (2 days per week onsite)
Role Category: Permanent

Overview:
We're recruiting for an experienced Senior Site Reliability Engineer to drive reliability, scalability and performance across critical banking systems. This role combines hands-on SRE engineering with technical leadership, with a strong focus on observability, automation, continuous improvement and optimisation.

Responsibilities:
* Build and maintain reliable, scalable and secure infrastructure platforms and solutions.
* Apply SRE and software engineering practices to improve reliability, availability and performance.
* Monitor systems, manage incidents and lead complex troubleshooting and root cause analysis.
* Develop automation using programming and scripting to reduce manual intervention and improve efficiency.
* Develop and improve observability, monitoring, instrumentation and performance capabilities.
* Use data and reliability metrics to drive continuous improvement and optimisation.
* Lead technical discussions, blameless retrospectives and problem-solving activities.
* Work with architects, engineers and stakeholders to define requirements and deliver effective solutions.
* Provide technical leadership, mentoring and guidance while helping to drive SRE maturity across teams.

Required Skills:
* 5+ years' experience in SRE, Production Engineering, Platform Engineering or a related discipline.
* Strong practical experience with SRE principles and production environments.
* Strong programming/scripting skills, such as Python, Go, Java, C# or Bash.
* Proven experience in incident management, troubleshooting and root cause analysis.
* Strong observability and monitoring experience.
* Experience with AWS, Azure or GCP.
* Good understanding of operating systems, networking, cloud infrastructure and automation.
* Experience with Infrastructure-as-Code.
* Strong communication and technical leadership skills.

Desirable Skills:
* SLOs, SLIs, SLAs and error budgets.
* Prometheus, Grafana, Elastic/ELK or OpenTelemetry.
* Kubernetes, containers and distributed systems.
* Performance and resilience engineering.
* Experience driving SRE maturity across engineering teams.
* Financial services or banking experience.

This is an excellent opportunity for a Senior SRE to combine hands-on engineering with technical leadership and help improve reliability, observability and performance across critical banking systems.

GCS is acting as an Employment Agency in relation to this vacancy.

£75,000.00 - £95,000.00
Per annum
GBP75000 - GBP95000 per annum + Bonus
Added 25/08/2026
Reference: SSRE2508_1787657463

Senior Site Reliability Engineer

City of Glasgow, United Kingdom Permanent DevOps

Other similar jobs

Senior Site Reliability Engineer

Added 05/08/2026

About the RoleClient is seeking a highly skilled Senior / Lead Site Reliability Engineer (SRE) to drive reliability, observability, and production excellence across critical business platforms. This role combines hands-on engineering expertise with leadership and governance responsibilities, ensuring services remain scalable, resilient, and aligned with business objectives.As a key member of the technology team, you will champion SRE best practices, improve operational efficiency through automation, enhance observability, and lead the organization's approach to service reliability and incident management. Key ResponsibilitiesReliability Engineering & Production ExcellenceDefine, implement, and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Error Budgets.Drive data-led decisions...

Learn more

Senior Site Reliability Engineer (Infrastructure Focus)

Added 27/07/2026

Hybrid Dublin | Enterprise Cloud PlatformsWe're hiring experienced Site Reliability Engineers to support the evolution of large-scale cloud platforms used globally across critical production environments.Unlike traditional SRE opportunities focused heavily on Kubernetes, this role places greater emphasis on cloud infrastructure, systems engineering and platform reliability.ResponsibilitiesEngineer highly available cloud infrastructure platformsImprove reliability, resilience and operational excellenceDrive automation across provisioning, deployment and operational workflowsDevelop tools and services using Python and modern scripting technologiesSupport multi-region production environmentsParticipate in incident analysis and root cause investigationsReduce operational overhead through automation and engineering improvementsRequired ExperienceStrong cloud infrastructure experience across AWS, Azure or GCPExperience operating customer-facing or...

Learn more

Azure Site Reliability Engineer

Added 04/08/2026

Azure Site Reliability Engineer (SRE)Location: Glasgow / Knutsford (Hybrid- 2 days a week in office)Team: 6 UK / 5 IndiaEnvironment: Part of a wider multi‑cloud engineering organisation (Azure, AWS, GCP)Growth: Significant technical development opportunities across cloud engineering, automation, and platform build Role OverviewWe are looking for a hands‑on Azure SRE who can design, build, and automate enterprise‑grade Azure Landing Zones and cloud governance frameworks. This is not an application development role - it is a platform engineering role focused on controls, policies, guardrails, IaC, and DevOps automation.You will work as part of a global SRE function, collaborating with engineers in...

Learn more

Site Reliability Engineer

Added 17/07/2026

Site Reliability Engineer (SRE) - Level 3Job SummaryWe are seeking a Site Reliability Engineer (SRE) to build, maintain, and support reliable, scalable, and secure cloud platforms. The ideal candidate will have experience with cloud infrastructure, Kubernetes, automation, monitoring, and production support while working closely with engineering teams to improve platform reliability and operational excellence.Key ResponsibilitiesManage and support cloud infrastructure and Kubernetes environments.Automate infrastructure using Infrastructure as Code (IaC).Monitor, troubleshoot, and optimize production systems.Support CI/CD pipelines and deployment automation.Participate in incident response, root cause analysis, and reliability improvements.Collaborate with engineering teams on platform enhancements.Develop automation using scripting and DevOps best practices.Participate...

Learn more

Site Reliability Engineer

Added 16/07/2026

Site Reliability Engineer III (AI Platform)Location: Mount Laurel, NJ (Onsite)Duration: ContractExperience: 4+ yearsAbout the RoleWe are seeking a Site Reliability Engineer (SRE) III to support a cutting-edge AI Platform Engineering team responsible for building and maintaining the infrastructure behind enterprise AI and machine learning applications. This is an exciting opportunity to work on large-scale distributed systems, Kubernetes environments, and cloud-native platforms that power next-generation AI solutions.The ideal candidate has a strong background in cloud infrastructure, Kubernetes, Infrastructure as Code, observability, and automation. Experience supporting production environments at scale is essential.ResponsibilitiesDesign, implement, and support highly available, scalable, and secure cloud infrastructure.Maintain...

Learn more

Site Reliability Engineer

Added 16/07/2026

Site Reliability Engineer III (AI Platform)Location: Mount Laurel, NJ (Onsite)Duration: ContractExperience: 4+ yearsAbout the RoleWe are seeking a Site Reliability Engineer (SRE) III to support a cutting-edge AI Platform Engineering team responsible for building and maintaining the infrastructure behind enterprise AI and machine learning applications. This is an exciting opportunity to work on large-scale distributed systems, Kubernetes environments, and cloud-native platforms that power next-generation AI solutions.The ideal candidate has a strong background in cloud infrastructure, Kubernetes, Infrastructure as Code, observability, and automation. Experience supporting production environments at scale is essential.ResponsibilitiesDesign, implement, and support highly available, scalable, and secure cloud infrastructure.Maintain...

Learn more

MongoDB Site Reliability Engineer

Added 02/06/2026

MongoDB SRE (AVP) - Knutsford (Hybrid)Are you a MongoDB expert ready to step into a true engineering role? Join a global team modernising a large‑scale database estate and move beyond repetitive DBA work.What You'll DoOwn MongoDB operations end‑to‑end (clusters, sharding, replica sets, backups).Troubleshoot and resolve complex production issues across L1-L3.Build automation using Python, Ansible, TDD, Agile.Improve observability with better monitoring, alerting, and performance insights.Reduce toil by engineering tools and automation that transform the platform.Required SkillsDeep MongoDB administration expertise.Strong experience with Ops Manager and backup tooling.Solid troubleshooting and production support capability.SRE fundamentals and an automation‑first mindset.Hands‑on Python and Ansible experience.Observability experience...

Learn more

Microsoft SQL Database Site Reliability Engineer

Added 02/06/2026

Step into a high‑impact engineering role where you'll shape the future of Microsoft SQL operations at enterprise scale. As a Database SRE, you'll combine deep SQL Server expertise with modern SRE practices to build a more reliable, automated, and observable database platform for one of the world's largest financial institutions. What You'll DoLead SQL Engineering - Solve complex SQL Server 2016-2022 challenges across availability, tuning, performance, and architecture.Shape the MSSQL SRE practice - Influence standards, patterns, SLIs/SLOs, and operational models for the SQL estate.Act as the top technical escalation - Provide expert‑level guidance on incidents, root cause, and long‑term fixes.Drive...

Learn more

Azure Site Reliability Engineer

Added 29/05/2026

Azure Site Reliability Engineer (SRE)Location: Glasgow / Knutsford (Hybrid- 2 days a week in office)Team: 6 UK / 5 IndiaEnvironment: Part of a wider multi‑cloud engineering organisation (Azure, AWS, GCP)Growth: Significant technical development opportunities across cloud engineering, automation, and platform build Role OverviewWe are looking for a hands‑on Azure SRE who can design, build, and automate enterprise‑grade Azure Landing Zones and cloud governance frameworks. This is not an application development role - it is a platform engineering role focused on controls, policies, guardrails, IaC, and DevOps automation.You will work as part of a global SRE function, collaborating with engineers in...

Learn more

Site Reliability Specialist

Added 17/07/2026

Site Reliability Engineer (SRE) - Level 3Job SummaryWe are seeking a Site Reliability Engineer (SRE) to build, maintain, and support reliable, scalable, and secure cloud platforms. The ideal candidate will have experience with cloud infrastructure, Kubernetes, automation, monitoring, and production support while working closely with engineering teams to improve platform reliability and operational excellence.Key ResponsibilitiesManage and support cloud infrastructure and Kubernetes environments.Automate infrastructure using Infrastructure as Code (IaC).Monitor, troubleshoot, and optimize production systems.Support CI/CD pipelines and deployment automation.Participate in incident response, root cause analysis, and reliability improvements.Collaborate with engineering teams on platform enhancements.Develop automation using scripting and DevOps best practices.Participate...

Learn more

Site Acquisition Specialist

Added 08/07/2026

Site Acquisition Specialist is responsible for managing wireless site acquisition activities from initial site identification through project completion. This includes obtaining properly leases, coordinating zoning and permitting, working with municipalities and landlords, and processing all project milestones through AT&T's LMPS system. The role supports wireless network deployment projects for AT&T and T-Mobile while partnering with internal teams, contractors, and customers to remain on schedule.Required experience:Site leasingZoning applicationsPermittingMunicipal approvalsLandlord negotiationsUnderstanding wireless cell site development processExperience coordinating with engineering, construction, and general contractorsDay To Day Duties:Manage the complete site acquisition process for assigned projectsSubmit and track projects through AT&T LMPS systemObtain and...

Learn more

Senior Fiber Engineer

Added 25/08/2026

Senior Fiber EngineerLocation: Remote Contract: 12+ MonthsAbout the RoleWe are seeking an experienced Fiber Engineer to join a dynamic field team supporting fiber network testing, troubleshooting, and remediation activities across ISP and OSP environments. This role requires hands-on expertise in OTDR analysis, fiber characterization, bidirectional testing, and fusion splicing to ensure the performance, reliability, and integrity of fiber optic networks.The ideal candidate will have a strong telecommunications background and be comfortable working directly with vendors, project managers, and technical teams while traveling to customer sites as needed.Key ResponsibilitiesPerform OTDR testing to validate fiber paths and identify faults, bends, splice losses,...

Learn more

Senior Modern Workplace Engineer

Added 25/08/2026

Senior Modern Workplace EngineerLocation: Dublin, IrelandWorking Model: HybridContract: 12 MonthsThe OpportunityA large international organisation is looking for a Senior Modern Workplace Engineer to join its Professional Services team on an initial 12-month contract.This is a hands-on technical position for an engineer who knows how to build, optimise and troubleshoot enterprise Modern Workspace environments.The focus is delivery - taking complex Modern Workplace requirements and turning them into stable, scalable solutions across Citrix, VDI, cloud infrastructure, identity, networking and user experience.What You'll Be DoingTake technical ownership of Modern Workplace and EUC project deliveryDesign and implement enterprise Citrix environmentsBuild and optimise Citrix Virtual...

Learn more

Senior Fiber Engineer

Added 20/08/2026

Senior Fiber EngineerWork Mode: Remote Contract: 12+ MonthsWe're looking for an experienced Fiber Engineer to support fiber network testing, troubleshooting, and remediation efforts across ISP and OSP environments. The ideal candidate will have a strong background in OTDR analysis, fiber characterization, bidirectional testing, and fusion splicing, with the ability to identify and resolve fiber-related network issues in both field and lab environments.ResponsibilitiesPerform OTDR testing and analyze results to identify faults, splice loss, bends, and attenuation issuesConduct fiber characterization, CD/PMD testing, bidirectional testing, and power/loss testingTroubleshoot fiber network performance issues and recommend corrective actionsPerform fusion splicing, fiber repairs, and remediation activities...

Learn more

Senior Software Engineer

Added 20/08/2026

Senior Software Engineer Locations: Pittsburgh, PA We're supporting an exciting robotics and automation project developing a vision-based quality control platform for robotic storage systems used at large scale across multiple operational sites.The platform combines NVIDIA Jetson edge devices, AWS IoT Core, MQTT, Lambda, and S3 to process inspection data and support ML-driven pass/fail decisions. This is a greenfield opportunity to help build and scale the device-to-cloud architecture from the ground up. ResponsibilitiesDesign and build scalable backend systems in PythonDevelop device-to-cloud pipelines using AWS IoT Core and MQTTBuild integrations with AWS services and internal APIsSupport system integration, testing, and deploymentCollaborate with...

Learn more
At least 8 characters, 1 uppercase, 1 lowercase and 1 special character or number
Your file must be a doc, docx or pdf. No larger than 5MB.