Skip to main content
Posted 11 August, 2026
Government Digital & Data

Mid and Senior Site Reliability Engineer - Government Digital Service - G7

England, UK Hybrid Full Time
Salary: £56,850 to £81,720 Annually
based on location and capability

Location

London, Manchester

About the job

Job summary

The Government Digital Service (GDS) is the digital centre of government. We are responsible for setting, leading and delivering the vision for a modern digital government.

Our priorities are to drive a modern digital government, by:

  1. joining up public sector services
  2. harnessing the power of AI for the public good
  3. strengthening and extending our digital and data public infrastructure
  4. elevating leadership and investing in talent
  5. funding for outcomes andprocuringfor growth and innovation
  6. committing to transparency and driving accountability

We are home to the Incubator for Artificial Intelligence (I.AI), the world-leading GOV.UK and at the forefront of coordinating the UK’s geospatial strategy and activity. We lead the Government Digital and Data function and champion the work of digital teams across government.

We’re part of the Department for Science, Innovation and Technology (DSIT) and employ more than 1,000 people all over the UK, with hubs in Manchester, London and Bristol.

The Government Digital Service is where talent translates into impact. From your first day, you’ll be working with some of the world’s most highly-skilled digital professionals, all contributing their knowledge to make change on a national scale.

Join us for rewarding work that makes a difference across the UK. You'll solve some of the nation’s highest-priority digital challenges, helping millions of people access services they need.

About Engineering Enablement

Part of the newly formed Platform Engineering, Resilience and Cyber directorate, the Engineering Enablement programme is boosting efficiency, strengthening security, and improving developer experience by providing infrastructure, tools and standards. We've been working for some time to manage our AWS estate and establish a robust privileged access management system for our most sensitive engineering access, based on Microsoft Entra.

Two roles (one Senior, one Mid) will go to our Engineering Access team. This team is rolling out a Privileged Access Management system based on Microsoft Entra across all GDS production services. The goal is a robust system fully integrated across GDS's engineering estate, reducing TOIL and human error associated with manual access management, and ensuring engineering productivity.

The third role (Senior) will go to our Cloud Platform team. The cloud platform team currently own and operate our AWS estate, and over the next 6-9 months will take over operational responsibility for the Azure Landing zone developed as part of the Engineering Access work. Your role will be to provide Azure experience and expertise to support the team in this journey.

Job description

As an Site Reliability Engineer (Platform/Identity/Azure) you will:

  • be part of a multidisciplinary team implementing and supporting our Azure cloud platform or Entra ID identity and access management system
  • write infrastructure as code using terraform to ensure our infrastructure is consistent, reusable and reliable
  • deploy and configure observability tools to enable our teams to identify and respond to operational issues quickly and effectively
  • build CI/CD pipelines to enable the team to get code into production quickly and reliably
  • provide day-to-day support for our platforms and tools to ensure they remain available, secure and robust
  • participate in on-call rotations when necessary
  • solve complex and interesting problems
  • share your knowledge and expertise with your peers and the wider team to drive consistency and develop a culture of openness and learning

In addition to the above, as a Senior Site Reliability Engineer, you will:

  • line manage 1-2 technologists, supporting their growth and development
  • provide technical leadership within a team, working with other team members to identify the best approaches and solutions

Read about a day in the life of a GDS Site Reliability Engineer on the GDS blog, and watch our video about Becoming a site reliability engineer at GDS.

Person specification

We’re interested in people who have:

  • strong experience with cloud services, architecture and best practices in Microsoft Azure
  • strong experience designing and implementing identity and access management systems using Entra ID, including Conditional Access, PIM and SCIM
  • a Platform Engineering mindset, with a proficiency in designing and implementing scalable, resilient, and secure cloud platforms
  • strong preference for infrastructure-as-code and automation using tools like Terraform
  • experience of using container orchestration systems such as Kubernetes, AKS or serverless application design
  • experience supporting large production services
  • proficiency in at least one programming language (we use Ruby, Typescript, and Python)
  • strong understanding of version control using Git
  • experience of creating pipelines in a CI/CD tool like Github Actions or AWS Codepipeline
  • a strong understanding of security principles and how to keep large operational services secure

In addition to the above, as a Senior Site Reliability Engineer you will have:

  • experience of technical leadership, for example acting as a tech lead for a team or leading a technical initiative
  • experience of line management, mentoring or coaching

If you meet a few of those criteria but think that you might not meet every last one then don’t let that stop you from submitting an application.

Sign up for Job Alerts