Leeds, Manchester, Telford
About the job Job summaryDiscover a career in your hands at HMRC. Whether you're seeking purpose, growth, or a workplace that gives you a true sense of belonging, hear from some of our employees as they share their story about what it's really like to work at HMRC.
Visit our YouTube channel to watch the full series and come and discover your potential.
Are you passionate about cloud platform operations, security, and automation?
Do you have experience in addressing risks, incidents and service quality across technical teams?
Do you want to play a key role in delivering reliable, secure, and efficient platforms that provide services for millions of citizens?
We are looking for a resourceful individual who can ensure our cloud platforms are robust, secure, and optimal. From monitoring performance and managing change to enabling automation and driving standardisation, you'll help keep HMRC's mission-critical cloud platforms running at their best.
Within HMRC's Chief Digital & Information Group (CDIO), specifically in the Enterprise Cloud Services (ECS) team we are redefining our offerings and growing the team of outstanding people to improve the Cloud Centre of Excellence. We are already a diverse team of 90+ technologists, creating a dynamic and inclusive working environment whose skills cover architecture, platform development, service design, platform operations and governance.
Important - Travel to Telford is required as part of this role, and 60% of your working time will need to be office based.
Additional Security Information:
This role requires the successful candidates to hold or be willing to hold Security Check (SC) clearance.
As a Senior Infrastructure Operations Engineer (Platform Operations Engineer) within HMRC's Enterprise Cloud Services (ECS), you will ensure the reliability, security, and efficiency of platforms hosting HMRC's mission-critical services. You will undertake day-to-day operations and drive continual improvement across our cloud platforms, maintaining a high level of availability through proactive monitoring and rapid incident response.
Security and compliance will be central to your role, applying access controls, patching, and vulnerability management to meet industry standards. You will apply change and configuration processes to minimise risk and maintain accurate system data, while using metrics, logging, and automation to optimise performance, reduce manual effort, and control costs.
Working closely with service desks and platform engineering teams, you will resolve incidents, support users, and maintain CI/CD pipelines and Infrastructure as Code. By enabling observability and feedback loops, you will help ensure our platforms continue to run reliably. You will also document procedures, standardise practices, and embrace good operational practises to strengthen resilience and consistency across the organisation.
Your responsibilities will include:
Platform Operations & Reliability
Automation & Infrastructure as Code
Security & Compliance
Performance & Cost Optimisation
Collaboration & Leadership
Mentoring & Technical Leadership
Critical Thinking & Technical Mindset
Essential Criteria:
List Desirable Criteria: