Site Reliability Engineer (Hyper-V Infrastructure)
Location: London (5 Days Onsite)
Contract: Inside IR35
Day Rate: Up to £450 per day
Duration: Initial 6 months (Likely Extension)
We're recruiting for an experienced Site Reliability Engineer (SRE)/Infrastructure Engineer to join a major Banking organisation, supporting a mission-critical Microsoft Hyper-V Private Cloud platform.
This is a hands-on engineering role focused on maintaining the reliability, availability and performance of a large-scale Hyper-V estate. You'll work within a highly regulated environment, delivering production support, infrastructure automation, platform upgrades and continuous service improvements.
This opportunity is ideal for an Infrastructure Engineer with strong Hyper-V, Windows Server and PowerShell expertise who enjoys solving complex production issues and improving operational resilience.
Key Responsibilities
- Administer and support enterprise Microsoft Hyper-V infrastructure.
- Manage Hyper-V Failover Clusters, storage, networking and high availability.
- Perform Windows Server administration, patching, upgrades and life cycle management.
- Support private cloud and VDI platforms within a production environment.
- Automate operational tasks using PowerShell.
- Manage disaster recovery, backup and business continuity activities.
- Perform proactive monitoring, incident management, root cause analysis and service improvement.
- Support infrastructure migrations including P2V, V2V and platform modernisation.
- Work closely with Infrastructure, Security and Platform Engineering teams to improve reliability and operational efficiency.
- Produce technical documentation, runbooks and operational procedures.
Essential Skills (must have)
- Strong Microsoft Hyper-V administration experience.
- Windows Server 2016/2019/2022 administration.
- Hyper-V Failover Clustering and High Availability.
- PowerShell Scripting and infrastructure automation.
- Storage technologies including SAN/NAS, Storage Spaces Direct (S2D) or Cluster Shared Volumes (CSV).
- SCVMM (System Center Virtual Machine Manager).
- Disaster Recovery, Backup and Hyper-V Replica.
- Experience supporting enterprise production infrastructure.
- Strong troubleshooting and incident management skills.
- ITIL environment experience
- Experience working within Banking or Financial Services.
Desirable Skills
- Azure Monitor, SCOM, Splunk, Grafana or Prometheus.
- Infrastructure as Code (Terraform, Ansible or similar).
- VDI technologies.
- Azure or Hybrid Cloud.
- Veeam Backup.
Experience Required
- 6+ years' Infrastructure Engineering experience.
- 4+ years' hands-on Microsoft Hyper-V administration.
- Experience supporting highly available production environments.
- Strong understanding of operational resilience, platform reliability and continuous service improvement.
- Previous experience within a regulated Financial Services or Banking environment is highly desirable.