• Your primary focus will be on the reliability and uptime of our customer-facing platforms
• You will be embedded within our unified CloudOps team - a team that brings together platform engineering, SRE, and DevOps functions under one roof
• Your day-to-day will centre on building stability, responding to active incidents, implementing short-term fixes, and helping to defend the uptime of our customer systems
• You will also feed into a broader platform, and DevOps work through cross-training and collaborative engineering
• This role offers a real opportunity for someone who comes with experience but wants to keep growing - learning new skills, shaping improvements, and contributing to how we build and run things for the future
Requirements
- Demonstrable experience in an SRE or similar operations role within a SaaS business
- Strong hands-on experience with at least one major cloud provider, with a preference for Microsoft Azure
- Infrastructure as Code (IaC) - for example, Terraform, Bicep, or ARM templates
- Container orchestration - for example, Kubernetes or Azure Kubernetes Service (AKS)
- Monitoring and observability tooling - for example, Prometheus, Grafana, Datadog, or Azure Monitor
- CI/CD pipelines and version control - for example, GitHub Actions, Azure DevOps
- AI tooling - proficient in leveraging AI tools to improve operational efficiency, accelerate troubleshooting, automate repetitive tasks, and augment day-to-day engineering workflows
- Windows Server administration - including Active Directory, DNS, Group Policy, and core Windows Server infrastructure fundamentals
- Advanced Networking - including TCP/IP, DNS, VPNs, firewalls, load balancers, and general network troubleshooting in cloud and hybrid environments
- Linux systems administration and networking fundamentals
- Documentation - proficient in producing clear technical documentation, including runbooks, incident reports, architecture diagrams, and process guides
- Atlassian suite - hands-on experience with Jira and Confluence for issue tracking, sprint management, and knowledge base management
ATS Optimization Keywords
Below are skills and terms extracted directly from this job posting to improve Applicant Tracking System (ATS) visibility. This unique feature helps candidates tailor their applications more effectively - a feature exclusive to JobTailor job listings.
Hard Skills
- SRE
- Infrastructure as Code
- Terraform
- Bicep
- Kubernetes
- Azure Kubernetes Service
- Prometheus
- Grafana
- GitHub Actions
- Windows Server administration
Soft Skills
- collaborative engineering
- problem-solving
- communication
- adaptability
- continuous learning