it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Career Advice
  • Contact us
  • Employers
    • Register as Employer
    • Pricing Plans
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Career Advice
  • Contact us
  • Employers
    • Register as Employer
    • Pricing Plans
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

6 jobs found

Email me jobs like this
Refine Search
Current Search
aws cloud operations engineer night shift
Spectrum IT Recruitment
Senior Site Reliability Engineer
Spectrum IT Recruitment
Site Reliability Engineer (SRE) AWS Kubernetes Fully Remote (UK) 24/7 Shift Pattern (28-day rota including days & nights) £ Competitive + Bonus + Excellent Benefits Build resilient cloud platforms that support critical national services. We're recruiting Site Reliability Engineers to join a global leader in AI-powered customer experience and cloud technology. Following the award of a major government programme, they're expanding their engineering teams to build and support highly secure, cloud-native platforms that deliver sensitive communication services. This is an opportunity to join an organisation investing heavily in modern cloud engineering, automation and reliability. Working as part of a collaborative SRE team, you'll help ensure large-scale production environments remain secure, available and resilient, whilst continuously improving the way they're operated through automation and engineering best practice. If you enjoy solving production challenges, improving reliability and automating away operational toil, we'd love to hear from you. What you'll be doing Monitoring and maintaining highly available production platforms running in AWS Responding to and managing production incidents across a 24/7 service Investigating complex technical issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn't essential. Why join? This is far more than a traditional NOC role. You'll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you'll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services. If you're passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we'd love to hear from you. Apply today or contact Dave Carlisle at Spectrum IT Recruitment for a confidential discussion. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
23/07/2026
Full time
Site Reliability Engineer (SRE) AWS Kubernetes Fully Remote (UK) 24/7 Shift Pattern (28-day rota including days & nights) £ Competitive + Bonus + Excellent Benefits Build resilient cloud platforms that support critical national services. We're recruiting Site Reliability Engineers to join a global leader in AI-powered customer experience and cloud technology. Following the award of a major government programme, they're expanding their engineering teams to build and support highly secure, cloud-native platforms that deliver sensitive communication services. This is an opportunity to join an organisation investing heavily in modern cloud engineering, automation and reliability. Working as part of a collaborative SRE team, you'll help ensure large-scale production environments remain secure, available and resilient, whilst continuously improving the way they're operated through automation and engineering best practice. If you enjoy solving production challenges, improving reliability and automating away operational toil, we'd love to hear from you. What you'll be doing Monitoring and maintaining highly available production platforms running in AWS Responding to and managing production incidents across a 24/7 service Investigating complex technical issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn't essential. Why join? This is far more than a traditional NOC role. You'll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you'll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services. If you're passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we'd love to hear from you. Apply today or contact Dave Carlisle at Spectrum IT Recruitment for a confidential discussion. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
Spectrum IT Recruitment
Senior Site Reliability Engineer
Spectrum IT Recruitment
Site Reliability Engineer (SRE) AWS Kubernetes Fully Remote (UK) 24/7 Shift Pattern (28-day rota including days & nights) Competitive + Bonus + Excellent Benefits Build resilient cloud platforms that support critical national services. We're recruiting Site Reliability Engineers to join a global leader in AI-powered customer experience and cloud technology. Following the award of a major government programme, they're expanding their engineering teams to build and support highly secure, cloud-native platforms that deliver sensitive communication services. This is an opportunity to join an organisation investing heavily in modern cloud engineering, automation and reliability. Working as part of a collaborative SRE team, you'll help ensure large-scale production environments remain secure, available and resilient, whilst continuously improving the way they're operated through automation and engineering best practice. If you enjoy solving production challenges, improving reliability and automating away operational toil, we'd love to hear from you. What you'll be doing Monitoring and maintaining highly available production platforms running in AWS Responding to and managing production incidents across a 24/7 service Investigating complex technical issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn't essential. Why join? This is far more than a traditional NOC role. You'll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you'll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services. If you're passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we'd love to hear from you. Apply today or contact Dave Carlisle at Spectrum IT Recruitment for a confidential discussion. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
21/07/2026
Full time
Site Reliability Engineer (SRE) AWS Kubernetes Fully Remote (UK) 24/7 Shift Pattern (28-day rota including days & nights) Competitive + Bonus + Excellent Benefits Build resilient cloud platforms that support critical national services. We're recruiting Site Reliability Engineers to join a global leader in AI-powered customer experience and cloud technology. Following the award of a major government programme, they're expanding their engineering teams to build and support highly secure, cloud-native platforms that deliver sensitive communication services. This is an opportunity to join an organisation investing heavily in modern cloud engineering, automation and reliability. Working as part of a collaborative SRE team, you'll help ensure large-scale production environments remain secure, available and resilient, whilst continuously improving the way they're operated through automation and engineering best practice. If you enjoy solving production challenges, improving reliability and automating away operational toil, we'd love to hear from you. What you'll be doing Monitoring and maintaining highly available production platforms running in AWS Responding to and managing production incidents across a 24/7 service Investigating complex technical issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn't essential. Why join? This is far more than a traditional NOC role. You'll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you'll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services. If you're passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we'd love to hear from you. Apply today or contact Dave Carlisle at Spectrum IT Recruitment for a confidential discussion. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
DevOps Engineer- Night Shift(10:00 PM - 6:00 AM)
Linuxconfig
United Kingdom - Remote At NiCE, we don't limit our challenges. We challenge our limits. Always. We're ambitious. We're game changers. And we play to win. We set the highest standards and execute beyond them. And if you're like us, we can offer you the ultimate career opportunity that will light a fire within you. So, what's the role all about? The DevOps Engineer is a hybrid, senior level role sitting at the intersection of operational reliability and software delivery automation. You will function as an integrated part of a cross functional engineering team, combining the proactive service management mindset of an Application Operations Engineer with the automation first philosophy of a DevOps practitioner. You will be responsible for keeping production environments healthy and performant, while simultaneously designing and maintaining the CI/CD pipelines, infrastructure as code frameworks, and tooling that enable rapid, high quality software delivery. You are the connective tissue between engineering, platform, and operations - someone who is equally comfortable in an incident bridge call and a sprint planning meeting. How will you make an impact? DevOps & Automation Design, build, and maintain continuous integration and continuous delivery (CI/CD) pipelines for rapid, quality assured deployment of software deliverables. Build and manage Infrastructure as Code (IaC) using tools such as CloudFormation, Ansible, Terraform, Chef, or Puppet. Manage day to day operations of release pipelines, build tools, artifact repositories, and source control systems. Coordinate build and release activities with engineering, QA, product, and other stakeholders across the organisation. Identify, research, and prototype new technologies and practices to continuously improve DevOps processes and team efficiency. Maintain and upgrade DevOps systems in both production and non production environments on an ongoing basis. Cloud & Application Operations Proactively monitor infrastructure and application health - including CPU, memory, file systems, databases, batch jobs, and network performance - and respond swiftly to anomalies. Identify and resolve operational issues including infrastructure failures, batch processing errors, network disruptions, and client data feed problems. Troubleshoot and respond to production downtime, performance degradation, and security related incidents in a timely, structured manner. Perform end to end operational duties covering application server health, service availability, and platform integrity in accordance with documented processes and runbooks. Review and manage client service request tickets in adherence to defined SLAs, ensuring accountability and timely resolution. Provide on call off hour support as part of a structured rotation, including during non prime and weekend shift windows as required. Documentation, Communication & Governance Maintain complete and accurate operational documentation including incident tracking, change logs, and runbooks. Produce metric reports and regular productivity/status updates for internal stakeholders and management. Communicate proactively and clearly - both written and verbal - with internal teams, leadership, and customers on a daily basis. Liaise with management to share feedback on existing and new processes, methodologies, best practices, and technology changes. Work efficiently under pressure to meet tight deadlines while maintaining the professionalism, accuracy, and consistency expected in a high availability environment. Demonstrate a high level of individual accountability and deliver service and support that consistently exceeds client expectations. Have you got what it takes? Education & Experience Bachelor's degree in Computer Science, Information Technology, Business Information Systems, or a related field (or equivalent practical experience). 2+ years of combined experience in application/production support, cloud operations, and/or software DevOps engineering in a high availability SLA environment. Demonstrated experience working as a contributor on a software engineering or platform team. Technical Skills - Required Strong proficiency with Linux and Unix environments; working knowledge of Windows Server administration. Experience writing scripting languages - Python, PowerShell, and/or Perl - for automation, monitoring, and tooling. Experience with distributed source control systems, preferably GitHub or BitBucket. Solid understanding of application server technologies including Tomcat and SSH based remote management. Database experience with one or more of: SQL Server, Oracle, or MySQL - including querying, performance tuning, backup/restore, and lifecycle management. Experience with application debugging, performance analysis, and scalability assessment. Familiarity with standard application security compliance and best practices. Knowledge of fault detection, RCA (Root Cause Analysis), and structured resolution processes. Experience with Amazon Web Services (AWS) - core services for compute, storage, networking, and monitoring. Technical Skills - Deep Knowledge in at Least One of: Database Administration: Structured and/or unstructured, indexing, performance tuning, backup/restore, data lifecycle management, scaling. Layer 2/3 Networking: DNS, SSL/TLS, Load Balancing, IPv4 subnetting, firewalling, and CDN configuration. Operating Systems & Virtualisation: Linux/Windows, containers, orchestration (Kubernetes), storage types and performance, monitoring, and capacity planning. VoIP Administration: Signalling, encoding/decoding, protocols including SIP, RTP, Media Gateway, security, border controllers, and QoS. You will have an advantage if you also have: Familiarity with CI/CD automation tools such as Jenkins, CircleCI, Bamboo, or TFS Build. Experience with release pipeline tooling - Concourse, Thoughtworks Go, Octopus Deploy, ElectricFlow, or XebiaLabs. Experience with Docker containers, microservices architecture, and container orchestration (Kubernetes). Experience with infrastructure automation tools: Ansible, Chef, Puppet, or AWS CloudFormation. Experience with Artifactory or similar artifact repository management. NICE product knowledge and/or implementation or support experience with NICE CXone or related platforms. Knowledge of ETL processes and data pipeline management. Call centre or telecoms industry experience. What's in it for you? Learn more about the Benefits at NICE NICE is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, age, sex, marital status, ancestry, neurotype, physical or mental disability, veteran status, gender identity, sexual orientation, or any other category protected by law.
18/07/2026
Full time
United Kingdom - Remote At NiCE, we don't limit our challenges. We challenge our limits. Always. We're ambitious. We're game changers. And we play to win. We set the highest standards and execute beyond them. And if you're like us, we can offer you the ultimate career opportunity that will light a fire within you. So, what's the role all about? The DevOps Engineer is a hybrid, senior level role sitting at the intersection of operational reliability and software delivery automation. You will function as an integrated part of a cross functional engineering team, combining the proactive service management mindset of an Application Operations Engineer with the automation first philosophy of a DevOps practitioner. You will be responsible for keeping production environments healthy and performant, while simultaneously designing and maintaining the CI/CD pipelines, infrastructure as code frameworks, and tooling that enable rapid, high quality software delivery. You are the connective tissue between engineering, platform, and operations - someone who is equally comfortable in an incident bridge call and a sprint planning meeting. How will you make an impact? DevOps & Automation Design, build, and maintain continuous integration and continuous delivery (CI/CD) pipelines for rapid, quality assured deployment of software deliverables. Build and manage Infrastructure as Code (IaC) using tools such as CloudFormation, Ansible, Terraform, Chef, or Puppet. Manage day to day operations of release pipelines, build tools, artifact repositories, and source control systems. Coordinate build and release activities with engineering, QA, product, and other stakeholders across the organisation. Identify, research, and prototype new technologies and practices to continuously improve DevOps processes and team efficiency. Maintain and upgrade DevOps systems in both production and non production environments on an ongoing basis. Cloud & Application Operations Proactively monitor infrastructure and application health - including CPU, memory, file systems, databases, batch jobs, and network performance - and respond swiftly to anomalies. Identify and resolve operational issues including infrastructure failures, batch processing errors, network disruptions, and client data feed problems. Troubleshoot and respond to production downtime, performance degradation, and security related incidents in a timely, structured manner. Perform end to end operational duties covering application server health, service availability, and platform integrity in accordance with documented processes and runbooks. Review and manage client service request tickets in adherence to defined SLAs, ensuring accountability and timely resolution. Provide on call off hour support as part of a structured rotation, including during non prime and weekend shift windows as required. Documentation, Communication & Governance Maintain complete and accurate operational documentation including incident tracking, change logs, and runbooks. Produce metric reports and regular productivity/status updates for internal stakeholders and management. Communicate proactively and clearly - both written and verbal - with internal teams, leadership, and customers on a daily basis. Liaise with management to share feedback on existing and new processes, methodologies, best practices, and technology changes. Work efficiently under pressure to meet tight deadlines while maintaining the professionalism, accuracy, and consistency expected in a high availability environment. Demonstrate a high level of individual accountability and deliver service and support that consistently exceeds client expectations. Have you got what it takes? Education & Experience Bachelor's degree in Computer Science, Information Technology, Business Information Systems, or a related field (or equivalent practical experience). 2+ years of combined experience in application/production support, cloud operations, and/or software DevOps engineering in a high availability SLA environment. Demonstrated experience working as a contributor on a software engineering or platform team. Technical Skills - Required Strong proficiency with Linux and Unix environments; working knowledge of Windows Server administration. Experience writing scripting languages - Python, PowerShell, and/or Perl - for automation, monitoring, and tooling. Experience with distributed source control systems, preferably GitHub or BitBucket. Solid understanding of application server technologies including Tomcat and SSH based remote management. Database experience with one or more of: SQL Server, Oracle, or MySQL - including querying, performance tuning, backup/restore, and lifecycle management. Experience with application debugging, performance analysis, and scalability assessment. Familiarity with standard application security compliance and best practices. Knowledge of fault detection, RCA (Root Cause Analysis), and structured resolution processes. Experience with Amazon Web Services (AWS) - core services for compute, storage, networking, and monitoring. Technical Skills - Deep Knowledge in at Least One of: Database Administration: Structured and/or unstructured, indexing, performance tuning, backup/restore, data lifecycle management, scaling. Layer 2/3 Networking: DNS, SSL/TLS, Load Balancing, IPv4 subnetting, firewalling, and CDN configuration. Operating Systems & Virtualisation: Linux/Windows, containers, orchestration (Kubernetes), storage types and performance, monitoring, and capacity planning. VoIP Administration: Signalling, encoding/decoding, protocols including SIP, RTP, Media Gateway, security, border controllers, and QoS. You will have an advantage if you also have: Familiarity with CI/CD automation tools such as Jenkins, CircleCI, Bamboo, or TFS Build. Experience with release pipeline tooling - Concourse, Thoughtworks Go, Octopus Deploy, ElectricFlow, or XebiaLabs. Experience with Docker containers, microservices architecture, and container orchestration (Kubernetes). Experience with infrastructure automation tools: Ansible, Chef, Puppet, or AWS CloudFormation. Experience with Artifactory or similar artifact repository management. NICE product knowledge and/or implementation or support experience with NICE CXone or related platforms. Knowledge of ETL processes and data pipeline management. Call centre or telecoms industry experience. What's in it for you? Learn more about the Benefits at NICE NICE is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, age, sex, marital status, ancestry, neurotype, physical or mental disability, veteran status, gender identity, sexual orientation, or any other category protected by law.
Spectrum IT Recruitment
AWS Cloud Operations Engineer (Night Shift)
Spectrum IT Recruitment City, Manchester
AWS Cloud Operations Engineer (Night Shift) Night Shift (10pm - 6am) (Fully Remote) Large Government Project Applicants must be eligible for Security Clearance Looking for an opportunity to make a real impact on a major government programme? We're recruiting an experienced AWS Cloud Operations Engineer to join a large-scale project focused on onboarding a major public sector client and delivering highly secure customer contact solutions. This is a key role within a growing engineering team, helping to build, support and optimise cloud platforms that underpin critical services. You'll work across AWS cloud infrastructure, Linux environments, container platforms and databases, helping to ensure secure, scalable and highly available systems. What we're looking for: Strong AWS cloud experience (EKS, ECS, EC2, RDS, IAM, VPC) Linux systems administration expertise Containerisation and Kubernetes experience Terraform and Infrastructure as Code knowledge Database administration experience (PostgreSQL, MySQL, Aurora or similar) A passion for reliability, security and continuous improvement Why apply? Join a major long-term government project Work on secure, mission-critical technology Excellent opportunity to influence architecture and operational excellence Be part of a global technology organisation investing heavily in growth and innovation If you enjoy solving complex infrastructure challenges and want to play a key role in a high-profile programme, we'd love to hear from you. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
15/07/2026
Full time
AWS Cloud Operations Engineer (Night Shift) Night Shift (10pm - 6am) (Fully Remote) Large Government Project Applicants must be eligible for Security Clearance Looking for an opportunity to make a real impact on a major government programme? We're recruiting an experienced AWS Cloud Operations Engineer to join a large-scale project focused on onboarding a major public sector client and delivering highly secure customer contact solutions. This is a key role within a growing engineering team, helping to build, support and optimise cloud platforms that underpin critical services. You'll work across AWS cloud infrastructure, Linux environments, container platforms and databases, helping to ensure secure, scalable and highly available systems. What we're looking for: Strong AWS cloud experience (EKS, ECS, EC2, RDS, IAM, VPC) Linux systems administration expertise Containerisation and Kubernetes experience Terraform and Infrastructure as Code knowledge Database administration experience (PostgreSQL, MySQL, Aurora or similar) A passion for reliability, security and continuous improvement Why apply? Join a major long-term government project Work on secure, mission-critical technology Excellent opportunity to influence architecture and operational excellence Be part of a global technology organisation investing heavily in growth and innovation If you enjoy solving complex infrastructure challenges and want to play a key role in a high-profile programme, we'd love to hear from you. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
Spectrum IT Recruitment
AWS Infrastructure Engineer (Night Shift)
Spectrum IT Recruitment
AWS Cloud Operations Engineer (Night Shift) Full Remote UK Only UK Security Clearance Eligible Night Shift 10pm - 6am Looking to build your career in AWS Cloud and Infrastructure? Whether you're an early-career Cloud Engineer, Infrastructure Engineer or Operations Engineer looking for the next step, this is an opportunity to join a major government programme where you'll gain hands-on experience supporting secure, large-scale cloud platforms used for mission-critical services. You'll be joining an experienced engineering team who will support your development as you grow your technical skills across AWS, cloud infrastructure, automation and platform operations. What you'll be doing Supporting AWS cloud infrastructure and production platforms Monitoring, troubleshooting and resolving infrastructure issues Working with Linux systems, containers and cloud services Learning and developing Infrastructure as Code and automation skills Collaborating with experienced engineers to improve platform reliability, performance and security Helping to support highly available services used by public sector customers We're looking for You don't need to tick every box, but you'll ideally have experience with some of the following: Commercial experience supporting AWS environments (EC2, IAM, VPC, RDS, ECS, EKS or similar) Linux administration or infrastructure support experience An understanding of networking, cloud platforms or virtualisation Exposure to Docker, Kubernetes or container technologies Some knowledge of Terraform or Infrastructure as Code would be beneficial A genuine interest in cloud engineering, automation and continuous learning Important This role operates on a permanent night shift (10pm-6am) , so we're looking for someone who enjoys working these hours or sees the benefits they can offer. Applicants must also be eligible to obtain UK Security Clearance , meaning you'll need to meet UK government residency and eligibility requirements. Why join? Excellent opportunity to accelerate your AWS and Cloud career Learn from experienced Cloud and Platform Engineers Work on a major long-term government programme Gain exposure to modern AWS technologies and enterprise-scale infrastructure Join a global technology organisation investing heavily in cloud, AI and innovation Clear opportunities for career progression as the programme continues to grow If you've already started your cloud journey and are looking for an opportunity to develop your AWS skills on a high-profile project, we'd love to hear from you. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
07/07/2026
Full time
AWS Cloud Operations Engineer (Night Shift) Full Remote UK Only UK Security Clearance Eligible Night Shift 10pm - 6am Looking to build your career in AWS Cloud and Infrastructure? Whether you're an early-career Cloud Engineer, Infrastructure Engineer or Operations Engineer looking for the next step, this is an opportunity to join a major government programme where you'll gain hands-on experience supporting secure, large-scale cloud platforms used for mission-critical services. You'll be joining an experienced engineering team who will support your development as you grow your technical skills across AWS, cloud infrastructure, automation and platform operations. What you'll be doing Supporting AWS cloud infrastructure and production platforms Monitoring, troubleshooting and resolving infrastructure issues Working with Linux systems, containers and cloud services Learning and developing Infrastructure as Code and automation skills Collaborating with experienced engineers to improve platform reliability, performance and security Helping to support highly available services used by public sector customers We're looking for You don't need to tick every box, but you'll ideally have experience with some of the following: Commercial experience supporting AWS environments (EC2, IAM, VPC, RDS, ECS, EKS or similar) Linux administration or infrastructure support experience An understanding of networking, cloud platforms or virtualisation Exposure to Docker, Kubernetes or container technologies Some knowledge of Terraform or Infrastructure as Code would be beneficial A genuine interest in cloud engineering, automation and continuous learning Important This role operates on a permanent night shift (10pm-6am) , so we're looking for someone who enjoys working these hours or sees the benefits they can offer. Applicants must also be eligible to obtain UK Security Clearance , meaning you'll need to meet UK government residency and eligibility requirements. Why join? Excellent opportunity to accelerate your AWS and Cloud career Learn from experienced Cloud and Platform Engineers Work on a major long-term government programme Gain exposure to modern AWS technologies and enterprise-scale infrastructure Join a global technology organisation investing heavily in cloud, AI and innovation Clear opportunities for career progression as the programme continues to grow If you've already started your cloud journey and are looking for an opportunity to develop your AWS skills on a high-profile project, we'd love to hear from you. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
Spectrum IT Recruitment
Site Reliability Engineer (AWS)
Spectrum IT Recruitment City, Birmingham
Site Reliability Engineer (SRE) AWS Kubernetes Fully Remote (UK) 24/7 Shift Pattern (28-day rota including days & nights) Competitive + Bonus + Excellent Benefits Build resilient cloud platforms that support critical national services. We're recruiting Site Reliability Engineers to join a global leader in AI-powered customer experience and cloud technology. Following the award of a major government programme, they're expanding their engineering teams to build and support highly secure, cloud-native platforms that deliver sensitive communication services. This is an opportunity to join an organisation investing heavily in modern cloud engineering, automation and reliability. Working as part of a collaborative SRE team, you'll help ensure large-scale production environments remain secure, available and resilient, whilst continuously improving the way they're operated through automation and engineering best practice. If you enjoy solving production challenges, improving reliability and automating away operational toil, we'd love to hear from you. What you'll be doing Monitoring and maintaining highly available production platforms running in AWS Responding to and managing production incidents across a 24/7 service Investigating complex technical issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn't essential. Why join? This is far more than a traditional NOC role. You'll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you'll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services. If you're passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we'd love to hear from you. Apply today or contact Dave Carlisle at Spectrum IT Recruitment for a confidential discussion. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.
01/07/2026
Full time
Site Reliability Engineer (SRE) AWS Kubernetes Fully Remote (UK) 24/7 Shift Pattern (28-day rota including days & nights) Competitive + Bonus + Excellent Benefits Build resilient cloud platforms that support critical national services. We're recruiting Site Reliability Engineers to join a global leader in AI-powered customer experience and cloud technology. Following the award of a major government programme, they're expanding their engineering teams to build and support highly secure, cloud-native platforms that deliver sensitive communication services. This is an opportunity to join an organisation investing heavily in modern cloud engineering, automation and reliability. Working as part of a collaborative SRE team, you'll help ensure large-scale production environments remain secure, available and resilient, whilst continuously improving the way they're operated through automation and engineering best practice. If you enjoy solving production challenges, improving reliability and automating away operational toil, we'd love to hear from you. What you'll be doing Monitoring and maintaining highly available production platforms running in AWS Responding to and managing production incidents across a 24/7 service Investigating complex technical issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous service improvements Supporting containerised workloads using Kubernetes and Docker What we're looking for You'll ideally have experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement and operational excellence Experience with Infrastructure as Code (Terraform), SRE principles (SLIs, SLOs), or regulated environments would be beneficial but isn't essential. Why join? This is far more than a traditional NOC role. You'll be joining an engineering-led organisation where reliability, automation and continuous improvement sit at the heart of the platform. Rather than simply responding to incidents, you'll work to prevent them by improving systems, automating operational processes and helping shape the future of highly resilient cloud services. If you're passionate about building reliable cloud platforms and enjoy solving complex technical problems in large-scale production environments, we'd love to hear from you. Apply today or contact Dave Carlisle at Spectrum IT Recruitment for a confidential discussion. Spectrum IT Recruitment (South) Limited is acting as an Employment Agency in relation to this vacancy.

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board