it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Career Advice
  • Contact us
  • Employers
    • Register as Employer
    • Pricing Plans
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Career Advice
  • Contact us
  • Employers
    • Register as Employer
    • Pricing Plans
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

8 jobs found

Email me jobs like this
Refine Search
Current Search
infrastructure compute architect 12 month ftc
Media Engineer
慨正橡扯 Winchester, Hampshire
Engineer - Audio & Video Processing - 12 months FTC Location:Crawley Court (Winchester), Emley Moor (Huddersfield), Newman Street Chalfont -London We operate a flexible, hybrid working environment - requirement to travel to either our Winchester or London office up to twice a week Total Package 12 months FTC Up to £54,000 base salary Generous pension scheme starting at 6% rising to 10% 25 days annual leave Private Medical A unique wellbeing programme that looks after the whole you Access to multiple learning platforms to support your individual development Active and diverse networks that build community, support wellbeing and advocate for change A comprehensive set of benefits including discounts on big brands, gymflex memberships and paid volunteering leave - see our full list of benefits here. Purpose - 12 months FTC Responsible for the design, development, implementation, operation, and troubleshooting of cloud-based and traditional Media platforms, with a particular focus on VOD, Cloud Playout, workflow automation, metadata-driven services, and AWS-hosted media solutions. The role is accountable for supporting, configuring, and optimising media workflows across production and non-production environments, ensuring reliable content delivery, operational excellence, and the successful integration of new services, partners, and technologies. The successful candidate will combine strong Media Engineering expertise with a solid understanding of cloud infrastructure, software-enabled workflows, automation, and modern deployment practices, enabling them to work effectively alongside Software, Platform, and Product Engineering teams. Accountabilities Design, configure, test, and support Media workflows and solutions to achieve maximum performance, reliability, and availability across VOD and Cloud Playout platforms. Provide technical expertise for the deployment, operation, and optimisation of Media services, ensuring efficient workflow execution and continual service improvement. Design and drive processes to facilitate the onboarding of new services, partners, channels, and content delivery workflows. Develop, analyse, troubleshoot, and modify Media systems, workflows, and configurations, ensuring compliance with operational, regulatory, and industry standards. Support the integration and operation of third-party Media products and vendor platforms hosted within AWS environments. Configure and maintain metadata-driven workflows, XML-based configurations, subtitle services, content delivery processes, and media processing pipelines. Support deployment and promotion activities through Git-based configuration management and CI/CD workflows. Troubleshoot issues across distributed cloud services, media workflows, and partner integrations, applying a structured and methodical approach to root-cause analysis and resolution. Support engagement with suppliers, vendors, contractors, and customers to ensure service delivery requirements are met and issues are effectively managed through to resolution. Standardise workflow configurations and engineering practices to improve maintainability, minimise disruption, and simplify integration with existing platforms and infrastructure. Document Media engineering principles, workflows, operational procedures, and platform configurations in line with industry best practices. Apply engineering, security, compliance, and operational standards across all aspects of Media platform design and support. Apply engineering best practices to deliver solutions that are reliable, scalable, maintainable, and supportable throughout their lifecycle. Act as a Design Authority (DA) within Media Engineering to ensure successful delivery of end-to-end solutions and support projects as Lead Design Authority (LDA) where required. Work closely with Product, Platform, Software Engineering, Architecture, Operations, and Network teams to deploy, optimise, and support Media solutions. Support AWS-based Media services including Lambda, Step Functions, S3, EventBridge, IAM, and related cloud-native technologies. Assist in the integration and migration of legacy Media systems into modern cloud-based architectures and workflows. Contribute to automation initiatives, operational tooling, and workflow orchestration to improve efficiency and reduce manual intervention. Provide technical leadership and mentoring to junior engineers, promoting engineering excellence, knowledge sharing, and continuous improvement. Stay current with emerging Media technologies, cloud platforms, video transport protocols, and industry best practices. Identify opportunities to improve engineering processes, platform performance, service resilience, and operational efficiency. Skills Media Engineering- Strong understanding of VOD, broadcast, playout, OTT, content delivery, metadata management, subtitle workflows, scheduling, and media processing technologies. Cloud Platforms & Operations- Practical understanding of AWS services including Lambda, Step Functions, S3, EventBridge, IAM, and cloud-native operational support. Configuration & Workflow Management- Experience working with XML, metadata-driven workflows, service onboarding, workflow configuration, and operational change management. Software & Automation Awareness- Ability to read, understand, troubleshoot, and modify code and scripts used within Media platforms and cloud workflows. Git & Deployment Practices- Understanding of Git-based workflows, CI/CD pipelines, configuration management, and deployment processes across multiple environments. Problem Solving- Strong analytical and troubleshooting skills with the ability to diagnose issues across distributed systems, cloud infrastructure, vendor products, and Media workflows. Vendor & Integration Management- Experience integrating with external partners and vendors using technologies such as FTP, SFTP, Apple Transporter, and S3-based delivery workflows. Video Connectivity & Distribution- Awareness of cloud ingest and distribution workflows and modern IP video transport technologies including SRT, RIST, and Zixi. Technical Leadership- Ability to mentor engineers, act as a subject matter expert, and share best practices across Media platforms and operational processes. Cross-Functional Collaboration- Proven ability to work effectively with Product, Engineering, Platform, Operations, Architecture, and external supplier teams. Operational Ownership- Ability to make informed decisions in live and time-sensitive operational environments while balancing technical, service, and commercial considerations. Communication & Stakeholder Management- Strong written and verbal communication skills with the ability to engage effectively with technical and non-technical stakeholders. Knowledge & Experience Experienced in Media Engineering within VOD, OTT, Broadcast, Playout, or Cloud Media environments. Demonstrable experience supporting Media workflows, platform configuration, metadata management, content delivery, service onboarding, and operational ownership of Media services. Experience working with cloud-hosted Media platforms and services running within AWS environments. Practical experience using AWS services such as Lambda, Step Functions, S3, EventBridge, and IAM for deployment, troubleshooting, and operational support. Experience working with Git-based source control, configuration management, and CI/CD deployment processes. Ability to understand, troubleshoot, and make changes to software-driven workflows, automation, and cloud-native services. Experience working with vendor-supported Media platforms and third-party integrations. Knowledge of XML and structured configuration formats used in Media workflows. Experience supporting or integrating cloud playout platforms, scheduling systems, playlist-driven workflows, or channel operations is beneficial. Good Python knowledge for workflow logic, automation support, AWS Lambda troubleshooting, and operational scripting is highly desirable. Knowledge of Infrastructure as Code tooling such as Terraform or Terragrunt is advantageous. Experience reverse-engineering legacy workflows and supporting migration to cloud-based platforms would be advantageous. Exposure to incident management, operational escalation, and live service support is desirable. Qualifications: A degree or equivalent experience in Media Technology, Computer Science, Software Engineering, Broadcast Engineering, Electrical Engineering, or a related discipline is advantageous. Important: This is not a traditional Playout Operations or Broadcast Support role. Successful candidates will have experience working with cloud-based media platforms and be comfortable engaging with software engineering concepts, source control, deployment pipelines, configuration management, automation, and code-based workflows, even where software development is not their primary discipline.
27/07/2026
Full time
Engineer - Audio & Video Processing - 12 months FTC Location:Crawley Court (Winchester), Emley Moor (Huddersfield), Newman Street Chalfont -London We operate a flexible, hybrid working environment - requirement to travel to either our Winchester or London office up to twice a week Total Package 12 months FTC Up to £54,000 base salary Generous pension scheme starting at 6% rising to 10% 25 days annual leave Private Medical A unique wellbeing programme that looks after the whole you Access to multiple learning platforms to support your individual development Active and diverse networks that build community, support wellbeing and advocate for change A comprehensive set of benefits including discounts on big brands, gymflex memberships and paid volunteering leave - see our full list of benefits here. Purpose - 12 months FTC Responsible for the design, development, implementation, operation, and troubleshooting of cloud-based and traditional Media platforms, with a particular focus on VOD, Cloud Playout, workflow automation, metadata-driven services, and AWS-hosted media solutions. The role is accountable for supporting, configuring, and optimising media workflows across production and non-production environments, ensuring reliable content delivery, operational excellence, and the successful integration of new services, partners, and technologies. The successful candidate will combine strong Media Engineering expertise with a solid understanding of cloud infrastructure, software-enabled workflows, automation, and modern deployment practices, enabling them to work effectively alongside Software, Platform, and Product Engineering teams. Accountabilities Design, configure, test, and support Media workflows and solutions to achieve maximum performance, reliability, and availability across VOD and Cloud Playout platforms. Provide technical expertise for the deployment, operation, and optimisation of Media services, ensuring efficient workflow execution and continual service improvement. Design and drive processes to facilitate the onboarding of new services, partners, channels, and content delivery workflows. Develop, analyse, troubleshoot, and modify Media systems, workflows, and configurations, ensuring compliance with operational, regulatory, and industry standards. Support the integration and operation of third-party Media products and vendor platforms hosted within AWS environments. Configure and maintain metadata-driven workflows, XML-based configurations, subtitle services, content delivery processes, and media processing pipelines. Support deployment and promotion activities through Git-based configuration management and CI/CD workflows. Troubleshoot issues across distributed cloud services, media workflows, and partner integrations, applying a structured and methodical approach to root-cause analysis and resolution. Support engagement with suppliers, vendors, contractors, and customers to ensure service delivery requirements are met and issues are effectively managed through to resolution. Standardise workflow configurations and engineering practices to improve maintainability, minimise disruption, and simplify integration with existing platforms and infrastructure. Document Media engineering principles, workflows, operational procedures, and platform configurations in line with industry best practices. Apply engineering, security, compliance, and operational standards across all aspects of Media platform design and support. Apply engineering best practices to deliver solutions that are reliable, scalable, maintainable, and supportable throughout their lifecycle. Act as a Design Authority (DA) within Media Engineering to ensure successful delivery of end-to-end solutions and support projects as Lead Design Authority (LDA) where required. Work closely with Product, Platform, Software Engineering, Architecture, Operations, and Network teams to deploy, optimise, and support Media solutions. Support AWS-based Media services including Lambda, Step Functions, S3, EventBridge, IAM, and related cloud-native technologies. Assist in the integration and migration of legacy Media systems into modern cloud-based architectures and workflows. Contribute to automation initiatives, operational tooling, and workflow orchestration to improve efficiency and reduce manual intervention. Provide technical leadership and mentoring to junior engineers, promoting engineering excellence, knowledge sharing, and continuous improvement. Stay current with emerging Media technologies, cloud platforms, video transport protocols, and industry best practices. Identify opportunities to improve engineering processes, platform performance, service resilience, and operational efficiency. Skills Media Engineering- Strong understanding of VOD, broadcast, playout, OTT, content delivery, metadata management, subtitle workflows, scheduling, and media processing technologies. Cloud Platforms & Operations- Practical understanding of AWS services including Lambda, Step Functions, S3, EventBridge, IAM, and cloud-native operational support. Configuration & Workflow Management- Experience working with XML, metadata-driven workflows, service onboarding, workflow configuration, and operational change management. Software & Automation Awareness- Ability to read, understand, troubleshoot, and modify code and scripts used within Media platforms and cloud workflows. Git & Deployment Practices- Understanding of Git-based workflows, CI/CD pipelines, configuration management, and deployment processes across multiple environments. Problem Solving- Strong analytical and troubleshooting skills with the ability to diagnose issues across distributed systems, cloud infrastructure, vendor products, and Media workflows. Vendor & Integration Management- Experience integrating with external partners and vendors using technologies such as FTP, SFTP, Apple Transporter, and S3-based delivery workflows. Video Connectivity & Distribution- Awareness of cloud ingest and distribution workflows and modern IP video transport technologies including SRT, RIST, and Zixi. Technical Leadership- Ability to mentor engineers, act as a subject matter expert, and share best practices across Media platforms and operational processes. Cross-Functional Collaboration- Proven ability to work effectively with Product, Engineering, Platform, Operations, Architecture, and external supplier teams. Operational Ownership- Ability to make informed decisions in live and time-sensitive operational environments while balancing technical, service, and commercial considerations. Communication & Stakeholder Management- Strong written and verbal communication skills with the ability to engage effectively with technical and non-technical stakeholders. Knowledge & Experience Experienced in Media Engineering within VOD, OTT, Broadcast, Playout, or Cloud Media environments. Demonstrable experience supporting Media workflows, platform configuration, metadata management, content delivery, service onboarding, and operational ownership of Media services. Experience working with cloud-hosted Media platforms and services running within AWS environments. Practical experience using AWS services such as Lambda, Step Functions, S3, EventBridge, and IAM for deployment, troubleshooting, and operational support. Experience working with Git-based source control, configuration management, and CI/CD deployment processes. Ability to understand, troubleshoot, and make changes to software-driven workflows, automation, and cloud-native services. Experience working with vendor-supported Media platforms and third-party integrations. Knowledge of XML and structured configuration formats used in Media workflows. Experience supporting or integrating cloud playout platforms, scheduling systems, playlist-driven workflows, or channel operations is beneficial. Good Python knowledge for workflow logic, automation support, AWS Lambda troubleshooting, and operational scripting is highly desirable. Knowledge of Infrastructure as Code tooling such as Terraform or Terragrunt is advantageous. Experience reverse-engineering legacy workflows and supporting migration to cloud-based platforms would be advantageous. Exposure to incident management, operational escalation, and live service support is desirable. Qualifications: A degree or equivalent experience in Media Technology, Computer Science, Software Engineering, Broadcast Engineering, Electrical Engineering, or a related discipline is advantageous. Important: This is not a traditional Playout Operations or Broadcast Support role. Successful candidates will have experience working with cloud-based media platforms and be comfortable engaging with software engineering concepts, source control, deployment pipelines, configuration management, automation, and code-based workflows, even where software development is not their primary discipline.
Lucy Group
Infrastructure & Cloud Operations Engineer - 12 month FTC
Lucy Group Oxford, Oxfordshire
Job Purpose The Cloud Infrastructure Engineer is responsible for implementing, supporting, and optimising Lucy Group's public and private cloud platforms. Working across public and private cloud environments, the role delivers secure, resilient, and high-performing infrastructure services to support global operations. This includes managing deployments, monitoring performance, ensuring compliance with governance frameworks, and collaborating with colleagues to deliver enterprise grade cloud capabilities. Job Context Operating within Group IS, the Cloud Infrastructure Engineer works alongside architects, other engineers, and cross functional technology teams to deliver cloud infrastructure services. Reporting directly to Lead Cloud Architect, this role plays a key part in the execution of cloud strategies, technical implementations, and operational excellence initiatives. While the role does not have budgetary sign off, it is instrumental in ensuring that cloud platforms meet business needs and governance standards. Business Overview Lucy Group is an international group that makes the built environment sustainable. Our electric businesses advance the transition to a carbon free world with infrastructure that enables renewable energy and smart cities. Our real estate businesses support sustainable living through responsible property development and investment. Job Dimensions Operational responsibility for cloud infrastructure supporting 1,000+ global users. Implementation and support of public cloud and private cloud environments. Contribution to disaster recovery, performance optimisation, and capacity planning. Collaboration with network, security, and application teams to deliver integrated solutions. Key Accountabilities Deploy, configure, and maintain cloud services and infrastructure components in our public and private cloud environments. Monitor cloud environments to ensure performance, availability, and security compliance. Implement governance aligned configurations and policies in line with architectural guidance. Troubleshoot and resolve cloud infrastructure issues, escalating as necessary. Contribute to automation initiatives using Infrastructure as Code (IaC) and scripting. Support backup, disaster recovery, and business continuity processes. Work with vendors and service providers to resolve incidents and implement improvements. Maintain accurate documentation of configurations, procedures, and changes. Participate in project delivery, providing technical input and operational handover. Stay current with cloud technology trends and recommend improvements where appropriate. Expand automation for cloud deployments to increase efficiency and reduce manual effort. Introduce cloud native monitoring tools to enhance observability and proactive issue detection. Gain cross training in DevOps practices to improve collaboration with development teams. Qualifications, Experience & Skills Degree or equivalent experience in IT, Computer Science, or related discipline. 3+ years in a cloud engineering or infrastructure operations role. Hands on experience with Microsoft Azure, AWS and/or Nutanix. Knowledge of virtualisation, networking, and storage technologies in cloud and hybrid contexts. Experience with Infrastructure as Code (IaC) tools such as Terraform, Bicep, or ARM templates. Familiarity with cloud governance, compliance, and cost optimisation practices. ITIL Foundation or equivalent service management experience. Relevant cloud certifications (e.g., Microsoft Certified: Azure Administrator Associate) are desirable. Behavioural Competencies Technical Expertise - Maintains up to date knowledge of cloud infrastructure technologies. Problem Solving - Applies logical and structured approaches to diagnosing and resolving issues. Collaboration - Works effectively within cross functional teams. Learning Agility - Adapts quickly to new tools, processes, and technologies. Resilience - Maintains focus and performance during incident resolution. Attention to Detail - Ensures quality and accuracy in configurations and documentation.
25/07/2026
Full time
Job Purpose The Cloud Infrastructure Engineer is responsible for implementing, supporting, and optimising Lucy Group's public and private cloud platforms. Working across public and private cloud environments, the role delivers secure, resilient, and high-performing infrastructure services to support global operations. This includes managing deployments, monitoring performance, ensuring compliance with governance frameworks, and collaborating with colleagues to deliver enterprise grade cloud capabilities. Job Context Operating within Group IS, the Cloud Infrastructure Engineer works alongside architects, other engineers, and cross functional technology teams to deliver cloud infrastructure services. Reporting directly to Lead Cloud Architect, this role plays a key part in the execution of cloud strategies, technical implementations, and operational excellence initiatives. While the role does not have budgetary sign off, it is instrumental in ensuring that cloud platforms meet business needs and governance standards. Business Overview Lucy Group is an international group that makes the built environment sustainable. Our electric businesses advance the transition to a carbon free world with infrastructure that enables renewable energy and smart cities. Our real estate businesses support sustainable living through responsible property development and investment. Job Dimensions Operational responsibility for cloud infrastructure supporting 1,000+ global users. Implementation and support of public cloud and private cloud environments. Contribution to disaster recovery, performance optimisation, and capacity planning. Collaboration with network, security, and application teams to deliver integrated solutions. Key Accountabilities Deploy, configure, and maintain cloud services and infrastructure components in our public and private cloud environments. Monitor cloud environments to ensure performance, availability, and security compliance. Implement governance aligned configurations and policies in line with architectural guidance. Troubleshoot and resolve cloud infrastructure issues, escalating as necessary. Contribute to automation initiatives using Infrastructure as Code (IaC) and scripting. Support backup, disaster recovery, and business continuity processes. Work with vendors and service providers to resolve incidents and implement improvements. Maintain accurate documentation of configurations, procedures, and changes. Participate in project delivery, providing technical input and operational handover. Stay current with cloud technology trends and recommend improvements where appropriate. Expand automation for cloud deployments to increase efficiency and reduce manual effort. Introduce cloud native monitoring tools to enhance observability and proactive issue detection. Gain cross training in DevOps practices to improve collaboration with development teams. Qualifications, Experience & Skills Degree or equivalent experience in IT, Computer Science, or related discipline. 3+ years in a cloud engineering or infrastructure operations role. Hands on experience with Microsoft Azure, AWS and/or Nutanix. Knowledge of virtualisation, networking, and storage technologies in cloud and hybrid contexts. Experience with Infrastructure as Code (IaC) tools such as Terraform, Bicep, or ARM templates. Familiarity with cloud governance, compliance, and cost optimisation practices. ITIL Foundation or equivalent service management experience. Relevant cloud certifications (e.g., Microsoft Certified: Azure Administrator Associate) are desirable. Behavioural Competencies Technical Expertise - Maintains up to date knowledge of cloud infrastructure technologies. Problem Solving - Applies logical and structured approaches to diagnosing and resolving issues. Collaboration - Works effectively within cross functional teams. Learning Agility - Adapts quickly to new tools, processes, and technologies. Resilience - Maintains focus and performance during incident resolution. Attention to Detail - Ensures quality and accuracy in configurations and documentation.
Infrastructure & Cloud Operations Engineer - 12 month FTC
Lucy Group Head Office Oxford, Oxfordshire
Job Purpose The Cloud Infrastructure Engineer is responsible for implementing, supporting, and optimising Lucy Group's public and private cloud platforms. Working across public and private cloud environments, the role delivers secure, resilient, and high-performing infrastructure services to support global operations. This includes managing deployments, monitoring performance, ensuring compliance with governance frameworks, and collaborating with colleagues to deliver enterprise grade cloud capabilities. Job Context Operating within Group IS, the Cloud Infrastructure Engineer works alongside architects, other engineers, and cross functional technology teams to deliver cloud infrastructure services. Reporting directly to Lead Cloud Architect, this role plays a key part in the execution of cloud strategies, technical implementations, and operational excellence initiatives. While the role does not have budgetary sign off, it is instrumental in ensuring that cloud platforms meet business needs and governance standards. Business Overview Lucy Group is an international group that makes the built environment sustainable. Our electric businesses advance the transition to a carbon free world with infrastructure that enables renewable energy and smart cities. Our real estate businesses support sustainable living through responsible property development and investment. Job Dimensions Operational responsibility for cloud infrastructure supporting 1,000+ global users. Implementation and support of public cloud and private cloud environments. Contribution to disaster recovery, performance optimisation, and capacity planning. Collaboration with network, security, and application teams to deliver integrated solutions. Key Accountabilities Deploy, configure, and maintain cloud services and infrastructure components in our public and private cloud environments. Monitor cloud environments to ensure performance, availability, and security compliance. Implement governance aligned configurations and policies in line with architectural guidance. Troubleshoot and resolve cloud infrastructure issues, escalating as necessary. Contribute to automation initiatives using Infrastructure as Code (IaC) and scripting. Support backup, disaster recovery, and business continuity processes. Work with vendors and service providers to resolve incidents and implement improvements. Maintain accurate documentation of configurations, procedures, and changes. Participate in project delivery, providing technical input and operational handover. Stay current with cloud technology trends and recommend improvements where appropriate. Expand automation for cloud deployments to increase efficiency and reduce manual effort. Introduce cloud native monitoring tools to enhance observability and proactive issue detection. Gain cross training in DevOps practices to improve collaboration with development teams. Qualifications, Experience & Skills Degree or equivalent experience in IT, Computer Science, or related discipline. 3+ years in a cloud engineering or infrastructure operations role. Hands on experience with Microsoft Azure, AWS and/or Nutanix. Knowledge of virtualisation, networking, and storage technologies in cloud and hybrid contexts. Experience with Infrastructure as Code (IaC) tools such as Terraform, Bicep, or ARM templates. Familiarity with cloud governance, compliance, and cost optimisation practices. ITIL Foundation or equivalent service management experience. Relevant cloud certifications (e.g., Microsoft Certified: Azure Administrator Associate) are desirable. Behavioural Competencies Technical Expertise - Maintains up-to-date knowledge of cloud infrastructure technologies. Problem Solving - Applies logical and structured approaches to diagnosing and resolving issues. Collaboration - Works effectively within cross functional teams. Learning Agility - Adapts quickly to new tools, processes, and technologies. Resilience - Maintains focus and performance during incident resolution. Attention to Detail - Ensures quality and accuracy in configurations and documentation.
25/07/2026
Full time
Job Purpose The Cloud Infrastructure Engineer is responsible for implementing, supporting, and optimising Lucy Group's public and private cloud platforms. Working across public and private cloud environments, the role delivers secure, resilient, and high-performing infrastructure services to support global operations. This includes managing deployments, monitoring performance, ensuring compliance with governance frameworks, and collaborating with colleagues to deliver enterprise grade cloud capabilities. Job Context Operating within Group IS, the Cloud Infrastructure Engineer works alongside architects, other engineers, and cross functional technology teams to deliver cloud infrastructure services. Reporting directly to Lead Cloud Architect, this role plays a key part in the execution of cloud strategies, technical implementations, and operational excellence initiatives. While the role does not have budgetary sign off, it is instrumental in ensuring that cloud platforms meet business needs and governance standards. Business Overview Lucy Group is an international group that makes the built environment sustainable. Our electric businesses advance the transition to a carbon free world with infrastructure that enables renewable energy and smart cities. Our real estate businesses support sustainable living through responsible property development and investment. Job Dimensions Operational responsibility for cloud infrastructure supporting 1,000+ global users. Implementation and support of public cloud and private cloud environments. Contribution to disaster recovery, performance optimisation, and capacity planning. Collaboration with network, security, and application teams to deliver integrated solutions. Key Accountabilities Deploy, configure, and maintain cloud services and infrastructure components in our public and private cloud environments. Monitor cloud environments to ensure performance, availability, and security compliance. Implement governance aligned configurations and policies in line with architectural guidance. Troubleshoot and resolve cloud infrastructure issues, escalating as necessary. Contribute to automation initiatives using Infrastructure as Code (IaC) and scripting. Support backup, disaster recovery, and business continuity processes. Work with vendors and service providers to resolve incidents and implement improvements. Maintain accurate documentation of configurations, procedures, and changes. Participate in project delivery, providing technical input and operational handover. Stay current with cloud technology trends and recommend improvements where appropriate. Expand automation for cloud deployments to increase efficiency and reduce manual effort. Introduce cloud native monitoring tools to enhance observability and proactive issue detection. Gain cross training in DevOps practices to improve collaboration with development teams. Qualifications, Experience & Skills Degree or equivalent experience in IT, Computer Science, or related discipline. 3+ years in a cloud engineering or infrastructure operations role. Hands on experience with Microsoft Azure, AWS and/or Nutanix. Knowledge of virtualisation, networking, and storage technologies in cloud and hybrid contexts. Experience with Infrastructure as Code (IaC) tools such as Terraform, Bicep, or ARM templates. Familiarity with cloud governance, compliance, and cost optimisation practices. ITIL Foundation or equivalent service management experience. Relevant cloud certifications (e.g., Microsoft Certified: Azure Administrator Associate) are desirable. Behavioural Competencies Technical Expertise - Maintains up-to-date knowledge of cloud infrastructure technologies. Problem Solving - Applies logical and structured approaches to diagnosing and resolving issues. Collaboration - Works effectively within cross functional teams. Learning Agility - Adapts quickly to new tools, processes, and technologies. Resilience - Maintains focus and performance during incident resolution. Attention to Detail - Ensures quality and accuracy in configurations and documentation.
AND Digital
Observability Architect - 12 Month FTC
AND Digital
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend enhancements to tooling, processes, and architecture where required. Key Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements or platform uplifts where current tooling does not provide sufficient visibility or resilience. Baseline Performance Analysis Analyse telemetry, monitoring data, dashboards, and operational trends from completed migration waves. Establish performance baselines for compute, storage, networking, application response times, and transaction throughput. Identify recurring operational issues and use historical insights to improve migration readiness. Define measurable service health indicators to compare pre- and post-migration performance. Monolithic Application Monitoring Design comprehensive monitoring for the tightly-coupled monolithic application estate, with particular emphasis on latency-sensitive interdependencies. Create real-time dashboards that provide operational visibility across infrastructure, middleware, databases, messaging, and application components. Ensure end-to-end transaction tracing is available to rapidly identify bottlenecks and service degradation. Validate monitoring coverage prior to each migration wave. Logging & Trace Management Review and standardise centralised logging across all migrated environments. Ensure consistent log formats, metadata, correlation IDs, and traceability across systems. Validate log ingestion, retention policies, indexing, and search performance. Ensure operational teams can rapidly investigate incidents using correlated logs and distributed traces. Alerting & Operational Readiness Review and optimise alert thresholds to minimise both missed events and unnecessary alert noise. Implement intelligent alerting aligned to business services and critical customer journeys. Define migration-specific alerting for infrastructure failures, application degradation, latency increases, replication issues, and capacity constraints. Support operational readiness activities including rehearsals and production cutover monitoring. Compliance & Audit Ensure observability solutions meet financial services regulatory requirements for auditability, log retention, security, and data governance. Validate access controls and security monitoring for observability platforms. Support evidence gathering for internal governance, audit, and regulatory reviews. Platform Improvement Evaluate the suitability of existing observability platforms and recommend improvements where required. Assess opportunities to improve automation, anomaly detection, service health monitoring, and predictive alerting. Define standards and best practices for observability across future migration phases. Stakeholder Collaboration Work closely with Infrastructure Architects, Application Architects, Platform Engineering, Security, Operations, and Migration teams. Provide technical guidance during migration planning, testing, dress rehearsals, and production cutovers. Produce architecture documentation, monitoring standards, operational runbooks, and knowledge transfer materials. Required Skills & Experience Extensive experience designing enterprise observability solutions within large scale infrastructure or data centre migration programmes. Strong knowledge of metrics, logging, distributed tracing, and application performance monitoring (APM). Experience monitoring latency sensitive, business critical enterprise applications. Strong understanding of infrastructure, virtualisation, networking, storage, databases, and middleware monitoring. Experience implementing centralised logging and observability best practices. Knowledge of financial services operational resilience, audit, and regulatory requirements. Ability to analyse complex operational telemetry and identify performance bottlenecks. Excellent stakeholder management and communication skills. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
21/07/2026
Full time
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend enhancements to tooling, processes, and architecture where required. Key Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements or platform uplifts where current tooling does not provide sufficient visibility or resilience. Baseline Performance Analysis Analyse telemetry, monitoring data, dashboards, and operational trends from completed migration waves. Establish performance baselines for compute, storage, networking, application response times, and transaction throughput. Identify recurring operational issues and use historical insights to improve migration readiness. Define measurable service health indicators to compare pre- and post-migration performance. Monolithic Application Monitoring Design comprehensive monitoring for the tightly-coupled monolithic application estate, with particular emphasis on latency-sensitive interdependencies. Create real-time dashboards that provide operational visibility across infrastructure, middleware, databases, messaging, and application components. Ensure end-to-end transaction tracing is available to rapidly identify bottlenecks and service degradation. Validate monitoring coverage prior to each migration wave. Logging & Trace Management Review and standardise centralised logging across all migrated environments. Ensure consistent log formats, metadata, correlation IDs, and traceability across systems. Validate log ingestion, retention policies, indexing, and search performance. Ensure operational teams can rapidly investigate incidents using correlated logs and distributed traces. Alerting & Operational Readiness Review and optimise alert thresholds to minimise both missed events and unnecessary alert noise. Implement intelligent alerting aligned to business services and critical customer journeys. Define migration-specific alerting for infrastructure failures, application degradation, latency increases, replication issues, and capacity constraints. Support operational readiness activities including rehearsals and production cutover monitoring. Compliance & Audit Ensure observability solutions meet financial services regulatory requirements for auditability, log retention, security, and data governance. Validate access controls and security monitoring for observability platforms. Support evidence gathering for internal governance, audit, and regulatory reviews. Platform Improvement Evaluate the suitability of existing observability platforms and recommend improvements where required. Assess opportunities to improve automation, anomaly detection, service health monitoring, and predictive alerting. Define standards and best practices for observability across future migration phases. Stakeholder Collaboration Work closely with Infrastructure Architects, Application Architects, Platform Engineering, Security, Operations, and Migration teams. Provide technical guidance during migration planning, testing, dress rehearsals, and production cutovers. Produce architecture documentation, monitoring standards, operational runbooks, and knowledge transfer materials. Required Skills & Experience Extensive experience designing enterprise observability solutions within large scale infrastructure or data centre migration programmes. Strong knowledge of metrics, logging, distributed tracing, and application performance monitoring (APM). Experience monitoring latency sensitive, business critical enterprise applications. Strong understanding of infrastructure, virtualisation, networking, storage, databases, and middleware monitoring. Experience implementing centralised logging and observability best practices. Knowledge of financial services operational resilience, audit, and regulatory requirements. Ability to analyse complex operational telemetry and identify performance bottlenecks. Excellent stakeholder management and communication skills. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
AND Digital
Observability Architect - 12 Month FTC
AND Digital Bristol, Gloucestershire
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend enhancements to tooling, processes, and architecture where required. Key Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements or platform uplifts where current tooling does not provide sufficient visibility or resilience. Baseline Performance Analysis Analyse telemetry, monitoring data, dashboards, and operational trends from completed migration waves. Establish performance baselines for compute, storage, networking, application response times, and transaction throughput. Identify recurring operational issues and use historical insights to improve migration readiness. Define measurable service health indicators to compare pre- and post-migration performance. Monolithic Application Monitoring Design comprehensive monitoring for the tightly-coupled monolithic application estate, with particular emphasis on latency-sensitive interdependencies. Create real-time dashboards that provide operational visibility across infrastructure, middleware, databases, messaging, and application components. Ensure end-to-end transaction tracing is available to rapidly identify bottlenecks and service degradation. Validate monitoring coverage prior to each migration wave. Logging & Trace Management Review and standardise centralised logging across all migrated environments. Ensure consistent log formats, metadata, correlation IDs, and traceability across systems. Validate log ingestion, retention policies, indexing, and search performance. Ensure operational teams can rapidly investigate incidents using correlated logs and distributed traces. Alerting & Operational Readiness Review and optimise alert thresholds to minimise both missed events and unnecessary alert noise. Implement intelligent alerting aligned to business services and critical customer journeys. Define migration-specific alerting for infrastructure failures, application degradation, latency increases, replication issues, and capacity constraints. Support operational readiness activities including rehearsals and production cutover monitoring. Compliance & Audit Ensure observability solutions meet financial services regulatory requirements for auditability, log retention, security, and data governance. Validate access controls and security monitoring for observability platforms. Support evidence gathering for internal governance, audit, and regulatory reviews. Platform Improvement Evaluate the suitability of existing observability platforms and recommend improvements where required. Assess opportunities to improve automation, anomaly detection, service health monitoring, and predictive alerting. Define standards and best practices for observability across future migration phases. Stakeholder Collaboration Work closely with Infrastructure Architects, Application Architects, Platform Engineering, Security, Operations, and Migration teams. Provide technical guidance during migration planning, testing, dress rehearsals, and production cutovers. Produce architecture documentation, monitoring standards, operational runbooks, and knowledge transfer materials. Required Skills & Experience Extensive experience designing enterprise observability solutions within large scale infrastructure or data centre migration programmes. Strong knowledge of metrics, logging, distributed tracing, and application performance monitoring (APM). Experience monitoring latency sensitive, business critical enterprise applications. Strong understanding of infrastructure, virtualisation, networking, storage, databases, and middleware monitoring. Experience implementing centralised logging and observability best practices. Knowledge of financial services operational resilience, audit, and regulatory requirements. Ability to analyse complex operational telemetry and identify performance bottlenecks. Excellent stakeholder management and communication skills. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
21/07/2026
Full time
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend enhancements to tooling, processes, and architecture where required. Key Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements or platform uplifts where current tooling does not provide sufficient visibility or resilience. Baseline Performance Analysis Analyse telemetry, monitoring data, dashboards, and operational trends from completed migration waves. Establish performance baselines for compute, storage, networking, application response times, and transaction throughput. Identify recurring operational issues and use historical insights to improve migration readiness. Define measurable service health indicators to compare pre- and post-migration performance. Monolithic Application Monitoring Design comprehensive monitoring for the tightly-coupled monolithic application estate, with particular emphasis on latency-sensitive interdependencies. Create real-time dashboards that provide operational visibility across infrastructure, middleware, databases, messaging, and application components. Ensure end-to-end transaction tracing is available to rapidly identify bottlenecks and service degradation. Validate monitoring coverage prior to each migration wave. Logging & Trace Management Review and standardise centralised logging across all migrated environments. Ensure consistent log formats, metadata, correlation IDs, and traceability across systems. Validate log ingestion, retention policies, indexing, and search performance. Ensure operational teams can rapidly investigate incidents using correlated logs and distributed traces. Alerting & Operational Readiness Review and optimise alert thresholds to minimise both missed events and unnecessary alert noise. Implement intelligent alerting aligned to business services and critical customer journeys. Define migration-specific alerting for infrastructure failures, application degradation, latency increases, replication issues, and capacity constraints. Support operational readiness activities including rehearsals and production cutover monitoring. Compliance & Audit Ensure observability solutions meet financial services regulatory requirements for auditability, log retention, security, and data governance. Validate access controls and security monitoring for observability platforms. Support evidence gathering for internal governance, audit, and regulatory reviews. Platform Improvement Evaluate the suitability of existing observability platforms and recommend improvements where required. Assess opportunities to improve automation, anomaly detection, service health monitoring, and predictive alerting. Define standards and best practices for observability across future migration phases. Stakeholder Collaboration Work closely with Infrastructure Architects, Application Architects, Platform Engineering, Security, Operations, and Migration teams. Provide technical guidance during migration planning, testing, dress rehearsals, and production cutovers. Produce architecture documentation, monitoring standards, operational runbooks, and knowledge transfer materials. Required Skills & Experience Extensive experience designing enterprise observability solutions within large scale infrastructure or data centre migration programmes. Strong knowledge of metrics, logging, distributed tracing, and application performance monitoring (APM). Experience monitoring latency sensitive, business critical enterprise applications. Strong understanding of infrastructure, virtualisation, networking, storage, databases, and middleware monitoring. Experience implementing centralised logging and observability best practices. Knowledge of financial services operational resilience, audit, and regulatory requirements. Ability to analyse complex operational telemetry and identify performance bottlenecks. Excellent stakeholder management and communication skills. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
Infrastructure (Compute) Architect - 12 Month FTC
慨正橡扯 Bristol, Gloucestershire
We're on a mission to close the world's tech skills gap. We help organisations navigate the future of technology, combining human expertise, emerging tech and AI to deliver better outcomes, faster. Since 2014, we've worked side-by-side with clients to solve complex challenges, build high-performing teams, and create lasting capability. As technology continues to evolve, we believe the most successful organisations will be those that combine the best of both: human ingenuity AND intelligent technology. That belief is embedded in everything we do. We call it the genius of the AND: deep expertise AND practical delivery, innovation AND responsibility, ambitious work AND sustainable careers. Through our Guide, Build and Equip approach, we help organisations embrace change, deliver meaningful impact, and develop the skills they need to thrive in an increasingly agentic world. About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learningYou communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams What you'll bring to the table: Audit the existing colocation infrastructure, reviewing High-Level Designs (HLDs) and Low-Level Designs (LLDs) to ensure alignment with the deployed environment Validate compute and storage environments following build and non-production migrations, ensuring performance, resilience and operational standards are met Conduct capacity planning and resource analysis to ensure sufficient compute and storage headroom for upcoming tightly coupled monolithic migrations Identify infrastructure risks, performance bottlenecks and optimisation opportunities to improve efficiency and readiness for production workloads Ensure physical security, host isolation and infrastructure controls comply with financial services regulatory requirements before production deployment Collaborate with infrastructure, architecture and migration teams, providing technical guidance, documentation and assurance throughout the migration programme Why join AND Digital? We're building a culture where talented people can do meaningful work, continue growing, AND enjoy the journey along the way. Our model is designed to build belonging and connection: we're organised into regional 'Clubs' of no more than 80 people, so you (and our clients) get the benefits of that small company feel while also being connected to a larger whole. Our Practice Areas act as a 'second home' for our ANDis - a place where you can innovate and learn with like-minded experts who care deeply about their craft, support each other generously, and believe the best ideas come from collaboration, not hierarchy. Our culture is rooted in our values of Wonder, Share and Delight, and shows up through five traits that we celebrate and nurture in all our ANDis: we value curiosity, ambition and a growth-mindset. We are deeply client-centric. Most important of all we are human - we create an inclusive environment where people feel trusted, supported and able to be themselves. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre-existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
16/07/2026
Full time
We're on a mission to close the world's tech skills gap. We help organisations navigate the future of technology, combining human expertise, emerging tech and AI to deliver better outcomes, faster. Since 2014, we've worked side-by-side with clients to solve complex challenges, build high-performing teams, and create lasting capability. As technology continues to evolve, we believe the most successful organisations will be those that combine the best of both: human ingenuity AND intelligent technology. That belief is embedded in everything we do. We call it the genius of the AND: deep expertise AND practical delivery, innovation AND responsibility, ambitious work AND sustainable careers. Through our Guide, Build and Equip approach, we help organisations embrace change, deliver meaningful impact, and develop the skills they need to thrive in an increasingly agentic world. About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learningYou communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams What you'll bring to the table: Audit the existing colocation infrastructure, reviewing High-Level Designs (HLDs) and Low-Level Designs (LLDs) to ensure alignment with the deployed environment Validate compute and storage environments following build and non-production migrations, ensuring performance, resilience and operational standards are met Conduct capacity planning and resource analysis to ensure sufficient compute and storage headroom for upcoming tightly coupled monolithic migrations Identify infrastructure risks, performance bottlenecks and optimisation opportunities to improve efficiency and readiness for production workloads Ensure physical security, host isolation and infrastructure controls comply with financial services regulatory requirements before production deployment Collaborate with infrastructure, architecture and migration teams, providing technical guidance, documentation and assurance throughout the migration programme Why join AND Digital? We're building a culture where talented people can do meaningful work, continue growing, AND enjoy the journey along the way. Our model is designed to build belonging and connection: we're organised into regional 'Clubs' of no more than 80 people, so you (and our clients) get the benefits of that small company feel while also being connected to a larger whole. Our Practice Areas act as a 'second home' for our ANDis - a place where you can innovate and learn with like-minded experts who care deeply about their craft, support each other generously, and believe the best ideas come from collaboration, not hierarchy. Our culture is rooted in our values of Wonder, Share and Delight, and shows up through five traits that we celebrate and nurture in all our ANDis: we value curiosity, ambition and a growth-mindset. We are deeply client-centric. Most important of all we are human - we create an inclusive environment where people feel trusted, supported and able to be themselves. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre-existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
Observability Architect - 12 Month FTC
慨正橡扯 Bristol, Gloucestershire
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend enhancements to tooling, processes, and architecture where required. Key Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements or platform uplifts where current tooling does not provide sufficient visibility or resilience. Baseline Performance Analysis Analyse telemetry, monitoring data, dashboards, and operational trends from completed migration waves. Establish performance baselines for compute, storage, networking, application response times, and transaction throughput. Identify recurring operational issues and use historical insights to improve migration readiness. Define measurable service health indicators to compare pre- and post-migration performance. Monolithic Application Monitoring Design comprehensive monitoring for the tightly-coupled monolithic application estate, with particular emphasis on latency-sensitive interdependencies. Create real-time dashboards that provide operational visibility across infrastructure, middleware, databases, messaging, and application components. Ensure end-to-end transaction tracing is available to rapidly identify bottlenecks and service degradation. Validate monitoring coverage prior to each migration wave. Logging & Trace Management Review and standardise centralised logging across all migrated environments. Ensure consistent log formats, metadata, correlation IDs, and traceability across systems. Validate log ingestion, retention policies, indexing, and search performance. Ensure operational teams can rapidly investigate incidents using correlated logs and distributed traces. Alerting & Operational Readiness Review and optimise alert thresholds to minimise both missed events and unnecessary alert noise. Implement intelligent alerting aligned to business services and critical customer journeys. Define migration-specific alerting for infrastructure failures, application degradation, latency increases, replication issues, and capacity constraints. Support operational readiness activities including rehearsals and production cutover monitoring. Compliance & Audit Ensure observability solutions meet financial services regulatory requirements for auditability, log retention, security, and data governance. Validate access controls and security monitoring for observability platforms. Support evidence gathering for internal governance, audit, and regulatory reviews. Platform Improvement Evaluate the suitability of existing observability platforms and recommend improvements where required. Assess opportunities to improve automation, anomaly detection, service health monitoring, and predictive alerting. Define standards and best practices for observability across future migration phases. Stakeholder Collaboration Work closely with Infrastructure Architects, Application Architects, Platform Engineering, Security, Operations, and Migration teams. Provide technical guidance during migration planning, testing, dress rehearsals, and production cutovers. Produce architecture documentation, monitoring standards, operational runbooks, and knowledge transfer materials. Required Skills & Experience Extensive experience designing enterprise observability solutions within large scale infrastructure or data centre migration programmes. Strong knowledge of metrics, logging, distributed tracing, and application performance monitoring (APM). Experience monitoring latency sensitive, business critical enterprise applications. Strong understanding of infrastructure, virtualisation, networking, storage, databases, and middleware monitoring. Experience implementing centralised logging and observability best practices. Knowledge of financial services operational resilience, audit, and regulatory requirements. Ability to analyse complex operational telemetry and identify performance bottlenecks. Excellent stakeholder management and communication skills. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
16/07/2026
Full time
Observability Architect 12 Month FTC About you: You care deeply about producing high-quality work that delivers real value You're comfortable navigating ambiguity and solving complex problems collaboratively You bring strong expertise in your craft, alongside a willingness to keep learning You communicate clearly and build trust quickly with clients and teammates You're pragmatic, adaptable and outcome-focused You enjoy sharing knowledge and helping others grow You value low-ego collaboration and enjoy working as part of multidisciplinary teams Role Objective Lead the assessment, design, and optimisation of the observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend enhancements to tooling, processes, and architecture where required. Key Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements or platform uplifts where current tooling does not provide sufficient visibility or resilience. Baseline Performance Analysis Analyse telemetry, monitoring data, dashboards, and operational trends from completed migration waves. Establish performance baselines for compute, storage, networking, application response times, and transaction throughput. Identify recurring operational issues and use historical insights to improve migration readiness. Define measurable service health indicators to compare pre- and post-migration performance. Monolithic Application Monitoring Design comprehensive monitoring for the tightly-coupled monolithic application estate, with particular emphasis on latency-sensitive interdependencies. Create real-time dashboards that provide operational visibility across infrastructure, middleware, databases, messaging, and application components. Ensure end-to-end transaction tracing is available to rapidly identify bottlenecks and service degradation. Validate monitoring coverage prior to each migration wave. Logging & Trace Management Review and standardise centralised logging across all migrated environments. Ensure consistent log formats, metadata, correlation IDs, and traceability across systems. Validate log ingestion, retention policies, indexing, and search performance. Ensure operational teams can rapidly investigate incidents using correlated logs and distributed traces. Alerting & Operational Readiness Review and optimise alert thresholds to minimise both missed events and unnecessary alert noise. Implement intelligent alerting aligned to business services and critical customer journeys. Define migration-specific alerting for infrastructure failures, application degradation, latency increases, replication issues, and capacity constraints. Support operational readiness activities including rehearsals and production cutover monitoring. Compliance & Audit Ensure observability solutions meet financial services regulatory requirements for auditability, log retention, security, and data governance. Validate access controls and security monitoring for observability platforms. Support evidence gathering for internal governance, audit, and regulatory reviews. Platform Improvement Evaluate the suitability of existing observability platforms and recommend improvements where required. Assess opportunities to improve automation, anomaly detection, service health monitoring, and predictive alerting. Define standards and best practices for observability across future migration phases. Stakeholder Collaboration Work closely with Infrastructure Architects, Application Architects, Platform Engineering, Security, Operations, and Migration teams. Provide technical guidance during migration planning, testing, dress rehearsals, and production cutovers. Produce architecture documentation, monitoring standards, operational runbooks, and knowledge transfer materials. Required Skills & Experience Extensive experience designing enterprise observability solutions within large scale infrastructure or data centre migration programmes. Strong knowledge of metrics, logging, distributed tracing, and application performance monitoring (APM). Experience monitoring latency sensitive, business critical enterprise applications. Strong understanding of infrastructure, virtualisation, networking, storage, databases, and middleware monitoring. Experience implementing centralised logging and observability best practices. Knowledge of financial services operational resilience, audit, and regulatory requirements. Ability to analyse complex operational telemetry and identify performance bottlenecks. Excellent stakeholder management and communication skills. By joining AND, we'll provide: 25 days bookable holiday + flexible Bank Holidays Pension: 6% of salary paid by AND Digital with a further 2% paid by you (can be increased by choice). Aviva healthcare cover (including pre existing condition cover) for you. Flexibenefit: £1000 assigned to you via our benefits portal to select or upgrade the benefits that suit you the most. Any unused allowance from the £1000 can be taken as cash. Life Assurance. Income Protection. Eye test + first pair of glasses. Parental Benefits: Generous Enhanced Maternity and Enhanced Partner (Paternity) Leave. Equal Opportunities Statement Diversity and inclusion are hugely important to us, and we're committed to providing equal opportunities for all. We're actively recruiting for a diverse and inclusive workforce so want to ensure we do everything we can to support your application. We want you to feel safe and empowered to let us know if you need any adjustments to be made to your application or interview process, so please speak to our recruitment team.
Infrastructure & Cloud Operations Engineer - 12 month FTC
Lucy Electric Oxford, Oxfordshire
Job Purpose The Cloud Infrastructure Engineer is responsible for implementing, supporting, and optimising Lucy Group's public and private cloud platforms. Working across public and private cloud environments, the role delivers secure, resilient, and high-performing infrastructure services to support global operations. This includes managing deployments, monitoring performance, ensuring compliance with governance frameworks, and collaborating with colleagues to deliver enterprise grade cloud capabilities. Job Context Operating within Group IS, the Cloud Infrastructure Engineer works alongside architects, other engineers, and cross functional technology teams to deliver cloud infrastructure services. Reporting directly to Lead Cloud Architect, this role plays a key part in the execution of cloud strategies, technical implementations, and operational excellence initiatives. While the role does not have budgetary sign off, it is instrumental in ensuring that cloud platforms meet business needs and governance standards. Business Overview Lucy Group is an international group that makes the built environment sustainable. Our electric businesses advance the transition to a carbon free world with infrastructure that enables renewable energy and smart cities. Our real estate businesses support sustainable living through responsible property development and investment. Job Dimensions Operational responsibility for cloud infrastructure supporting 1,000+ global users. Implementation and support of public cloud and private cloud environments. Contribution to disaster recovery, performance optimisation, and capacity planning. Collaboration with network, security, and application teams to deliver integrated solutions. Key Accountabilities Deploy, configure, and maintain cloud services and infrastructure components in our public and private cloud environments. Monitor cloud environments to ensure performance, availability, and security compliance. Implement governance aligned configurations and policies in line with architectural guidance. Troubleshoot and resolve cloud infrastructure issues, escalating as necessary. Contribute to automation initiatives using Infrastructure as Code (IaC) and scripting. Support backup, disaster recovery, and business continuity processes. Work with vendors and service providers to resolve incidents and implement improvements. Maintain accurate documentation of configurations, procedures, and changes. Participate in project delivery, providing technical input and operational handover. Stay current with cloud technology trends and recommend improvements where appropriate. Expand automation for cloud deployments to increase efficiency and reduce manual effort. Introduce cloud native monitoring tools to enhance observability and proactive issue detection. Gain cross training in DevOps practices to improve collaboration with development teams. Qualifications, Experience & Skills Degree or equivalent experience in IT, Computer Science, or related discipline. 3+ years in a cloud engineering or infrastructure operations role. Hands on experience with Microsoft Azure, AWS and/or Nutanix. Knowledge of virtualisation, networking, and storage technologies in cloud and hybrid contexts. Experience with Infrastructure as Code (IaC) tools such as Terraform, Bicep, or ARM templates. Familiarity with cloud governance, compliance, and cost optimisation practices. ITIL Foundation or equivalent service management experience. Relevant cloud certifications (e.g., Microsoft Certified: Azure Administrator Associate) are desirable. Behavioural Competencies Technical Expertise - Maintains up-to-date knowledge of cloud infrastructure technologies. Problem Solving - Applies logical and structured approaches to diagnosing and resolving issues. Collaboration - Works effectively within cross functional teams. Learning Agility - Adapts quickly to new tools, processes, and technologies. Resilience - Maintains focus and performance during incident resolution. Attention to Detail - Ensures quality and accuracy in configurations and documentation.
06/07/2026
Full time
Job Purpose The Cloud Infrastructure Engineer is responsible for implementing, supporting, and optimising Lucy Group's public and private cloud platforms. Working across public and private cloud environments, the role delivers secure, resilient, and high-performing infrastructure services to support global operations. This includes managing deployments, monitoring performance, ensuring compliance with governance frameworks, and collaborating with colleagues to deliver enterprise grade cloud capabilities. Job Context Operating within Group IS, the Cloud Infrastructure Engineer works alongside architects, other engineers, and cross functional technology teams to deliver cloud infrastructure services. Reporting directly to Lead Cloud Architect, this role plays a key part in the execution of cloud strategies, technical implementations, and operational excellence initiatives. While the role does not have budgetary sign off, it is instrumental in ensuring that cloud platforms meet business needs and governance standards. Business Overview Lucy Group is an international group that makes the built environment sustainable. Our electric businesses advance the transition to a carbon free world with infrastructure that enables renewable energy and smart cities. Our real estate businesses support sustainable living through responsible property development and investment. Job Dimensions Operational responsibility for cloud infrastructure supporting 1,000+ global users. Implementation and support of public cloud and private cloud environments. Contribution to disaster recovery, performance optimisation, and capacity planning. Collaboration with network, security, and application teams to deliver integrated solutions. Key Accountabilities Deploy, configure, and maintain cloud services and infrastructure components in our public and private cloud environments. Monitor cloud environments to ensure performance, availability, and security compliance. Implement governance aligned configurations and policies in line with architectural guidance. Troubleshoot and resolve cloud infrastructure issues, escalating as necessary. Contribute to automation initiatives using Infrastructure as Code (IaC) and scripting. Support backup, disaster recovery, and business continuity processes. Work with vendors and service providers to resolve incidents and implement improvements. Maintain accurate documentation of configurations, procedures, and changes. Participate in project delivery, providing technical input and operational handover. Stay current with cloud technology trends and recommend improvements where appropriate. Expand automation for cloud deployments to increase efficiency and reduce manual effort. Introduce cloud native monitoring tools to enhance observability and proactive issue detection. Gain cross training in DevOps practices to improve collaboration with development teams. Qualifications, Experience & Skills Degree or equivalent experience in IT, Computer Science, or related discipline. 3+ years in a cloud engineering or infrastructure operations role. Hands on experience with Microsoft Azure, AWS and/or Nutanix. Knowledge of virtualisation, networking, and storage technologies in cloud and hybrid contexts. Experience with Infrastructure as Code (IaC) tools such as Terraform, Bicep, or ARM templates. Familiarity with cloud governance, compliance, and cost optimisation practices. ITIL Foundation or equivalent service management experience. Relevant cloud certifications (e.g., Microsoft Certified: Azure Administrator Associate) are desirable. Behavioural Competencies Technical Expertise - Maintains up-to-date knowledge of cloud infrastructure technologies. Problem Solving - Applies logical and structured approaches to diagnosing and resolving issues. Collaboration - Works effectively within cross functional teams. Learning Agility - Adapts quickly to new tools, processes, and technologies. Resilience - Maintains focus and performance during incident resolution. Attention to Detail - Ensures quality and accuracy in configurations and documentation.

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board