Harmonic Security lets teams adopt AI tools safely by protecting sensitive data in real time with minimal effort. It gives enterprises full control and stops leaks so that their teams can innovate confidently. We are led by cybersecurity experts and backed by top investors including N47, Ten Eleven Ventures, and In-Q-Tel. We have achieved early traction and product-market fit with a world-class team, and we are now focused on building a category-defining company. This is your opportunity to join us early and shape not just a product, but a category. How We Work: AI-First by Design Harmonic exists to help enterprises adopt AI safely and at scale. We hold ourselves to that same standard. Everyone here actively leverages AI tools to perform their best work - from deep research and writing to building robust processes and automating complex workflows. We expect every new hire to bring a genuine curiosity for AI and a commitment to using it to work smarter, faster, and with greater creativity. For some, this involves tinkering and remaining open to emerging tools; for others, it means architecting entirely new systems with AI at the core. We will be transparent about expectations for every role and provide the tools and support needed for you to thrive. About the Team Our Product Delivery team is the engine that turns vision into impact. We ship early and often, getting valuable features into the hands of customers quickly and iterating from there. We work in the open by default, sharing progress and ideas, and we trust each other to own outcomes. We're a small but mighty crew where every person plays a critical role and we're committed to using AI to work smarter and faster. About the Role We're looking for a Backend Engineer who has an appreciation for all things security and system interconnectivity. This is a very hands-on technical role where you will be responsible for designing, building and maintaining scalable integrations between our core applications, identity providers and data security platforms. What You Will Do Own deliveries and features from start to finish Design, build, and own backend services in a microservices-based, event-driven architecture Implement core domain logic that drives key product workflows and system behavior Ingest and process high-volume usage and event data from customer-facing endpoints Build backend services that support operational, analytics, and reporting use cases Model, persist, and manage domain data, enabling querying, filtering, and lifecycle management Expose well-designed APIs and webhooks consumed by customers, UIs, and external systems Power customer-facing UI experiences by providing backend APIs for visibility, analytics, and configuration What You Bring 3-5+ years of professional software engineering experience building backend services and distributed systems in the Java ecosystem (Java, Spring Boot, Maven/Gradle) Strong understanding of event-driven architectures, asynchronous processing patterns, messaging systems, and distributed system design Proven experience designing and maintaining RESTful APIs, web hooks, and service-to-service integrations Comfortable operating in AWS, GCP, or Azure environments and working with modern observability tooling, logging, and monitoring platforms A track record of delivering performant, reliable and scalable applications Excellent collaboration and communication skills in cross-functional teams You Might Be a Fit if You Have experience building products in cybersecurity, identity, governance, compliance, data protection, or enterprise SaaS environments You understand not just how to build systems, but why architectural decisions matter You can balance scalability, reliability, maintainability, and speed of execution You think critically about security, observability, and operational excellence from day one Love the idea of blending software development, distributed systems and data-intensive applications Strong familiarity with authentication and identity technologies such as OAuth, OpenID Connect (OIDC), SAML, SCIM, Entra ID (Azure AD), Okta, or other enterprise identity providers Thrive in fast-paced startup environments where ambiguity is the norm Enjoy shaping culture and engineering practices, not just writing code Leverage AI tools as an engineer to help you build smarter, faster and better Why Join Us This isn't just a job; it's an opportunity to be part of a team that is redefining cybersecurity. We believe today's talent is tomorrow's success, and we're committed to creating an environment where you can do the best work of your life. Competitive pay and meaningful equity with a direct stake in Harmonic's success Comprehensive benefits, pension plan, generous PTO, and flexible hybrid work A small, passionate team that values transparency, creativity, and learning Thoughtful leadership that cares deeply about growth, impact, and people Annual global off-sites (past trips include Lisbon and Nashville) The chance to directly shape both our product and our culture as we build a category-defining company Harmonic's Core Values Flourish in the Unknown We embrace new, unfamiliar situations that require initiative and rapid decision-making. We orient ourselves quickly and deliver results with minimal guidance. Never Full We raise our hands, take on challenges, and assist others whenever possible. We hunger for opportunities to learn and do more. Perfect Harmony We support one another to create cohesion and unity. We collaborate openly, share feedback honestly, and help everyone produce their best work
25/07/2026
Full time
Harmonic Security lets teams adopt AI tools safely by protecting sensitive data in real time with minimal effort. It gives enterprises full control and stops leaks so that their teams can innovate confidently. We are led by cybersecurity experts and backed by top investors including N47, Ten Eleven Ventures, and In-Q-Tel. We have achieved early traction and product-market fit with a world-class team, and we are now focused on building a category-defining company. This is your opportunity to join us early and shape not just a product, but a category. How We Work: AI-First by Design Harmonic exists to help enterprises adopt AI safely and at scale. We hold ourselves to that same standard. Everyone here actively leverages AI tools to perform their best work - from deep research and writing to building robust processes and automating complex workflows. We expect every new hire to bring a genuine curiosity for AI and a commitment to using it to work smarter, faster, and with greater creativity. For some, this involves tinkering and remaining open to emerging tools; for others, it means architecting entirely new systems with AI at the core. We will be transparent about expectations for every role and provide the tools and support needed for you to thrive. About the Team Our Product Delivery team is the engine that turns vision into impact. We ship early and often, getting valuable features into the hands of customers quickly and iterating from there. We work in the open by default, sharing progress and ideas, and we trust each other to own outcomes. We're a small but mighty crew where every person plays a critical role and we're committed to using AI to work smarter and faster. About the Role We're looking for a Backend Engineer who has an appreciation for all things security and system interconnectivity. This is a very hands-on technical role where you will be responsible for designing, building and maintaining scalable integrations between our core applications, identity providers and data security platforms. What You Will Do Own deliveries and features from start to finish Design, build, and own backend services in a microservices-based, event-driven architecture Implement core domain logic that drives key product workflows and system behavior Ingest and process high-volume usage and event data from customer-facing endpoints Build backend services that support operational, analytics, and reporting use cases Model, persist, and manage domain data, enabling querying, filtering, and lifecycle management Expose well-designed APIs and webhooks consumed by customers, UIs, and external systems Power customer-facing UI experiences by providing backend APIs for visibility, analytics, and configuration What You Bring 3-5+ years of professional software engineering experience building backend services and distributed systems in the Java ecosystem (Java, Spring Boot, Maven/Gradle) Strong understanding of event-driven architectures, asynchronous processing patterns, messaging systems, and distributed system design Proven experience designing and maintaining RESTful APIs, web hooks, and service-to-service integrations Comfortable operating in AWS, GCP, or Azure environments and working with modern observability tooling, logging, and monitoring platforms A track record of delivering performant, reliable and scalable applications Excellent collaboration and communication skills in cross-functional teams You Might Be a Fit if You Have experience building products in cybersecurity, identity, governance, compliance, data protection, or enterprise SaaS environments You understand not just how to build systems, but why architectural decisions matter You can balance scalability, reliability, maintainability, and speed of execution You think critically about security, observability, and operational excellence from day one Love the idea of blending software development, distributed systems and data-intensive applications Strong familiarity with authentication and identity technologies such as OAuth, OpenID Connect (OIDC), SAML, SCIM, Entra ID (Azure AD), Okta, or other enterprise identity providers Thrive in fast-paced startup environments where ambiguity is the norm Enjoy shaping culture and engineering practices, not just writing code Leverage AI tools as an engineer to help you build smarter, faster and better Why Join Us This isn't just a job; it's an opportunity to be part of a team that is redefining cybersecurity. We believe today's talent is tomorrow's success, and we're committed to creating an environment where you can do the best work of your life. Competitive pay and meaningful equity with a direct stake in Harmonic's success Comprehensive benefits, pension plan, generous PTO, and flexible hybrid work A small, passionate team that values transparency, creativity, and learning Thoughtful leadership that cares deeply about growth, impact, and people Annual global off-sites (past trips include Lisbon and Nashville) The chance to directly shape both our product and our culture as we build a category-defining company Harmonic's Core Values Flourish in the Unknown We embrace new, unfamiliar situations that require initiative and rapid decision-making. We orient ourselves quickly and deliver results with minimal guidance. Never Full We raise our hands, take on challenges, and assist others whenever possible. We hunger for opportunities to learn and do more. Perfect Harmony We support one another to create cohesion and unity. We collaborate openly, share feedback honestly, and help everyone produce their best work
Location: Slough (Hybrid - 2 days office, 3 days home) We are working with a fast-growing B2B SaaS business operating within the parts and inventory management sector, serving major names in food manufacturing and automotive engineering. The business operates in a collaborative, hands on environment, with a strong focus on reliability, security and scaling a proven, enterprise ready platform. This is a genuinely hands on leadership role covering architecture, development, integrations, security and operations. You'll work alongside two developers, one focused on full stack delivery and one focused on AI and machine learning features, while still writing code yourself, owning the Azure DevOps and AKS infrastructure, and setting the technical direction as the business works towards ISO 27001 certification. The role is based in Slough, with two days a week in the office and the rest working from home. This is a great opportunity for a hands on technical leader looking for genuine ownership and a clear route to CTO as the team and business grow. It would suit an IT Manager with a software development background, or a lead engineer with strong front end skills and coaching experience who is ready to step up into architecture and people leadership. Responsibilities of a Head of Engineering Own and evolve the platform architecture across a multi-tenant SaaS environment, balancing day to day reliability with long term scalability Stay hands on across the full stack, contributing to front end, back end, APIs and infrastructure Lead, mentor and grow a small team of senior developers, setting engineering standards and building a culture of ownership Manage the Azure DevOps and AKS estate, including CI/CD pipelines, environment management and cost optimisation Lead the design and delivery of integrations with customer systems such as CMMS, ERP and SharePoint Drive security posture across infrastructure, code and process, supporting ISO 27001 certification Oversee day to day platform operations, including uptime, incident response and performance Skills & Qualifications of a Head of Engineering Strong hands on full stack development experience across .NET, JavaScript and React Deep expertise in Azure DevOps and AKS, including CI/CD pipelines and production Kubernetes operations Experience owning the architecture of a production SaaS platform, ideally multi tenant A background in DevOps and security, with exposure to ISO 27001 or similar compliance frameworks Experience leading or mentoring developers; exposure to supply chain, procurement or inventory systems is a plus The benefits of Head of Engineering £90,000 - £110,000 salary, dependent on experience Hybrid working - two days a week from the Slough office, the rest from home A genuine pathway to CTO as the business and team continue to grow Direct influence over technical strategy, working closely with the founder Company performance related financial incentives Join a business with proven traction, enterprise customers and strong product market fit A small, low bureaucracy team where decisions are made quickly If you feel this Head of Engineering role is right for you, please contact Becky Prince or Emma Devereux at Maintech Recruitment for more information or click apply. Maintech Recruitment - Engineering Great Careers Maintech Recruitment are an equal opportunities agency and welcome applications from all suitably qualified persons regardless of sex, religion, belief, political opinion, race, age, sexual orientation, marital status or disability. Please note by applying for this role your data will be processed and stored in line with our privacy policy, full details of which are held on our website, and a copy can be provided if you wish.
24/07/2026
Full time
Location: Slough (Hybrid - 2 days office, 3 days home) We are working with a fast-growing B2B SaaS business operating within the parts and inventory management sector, serving major names in food manufacturing and automotive engineering. The business operates in a collaborative, hands on environment, with a strong focus on reliability, security and scaling a proven, enterprise ready platform. This is a genuinely hands on leadership role covering architecture, development, integrations, security and operations. You'll work alongside two developers, one focused on full stack delivery and one focused on AI and machine learning features, while still writing code yourself, owning the Azure DevOps and AKS infrastructure, and setting the technical direction as the business works towards ISO 27001 certification. The role is based in Slough, with two days a week in the office and the rest working from home. This is a great opportunity for a hands on technical leader looking for genuine ownership and a clear route to CTO as the team and business grow. It would suit an IT Manager with a software development background, or a lead engineer with strong front end skills and coaching experience who is ready to step up into architecture and people leadership. Responsibilities of a Head of Engineering Own and evolve the platform architecture across a multi-tenant SaaS environment, balancing day to day reliability with long term scalability Stay hands on across the full stack, contributing to front end, back end, APIs and infrastructure Lead, mentor and grow a small team of senior developers, setting engineering standards and building a culture of ownership Manage the Azure DevOps and AKS estate, including CI/CD pipelines, environment management and cost optimisation Lead the design and delivery of integrations with customer systems such as CMMS, ERP and SharePoint Drive security posture across infrastructure, code and process, supporting ISO 27001 certification Oversee day to day platform operations, including uptime, incident response and performance Skills & Qualifications of a Head of Engineering Strong hands on full stack development experience across .NET, JavaScript and React Deep expertise in Azure DevOps and AKS, including CI/CD pipelines and production Kubernetes operations Experience owning the architecture of a production SaaS platform, ideally multi tenant A background in DevOps and security, with exposure to ISO 27001 or similar compliance frameworks Experience leading or mentoring developers; exposure to supply chain, procurement or inventory systems is a plus The benefits of Head of Engineering £90,000 - £110,000 salary, dependent on experience Hybrid working - two days a week from the Slough office, the rest from home A genuine pathway to CTO as the business and team continue to grow Direct influence over technical strategy, working closely with the founder Company performance related financial incentives Join a business with proven traction, enterprise customers and strong product market fit A small, low bureaucracy team where decisions are made quickly If you feel this Head of Engineering role is right for you, please contact Becky Prince or Emma Devereux at Maintech Recruitment for more information or click apply. Maintech Recruitment - Engineering Great Careers Maintech Recruitment are an equal opportunities agency and welcome applications from all suitably qualified persons regardless of sex, religion, belief, political opinion, race, age, sexual orientation, marital status or disability. Please note by applying for this role your data will be processed and stored in line with our privacy policy, full details of which are held on our website, and a copy can be provided if you wish.
Location; Slough (Hybrid 2 days office, 3 days home) We are working with a fast-growing B2B SaaS business operating within the parts and inventory management sector, serving major names in food manufacturing and automotive engineering. The business operates in a collaborative, hands-on environment, with a strong focus on reliability, security and scaling a proven, enterprise-ready platform. This is a genuinely hands-on leadership role covering architecture, development, integrations, security and operations. You'll work alongside two developers, one focused on full stack delivery and one focused on AI and machine learning features, while still writing code yourself, owning the Azure DevOps and AKS infrastructure, and setting the technical direction as the business works towards ISO 27001 certification. The role is based in Slough, with two days a week in the office and the rest working from home. This is a great opportunity for a hands-on technical leader looking for genuine ownership and a clear route to CTO as the team and business grow. It would suit an IT Manager with a software development background, or a lead engineer with strong front-end skills and coaching experience who is ready to step up into architecture and people leadership. Responsibilities of a Head of Engineering: Own and evolve the platform architecture across a multi-tenant SaaS environment, balancing day-to-day reliability with long-term scalability Stay hands-on across the full stack, contributing to front end, back end, APIs and infrastructure Lead, mentor and grow a small team of senior developers, setting engineering standards and building a culture of ownership Manage the Azure DevOps and AKS estate, including CI/CD pipelines, environment management and cost optimisation Lead the design and delivery of integrations with customer systems such as CMMS, ERP and SharePoint Drive security posture across infrastructure, code and process, supporting ISO 27001 certification Oversee day-to-day platform operations, including uptime, incident response and performance Skills & Qualifications of a Head of Engineering: Strong hands-on full stack development experience across .NET, JavaScript and React Deep expertise in Azure DevOps and AKS, including CI/CD pipelines and production Kubernetes operations Experience owning the architecture of a production SaaS platform, ideally multi-tenant A background in DevOps and security, with exposure to ISO 27001 or similar compliance frameworks Experience leading or mentoring developers; exposure to supply chain, procurement or inventory systems is a plus The benefits of Head of Engineering; £90,000 - £110,000 salary, dependent on experience Hybrid working two days a week from the Slough office, the rest from home A genuine pathway to CTO as the business and team continue to grow Direct influence over technical strategy, working closely with the founder Company performance-related financial incentives Join a business with proven traction, enterprise customers and strong product-market fit A small, low-bureaucracy team where decisions are made quickly If you feel this Head of Engineering role is right for you, please contact Becky Prince or Emma Devereux at Maintech Recruitment for more information or click apply. Maintech Recruitment Engineering Great Careers Maintech Recruitment are an equal opportunities agency and welcome applications from all suitably qualified persons regardless of sex, religion, belief, political opinion, race, age, sexual orientation, marital status or disability. Please note by applying for this role your data will be processed and stored in line with our privacy policy, full details of which are held on our website, and a copy can be provided if you wish.
22/07/2026
Full time
Location; Slough (Hybrid 2 days office, 3 days home) We are working with a fast-growing B2B SaaS business operating within the parts and inventory management sector, serving major names in food manufacturing and automotive engineering. The business operates in a collaborative, hands-on environment, with a strong focus on reliability, security and scaling a proven, enterprise-ready platform. This is a genuinely hands-on leadership role covering architecture, development, integrations, security and operations. You'll work alongside two developers, one focused on full stack delivery and one focused on AI and machine learning features, while still writing code yourself, owning the Azure DevOps and AKS infrastructure, and setting the technical direction as the business works towards ISO 27001 certification. The role is based in Slough, with two days a week in the office and the rest working from home. This is a great opportunity for a hands-on technical leader looking for genuine ownership and a clear route to CTO as the team and business grow. It would suit an IT Manager with a software development background, or a lead engineer with strong front-end skills and coaching experience who is ready to step up into architecture and people leadership. Responsibilities of a Head of Engineering: Own and evolve the platform architecture across a multi-tenant SaaS environment, balancing day-to-day reliability with long-term scalability Stay hands-on across the full stack, contributing to front end, back end, APIs and infrastructure Lead, mentor and grow a small team of senior developers, setting engineering standards and building a culture of ownership Manage the Azure DevOps and AKS estate, including CI/CD pipelines, environment management and cost optimisation Lead the design and delivery of integrations with customer systems such as CMMS, ERP and SharePoint Drive security posture across infrastructure, code and process, supporting ISO 27001 certification Oversee day-to-day platform operations, including uptime, incident response and performance Skills & Qualifications of a Head of Engineering: Strong hands-on full stack development experience across .NET, JavaScript and React Deep expertise in Azure DevOps and AKS, including CI/CD pipelines and production Kubernetes operations Experience owning the architecture of a production SaaS platform, ideally multi-tenant A background in DevOps and security, with exposure to ISO 27001 or similar compliance frameworks Experience leading or mentoring developers; exposure to supply chain, procurement or inventory systems is a plus The benefits of Head of Engineering; £90,000 - £110,000 salary, dependent on experience Hybrid working two days a week from the Slough office, the rest from home A genuine pathway to CTO as the business and team continue to grow Direct influence over technical strategy, working closely with the founder Company performance-related financial incentives Join a business with proven traction, enterprise customers and strong product-market fit A small, low-bureaucracy team where decisions are made quickly If you feel this Head of Engineering role is right for you, please contact Becky Prince or Emma Devereux at Maintech Recruitment for more information or click apply. Maintech Recruitment Engineering Great Careers Maintech Recruitment are an equal opportunities agency and welcome applications from all suitably qualified persons regardless of sex, religion, belief, political opinion, race, age, sexual orientation, marital status or disability. Please note by applying for this role your data will be processed and stored in line with our privacy policy, full details of which are held on our website, and a copy can be provided if you wish.
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
21/07/2026
Full time
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Your profile Are you a senior DevOps and infrastructure engineer who enjoys building reliable cloud platforms, enabling engineering teams, and using AI to remove operational complexity? At Cenosco, we buildAsset Integrity Management software trusted by global leaders in Oil & Gas, Chemicals, and Energy. We are looking for someone who can strengthen our DevOps foundation, evolve our infrastructure, and help us use AI-powered tools to automate how we build, deploy, monitor, and operate our platform. This role is primarily focused onDevOps, infrastructure enablement, IaC, deployment automation, and platform reliability. You will also play an important part in shaping how we use AI and agentic tooling (i.e. Copilot, Cursor, Codex, and similar tools) to make DevOps work faster, smarter, and more scalable. What You'll Do DevOps, Infrastructure & Platform Enablement Design, build, andmaintainscalable cloud infrastructure on Azure, with a strong focus on reliability, security, performance, and operational simplicity. Own and improveInfrastructure as Code practices usingTerraform, Pulumi, or similar tools, creating reusable patterns that help teams provision and manage environments consistently. Build,optimize, and automate CI/CD pipelines for application, infrastructure, and platform deployments across environments. Enable engineering teams by creating self-service infrastructure, deployment templates, automation scripts, documentation, and platform standards. Drive containerization and orchestration practices usingKubernetesand Talos, ensuring our platform can scale reliably and efficiently. Embed security, compliance, monitoring, and disaster recovery considerations into infrastructure and deployment workflows from the start. AI-Assisted DevOps Automation Use AI-assisted and agentic tools to automate repetitive DevOps work, including pipeline creation,IaCgeneration, environment provisioning, deployment workflows, and operational runbooks. Experiment with tools such asGitHub Copilot,Cursor,Codex, and similar solutions to improve engineering productivity and reduce manual infrastructure effort. Identifyopportunities to apply AI to troubleshooting, incident response, monitoring, release management, and platform operations. Help define practical, secure, and scalable ways of using AI in DevOps processes, whilemaintainingstrong engineering standards and governance. Observability, Reliability & Continuous Improvement Improve monitoring, alerting, logging, and observability practices to help teams detect issues earlier and resolve incidents faster. Analyze platform performance, deployment quality, incidents, and recurring operational patterns to recommend practical improvements. Collaborate closely with software engineering, security, product, and architecture teams to continuously improve our delivery andinfrastructure standards. What You'll Need 8+ years of professional experience in DevOps&Infrastructure engineering,Platform engineering, or a similar role in a SaaS or product-based environment. Hands-on experience with Azure(or similar)cloud services and cloud-native infrastructure design Strong practical knowledge of CI/CD, deployment automation, release pipelines, and environment management. Experience with Infrastructure as Code tools such as Terraform,Pulumi,or similar technologies. Good understanding ofcontainers,Kubernetes, AKS,Ansilbe, Packer, networking, security, monitoring, and reliability practices. Practical experience using AI-assisted coding or automation tools, or a strong interest in applying them to DevOps and infrastructure workflows. A strong enablement mindset: you enjoy creating standards, reusable solutions, documentation, and automation that help engineering teams move faster andmore safely. Bonus Points Experience using machine learning, heuristic analysis, or AI-assisted techniques to analyze logs, detect anomalies, spot trends, and support proactive troubleshooting. Experience building internal developer platforms, self-service deployment workflows, golden paths, or platform engineering tooling. Experience with security automation, policy-as-code, cost optimization, or FinOps practices in cloud environments. Exposure to AI infrastructure,MLOps, data pipelines, or supporting engineering teams that build AI-enabled products. Working Locations Croatia, Zagreb - Hybrid Croatia, Pula - Hybrid UK - Remote Why Join Us? Join a product-focused engineering environment where DevOps and infrastructure directly enable scale, reliability, and innovation. Shape modern cloud, infrastructure, and deployment practices across a growing SaaS platform. Introduce practical AI-driven automation that helps engineering teams work faster, smarter, and with greater confidence. Work on meaningful platform improvements that reduce manual effort, improve reliability, and create a better developer experience.
16/07/2026
Full time
Your profile Are you a senior DevOps and infrastructure engineer who enjoys building reliable cloud platforms, enabling engineering teams, and using AI to remove operational complexity? At Cenosco, we buildAsset Integrity Management software trusted by global leaders in Oil & Gas, Chemicals, and Energy. We are looking for someone who can strengthen our DevOps foundation, evolve our infrastructure, and help us use AI-powered tools to automate how we build, deploy, monitor, and operate our platform. This role is primarily focused onDevOps, infrastructure enablement, IaC, deployment automation, and platform reliability. You will also play an important part in shaping how we use AI and agentic tooling (i.e. Copilot, Cursor, Codex, and similar tools) to make DevOps work faster, smarter, and more scalable. What You'll Do DevOps, Infrastructure & Platform Enablement Design, build, andmaintainscalable cloud infrastructure on Azure, with a strong focus on reliability, security, performance, and operational simplicity. Own and improveInfrastructure as Code practices usingTerraform, Pulumi, or similar tools, creating reusable patterns that help teams provision and manage environments consistently. Build,optimize, and automate CI/CD pipelines for application, infrastructure, and platform deployments across environments. Enable engineering teams by creating self-service infrastructure, deployment templates, automation scripts, documentation, and platform standards. Drive containerization and orchestration practices usingKubernetesand Talos, ensuring our platform can scale reliably and efficiently. Embed security, compliance, monitoring, and disaster recovery considerations into infrastructure and deployment workflows from the start. AI-Assisted DevOps Automation Use AI-assisted and agentic tools to automate repetitive DevOps work, including pipeline creation,IaCgeneration, environment provisioning, deployment workflows, and operational runbooks. Experiment with tools such asGitHub Copilot,Cursor,Codex, and similar solutions to improve engineering productivity and reduce manual infrastructure effort. Identifyopportunities to apply AI to troubleshooting, incident response, monitoring, release management, and platform operations. Help define practical, secure, and scalable ways of using AI in DevOps processes, whilemaintainingstrong engineering standards and governance. Observability, Reliability & Continuous Improvement Improve monitoring, alerting, logging, and observability practices to help teams detect issues earlier and resolve incidents faster. Analyze platform performance, deployment quality, incidents, and recurring operational patterns to recommend practical improvements. Collaborate closely with software engineering, security, product, and architecture teams to continuously improve our delivery andinfrastructure standards. What You'll Need 8+ years of professional experience in DevOps&Infrastructure engineering,Platform engineering, or a similar role in a SaaS or product-based environment. Hands-on experience with Azure(or similar)cloud services and cloud-native infrastructure design Strong practical knowledge of CI/CD, deployment automation, release pipelines, and environment management. Experience with Infrastructure as Code tools such as Terraform,Pulumi,or similar technologies. Good understanding ofcontainers,Kubernetes, AKS,Ansilbe, Packer, networking, security, monitoring, and reliability practices. Practical experience using AI-assisted coding or automation tools, or a strong interest in applying them to DevOps and infrastructure workflows. A strong enablement mindset: you enjoy creating standards, reusable solutions, documentation, and automation that help engineering teams move faster andmore safely. Bonus Points Experience using machine learning, heuristic analysis, or AI-assisted techniques to analyze logs, detect anomalies, spot trends, and support proactive troubleshooting. Experience building internal developer platforms, self-service deployment workflows, golden paths, or platform engineering tooling. Experience with security automation, policy-as-code, cost optimization, or FinOps practices in cloud environments. Exposure to AI infrastructure,MLOps, data pipelines, or supporting engineering teams that build AI-enabled products. Working Locations Croatia, Zagreb - Hybrid Croatia, Pula - Hybrid UK - Remote Why Join Us? Join a product-focused engineering environment where DevOps and infrastructure directly enable scale, reliability, and innovation. Shape modern cloud, infrastructure, and deployment practices across a growing SaaS platform. Introduce practical AI-driven automation that helps engineering teams work faster, smarter, and with greater confidence. Work on meaningful platform improvements that reduce manual effort, improve reliability, and create a better developer experience.
Description As a Senior Software Engineer (SaaS), you will join a substantial global engineering organisation and provide leadership to highly skilled engineers in the software platform area. You will be part of a team that follows Agile methodologies to deliver market leading insurance solutions. This is a new position created as the organisation continues to grow and accelerate the delivery of innovative solutions that harness the latest advances in technology. The role is a dedicated focus on the evolution of ICT Platform Infrastructure capabilities which are the foundation of many of the company's market leading products. You will have worked on large scale and globally deployed SaaS applications that have a mature stack management capability. A sound technical background is essential to be successful in this role with experience in delivering large scale Cloud services. The Role Provide technical input, guidance and leadership to the team (including code quality, best practices, processes, some aspects of release management, etc.). Assist in the design and documenting of solutions meeting functional and non functional requirements. Lead by example getting directly involved in the day to day delivery of work with hands on development, following best practices for maintainability, testability and performance. Participate in sprint planning meetings, daily stand ups and sprint retrospectives, striving to continuously improve the team velocity, its processes and engineering practices. Coach and mentor colleagues where appropriate, fostering a collaborative and quality focused engineering culture. Qualifications What you'll bring Platform and control plane engineering + Practical experience designing and developing management and control plane capabilities for line of business or SaaS applications, including stack/environment management, API design for platform services, and developer/operator UX. Experience working with or designing Azure stamp based or scale unit architectures, including multi region or multi tenant deployments, isolation and blast radius reduction, and automated stamp lifecycle management. SaaS product lifecycle + Demonstrated experience across the full SaaS development lifecycle: requirements analysis, estimation, architecture/design, implementation, unit and system testing, deployment, operations, monitoring and incident management. Software engineering fundamentals + Strong grounding in object oriented design, design patterns, SOLID principles, clean code practices, code review, documentation and Agile delivery. Core technical stack + Hands on experience with C# and .NET (Core/6+) for backend and service development. + Strong practical experience with Azure cloud and SaaS services (e.g., Container Apps, AKS, Azure Storage, Azure SQL/Cosmos DB, Key Vault, Azure Monitor, Log Analytics). Containers, Container Apps, Kubernetes and microservices + Experience designing, building and operating containerised microservices, including Docker image design, optimisation, Container Apps, Kubernetes (preferably AKS) deployment and scaling strategies, and Kubernetes configuration, secrets and traffic management. Infrastructure as Code and platform automation + Experience with IaC using OpenTofu or Terraform on Azure, producing modular, reusable, and maintainable infrastructure definitions. + Experience automating environment and stamp provisioning, implementing blue/green or canary deployments, and detecting/remediating infrastructure drift. CI/CD and DevOps + Hands on experience building CI/CD pipelines in Azure DevOps (Pipelines, Repos), including environments and quality gates. + Familiarity with GitHub and GitHub Actions for reusable workflows, build/test/deploy automation and integration with platform tooling. Automation and testing of infrastructure and platforms + Experience implementing automated testing for infrastructure and platform components, including IaC unit/integration tests, deployment validation, policy as code (Azure Policy, OPA) and IaC linting. Security and compliance + Strong understanding of cloud and SaaS security best practices: least privilege, managed identities, secret management, network security, private endpoints, and secure software supply chain practices (image scanning, SBOM, dependency governance). + Awareness of compliance, data residency and multi region considerations in multi tenant Azure environments. Technical leadership and communication + Ability to run proofs of concept with emerging technologies and present outcomes to both technical and non technical audiences. + Comfortable partnering with application teams to gather requirements and shape platform capabilities. Other highly desirable, but not essential skills are Microsoft Azure certifications (e.g., AZ 204, AZ 400, AZ 305). Kubernetes certifications (e.g., CKA, CKAD, CKS). Degree in Computer Science, Engineering, Mathematics or equivalent industry experience. Experience with high volume, low latency SaaS systems. Experience with distributed and event driven architectures (Azure Service Bus, Event Hubs, Event Grid, Kafka). Experience building multi tenant architectures, including tenant provisioning, routing and per tenant configuration/quota management. Experience with Azure ingress and global routing solutions such as Azure Front Door, Application Gateway, and Azure API Management. Experience with identity and access management using Azure AD/Entra ID, B2B, RBAC, OAuth2, OIDC, JWT and claims based authorisation. Experience in platform engineering or SRE roles: building internal platforms, defining SLI/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing). Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an understanding of their applicability to SaaS platforms. General knowledge of the insurance industry or other regulated domains. What we offer Enjoy a benefits package designed to help you thrive, both professionally and personally. You'll receive 25 days of annual leave plus an extra company day to relax and recharge. Our comprehensive health and wellbeing offering includes private healthcare, life insurance, group income protection, and regular health assessments, all giving you peace of mind. Secure your future with our defined contribution pension scheme, featuring matched contributions up to 10% from the company. We support your growth and balance with hybrid working options, access to an employee assistance programme, and a fully paid volunteer day to make a difference in your community. On top of these, you can opt into a variety of additional perks including an electric vehicle car scheme, share scheme, cycle to work programme, dental and optical cover, critical illness protection, and much more. Start making the most of your career and wellbeing with a range of benefits tailored for you. Equal Opportunity Employer We're committed to equal employment opportunity and provide application, interview and workplace adjustments and accommodations to all applicants. If you foresee any barriers, from the application process through to joining the company, please email .
12/07/2026
Full time
Description As a Senior Software Engineer (SaaS), you will join a substantial global engineering organisation and provide leadership to highly skilled engineers in the software platform area. You will be part of a team that follows Agile methodologies to deliver market leading insurance solutions. This is a new position created as the organisation continues to grow and accelerate the delivery of innovative solutions that harness the latest advances in technology. The role is a dedicated focus on the evolution of ICT Platform Infrastructure capabilities which are the foundation of many of the company's market leading products. You will have worked on large scale and globally deployed SaaS applications that have a mature stack management capability. A sound technical background is essential to be successful in this role with experience in delivering large scale Cloud services. The Role Provide technical input, guidance and leadership to the team (including code quality, best practices, processes, some aspects of release management, etc.). Assist in the design and documenting of solutions meeting functional and non functional requirements. Lead by example getting directly involved in the day to day delivery of work with hands on development, following best practices for maintainability, testability and performance. Participate in sprint planning meetings, daily stand ups and sprint retrospectives, striving to continuously improve the team velocity, its processes and engineering practices. Coach and mentor colleagues where appropriate, fostering a collaborative and quality focused engineering culture. Qualifications What you'll bring Platform and control plane engineering + Practical experience designing and developing management and control plane capabilities for line of business or SaaS applications, including stack/environment management, API design for platform services, and developer/operator UX. Experience working with or designing Azure stamp based or scale unit architectures, including multi region or multi tenant deployments, isolation and blast radius reduction, and automated stamp lifecycle management. SaaS product lifecycle + Demonstrated experience across the full SaaS development lifecycle: requirements analysis, estimation, architecture/design, implementation, unit and system testing, deployment, operations, monitoring and incident management. Software engineering fundamentals + Strong grounding in object oriented design, design patterns, SOLID principles, clean code practices, code review, documentation and Agile delivery. Core technical stack + Hands on experience with C# and .NET (Core/6+) for backend and service development. + Strong practical experience with Azure cloud and SaaS services (e.g., Container Apps, AKS, Azure Storage, Azure SQL/Cosmos DB, Key Vault, Azure Monitor, Log Analytics). Containers, Container Apps, Kubernetes and microservices + Experience designing, building and operating containerised microservices, including Docker image design, optimisation, Container Apps, Kubernetes (preferably AKS) deployment and scaling strategies, and Kubernetes configuration, secrets and traffic management. Infrastructure as Code and platform automation + Experience with IaC using OpenTofu or Terraform on Azure, producing modular, reusable, and maintainable infrastructure definitions. + Experience automating environment and stamp provisioning, implementing blue/green or canary deployments, and detecting/remediating infrastructure drift. CI/CD and DevOps + Hands on experience building CI/CD pipelines in Azure DevOps (Pipelines, Repos), including environments and quality gates. + Familiarity with GitHub and GitHub Actions for reusable workflows, build/test/deploy automation and integration with platform tooling. Automation and testing of infrastructure and platforms + Experience implementing automated testing for infrastructure and platform components, including IaC unit/integration tests, deployment validation, policy as code (Azure Policy, OPA) and IaC linting. Security and compliance + Strong understanding of cloud and SaaS security best practices: least privilege, managed identities, secret management, network security, private endpoints, and secure software supply chain practices (image scanning, SBOM, dependency governance). + Awareness of compliance, data residency and multi region considerations in multi tenant Azure environments. Technical leadership and communication + Ability to run proofs of concept with emerging technologies and present outcomes to both technical and non technical audiences. + Comfortable partnering with application teams to gather requirements and shape platform capabilities. Other highly desirable, but not essential skills are Microsoft Azure certifications (e.g., AZ 204, AZ 400, AZ 305). Kubernetes certifications (e.g., CKA, CKAD, CKS). Degree in Computer Science, Engineering, Mathematics or equivalent industry experience. Experience with high volume, low latency SaaS systems. Experience with distributed and event driven architectures (Azure Service Bus, Event Hubs, Event Grid, Kafka). Experience building multi tenant architectures, including tenant provisioning, routing and per tenant configuration/quota management. Experience with Azure ingress and global routing solutions such as Azure Front Door, Application Gateway, and Azure API Management. Experience with identity and access management using Azure AD/Entra ID, B2B, RBAC, OAuth2, OIDC, JWT and claims based authorisation. Experience in platform engineering or SRE roles: building internal platforms, defining SLI/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing). Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an understanding of their applicability to SaaS platforms. General knowledge of the insurance industry or other regulated domains. What we offer Enjoy a benefits package designed to help you thrive, both professionally and personally. You'll receive 25 days of annual leave plus an extra company day to relax and recharge. Our comprehensive health and wellbeing offering includes private healthcare, life insurance, group income protection, and regular health assessments, all giving you peace of mind. Secure your future with our defined contribution pension scheme, featuring matched contributions up to 10% from the company. We support your growth and balance with hybrid working options, access to an employee assistance programme, and a fully paid volunteer day to make a difference in your community. On top of these, you can opt into a variety of additional perks including an electric vehicle car scheme, share scheme, cycle to work programme, dental and optical cover, critical illness protection, and much more. Start making the most of your career and wellbeing with a range of benefits tailored for you. Equal Opportunity Employer We're committed to equal employment opportunity and provide application, interview and workplace adjustments and accommodations to all applicants. If you foresee any barriers, from the application process through to joining the company, please email .
VIQU IT Recruitment
Milton Keynes, Buckinghamshire
Senior Site Reliability Engineer Up to £70,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation. This is a genuine opportunity to own and operate how the cloud function works, and progress into a team lead position as the team grows. Experience required for the Senior Site Reliability Engineer Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment - e.g SaaS or MSP. Strong hands-on experience with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code / Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working. Ability to communicate across internal teams and external customers. Skilled in networking across both cloud (Azure) and on premise environments. Either Windows or Linux systems administration skills (Linux preferred). Previous use of AI tools to enhance efficiency. Job Duties of the Senior Site Reliability Engineer Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure servers and networks, and automate application lifecycles. Regularly use Datadog and other observability tools for application performance monitoring. Implement new ways of working, helping to shape how the organisation responds and recovers to incidents. Take ownership of incident resolutions. Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents. Work on an a on call rota, ensuring you are available to respond to incidents during this time. Identify areas for automation and help implement changes that raise the bar for reliability. Apply now to speak with VIQU IT in confidence. Or reach out to Jack McManus via the Do you know someone great? We'll thank you with up to £1,000 if your referral is successful (terms apply). For more exciting roles and opportunities like this, please follow us on IT Recruitment
11/07/2026
Full time
Senior Site Reliability Engineer Up to £70,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation. This is a genuine opportunity to own and operate how the cloud function works, and progress into a team lead position as the team grows. Experience required for the Senior Site Reliability Engineer Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment - e.g SaaS or MSP. Strong hands-on experience with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code / Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working. Ability to communicate across internal teams and external customers. Skilled in networking across both cloud (Azure) and on premise environments. Either Windows or Linux systems administration skills (Linux preferred). Previous use of AI tools to enhance efficiency. Job Duties of the Senior Site Reliability Engineer Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure servers and networks, and automate application lifecycles. Regularly use Datadog and other observability tools for application performance monitoring. Implement new ways of working, helping to shape how the organisation responds and recovers to incidents. Take ownership of incident resolutions. Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents. Work on an a on call rota, ensuring you are available to respond to incidents during this time. Identify areas for automation and help implement changes that raise the bar for reliability. Apply now to speak with VIQU IT in confidence. Or reach out to Jack McManus via the Do you know someone great? We'll thank you with up to £1,000 if your referral is successful (terms apply). For more exciting roles and opportunities like this, please follow us on IT Recruitment
Senior Site Reliability Engineer Up to £70,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation. This is a genuine opportunity to own and operate how the cloud function works, and progress into a team lead position as the team grows. Experience required for the Senior Site Reliability Engineer Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment e.g SaaS or MSP. Strong hands-on experience with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code / Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working. Ability to communicate across internal teams and external customers. Skilled in networking across both cloud (Azure) and on premise environments. Either Windows or Linux systems administration skills (Linux preferred). Previous use of AI tools to enhance efficiency. Job Duties of the Senior Site Reliability Engineer Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure servers and networks, and automate application lifecycles. Regularly use Datadog and other observability tools for application performance monitoring. Implement new ways of working, helping to shape how the organisation responds and recovers to incidents. Take ownership of incident resolutions. Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents. Work on an a on call rota, ensuring you are available to respond to incidents during this time. Identify areas for automation and help implement changes that raise the bar for reliability. Apply now to speak with VIQU IT in confidence. Or reach out to Jack McManus via the (url removed) Do you know someone great? We ll thank you with up to £1,000 if your referral is successful (terms apply). For more exciting roles and opportunities like this, please follow us on IT Recruitment
08/07/2026
Full time
Senior Site Reliability Engineer Up to £70,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation. This is a genuine opportunity to own and operate how the cloud function works, and progress into a team lead position as the team grows. Experience required for the Senior Site Reliability Engineer Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment e.g SaaS or MSP. Strong hands-on experience with both Azure, and on-premise virtual machines. Experience withInfrastructure as Code / Terraform, Container orchestration (Kubernetes or AKS), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor). Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working. Ability to communicate across internal teams and external customers. Skilled in networking across both cloud (Azure) and on premise environments. Either Windows or Linux systems administration skills (Linux preferred). Previous use of AI tools to enhance efficiency. Job Duties of the Senior Site Reliability Engineer Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure servers and networks, and automate application lifecycles. Regularly use Datadog and other observability tools for application performance monitoring. Implement new ways of working, helping to shape how the organisation responds and recovers to incidents. Take ownership of incident resolutions. Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents. Work on an a on call rota, ensuring you are available to respond to incidents during this time. Identify areas for automation and help implement changes that raise the bar for reliability. Apply now to speak with VIQU IT in confidence. Or reach out to Jack McManus via the (url removed) Do you know someone great? We ll thank you with up to £1,000 if your referral is successful (terms apply). For more exciting roles and opportunities like this, please follow us on IT Recruitment
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
06/07/2026
Full time
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Senior Azure DevOps Engineer - Shared Records Reporting To: Chief Software Engineer Department: Development Location: Milton Keynes/Homebased - with occasional travel to the office in Milton Keynes Overview Graphnet Health is the leading UK supplier of integrated healthcare solutions to the NHS. Our software solutions are helping to revolutionise patient care across the country, driving improved patient outcomes and patient satisfaction and delivering greater efficiency across the healthcare system. We have high ambitions and looking to bring in highly skilled individuals to help continue our adoption of Azure-native services and building a best-of-breed scalable SaaS platform for our solutions. The person we are looking for will need to have 5+ years enterprise experience with Azure, who can demonstrate a real passion for cloud technologies, with strong problem solving and critical thinking, good communication skills and a strong work ethic. Working within an existing DevOps team and part of a wider DevOps Consortium, the individual will have a diverse workload covering (but not limited to): Automation of infrastructure and software rollout using Azure DevOps pipelines, including quality gates such as approvals and code quality scanning and reusable templates to enable consistent scalable delivery. Design and maintain Infrastructure as Code using Terraform, including modular architectures, remote state management, and environment segregation with enterprise level governance. Support and operate enterprise scale cloud platforms across multiple regions and environments, ensuring resilience, consistency, and high availability. Design and develop automated scaling and failover using deep health checks and insights. Architect and implement cloud infrastructure aligned with the Azure Well Architected Framework, ensuring adherence to reliability, security, performance, operational excellence, and cost optimisation principles. Work closely with others, such as Development or Platform teams, to agree concepts/solutions that fulfil complex requirements yet remain achievable. Work closely with other departments, such as Ops and TechOps, to facilitate the smooth running of the cloud services by providing utilities/dashboards/scripts. Work closely with the Security teams to analyse and help remediate on internal and SoC flagged issues, create secure pipeline design and assisting with vulnerability remediation. Design, test, and continuously improve disaster recovery and business continuity strategies aligned to defined RTO/RPO targets. Education & Skills Required A demonstrable understanding of Azure PaaS components. Key areas: Containerisation and orchestration with AKS and Docker (design, scaling, network and security). API Management. App Services and Azure Functions. Observability using Azure Monitor, Log Analytics and Application Insights, Prometheus and Grafana. Implementation of governance and compliance controls using Azure Policy, Management Groups, and Landing Zone principles. Azure SQL and Managed Instance. Event driven architecture (Service Bus, Event Grid). Data integration and processing using Data Factory / Databricks. Identity and Access Management (RBAC, Key Vault and Managed Identities). Azure networking (e.g. VNets, NSGs, Private endpoints, DNS, load balancing, Front Door) and understanding of the hub spoke topology. Proven experience designing and implementing CI/CD pipelines in GitHub Actions and/or Azure DevOps using YAML, including release strategies, approvals, and artifact management as well as testing integration. Strong experience with Git based workflows (branching strategies, repo management including pull requests and code reviews). Hands on experience and scripting skills in: HELM and Flux. Powershell and Azure CLI. Terraform. Experience in producing and maintaining technical documentation, standards, and reusable patterns to support knowledge sharing and team scalability. Advantageous Healthcare or Government related industry experience. Understanding of JIRA and Confluence. Exposure to AIOps practices, including intelligent monitoring and experience using AI assisted development and operations tooling e.g. GitHub Copilot, intelligent automation. Qualifications Microsoft certification(s) in Azure, such as AZ 400, AZ 305. Experience is valued over accreditation, however, there will be encouragement to gain accreditation during employment.
02/07/2026
Full time
Senior Azure DevOps Engineer - Shared Records Reporting To: Chief Software Engineer Department: Development Location: Milton Keynes/Homebased - with occasional travel to the office in Milton Keynes Overview Graphnet Health is the leading UK supplier of integrated healthcare solutions to the NHS. Our software solutions are helping to revolutionise patient care across the country, driving improved patient outcomes and patient satisfaction and delivering greater efficiency across the healthcare system. We have high ambitions and looking to bring in highly skilled individuals to help continue our adoption of Azure-native services and building a best-of-breed scalable SaaS platform for our solutions. The person we are looking for will need to have 5+ years enterprise experience with Azure, who can demonstrate a real passion for cloud technologies, with strong problem solving and critical thinking, good communication skills and a strong work ethic. Working within an existing DevOps team and part of a wider DevOps Consortium, the individual will have a diverse workload covering (but not limited to): Automation of infrastructure and software rollout using Azure DevOps pipelines, including quality gates such as approvals and code quality scanning and reusable templates to enable consistent scalable delivery. Design and maintain Infrastructure as Code using Terraform, including modular architectures, remote state management, and environment segregation with enterprise level governance. Support and operate enterprise scale cloud platforms across multiple regions and environments, ensuring resilience, consistency, and high availability. Design and develop automated scaling and failover using deep health checks and insights. Architect and implement cloud infrastructure aligned with the Azure Well Architected Framework, ensuring adherence to reliability, security, performance, operational excellence, and cost optimisation principles. Work closely with others, such as Development or Platform teams, to agree concepts/solutions that fulfil complex requirements yet remain achievable. Work closely with other departments, such as Ops and TechOps, to facilitate the smooth running of the cloud services by providing utilities/dashboards/scripts. Work closely with the Security teams to analyse and help remediate on internal and SoC flagged issues, create secure pipeline design and assisting with vulnerability remediation. Design, test, and continuously improve disaster recovery and business continuity strategies aligned to defined RTO/RPO targets. Education & Skills Required A demonstrable understanding of Azure PaaS components. Key areas: Containerisation and orchestration with AKS and Docker (design, scaling, network and security). API Management. App Services and Azure Functions. Observability using Azure Monitor, Log Analytics and Application Insights, Prometheus and Grafana. Implementation of governance and compliance controls using Azure Policy, Management Groups, and Landing Zone principles. Azure SQL and Managed Instance. Event driven architecture (Service Bus, Event Grid). Data integration and processing using Data Factory / Databricks. Identity and Access Management (RBAC, Key Vault and Managed Identities). Azure networking (e.g. VNets, NSGs, Private endpoints, DNS, load balancing, Front Door) and understanding of the hub spoke topology. Proven experience designing and implementing CI/CD pipelines in GitHub Actions and/or Azure DevOps using YAML, including release strategies, approvals, and artifact management as well as testing integration. Strong experience with Git based workflows (branching strategies, repo management including pull requests and code reviews). Hands on experience and scripting skills in: HELM and Flux. Powershell and Azure CLI. Terraform. Experience in producing and maintaining technical documentation, standards, and reusable patterns to support knowledge sharing and team scalability. Advantageous Healthcare or Government related industry experience. Understanding of JIRA and Confluence. Exposure to AIOps practices, including intelligent monitoring and experience using AI assisted development and operations tooling e.g. GitHub Copilot, intelligent automation. Qualifications Microsoft certification(s) in Azure, such as AZ 400, AZ 305. Experience is valued over accreditation, however, there will be encouragement to gain accreditation during employment.
About StackOne: StackOne is the AI Integration Gateway for SaaS products and AI Agents. Backed by GV and Workday Ventures ($24M raised), we help builders of SaaS platforms and AI Agents orchestrate hundreds of scalable, accurate, and enterprise-grade integrations. Our platform combines 25,000 pre-mapped actions on 200 connectors, an AI-powered integration development toolkit, plus security by design: a real-time architecture, managed authentication and permissions, and end-to-end observability. Join us on our fast trajectory to build the future of agentic integrations. About the role We're looking for a Senior Platform Engineer to own how StackOne is built, shipped, and run, as we scale across our own cloud and into our customers' clouds. You'll own the infrastructure behind the platform, our deployment pipeline and developer tooling, and how we package StackOne to run inside customers' own AWS, GCP, or Azure accounts. It's a hands on role with broad scope. You write code and tooling, you own the IaC other engineers depend on, and you set the standard for how every new repository gets deployed and secured. You'll report directly to the CTO and work closely with our Security Engineer and tech leads. Responsibilities Own our infrastructure at scale: the AWS estate today (ECS Fargate, Aurora, ElastiCache, MSK, OpenSearch, Lambda, KMS) and the AWS CDK to Terraform migration. Keep it reliable, observable, and cost aware. Build out the deployment pipeline as we consolidate toward a monorepo: the CI/CD that ships every service, plus an automated end to end testing harness with incremental (affected only) testing that stays fast as we grow. Ship into customers' own clouds (self hosted / BYOC): the Terraform modules, container images, runbooks, and documentation for internal teams and customers. Own the release and upgrade path for self hosted customers: versioned, signed releases, a supported version policy, and a call home for usage and version reporting. Set the standard for new repositories: partner with the Security Engineer and tech leads so new projects (product, internal tools, and vibe coded prototypes) ship secure and deployable from day one. Templates and golden paths, not gate keeping. Raise reliability: SLOs, observability, and incident response. Make the system easy to operate when something breaks at 3am. Treat infrastructure as a product: paved roads and self serve tooling so product engineers ship without waiting on you. Use AI in the workflow: lean on LLMs and agents for IaC generation, test scaffolding, and runbook drafting, with guardrails you trust. What we're looking for 4+ years in platform, infrastructure, SRE, or DevOps, with hands on AWS at scale: ECS/Fargate, RDS/Aurora, Lambda, networking, IAM, KMS. Deep IaC ability: Terraform and/or AWS CDK, comfortable owning modules other teams build on. Strong CI/CD and monorepo build experience: caching and incremental/affected only test execution. You've made a slow pipeline fast. Strong coding in TypeScript, Python, or Go. You build tooling, not just YAML and configs. Containers in production (Docker); Kubernetes a plus. Security minded: you bake secure defaults into pipelines and repo templates, and partner with security rather than routing around it. A clear writer: docs and runbooks that internal engineers and external customers can follow. End to end ownership: you scope, ship, and measure, and automate instead of running manual checklists. Nice to have Shipping software into customers' own cloud or on prem (BYOC / self hosted), or a platform like Nuon, Replicated, or Omnistrate. Multi cloud (GCP or Azure) IaC, and exposure to Temporal, Kafka/MSK, ClickHouse, or OpenSearch. Our stack Cloud & infra: AWS (ECS Fargate, Aurora Postgres, ElastiCache, MSK, OpenSearch, Lambda, S3, KMS, CloudFront, WAF), Cloudflare (Workers, WAF) IaC: AWS CDK today, migrating to Terraform Data & messaging: Postgres, Redis, Kafka, OpenSearch, ClickHouse CI/CD: GitHub Actions Observability & analytics: Datadog, Sentry, Metabase Languages: TypeScript (Node.js), Python Benefits Meaningful share options (EMI) 25 days holiday + 1 additional day per year of tenure Private health insurance, including dental & optical £15/day London office lunch budget, up to £120/month £1,000 home office setup + £500/year top up Annual team offsite to sunny spots Join one of Europe's fastest growing startups Work with a veteran team of ex Google, Microsoft, Oracle, Coinbase, JP Morgan and more Health, fitness and gift card discounts; Cycle2Work and Electric Cars scheme London (hybrid, 2 days/week) preferred; open to remote within the UK We believe diversity drives innovation. We encourage individuals from all backgrounds to apply. As an equal opportunity employer, we celebrate diversity and are committed to creating an inclusive environment for all employees.
30/06/2026
Full time
About StackOne: StackOne is the AI Integration Gateway for SaaS products and AI Agents. Backed by GV and Workday Ventures ($24M raised), we help builders of SaaS platforms and AI Agents orchestrate hundreds of scalable, accurate, and enterprise-grade integrations. Our platform combines 25,000 pre-mapped actions on 200 connectors, an AI-powered integration development toolkit, plus security by design: a real-time architecture, managed authentication and permissions, and end-to-end observability. Join us on our fast trajectory to build the future of agentic integrations. About the role We're looking for a Senior Platform Engineer to own how StackOne is built, shipped, and run, as we scale across our own cloud and into our customers' clouds. You'll own the infrastructure behind the platform, our deployment pipeline and developer tooling, and how we package StackOne to run inside customers' own AWS, GCP, or Azure accounts. It's a hands on role with broad scope. You write code and tooling, you own the IaC other engineers depend on, and you set the standard for how every new repository gets deployed and secured. You'll report directly to the CTO and work closely with our Security Engineer and tech leads. Responsibilities Own our infrastructure at scale: the AWS estate today (ECS Fargate, Aurora, ElastiCache, MSK, OpenSearch, Lambda, KMS) and the AWS CDK to Terraform migration. Keep it reliable, observable, and cost aware. Build out the deployment pipeline as we consolidate toward a monorepo: the CI/CD that ships every service, plus an automated end to end testing harness with incremental (affected only) testing that stays fast as we grow. Ship into customers' own clouds (self hosted / BYOC): the Terraform modules, container images, runbooks, and documentation for internal teams and customers. Own the release and upgrade path for self hosted customers: versioned, signed releases, a supported version policy, and a call home for usage and version reporting. Set the standard for new repositories: partner with the Security Engineer and tech leads so new projects (product, internal tools, and vibe coded prototypes) ship secure and deployable from day one. Templates and golden paths, not gate keeping. Raise reliability: SLOs, observability, and incident response. Make the system easy to operate when something breaks at 3am. Treat infrastructure as a product: paved roads and self serve tooling so product engineers ship without waiting on you. Use AI in the workflow: lean on LLMs and agents for IaC generation, test scaffolding, and runbook drafting, with guardrails you trust. What we're looking for 4+ years in platform, infrastructure, SRE, or DevOps, with hands on AWS at scale: ECS/Fargate, RDS/Aurora, Lambda, networking, IAM, KMS. Deep IaC ability: Terraform and/or AWS CDK, comfortable owning modules other teams build on. Strong CI/CD and monorepo build experience: caching and incremental/affected only test execution. You've made a slow pipeline fast. Strong coding in TypeScript, Python, or Go. You build tooling, not just YAML and configs. Containers in production (Docker); Kubernetes a plus. Security minded: you bake secure defaults into pipelines and repo templates, and partner with security rather than routing around it. A clear writer: docs and runbooks that internal engineers and external customers can follow. End to end ownership: you scope, ship, and measure, and automate instead of running manual checklists. Nice to have Shipping software into customers' own cloud or on prem (BYOC / self hosted), or a platform like Nuon, Replicated, or Omnistrate. Multi cloud (GCP or Azure) IaC, and exposure to Temporal, Kafka/MSK, ClickHouse, or OpenSearch. Our stack Cloud & infra: AWS (ECS Fargate, Aurora Postgres, ElastiCache, MSK, OpenSearch, Lambda, S3, KMS, CloudFront, WAF), Cloudflare (Workers, WAF) IaC: AWS CDK today, migrating to Terraform Data & messaging: Postgres, Redis, Kafka, OpenSearch, ClickHouse CI/CD: GitHub Actions Observability & analytics: Datadog, Sentry, Metabase Languages: TypeScript (Node.js), Python Benefits Meaningful share options (EMI) 25 days holiday + 1 additional day per year of tenure Private health insurance, including dental & optical £15/day London office lunch budget, up to £120/month £1,000 home office setup + £500/year top up Annual team offsite to sunny spots Join one of Europe's fastest growing startups Work with a veteran team of ex Google, Microsoft, Oracle, Coinbase, JP Morgan and more Health, fitness and gift card discounts; Cycle2Work and Electric Cars scheme London (hybrid, 2 days/week) preferred; open to remote within the UK We believe diversity drives innovation. We encourage individuals from all backgrounds to apply. As an equal opportunity employer, we celebrate diversity and are committed to creating an inclusive environment for all employees.