Attio is the CRM built for the AI era. Designed for the most ambitious go-to-market teams, it gives companies the power to understand every customer, automate at scale, and build their go-to-market motion exactly as they need. We've raised $116M from some of the world's best investors: GV (Google Ventures), Redpoint, Balderton, Point Nine, and 01A. We hire builders who thrive on complex technical challenges, hold themselves to a high bar, and genuinely care about delighting the people who use what they build. The team here brings sharp judgement, real craft, and the drive to do exceptional work. We're obsessed about the details and energized by the frontier. If you want to do the best work of your career, this is the right place. About the Role We are seeking highly skilled and experienced Platform Product Engineers to join our Security, Infrastructure and Performance team. This is a crucial, dual faceted role that combines high level engineering strategy with hands on operational excellence. The successful candidates will be responsible for building, operating, and continuously enhancing the internal technology platform, fundamentally treating this platform as a product with all development teams as its primary customers. The Platform Product Engineer role is centered around embodying and executing DevOps principles, specifically focusing on: Automation: Systematically removing manual toil from the software development lifecycle (SDLC) through the creation of robust tooling, CI/CD pipelines, and infrastructure as code (IaC). Collaboration: Fostering a tight, cooperative partnership with product development teams, gathering requirements, and delivering solutions that accelerate their productivity and time to market. Continuous Improvement: Instilling a culture of iterative enhancement for the platform's reliability, cost efficiency, and developer experience. This mandate is underpinned by a rigorous Site Reliability Engineering (SRE) mindset. The successful candidates will be instrumental in defining and upholding Service Level Objectives (SLOs) and Service Level Indicators (SLIs), implementing effective monitoring and alerting strategies, and leading operational incident response processes. What you'll do The core responsibility is to implement, maintain, and continuously improve the foundational platform infrastructure that powers all engineering services. This necessitates a relentless focus on ensuring high reliability, exceptional scalability, and optimal performance across the entire stack. Platform Infrastructure: Build and maintain platform infrastructure using declarative IaC tools (e.g., Terraform, Pulumi), ensuring all environments are reproducible, version controlled, and auditable. Proactively manage the capacity of the infrastructure to consistently meet or exceed Service Level Objectives for latency, error rates, and availability. Incident Response and Post Mortems: Act as first line responders for critical system incidents. Triage, diagnose, and resolve complex production issues rapidly. Drive a culture of blameless post mortems, ensuring root causes are identified, and long term preventative measures are implemented as code (e.g., via runbooks, automation, or system design changes). Tooling & Automation: Own the stack of supporting tools necessary for operational excellence and developer enablement, including: Continuous Integration and Continuous Delivery (CI/CD) Pipelines: Implement, maintain, and evolve the fully automated CI and CD pipelines. This includes establishing best practices for fast, reliable, and secure build, test, and deployment processes. Observability: Implement and manage robust systems for monitoring (metrics), logging (centralised log aggregation), and distributed tracing to provide deep insights into application and infrastructure health. What you'll bring Applied DevOps and SRE Principles: Must have : Demonstrable, hands on experience applying core DevOps and Site Reliability Engineering (SRE) principles to manage, monitor, and scale production systems. Must have: A deep understanding of the SRE mindset, including SLO/SLA creation and monitoring, error budget management, toil reduction, and post incident review (blameless postmortems). Desirable: Proven ability to drive cultural and process change that fosters a collaborative approach between development and operations teams. Cloud Infrastructure and Containerisation Expertise: Must have: Expertise in one or more major public cloud providers (AWS, GCP, or Azure), encompassing network configuration, security best practices (IAM, security groups, etc.), compute services (EC2, GKE, ECS, etc.), and managed services (databases, queues, serverless functions). Must have: In depth knowledge of container technologies, specifically Docker, and extensive experience orchestrating them at scale using Kubernetes (K8s). This includes designing, deploying, and managing Kubernetes clusters, understanding networking (CNI), storage (CSI), and security configurations within the Kubernetes ecosystem. Automation and Programming Skills: Must have: Proficiency in one or more modern software languages (e.g., Typescript, Go, Python, Rust) and associated frameworks used for building high performance, resilient production systems. Must have: Proven experience developing robust, maintainable, and well tested automation scripts, services and pipelines to manage infrastructure, deployments, and operational tasks. Operational Tooling and Observability Management: Must have: Experience owning, managing, and maintaining mission critical operational tooling. Desirable: Proven background in implementing and managing centralised logging solutions or similar platforms (e.g., Splunk, DataDog). Desirable: Familiarity with distributed tracing tools (e.g., Jaeger, Zipkin) and Application Performance Monitoring (APM) solutions. What we offer Competitive salary of £95,000 to £125,000 Equity in an early stage tech company on an incredible trajectory 25 days holiday plus local public holidays Apple hardware Private medical insurance through AXA Pension contribution through Hargreaves Lansdown Enhanced family leave Team off site in fun places! (We've been to Barcelona, Lisbon, Malta, and Split so far)
11/07/2026
Full time
Attio is the CRM built for the AI era. Designed for the most ambitious go-to-market teams, it gives companies the power to understand every customer, automate at scale, and build their go-to-market motion exactly as they need. We've raised $116M from some of the world's best investors: GV (Google Ventures), Redpoint, Balderton, Point Nine, and 01A. We hire builders who thrive on complex technical challenges, hold themselves to a high bar, and genuinely care about delighting the people who use what they build. The team here brings sharp judgement, real craft, and the drive to do exceptional work. We're obsessed about the details and energized by the frontier. If you want to do the best work of your career, this is the right place. About the Role We are seeking highly skilled and experienced Platform Product Engineers to join our Security, Infrastructure and Performance team. This is a crucial, dual faceted role that combines high level engineering strategy with hands on operational excellence. The successful candidates will be responsible for building, operating, and continuously enhancing the internal technology platform, fundamentally treating this platform as a product with all development teams as its primary customers. The Platform Product Engineer role is centered around embodying and executing DevOps principles, specifically focusing on: Automation: Systematically removing manual toil from the software development lifecycle (SDLC) through the creation of robust tooling, CI/CD pipelines, and infrastructure as code (IaC). Collaboration: Fostering a tight, cooperative partnership with product development teams, gathering requirements, and delivering solutions that accelerate their productivity and time to market. Continuous Improvement: Instilling a culture of iterative enhancement for the platform's reliability, cost efficiency, and developer experience. This mandate is underpinned by a rigorous Site Reliability Engineering (SRE) mindset. The successful candidates will be instrumental in defining and upholding Service Level Objectives (SLOs) and Service Level Indicators (SLIs), implementing effective monitoring and alerting strategies, and leading operational incident response processes. What you'll do The core responsibility is to implement, maintain, and continuously improve the foundational platform infrastructure that powers all engineering services. This necessitates a relentless focus on ensuring high reliability, exceptional scalability, and optimal performance across the entire stack. Platform Infrastructure: Build and maintain platform infrastructure using declarative IaC tools (e.g., Terraform, Pulumi), ensuring all environments are reproducible, version controlled, and auditable. Proactively manage the capacity of the infrastructure to consistently meet or exceed Service Level Objectives for latency, error rates, and availability. Incident Response and Post Mortems: Act as first line responders for critical system incidents. Triage, diagnose, and resolve complex production issues rapidly. Drive a culture of blameless post mortems, ensuring root causes are identified, and long term preventative measures are implemented as code (e.g., via runbooks, automation, or system design changes). Tooling & Automation: Own the stack of supporting tools necessary for operational excellence and developer enablement, including: Continuous Integration and Continuous Delivery (CI/CD) Pipelines: Implement, maintain, and evolve the fully automated CI and CD pipelines. This includes establishing best practices for fast, reliable, and secure build, test, and deployment processes. Observability: Implement and manage robust systems for monitoring (metrics), logging (centralised log aggregation), and distributed tracing to provide deep insights into application and infrastructure health. What you'll bring Applied DevOps and SRE Principles: Must have : Demonstrable, hands on experience applying core DevOps and Site Reliability Engineering (SRE) principles to manage, monitor, and scale production systems. Must have: A deep understanding of the SRE mindset, including SLO/SLA creation and monitoring, error budget management, toil reduction, and post incident review (blameless postmortems). Desirable: Proven ability to drive cultural and process change that fosters a collaborative approach between development and operations teams. Cloud Infrastructure and Containerisation Expertise: Must have: Expertise in one or more major public cloud providers (AWS, GCP, or Azure), encompassing network configuration, security best practices (IAM, security groups, etc.), compute services (EC2, GKE, ECS, etc.), and managed services (databases, queues, serverless functions). Must have: In depth knowledge of container technologies, specifically Docker, and extensive experience orchestrating them at scale using Kubernetes (K8s). This includes designing, deploying, and managing Kubernetes clusters, understanding networking (CNI), storage (CSI), and security configurations within the Kubernetes ecosystem. Automation and Programming Skills: Must have: Proficiency in one or more modern software languages (e.g., Typescript, Go, Python, Rust) and associated frameworks used for building high performance, resilient production systems. Must have: Proven experience developing robust, maintainable, and well tested automation scripts, services and pipelines to manage infrastructure, deployments, and operational tasks. Operational Tooling and Observability Management: Must have: Experience owning, managing, and maintaining mission critical operational tooling. Desirable: Proven background in implementing and managing centralised logging solutions or similar platforms (e.g., Splunk, DataDog). Desirable: Familiarity with distributed tracing tools (e.g., Jaeger, Zipkin) and Application Performance Monitoring (APM) solutions. What we offer Competitive salary of £95,000 to £125,000 Equity in an early stage tech company on an incredible trajectory 25 days holiday plus local public holidays Apple hardware Private medical insurance through AXA Pension contribution through Hargreaves Lansdown Enhanced family leave Team off site in fun places! (We've been to Barcelona, Lisbon, Malta, and Split so far)
Sr. Applied AI Solutions Architect, Amazon Connect Job ID: AWS EMEA SARL (UK Branch) Are you a customer obsessed builder with a passion for helping customers achieve their full potential? Do you have the technical background, customer experience, and skills necessary to help accelerate customer adoption of Amazon Connect's AI capabilities? Do you love building new strategic and data driven businesses? Join the Applied AI Solutions team as an Amazon Connect Specialist Solutions Architect! The Applied AI Solutions Architecture team is seeking a hands on, customer obsessed Solutions Architect to accelerate customer adoption of Amazon Connect's AI capabilities. Applied AI Solutions is part of the AWS Specialist & Partner (ASP) org, which works backwards from our customer's most complex and business critical problems to build and execute go to market plans that turn AWS ideas into multi billion dollar businesses. We pride ourselves on thinking big, delivering exceptional results for our customers, and working across AWS as . A critical dimension of this role is Customer Data Readiness - assessing, preparing, and structuring customer data assets so that AI agents can reliably access, retrieve, and act on the right information. You will help customers evaluate their data landscape, identify gaps, establish data pipelines, and ensure their knowledge bases, CRMs, and backend systems are AI ready before agents go live. You will work at the intersection of contact center operations and applied AI, helping customers move from proof of concept to pre production for their Amazon Connect deployments. We stay closely connected to our customers and bring valuable data and insights to our product teams, strengthening the product roadmap. Our team is at its best when a customer is thinking big and needs specialized experience to innovate for their business. Key job responsibilities Customer Engagement: Lead technical discovery sessions with customer teams to understand business requirements, existing contact center architecture, and AI readiness. Translate findings into actionable implementation plans. Customer Data Readiness: Conduct data readiness assessments to evaluate the quality, accessibility, structure, and governance of customer data assets (CRMs, knowledge bases, ticketing systems, order management, etc.). Identify data gaps, recommend remediation strategies, and help customers build the data foundation required for effective AI agent tool use and RAG powered responses. Agentic AI Implementation: Design and configure agentic AI solutions within Amazon Connect, including AI agent creation, AI prompt engineering, model selection, guardrail configuration, and tool/action integration. A2A (Agent to Agent) Integration: Architect Agent to Agent communication patterns that allow Amazon Connect AI agents to collaborate with specialized agents across the enterprise (e.g., billing agents, order management agents, IT support agents), enabling multi agent workflows that span organizational boundaries. Integration Development: Build serverless integrations using AWS Lambda, API Gateway, Step Functions, and scripting (Python, Node.js) to connect Amazon Connect AI agents with customer data systems (CRMs, ERPs, databases, knowledge bases). Cloud Data Access: Architect secure access patterns to cloud based data systems to power AI agent tool use and retrieval augmented generation (RAG). Agentic IDE Proficiency: Leverage agentic development environments such as Kiro (and similar AI assisted IDEs) to accelerate development workflows, including spec driven development, agent hooks, MCP server configuration, and AI assisted code generation. Pre Production Validation: Guide customers through testing, evaluation, and validation of AI agent performance against defined success criteria before production deployment. Field Enablement: Share learnings, delivering technical deep dives, and mentoring other SAs on agentic AI implementation patterns. A day in the life Conducting data readiness assessments, identifying gaps in knowledge base coverage, and recommending data preparation steps before AI agent configuration. Designing prompt strategies and evaluating model performance across different foundation models. Building Lambda functions and API integrations that serve as tools for AI agents. Configuring MCP servers to expose customer APIs, databases, and tools in a standardized format for agent consumption. Designing A2A workflows where Amazon Connect agents hand off to or collaborate with specialized agents across the customer's enterprise. Configuring knowledge bases and data connectors for RAG powered agent responses. Running evaluation frameworks to measure AI agent accuracy, latency, and customer satisfaction. Conducting architecture reviews and providing prescriptive guidance for production readiness. Documenting implementation patterns and contributing to the team's knowledge base. About the team The Applied AI Solutions Architecture team is part of the AWS Specialist and Partner Organization (ASP). We are the technical bridge between Amazon Connect customers and the service teams building the next generation of AI powered contact center capabilities. Our team operates at the forefront of agentic AI adoption, helping customers become production ready with Amazon Connect's Unlimited AI features. Basic Qualifications Experience within specific technology domain areas (e.g., software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics). Experience in design, implementation, or consulting in applications and infrastructures. Experience communicating across technical and non technical audiences, including executive level stakeholders or clients. Preferred Qualifications Experience working with and presenting to C level executives, IT, and lines of businesses across organizations or equivalent. AWS certification, such as AWS Solutions Architect, or a similar cloud certification. Knowledge of data structures, data modeling, and database schema. Experience architecting, migrating, transforming or modernizing customer requirements to the cloud. Experience with Amazon Connect or other enterprise contact center platforms (Genesys, Avaya, Cisco, NICE, Five9, etc.). Hands on experience with Amazon Bedrock, including model invocation, agent creation, knowledge base configuration, and guardrails. Amazon is an equal opportunity employer. We believe passionately that employing a diverse workforce is central to our success. We make recruiting decisions based on your experience and skills. Protecting your privacy and the security of your data is a longstanding top priority for Amazon. Please consult our Privacy Notice () to know more about how we collect, use and transfer the personal data of our candidates. Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. Posted: June 10, 2026
11/07/2026
Full time
Sr. Applied AI Solutions Architect, Amazon Connect Job ID: AWS EMEA SARL (UK Branch) Are you a customer obsessed builder with a passion for helping customers achieve their full potential? Do you have the technical background, customer experience, and skills necessary to help accelerate customer adoption of Amazon Connect's AI capabilities? Do you love building new strategic and data driven businesses? Join the Applied AI Solutions team as an Amazon Connect Specialist Solutions Architect! The Applied AI Solutions Architecture team is seeking a hands on, customer obsessed Solutions Architect to accelerate customer adoption of Amazon Connect's AI capabilities. Applied AI Solutions is part of the AWS Specialist & Partner (ASP) org, which works backwards from our customer's most complex and business critical problems to build and execute go to market plans that turn AWS ideas into multi billion dollar businesses. We pride ourselves on thinking big, delivering exceptional results for our customers, and working across AWS as . A critical dimension of this role is Customer Data Readiness - assessing, preparing, and structuring customer data assets so that AI agents can reliably access, retrieve, and act on the right information. You will help customers evaluate their data landscape, identify gaps, establish data pipelines, and ensure their knowledge bases, CRMs, and backend systems are AI ready before agents go live. You will work at the intersection of contact center operations and applied AI, helping customers move from proof of concept to pre production for their Amazon Connect deployments. We stay closely connected to our customers and bring valuable data and insights to our product teams, strengthening the product roadmap. Our team is at its best when a customer is thinking big and needs specialized experience to innovate for their business. Key job responsibilities Customer Engagement: Lead technical discovery sessions with customer teams to understand business requirements, existing contact center architecture, and AI readiness. Translate findings into actionable implementation plans. Customer Data Readiness: Conduct data readiness assessments to evaluate the quality, accessibility, structure, and governance of customer data assets (CRMs, knowledge bases, ticketing systems, order management, etc.). Identify data gaps, recommend remediation strategies, and help customers build the data foundation required for effective AI agent tool use and RAG powered responses. Agentic AI Implementation: Design and configure agentic AI solutions within Amazon Connect, including AI agent creation, AI prompt engineering, model selection, guardrail configuration, and tool/action integration. A2A (Agent to Agent) Integration: Architect Agent to Agent communication patterns that allow Amazon Connect AI agents to collaborate with specialized agents across the enterprise (e.g., billing agents, order management agents, IT support agents), enabling multi agent workflows that span organizational boundaries. Integration Development: Build serverless integrations using AWS Lambda, API Gateway, Step Functions, and scripting (Python, Node.js) to connect Amazon Connect AI agents with customer data systems (CRMs, ERPs, databases, knowledge bases). Cloud Data Access: Architect secure access patterns to cloud based data systems to power AI agent tool use and retrieval augmented generation (RAG). Agentic IDE Proficiency: Leverage agentic development environments such as Kiro (and similar AI assisted IDEs) to accelerate development workflows, including spec driven development, agent hooks, MCP server configuration, and AI assisted code generation. Pre Production Validation: Guide customers through testing, evaluation, and validation of AI agent performance against defined success criteria before production deployment. Field Enablement: Share learnings, delivering technical deep dives, and mentoring other SAs on agentic AI implementation patterns. A day in the life Conducting data readiness assessments, identifying gaps in knowledge base coverage, and recommending data preparation steps before AI agent configuration. Designing prompt strategies and evaluating model performance across different foundation models. Building Lambda functions and API integrations that serve as tools for AI agents. Configuring MCP servers to expose customer APIs, databases, and tools in a standardized format for agent consumption. Designing A2A workflows where Amazon Connect agents hand off to or collaborate with specialized agents across the customer's enterprise. Configuring knowledge bases and data connectors for RAG powered agent responses. Running evaluation frameworks to measure AI agent accuracy, latency, and customer satisfaction. Conducting architecture reviews and providing prescriptive guidance for production readiness. Documenting implementation patterns and contributing to the team's knowledge base. About the team The Applied AI Solutions Architecture team is part of the AWS Specialist and Partner Organization (ASP). We are the technical bridge between Amazon Connect customers and the service teams building the next generation of AI powered contact center capabilities. Our team operates at the forefront of agentic AI adoption, helping customers become production ready with Amazon Connect's Unlimited AI features. Basic Qualifications Experience within specific technology domain areas (e.g., software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics). Experience in design, implementation, or consulting in applications and infrastructures. Experience communicating across technical and non technical audiences, including executive level stakeholders or clients. Preferred Qualifications Experience working with and presenting to C level executives, IT, and lines of businesses across organizations or equivalent. AWS certification, such as AWS Solutions Architect, or a similar cloud certification. Knowledge of data structures, data modeling, and database schema. Experience architecting, migrating, transforming or modernizing customer requirements to the cloud. Experience with Amazon Connect or other enterprise contact center platforms (Genesys, Avaya, Cisco, NICE, Five9, etc.). Hands on experience with Amazon Bedrock, including model invocation, agent creation, knowledge base configuration, and guardrails. Amazon is an equal opportunity employer. We believe passionately that employing a diverse workforce is central to our success. We make recruiting decisions based on your experience and skills. Protecting your privacy and the security of your data is a longstanding top priority for Amazon. Please consult our Privacy Notice () to know more about how we collect, use and transfer the personal data of our candidates. Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. Posted: June 10, 2026
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
06/07/2026
Full time
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
23/06/2026
Full time
Senior Site Reliability Engineer - iManage SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams - SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in office collaboration, rapid innovation, and developing a sense of belonging at iManage. Mondays and Fridays are reserved for focus time to get things done. Have the best of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage means You are an engineer, a builder, and a systems thinker. You'll create middleware and platform guardrails that empower developers to innovate quickly and reliably. You combine deep technical judgment with empathy to eliminate customer pain, especially when working with enthusiastic teams stewarding the world's most privileged data. You uplift those around you, act as a subject matter expert, mentor others, and drive change. You chase contributing factors over root causes, value code over documentation, and documentation over process. You'll engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you'll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on call rotations. You'll become a key voice in observability, change management, and service scalability, providing guidance during complex technical decisions and high impact events. iManage is experiencing explosive growth in its flagship cloud product. We're seeking senior software and systems engineers specializing in reliability and platform services to join our transformative cloud journey. This requires rethinking technical decisions with a beginner's mindset and a focus on resilience and sustainability. If you write code, think in systems, embrace complexity and automation, and are passionate about service resilience and scalability - we want to talk to you. sRE Responsibilities Eliminate TOIL through automation and software development. Partner cross functionally with application teams and internal stakeholders. Create a modern, cloud native platform that is resilient, cost effective, and secure by default. Scale cloud infrastructure to support our Kubernetes based ecosystem. Maintain the freshness and utility of platform services. Improve the security posture of our products. Design automation, orchestration, observability, and disaster readiness into our products. Participate in production support and on call rotations, providing senior level guidance during critical events. Lead incident management and post incident retrospectives, coaching teams in these practices. Qualifications Experience writing design documents, postmortems, and refactoring application code. Built automation to reduce operational burden or developed internal SaaS tools. Ability to advocate for SRE principles (e.g., SLOs vs SLAs) and introduce them effectively. Experience in public cloud or hosted datacenter environments (Azure and AKS preferred). A passion for collaborative teamwork and influencing reliability best practices across teams. Bonus Points Hands on experience with Linux server stacks (Ubuntu/Debian preferred). Knowledge of cloud provisioning platforms (Terraform preferred). Exposure to configuration management tools (Chef preferred). Experience with containerization/clustering technologies (Docker preferred). Familiarity with observability and alerting tools (Prometheus/Grafana or ELK/EFK). Practical experience with CI/CD pipelines and rollout strategies. A bachelor's degree (or equivalent experience) in Computer Engineering or related field. Proficiency in one or more programming languages (e.g., Java, Python, Golang). Familiarity with scripting languages (e.g., PowerShell, Bash, Python, Ruby). Benefits Creating an inclusive environment where you're encouraged to help shape the culture. Market leading salary determined through a fair and consistent process, equitable for all employees. Annual performance based bonus. Enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay). Matching pension contribution (up to 6%). Private medical insurance and cash plan. Group life cover, income protection, and critical illness protection. Flexible time off policy, 25 days of annual leave with additional flexibility. Wellness days each year to prioritize mental health and well being. Access to RethinkCare, a global behavioral health platform. We welcome those who come with a growth mindset and a hunger for learning; if you are excited about this role but your past experience doesn't align perfectly with every qualification, we encourage you to apply anyway. iManage is committed to providing an excellent candidate experience and will never ask you to engage in recruitment activity via text and exclusively communicate from emails using domain. If you have any concerns or questions about communications you have received, please send them to so our team members can review. iManage provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.