Vercel in London, UK is seeking a Software Engineer for the Persistence team to help scale our compute platform. You will own infrastructure across regions, rethink primitives, and contribute to features like Fluid Compute and Persistent Sandboxes. The role requires 5+ years of software engineering, strong Golang, and deep Linux/VM expertise. You'll write Golang daily, use Terraform, and collaborate to improve reliability and developer experience for thousands of developers globally.
21/07/2026
Full time
Vercel in London, UK is seeking a Software Engineer for the Persistence team to help scale our compute platform. You will own infrastructure across regions, rethink primitives, and contribute to features like Fluid Compute and Persistent Sandboxes. The role requires 5+ years of software engineering, strong Golang, and deep Linux/VM expertise. You'll write Golang daily, use Terraform, and collaborate to improve reliability and developer experience for thousands of developers globally.
About Vercel: Vercel is the agentic infrastructure company. We free people and agents to ship what's next. For more than a decade, Vercel has shaped how the web is built. As the team behind Next.js, v0, and AI SDK, we create products that help builders move from idea to production with speed, security, and exceptional developer experience. Now, software is entering a new era, and the next generation of products will not just be used by people. They will be built, extended, and operated by agents. We are building the platform for that future, trusted by companies like OpenAI, PayPal, Ramp, Supreme, and millions of developers worldwide. Whether you're building our products, supporting our customers, growing our community, or shaping our story, you'll help define what comes next. About the role: Developers are building faster than ever. Humans and agents are deploying more than 6 million times a day on Vercel, spinning up over 2 million Sandboxes and serving a trillion requests per month. Every single one is powered by our own compute infrastructure. The Persistence team sits within Vercel's Compute division. This team owns the deep infrastructure within Hive (our compute platform) focusing on storage and state for all our workloads (Builds, Sandboxes and Functions). This work is redefining compute with features like Fluid Compute, Persistent Sandboxes, and Custom Images. As a Software Engineer on the Persistence team, you'll work on the infrastructure that runs 100s of instances across every region where our customers deploy code, rethinking low-level primitives and contributing to platform features that improve the experience for thousands of developers every day. This role is based in London, UK. If you're within commuting distance of our London office, the role includes in-office anchor days. If you're located further afield within the UK, the role is fully remote. For location-specific details, please connect with our recruiting team. What you will do: Manage and improve our fleet of clusters, running 100s of instances deployed in every region where our customers deploy and run their code. Write Golang on a daily basis and use Terraform to provision infrastructure; you'll get to know Nomad as our scheduler for managing workloads. Rethink the primitives of our infrastructure - working with virtual filesystems, Linux primitives, and low-level virtualization. Own the reliability and performance of our compute platform, including on-call coverage for a team where every improvement has outsized impact at scale. Collaborate across teams to drive the convergence of all compute at Vercel. About you: 5+ years of software engineering experience, with Golang strongly preferred. Deep experience with virtual machines, file systems and Linux - you like knowing how things work under the hood (tcpdump, strace, and iptables are familiar tools). Experience building and operating distributed systems at scale; you design for performance and reliability. Experience with schedulers and orchestrators for managing containers and non-containerised workloads (e.g. Nomad, Kubernetes). Excellent problem-solving and communication skills, with a genuine enthusiasm for digging into problems that don't have obvious solutions. Bonus if you: Have worked on low-level virtualization or sandbox execution environments. Have product engineering experience and care about the developer-facing impact of infrastructure decisions. Have experience with on-call operations for large-scale distributed systems. Benefits: Competitive compensation package, including equity. Inclusive Healthcare Package. Learn and Grow - we provide mentorship and send you to events that help you build your network and skills. Flexible Time Off. We will provide you the gear you need to do your role, and a WFH budget for you to outfit your space as needed. Vercel is committed to fostering and empowering an inclusive community within our organization. We do not discriminate on the basis of race, religion, color, gender expression or identity, sexual orientation, national origin, citizenship, age, marital status, veteran status, disability status, or any other characteristic protected by law. Vercel encourages everyone to apply for our available positions, even if they don't necessarily check every box on the job description.
21/07/2026
Full time
About Vercel: Vercel is the agentic infrastructure company. We free people and agents to ship what's next. For more than a decade, Vercel has shaped how the web is built. As the team behind Next.js, v0, and AI SDK, we create products that help builders move from idea to production with speed, security, and exceptional developer experience. Now, software is entering a new era, and the next generation of products will not just be used by people. They will be built, extended, and operated by agents. We are building the platform for that future, trusted by companies like OpenAI, PayPal, Ramp, Supreme, and millions of developers worldwide. Whether you're building our products, supporting our customers, growing our community, or shaping our story, you'll help define what comes next. About the role: Developers are building faster than ever. Humans and agents are deploying more than 6 million times a day on Vercel, spinning up over 2 million Sandboxes and serving a trillion requests per month. Every single one is powered by our own compute infrastructure. The Persistence team sits within Vercel's Compute division. This team owns the deep infrastructure within Hive (our compute platform) focusing on storage and state for all our workloads (Builds, Sandboxes and Functions). This work is redefining compute with features like Fluid Compute, Persistent Sandboxes, and Custom Images. As a Software Engineer on the Persistence team, you'll work on the infrastructure that runs 100s of instances across every region where our customers deploy code, rethinking low-level primitives and contributing to platform features that improve the experience for thousands of developers every day. This role is based in London, UK. If you're within commuting distance of our London office, the role includes in-office anchor days. If you're located further afield within the UK, the role is fully remote. For location-specific details, please connect with our recruiting team. What you will do: Manage and improve our fleet of clusters, running 100s of instances deployed in every region where our customers deploy and run their code. Write Golang on a daily basis and use Terraform to provision infrastructure; you'll get to know Nomad as our scheduler for managing workloads. Rethink the primitives of our infrastructure - working with virtual filesystems, Linux primitives, and low-level virtualization. Own the reliability and performance of our compute platform, including on-call coverage for a team where every improvement has outsized impact at scale. Collaborate across teams to drive the convergence of all compute at Vercel. About you: 5+ years of software engineering experience, with Golang strongly preferred. Deep experience with virtual machines, file systems and Linux - you like knowing how things work under the hood (tcpdump, strace, and iptables are familiar tools). Experience building and operating distributed systems at scale; you design for performance and reliability. Experience with schedulers and orchestrators for managing containers and non-containerised workloads (e.g. Nomad, Kubernetes). Excellent problem-solving and communication skills, with a genuine enthusiasm for digging into problems that don't have obvious solutions. Bonus if you: Have worked on low-level virtualization or sandbox execution environments. Have product engineering experience and care about the developer-facing impact of infrastructure decisions. Have experience with on-call operations for large-scale distributed systems. Benefits: Competitive compensation package, including equity. Inclusive Healthcare Package. Learn and Grow - we provide mentorship and send you to events that help you build your network and skills. Flexible Time Off. We will provide you the gear you need to do your role, and a WFH budget for you to outfit your space as needed. Vercel is committed to fostering and empowering an inclusive community within our organization. We do not discriminate on the basis of race, religion, color, gender expression or identity, sexual orientation, national origin, citizenship, age, marital status, veteran status, disability status, or any other characteristic protected by law. Vercel encourages everyone to apply for our available positions, even if they don't necessarily check every box on the job description.
Waymo is expanding its Software Quality Operations (SWQOps) team to scale its autonomous driving efforts, including international expansion to London. You will partner with Engineering to deploy ML/Gen-AI models, drive issue discovery, and establish SOPs that integrate AI insights into vendor workflows. The role emphasizes policy leadership, data-driven monitoring, and cross-functional collaboration. The position targets growth in London/Tokyo data signals and driving rule compliance, with a
17/07/2026
Full time
Waymo is expanding its Software Quality Operations (SWQOps) team to scale its autonomous driving efforts, including international expansion to London. You will partner with Engineering to deploy ML/Gen-AI models, drive issue discovery, and establish SOPs that integrate AI insights into vendor workflows. The role emphasizes policy leadership, data-driven monitoring, and cross-functional collaboration. The position targets growth in London/Tokyo data signals and driving rule compliance, with a
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver-The World's Most Experienced Driver -to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo's fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. The Software Quality Operations (SWQOps) team is at the heart of enabling safe and successful scale-up into new international markets and incubating mission-critical, high-impact workflows. Our mission is to build an adaptable and scalable operation, increasingly powered by AI, to deliver the crucial insights necessary to confidently deploy and grow Waymo's autonomous vehicle service. Why This Team is Essential to Waymo's Success Waymo is undergoing unprecedented growth, rapidly expanding into new cities (targeting 20 new cities by EOY 2026, including international expansion to London and Tokyo) and launching new vehicle platforms. SWQOps, especially our Technical Specialists, plays a critical role in this expansion, making it possible to scale safely and efficiently. We are on the front lines of: International RO & New Market Readiness: Concerted effort to update triage policy, standing Voice of Customer in new countries. Through meticulous triage of driving events, issue discovery, and continuous field monitoring, SWQOps provides early warnings and critical insights to Product and Data Science partners. Supporting the development of a single, automated, end-to-end machine learning flywheel for the entire Waymo Driver. A successful flywheel will be the core engine for scaling our technology, enabling faster international expansion, quicker remediation of driving issues, and a significant reduction in the engineering effort required to maintain and improve the driver. Driving Engineering Velocity: By handling the vital work of performance evaluation, issue deep-dives, and data set curation, SWQOps collaborates heavily and allows Waymo's Engineering, SysEng, Simulation, and Data Science teams to focus on their core tasks of developing and improving the Waymo Driver. Workflow Incubation & Development: Incubating high-impact workflows and ensuring readiness for automation and long-term operational excellence. You will: Partner with Engineering to design, test, and deploy cutting-edge Machine Learning (ML) and Generative AI (Gen-AI) models and tools to drive step-change improvements in issue discovery & detection, triage efficiency, and quality assurance. Provides technical requirements to Engineering for tooling that supports new market workflows (e.g., localization of ML/Gen-AI models, new data pipelines). Leverage AI-powered insights and traditional triage signals to proactively identify emerging on-road issue trends, new risk scenarios, and edge cases. Develop and refine data-driven strategies for issue discovery and monitoring, enhanced by ML model outputs. Serve as the key link between AI/ML development and operational execution. Define and document new policies, guidelines, and Standard Operating Procedures (SOPs) that integrate AI tools and insights into daily vendor workflows. Design and implement robust quality control processes for both human and AI-generated outputs. Designs and validates new triage policies, SOPs, and quality control processes specifically for London/Tokyo driving rules and data signals. Act as the Subject Matter Expert for technical policies and guidelines in the new market. Consult with local/international stakeholders to ensure seamless adoption of technical policies. Provide technical leadership and consultation to stakeholders to enhance our workflows and quality. You'll be at the forefront of identifying and escalating issues with our tools, providing technical requirements to engineering, and driving user testing to support the development and deployment of new tooling features. You have: BS/BA degree or 8+ years of relevant work experience in AV Software Quality Operations Strong understanding of driving rules and regulations in European (London) metropolitan areas. Increased competency in supporting all phases of the machine learning development lifecycle, from data preparation and training to validation, deployment, and continuous monitoring. Experience with ML testing and validation, including dataset quality assurance, bias detection, edge-case scenario testing, and performance evaluation using statistical metrics. Ability to quickly learn and implement new concepts and utilize proprietary tools. Strong understanding of driving rules and regulations. A proven ability to work in a fast-paced, high-stress environment while maintaining good judgment Excellent communication and interpersonal skills to effectively collaborate with a wide range of individuals in a diverse and dynamic work environment We prefer: Demonstrated strong execution with ability to drive outcomes Demonstrated experience working with international and/or offshore vendor teams to drive global process standardization and quality Basic SQL querying Competency in LLM / transformer models, and / or ML for robotics domain experience A greater focus on using your subject matter expertise for results analysis and direct customer consultation in the development of new and improved solutions Self-motivated with basic skills in task planning and time management Experience with project management or program management in a global, distributed environment The expected base salary range for this full-time position is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Waymo employees are also eligible to participate in Waymo's discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements. Salary Range £97,000 - £105,000 GBP
17/07/2026
Full time
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver-The World's Most Experienced Driver -to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo's fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. The Software Quality Operations (SWQOps) team is at the heart of enabling safe and successful scale-up into new international markets and incubating mission-critical, high-impact workflows. Our mission is to build an adaptable and scalable operation, increasingly powered by AI, to deliver the crucial insights necessary to confidently deploy and grow Waymo's autonomous vehicle service. Why This Team is Essential to Waymo's Success Waymo is undergoing unprecedented growth, rapidly expanding into new cities (targeting 20 new cities by EOY 2026, including international expansion to London and Tokyo) and launching new vehicle platforms. SWQOps, especially our Technical Specialists, plays a critical role in this expansion, making it possible to scale safely and efficiently. We are on the front lines of: International RO & New Market Readiness: Concerted effort to update triage policy, standing Voice of Customer in new countries. Through meticulous triage of driving events, issue discovery, and continuous field monitoring, SWQOps provides early warnings and critical insights to Product and Data Science partners. Supporting the development of a single, automated, end-to-end machine learning flywheel for the entire Waymo Driver. A successful flywheel will be the core engine for scaling our technology, enabling faster international expansion, quicker remediation of driving issues, and a significant reduction in the engineering effort required to maintain and improve the driver. Driving Engineering Velocity: By handling the vital work of performance evaluation, issue deep-dives, and data set curation, SWQOps collaborates heavily and allows Waymo's Engineering, SysEng, Simulation, and Data Science teams to focus on their core tasks of developing and improving the Waymo Driver. Workflow Incubation & Development: Incubating high-impact workflows and ensuring readiness for automation and long-term operational excellence. You will: Partner with Engineering to design, test, and deploy cutting-edge Machine Learning (ML) and Generative AI (Gen-AI) models and tools to drive step-change improvements in issue discovery & detection, triage efficiency, and quality assurance. Provides technical requirements to Engineering for tooling that supports new market workflows (e.g., localization of ML/Gen-AI models, new data pipelines). Leverage AI-powered insights and traditional triage signals to proactively identify emerging on-road issue trends, new risk scenarios, and edge cases. Develop and refine data-driven strategies for issue discovery and monitoring, enhanced by ML model outputs. Serve as the key link between AI/ML development and operational execution. Define and document new policies, guidelines, and Standard Operating Procedures (SOPs) that integrate AI tools and insights into daily vendor workflows. Design and implement robust quality control processes for both human and AI-generated outputs. Designs and validates new triage policies, SOPs, and quality control processes specifically for London/Tokyo driving rules and data signals. Act as the Subject Matter Expert for technical policies and guidelines in the new market. Consult with local/international stakeholders to ensure seamless adoption of technical policies. Provide technical leadership and consultation to stakeholders to enhance our workflows and quality. You'll be at the forefront of identifying and escalating issues with our tools, providing technical requirements to engineering, and driving user testing to support the development and deployment of new tooling features. You have: BS/BA degree or 8+ years of relevant work experience in AV Software Quality Operations Strong understanding of driving rules and regulations in European (London) metropolitan areas. Increased competency in supporting all phases of the machine learning development lifecycle, from data preparation and training to validation, deployment, and continuous monitoring. Experience with ML testing and validation, including dataset quality assurance, bias detection, edge-case scenario testing, and performance evaluation using statistical metrics. Ability to quickly learn and implement new concepts and utilize proprietary tools. Strong understanding of driving rules and regulations. A proven ability to work in a fast-paced, high-stress environment while maintaining good judgment Excellent communication and interpersonal skills to effectively collaborate with a wide range of individuals in a diverse and dynamic work environment We prefer: Demonstrated strong execution with ability to drive outcomes Demonstrated experience working with international and/or offshore vendor teams to drive global process standardization and quality Basic SQL querying Competency in LLM / transformer models, and / or ML for robotics domain experience A greater focus on using your subject matter expertise for results analysis and direct customer consultation in the development of new and improved solutions Self-motivated with basic skills in task planning and time management Experience with project management or program management in a global, distributed environment The expected base salary range for this full-time position is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Waymo employees are also eligible to participate in Waymo's discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements. Salary Range £97,000 - £105,000 GBP
Snowflake is seeking an entrepreneurial Staff Applied AI Product Manager to lead the vision and execution for AI Solutions and Services within the Cortex team. You will own end-to-end product lifecycle, understand customer needs, and collaborate with engineering, design, and GTM to deliver simple, powerful AI solutions. You will work at the intersection of customers, advanced engineers, and leadership to translate bespoke deployments into scalable platform features, shaping the AI strategy
17/07/2026
Full time
Snowflake is seeking an entrepreneurial Staff Applied AI Product Manager to lead the vision and execution for AI Solutions and Services within the Cortex team. You will own end-to-end product lifecycle, understand customer needs, and collaborate with engineering, design, and GTM to deliver simple, powerful AI solutions. You will work at the intersection of customers, advanced engineers, and leadership to translate bespoke deployments into scalable platform features, shaping the AI strategy
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver-The World's Most Experienced Driver -to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo's fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. The DUE ML Core London team builds and operates scalable machine learning systems, simulation workflows, and insight tools designed to improve the evaluation and developer onboarding journeys. By combining expert human judgement with advanced machine learning models, we deliver training and evaluation data for hundreds of metrics and components that comprise the Waymo Driver. We are looking for researchers and software engineers passionate about developing ML techniques for evaluation systems and driving performance improvements across our technology stack. You will Build scalable systems for training and fine-tuning large-scale generative models to produce realistic and evaluate interesting driving behaviors. Lead the implementation, and iteration of novel RL algorithms, reward functions, and training paradigms tailored for generating high-fidelity and insightful driving behaviors. Lead the development of cutting-edge deep learning models and generative AI (LLM/VLM) solutions to enhance human-led triaging, introduce automation for high-volume workflows, and perform nuanced analysis of self-driving behavior to detect critical anomalies. Oversee the production and optimization of machine learning models aiming to assess Waymo's expansive fleet of vehicles that cumulatively travel millions of miles. Proactively monitor and assimilate best practices from within Alphabet and the broader industry to develop a novel reinforcement learning from human preference (RLHF) based data collection and evaluation system. Collaborate closely with multiple teams (e.g., Prediction, Planning, Research), other technical leads, and senior leaderships across Waymo to deliver on key strategic efforts. You have M.S. or Ph.D. degree in Computer Science, Machine Learning, Artificial Intelligence, or a related technical field; or equivalent practical experience. 7+ years of hands on experience in developing and applying Machine Learning models, with a significant focus on reinforcement learning. Demonstrated expertise in deep learning, sequence modeling, and generative models. Strong publication record or history of impactful project delivery in RL or related areas. Proficiency in Python and standard ML frameworks (e.g., JAX, TensorFlow). Experience with large scale distributed training and data processing. Proven ability to lead complex and ambiguous technical projects from conception to completion. We prefer 10+ years of relevant experience in ML/RL research and application. Experience in the autonomous vehicles domain, robotics, or complex simulation environments. Deep understanding of state of the art RL techniques, including those used for fine tuning large models (e.g., from human feedback/preferences). Familiarity with large scale simulation platforms and their integration with ML training workflows. Experience designing and using metrics for evaluating complex AI systems. Track record of technical leadership, influencing senior stakeholders, and driving innovation across team boundaries. Excellent communication skills, with the ability to articulate complex technical concepts clearly. Salary Range The expected base salary range for this full time position is listed below. Actual starting pay will be based on job related factors, including exact work location, experience, relevant training and education, and skill level. Waymo employees are also eligible to participate in Waymo's discretionary annual bonus program, equity incentive plan, and generous company benefits program, subject to eligibility requirements. £155,000 - £163,000 GBP
17/07/2026
Full time
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver-The World's Most Experienced Driver -to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo's fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states. The DUE ML Core London team builds and operates scalable machine learning systems, simulation workflows, and insight tools designed to improve the evaluation and developer onboarding journeys. By combining expert human judgement with advanced machine learning models, we deliver training and evaluation data for hundreds of metrics and components that comprise the Waymo Driver. We are looking for researchers and software engineers passionate about developing ML techniques for evaluation systems and driving performance improvements across our technology stack. You will Build scalable systems for training and fine-tuning large-scale generative models to produce realistic and evaluate interesting driving behaviors. Lead the implementation, and iteration of novel RL algorithms, reward functions, and training paradigms tailored for generating high-fidelity and insightful driving behaviors. Lead the development of cutting-edge deep learning models and generative AI (LLM/VLM) solutions to enhance human-led triaging, introduce automation for high-volume workflows, and perform nuanced analysis of self-driving behavior to detect critical anomalies. Oversee the production and optimization of machine learning models aiming to assess Waymo's expansive fleet of vehicles that cumulatively travel millions of miles. Proactively monitor and assimilate best practices from within Alphabet and the broader industry to develop a novel reinforcement learning from human preference (RLHF) based data collection and evaluation system. Collaborate closely with multiple teams (e.g., Prediction, Planning, Research), other technical leads, and senior leaderships across Waymo to deliver on key strategic efforts. You have M.S. or Ph.D. degree in Computer Science, Machine Learning, Artificial Intelligence, or a related technical field; or equivalent practical experience. 7+ years of hands on experience in developing and applying Machine Learning models, with a significant focus on reinforcement learning. Demonstrated expertise in deep learning, sequence modeling, and generative models. Strong publication record or history of impactful project delivery in RL or related areas. Proficiency in Python and standard ML frameworks (e.g., JAX, TensorFlow). Experience with large scale distributed training and data processing. Proven ability to lead complex and ambiguous technical projects from conception to completion. We prefer 10+ years of relevant experience in ML/RL research and application. Experience in the autonomous vehicles domain, robotics, or complex simulation environments. Deep understanding of state of the art RL techniques, including those used for fine tuning large models (e.g., from human feedback/preferences). Familiarity with large scale simulation platforms and their integration with ML training workflows. Experience designing and using metrics for evaluating complex AI systems. Track record of technical leadership, influencing senior stakeholders, and driving innovation across team boundaries. Excellent communication skills, with the ability to articulate complex technical concepts clearly. Salary Range The expected base salary range for this full time position is listed below. Actual starting pay will be based on job related factors, including exact work location, experience, relevant training and education, and skill level. Waymo employees are also eligible to participate in Waymo's discretionary annual bonus program, equity incentive plan, and generous company benefits program, subject to eligibility requirements. £155,000 - £163,000 GBP
About the role As a Model Behavior Architect on the Function Calling team, you are at the forefront of defining and measuring how LLMs use tools, invoke functions, and orchestrate complex agentic workflows. We are looking for people who have built a career in engineering, machine learning, and large language models and are experts in model evaluation, policy writing, and creating eval pipelines for tool use and function calling. Your role is to work hand in hand with our Science team to define what "good" looks like for function calling-from accurate parameter selection and schema adherence to multi step tool orchestration, error recovery, and agentic reasoning. Join us if you are passionate about tackling cutting edge, open ended research challenges and transforming your insights into best in class models. What you will do Interact with models to identify where function calling and tool use behaviour can be improved Gather internal and external feedback on tool calling behaviour to scope areas for improvement Design and implement evals, data guidelines, data generation, and synthetic tool environments and APIs Identify and fix edge case behaviours, such as malformed arguments, hallucinated functions, and incorrect tool selection-through rigorous testing Develop robust evaluation pipelines for the function calling capabilities of our model candidates Work collaboratively with AI Scientists About you You have a deep understanding of either 1) API design, structured outputs, and schema specification (e.g. JSON Schema), 2) engineering and code behaviour, 3) LLM agents at work, including reasoning, planning, and multi step tool use You have prior knowledge in training and optimising model behaviour You are an expert at building robust evaluations You thrive in dynamic and technically complex environments You have a track record of delivering innovative, out of the box solutions to address real world constraints
16/07/2026
Full time
About the role As a Model Behavior Architect on the Function Calling team, you are at the forefront of defining and measuring how LLMs use tools, invoke functions, and orchestrate complex agentic workflows. We are looking for people who have built a career in engineering, machine learning, and large language models and are experts in model evaluation, policy writing, and creating eval pipelines for tool use and function calling. Your role is to work hand in hand with our Science team to define what "good" looks like for function calling-from accurate parameter selection and schema adherence to multi step tool orchestration, error recovery, and agentic reasoning. Join us if you are passionate about tackling cutting edge, open ended research challenges and transforming your insights into best in class models. What you will do Interact with models to identify where function calling and tool use behaviour can be improved Gather internal and external feedback on tool calling behaviour to scope areas for improvement Design and implement evals, data guidelines, data generation, and synthetic tool environments and APIs Identify and fix edge case behaviours, such as malformed arguments, hallucinated functions, and incorrect tool selection-through rigorous testing Develop robust evaluation pipelines for the function calling capabilities of our model candidates Work collaboratively with AI Scientists About you You have a deep understanding of either 1) API design, structured outputs, and schema specification (e.g. JSON Schema), 2) engineering and code behaviour, 3) LLM agents at work, including reasoning, planning, and multi step tool use You have prior knowledge in training and optimising model behaviour You are an expert at building robust evaluations You thrive in dynamic and technically complex environments You have a track record of delivering innovative, out of the box solutions to address real world constraints
About the role As a Model Behavior Architect on the Function Calling team, you are at the forefront of defining and measuring how LLMs use tools, invoke functions, and orchestrate complex agentic workflows. We are looking for people who have built a career in engineering, machine learning, and large language models and are experts in model evaluation, policy writing, and creating eval pipelines for tool use and function calling. Your role is to work hand in hand with our Science team to define what "good" looks like for function calling-from accurate parameter selection and schema adherence to multi step tool orchestration, error recovery, and agentic reasoning. Join us if you are passionate about tackling cutting edge, open ended research challenges and transforming your insights into best in class models. What you will do Interact with models to identify where function calling and tool use behaviour can be improved Gather internal and external feedback on tool calling behaviour to scope areas for improvement Design and implement evals, data guidelines, data generation, and synthetic tool environments and APIs Identify and fix edge case behaviours, such as malformed arguments, hallucinated functions, and incorrect tool selection-through rigorous testing Develop robust evaluation pipelines for the function calling capabilities of our model candidates Work collaboratively with AI Scientists About you You have a deep understanding of either 1) API design, structured outputs, and schema specification (e.g. JSON Schema), 2) engineering and code behaviour, 3) LLM agents at work, including reasoning, planning, and multi step tool use You have prior knowledge in training and optimising model behaviour You are an expert at building robust evaluations You thrive in dynamic and technically complex environments You have a track record of delivering innovative, out of the box solutions to address real world constraints
12/07/2026
Full time
About the role As a Model Behavior Architect on the Function Calling team, you are at the forefront of defining and measuring how LLMs use tools, invoke functions, and orchestrate complex agentic workflows. We are looking for people who have built a career in engineering, machine learning, and large language models and are experts in model evaluation, policy writing, and creating eval pipelines for tool use and function calling. Your role is to work hand in hand with our Science team to define what "good" looks like for function calling-from accurate parameter selection and schema adherence to multi step tool orchestration, error recovery, and agentic reasoning. Join us if you are passionate about tackling cutting edge, open ended research challenges and transforming your insights into best in class models. What you will do Interact with models to identify where function calling and tool use behaviour can be improved Gather internal and external feedback on tool calling behaviour to scope areas for improvement Design and implement evals, data guidelines, data generation, and synthetic tool environments and APIs Identify and fix edge case behaviours, such as malformed arguments, hallucinated functions, and incorrect tool selection-through rigorous testing Develop robust evaluation pipelines for the function calling capabilities of our model candidates Work collaboratively with AI Scientists About you You have a deep understanding of either 1) API design, structured outputs, and schema specification (e.g. JSON Schema), 2) engineering and code behaviour, 3) LLM agents at work, including reasoning, planning, and multi step tool use You have prior knowledge in training and optimising model behaviour You are an expert at building robust evaluations You thrive in dynamic and technically complex environments You have a track record of delivering innovative, out of the box solutions to address real world constraints
We are an applied AI lab building end-to-end software agents. We're the makers of Devin, the first AI software engineer, and Windsurf, the AI-native IDE. Together, they represent our vision for collaborative AI teammates that enable engineers to focus on more interesting problems and empower teams to strive for more ambitious goals. Our team is small and talent-dense. Among our founding team, we have world-class competitive programmers, former founders, and leaders from companies at the cutting edge of AI including Scale AI, Palantir, Cursor, Waymo, Tesla, Lunchclub, Modal, Google DeepMind, and Nuro. Building Devin and Windsurf is just the first step-our hardest challenges still lie ahead. If you're excited to solve some of the world's biggest problems and build AI that can reason on real-world tasks, apply to join us. About the Role Applied AI Transformation Managers are technically oriented strategic advisors and operators who ensure value alignment to our customer's most strategic projects and drive program delivery across several strategic accounts. You'll be responsible for identifying strategic opportunities for our customers to realize value from Cognition's platform and ensuring strong value capture and reporting. You will support delivery of enterprise wide programs to implement agentic AI, specifically by designing and building the operating model and delivery model for agentic AI. You'll act as a trusted advisor to senior and functional leaders, designing value realization targets and delivering transformation programs. This role embeds with a select set of Cognition's strategic accounts for upfront strategy work through in-depth program management and delivery. You will be on the frontlines of supporting customers turning productivity gains and quantifiable financial impact. You will pioneer the playbook for how the world's largest enterprises get value from AI Agents. In this role, you will: Working with customer leadership to identify, quantify and report value targets for productivity gains and financial impact Ensuring seamless operations and delivery for the overall customer program, in close partnership with the Cognition account team Ensuring everyone at customer and Cog has visibility on playbook implementation, consumption metrics and value metrics Maintaining great relationships and customer trust via delivery Oversee onboarding and rollout to thousands of engineers, ensuring successful deployment Apply world class analytical and technical program management skills to support customer executives and teams, particularly with value realization and tracking Build centers of excellence within accounts and empower them through tailored enablement programs Requirements for the role: Computer science or EE undergrad SWE intern or full time work experience 3-5 years at a big 3 strategy consulting firm with a strong background in value targeting and realization and technology strategy as well as a proven track record building and expanding client relationships Ideally 2+ years experience at a startup in a customer facing role preferred Thrive in ambiguous, fast-changing environments-you move quickly and grow quickly Demonstrated ability to learn exceptionally fast You might excel if you have successfully been a part of complex enterprise teams, especially sourcing and managing account growth have worked as a software engineer, SE, technical account manager or other technical role previously founded a startup or were early stage at a high growth startup are a competitive, highly ambitious person who loves working in high-intensity environments Equal Opportunity Cognition is an equal opportunity employer. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected characteristic under applicable law. We are committed to providing reasonable accommodations for candidates with disabilities throughout the hiring process - please let us know if you need any.
08/07/2026
Full time
We are an applied AI lab building end-to-end software agents. We're the makers of Devin, the first AI software engineer, and Windsurf, the AI-native IDE. Together, they represent our vision for collaborative AI teammates that enable engineers to focus on more interesting problems and empower teams to strive for more ambitious goals. Our team is small and talent-dense. Among our founding team, we have world-class competitive programmers, former founders, and leaders from companies at the cutting edge of AI including Scale AI, Palantir, Cursor, Waymo, Tesla, Lunchclub, Modal, Google DeepMind, and Nuro. Building Devin and Windsurf is just the first step-our hardest challenges still lie ahead. If you're excited to solve some of the world's biggest problems and build AI that can reason on real-world tasks, apply to join us. About the Role Applied AI Transformation Managers are technically oriented strategic advisors and operators who ensure value alignment to our customer's most strategic projects and drive program delivery across several strategic accounts. You'll be responsible for identifying strategic opportunities for our customers to realize value from Cognition's platform and ensuring strong value capture and reporting. You will support delivery of enterprise wide programs to implement agentic AI, specifically by designing and building the operating model and delivery model for agentic AI. You'll act as a trusted advisor to senior and functional leaders, designing value realization targets and delivering transformation programs. This role embeds with a select set of Cognition's strategic accounts for upfront strategy work through in-depth program management and delivery. You will be on the frontlines of supporting customers turning productivity gains and quantifiable financial impact. You will pioneer the playbook for how the world's largest enterprises get value from AI Agents. In this role, you will: Working with customer leadership to identify, quantify and report value targets for productivity gains and financial impact Ensuring seamless operations and delivery for the overall customer program, in close partnership with the Cognition account team Ensuring everyone at customer and Cog has visibility on playbook implementation, consumption metrics and value metrics Maintaining great relationships and customer trust via delivery Oversee onboarding and rollout to thousands of engineers, ensuring successful deployment Apply world class analytical and technical program management skills to support customer executives and teams, particularly with value realization and tracking Build centers of excellence within accounts and empower them through tailored enablement programs Requirements for the role: Computer science or EE undergrad SWE intern or full time work experience 3-5 years at a big 3 strategy consulting firm with a strong background in value targeting and realization and technology strategy as well as a proven track record building and expanding client relationships Ideally 2+ years experience at a startup in a customer facing role preferred Thrive in ambiguous, fast-changing environments-you move quickly and grow quickly Demonstrated ability to learn exceptionally fast You might excel if you have successfully been a part of complex enterprise teams, especially sourcing and managing account growth have worked as a software engineer, SE, technical account manager or other technical role previously founded a startup or were early stage at a high growth startup are a competitive, highly ambitious person who loves working in high-intensity environments Equal Opportunity Cognition is an equal opportunity employer. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected characteristic under applicable law. We are committed to providing reasonable accommodations for candidates with disabilities throughout the hiring process - please let us know if you need any.
Neura Market is seeking a Senior Backend Engineer for a remote opportunity, focused on managing alerts in the Grafana open-source project. You will work with a backend team to design and maintain critical systems. The ideal candidate has strong programming skills, is self-motivated, and is passionate about creating user-focused products. The role offers a competitive salary between £91,755 and £110,106, including RSUs for ownership in the company's success.
02/07/2026
Full time
Neura Market is seeking a Senior Backend Engineer for a remote opportunity, focused on managing alerts in the Grafana open-source project. You will work with a backend team to design and maintain critical systems. The ideal candidate has strong programming skills, is self-motivated, and is passionate about creating user-focused products. The role offers a competitive salary between £91,755 and £110,106, including RSUs for ownership in the company's success.