H is hiring to advance agentic AI by building and refining large language and vision-language models. You will lead research and engineering efforts across data pipelines, backends, and multi-node distributed training to push capabilities in complex environments.
The role emphasizes SFT, RLHF/RLVR and reward modelling, with a focus on practical deployment and cross-functional collaboration in a hybrid London/Paris setting.