Apollo Research
London & San Francisco
Research Science
Posted 6 months ago
We’re looking for Research Scientists/Engineers for our pre-deployment team to work on Training-Run Assessments (TRAs). You will design and build automated pipelines for assessing whether egregious misalignment or scheming are emerging at any point of frontier post-training.
This will involve evaluating and red-teaming of checkpoints at various stages of post-training as well as automated analysis of post-training data.
You will get to work with frontier labs like OpenAI, Anthropic, and Google DeepMind and be among the first to interact with new models before anyone else. Our ideal candidate loves rigorously testing frontier AI models, and enjoys building efficient pipelines for automated analysis.
KEY RESPONSIBILITIES
Not ready to apply?
Hiring for a role like this?
Reach AI professionals browsing the board - your listing goes live instantly.
Stay ahead of the curve. Get new AI jobs in your inbox.