HHiring Reality
← AI Security Institute

Alignment Red Team - Research Engineer/Research Scientist

AI Security Institute
Location
London, UK
Posted
5 days ago
Department
Red Team
What they actually want (must-haves)
  • Ability to work autonomously on complex research projects involving substantial engineering.
  • Completed at least one significant research project in AI safety, security or alignment involving engineering, experiment design and analysis on frontier LLMs.
  • Strong software engineering and ML experience writing complex projects involving language models and ML, beyond just research code.
  • 1+ years professional experience programming in Python for ML or SWE work.
  • Experience writing clean, documented research code for machine learning experiments, including experience with ML frameworks like PyTorch or evaluation frameworks like Inspect.
  • Proven ability in a team environment – flexible, adaptive to needs, and willing to contribute wherever necessary.
Nice to have
  • Ability to make high-quality decisions by identifying risks and testing assumptions, demonstrated through strong prioritisation of research projects using clear, systematic criteria.
  • Familiarity with alignment literature, current methods for post-training and aligning LLMs, loss-of-control risks and threat models.
  • High-quality research papers (first author at top ML venues such as NeurIPS, ICLR or ICML), particularly in relevant areas.
  • Professional experience working on alignment or evaluations, especially at frontier labs or other frontier 3rd party evaluators.
  • Strong open-source software projects, particularly related to LLMs.
What the job really is

The Alignment Red Team at the AI Security Institute is looking for Research Engineers and Research Scientists to focus on detecting and evaluating misalignment in frontier AI systems. Daily tasks include researching methods for identifying misalignment, conducting pre-deployment evaluations, and contributing to public-facing research publications. The role also involves designing software tools for alignment evaluations and mentoring external collaborators.

Benefits
  • Direct influence on how frontier AI is governed and deployed globally.
  • Opportunity to shape the first & best-resourced public-interest research team focused on AI security.
  • Pre-release access to multiple frontier models and ample compute.
  • Extensive operational support so you can focus on research and ship quickly.
  • If you’re talented and driven, you’ll own important problems early.
  • 5 days off and annual stipends for learning and development, and funding for conferences and external collaborations.
Things to weigh
  • No specific mention of the technologies or frameworks beyond Python and ML frameworks, which may limit insight into the tech stack.
  • Candidates should be prepared for a potentially rigorous interview process with multiple stages, including technical assessments and interviews with senior leadership.
  • The role requires a high level of autonomy and may involve complex problem-solving without direct supervision, which may not suit everyone.
Job score4/5
Benefits5/5
Freshness5/5
Career value5/5
Role clarity5/5
Pay transparency0/5

Applying to AI Security Institute?

See how your résumé matches this role — and tailor it from what actually gets interviews.