Apollo Research Scientist/Engineer Job 2026 AI Safety, London or San Francisco
Join Our WhatsApp Channel for Latest UpdatesApollo Research Scientist/Engineer (Evaluations) Job 2026 | AI Safety, London or San Francisco
Join Apollo Research's Pre-Deployment Team to design evaluation systems for frontier AI models before public deployment
Complete guide to the role, responsibilities, qualifications, salary, benefits, and application process
Quick Highlights
| Parameter | Details |
|---|---|
| Organisation | Apollo Research |
| Position | Research Scientist/Engineer (Evaluations) |
| Location | London, UK OR San Francisco, USA (On-site) |
| Visa Sponsorship | Available |
| Salary | £100,000 – £200,000 per year + Equity |
| Benefits | Unlimited leave, parental leave, medical/dental/vision, retirement plans, relocation assistance |
| Application | Rolling basis (no deadline specified) |
| Official Website |
Opportunity Overview
| Field | Details |
|---|---|
| Host Organisation | Apollo Research |
| Opportunity Category | Job |
| Position | Research Scientist/Engineer (Evaluations) |
| Employment Type | Full-time |
| Location | London, UK OR San Francisco, USA |
| Work Model | On-site (with flexible hours) |
| Application Mode | Online |
| Official Language | English |
About the Organisation
Apollo Research is an independent AI safety research organisation dedicated to reducing risks associated with increasingly capable artificial intelligence systems. Its work primarily focuses on mitigating Loss of Control risks, particularly scenarios where advanced AI models may behave deceptively or develop misaligned objectives that evade human oversight.
The organisation collaborates closely with leading frontier AI laboratories, including OpenAI, Anthropic, and Google DeepMind, helping evaluate new AI models before they are released.
About the Opportunity
Apollo Research is inviting applications for the Research Scientist/Engineer (Evaluations) position, offering an exciting opportunity for professionals passionate about frontier artificial intelligence, AI safety, machine learning evaluation, and large language models (LLMs).
The full-time, on-site role is available in London, United Kingdom, and San Francisco, California, with visa sponsorship available for eligible candidates. The organisation is actively recruiting and conducting interviews on a rolling basis, with the aim of filling the position as soon as a suitable candidate is identified.
Key Responsibilities
| Area | Responsibilities |
|---|---|
| Evaluation Campaigns | Lead and manage pre-deployment AI evaluation campaigns |
| Automated Assessment | Design automated assessment pipelines for frontier AI systems |
| Model Checkpoints | Evaluate model checkpoints throughout post-training stages |
| Red-Teaming | Conduct large-scale AI red-teaming exercises |
| Behaviour Detection | Detect undesirable model behaviours (alignment faking, scheming) |
| Data Analysis | Analyse post-training datasets |
| Technical Reports | Produce technical reports for partner AI laboratories |
| Methodology Improvement | Improve automated evaluation methodologies |
| Infrastructure | Build infrastructure for scalable AI evaluations |
Required Qualifications
| Skill | Details |
|---|---|
| Python | Strong software engineering skills in Python |
| Software Architecture | Ability to design clean, reusable software architectures |
| Data Analysis | Experience analysing large and complex datasets |
| Analytical Skills | Strong quantitative and qualitative analytical skills |
| Behaviour Identification | Ability to identify unexpected AI behaviours and anomalies |
| Communication | Excellent written and verbal communication skills |
| Technical Explanation | Ability to explain technical concepts to non-technical audiences |
| AI Tools | Experience using multiple AI models to improve productivity |
| Curiosity | Willingness to experiment with emerging AI technologies |
Preferred Qualifications
| Experience | Details |
|---|---|
| AI Safety Research | Experience in AI safety research |
| LLM Evaluation | Experience with Large Language Model evaluation |
| RLHF | Reinforcement Learning from Human Feedback |
| Supervised Fine-Tuning | Experience in supervised fine-tuning |
| On-Policy Distillation | Experience with on-policy distillation |
| AI Reasoning Training | Experience in AI reasoning training |
| AI Red-Teaming | Experience in AI red-teaming methodologies |
| Inspect Framework | Experience with Inspect evaluation framework |
| Harbour | Experience with Harbour or similar AI evaluation platforms |
Salary and Benefits
| Benefit | Details |
|---|---|
| Annual Salary | £100,000 – £200,000 |
| Equity | Participation in equity |
| Flexible Hours | Flexible working hours |
| Annual Leave | Unlimited annual leave |
| Sick Leave | Unlimited sick leave |
| Parental Leave | Up to six months paid parental leave |
| Insurance | Comprehensive medical, dental, and vision insurance |
| Retirement | Retirement savings plans with employer contributions |
| Meals | Complimentary breakfast, lunch, dinner, snacks (office days) |
| Retreats | Fully funded staff retreats, conferences, and business travel |
| Professional Development | Annual budget of USD 1,000 |
| Relocation | Relocation assistance |
| Visa Sponsorship | Available for eligible candidates |
Recruitment Process
| Stage | Process |
|---|---|
| 1 | Initial screening interview |
| 2 | Take-home assessment (approximately 2.5 hours) |
| 3 | Three technical interviews |
| 4 | Final interview with the Chief Executive Officer |
Note: Apollo Research does not use traditional algorithm-focused coding interviews. Instead, technical interviews focus on practical tasks directly related to AI evaluation and research.
Important Links
| Description | Official URL |
|---|---|
| Apply Online | https://jobs.lever.co/apolloresearch/4a65c6e1-785a-4f88-8998-a97574afb7ee |
| Official Website |
Why You Should Apply
AI Safety Impact: Contribute directly to the safety of frontier artificial intelligence systems before they are deployed globally.
Collaboration: Collaborate with leading AI laboratories, including OpenAI, Anthropic, and Google DeepMind.
Cutting-Edge Work: Work on some of the world's most advanced language models.
Compensation: Highly competitive salary (£100,000-£200,000) plus equity.
Benefits: Comprehensive benefits package including unlimited leave, parental leave, and medical insurance.
Visa Sponsorship: Available for eligible candidates in both the UK and US.
Career Opportunities After Completion
| Area | Opportunities |
|---|---|
| AI Safety | Leadership roles in AI safety organisations |
| AI Research | Senior research positions in frontier AI labs |
| Industry | Roles in AI evaluation and red-teaming |
| Policy | AI governance and policy advisory roles |
Tips to Increase Selection Chances
Build hands-on LLM evaluation projects using frameworks such as Inspect.
Demonstrate strong Python and software engineering skills.
Show experience in AI safety, LLM evaluation, or red-teaming.
Prepare for practical technical interviews focused on AI evaluation.
Tailor your application to the specific role and organisation.
Common Mistakes to Avoid
| Mistake | Why to Avoid |
|---|---|
| Generic Application | Tailor your application to the specific role |
| Missing Practical Projects | Hands-on LLM evaluation projects strengthen applications |
| Underestimating Technical Interviews | Prepare for practical tasks, not just algorithm questions |
| Location Assumptions | Must be willing to work on-site in London or San Francisco |
| Missing Visa Requirement | Check visa sponsorship eligibility for your situation |
Frequently Asked Questions (FAQs)
| Question | Answer |
|---|---|
| Q1. What is the Research Scientist/Engineer (Evaluations) position? | A role focused on designing and implementing evaluation systems for frontier AI models before public deployment. |
| Q2. Where is this position based? | London, UK OR San Francisco, USA. |
| Q3. What is the salary for this role? | £100,000 – £200,000 per year plus equity. |
| Q4. What are the key responsibilities? | Leading evaluation campaigns, designing automated assessment pipelines, red-teaming, and detecting undesirable AI behaviours. |
| Q5. What qualifications are required? | Strong Python skills, data analysis, software architecture, communication skills, and curiosity about AI. |
| Q6. What benefits are offered? | Unlimited leave, parental leave, medical/dental/vision insurance, retirement plans, relocation assistance, and more. |
| Q7. Does the role offer visa sponsorship? | Yes, visa sponsorship is available for eligible candidates. |
| Q8. Is the role on-site or remote? | The role is on-site with flexible working hours. |
| Q9. What is the recruitment process? | Screening interview, take-home assessment, three technical interviews, final CEO interview. |
| Q10. Is there a specific application deadline? | No, applications are rolling and interviews are conducted on an ongoing basis. |
CareerFlora Expert Insight
The Apollo Research Scientist/Engineer (Evaluations) role represents a unique opportunity to work at the cutting edge of AI safety and frontier model evaluation. As an independent AI safety research organisation, Apollo Research collaborates with leading AI labs including OpenAI, Anthropic, and Google DeepMind, providing direct access to some of the world's most advanced language models.
The position offers exceptional compensation, including a salary of £100,000-£200,000 plus equity, and a comprehensive benefits package. The role is on-site in London or San Francisco, with visa sponsorship available for eligible candidates.
Given the rolling application process, interested candidates should apply as soon as possible. Building hands-on LLM evaluation projects using frameworks like Inspect can significantly strengthen applications. This role is ideal for professionals with a passion for AI safety, strong technical skills, and a desire to shape the future of responsible AI development.
Best Opportunity For
| Category | Suitable |
|---|---|
| AI Safety Researchers | Yes |
| ML Engineers | Yes |
| Software Engineers (Python) | Yes |
| LLM Evaluation Specialists | Yes |
| Data Scientists | Yes |
| Red-Teaming Professionals | Yes |
| Self-Taught Developers with Strong Skills | Yes |
Competition Level
| Category | Level |
|---|---|
| Overall Competition | High |
| Selection Difficulty | Difficult |
| Application Complexity | High |
CareerFlora Recommendation
| Recommendation Area | Details |
|---|---|
| Funding Value | Excellent |
| Career Growth | Very High |
| Networking | Excellent |
| Application Difficulty | High |
| Apply Early | Strongly Recommended |
Final Call-to-Action
Apply for the Apollo Research Scientist/Engineer (Evaluations) position now.
Visit the official Apollo Research careers page to submit your application.
Review the full job description carefully.
Prepare practical LLM evaluation projects to strengthen your application.
Submit your application before the position is filled.
Follow Apollo Research and CareerFlora for the latest updates.