Senior Research Scientist - AI Safety Evaluations

Faculty
London

Why Faculty?


We established Faculty in 2014 because we thought that AI would be the most important technology of our time. Since then, we’ve worked with over 350 global customers to transform their performance through human-centric AI. You can read about our real-world impact here .

We don’t chase hype cycles. We innovate, build and deploy responsible AI which moves the needle - and we know a thing or two about doing it well. We bring an unparalleled depth of technical, product and delivery expertise to our clients who span government, finance, retail, energy, life sciences and defence.

Our business, and reputation, is growing fast and we’re always on the lookout for individuals who share our intellectual curiosity and desire to build a positive legacy through technology.

AI is an epoch-defining technology, join a company where you’ll be empowered to envision its most powerful applications, and to make them happen.

About the Team

We are the central research and development team within Faculty's broader AI safety charter. Our group focuses on fundamental and applied technical AI safety research, producing rigorous scientific outputs - publications, tooling, technical reports & evaluations - that advance the theory, practice and understanding of AI risks, and directly inform the work of frontier AI labs, government agencies, and national security institutes.

Our research spans fundamental research using black-box and white-box approaches to understand and steer AI systems, to building and advancing safety evaluations which better understand and quantify risks in AI. We care deeply about mechanistic understanding and scientific rigour in measuring risks in AI systems. Current research threads include (but are not limited to) uncertainty calibration, goal drift, misinformation mitigation, steering vectors, robust safeguard measurement and the science of evaluations.

We collaborate closely with the Faculty's wider AI safety team, a group with a well-established track record in capability evaluations and red-teaming for misuse risk across CBRN, cybersecurity, societal and psychosocial harms. That work has been conducted for several leading frontier model developers and national safety institutes, and has been featured in model cards and safety reports from Anthropic, GDM, Meta & OpenAI.

About the role

As a Senior Research Scientist at Faculty, you will lead the development of cutting-edge safety evaluations to quantify AI risks in critical domains like CBRN and Cyber. Joining our high-impact R&D team, you will drive original research that advances safety methodology while collaborating with delivery teams building evaluations and red-teaming for frontier labs. This is a high-agency opportunity to conduct technical AI safety research that produces scientific outputs shaping the future of safe real-world AI deployment.

What you'll be doing

  • Leading the development of novel safety evaluations in high-impact domains such as CBRN and Cyber to quantify emerging risks.

  • Executing original technical research in AI safety evaluation methods, taking ideas from concept to publication.

  • Shaping the R&D agenda by identifying strategic opportunities to advance safety evaluation methodology across Faculty and the broader ecosystem.

  • Contributing thought leadership and deep technical expertise to client delivery projects, evaluation work, and red-teaming for frontier labs.

  • Representing Faculty’s scientific leadership through active external engagement with the global research community, frontier labs, and government stakeholders.

Who we're looking for

  • A track record of owning research end-to-end —from identifying novel problems to publication—driven by scientific curiosity and tenacity.

  • Hands-on experience designing and building AI evaluations or benchmarks, alongside a strong ability to reason about construct validity and mitigate confounds.

  • Expertise in red-teaming , adversarial testing, jailbreaking, indirect prompt injection, and assessing the robustness of model safeguards.

  • Strong foundational skills in experimental design, statistical analysis, and uncertainty quantification, including Bayesian methods.

  • Deep knowledge of language models, generative AI architectures, training methodologies, and safety mitigation techniques.

  • Solid Python proficiency combined with the engineering discipline required to build robust, reproducible research.

Nice to haves

  • Experience in threat and risk modelling.

  • Background or knowledge in high-risk domains such as CBRN or Cybersecurity.

  • A history of high-impact AI research, evidenced by top-tier publications or equivalent practical achievements.

Our Recruitment Ethos

We aim to grow the best team - not the most similar one. We know that diversity of individuals fosters diversity of thought, and that strengthens our principle of seeking truth. And we know from experience that diverse teams deliver better work, relevant to the world in which we live. We’re united by a deep intellectual curiosity and desire to use our abilities for measurable positive impact. We strongly encourage applications from people of all backgrounds, ethnicities, genders, religions and sexual orientations.

Some of our standout benefits:

  • Unlimited Annual Leave Policy

  • Private healthcare and dental

  • Enhanced parental leave

  • Family-Friendly Flexibility & Flexible working

  • Sanctus Coaching

  • Hybrid Working

If you don’t feel you meet all the requirements, but are excited by the role and know you bring some key strengths, please don't hesitate in applying as you might be right for this role, or other roles. We are open to conversations about part-time hours.

Posted 2026-07-27

Recommended Jobs

Competition Litigation Paralegal

Michael Page
City of London, Greater London

Support lawyers on a range of competition and commercial litigation matters Assist with disclosure exercises and document review projects Manage and organise large volumes of case documentation…

View Details
Posted 2026-07-25

CRM Executive

Harvey Nichols
London

Here at Harvey Nichols we are looking for a CRM Executive to join our Marketing Team at Head Office - the home of modern luxury, exceptional client experiences and some of the world's most coveted br…

View Details
Posted 2026-07-25

Geography Teaching Role in Enfield (Independent School)

Marchant Recruitment
Enfield, Greater London

School Status & Location Sector: Leading Independent School, Outer London. Borough: Enfield. Start Date: Permanent, full-time role commencing January 2026. The Opportunity & School Prof…

View Details
Posted 2025-11-13

BIM Technician - London

Johnson BIM
London

We are looking for an ‘on-the-way-up’, client facing, engaging, passionate BIM technician who has been involved in v alidating BIM models derived from point cloud data, ensuring quality, accuracy, a…

View Details
Posted 2026-03-10

Senior Product Designer (FTC) (Hiring Immediately)

AKQA
London

At AKQA (part of WPP), we believe in the imaginative application of art and science to create beautiful ideas, products and services. We work at the intersection of design, technology and culture to …

View Details
Posted 2026-06-03

Tax Senior

London

Our client is a Leading London Firm, who prize themselves on meeting the requirements of their varying clients and the development of their staff.  The Mixed Tax Senior role will be based in the …

View Details
Posted 2025-09-10

Senior Policy Lawyer

Michael Page
London

Lead and develop a team of lawyers delivering critical policy and legal workstreams linked to complex enforcement cases Provide expert legal and policy advice to support investigation teams on hig…

View Details
Posted 2026-05-27

Science Technician - Phenomenal School - Ealing

Marchant Recruitment
London

Forward-thinking Secondary School in Ealing with a strong commitment to high-quality science education. • Full-time Science Technician required from January 2026 • Well-resourced Science departm…

View Details
Posted 2026-01-27

Automation Engineer

level-zero-health
London

About Level Zero Level Zero is pioneering next-generation biosensor technology for precise, real-time monitoring of critical endocrine biomarkers. Recognised as one of the top 20 startups across a…

View Details
Posted 2026-07-21