Henry Papadatos

Executive Director, SaferAI

prof_pic.jpg

Hi, I’m Henry,

As a technical AI risk management expert, my focus lies in addressing the growing gap between rapid AI advancement and our ability to manage its associated risks.

I am the Executive Director of SaferAI where I work on both AI governance and technical solutions to improve the risk management for frontier AI systems.

Governance

I am contributing to various initiatives in the governance space. I was in an expert working group for the EU AI Act’s General-Purpose AI Code of Practice, focused on the risk taxonomy and on risk identification and assessment. My team took part in all four working groups.

I also took part in the OECD task force in charge of drafting the G7 Hiroshima AI Process reporting framework. Just before the Paris AI action summit, I advocated for international AI safety standards through an op-ed published in TIME.

Technical research

My team and I are building Europe’s capacity to evaluate frontier AI, working directly with model developers to evaluate their systems.

We also model AI risks across cyber, CBRN, and loss of control, with several AI safety institutes and the European Commission. Our work connects empirical measures of AI capabilities to estimates of real-world harm.

We built an AI risk management tracker for AI developers, featured in TIME and Euractiv, then a risk management framework that brings proven practices from other industries into AI.

Prior to joining SaferAI, I conducted technical research on large language models alignment at the Center for Human-Compatible AI at UC Berkeley.

Selected publications

  1. Report
    GLM-5.2 Risk Evaluation Report
    Chinmayi Dixit*, Jacob Davies*, Jasmine Li, Ben Snodin, Jack Kengott, and 1 more author
    2026
  2. Paper
    Lessons from External Review of DeepMind’s Scheming Inability Safety Case
    Stephen Barrett, Francisco Javier Campos Zabala, Sean P. Fillingham, Umair Siddique, James Walpole, and 2 more authors
    TAIGR Workshop, ICML, 2026
  3. Paper
    Toward Quantitative Modeling of Cybersecurity Risks Due to AI Misuse
    Steve Barrett, Malcolm Murray, Otter Quarks, Matthew Smith, Jakub Kryś, and 15 more authors
    2025
  4. Paper
    A Frontier AI Risk Management Framework: Bridging the Gap Between Current AI Practices and Established Risk Management
    Simeon Campos*Henry Papadatos*, Fabien Roger, Chloé Touzet, Otter Quarks, and 1 more author
    Conference on frontier AI safety frameworks (2024), 2025
  5. Paper
    Linear Probe Penalties Reduce LLM Sycophancy
    Henry Papadatos, and Rachel Freedman
    SoLaR workshop, NeurIPS, 2024