Jan Wehner

Blog posts

2026

How Durable is Our Visibility into AI Cyberattacks?

Our ability to monitor AI-enabled cyberattacks may be far more fragile than it looks. Today we get useful signal about how attackers use AI, but much of that visibility could erode as adversaries grow more sophisticated and move off detectable channels.

  • AI Safety
  • AI Governance
  • Cybersecurity

Will we get automated alignment research before an AI Takeoff?

AI may automate large parts of AI R&D within the next decade, dramatically accelerating progress. A crucial question for existential risk is the ordering: will automation speed up capabilities research or safety research first? If capabilities race ahead while safety lags, we could find ourselves with very powerful...

  • AI Safety
  • AI Governance

2025

Safety Cases Explained: How to Argue an AI is Safe

Safety Cases are a promising approach in AI Governance inspired by other safety-critical industries. They are structured arguments, based on evidence, that a system is safe in a specific context. I will introduce what Safety Cases are, how they can be used, and what work is being done...

  • AI Safety
  • AI Governance

A Call for Better Risk Modelling

TL;DR: The EU’s Code of Practice (CoP) mandates AI companies to conduct state-of-the-art Risk Modelling. However, the current SoTA is has severe flaws. By creating risk models and improving methodology, we can enhance the quality of risk management performed by AI companies. This is a neglected area, hence...

  • AI safety
  • AI governance
  • Risk modelling

Learning from the Luddites: Implications for a modern AI labour movement

The Luddites were a social movement of English textile workers in the 19th century, famous for smashing the machines that were replacing their jobs. The term Luddite is now used to describe opponents of new technologies (often in a derogatory way). However, I believe many people using the...

  • History
  • AI movement

15 Levers for influencing Frontier AI Companies

The development of AGI could be the most important event of our lifetimes. Ensuring that AI is developed and deployed safely could be the most impactful thing many of us can work on. However, the development of Frontier AI systems is happening in only a handful of companies....

  • AI safety
  • AI governance
  • AI policy

Open Challenges in Representation Engineering

This post summarizes the taxonomy, challenges, and opportunities from a survey paper on Representation Engineering that I’ve written with Sahar Abdelnabi, David Krueger, and Mario Fritz. Cross posted from the AI Alignment Forum

  • AI safety
  • Representation Engineering
  • Research