Florian Mai
Welcome
I am the Junior Research Group Leader of the mAI-alignment group at the CAISA group at University of Bonn as part of The Lamarr Institute for Machine Learning and Artificial Intelligence.
My current research focuses on AI alignment and safety issues, exploring how to ensure that current and future advanced AI systems are beneficial and safe for humanity.
Read more about my research and background →
News
For the latest updates from my group, see the mAI alignment lab news page →
Selected Publications
In-Training Defenses against Emergent Misalignment in Language Models
ICML, 2026
We study practical in-training safeguards against emergent misalignment, evaluating regularization, safe subspaces, and data interleaving.
ICML, 2026
We study practical in-training safeguards against emergent misalignment, evaluating regularization, safe subspaces, and data interleaving.
AI Alignment Strategies from a Risk Perspective: Independent Safety Mechanisms or Shared Failures?
IASEAI'26: International Association for Safe and Ethical AI Conference, 2026
We analyze overlap in failure modes across alignment techniques to assess the limits of defense-in-depth risk mitigation.
IASEAI'26: International Association for Safe and Ethical AI Conference, 2026
We analyze overlap in failure modes across alignment techniques to assess the limits of defense-in-depth risk mitigation.
Learning to Plan for Language Modeling from Unlabeled Data
COLM, 2024
We propose a method to learn planning for language modeling using unlabeled data.
COLM, 2024
We propose a method to learn planning for language modeling using unlabeled data.
