ELLIS Institute Tübingen · Max Planck Institute for Intelligent Systems
Group page · Updates · Datasets
AI safety · alignment · evaluation
We develop algorithmic approaches to reduce harms from increasingly capable general-purpose AI systems. Our work focuses on the alignment and evaluation of autonomous language-model agents, frontier-model risks and capabilities, and model generalisation and steerability.
The AI Safety course at the University of Tübingen is openly available.
Browse all repositories or follow the group on Substack.