Marius Hobbhahn
x · Follow ↗
Evals & red teamingAlignment & interpretabilityGovernance & standards
Model evaluations, deception, preparedness and governance
- Type
- Individual expert
- Priority
- Tier 1 · core
- Perspective
- Evaluations / research
- Best for
- Evaluation results and methodological debate
- Publishes
- Frequent
- Editor's note
- Included in the AI Safety Guide expert directory.