| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
ELLIS Institute Tübingen · Max Planck Institute for Intelligent Systems
Group page · Updates · Datasets
AI safety · alignment · evaluation
We develop algorithmic approaches to reduce harms from increasingly capable general-purpose AI systems. Our work focuses on the alignment and evaluation of autonomous language-model agents, frontier-model risks and capabilities, and model generalisation and steerability.
The AI Safety course at the University of Tübingen is openly available.
Browse all repositories or follow the group on Substack.
Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours
AI safety course at the University of Tübingen (Summer Semester 2026)
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
Benchmarking Open-Ended Inference Optimization by AI Agents
Agent Skills Enable a New Class of Realistic and Trivially Simple Prompt Injections
Decomposing and measuring evaluation awareness in existing benchmarks and our proposed EvalAwareBench.
Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |