arxiv:2606.10747
Lukas Galke Poech
lgalke
AI & ML interests
LLM interpretability, agentic/multi-agent safety
Recent Activity
liked a model about 6 hours ago
danish-foundation-models/DFM-Mimir authored a paper 2 months ago
The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment