arXivYuqiao Tan, Shizhu He, Jun Zhao, Kang LiuTue, Sep 8, 2026, 10:45 AM PDT
score 17.1
Benchmark tests AI agents' ability to autonomously probe AI models
Original: SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
Source: arxiv.org ↗
Writing ELI5 summary…