The OpenInterpretability research arc on why capable LLM agents loop forever and never finish — and whether their internals can tell us, or change it. All on Qwen3.6-27B over SWE-bench Pro, with cross-architecture replications (Mistral, Llama-3.1, gpt-oss-20b). Author: Caio Vicentino (OpenInterpretability, ORCID 0009-0003-4331-6259), CC-BY-4.0.
Thesis: interpretability as AUDIT — the rigor that tells a real signal from a confound. The arc… See the full description on the dataset page:
https://huggingface.co/datasets/caiovicentino1/wandering-arc-papers.