The Frontier
Sparse weight decomposition — LLM circuits from under 1% of the data
Mechanistic interpretability has always had a cost problem: explaining a model meant training a second model. A new preprint from IQuest Research and collaborators at Oxford, Stanford, Tsinghua, and the Safe AI Forum argues the circuits are already sitting in the weights — you just have to pull them out. A