Research explainers on J-Space, the Jacobian Lens, and interpretability inside modern LLMs.
Deep dives on Anthropic’s interpretability work, global workspace ideas in LLMs, and what J-Space means for understanding model behavior.