Category AI Safety

Using J-space for monitoring misalignment, detecting off-rail behavior, sandbagging, jailbreaks, scalable oversight, implications for future safety research.

Jacobian Lens J-Lens Claude

Abstract Jacobian lens visualization

On July 6, 2026, Anthropic published a paper that changed how we think about what language models know but never say aloud. The jacobian lens—named after the Jacobian matrix in calculus—gives researchers a tool to read the silent, internal neural…