Anthropic researchers have discovered a hidden space within its Claude model, dubbed “J-space,” containing words that influence the model’s reasoning without appearing in its output. Senior editor Will Douglas Heaven writes that while this represents genuine progress in mechanistic interpretability research, the finding should be understood as one more step toward understanding the technology rather than an immediate solution to AI opacity. Heaven cautions against anthropomorphizing AI systems, noting that large language models are “not brains” despite the convenient but potentially misleading terminology commonly used to describe them.