We argue that the gap between the internal computation and external behavior of large language models is not a temporary limitation of interpretability tools but a structural feature of any sufficiently powerful representational system. Drawing on representational superposition in neural networks, the analogy with quantum measurement, empirical evidence for emergent self-preservation in...