importantSYS.SOURCE: arXiv• 2026-08-11T21:14:10Z
Emergent Introspective Awareness in Large Language Models
This study investigates whether large language models can introspect on their internal states by injecting known concepts into activations and measuring self-reported state changes. Results show some models can detect injected concepts, recall prior representations, and modulate activations when prompted, indicating limited functional introspective awareness.
*** END OF TRANSMISSION ***