Research investigates the ability of large language models to introspect on their internal states. The study employs injected representations of known concepts into model activations to differentiate genuine introspection from confabulations.
arxiv.org
2 min
8/11/2026
Research investigates the ability of large language models to introspect on their internal states. The study employs injected representations of known concepts into model activations to differentiate genuine introspection from confabulations.
arxiv.org
2 min
8/11/2026
Research investigates the ability of large language models to introspect on their internal states. The study employs injected representations of known concepts into model activations to differentiate genuine introspection from confabulations.
arxiv.org
2 min
8/11/2026
No more articles to load