Research investigates the ability of large language models to introspect on their internal states. The study employs injected representations of known concepts into model activations to differentiate genuine introspection from confabulations.
arxiv.org
2 min
17h ago
Research investigates the ability of large language models to introspect on their internal states. The study employs injected representations of known concepts into model activations to differentiate genuine introspection from confabulations.
arxiv.org
2 min
17h ago
Research investigates the ability of large language models to introspect on their internal states. The study employs injected representations of known concepts into model activations to differentiate genuine introspection from confabulations.
arxiv.org
2 min
17h ago
No more articles to load