PaperScope
LIVE · 2026-09-11 05:40 UTC

From Parameters to Answers: How LLMs Retrieve and Use Their Internal Knowledge

Wenkang Wei, Yuan Fang, Renhe Jiang, Hong Cheng, Xingtong Yu

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.11859 v1
Category
Submitted
2026-09-10

Abstract

How does a language model's dependence on query-routing information and target knowledge change as it answers a question? We study this question through layerwise interventions on the hidden state at the end of the question. Across Qwen, Llama, and Gemma, we compare country-continent questions with noun, adjective, and code answers while keeping several fitted measurements distinct. A pair-conditioned request direction describes which country is queried in natural single-country questions; a global request direction describes first- versus second-country requests in paired questions; separate selection candidates test control among contents already available in the hidden state. A diagnostic reanalysis of frozen Qwen natural-question states shows that the pair-conditioned direction grows stronger before interventions on it begin to alter later fitted knowledge, with this causal window opening while answer-supporting content is still forming. The paired three-model trajectories are not uniform: Gemma shows a partially overlapping mid-layer routing-content profile, whereas Llama has no sustained routing-effect window under the same gates. In the paired protocol, dependence on the global request direction decreases from fixed earlier to later layer sets while dependence on fitted content persists. A matched Qwen comparison shows that the pair-conditioned direction retains a late effect, so this operational handoff concerns the global fitted direction rather than all request information. These results separate early readability, natural strength, causal steering, and later content dependence.

Comment: 53 pages, 13 figures, including appendices

arXiv abs page · PDF · same-day batch