papersSEP 12 04:00 UTC
Layerwise study probes how LLMs route queries and draw on internal knowledge
A new arXiv paper examines how much a language model relies on query-routing signals versus stored knowledge as it produces an answer. The authors apply layerwise interventions to the hidden state at the end of the question and test the approach across several model families, including Qwen and Llama. The work aims to clarify where and when internal knowledge is retrieved during generation.