
research note
Peak-Then-Collapse and the Four Interface Channels of Knowledge-Graph Tool Use
This paper investigates the limits of reinforcement learning with verifiable rewards (RLVR) for teaching large language models (LLMs) to use knowledge graph (KG) navigation tools under a deliberate…










