
In-Process Retrieval as Extended Working Memory for AI Agents
This research proposes moving memory retrieval inside the agent's reasoning loop to eliminate network latency. By using an in-process store, retrieval time drops from milliseconds to microseconds, enabling agents to maintain a persistent, high-speed working memory.



