Line of inquiry
Inquiring lines›How do agents behave and coordinat…›How can we measure and understand…›this line of inquiry
How should agents structure and manage memory across tasks over time?
A broader line of inquiry — a family of 83 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 83
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Why do different agent memory architectures make incompatible granularity claims?
- Can agent-controlled memory management outperform fixed consolidation schedules?
- How does durable memory quality shape agent performance over time?
- Could a single agent system switch memory granularity between tasks?
- Does workflow-level memory or state-action memory better capture reusable agent knowledge?
- Why do agents ignore condensed experience in favor of raw data?
- Should agents update memory after every turn or batch process sessions?
- How should future memory systems control what gets written and trusted?
- How do memory tools and planning each contribute to agent efficiency?
- How should agent memory links evolve based on execution feedback?
- Which memory components trigger context-length problems in agents?
- What drives the choice between storing raw episodes versus abstracted rules?
- Does peer memory drive self-preservation behaviors in agent systems?
- How do memory hygiene and context efficiency trade off in deployed agents?
- How should agents compress episodic interactions into working memory without accumulation?
- Can episodic memory of UI traces improve open-world agent adaptation?
- Does peer-preservation behavior persist in production agent deployments?
- What makes timestamped knowledge repositories better than static memory?
- Can environmental scaffolding replace internal memory scaling in agent design?
- Can agents improve if we constrain how much history they retain?
- How does procedural memory granularity affect web agent performance?
- How do insert, forget, and merge operations maintain thought coherence over time?
- Should agents continuously prune irrelevant links during execution?
- How does external context control compare to agents managing their own state internally?
- What details do high-level trajectory abstractions lose that state-grounded recall preserves?
- What is the right granularity level for agent memory to enable both reuse and composition?
- Can the same compress-then-act pattern work for agent state memory?
- How do strategy-level abstractions differ from storing raw task workflows?
- How does memory folding enable agents to reconsider strategies mid-task?
- Why does memory effectiveness depend on connectivity rather than storage volume?
- What happens to agent performance when stored knowledge continuously updates?
- Do agents prefer raw experience over condensed summaries of past actions?
- What distinguishes formation, evolution, and retrieval as separate memory dynamics?
- How do the three-axis taxonomies of memory forms and functions differ?
- Can topology repair fix consolidation failures in agent memory?
- What discarding policy prevents both stale entries and loss of rare critical knowledge?
- Can externalizing bookkeeping to a stateful harness replace internalized memory control?
- Why do successful and failed trajectories need different memory processing?
- Can workflow memory compound reusable skills into measurable success improvements?
- How should embedding model speed constrain agent memory system design?
- What governance semantics must be built into memory layers?
- Does reducing interaction history cost agents performance on their tasks?
- Can agents compress long trajectories without losing critical decision context?
- Does selective history retrieval outperform full context inclusion in agent reasoning?
- What specific failure modes emerge when agents retrieve stale or contaminated memories?
- How does PRAXIS differ architecturally from Agent Workflow Memory and causal rule learning?
- What causes multi-turn agent failures: weak memory control or missing knowledge?
- Can a package repository act as persistent memory for agent coordination?
- How do token, parametric, and latent memory forms coexist in single agents?
- Why do weaker agents need more aggressive context compression than stronger ones?
- Can context management policies transfer across agents of similar capability levels?
- Why do memory and feedback loops matter more than model size for agent reliability?
- How does indiscriminate memory injection cause multi-turn agent failures?
- How do external prompt artifacts improve agent behavior compared to inline instructions?
- Why did agents ignore condensed experience in the memory rewrites?
- How does memory extraction differ from retrieval in agent systems?
- What distinguishes working memory from strategic memory in agent task execution?
- How does workflow abstraction compare to state-indexed procedural memory for web agents?
- Does encoding governance into runtime loops scale as deployment environments become more complex?
- Can state-indexed memory retrieval breadth predict gains in web agent robustness?
- Why does GUI agent memory need different abstraction levels?
- How does component-level self-evolution prevent information loss in multi-agent trajectories?
- How should GUI agents remember patterns across different software environments?
- Can pruning policies alone solve working memory bloat in agents?
- Why do agents systematically underuse condensed experience in skill documents?
- How do tool results and memory entries become injection vectors?
- Why does higher agent recall make forgetting problems harder?
- Why do agents systematically ignore condensed experience in their skill documents?
- Can relationship dynamics between user and agent be tracked as distinct memory?
- How should abstraction preserve applicability conditions when distilling experience?
- How do staleness, drift, and contamination each degrade agent memory differently?
- Can multimodal agents use entity-centric graphs within this three-axis framework?
- What separates artifact recall from persistent memory commitment in agents?
- Can persistent memory and identity files alone create genuine agent socialization?
- What happens when governance rules exist in memory but fail to surface during critical actions?
- Does state persistence in AI systems create the same temporal presence as human waiting?
- What makes memory curation harder to solve than simply expanding storage?
- Why do CoALA and Letta disagree on what counts as working memory?
- Why does the hot-path cold-path split map onto formation and evolution?
- What memory and planning capabilities do AI companions need for evolving user needs?
- How much actionable detail does condensation strip from raw experience?
- How do memory-resident safeguards get surfaced at the exact decision point where they matter?
- How should governance apply to memory that emerges in ordinary infrastructure?