#arxiv
21 briefs tagged #arxiv.
Study tests cross-model transfer for hashed LLM memory
A new paper finds frozen external memories need target-aligned readers to work across backbones.
Paper proposes common equation for graph neural networks
The arXiv paper unifies GNN layer descriptions across seven architectural families.
Researchers propose capability-centric image data design
The framework organizes image-generation training data around capability dependencies.
Researchers propose hybrid-policy self-editing for LLMs
The arXiv paper targets composability in unstructured knowledge editing for language models.
Researchers push structured memory for AI agents
Six new papers propose schema, graph and skill-based systems for more reliable long-term agent memory.
Researchers target memory failures in long-term LLM agents
New papers propose graph, segment and interference methods for maintaining agent memory.
Paper bounds GPU opportunity in LLM-agent control
The study models when agent control paths expose enough concurrent work for GPU execution.
Paper argues agent safety needs runtime contracts
The arXiv paper says training-time alignment is insufficient for autonomous agents.
Researchers improve logical compound-answer reasoning
A new framework decomposes AND, OR and NEITHER/NOR options before composing predictions.
Paper proposes power law graph attention for LLMs
PLGA generalizes scaled dot-product attention and reports inference-collapse results.
Spark-to-Paper automates research paper generation
The arXiv system turns a research idea into a full paper using 13 composable coding-assistant skills.
Survey maps co-evolution in agentic systems
The arXiv paper organizes agentic self-evolution into agent, environment and meta co-evolution stages.
Paper maps blueprint for economic world models
The arXiv paper outlines a six-level roadmap for building generative economic simulations.
Researchers target visual evidence gaps in VLMs
New arXiv papers propose evidence selection, retrieval and token-pruning methods for multimodal QA.
Activity Frames compiles screen activity for agent memory
The arXiv paper describes a deterministic pipeline for auditable computer-use agent memory.
Researchers outline blueprint for economic world models
The arXiv paper maps a capability ladder for agent-based simulations of economies.
Researchers target credit assignment for search agents
New papers propose denser training signals for multi-step search, retrieval and reasoning agents.
Researchers target VLA reliability in robot manipulation
New arXiv papers propose fixes for recovery, planning, training and physical robustness in VLA robots.
Researchers target weak spots in on-policy distillation
New arXiv papers propose OPD variants for multimodal, agent and generator training.
Researchers release framework for clinical AI modality failures
The harness analyzes which missing modalities cause errors and whether failures are detectable.
Researchers propose DeepVoyager-VL for multimodal agents
The framework targets long-horizon search where visual evidence guides intermediate reasoning.