RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published 8 days ago • 110 • 3
Rethinking the Evaluation of Harness Evolution for Agents Paper • 2607.12227 • Published 7 days ago • 8 • 3
Spectral Rewiring for Exploration, Purification, and Model Merging Paper • 2607.03065 • Published 18 days ago • 24 • 5
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published 5 days ago • 182 • 3
From Noisy Traces to Root Causes: Structural Trajectory Analysis and Causal Extraction for Agent Optimization Paper • 2607.07702 • Published 13 days ago • 11 • 3
Weak-to-Strong Generalization via Direct On-Policy Distillation Paper • 2607.05394 • Published 13 days ago • 133 • 3
KronQ: LLM Quantization via Kronecker-Factored Hessian Paper • 2607.07964 • Published 13 days ago • 32 • 4
Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents Paper • 2607.08716 • Published 12 days ago • 14 • 6
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition Paper • 2601.16211 • Published 19 days ago • 54 • 3
A Quantized Native Runtime for On-Device Semantic Audio Generation Paper • 2607.08526 • Published 12 days ago • 4 • 4
Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training Paper • 2607.01232 • Published 19 days ago • 6 • 3
CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration Paper • 2607.05465 • Published 15 days ago • 12 • 3
KVpop -- Key-Value Cache Compression with Predictive Online Pruning Paper • 2607.05061 • Published 15 days ago • 23 • 3
VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Paper • 2607.01804 • Published 19 days ago • 31 • 4
AutoMem: Automated Learning of Memory as a Cognitive Skill Paper • 2607.01224 • Published 20 days ago • 21 • 3
AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents Paper • 2607.02255 • Published 19 days ago • 64 • 3
ELDR: Expert-Locality-Aware Decode Routing for PD-Disaggregated MoE Serving Paper • 2607.00466 • Published 20 days ago • 32 • 3
Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation Paper • 2606.23127 • Published 29 days ago • 25 • 3
Agentic Abstention: Do Agents Know When to Stop Instead of Act? Paper • 2606.28733 • Published 24 days ago • 148 • 9