Prefill-Free Cross-Family KV Cache Transfer for Heterogeneous Multi-Agent LLMs Paper • 2609.32259 • Published 8 days ago • 95
On-Policy or Off-Policy Learning? A Systematic Study of Distillation Dynamics Paper • 2609.35259 • Published 9 days ago • 194
WUSH-KV: KV Cache Quantization with Data-Adaptive Transforms Paper • 2609.38121 • Published 8 days ago • 25
OSWorld-Science: A Benchmark of Computer Use Agents for Learning and Using Scientific Software Paper • 2609.39903 • Published 7 days ago • 63
RGBD20K: A Large-Scale Benchmark for RGB-D Semantic Segmentation Paper • 2609.29028 • Published 13 days ago • 16
Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR Paper • 2609.37868 • Published 8 days ago • 63
WorldAttention: An Efficient Attention Architecture for Interactive Video World Models Paper • 2609.34606 • Published 9 days ago • 49
In-Flight KV Cache with Clean Anchors for Faster Autoregressive Video Diffusion Paper • 2609.32540 • Published 11 days ago • 35
How Far Are We from Removing the Visual Encoder? Scaling Laws for Encoder-Free Multimodal Pretraining Paper • 2609.35457 • Published 9 days ago • 66
Coding Agents for Generalized Task and Motion Planning Problems Paper • 2609.30233 • Published 13 days ago • 28
PackLab: A Comprehensive Framework for Developing, Training, and Evaluating MLLMs in Robotic Bin Packing Paper • 2609.23784 • Published 17 days ago • 16
Blaming Across the Aisle: Political Contrasting and Blame Attribution in the Danish Parliament Paper • 2609.26346 • Published 15 days ago • 11
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 20 days ago • 224
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published about 1 month ago • 17
All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts Paper • 2609.24058 • Published 16 days ago • 55
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 15 days ago • 164
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 19 days ago • 79
Grounded Skill Synthesis from Code at Scale for Agentic Intelligence Paper • 2609.05571 • Published Sep 4 • 119