VTR-Bench: A Systematic Benchmark for Evaluating Visual Text Rendering in Video Generation Paper • 2610.01499 • Published 8 days ago • 12
Prefill-Free Cross-Family KV Cache Transfer for Heterogeneous Multi-Agent LLMs Paper • 2609.32259 • Published 10 days ago • 95
Devils in Question Relay: Source-Conditioned Relay Steering to Mitigate Hallucinations in Audio-visual Large Language Models Paper • 2609.37568 • Published 10 days ago • 11
Learning What to Recall: Adaptive Multi-Cue Episodic Memory for World Models Paper • 2609.34677 • Published 11 days ago • 12
DataMagic: Authoring Data Videos through Declarative Multi-Agent Orchestration Paper • 2609.33403 • Published 12 days ago • 15
4Director: Controlling Video World Models with Rigid 3D Geometry Paper • 2610.02160 • Published 8 days ago • 38