-
Cosmos 3: Omnimodal World Models for Physical AI
Paper • 2606.02800 • Published • 142 -
Robots Need More than VLA and World Models
Paper • 2606.06556 • Published • 31 -
VLANeXt: Recipes for Building Strong VLA Models
Paper • 2602.18532 • Published • 52 -
Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models
Paper • 2606.11025 • Published • 41
Collections
Discover the best community collections!
Collections including paper arxiv:2606.06556
-
Robots Need More than VLA and World Models
Paper • 2606.06556 • Published • 31 -
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models
Paper • 2606.03988 • Published • 126 -
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation
Paper • 2606.28128 • Published • 53 -
EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots
Paper • 2607.02646 • Published • 27
-
Robots Need More than VLA and World Models
Paper • 2606.06556 • Published • 31 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
LLM Explainability with Counterfactual Chains and Causal Graphs
Paper • 2606.05972 • Published • 18 -
Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings
Paper • 2606.07502 • Published • 100
-
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
Paper • 2605.28742 • Published • 4 -
Reinforcement Learning from Rich Feedback with Distributional DAgger
Paper • 2606.05152 • Published • 3 -
Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development
Paper • 2606.07207 • Published • 4 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
AI for Auto-Research: Roadmap & User Guide
Paper • 2605.18661 • Published • 73 -
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data
Paper • 2605.18287 • Published • 15 -
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
Paper • 2605.16865 • Published • 10 -
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
Paper • 2603.28069 • Published • 9
-
Cosmos 3: Omnimodal World Models for Physical AI
Paper • 2606.02800 • Published • 142 -
Robots Need More than VLA and World Models
Paper • 2606.06556 • Published • 31 -
VLANeXt: Recipes for Building Strong VLA Models
Paper • 2602.18532 • Published • 52 -
Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models
Paper • 2606.11025 • Published • 41
-
Robots Need More than VLA and World Models
Paper • 2606.06556 • Published • 31 -
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models
Paper • 2606.03988 • Published • 126 -
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation
Paper • 2606.28128 • Published • 53 -
EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots
Paper • 2607.02646 • Published • 27
-
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
Paper • 2605.28742 • Published • 4 -
Reinforcement Learning from Rich Feedback with Distributional DAgger
Paper • 2606.05152 • Published • 3 -
Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development
Paper • 2606.07207 • Published • 4 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
Robots Need More than VLA and World Models
Paper • 2606.06556 • Published • 31 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
LLM Explainability with Counterfactual Chains and Causal Graphs
Paper • 2606.05972 • Published • 18 -
Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings
Paper • 2606.07502 • Published • 100
-
AI for Auto-Research: Roadmap & User Guide
Paper • 2605.18661 • Published • 73 -
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data
Paper • 2605.18287 • Published • 15 -
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
Paper • 2605.16865 • Published • 10 -
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
Paper • 2603.28069 • Published • 9