Collections
Discover the best community collections!
Collections including paper arxiv:2604.09132
-
GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Paper • 2503.14734 • Published • 9 -
Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
Paper • 2401.02117 • Published • 33 -
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Paper • 2506.01844 • Published • 165 -
Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding
Paper • 2506.16035 • Published • 89
-
WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes
Paper • 2605.15843 • Published • 7 -
PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World
Paper • 2605.05163 • Published • 38 -
SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects
Paper • 2605.19587 • Published • 10 -
Code-as-Room: Generating 3D Rooms from Top-Down View Images via Agentic Code Synthesis
Paper • 2605.18451 • Published • 41
-
WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes
Paper • 2605.15843 • Published • 7 -
PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World
Paper • 2605.05163 • Published • 38 -
SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects
Paper • 2605.19587 • Published • 10 -
Code-as-Room: Generating 3D Rooms from Top-Down View Images via Agentic Code Synthesis
Paper • 2605.18451 • Published • 41
-
GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Paper • 2503.14734 • Published • 9 -
Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
Paper • 2401.02117 • Published • 33 -
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Paper • 2506.01844 • Published • 165 -
Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding
Paper • 2506.16035 • Published • 89