RationalVLA: A Rational Vision-Language-Action Model with Dual System Paper • 2506.10826 • Published Jun 13, 2025
OmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation Paper • 2609.22069 • Published 12 days ago • 37
OmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation Paper • 2609.22069 • Published 12 days ago • 37
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 13 days ago • 137
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published Jul 27 • 39
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment Paper • 2607.07820 • Published Jul 8 • 95
Generalizing to Unseen Domains in Diabetic Retinopathy with Disentangled Representations Paper • 2406.06384 • Published Jun 10, 2024 • 1
PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution Paper • 2605.25801 • Published May 25 • 1
PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution Paper • 2605.25801 • Published May 25 • 1
Running Agents 362 VBench Leaderboard 📊 362 Submit video model evaluation results to a public benchmark