VoxMem: Benchmarking Multimodal Memory in Large Audio Language Models Paper • 2609.32607 • Published 8 days ago • 150
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 7 days ago • 43
RayOrch: Programming and Executing Lineage-Controlled Multi-Grain Dataflows for Foundation-Model Data Preparation Paper • 2609.18703 • Published 18 days ago • 54
Tri-PvP: Exposing Modality Bias in Omni-Modal Large Language Models through Perceptual-Propositional Evidence Conflicts Paper • 2609.06011 • Published 29 days ago • 18
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 13 days ago • 102
Emergent Collusion in Long-Horizon LLM Agent Interaction Paper • 2609.24967 • Published 13 days ago • 19
Why Do Video Diffusion Models Violate Physics? Unveiling the Flaws in Attention Mechanisms Paper • 2609.23658 • Published 14 days ago • 30
HuRo: Robotizing Human Videos for Scalable VLA Pretraining Paper • 2609.10706 • Published 16 days ago • 29
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 13 days ago • 157
onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction Paper • 2609.24983 • Published 13 days ago • 55
Paint-Anything: Unified Any-Color Control for Image Generation and Editing Paper • 2609.20816 • Published 17 days ago • 57
Sample Count Is Not Enough: Candidate-Generation Strategy Shapes the Energy and Performance of LLM Test-Time Scaling Paper • 2609.19499 • Published 18 days ago • 37
The Other Half of the Memory Wall: Serving 35B MoEs from SSD with Trained Routing Prediction Paper • 2609.18063 • Published 18 days ago • 19