UNREAL: Unifying Retrieval and Long-Context with a Single Model Paper • 2610.08463 • Published 6 days ago • 30
Rationale-Guided Policy Optimization: Learning to Reason with Adaptive Rationale Scaffolding Paper • 2610.07342 • Published 7 days ago • 21
Kinematic MeanFlow: One-Step Action Generation Policy for Robotic Foundation Models Paper • 2610.00864 • Published 11 days ago • 47
Learning to Read the Contextual Tokens in Diffusion Transformers Paper • 2610.06844 • Published 7 days ago • 9
Harness-Aware Distillation for Small Language Model Agents Paper • 2610.02858 • Published 10 days ago • 15
VeriHarness: Scaling Agentic Verification for Long-Horizon Tasks Paper • 2610.00972 • Published 11 days ago • 61
Learning from Teacher Continuations at Student States Paper • 2609.36246 • Published 14 days ago • 41
Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation Paper • 2609.35347 • Published 14 days ago • 180
VideoPhysEdit: Physical Counterfactual Video Editing via Rigid-Body Physical Scene Reconstruction Paper • 2609.35134 • Published 14 days ago • 19
IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis Paper • 2609.29444 • Published 18 days ago • 21
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 19 days ago • 43
Knowledge Pull Requests for Continual Document Authoring Paper • 2609.26634 • Published 20 days ago • 12
StudentBench: AI and human tutoring yield equivalent GRE learning gains Paper • 2609.28470 • Published 19 days ago • 11
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 25 days ago • 78
Grounded Skill Synthesis from Code at Scale for Agentic Intelligence Paper • 2609.05571 • Published Sep 4 • 119
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training Paper • 2609.26774 • Published 20 days ago • 56
Circuit Hypernetworks for Quantum-Augmented Diffusion Language Models Paper • 2609.24657 • Published 21 days ago • 40
OmniEdu: Open Foundation Models for Learning and Teaching Paper • 2609.23088 • Published 23 days ago • 239
Harness-Zero: Harness Distillation via Agent-as-Harness Paper • 2609.24974 • Published 21 days ago • 38