Do Implicit Personalization and Explicit Styles Conflict? PsPLUG: A Lightweight Plug-in for Balancing Personalization and Style in Customized LLMs Paper • 2601.06362 • Published 11 days ago • 10
StudentSim: Training LLM-based Student Simulators Paper • 2609.01591 • Published about 1 month ago • 494
AV-GRPO: Modality-Anchored Decoupling Diffusion Reinforcement Learning for Joint Audio-Video Generation Paper • 2609.29816 • Published 7 days ago • 12
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 7 days ago • 32
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 12 days ago • 41
Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL Paper • 2609.20715 • Published 14 days ago • 44
ImIR: Image-Instruction Tuning for All-in-One Image Restoration Paper • 2609.25267 • Published 10 days ago • 14
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 10 days ago • 101
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 13 days ago • 138
ALPINE: Adaptive Localization for Parameter- and Sample-Efficient Few-Shot Learning Paper • 2609.22323 • Published 15 days ago • 10