-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 88 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 234 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 161 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
Collections
Discover the best community collections!
Collections including paper arxiv:2608.06296
-
Memory is Reconstructed, Not Retrieved: Graph Memory for LLM Agents
Paper • 2606.06036 • Published • 77 -
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments
Paper • 2606.13681 • Published • 143 -
The Personalization Mirage: How LLMs Fabricate User Profiles, and Why Self-Monitoring Misleads
Paper • 2608.04570 • Published • 41 -
To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing
Paper • 2607.28887 • Published • 20
-
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning
Paper • 2607.07508 • Published • 32 -
Proximal Policy Optimization Algorithms
Paper • 1707.06347 • Published • 12 -
Scaling Laws for Neural Language Models
Paper • 2001.08361 • Published • 12 -
Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning
Paper • 2012.13255 • Published • 6
-
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
Paper • 2608.09888 • Published • 766 -
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination
Paper • 2608.14391 • Published • 281 -
EnvHarness: Awakening Static Worlds for Agent Learning
Paper • 2608.19880 • Published • 273 -
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution
Paper • 2608.00677 • Published • 263
-
SAF-OPD: Stable Advantage Fusion for On-Policy Distillation
Paper • 2607.29209 • Published • 34 -
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning
Paper • 2608.05987 • Published • 100 -
OPD-V: Visual On-Policy Self-Distillation with Modality Balance
Paper • 2608.05131 • Published • 14 -
On-Policy Self-Distillation without Any Supervision
Paper • 2608.06296 • Published • 218
-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 88 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 234 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 161 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
-
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
Paper • 2608.09888 • Published • 766 -
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination
Paper • 2608.14391 • Published • 281 -
EnvHarness: Awakening Static Worlds for Agent Learning
Paper • 2608.19880 • Published • 273 -
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution
Paper • 2608.00677 • Published • 263
-
Memory is Reconstructed, Not Retrieved: Graph Memory for LLM Agents
Paper • 2606.06036 • Published • 77 -
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments
Paper • 2606.13681 • Published • 143 -
The Personalization Mirage: How LLMs Fabricate User Profiles, and Why Self-Monitoring Misleads
Paper • 2608.04570 • Published • 41 -
To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing
Paper • 2607.28887 • Published • 20
-
SAF-OPD: Stable Advantage Fusion for On-Policy Distillation
Paper • 2607.29209 • Published • 34 -
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning
Paper • 2608.05987 • Published • 100 -
OPD-V: Visual On-Policy Self-Distillation with Modality Balance
Paper • 2608.05131 • Published • 14 -
On-Policy Self-Distillation without Any Supervision
Paper • 2608.06296 • Published • 218
-
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning
Paper • 2607.07508 • Published • 32 -
Proximal Policy Optimization Algorithms
Paper • 1707.06347 • Published • 12 -
Scaling Laws for Neural Language Models
Paper • 2001.08361 • Published • 12 -
Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning
Paper • 2012.13255 • Published • 6