Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 7 days ago • 51
FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation Paper • 2609.27657 • Published 7 days ago • 9
SamsungSDS-Research/SGuard-JailbreakFilter-2B-v1 Text Generation • 3B • Updated Dec 15, 2025 • 545 • 20
Heoni/llama-3-KoEn-8b_sft_ep3_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 27 • 3
ALPINE: Adaptive Localization for Parameter- and Sample-Efficient Few-Shot Learning Paper • 2609.22323 • Published 14 days ago • 10
Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents Paper • 2609.23986 • Published 9 days ago • 28
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 12 days ago • 137
andersonbcdefg/red_teaming_reward_modeling_pairwise_no_as_an_ai Viewer • Updated Jun 1, 2023 • 35.3k • 162 • 8
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation Paper • 2609.20511 • Published 13 days ago • 110