ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks Paper • 2609.18805 • Published 6 days ago • 57
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 22 days ago • 97
Are We Measuring Strategy or Phrasing? The Gap Between Surface- and Approach-Level Diversity in LLM Math Reasoning Paper • 2606.29985 • Published Jun 29 • 19
Are We Measuring Strategy or Phrasing? The Gap Between Surface- and Approach-Level Diversity in LLM Math Reasoning Paper • 2606.29985 • Published Jun 29 • 19
LiSA: Lifelong Safety Adaptation via Conservative Policy Induction Paper • 2605.14454 • Published May 14 • 5
LiSA: Lifelong Safety Adaptation via Conservative Policy Induction Paper • 2605.14454 • Published May 14 • 5
LiSA: Lifelong Safety Adaptation via Conservative Policy Induction Paper • 2605.14454 • Published May 14 • 5
ReflectCAP: Detailed Image Captioning with Reflective Memory Paper • 2604.12357 • Published Apr 14 • 2
ReflectCAP: Detailed Image Captioning with Reflective Memory Paper • 2604.12357 • Published Apr 14 • 2
Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs? Paper • 2603.24472 • Published Mar 25 • 58
Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs? Paper • 2603.24472 • Published Mar 25 • 58
Understanding Reasoning in LLMs through Strategic Information Allocation under Uncertainty Paper • 2603.15500 • Published Mar 16 • 12
Understanding Reasoning in LLMs through Strategic Information Allocation under Uncertainty Paper • 2603.15500 • Published Mar 16 • 12
Beyond Normalization: Rethinking the Partition Function as a Difficulty Scheduler for RLVR Paper • 2602.12642 • Published Feb 13 • 1
CausalArmor: Efficient Indirect Prompt Injection Guardrails via Causal Attribution Paper • 2602.07918 • Published Feb 8 • 6
Critic-Guided Decoding for Controlled Text Generation Paper • 2212.10938 • Published Dec 21, 2022 • 2
VLind-Bench: Measuring Language Priors in Large Vision-Language Models Paper • 2406.08702 • Published Jun 13, 2024
AdvisorQA: Towards Helpful and Harmless Advice-seeking Question Answering with Collective Intelligence Paper • 2404.11826 • Published Apr 18, 2024
Drift: Decoding-time Personalized Alignments with Implicit User Preferences Paper • 2502.14289 • Published Feb 20, 2025 • 1