OPD-V: Visual On-Policy Self-Distillation with Modality Balance Paper • 2608.05131 • Published 3 days ago • 8
ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning Paper • 2608.03972 • Published 4 days ago • 3
How Do Decoder-Only LLMs Perceive Users? Rethinking Attention Masking for User Representation Learning Paper • 2602.10622 • Published Feb 11 • 28
KORE: Enhancing Knowledge Injection for Large Multimodal Models via Knowledge-Oriented Augmentations and Constraints Paper • 2510.19316 • Published Oct 22, 2025 • 12
Backdoor Cleaning without External Guidance in MLLM Fine-tuning Paper • 2505.16916 • Published May 22, 2025 • 17