arxiv:2609.39687
Tom Lu
eigentom
AI & ML interests
MLLM, Reinforcement Learning, Agentic RL
Recent Activity
updated a dataset about 19 hours ago
eigentom/minicpm5-swe-native-eval-archive published a dataset 1 day ago
eigentom/minicpm5-swe-native-eval-archive authored a paper 1 day ago
Better Supervision Is Nearby: Neighborhood On-Policy Self-Distillation