FinanceHarness: Autonomous Financial Deep Research Framework Paper • 2607.27853 • Published 8 days ago • 5 • 2
When Teachers Mislead: Spurious-Signal-Aware On-Policy Distillation Paper • 2608.03632 • Published 3 days ago • 20 • 2
TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex Paper • 2607.22143 • Published 3 days ago • 5 • 2
ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment Paper • 2608.05102 • Published 2 days ago • 57 • 2
OPD-V: Visual On-Policy Self-Distillation with Modality Balance Paper • 2608.05131 • Published 2 days ago • 8 • 1
Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance Paper • 2608.00782 • Published 6 days ago • 14 • 2
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning Paper • 2608.05139 • Published 2 days ago • 23 • 2
NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap Paper • 2608.04397 • Published 2 days ago • 20 • 4
OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents Paper • 2608.05013 • Published 3 days ago • 29 • 3
When Memory Lies: An Empirical Study of Spatial Memory Staleness in VLM Agents Paper • 2608.04574 • Published 2 days ago • 12 • 2
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Paper • 2608.05000 • Published 2 days ago • 49 • 2
DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack Paper • 2608.03207 • Published 3 days ago • 3 • 2
BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Paper • 2608.05042 • Published 2 days ago • 7 • 2
UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Paper • 2608.04701 • Published 2 days ago • 6 • 2
Agent Against Agent: An Agentic System for Automatic Prompt Injection Red Teaming Paper • 2608.05108 • Published 2 days ago • 5 • 2
SKILL-KD: Contrastive Skill Distillation for LLM Agents Paper • 2607.28048 • Published 3 days ago • 9 • 2
SIGNPOST-Bench: Benchmarking Text-Vision Conflict Resolution in Multimodal Large Language Models Paper • 2608.04244 • Published 3 days ago • 2 • 2