Penghui Qi
QPHutu
AI & ML interests
None yet
Recent Activity
upvoted a paper about 23 hours ago
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements authored a paper 2 months ago
Rethinking the Divergence Regularization in LLM RL upvoted a paper 2 months ago
Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models