arxiv:2606.11025
Tianyu Pang
P2333
AI & ML interests
Machine Learning
Recent Activity
upvoted a paper 4 days ago
WorldReward: Reward Modeling for Camera-Conditioned World Models authored a paper 3 months ago
Rethinking the Divergence Regularization in LLM RL authored a paper 3 months ago
Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models