Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards Paper • 2610.02967 • Published 9 days ago • 29
RL-Native Distillation: Exploiting Scored Trajectories for Few-Step Image Generation Paper • 2608.09226 • Published 13 days ago • 1
Scaling Reinforcement Learning for Diffusion Models via Velocity Matching Paper • 2608.23664 • Published 14 days ago • 1
Video-MOPD: Multi-Teacher On-Policy Distillation for Video Understanding Paper • 2609.09300 • Published Sep 8 • 1
Video Generation Models: A Survey of Post-Training and Alignment Paper • 2610.00812 • Published 11 days ago • 63
Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL Paper • 2609.37200 • Published 12 days ago • 140
LongLive-Plug: Once-for-All Distillation for Video Generation Paper • 2609.38154 • Published 12 days ago • 42