Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards Paper • 2610.02967 • Published 10 days ago • 29
RL-Native Distillation: Exploiting Scored Trajectories for Few-Step Image Generation Paper • 2608.09226 • Published 14 days ago • 1
Scaling Reinforcement Learning for Diffusion Models via Velocity Matching Paper • 2608.23664 • Published 15 days ago • 1
Video-MOPD: Multi-Teacher On-Policy Distillation for Video Understanding Paper • 2609.09300 • Published Sep 8 • 1
Video Generation Models: A Survey of Post-Training and Alignment Paper • 2610.00812 • Published 12 days ago • 63
Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL Paper • 2609.37200 • Published 13 days ago • 140
LongLive-Plug: Once-for-All Distillation for Video Generation Paper • 2609.38154 • Published 13 days ago • 42
Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision Paper • 2608.16812 • Published Aug 17 • 50
CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation Paper • 2602.24286 • Published Feb 27 • 101
Beyond Drug Discovery: The Nanotechnology Molecular Optimization (NMO) Benchmark Paper • 2606.30170 • Published Jun 29 • 6
DrugGen 2: A disease-aware language model for enhancing drug discovery Paper • 2607.08404 • Published Jul 9 • 19
DiFA: Inference-Time Forward-Process Alignment for Diffusion Models Paper • 2607.17972 • Published Jul 20 • 8
Parallel Decoding Distillation for Fast Image and Video Generation Paper • 2607.26004 • Published Jul 28 • 17
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published Jul 27 • 40