Live page · Day archive

Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards

Read the original at HF Daily Papers

Summary

Researchers develop a post-training method for text-to-image models that combines preference and rubric rewards, raising the Flux2dev model's Elo rating by 69 points.

Carried by: HF Daily Papers. First seen: .