Official codebase for Margin-aware Preference Optimization for Aligning Diffusion Models without Reference (MaPO).
-
Updated
Jun 11, 2024 - Python
Official codebase for Margin-aware Preference Optimization for Aligning Diffusion Models without Reference (MaPO).
Video Generation Benchmark
Conditional VAE experiments: contextual bandit regret minimization and BERT-embedding human preference / reward prediction on WebGPT comparisons.
To associate your repository with the human-preference topic, visit your repo's landing page and select "manage topics."