Keanggotaan PixAI
PixAI
PixAI
Generate
Sign in
Beranda
Buat
Model
Studio
Kontes
Mio.2

Conversational creation

Toolbox
Dokumen PixAI

LoRA Pengguna

XL

DPO (Direct Preference Optimization) LoRA for XL and 1.5 - OpenRail++

Alat AI

  • OC Maker
  • AI Anime Generator
  • VTuber Maker
  • Sonic OC Maker
  • Hazbin Hotel OC Maker
  • Anime Character Creator
  • AI Figure Generator
  • Genshin OC Maker
  • All AI Tools

Temukan

  • Tag Populer
  • Peringkat
  • Model Market
  • Kontes
  • Berita

Tentang

  • Dokumen PixAI
  • Cara Menggunakan PixAI
  • Tsubaki.2
  • Kenalan dengan Mio
  • Aturan Konten

Harga & Bantuan

  • Keanggotaan
  • Paket Kredit
  • Kontak

Aplikasi Seluler

  • Dapatkan aplikasi PixAI
  • App Store
  • Google Play
© 2026 PixAI
  • Ketentuan Layanan
  • Kebijakan Privasi
  • Kebijakan Hak Cipta
  • Atribusi
  • Undang-Undang Transaksi Komersial Tertentu
DPO (Direct Preference Optimization) LoRA for XL and 1.5 - OpenRail++ - AI Model cover image
Diunggah pada
5 Mar 2024, 18.51
Penggunaan
352k
Ulasan
Luar Biasa  (7)
Izin
  • Izinkan pembuatan & berbagi gambar
  • Izinkan pengguna mengunduh model Anda
  • Penggunaan komersial

LoRA yang Direkomendasikan

Candy White Ardley from Candy Candy - AI Model cover image

LoRA Pengguna

Candy White Ardley from Candy Candy
8
Cute expression Pose-SDXL - AI Model cover image

LoRA Pengguna

Cute expression Pose-SDXL
42
Brunhild - Taimanin / ブリュンヒルド - 対魔忍シリーズ - AI Model cover image

LoRA Pengguna

Brunhild - Taimanin / ブリュンヒルド - 対魔忍シリーズ
27
BrushStroke - AI Model cover image

LoRA Pengguna

BrushStroke
14

Deskripsi

What is DPO?DPO is Direct Preference Optimization, the name given to the process whereby a diffusion model is finetuned based on human-chosen images. Meihua Dang et. al. have trained Stable Diffusion 1.5 and Stable Diffusion XL using this method and the Pick-a-Pic v2 Dataset, which can be found at https://huggingface.co/datasets/yuvalkirstain/pickapic_v2, and wrote a paper about it at https://huggingface.co/papers/2311.12908.What does it Do?The trained DPO models have been observed to produce higher quality images than their untuned counterparts, with a significant emphasis on the adherence of the model to your prompt. These LoRA can bring that prompt adherence to other fine-tuned Stable Diffusion models.Who Trained This?These LoRA are based on the works of Meihua Dang (https://huggingface.co/mhdang) athttps://huggingface.co/mhdang/dpo-sdxl-text2image-v1 and https://huggingface.co/mhdang/dpo-sd1.5-text2image-v1, licensed under OpenRail++.How were these LoRA Made?They were created using Kohya SS by extracting them from other OpenRail++ licensed checkpoints on CivitAI and HuggingFace.1.5: https://civitai.com/models/240850/sd15-direct-preference-optimization-dpo extracted from https://huggingface.co/fp16-guy/Stable-Diffusion-v1-5_fp16_cleaned/blob/main/sd_1.5.safetensors.XL: https://civitai.com/models/238319/sd-xl-dpo-finetune-direct-preference-optimization extracted from https://huggingface.co/stabilityai/stable-diffusion-xl-base-1.0/blob/main/sd_xl_base_1.0_0.9vae.safetensorsThese are also hosted on HuggingFace at https://huggingface.co/benjamin-paine/sd-dpo-offsets/

Komentar2