Członkostwo PixAI
PixAI
PixAI
Generate
Sign in
Strona główna
Generowanie
Modele
Recipe
NEW
Studio
Konkurs
Mio.2

Conversational creation

Skrzynka narzędziowa
Guide

User LoRA

XL

DPO (Direct Preference Optimization) LoRA for XL and 1.5 - OpenRail++

Strona głównaModeleMio.2Profil

Narzędzia AI

  • OC Maker
  • AI Anime Generator
  • VTuber Maker
  • Sonic OC Maker
  • Hazbin Hotel OC Maker
  • Anime Character Creator
  • AI Figure Generator
  • Genshin OC Maker
  • All AI Tools

Odkrywaj

  • Popularne tagi
  • Ranking
  • Rynek Modeli
  • Konkurs
  • Aktualności

O nas

  • O PixAI
  • Guide
  • Jak używać PixAI
  • Tsubaki.2
  • Poznaj Mio
  • Zasady treści

Cennik i pomoc

  • Członkostwo
  • Pakiety kredytów
  • Kontakt

Aplikacja mobilna

  • Pobierz aplikację PixAI
  • App Store
  • Google Play
© 2026 PixAI
  • Regulamin
  • Polityka prywatności
  • Polityka praw autorskich
  • Attributions
  • Ustawa o Określonych Transakcjach Handlowych
DPO (Direct Preference Optimization) LoRA for XL and 1.5 - OpenRail++ - AI Model cover image
Przesłano
5 mar 2024, 18:51
Użycia
354k
Opinie
Doskonały  (7)
Uprawnienia
  • Allow image generation & sharing
  • Zezwól użytkownikom pobierać Twój model
  • Zastosowania komercyjne

Recommended LoRAs

Candy White Ardley from Candy Candy - AI Model cover image

User LoRA

Candy White Ardley from Candy Candy
8
Cute expression Pose-SDXL - AI Model cover image

User LoRA

Cute expression Pose-SDXL
42
Brunhild - Taimanin / ブリュンヒルド - 対魔忍シリーズ - AI Model cover image

User LoRA

Brunhild - Taimanin / ブリュンヒルド - 対魔忍シリーズ
27
BrushStroke - AI Model cover image

User LoRA

BrushStroke
14

Opis

What is DPO?DPO is Direct Preference Optimization, the name given to the process whereby a diffusion model is finetuned based on human-chosen images. Meihua Dang et. al. have trained Stable Diffusion 1.5 and Stable Diffusion XL using this method and the Pick-a-Pic v2 Dataset, which can be found at https://huggingface.co/datasets/yuvalkirstain/pickapic_v2, and wrote a paper about it at https://huggingface.co/papers/2311.12908.What does it Do?The trained DPO models have been observed to produce higher quality images than their untuned counterparts, with a significant emphasis on the adherence of the model to your prompt. These LoRA can bring that prompt adherence to other fine-tuned Stable Diffusion models.Who Trained This?These LoRA are based on the works of Meihua Dang (https://huggingface.co/mhdang) athttps://huggingface.co/mhdang/dpo-sdxl-text2image-v1 and https://huggingface.co/mhdang/dpo-sd1.5-text2image-v1, licensed under OpenRail++.How were these LoRA Made?They were created using Kohya SS by extracting them from other OpenRail++ licensed checkpoints on CivitAI and HuggingFace.1.5: https://civitai.com/models/240850/sd15-direct-preference-optimization-dpo extracted from https://huggingface.co/fp16-guy/Stable-Diffusion-v1-5_fp16_cleaned/blob/main/sd_1.5.safetensors.XL: https://civitai.com/models/238319/sd-xl-dpo-finetune-direct-preference-optimization extracted from https://huggingface.co/stabilityai/stable-diffusion-xl-base-1.0/blob/main/sd_xl_base_1.0_0.9vae.safetensorsThese are also hosted on HuggingFace at https://huggingface.co/benjamin-paine/sd-dpo-offsets/

Komentarze2