SD微调模型DPO
这个模型使用了类似 LLM 的 RlHF 的技术对模型进行微调来提高其美学表现,从他们的测试来看人工评价的效果比原始的 SDXL 和 SD1.5 模型要好 70%。
https://www.tarogoing.uk/2024/01/10/sd%e5%be%ae%e8%b0%83%e6%a8%a1%e5%9e%8bdpo/
CleverLibre Social is an inclusive social instance for open discussion, learning, and community. All cultures welcome. Hate speech and harassment strictly forbidden.