Follow

SD微调模型DPO

这个模型使用了类似 LLM 的 RlHF 的技术对模型进行微调来提高其美学表现,从他们的测试来看人工评价的效果比原始的 SDXL 和 SD1.5 模型要好 70%。

tarogoing.uk/2024/01/10/sd%e5%

Sign in to participate in the conversation
CleverLibre Social

CleverLibre Social is an inclusive social instance for open discussion, learning, and community.
All cultures welcome.
Hate speech and harassment strictly forbidden.