系列:Multimodal Decoding Notes

Multimodal Decoding Notes (3): Generative Models — Diffusion vs GAN (Stable Diffusion & GAN)

1. Two Routes to the Same Goal

Generate realistic samples from noise / a random vector.

Different in spirit, but both are now the substrate of industrial models like SDXL / FLUX / StyleGAN / video generation.

2. Stable Diffusion (LDM, CVPR 2022)

3. GAN (Generative Adversarial Nets, Goodfellow 2014)

4. Why Generative Models Are a Key Multimodal Piece

5. Practical Takeaways

觉得有用?欢迎点赞、收藏,或请作者喝咖啡 ☕️

支付宝收款码

支付宝

微信收款码

微信

💬 留言

评论由 Giscus 驱动(基于 GitHub Discussions)。 当前仓库 NaphJohn/LLM-blog 尚未启用 Discussions:请在 GitHub 仓库 Settings → General → Features 勾选 Discussions 后刷新本页,评论区即自动显示。