Page 416 - 《软件学报》2026年第4期
P. 416
何子健 等: 基于扩散模型的个性化图像生成方法综述 1857
(a) 豆包 (b) Gemini
(c) 可灵 (d) ChatGPT4o
图 3 企业发布的个性化内容生成模型交互界面
生成任务 参考图像 文本指令 生成图像/视频
1. A dog is sleeping
物体驱动
2. A dog is in a bucket
1. A man wearing
人像驱动 headphones with red hair
2. A woman wearing
sunglasses and necklace
图像驱动个性化 1. A woman is dancing 相
视频生成 关
个
单主体个性化生成 性
1. A dog wearing a shirt and 化
necklace having a picnic 生
多概念生成与组合 2. A dog wearing a shirt, 成
glasses flying in the sky 方
tied to balloons 法
多概念属性可控 1. Replace the fork
编辑 with a knife
2. Remove the fork
多概念服装组合 1. A girl with a bag in
试衣 the classroom
多概念个性化生成
图 4 可视化基于扩散模型的个性化图像生成相关任务

