用 Seedream 生成 AI 图片 — 版本、能力与示例

What is Seedream

Seedream is a series of text-to-image models developed by the Seed team at ByteDance — the company behind the Doubao assistant and a number of consumer AI products. The model turns a natural-language description into a finished image: it is enough to describe what you want to see, and Seedream builds the scene, the objects, the light and the composition for you.

The series has moved through several successive versions, with each new generation improving the level of detail, the match between prompt and final result, and the quality of the text that appears in the image itself. On genkiki.com you will find Seedream 5.0 Pro, Seedream V5 Lite and Seedream V4.5 — each of them in both modes: creating an image from text and editing a photo you have uploaded.

Seedream is aimed at creating images — from realistic photographic frames to illustrations, concept art and graphics with lettering. That makes it a handy tool anywhere you want to turn an idea or a short description into a finished visual quickly, without drawing or shooting anything.

Versions and capabilities in generiram

At the moment Seedream is available in the AI Studio on genkiki.com in the following versions:

VersionModeWhat it does
Seedream V4.5Text to imageCreates an image from a text description (prompt)
Seedream V5 LiteText to imageThe newer generation — images up to 4K from a text description
Seedream V5 LiteImage editingChanges uploaded photos from a description — up to 10 input images at once
Seedream 5.0 ProText to imageThe strongest version in the series — images up to 2K from a text description
Seedream 5.0 ProImage editingEdits uploaded photos from a description — up to 10 input images at once
Seedream V4.5Image editingChanges uploaded photos from a description, with output up to 4K

Text to image

The “text to image” mode is the main way to work with Seedream. You write a description in your own language — for instance what object or scene you want, in what style, with what light and mood — and the model generates an image that matches your words. The more specific the description (details about composition, angle, colours and atmosphere), the closer the result comes to what you pictured. You can experiment with different wordings and run a new generation until you get the look you want.

Image editing

Besides creating images from scratch, Seedream can also work on a photo you have already uploaded: you describe what should change — background, clothing, colour, an added or removed object — and the model returns the new version, keeping everything else. It takes up to 10 input images at once, which lets you combine elements from several photos in a single frame. The mode is handy for product visuals, fixes to a finished frame and variations on one and the same composition.

Strengths

  • Photorealistic detail — Seedream is known for clean, detailed frames that in many cases look close to real photography.
  • Lettering inside the image — one of the hallmarks of the series is better reproduction of captions and letters, which is useful for posters, logos and graphics.
  • Following the prompt — the model tries to stick to the description, including more complex scenes with several elements and specific requirements.
  • High image quality — suited to detailed visuals where clarity and a clean frame are valued.
  • Flexible style — it works for realistic results as well as illustrative and stylised ones, depending on the description.

What it is good for

Seedream is handy for a wide range of tasks where you need an image made from an idea:

  • visuals for social media, posts and ads;
  • concept art, idea sketches and mood board visuals;
  • illustrations for articles, presentations and blogs;
  • posters and graphics where legible text matters;
  • product and advertising compositions from a description;
  • fast prototyping of visual ideas before final production.

How to get started

To try Seedream, open the AI Studio on genkiki.com and pick the Seedream V4.5 model in “text to image” mode. Write a description of what you want to see — the clearer you describe the subject, the style and the mood, the more accurate the result will be — and start the generation. If needed, change the wording and try again until you get the image you are after.

Seedream 的模型 — 详解

哪个适合做什么、弱在哪里、不该拿来做什么。

Seedream V4.5

文字照片
画面比例: 1:1, 9:16, 16:9 价格: 78 积分

擅长

  • The previous generation of the family — tried and predictable
  • Strong on text INSIDE the image, Cyrillic included
  • Good price for the quality

不擅长

  • Falls behind the new generation on detail and on resolution

不要用它来做

  • Don’t pick it when the new generation is available

典型用途

  • An image with a caption
  • Everyday photo work

Seedream V5 Lite

文字照片
画面比例: 1:1, 9:16, 16:9 价格: 68 积分

擅长

  • 先想后画——它会按描述一步步推理,所以能理解复杂的、由多部分组成的任务
  • 可以实时上网搜索:你描述今天的话题,它就把它画出来
  • 我们这里同系列中尺寸最大的图——约 3072×3072(最高约 940 万像素)
  • 一次调用最多 6 次独立生成,外加多个变体——一口气得到一组相互关联的画面
  • 最多接受 14 张参考图片
  • 版面设计很强:海报、文字排版、图表、产品样机、电商图
  • 三者中最便宜也最快的

不擅长

  • 照片真实感不如 4.5 — 这是 ByteDance 有意的取舍:「智能优先于画面」
  • 在复杂场景中有时会弄错人体结构和空间关系
  • 细小密集的文字在 4.5 上更可靠

不要用它来做

  • 会被近距离细看的高端人像,别用它
  • 别用它做字体很细小的标签
  • 别以为它来源可靠,真人脸就一定能往下过到视频 —— 内容还是会另外再查一遍

典型用途

  • 有版式和文字的海报或横幅
  • 网店用的产品样图
  • 按描述做的图表或示意图
  • 分镜用的一组连贯画面
  • 用于印刷的大尺寸图像

Seedream V5 Lite

照片照片
画面比例: 1:1, 9:16, 16:9 价格: 68 积分

擅长

  • 以同系列最高的分辨率改写你给的图片——约 3072×3072
  • 会先想一想任务:能听懂「放到别的环境里,但保持姿势不变」,不用把每一点都写死
  • 最多 14 张参考图片——把几个来源拼到一起
  • 换环境或换背景最便宜的办法
  • 适合做视频的首帧——图片 → 视频 这条链就从这里开始

不擅长

  • 做脸部精细活时比 4.5 和 Pro 粗糙
  • 图片上的文字不如 Pro
  • 场景复杂时会把空间关系搞错

不要用它来做

  • 要求高的人像修图,别用它
  • 别给它同一张脸多个角度拼成的拼图
  • 图上的文字必须无可挑剔时,别用它

典型用途

  • 按描述生成视频的首帧
  • 换背景,同时保留人脸
  • 把现有照片放大重做
  • 一次生成同一画面的多个版本

Seedream 4.5

照片照片
画面比例: 1:1, 9:16, 16:9 价格: 78 积分

擅长

  • 同系列中人像真实感最强——皮肤、布料和材质经得起特写
  • 小字和密集文字上最可靠
  • 完整 4K 输出——像素比 Pro 多
  • 做同样的活比 Pro 便宜
  • 过滤没那么严——Pro 拒绝的描述这里能过
  • 实战检验过:结果可预期,没有意外

不擅长

  • 没有分层编辑——整张图一次性重做
  • 不像 Pro 那样能在图里写多语言文字
  • 在推理和遵循超长描述上都要差一些
  • 上一代产品——不会再加新功能

不要用它来做

  • 别用它在图里写西里尔字母的文字 —— 那是 Pro 的活
  • 别用它做元素很多、描述很细的复杂场景
  • 需要在某一处做精准修补时,别用它

典型用途

  • 会被近距离细看的人像修图
  • 需要让人感受到质感的产品照
  • 品牌的高端视觉
  • 小字必须清晰可读的海报

Seedream 5.0 Pro

文字照片
画面比例: 1:1, 9:16, 16:9 价格: 131 积分

擅长

  • 直接在图里写文字,14 种以上语言——同系列中唯一能把字写好的,西里尔字母也行
  • 一次最多接受 10 张参考图片——同一个人物或同一种风格出现在好几个不同画面里
  • 分层编辑:把图像拆成 2–20 个独立图层,可以分别移动和修改
  • 长而详细的描述它最听话——元素的摆放就按你要求的来
  • 它会考虑空间关系:什么在什么前面,哪个物体相对人有多大
  • 干净、专业的光线,没有那种「塑料感」

不擅长

  • 人像真实感不如 4.5 — 皮肤显得更平滑,脸更像「画出来的」
  • 细小密集的文字(标签、长标题、小字备注)会散乱 — ByteDance 自己也承认这里还有不足
  • 在我们这里输出最高 2048×2048 — 比 Lite 和 4.5 都小
  • 过滤比 4.5 更严:会拒绝 4.5 放行的描述
  • 价格大约是 Lite 的两倍

不要用它来做

  • 皮肤要经得起特写的人像,别用它 —— 那种情况 4.5 更好
  • 别用它做字体很小的海报 —— 字会散掉
  • 需要打印的最大尺寸时,别用它 —— Lite 给的像素更多

典型用途

  • 画面里直接带保加利亚语文字的广告
  • 同一个角色出现在几个不同的画面里
  • 元素很多、描述很细的复杂场景
  • 之后还要分图层编辑的图像
  • 大字清晰可读的包装或标签

Seedream 5.0 Pro

照片照片
画面比例: 1:1, 9:16, 16:9 价格: 131 积分

擅长

  • 同系列中最可控的编辑——它是为改图而生的,不是为从零画图
  • 点选、画箭头、随手涂画都能改:你指出在哪里,不用把整张图重新描述一遍
  • 最多 10 张参考图片——把几个来源拼进同一个画面
  • 换背景、换衣服、换环境时,保持图上的人还是同一个人
  • 在现成图像上加清晰可读的文字,14 种以上语言都行

不擅长

  • 对脸改得太狠就会离原样越来越远——要求改得越多,越像另一个人
  • 图片上的小字会散掉
  • 在我们这里输出最高 2048×2048
  • 做同样的活比 Lite 贵

不要用它来做

  • 别用它从零开始造场景 —— 那是文字生成那个版本的活
  • 别给它同一张脸多个角度拼成的拼图 —— 后面走到视频那一步,这是被拒的导火索
  • 精细的人像修图别指望它 —— 4.5 对皮肤的保护更好

典型用途

  • 换掉人物身后的背景,脸不变
  • 把产品和模特从两张照片合成一张
  • 在现成画面上加文字或标志
  • 只改一个细节,其余都不动

立即试用 Seedream

直接在浏览器里用 Seedream 生成你自己的 图片 — 无需安装,中文界面。