用 Seedream 生成 AI 圖片 — 版本、能力與範例

What is Seedream

Seedream is a series of text-to-image models developed by the Seed team at ByteDance — the company behind the Doubao assistant and a number of consumer AI products. The model turns a natural-language description into a finished image: it is enough to describe what you want to see, and Seedream builds the scene, the objects, the light and the composition for you.

The series has moved through several successive versions, with each new generation improving the level of detail, the match between prompt and final result, and the quality of the text that appears in the image itself. On genkiki.com you will find Seedream 5.0 Pro, Seedream V5 Lite and Seedream V4.5 — each of them in both modes: creating an image from text and editing a photo you have uploaded.

Seedream is aimed at creating images — from realistic photographic frames to illustrations, concept art and graphics with lettering. That makes it a handy tool anywhere you want to turn an idea or a short description into a finished visual quickly, without drawing or shooting anything.

Versions and capabilities in generiram

At the moment Seedream is available in the AI Studio on genkiki.com in the following versions:

VersionModeWhat it does
Seedream V4.5Text to imageCreates an image from a text description (prompt)
Seedream V5 LiteText to imageThe newer generation — images up to 4K from a text description
Seedream V5 LiteImage editingChanges uploaded photos from a description — up to 10 input images at once
Seedream 5.0 ProText to imageThe strongest version in the series — images up to 2K from a text description
Seedream 5.0 ProImage editingEdits uploaded photos from a description — up to 10 input images at once
Seedream V4.5Image editingChanges uploaded photos from a description, with output up to 4K

Text to image

The “text to image” mode is the main way to work with Seedream. You write a description in your own language — for instance what object or scene you want, in what style, with what light and mood — and the model generates an image that matches your words. The more specific the description (details about composition, angle, colours and atmosphere), the closer the result comes to what you pictured. You can experiment with different wordings and run a new generation until you get the look you want.

Image editing

Besides creating images from scratch, Seedream can also work on a photo you have already uploaded: you describe what should change — background, clothing, colour, an added or removed object — and the model returns the new version, keeping everything else. It takes up to 10 input images at once, which lets you combine elements from several photos in a single frame. The mode is handy for product visuals, fixes to a finished frame and variations on one and the same composition.

Strengths

  • Photorealistic detail — Seedream is known for clean, detailed frames that in many cases look close to real photography.
  • Lettering inside the image — one of the hallmarks of the series is better reproduction of captions and letters, which is useful for posters, logos and graphics.
  • Following the prompt — the model tries to stick to the description, including more complex scenes with several elements and specific requirements.
  • High image quality — suited to detailed visuals where clarity and a clean frame are valued.
  • Flexible style — it works for realistic results as well as illustrative and stylised ones, depending on the description.

What it is good for

Seedream is handy for a wide range of tasks where you need an image made from an idea:

  • visuals for social media, posts and ads;
  • concept art, idea sketches and mood board visuals;
  • illustrations for articles, presentations and blogs;
  • posters and graphics where legible text matters;
  • product and advertising compositions from a description;
  • fast prototyping of visual ideas before final production.

How to get started

To try Seedream, open the AI Studio on genkiki.com and pick the Seedream V4.5 model in “text to image” mode. Write a description of what you want to see — the clearer you describe the subject, the style and the mood, the more accurate the result will be — and start the generation. If needed, change the wording and try again until you get the image you are after.

Seedream 的模型 — 詳解

哪個適合做什麼、弱在哪裡、不該拿來做什麼。

Seedream V4.5

文字照片
畫面比例: 1:1, 9:16, 16:9 價格: 78 點數

擅長

  • The previous generation of the family — tried and predictable
  • Strong on text INSIDE the image, Cyrillic included
  • Good price for the quality

不擅長

  • Falls behind the new generation on detail and on resolution

不要用它來做

  • Don’t pick it when the new generation is available

典型用途

  • An image with a caption
  • Everyday photo work

Seedream V5 Lite

文字照片
畫面比例: 1:1, 9:16, 16:9 價格: 68 點數

擅長

  • 先想後畫——它會按描述一步步推理,所以能理解複雜的、由多部分組成的任務
  • 可以即時上網搜尋:你描述今天的話題,它就把它畫出來
  • 我們這裡同系列中尺寸最大的圖——約 3072×3072(最高約 940 萬像素)
  • 一次呼叫最多 6 次獨立生成,外加多個變體——一口氣得到一組相互關聯的畫面
  • 最多接受 14 張參考圖片
  • 版面設計很強:海報、文字排版、圖表、產品樣機、電商圖
  • 三者中最便宜也最快的

不擅長

  • 照片真實感不如 4.5 — 這是 ByteDance 有意的取捨:「智慧優先於畫面」
  • 在複雜場景中有時會弄錯人體結構和空間關係
  • 細小密集的文字在 4.5 上更可靠

不要用它來做

  • 會被近距離細看的高階人像,別用它
  • 別用它做字體很細小的標籤
  • 別以為它來源可靠,真人臉就一定能往下過到影片 —— 內容還是會另外再查一遍

典型用途

  • 有版式和文字的海報或橫幅
  • 網路商店用的產品樣圖
  • 按描述做的圖表或示意圖
  • 分鏡用的一組連貫畫面
  • 用於印刷的大尺寸圖像

Seedream V5 Lite

照片照片
畫面比例: 1:1, 9:16, 16:9 價格: 68 點數

擅長

  • 以同系列最高的解析度改寫你給的圖片——約 3072×3072
  • 會先想一想任務:能聽懂「放到別的環境裡,但保持姿勢不變」,不用把每一點都寫死
  • 最多 14 張參考圖片——把幾個來源拼到一起
  • 換環境或換背景最便宜的辦法
  • 適合做影片的首幀——圖片 → 影片 這條鏈就從這裡開始

不擅長

  • 做臉部精細活時比 4.5 和 Pro 粗糙
  • 圖片上的文字不如 Pro
  • 場景複雜時會把空間關係搞錯

不要用它來做

  • 要求高的人像修圖,別用它
  • 別給它同一張臉多個角度拼成的拼圖
  • 圖上的文字必須無可挑剔時,別用它

典型用途

  • 按描述生成影片的首幀
  • 換背景,同時保留人臉
  • 把現有照片放大重做
  • 一次生成同一畫面的多個版本

Seedream 4.5

照片照片
畫面比例: 1:1, 9:16, 16:9 價格: 78 點數

擅長

  • 同系列中人像真實感最強——皮膚、布料和材質經得起特寫
  • 小字和密集文字上最可靠
  • 完整 4K 輸出——像素比 Pro 多
  • 做同樣的活比 Pro 便宜
  • 過濾沒那麼嚴——Pro 拒絕的描述這裡能過
  • 實戰檢驗過:結果可預期,沒有意外

不擅長

  • 沒有分層編輯——整張圖一次性重做
  • 不像 Pro 那樣能在圖裡寫多語言文字
  • 在推理和遵循超長描述上都要差一些
  • 上一代產品——不會再加新功能

不要用它來做

  • 別用它在圖裡寫西里爾字母的文字 —— 那是 Pro 的活
  • 別用它做元素很多、描述很細的複雜場景
  • 需要在某一處做精準修補時,別用它

典型用途

  • 會被近距離細看的人像修圖
  • 需要讓人感受到質感的產品照
  • 品牌的頂級視覺
  • 小字必須清晰可讀的海報

Seedream 5.0 Pro

文字照片
畫面比例: 1:1, 9:16, 16:9 價格: 131 點數

擅長

  • 直接在圖裡寫文字,14 種以上語言——同系列中唯一能把字寫好的,西里爾字母也行
  • 一次最多接受 10 張參考圖片——同一個人物或同一種風格出現在好幾個不同畫面裡
  • 分層編輯:把圖像拆成 2–20 個獨立圖層,可以分別移動和修改
  • 長而詳細的描述它最聽話——元素的擺放就按你要求的來
  • 它會考慮空間關係:什麼在什麼前面,哪個物體相對人有多大
  • 乾淨、專業的光線,沒有那種「塑膠感」

不擅長

  • 人像真實感不如 4.5 — 皮膚顯得更平滑,臉更像「畫出來的」
  • 細小密集的文字(標籤、長標題、小字備註)會散亂 — ByteDance 自己也承認這裡還有不足
  • 在我們這裡輸出最高 2048×2048 — 比 Lite 和 4.5 都小
  • 過濾比 4.5 更嚴:會拒絕 4.5 放行的描述
  • 價格大約是 Lite 的兩倍

不要用它來做

  • 皮膚要經得起特寫的人像,別用它 —— 那種情況 4.5 更好
  • 別用它做字體很小的海報 —— 字會散掉
  • 需要列印的最大尺寸時,別用它 —— Lite 給的像素更多

典型用途

  • 畫面裡直接帶保加利亞語文字的廣告
  • 同一個角色出現在幾個不同的畫面裡
  • 元素很多、描述很細的複雜場景
  • 之後還要分圖層編輯的圖像
  • 大字清晰可讀的包裝或標籤

Seedream 5.0 Pro

照片照片
畫面比例: 1:1, 9:16, 16:9 價格: 131 點數

擅長

  • 同系列中最可控的編輯——它是為改圖而生的,不是為從零畫圖
  • 點選、畫箭頭、隨手塗畫都能改:你指出在哪裡,不用把整張圖重新描述一遍
  • 最多 10 張參考圖片——把幾個來源拼進同一個畫面
  • 換背景、換衣服、換環境時,保持圖上的人還是同一個人
  • 在現成圖像上加清晰可讀的文字,14 種以上語言都行

不擅長

  • 對臉改得太狠就會離原樣越來越遠——要求改得越多,越像另一個人
  • 圖片上的小字會散掉
  • 在我們這裡輸出最高 2048×2048
  • 做同樣的活比 Lite 貴

不要用它來做

  • 別用它從零開始造場景 —— 那是文字生成那個版本的活
  • 別給它同一張臉多個角度拼成的拼圖 —— 後面走到影片那一步,這是被拒的導火線
  • 精細的人像修圖別指望它 —— 4.5 對皮膚的保護更好

典型用途

  • 換掉人物身後的背景,臉不變
  • 把產品和模特兒從兩張照片合成一張
  • 在現成畫面上加文字或標誌
  • 只改一個細節,其餘都不動

立即試用 Seedream

直接在瀏覽器裡用 Seedream 生成你自己的 圖片 — 無需安裝,中文介面。