AI 生圖

用 AI 給 AI 寫 Prompt

用戶講「Alice 喺露台發呆」。生圖模型需要的是一長段精確的視覺描述,光這句話遠遠不夠。中間有一個翻譯步驟:用 LLM 把人話翻譯成生圖語言。

示例一:陽台發呆
用戶講
「Alice 喺露台發呆」
LLM 翻譯成生圖 Prompt
A young woman named Alice standing on a sunlit balcony, leaning against the railing, gazing into the distance with a dreamy expression. She has shoulder-length dark hair, wearing a white blouse with a small star necklace. Soft afternoon golden hour lighting, potted plants on the balcony, blurred city skyline in background. Illustration style, warm color palette, peaceful mood. Upper body to full body composition.
生圖模型輸出
Alice 在陽台發呆
再來一個例子
示例二:早晨做飯
用戶講
「Alice 喺度煮飯」
LLM 翻譯成生圖 Prompt
A young woman named Alice in a bright modern kitchen during morning time, cooking breakfast. She has shoulder-length dark hair, wearing a casual cardigan over a white top with a star necklace. Warm natural light streaming through windows, kitchen utensils and ingredients on counter, steam rising from pan. Illustration style, cozy domestic atmosphere, soft warm tones. Wide shot showing kitchen environment.
生圖模型輸出
Alice 在廚房做早餐
為什麼需要這一層翻譯?
① 用戶不會寫生圖 Prompt:他不知道要指定 golden hour lighting 還是 illustration style
② 生圖模型無法理解模糊意圖:「發呆」對模型來説唔係一個視覺描述
③ 每個生圖模型的方言不同:Midjourney、DALL-E、Stable Diffusion 偏好的 Prompt 風格各異
用戶想的和生圖模型需要的是完全不同的語言。在 Alice 裏,每一次出圖背後都有一個 LLM 在翻譯:把用戶的一句話擴展成幾百 token 的詳細視覺描述。這不是錦上添花,這是必要架構。