
將草圖轉為完整圖片
把簡單草圖轉為產品照、室內設計或 3D 風格場景。上傳草圖並描述想呈現的樣貌,利用草圖引導物件位置與比例,再由 GPT Image 2.5 加入色彩、光線及紋理。
GPT Image 2.5 提供更自然的細節、更貼近參考圖的成果,以及更好的編輯控制。Flare 也縮短了生成時間,Sunburst 則著重細膩的編輯精準度。
| 功能 | GPT Image 2 | GPT Image 2.5 |
|---|---|---|
| 釋出日期 | 2026 年 4 月 21 日 | 2026 年 9 月 8 日 |
| 影像細節 | 寫實影像生成 | 更自然的光線與細緻紋理 |
| 參考圖片 | 使用參考圖片引導成果 | 更能保留臉孔、產品與其他主體的辨識度 |
| 局部編輯 | 支援文字指令修改 | 修改更精準,減少其他區域不必要的變動 |
| 多次編輯的一致性 | 支援後續編輯 | 更能保留先前的修改與細節 |
| 生成速度 | OpenAI 比較基準 | Flare:根據 OpenAI,延遲降低 50% |
| Arena 文字生圖排名 | 第 3 名 · 分數:1381 | Sunburst 第 1 名 · 1421;Flare 第 2 名 · 1399 |
| Arena 單張圖片編輯排名 | 第 3 名 · 分數:1461 | Sunburst 第 1 名 · 1520;Flare 第 2 名 · 1491 |
排名取自 Arena 2026 年 9 月 7 日排行榜快照,於 9 月 9 日核對。GPT Image 2.5 的結果標示為 Preliminary(初步結果),可能隨新增票數而變動。GPT Image 2 項目採用 medium 品質。速度提升適用於 Flare,不適用於 Sunburst。
從提示詞、草圖或參考圖片開始,完成視覺作品後再微調細節,無須重新製作整張圖片。

把簡單草圖轉為產品照、室內設計或 3D 風格場景。上傳草圖並描述想呈現的樣貌,利用草圖引導物件位置與比例,再由 GPT Image 2.5 加入色彩、光線及紋理。


建立一系列姿勢、表情或物件位置略有變化的圖片,再透過其他編輯工具組合成循環 GIF 或停格風格動畫。各影格使用相同參考圖,有助於維持角色及場景一致。

替換背景、調整產品顏色,或更新海報文字。說明要修改與保留的內容,GPT Image 2.5 能更精準地編輯,協助保留主體、版面配置與周圍細節。

將同一產品放入不同情境,或嘗試新的人像風格。參考圖片能引導成果,更貼近原本的獨特特徵,協助建立風格一致的產品照或角色圖片。

設計活動海報、邀請函與社群圖片,將文字融入畫面。指定確切文案、顯示位置及要凸顯的元素。發布前請檢查拼字與小字。

生成人像自然的膚質、清晰的衣料細節,以及可信的產品反射效果。描述材質、光源與場景,引導生成柔光攝影棚照片或戶外影像。

An expedition team in weathered EVA suits discovering cyclopean alien architecture half-buried in blue ice on a distant moon, a colossal stone face staring from a cliff, aurora lights dancing overhead, frost crystals catching the light of a distant binary star, footprints in untouched snow, sense of ancient dread, cold blue-cyan palette with warm suit lights, National Geographic meets Alien, cinematic documentary still.

Turn the reference sketch into a photorealistic autumn portrait, preserving its pose and composition. A young adult woman with long brown hair sits on a wooden swing, holding both ropes and smiling gently. She wears a cream knit sweater, burgundy plaid scarf, and blue jeans. Shoot from a close, slightly low angle beneath orange-red maple leaves, with a rustic fence and wooded hills behind her. Soft daylight, natural skin and fabric texture, windblown hair, subtle film grain. 16:9 landscape, finished photo only, no sketch or text.

A whimsical cinematic 3D illustration of frightened, dirty fruits and vegetables on a wooden kitchen counter. A large red apple stands in the center, surrounded by a tomato, lettuce, carrot, and broccoli, all with expressive eyes and worried faces. Soil, caterpillars, flies, and a snail cover their surfaces. A hand holds a magnifying glass in the foreground, revealing colorful cartoon microbes on the apple’s skin. Warm sunlight, a cozy rustic kitchen, detailed textures, shallow depth of field, playful educational advertising style. Background signs read "FRUTAS Y VERDURAS MÁS LIMPIAS = UNA FAMILIA MÁS SANA" and "LO QUE NO VES TAMBIÉN IMPORTA." Landscape composition.

A candid full-body photo of a young adult woman at an outdoor flower market, holding a bouquet wrapped in brown paper. White camisole, loose light-blue jeans, ivory shoulder bag, tousled brown hair, looking down with a gentle smile. Colorful flower stalls, soft daylight, natural skin texture, subtle film grain.

Find exactly 100 distinct, real, drawable physical things whose established English names start with X. Include varied objects, organisms, foods, materials, and artifacts. Exclude invented names, "Xmas," abstract terms, symbols, software, duplicates, synonyms, trivial variants, and generic objects padded with X-word modifiers. Allow established compounds, common object abbreviations such as "XLR connector," accepted loanwords, and accurately drawable scientific terms. Verify obscure terms and avoid misleading names or illustrations. For each entry, provide: number, exact name, category, brief definition, visual description, and High/Medium confidence. Check every entry twice for validity, tangibility, visual distinctness, accurate depiction, and suitability for a children’s educational chart. Exclude borderline entries where possible; identify any retained ones. Add a "Rejected candidates" section with brief reasons. Prioritize accuracy over quantity. If fewer than 100 qualify, state this and provide the best verified set without inventing fillers. Only after validating 100 entries, create a 10 × 10 picture grid. Each cell must show exactly one approved item, its number, and its unchanged name. Use a white background, consistent illustrations, clear cell divisions, large readable labels, and no decorative objects. Depictions must match the validated definitions precisely.

Photorealistic Paris street portrait, vertical 9:16. A young Southeast Asian hijabi woman walks forward, turning back toward the camera with a calm, curious expression. She wears a flowing taupe hijab, oversized emerald blazer, beige tulle maxi skirt, and brown leather shoulder bag, carrying blush roses wrapped in kraft paper. The Eiffel Tower is recognizable in the upper-right background. Slow-shutter panning creates horizontal background blur while her face stays relatively sharp. Wind moves her hijab and skirt. Slight Dutch angle, 50mm perspective, soft afternoon backlight, muted natural colors, creamy highlights, and subtle film grain. Romantic, candid, and realistic. No aged-photo effects, heavy grain, orange grading, plastic skin, CGI, text, or watermark.

Photorealistic travel portrait of a young East Asian woman on a sandy shore beside moss-covered rocks, with a grand medieval stone abbey rising on a rocky island behind her. Long dark brown hair, gentle smile, looking at the camera. She wears an oversized black coat with hands in pockets, a cream scarf, and a subtle chain shoulder bag. Soft evening light, pale blue sky, detailed stone architecture, natural skin and fabric textures, candid smartphone-photo feel.

Makoto Shinkai-style anime illustration of a boy with a backpack skateboarding down a steep coastal street, viewed from behind. A sparkling turquoise ocean and crescent beach on the left, a detailed seaside town on the right, towering white clouds in a brilliant blue sky. Vivid summer colors, cinematic sunlight, nostalgic atmosphere.

Create a charming handcrafted miniature travel scene featuring [ICONIC STRUCTURE] as the main focal point. Show the landmark as a beautifully sculpted tiny 3D model, with soft rounded details, handmade textures, delicate imperfections, and a whimsical storybook feeling. Surround it with a few subtle elements that represent its location—such as tiny trees, flowers, streets, boats, mountains, clouds, or local objects—without making the scene crowded. Place everything on a clean warm-white textured paper background, with plenty of elegant negative space. Add a small tasteful wooden or paper travel plaque containing: [STRUCTURE NAME] [CITY, COUNTRY] Famous for: [SHORT UNIQUE FACT] Use soft natural lighting, gentle shadows, pastel yet realistic colors, miniature diorama depth, handcrafted clay/paper textures, and a premium cute travel-journal aesthetic. Centered composition, highly detailed landmark, adorable but sophisticated, clean and collectible travel-card design, no photorealistic people, no clutter.

A candid fashion portrait of a young man with tousled dark hair sitting among empty blue velvet theater seats. Brown suede jacket, cream shirt, loose burgundy tie, ivory wide-leg trousers, and dark leather boots. Holding a small silver camera, looking off-camera. Direct flash, natural textures, subtle film grain, vintage editorial style. Landscape, approximately 5:4.

Create image of Magazine feature article [travel] guide page, cute, information dense photo book style magazine feature article page. Add all necessary sections, tips, recommendations, information. add photos for any sections and recommendations if you like. Place the attached person at the precise location of [Plateau of Gorgoroth in Mordor]. Seamlessly blend the attached person as if they are sightseeing. Approach this task with the understanding that this is a critical, information rich page that will significantly influence visitor numbers, text accuracy is important.

A minimalist illustrated poster inspired by Léon: The Professional. A tall man in sunglasses and a long black coat holds a suppressed pistol upright beside a young girl with a short black bob, an oversized green jacket, shorts, and black boots. She holds a potted plant; a suitcase rests between them. Huge red "LEON" lettering fills the upper background, with a simplified gray city skyline below. Angular geometric shapes, elongated proportions, dramatic shadows, subtle paper texture, and a restrained red, black, cream, and green palette.

Transform my photo into an authentic 1980s retro Indian fashion portrait, inspired by the reference image. Keep the exact face, facial features, face structure, skin tone, identity, and natural proportions unchanged. Do not alter or beautify the face. Dress her in a full-sleeve pink shirt tucked neatly into a black high-waisted skirt with a stylish belt. Remove the scarf completely. Add simple 1980s-inspired accessories. Give her bouncy, voluminous retro hair styled with a cute ribbon. Create an authentic 1980s background and atmosphere. Use a Kodak film-camera look with warm faded colors, subtle film grain, soft analog texture, gentle vintage lighting, and natural skin texture. Make it look like a genuine photograph taken in the 1980s, not a modern photo with a vintage filter. No watermark, no text, no distortion.

A documentary wildlife photograph of dozens of seagulls resting on brown seaweed-covered tidal flats, scattered between winding pools of vivid blue seawater. Some stand with clear reflections, others rest or stretch their wings. Elevated viewpoint, layered horizontal composition, strong natural sunlight, crisp shadows, detailed coastal textures, realistic colors. Landscape format, no text or interface elements.

A dramatic low-angle photograph of a male cyclist resting astride a mountain bike on a rocky countryside trail, drinking from a water bottle. Black cycling kit, red-and-black helmet, sunglasses. The large front tire dominates the foreground, surrounded by dry golden grass and distant wind turbines beneath a turquoise sky with towering clouds. Warm late-afternoon sunlight, long shadows, rich earthy colors, detailed gravel, wide-angle perspective. Vertical composition, realistic outdoor sports photography.

Create a professional 16:9 character reference sheet based strictly on the uploaded image. Preserve the character’s exact face, age, hairstyle, body proportions, costume, accessories, colors, and distinctive details without redesign. Include: One large hero portrait. Five aligned full-body views: front, front three-quarter, side, back three-quarter, and back. Six expressions: neutral, happy, serious, angry, surprised, and sad. 4–6 natural full-body poses. Costume, accessory, and material close-ups. A faithful color palette, proportion guides, and silhouette thumbnail. Use a clean off-white background, organized grid, generous spacing, readable labels, consistent scale, neutral studio lighting, and detailed materials. Keep identity and left/right design details consistent throughout. No overlapping or cropped figures, anatomical errors, extra props, clutter, logos, or watermarks. Target resolution: 15360 × 8640.
影片成果
提示詞
Create ONE 16:9 landscape storyboard image with exactly 16 panels in a clean 4×4 grid, read left to right, top to bottom. Polished 3D animation style, expressive comedy, detailed food, cinematic lighting. Keep two consistent characters: a tall, lean, long-limbed chef with a dark moustache, white jacket and apron; and a short, round, exhausted guest with messy hair, a wrinkled shirt and loose tie. Setting: a closed restaurant kitchen with steel counters, copper pans and large wall switches. Panels: 1. The chef wipes the counter; the guest slumps beside cold takeaway noodles under one spotlight. 2. The guest pokes the noodles, then drops his forehead onto the counter. 3. The chef notices him and sets down his clothes. 4. The chef rolls up his sleeves and tightens his apron. 5. He switches on the kitchen lights; the burners ignite. 6. He rapidly slices beef into neat pieces. 7. Orange flames rise from a copper pan; the chef remains calm. 8. The guest ducks as a pan flies past, his tie blown backward. 9. The chef juggles cooking tasks and tosses a pan overhead. 10. A rapidly spinning pepper grinder releases a huge pepper cloud. 11. The guest sneezes, then watches wide-eyed. 12. The chef catches two falling pans without spilling anything. 13. Overhead view of beautifully plated beef medallions, dark sauce and three tiny herbs. 14. The guest tastes the dish, closes his eyes and sheds a joyful tear. 15. The guest grins; the chef smirks and discards the takeaway box. 16. The kitchen goes dark except for the guest’s spotlight as he keeps eating. Maintain matching faces, costumes, props and kitchen layout. Lighting shifts from cool darkness to warm cooking light, then back to darkness. Clear panel borders, small numbers 1–16 only, no dialogue, captions, logos or watermark.
影片成果
提示詞
Create a premium cinematic character bible sheet for CALVIN & BOB from central intelligence. LAYOUT: Split screen partner format. Two halves divided by a bold dramatic dividing element in the center. LEFT SIDE — CALVIN: Deep blue watercolor splash behind him fading into center. Large bold brushstroke text CALVIN top left in deep blue. Below small text: THE SPY / CENTRAL INTELLIGENCE. One massive dramatic cropped hero image of Calvin from mid-thigh up — enormous muscular build, light blue unicorn t-shirt, no fanny pack, serious focused expression, fist raised, deep blue watercolor splash radiating behind him. CENTER: Bold dramatic CENTRAL INTELLIGENCE in deep blue slightly worn. Below it small text: OFFICE PARK. UNLIKELY HEROES. RIGHT SIDE — BOB: Warm yellow and white watercolor splash behind him fading into center. Large bold brushstroke text BOB top right in deep yellow. Below small text: THE ACCOUNTANT / CENTRAL INTELLIGENCE. One massive dramatic cropped hero image of Bob from mid-thigh up — wide nervous eyes, white striped shirt, normal grey blue tie loosened, mouth open mid scream, warm yellow and white watercolor splash radiating behind him. BOTTOM CENTER: Color palette — deep blue, warm yellow, white, grey. Tagline centered: ONE WAS READY. ONE WAS NOT. OVERALL: Clean white background, deep blue watercolor left side, warm yellow watercolor right side, dramatic blue center divider, bold flat color blocking, chunky simplified forms, hard edge shadows, thick black outlines, vibrant saturated colors, minimal clean typography, cinematic cel-shaded 3D anime, hand-painted textures, not cartoon not Disney not Pixar, print ready.
影片成果
提示詞
Create a 15-second cinematic anime cooking sequence in a stainless-steel kitchen with warm dramatic lighting. Follow the uploaded character sheet exactly: a tall, athletic, middle-aged chef with spiky blonde hair, blue eyes, a white chef uniform and black apron. No other chefs. 0–2s: Sear beef fillet, turning each side until golden. 2–4s: Finely chop mushrooms and cook into dark duxelles. 4–6s: Layer Parma ham and duxelles, then tightly roll around the beef using cling film. 6–8s: Remove the film, wrap in pastry, seal and brush with egg wash. 8–10s: Place in the oven; jump cut to the golden baked Wellington. 10–13s: Sound briefly drops as he slices through, revealing pink beef beneath crisp pastry. 13–15s: Wide hero shot of the steaming Wellington; the chef steps back proudly. Cel-shaded 3D anime, hand-painted textures, strong shadows and subtle film grain. Food-focused close-ups, occasional chef reactions, cuts every 2–3 seconds. Preserve cooking order, compressing preparation and baking through montage. Build orchestral music with sizzling and chopping sounds toward the finale.
在工作台、Playground 或 API 中使用 Flare 或 Sunburst。每張生成的圖片單獨計費。
描述想從零建立的圖片,或加入參考圖片並說明要建立或修改的內容。
選擇圖片尺寸與品質,確認點數費用,再點選生成。
檢視並下載圖片。若想繼續編輯,請上傳成果並說明要變更與保留的內容。
探索 Reddit、YouTube 與 X 上的使用者評測、圖片範例及實測展示,了解大家嘗試的提示詞、分享的成果及遇到的限制。
r/ChatGPT
reddit
Image Gen 2.5 產生的圖片太精彩了!生圖模型進步好多!
r/codex
reddit
這些對照表使用一系列 UI 生成提示詞,比較 gpt-image-2、2.5 Flare 與 2.5 Sunburst。參考圖及提示詞都很有挑戰性,需要模型協調處理。我認為 2.5 版本明顯更好。
r/ChatGPT
reddit
大家目前覺得如何?我剛收到開放使用的通知,至少感覺速度快了很多。

OpenAI
ChatGPT Images 2.5 生圖更快,成果更自然、更有辨識度,多次編輯也能維持細節一致。透過留言指定修改,並保留圖片其餘部分。

Brock Mesarich
ChatGPT Image 2.5 剛推出!這支影片會介紹所有新功能,並即時測試。

Riko Nazza AI
介紹 OpenAI 新影像模型 ChatGPT Images 2.5 的所有更新,以及我實際如何用它製作 YouTube 縮圖。完整說明生圖工具的改變,再分享我的具體提示詞與編輯流程。
從提示詞、草圖或照片出發,在 SeeGen AI 使用 GPT Image 2.5 創作新作品,或調整已有圖片。