
将草图变成完整图像
把简单绘画转化为产品照片、室内设计或 3D 风格场景。上传草图并描述期望的效果,用草图引导物体位置和比例,再由 GPT Image 2.5 添加颜色、光线和纹理。
GPT Image 2.5 的细节更自然,更贴近参考图,编辑控制更好。Flare 还缩短了生成时间,Sunburst 则注重更精细的编辑。
| 功能 | GPT Image 2 | GPT Image 2.5 |
|---|---|---|
| 发布日期 | 2026 年 4 月 21 日 | 2026 年 9 月 8 日 |
| 图像细节 | 写实图像生成 | 更自然的光线和更精细的纹理 |
| 参考图片 | 使用参考图引导生成结果 | 更好地保留人脸、产品及其他主体的辨识度 |
| 局部编辑 | 支持通过文本指令修改 | 修改更精准,减少对图像其他部分的意外改动 |
| 多轮编辑的一致性 | 支持后续编辑 | 更好地保留此前的修改和细节 |
| 生成速度 | OpenAI 对比的基准 | Flare:据 OpenAI 表示,延迟降低 50% |
| Arena 文生图排名 | 第 3 名 · 评分:1381 | Sunburst 第 1 名 · 1421;Flare 第 2 名 · 1399 |
| Arena 单图编辑排名 | 第 3 名 · 评分:1461 | Sunburst 第 1 名 · 1520;Flare 第 2 名 · 1491 |
排名来自 Arena 2026 年 9 月 7 日榜单快照,于 9 月 9 日核对。GPT Image 2.5 的结果标记为 Preliminary(初步结果),可能随投票增加而变化。GPT Image 2 条目使用 medium 质量。速度提升仅针对 Flare,不适用于 Sunburst。
从提示词、草图或参考图开始,创作完整画面,再调整细节,无需重做整张图。

把简单绘画转化为产品照片、室内设计或 3D 风格场景。上传草图并描述期望的效果,用草图引导物体位置和比例,再由 GPT Image 2.5 添加颜色、光线和纹理。


生成一组姿势、表情或物体位置略有变化的图像,再使用独立编辑工具将这些帧合成为循环 GIF 或定格风格动画。各帧沿用同一参考图,有助于保持角色和场景一致。

替换背景、改变产品颜色,或更新海报文字。描述需要修改和保留的内容,GPT Image 2.5 可更精准地编辑,帮助保留主体、布局和周围细节。

将同一产品置于不同场景,或为人像探索新风格。参考图可引导结果,更好地保留原始特征,帮助制作风格统一的产品照片或角色图。

设计活动海报、邀请函和社交媒体图片,让文字融入画面。指定准确文案、出现位置及重点元素。发布前请检查拼写和小字。

生成人像自然的皮肤纹理、清晰的服装面料细节和真实的产品反光。描述材质、光源和环境,引导生成柔光影棚照片或户外场景。

An expedition team in weathered EVA suits discovering cyclopean alien architecture half-buried in blue ice on a distant moon, a colossal stone face staring from a cliff, aurora lights dancing overhead, frost crystals catching the light of a distant binary star, footprints in untouched snow, sense of ancient dread, cold blue-cyan palette with warm suit lights, National Geographic meets Alien, cinematic documentary still.

Turn the reference sketch into a photorealistic autumn portrait, preserving its pose and composition. A young adult woman with long brown hair sits on a wooden swing, holding both ropes and smiling gently. She wears a cream knit sweater, burgundy plaid scarf, and blue jeans. Shoot from a close, slightly low angle beneath orange-red maple leaves, with a rustic fence and wooded hills behind her. Soft daylight, natural skin and fabric texture, windblown hair, subtle film grain. 16:9 landscape, finished photo only, no sketch or text.

A whimsical cinematic 3D illustration of frightened, dirty fruits and vegetables on a wooden kitchen counter. A large red apple stands in the center, surrounded by a tomato, lettuce, carrot, and broccoli, all with expressive eyes and worried faces. Soil, caterpillars, flies, and a snail cover their surfaces. A hand holds a magnifying glass in the foreground, revealing colorful cartoon microbes on the apple’s skin. Warm sunlight, a cozy rustic kitchen, detailed textures, shallow depth of field, playful educational advertising style. Background signs read "FRUTAS Y VERDURAS MÁS LIMPIAS = UNA FAMILIA MÁS SANA" and "LO QUE NO VES TAMBIÉN IMPORTA." Landscape composition.

A candid full-body photo of a young adult woman at an outdoor flower market, holding a bouquet wrapped in brown paper. White camisole, loose light-blue jeans, ivory shoulder bag, tousled brown hair, looking down with a gentle smile. Colorful flower stalls, soft daylight, natural skin texture, subtle film grain.

Find exactly 100 distinct, real, drawable physical things whose established English names start with X. Include varied objects, organisms, foods, materials, and artifacts. Exclude invented names, "Xmas," abstract terms, symbols, software, duplicates, synonyms, trivial variants, and generic objects padded with X-word modifiers. Allow established compounds, common object abbreviations such as "XLR connector," accepted loanwords, and accurately drawable scientific terms. Verify obscure terms and avoid misleading names or illustrations. For each entry, provide: number, exact name, category, brief definition, visual description, and High/Medium confidence. Check every entry twice for validity, tangibility, visual distinctness, accurate depiction, and suitability for a children’s educational chart. Exclude borderline entries where possible; identify any retained ones. Add a "Rejected candidates" section with brief reasons. Prioritize accuracy over quantity. If fewer than 100 qualify, state this and provide the best verified set without inventing fillers. Only after validating 100 entries, create a 10 × 10 picture grid. Each cell must show exactly one approved item, its number, and its unchanged name. Use a white background, consistent illustrations, clear cell divisions, large readable labels, and no decorative objects. Depictions must match the validated definitions precisely.

Photorealistic Paris street portrait, vertical 9:16. A young Southeast Asian hijabi woman walks forward, turning back toward the camera with a calm, curious expression. She wears a flowing taupe hijab, oversized emerald blazer, beige tulle maxi skirt, and brown leather shoulder bag, carrying blush roses wrapped in kraft paper. The Eiffel Tower is recognizable in the upper-right background. Slow-shutter panning creates horizontal background blur while her face stays relatively sharp. Wind moves her hijab and skirt. Slight Dutch angle, 50mm perspective, soft afternoon backlight, muted natural colors, creamy highlights, and subtle film grain. Romantic, candid, and realistic. No aged-photo effects, heavy grain, orange grading, plastic skin, CGI, text, or watermark.

Photorealistic travel portrait of a young East Asian woman on a sandy shore beside moss-covered rocks, with a grand medieval stone abbey rising on a rocky island behind her. Long dark brown hair, gentle smile, looking at the camera. She wears an oversized black coat with hands in pockets, a cream scarf, and a subtle chain shoulder bag. Soft evening light, pale blue sky, detailed stone architecture, natural skin and fabric textures, candid smartphone-photo feel.

Makoto Shinkai-style anime illustration of a boy with a backpack skateboarding down a steep coastal street, viewed from behind. A sparkling turquoise ocean and crescent beach on the left, a detailed seaside town on the right, towering white clouds in a brilliant blue sky. Vivid summer colors, cinematic sunlight, nostalgic atmosphere.

Create a charming handcrafted miniature travel scene featuring [ICONIC STRUCTURE] as the main focal point. Show the landmark as a beautifully sculpted tiny 3D model, with soft rounded details, handmade textures, delicate imperfections, and a whimsical storybook feeling. Surround it with a few subtle elements that represent its location—such as tiny trees, flowers, streets, boats, mountains, clouds, or local objects—without making the scene crowded. Place everything on a clean warm-white textured paper background, with plenty of elegant negative space. Add a small tasteful wooden or paper travel plaque containing: [STRUCTURE NAME] [CITY, COUNTRY] Famous for: [SHORT UNIQUE FACT] Use soft natural lighting, gentle shadows, pastel yet realistic colors, miniature diorama depth, handcrafted clay/paper textures, and a premium cute travel-journal aesthetic. Centered composition, highly detailed landmark, adorable but sophisticated, clean and collectible travel-card design, no photorealistic people, no clutter.

A candid fashion portrait of a young man with tousled dark hair sitting among empty blue velvet theater seats. Brown suede jacket, cream shirt, loose burgundy tie, ivory wide-leg trousers, and dark leather boots. Holding a small silver camera, looking off-camera. Direct flash, natural textures, subtle film grain, vintage editorial style. Landscape, approximately 5:4.

Create image of Magazine feature article [travel] guide page, cute, information dense photo book style magazine feature article page. Add all necessary sections, tips, recommendations, information. add photos for any sections and recommendations if you like. Place the attached person at the precise location of [Plateau of Gorgoroth in Mordor]. Seamlessly blend the attached person as if they are sightseeing. Approach this task with the understanding that this is a critical, information rich page that will significantly influence visitor numbers, text accuracy is important.

A minimalist illustrated poster inspired by Léon: The Professional. A tall man in sunglasses and a long black coat holds a suppressed pistol upright beside a young girl with a short black bob, an oversized green jacket, shorts, and black boots. She holds a potted plant; a suitcase rests between them. Huge red "LEON" lettering fills the upper background, with a simplified gray city skyline below. Angular geometric shapes, elongated proportions, dramatic shadows, subtle paper texture, and a restrained red, black, cream, and green palette.

Transform my photo into an authentic 1980s retro Indian fashion portrait, inspired by the reference image. Keep the exact face, facial features, face structure, skin tone, identity, and natural proportions unchanged. Do not alter or beautify the face. Dress her in a full-sleeve pink shirt tucked neatly into a black high-waisted skirt with a stylish belt. Remove the scarf completely. Add simple 1980s-inspired accessories. Give her bouncy, voluminous retro hair styled with a cute ribbon. Create an authentic 1980s background and atmosphere. Use a Kodak film-camera look with warm faded colors, subtle film grain, soft analog texture, gentle vintage lighting, and natural skin texture. Make it look like a genuine photograph taken in the 1980s, not a modern photo with a vintage filter. No watermark, no text, no distortion.

A documentary wildlife photograph of dozens of seagulls resting on brown seaweed-covered tidal flats, scattered between winding pools of vivid blue seawater. Some stand with clear reflections, others rest or stretch their wings. Elevated viewpoint, layered horizontal composition, strong natural sunlight, crisp shadows, detailed coastal textures, realistic colors. Landscape format, no text or interface elements.

A dramatic low-angle photograph of a male cyclist resting astride a mountain bike on a rocky countryside trail, drinking from a water bottle. Black cycling kit, red-and-black helmet, sunglasses. The large front tire dominates the foreground, surrounded by dry golden grass and distant wind turbines beneath a turquoise sky with towering clouds. Warm late-afternoon sunlight, long shadows, rich earthy colors, detailed gravel, wide-angle perspective. Vertical composition, realistic outdoor sports photography.

Create a professional 16:9 character reference sheet based strictly on the uploaded image. Preserve the character’s exact face, age, hairstyle, body proportions, costume, accessories, colors, and distinctive details without redesign. Include: One large hero portrait. Five aligned full-body views: front, front three-quarter, side, back three-quarter, and back. Six expressions: neutral, happy, serious, angry, surprised, and sad. 4–6 natural full-body poses. Costume, accessory, and material close-ups. A faithful color palette, proportion guides, and silhouette thumbnail. Use a clean off-white background, organized grid, generous spacing, readable labels, consistent scale, neutral studio lighting, and detailed materials. Keep identity and left/right design details consistent throughout. No overlapping or cropped figures, anatomical errors, extra props, clutter, logos, or watermarks. Target resolution: 15360 × 8640.
视频结果
提示词
Create ONE 16:9 landscape storyboard image with exactly 16 panels in a clean 4×4 grid, read left to right, top to bottom. Polished 3D animation style, expressive comedy, detailed food, cinematic lighting. Keep two consistent characters: a tall, lean, long-limbed chef with a dark moustache, white jacket and apron; and a short, round, exhausted guest with messy hair, a wrinkled shirt and loose tie. Setting: a closed restaurant kitchen with steel counters, copper pans and large wall switches. Panels: 1. The chef wipes the counter; the guest slumps beside cold takeaway noodles under one spotlight. 2. The guest pokes the noodles, then drops his forehead onto the counter. 3. The chef notices him and sets down his clothes. 4. The chef rolls up his sleeves and tightens his apron. 5. He switches on the kitchen lights; the burners ignite. 6. He rapidly slices beef into neat pieces. 7. Orange flames rise from a copper pan; the chef remains calm. 8. The guest ducks as a pan flies past, his tie blown backward. 9. The chef juggles cooking tasks and tosses a pan overhead. 10. A rapidly spinning pepper grinder releases a huge pepper cloud. 11. The guest sneezes, then watches wide-eyed. 12. The chef catches two falling pans without spilling anything. 13. Overhead view of beautifully plated beef medallions, dark sauce and three tiny herbs. 14. The guest tastes the dish, closes his eyes and sheds a joyful tear. 15. The guest grins; the chef smirks and discards the takeaway box. 16. The kitchen goes dark except for the guest’s spotlight as he keeps eating. Maintain matching faces, costumes, props and kitchen layout. Lighting shifts from cool darkness to warm cooking light, then back to darkness. Clear panel borders, small numbers 1–16 only, no dialogue, captions, logos or watermark.
视频结果
提示词
Create a premium cinematic character bible sheet for CALVIN & BOB from central intelligence. LAYOUT: Split screen partner format. Two halves divided by a bold dramatic dividing element in the center. LEFT SIDE — CALVIN: Deep blue watercolor splash behind him fading into center. Large bold brushstroke text CALVIN top left in deep blue. Below small text: THE SPY / CENTRAL INTELLIGENCE. One massive dramatic cropped hero image of Calvin from mid-thigh up — enormous muscular build, light blue unicorn t-shirt, no fanny pack, serious focused expression, fist raised, deep blue watercolor splash radiating behind him. CENTER: Bold dramatic CENTRAL INTELLIGENCE in deep blue slightly worn. Below it small text: OFFICE PARK. UNLIKELY HEROES. RIGHT SIDE — BOB: Warm yellow and white watercolor splash behind him fading into center. Large bold brushstroke text BOB top right in deep yellow. Below small text: THE ACCOUNTANT / CENTRAL INTELLIGENCE. One massive dramatic cropped hero image of Bob from mid-thigh up — wide nervous eyes, white striped shirt, normal grey blue tie loosened, mouth open mid scream, warm yellow and white watercolor splash radiating behind him. BOTTOM CENTER: Color palette — deep blue, warm yellow, white, grey. Tagline centered: ONE WAS READY. ONE WAS NOT. OVERALL: Clean white background, deep blue watercolor left side, warm yellow watercolor right side, dramatic blue center divider, bold flat color blocking, chunky simplified forms, hard edge shadows, thick black outlines, vibrant saturated colors, minimal clean typography, cinematic cel-shaded 3D anime, hand-painted textures, not cartoon not Disney not Pixar, print ready.
视频结果
提示词
Create a 15-second cinematic anime cooking sequence in a stainless-steel kitchen with warm dramatic lighting. Follow the uploaded character sheet exactly: a tall, athletic, middle-aged chef with spiky blonde hair, blue eyes, a white chef uniform and black apron. No other chefs. 0–2s: Sear beef fillet, turning each side until golden. 2–4s: Finely chop mushrooms and cook into dark duxelles. 4–6s: Layer Parma ham and duxelles, then tightly roll around the beef using cling film. 6–8s: Remove the film, wrap in pastry, seal and brush with egg wash. 8–10s: Place in the oven; jump cut to the golden baked Wellington. 10–13s: Sound briefly drops as he slices through, revealing pink beef beneath crisp pastry. 13–15s: Wide hero shot of the steaming Wellington; the chef steps back proudly. Cel-shaded 3D anime, hand-painted textures, strong shadows and subtle film grain. Food-focused close-ups, occasional chef reactions, cuts every 2–3 seconds. Preserve cooking order, compressing preparation and baking through montage. Build orchestral music with sizzling and chopping sounds toward the finale.
在工作台、Playground 或 API 中使用 Flare 或 Sunburst。每张生成的图片单独计费。
描述想从零创建的图像,或添加参考图并说明想创建或修改的内容。
选择图片尺寸和质量,查看所需积分,然后点击生成。
查看并下载图片。如需继续编辑,上传结果并描述要修改和保留的内容。
浏览 Reddit、YouTube 和 X 上的用户评测、图片示例与实测演示,了解大家使用的提示词、分享的结果和遇到的限制。
r/ChatGPT
reddit
Image Gen 2.5 生成的图片太惊艳了!生图模型进步很大!
r/codex
reddit
这些对比图使用一系列 UI 生成提示词,比较 gpt-image-2、2.5 Flare 和 2.5 Sunburst。参考图和提示词都很有挑战性,需要模型协调处理。我觉得 2.5 版本明显更好。
r/ChatGPT
reddit
大家目前觉得怎么样?我刚收到可用通知,至少感觉速度快了很多。

OpenAI
ChatGPT Images 2.5 生图更快,结果更自然、更有辨识度,多轮编辑也能保持细节一致。用评论指定修改,同时保留图片其余部分。

Brock Mesarich
ChatGPT Image 2.5 刚刚发布!这期视频将讲解全部新变化,并进行实时测试。

Riko Nazza AI
介绍 OpenAI 新图像模型 ChatGPT Images 2.5 的全部更新,以及我如何用它制作 YouTube 缩略图。先完整解析生图工具的变化,再分享我的具体提示词和编辑流程。
从提示词、草图或照片出发,在 SeeGen AI 用 GPT Image 2.5 创作新作品,或完善已有图片。