30 秒视频,50+ 参考图。 立即体验
GPT Image 2.5

ChatGPT Image 2.5 AI 图片生成器

OpenAI 于 2026 年 9 月 8 日发布 GPT Image 2.5。通过文本或参考图创建产品照片、海报和人像,更精确地调整背景、文字和细节,控制哪些内容改变、哪些内容保留。

Flare 和 Sunburst 已支持 medium/high/xhigh/max 画质、1K/2K/4K 预设及 PNG/JPEG/WebP 输出。实际尺寸可能与所选预设不同。

简单人物草图转化为粉色头发的插画角色

GPT Image 2.5 与 GPT Image 2:有哪些变化?

GPT Image 2.5 的细节更自然,更贴近参考图,编辑控制更好。Flare 还缩短了生成时间,Sunburst 则注重更精细的编辑。

功能GPT Image 2GPT Image 2.5
发布日期2026 年 4 月 21 日2026 年 9 月 8 日
图像细节写实图像生成更自然的光线和更精细的纹理
参考图片使用参考图引导生成结果更好地保留人脸、产品及其他主体的辨识度
局部编辑支持通过文本指令修改修改更精准,减少对图像其他部分的意外改动
多轮编辑的一致性支持后续编辑更好地保留此前的修改和细节
生成速度OpenAI 对比的基准Flare:据 OpenAI 表示,延迟降低 50%
Arena 文生图排名第 3 名 · 评分:1381Sunburst 第 1 名 · 1421;Flare 第 2 名 · 1399
Arena 单图编辑排名第 3 名 · 评分:1461Sunburst 第 1 名 · 1520;Flare 第 2 名 · 1491

排名来自 Arena 2026 年 9 月 7 日榜单快照,于 9 月 9 日核对。GPT Image 2.5 的结果标记为 Preliminary(初步结果),可能随投票增加而变化。GPT Image 2 条目使用 medium 质量。速度提升仅针对 Flare,不适用于 Sunburst。

ChatGPT Image 2.5 核心功能

从提示词、草图或参考图开始,创作完整画面,再调整细节,无需重做整张图。

将草图变成完整图像
01

将草图变成完整图像

把简单绘画转化为产品照片、室内设计或 3D 风格场景。上传草图并描述期望的效果,用草图引导物体位置和比例,再由 GPT Image 2.5 添加颜色、光线和纹理。

图像帧
图像帧
动态 GIF
动态 GIF
02

为 GIF 和定格动画创建帧图

生成一组姿势、表情或物体位置略有变化的图像,再使用独立编辑工具将这些帧合成为循环 GIF 或定格风格动画。各帧沿用同一参考图,有助于保持角色和场景一致。

编辑图片,无需从头开始
03

编辑图片,无需从头开始

替换背景、改变产品颜色,或更新海报文字。描述需要修改和保留的内容,GPT Image 2.5 可更精准地编辑,帮助保留主体、布局和周围细节。

保留参考主体的辨识度
04

保留参考主体的辨识度

将同一产品置于不同场景,或为人像探索新风格。参考图可引导结果,更好地保留原始特征,帮助制作风格统一的产品照片或角色图。

创建带文字与排版的海报
05

创建带文字与排版的海报

设计活动海报、邀请函和社交媒体图片,让文字融入画面。指定准确文案、出现位置及重点元素。发布前请检查拼写和小字。

生成细节自然的图像
06

生成细节自然的图像

生成人像自然的皮肤纹理、清晰的服装面料细节和真实的产品反光。描述材质、光源和环境,引导生成柔光影棚照片或户外场景。

GPT Image 2.5 示例与提示词

GPT Image 2.5 示例 1

An expedition team in weathered EVA suits discovering cyclopean alien architecture half-buried in blue ice on a distant moon, a colossal stone face staring from a cliff, aurora lights dancing overhead, frost crystals catching the light of a distant binary star, footprints in untouched snow, sense of ancient dread, cold blue-cyan palette with warm suit lights, National Geographic meets Alien, cinematic documentary still.

GPT Image 2.5 示例 2

Turn the reference sketch into a photorealistic autumn portrait, preserving its pose and composition. A young adult woman with long brown hair sits on a wooden swing, holding both ropes and smiling gently. She wears a cream knit sweater, burgundy plaid scarf, and blue jeans. Shoot from a close, slightly low angle beneath orange-red maple leaves, with a rustic fence and wooded hills behind her. Soft daylight, natural skin and fabric texture, windblown hair, subtle film grain. 16:9 landscape, finished photo only, no sketch or text.

GPT Image 2.5 示例 3

A whimsical cinematic 3D illustration of frightened, dirty fruits and vegetables on a wooden kitchen counter. A large red apple stands in the center, surrounded by a tomato, lettuce, carrot, and broccoli, all with expressive eyes and worried faces. Soil, caterpillars, flies, and a snail cover their surfaces. A hand holds a magnifying glass in the foreground, revealing colorful cartoon microbes on the apple’s skin. Warm sunlight, a cozy rustic kitchen, detailed textures, shallow depth of field, playful educational advertising style. Background signs read "FRUTAS Y VERDURAS MÁS LIMPIAS = UNA FAMILIA MÁS SANA" and "LO QUE NO VES TAMBIÉN IMPORTA." Landscape composition.

GPT Image 2.5 示例 4

A candid full-body photo of a young adult woman at an outdoor flower market, holding a bouquet wrapped in brown paper. White camisole, loose light-blue jeans, ivory shoulder bag, tousled brown hair, looking down with a gentle smile. Colorful flower stalls, soft daylight, natural skin texture, subtle film grain.

GPT Image 2.5 示例 5

Find exactly 100 distinct, real, drawable physical things whose established English names start with X. Include varied objects, organisms, foods, materials, and artifacts. Exclude invented names, "Xmas," abstract terms, symbols, software, duplicates, synonyms, trivial variants, and generic objects padded with X-word modifiers. Allow established compounds, common object abbreviations such as "XLR connector," accepted loanwords, and accurately drawable scientific terms. Verify obscure terms and avoid misleading names or illustrations. For each entry, provide: number, exact name, category, brief definition, visual description, and High/Medium confidence. Check every entry twice for validity, tangibility, visual distinctness, accurate depiction, and suitability for a children’s educational chart. Exclude borderline entries where possible; identify any retained ones. Add a "Rejected candidates" section with brief reasons. Prioritize accuracy over quantity. If fewer than 100 qualify, state this and provide the best verified set without inventing fillers. Only after validating 100 entries, create a 10 × 10 picture grid. Each cell must show exactly one approved item, its number, and its unchanged name. Use a white background, consistent illustrations, clear cell divisions, large readable labels, and no decorative objects. Depictions must match the validated definitions precisely.

GPT Image 2.5 示例 6

Photorealistic Paris street portrait, vertical 9:16. A young Southeast Asian hijabi woman walks forward, turning back toward the camera with a calm, curious expression. She wears a flowing taupe hijab, oversized emerald blazer, beige tulle maxi skirt, and brown leather shoulder bag, carrying blush roses wrapped in kraft paper. The Eiffel Tower is recognizable in the upper-right background. Slow-shutter panning creates horizontal background blur while her face stays relatively sharp. Wind moves her hijab and skirt. Slight Dutch angle, 50mm perspective, soft afternoon backlight, muted natural colors, creamy highlights, and subtle film grain. Romantic, candid, and realistic. No aged-photo effects, heavy grain, orange grading, plastic skin, CGI, text, or watermark.

GPT Image 2.5 示例 7

Photorealistic travel portrait of a young East Asian woman on a sandy shore beside moss-covered rocks, with a grand medieval stone abbey rising on a rocky island behind her. Long dark brown hair, gentle smile, looking at the camera. She wears an oversized black coat with hands in pockets, a cream scarf, and a subtle chain shoulder bag. Soft evening light, pale blue sky, detailed stone architecture, natural skin and fabric textures, candid smartphone-photo feel.

GPT Image 2.5 示例 8

Makoto Shinkai-style anime illustration of a boy with a backpack skateboarding down a steep coastal street, viewed from behind. A sparkling turquoise ocean and crescent beach on the left, a detailed seaside town on the right, towering white clouds in a brilliant blue sky. Vivid summer colors, cinematic sunlight, nostalgic atmosphere.

GPT Image 2.5 示例 9

Create a charming handcrafted miniature travel scene featuring [ICONIC STRUCTURE] as the main focal point. Show the landmark as a beautifully sculpted tiny 3D model, with soft rounded details, handmade textures, delicate imperfections, and a whimsical storybook feeling. Surround it with a few subtle elements that represent its location—such as tiny trees, flowers, streets, boats, mountains, clouds, or local objects—without making the scene crowded. Place everything on a clean warm-white textured paper background, with plenty of elegant negative space. Add a small tasteful wooden or paper travel plaque containing: [STRUCTURE NAME] [CITY, COUNTRY] Famous for: [SHORT UNIQUE FACT] Use soft natural lighting, gentle shadows, pastel yet realistic colors, miniature diorama depth, handcrafted clay/paper textures, and a premium cute travel-journal aesthetic. Centered composition, highly detailed landmark, adorable but sophisticated, clean and collectible travel-card design, no photorealistic people, no clutter.

GPT Image 2.5 示例 10

A candid fashion portrait of a young man with tousled dark hair sitting among empty blue velvet theater seats. Brown suede jacket, cream shirt, loose burgundy tie, ivory wide-leg trousers, and dark leather boots. Holding a small silver camera, looking off-camera. Direct flash, natural textures, subtle film grain, vintage editorial style. Landscape, approximately 5:4.

GPT Image 2.5 示例 11

Create image of Magazine feature article [travel] guide page, cute, information dense photo book style magazine feature article page. Add all necessary sections, tips, recommendations, information. add photos for any sections and recommendations if you like. Place the attached person at the precise location of [Plateau of Gorgoroth in Mordor]. Seamlessly blend the attached person as if they are sightseeing. Approach this task with the understanding that this is a critical, information rich page that will significantly influence visitor numbers, text accuracy is important.

GPT Image 2.5 示例 12

A minimalist illustrated poster inspired by Léon: The Professional. A tall man in sunglasses and a long black coat holds a suppressed pistol upright beside a young girl with a short black bob, an oversized green jacket, shorts, and black boots. She holds a potted plant; a suitcase rests between them. Huge red "LEON" lettering fills the upper background, with a simplified gray city skyline below. Angular geometric shapes, elongated proportions, dramatic shadows, subtle paper texture, and a restrained red, black, cream, and green palette.

GPT Image 2.5 示例 13

Transform my photo into an authentic 1980s retro Indian fashion portrait, inspired by the reference image. Keep the exact face, facial features, face structure, skin tone, identity, and natural proportions unchanged. Do not alter or beautify the face. Dress her in a full-sleeve pink shirt tucked neatly into a black high-waisted skirt with a stylish belt. Remove the scarf completely. Add simple 1980s-inspired accessories. Give her bouncy, voluminous retro hair styled with a cute ribbon. Create an authentic 1980s background and atmosphere. Use a Kodak film-camera look with warm faded colors, subtle film grain, soft analog texture, gentle vintage lighting, and natural skin texture. Make it look like a genuine photograph taken in the 1980s, not a modern photo with a vintage filter. No watermark, no text, no distortion.

GPT Image 2.5 示例 14

A documentary wildlife photograph of dozens of seagulls resting on brown seaweed-covered tidal flats, scattered between winding pools of vivid blue seawater. Some stand with clear reflections, others rest or stretch their wings. Elevated viewpoint, layered horizontal composition, strong natural sunlight, crisp shadows, detailed coastal textures, realistic colors. Landscape format, no text or interface elements.

GPT Image 2.5 示例 15

A dramatic low-angle photograph of a male cyclist resting astride a mountain bike on a rocky countryside trail, drinking from a water bottle. Black cycling kit, red-and-black helmet, sunglasses. The large front tire dominates the foreground, surrounded by dry golden grass and distant wind turbines beneath a turquoise sky with towering clouds. Warm late-afternoon sunlight, long shadows, rich earthy colors, detailed gravel, wide-angle perspective. Vertical composition, realistic outdoor sports photography.

GPT Image 2.5 示例 16

Create a professional 16:9 character reference sheet based strictly on the uploaded image. Preserve the character’s exact face, age, hairstyle, body proportions, costume, accessories, colors, and distinctive details without redesign. Include: One large hero portrait. Five aligned full-body views: front, front three-quarter, side, back three-quarter, and back. Six expressions: neutral, happy, serious, angry, surprised, and sad. 4–6 natural full-body poses. Costume, accessory, and material close-ups. A faithful color palette, proportion guides, and silhouette thumbnail. Use a clean off-white background, organized grid, generous spacing, readable labels, consistent scale, neutral studio lighting, and detailed materials. Keep identity and left/right design details consistent throughout. No overlapping or cropped figures, anatomical errors, extra props, clutter, logos, or watermarks. Target resolution: 15360 × 8640.

ChatGPT Image 2.5 x Seedance 2.5

最后一道菜

视频结果

提示词

Create ONE 16:9 landscape storyboard image with exactly 16 panels in a clean 4×4 grid, read left to right, top to bottom. Polished 3D animation style, expressive comedy, detailed food, cinematic lighting. Keep two consistent characters: a tall, lean, long-limbed chef with a dark moustache, white jacket and apron; and a short, round, exhausted guest with messy hair, a wrinkled shirt and loose tie. Setting: a closed restaurant kitchen with steel counters, copper pans and large wall switches. Panels: 1. The chef wipes the counter; the guest slumps beside cold takeaway noodles under one spotlight. 2. The guest pokes the noodles, then drops his forehead onto the counter. 3. The chef notices him and sets down his clothes. 4. The chef rolls up his sleeves and tightens his apron. 5. He switches on the kitchen lights; the burners ignite. 6. He rapidly slices beef into neat pieces. 7. Orange flames rise from a copper pan; the chef remains calm. 8. The guest ducks as a pan flies past, his tie blown backward. 9. The chef juggles cooking tasks and tosses a pan overhead. 10. A rapidly spinning pepper grinder releases a huge pepper cloud. 11. The guest sneezes, then watches wide-eyed. 12. The chef catches two falling pans without spilling anything. 13. Overhead view of beautifully plated beef medallions, dark sauce and three tiny herbs. 14. The guest tastes the dish, closes his eyes and sheds a joyful tear. 15. The guest grins; the chef smirks and discards the takeaway box. 16. The kitchen goes dark except for the guest’s spotlight as he keeps eating. Maintain matching faces, costumes, props and kitchen layout. Lighting shifts from cool darkness to warm cooking light, then back to darkness. Clear panel borders, small numbers 1–16 only, no dialogue, captions, logos or watermark.

意想不到的英雄

视频结果

提示词

Create a premium cinematic character bible sheet for CALVIN & BOB from central intelligence. LAYOUT: Split screen partner format. Two halves divided by a bold dramatic dividing element in the center. LEFT SIDE — CALVIN: Deep blue watercolor splash behind him fading into center. Large bold brushstroke text CALVIN top left in deep blue. Below small text: THE SPY / CENTRAL INTELLIGENCE. One massive dramatic cropped hero image of Calvin from mid-thigh up — enormous muscular build, light blue unicorn t-shirt, no fanny pack, serious focused expression, fist raised, deep blue watercolor splash radiating behind him. CENTER: Bold dramatic CENTRAL INTELLIGENCE in deep blue slightly worn. Below it small text: OFFICE PARK. UNLIKELY HEROES. RIGHT SIDE — BOB: Warm yellow and white watercolor splash behind him fading into center. Large bold brushstroke text BOB top right in deep yellow. Below small text: THE ACCOUNTANT / CENTRAL INTELLIGENCE. One massive dramatic cropped hero image of Bob from mid-thigh up — wide nervous eyes, white striped shirt, normal grey blue tie loosened, mouth open mid scream, warm yellow and white watercolor splash radiating behind him. BOTTOM CENTER: Color palette — deep blue, warm yellow, white, grey. Tagline centered: ONE WAS READY. ONE WAS NOT. OVERALL: Clean white background, deep blue watercolor left side, warm yellow watercolor right side, dramatic blue center divider, bold flat color blocking, chunky simplified forms, hard edge shadows, thick black outlines, vibrant saturated colors, minimal clean typography, cinematic cel-shaded 3D anime, hand-painted textures, not cartoon not Disney not Pixar, print ready.

惠灵顿牛排

视频结果

提示词

Create a 15-second cinematic anime cooking sequence in a stainless-steel kitchen with warm dramatic lighting. Follow the uploaded character sheet exactly: a tall, athletic, middle-aged chef with spiky blonde hair, blue eyes, a white chef uniform and black apron. No other chefs. 0–2s: Sear beef fillet, turning each side until golden. 2–4s: Finely chop mushrooms and cook into dark duxelles. 4–6s: Layer Parma ham and duxelles, then tightly roll around the beef using cling film. 6–8s: Remove the film, wrap in pastry, seal and brush with egg wash. 8–10s: Place in the oven; jump cut to the golden baked Wellington. 10–13s: Sound briefly drops as he slices through, revealing pink beef beneath crisp pastry. 13–15s: Wide hero shot of the steaming Wellington; the chef steps back proudly. Cel-shaded 3D anime, hand-painted textures, strong shadows and subtle film grain. Food-focused close-ups, occasional chef reactions, cuts every 2–3 seconds. Preserve cooking order, compressing preparation and baking through montage. Build orchestral music with sizzling and chopping sounds toward the finale.

如何在 SeeGen AI 使用 GPT Image 2.5

在工作台、Playground 或 API 中使用 Flare 或 Sunburst。每张生成的图片单独计费。

1

添加提示词或图片

描述想从零创建的图像,或添加参考图并说明想创建或修改的内容。

2

选择设置并生成

选择图片尺寸和质量,查看所需积分,然后点击生成。

3

查看并下载

查看并下载图片。如需继续编辑,上传结果并描述要修改和保留的内容。

社区热议

GPT Image 2.5 评测与社区讨论

浏览 Reddit、YouTube 和 X 上的用户评测、图片示例与实测演示,了解大家使用的提示词、分享的结果和遇到的限制。

Reddit 用户体验与讨论

YouTube 评测与教程

X 示例与初步体验

ChatGPT Image 2.5 图片生成器常见问题

Flare 和 Sunburst 有什么区别?+
Flare 注重更快的图片生成;Sunburst 更适合注重精度的细节编辑。两者都能通过文本创建图片,并使用参考图。
可以同时使用草图、照片和风格参考图吗?+
可以。一起上传并说明每张图的用途,例如:“遵循草图的布局,保留照片中的人物,使用风格参考图的配色。”
如何告诉模型该参考哪张图?+
按上传顺序引用图片,说明从每张图中采用什么。例如:“用 @Image1 的脸和 @Image2 的服装,保留 @Image1 的背景。”
如何只修改图片的一个部分?+
说明需要修改和保留的内容。例如:“把外套改成蓝色,保持脸、姿势、光线和背景不变。”仍可能出现细微的意外改动,请检查结果。
为什么多次编辑后细节会改变?+
每次编辑都可能引入小差异,并在多轮编辑中累积。重复使用原始参考图,并重申需要保留的细节,例如人脸、产品外形或标志位置,可减少这种变化。
GPT Image 2.5 能添加不同语言的文字吗?+
可以,但准确度会受到语言、字体和文字大小影响。用引号提供准确文案,并说明位置。发布前请检查拼写和小字。
GPT Image 2.5 能创建 GIF 吗?+
它生成静态图片,不能直接生成完整动态 GIF。可以生成一组帧图,再在 GIF 编辑器中合成。保持背景和构图一致,让帧之间只有少量变化。
如何将图片变成 Seedance 2.5 视频?+
下载图片并上传到 Seedance 2.5,或直接从图库选择,作为参考图或起始帧。描述需要的运动、镜头动作和时序。GPT Image 2.5 创作画面,Seedance 2.5 添加动态。

用 GPT Image 2.5 实现你的创意

从提示词、草图或照片出发,在 SeeGen AI 用 GPT Image 2.5 创作新作品,或完善已有图片。