OpenAI가 ChatGPT Images 2.5를 출시하며 더 선명한 디테일, 정밀한 편집, 최대 50% 빨라진 생성 속도를 약속했습니다. 두 모델(GPT-Image-2.5 Flare와 Sunburst)이 API에 공개되었으며, 요청한 부분만 수정하는 편집 기능이 크게 개선되었습니다. 다만 품질 등급별 가격 차이가 크고, ChatGPT 사용자에게 두 모델이 어떻게 배분되는지는 불분명합니다.
번역된 본문
챗GPT 이미지 2.5: 더 빠르고 정확하지만 모두에게 같지는 않다
막시밀리안 슈라이너 (2026년 9월 9일)
OpenAI가 'ChatGPT Images 2.5'라는 이름으로 두 개의 새 이미지 모델을 출시했다. 더 선명한 디테일, 더 정밀한 편집, 더 빠른 생성 속도를 약속한다. 그리기 도구와 공유 가능한 프롬프트 같은 새 기능은 창작 과정을 확장하는 것이 목표다.
OpenAI에 따르면 사용자는 현재 ChatGPT Images와 API의 GPT-Image 모델을 통해 매주 30억 장 이상의 이미지를 생성한다. ChatGPT Images 2.5에서 회사는 더 자연스러운 조명과 더 섬세한 질감을 약속하는 새로운 생성 방식을 선보였다. 새 모델은 참조 사진 속 피사체를 더 잘 보존하고, 여러 차례에 걸친 편집 지시를 더 안정적으로 따르도록 설계되었다. OpenAI에 따르면 빠른 버전은 Images 2.0 대비 이미지 생성 지연시간을 최대 50% 줄였다.
서로 다른 용도로 만들어진 두 모델
개발자를 위해 OpenAI는 두 모델을 모두 API에 제공한다. GPT-Image-2.5 Flare는 대부분의 용도에 적합한 기본 선택으로, GPT-Image-2보다 높은 화질을 50% 낮은 지연시간으로 제공한다. 더 강력한 모델인 GPT-Image-2.5 Sunburst는 편집에 대한 더 정밀한 제어가 필요한 까다로운 시각 작업을 겨냥하며, 그만큼 생성 시간이 더 걸린다.
두 모델의 토큰 요금은 동일하다. 이미지 입력 토큰 100만 개당 8달러, 출력 토큰 100만 개당 30달러다. 하지만 모델과 품질 등급에 따라 토큰 사용량이 다르기 때문에 같은 요금이라고 이미지당 비용이 같은 것은 아니다.
Images 2.5에는 기존 최고 등급인 'high'를 넘어서는 'xhigh'와 'max' 품질 등급이 새로 추가됐다. 1024x1024 이미지 기준으로 'low' 등급은 약 0.006달러로 전작과 동일하고, 'high'는 약 0.053달러, 새로운 'max' 등급은 약 7,024개의 출력 토큰으로 대략 0.21달러다. 즉 Images 2.5의 'max' 등급은 GPT-Image-2의 'high' 등급과 같은 가격이다. 전작과 달리 Images 2.5에는 아직 더 저렴한 배치(batch) 요금이 없다.
초기 테스트에서는 동일한 토큰 가격에도 불구하고 Sunburst가 Flare보다 이미지당 비용이 보통 더 높게 나왔는데, 추론 과정이 더 길어서일 가능성이 크다. 또 이번에는 OpenAI가 평균 이미지당 가격을 제공하지 않았다.
ChatGPT 사용자에게 두 모델이 어떻게 배분되는지도 불분명하다. 발표 자료와 문서 어디에도 ChatGPT가 언제 빠른 Flare를 쓰고 언제 더 정밀한 Sunburst를 쓰는지 나와 있지 않다. API에서는 모델을 명시적으로 선택할 수 있지만, ChatGPT 인터페이스에는 아직 그런 제어 기능이 없다. 테스트 결과, 현재 그 경계는 주로 Chat과 Work 사이에 있는 것으로 보인다. Work에서는 추론 설정과 무관하게 프롬프트가 요청한 부분만 실제로 변경됐다. 이러한 정밀 편집이 이번 모델 개선의 핵심이다. 반면 Chat에서는 추론을 높음으로 설정해도 후속 이미지에서 다른 디테일까지 계속 바뀌었다. '6 Pro' 설정에서만 가끔 더 강력한 모델이 작동하는 듯했다.
요청한 부분만 변경하는 편집
언급했듯이 새 모델의 핵심은 편집이다. 더 복잡한 피사체와 배경에서도 요청된 요소만 변경하고 이미지의 나머지 부분은 그대로 유지하도록 설계되었다. 긴 대화에서도 여러 편집 단계를 거쳐도 이미지 품질이 떨어지지 않고 이전 변경 사항이 일관되게 유지되어야 한다.
OpenAI는 예를 들어 방 재구성 시연으로 이를 보여줬다. 이전 모델이 편집할 때마다 다른 디테일까지 바꿨다면, 새 모델은 여러 차례 반복해도 안정적으로 유지된다.
GPT-6 Astra(Max)를 통한 ChatGPT Work에서의 간단한 테스트로 이 기능이 얼마나 잘 작동하는지 확인했다. 다음 프롬프트를 사용한 뒤 작은 디테일 하나(바나나 색)와 큰 요소 하나(큰 고양이)를 반복적으로 수정했다.
"초현실적인 DSLR 사진. 전경에 분홍색 바나나를 든 원숭이가 호랑이 위에 앉아 있다. 배경에는 말이 우주비행사를 타고 있다. 우주비행사는 아래에 있으며 살아있는 '우주복 말 안장'처럼 보이고, 말은 명확히 위에 있어 통제권을 쥔 기수처럼 보인다. 100% 명확하게 만들어라."
ChatGPT Images 2.5: Faster, more precise, but not the same for everyone Maximilian Schreiner View the LinkedIn Profile of Maximilian Schreiner Sep 9, 2026 GPT-Images-2.5 prompted by OpenAI OpenAI has released two new image models under ChatGPT Images 2.5, promising sharper details, more precise editing, and faster generation. New features like a drawing tool and shareable prompts aim to expand the creative process. OpenAI says users now generate more than three billion images each week through ChatGPT Images and the GPT-Image models in the API. With ChatGPT Images 2.5, the company is rolling out a reworked generation that promises more natural lighting and finer textures. The models are meant to preserve subjects from reference photos better and follow editing instructions more reliably across multiple rounds. According to OpenAI, the faster variant cuts image generation latency by up to 50 percent compared to Images 2.0. Two new models built for different jobs For developers, OpenAI brings both models to the API. GPT-Image-2.5 Flare is the default pick for most uses, with higher image quality than GPT-Image-2 at 50 percent lower latency. The stronger model, GPT-Image-2.5 Sunburst, targets more demanding visual work with tighter control over edits, and it needs longer generation times to deliver. Both models use the same token rates. That's eight dollars per one million image input tokens and 30 dollars per one million output tokens. But since token use varies by model and quality tier, the same rates don't mean the same cost per image. New in Images 2.5 are the "xhigh" and "max" quality tiers, which go beyond the previous ceiling of "high." A 1024x1024 image at the "low" tier costs about 0.006 dollars, same as the predecessor, while "high" runs about 0.053 dollars, and the new "max" tier lands at roughly 0.21 dollars with around 7,024 output tokens. That puts the "max" tier of Images 2.5 at the same price as the "high" tier of GPT-Image-2. Unlike the predecessor, Images 2.5 has no cheaper batch rate so far. In early tests, though, Sunburst usually costs more per image than Flare despite identical token prices, probably because of longer reasoning runs. And unlike last time, OpenAI gives no average price-per-image figure. It's also unclear how OpenAI routes ChatGPT users between the two models. Neither the announcement nor the documentation says when ChatGPT reaches for the faster Flare or the more precise Sunburst. In the API you can pick the model explicitly. The ChatGPT interface offers no such control yet. In our tests, the line currently seems to run mainly between Chat and Work. In Work, our prompts really do change only what's asked, no matter the reasoning settings. These targeted edits are the focus of the model improvements. In Chat, though, more details keep shifting in the follow-up images, even with reasoning set to high. Only at the "6 Pro" setting does the stronger model sometimes appear to kick in. Editing that changes only what you ask As noted, the new model's focus is editing. It's meant to change only the requested elements and leave the rest of an image untouched, even with more complex subjects and backgrounds. In longer conversations, earlier changes should stay consistent without image quality dropping over multiple editing steps. OpenAI shows this with a room redesign, for example. Where the older model changed other details with every edit, the new model stays stable even across several iterations. A quick test through ChatGPT Work with GPT-6 Astra (Max) shows how well this works. We used the following prompt and then iteratively adjusted one small detail (banana color) and one large one (the big cat). A hyper-realistic DSLR photo. A monkey holding a pink banana is sitting on a tiger in the foreground. In the background, a HORSE is RIDING AN ASTRONAUT. The astronaut is underneath, like a living “spacesuit horse saddle,” and the HORSE is clearly on top, in control, as the rider. Make it 100% unambiguous: the HORSE is the rider and the ASTRONAUT is being ridden, NOT the other way around. High resolution, sharp focus, realistic lighting. Image: GPT-Images-2.5 / Astra (Max) On a side note, this is probably the best version of a horse riding an astronaut that an OpenAI image model has produced in our tests. Here's how it looked with Image 2.0 (Thinking variant). Tests in ChatGPT's Chat mode, by contrast, always changed the rest of the image when we adjusted the banana color. Only in "6 Pro" mode did the image stay consistent in one run, but not in another. This may change over the course of the week as the rollout continues. Images 2.5 is also meant to handle complex visual instructions better, deliver more accurate content for real-world information, and work with transparent backgrounds and more demanding layouts. In our last article, for instance, we had ChatGPT turn the piece into an 80s magazine spread. It looked like this. GPT-Images-2.5 via Astra (Max) also produced a detailed magazine with a sample image based on our monkey-astronaut prompt, in a weaker and a stronger variant, without losing sight of the original instruction for worse quality. The model corrected an error we planted (GPT-Images-2.5 instead of 2.0 right above the images on the first page) on command, showing how well it preserves the rest of the content, and it did the same for a translation on the simple request "Create a US English version." Image: GPT-Images-2.5 / Astra (Max) For comparison, the results through plain ChatGPT Chat are also worth a look, but they show that the weaker model changes other details during translation too. In one case, though, it nailed the current Microsoft CEO better. Image: GPT-Images-2.5 / ChatGPT Chat (Medium and Instant) Drawing, templates, and shareable prompts In ChatGPT, OpenAI is adding several new features alongside the new models. The most important one is called "Sketch." It lets users draw directly in ChatGPT and use the sketch as a visual template for the finished image. You activate it with the "@Sketch" command, and OpenAI says it works well for diagrams, room layouts, or posters. Templates are ready-made prompts meant to make it easier to get started with formats like posters, logos, infographics, thumbnails, illustrations, or ads. They give you a structure instead of a blank canvas and pin down your requirements through targeted follow-up questions. Users can also place comments directly on images and share the prompts they used, so others can try the same idea with their own photos and details. As an example, OpenAI points to a currently viral prompt that generates portraits in 1980s style. Both models top the Arena ranking In Arena's text-to-image leaderboard, the new models hold the top two spots for now. GPT-Image-2.5 Sunburst leads with a score of 1421, followed by GPT-Image-2.5 Flare at 1399. Both scores carry a "Preliminary" label and rest on still-low vote counts of around 3,100 and 2,900. In third place is the predecessor GPT-Image-2 with 1381 points from about 78,700 votes. Behind them come Microsoft's mai-image-2.6 (1331), SpaceXAI's grok-imagine-image-2.0 (1315), and several models from Reve, Meta, Google, and Bytedance. Since the ratings for the two new OpenAI models are preliminary, their position could still shift as more votes come in. Watermarking in partnership with Google DeepMind For provenance labeling, OpenAI still relies on the C2PA industry standard, which embeds metadata for tracking. To make provenance more resistant, OpenAI also adds an invisible watermark via Google DeepMind's SynthID, in ChatGPT, Codex, and the API. OpenAI says there's no single solution for provenance labeling, which is why it takes a layered approach. Images 2.5 is available now worldwide for all ChatGPT, ChatGPT Work, and Codex users across desktop, mobile, and web, including the free version. OpenAI says shorter wait times and higher usage limits apply. Our tests suggest the Chat mode currently uses the wea