GPT-image-2 API de imagem com IA
GPT Image 2 é o modelo de imagem da OpenAI exposto no fluxo de geração da Flatkey. Use /v1/images/generations e configure quantidade, tamanho, qualidade, formato, fundo e moderação antes de enviar a solicitação.
Saída
Prévia
Resumo da solicitação
- Endpoint da API
- POST /v1/images/generations
- ID do modelo
- GPT-image-2
- Proporção
- 1:1
- Resolução
- 1024x1024
- Qualidade
- high
- Saídas
- 1
Somente prévia. Entre para enviar esta solicitação à Flatkey.
{
"model": "gpt-image-2",
"prompt": "Goal: Create a glossy miniature 3D diorama, like a handcrafted ceramic toy “box garden,” filled with salsa dancing, Cuban travel, music, plants, and colorful personal-favorite objects.\n\nCanvas: Square 1:1 image, front-facing three-quarter view of a small white corner room on a white base, photographed like a high-end product shot against a clean pale background. Use soft studio lighting, crisp focus, bright saturated colors, shiny enamel/plastic/ceramic textures, and charming miniature scale.\n\nMain subject: In the center, place exactly 2 salsa dancers: a woman in a flowing glossy {argument name=\"dress color\" default=\"red\"} salsa dress with matching heels, hair in a bun with red flowers, and a man in a shiny black suit with a red shirt and black fedora. They are mid-dance in a dramatic close embrace, one arm raised, posed on a small round patterned rug.\n\nRoom layout and counted objects: Build a cozy L-shaped corner room with exactly 2 white walls and 1 white floor base. Include exactly 1 blue shuttered window on the left wall with flower boxes. Include exactly 1 black wall lantern, exactly 1 framed tropical beach picture, exactly 1 small chalkboard sign reading “¡BAILA!” with a heart, exactly 1 blue folding chair, exactly 1 round blue café table, exactly 1 white coffee cup on a saucer, and exactly 1 yellow potted pink flower on or beside the table. On the back wall, hang exactly 1 clothesline with exactly 5 clipped travel photos/postcards, including one that reads “CUBA.” Add exactly 1 large yellow poster on the right back wall reading “SALSA” with silhouettes of dancers and the word “LIVE” near the bottom.\n\nMusic and dance details: At the front left, include exactly 1 black DJ controller/turntable console with exactly 2 visible vinyl records, one red-centered and one blue-centered. Near the front center, place exactly 1 glossy black cat figurine with bright yellow eyes, sitting beside exactly 2 golden music notes and exactly 1 black maraca-like handheld instrument. Keep the vibe playful and musical, like a tiny salsa party scene.\n\nBookshelves and decor: On the right wall, place exactly 1 blue bookshelf with exactly 2 shelves of colorful books and small decorative items. Put exactly 3 potted plants on top of the shelf: one yellow pot with green cactus and pink flower, one red mug-shaped pot with yellow flowers, and one small green pot. Include another shelf section with books and a small face-like pot. In the room overall, include exactly 9 visible potted plants or succulent arrangements: the yellow pink-flower pot near the café table, a green plant near the dancers, a pink succulent, a red cactus pot on the back shelf, a green succulent mound behind the dancers, a yellow pot with small plant near the stereo shelf, the yellow cactus pot on the bookshelf, the red mug flower pot on the bookshelf, and a large red pot with leafy green plant at the front right.\n\nTravel and beach details: At the front right, add exactly 1 small toy sailboat with a blue triangular sail, white triangular sail, brown mast, and white hull, floating on exactly 1 stylized blue wavy water base. Near the lower right foreground, place exactly 1 fan of color swatches containing many rainbow rectangles and exactly 1 black pen or marker laid across it.\n\nStyle constraints: Make everything miniature, rounded, glossy, tactile, and highly detailed, like a joyful collectible clay/ceramic diorama. Use a cheerful palette dominated by red, blue, yellow, green, and white. Ensure the scene feels packed but organized, with no human-scale realism, no blur, no dark mood, no extra text beyond the visible signs and posters, and no watermark.",
"n": 1,
"size": "1024x1024",
"quality": "high",
"output_format": "png",
"background": "opaque",
"moderation": "auto"
}capacidades
Desempenho e disponibilidade da API de geração de imagens GPT Image 2
A telemetria de solicitações da Flatkey para GPT Image 2 aparece quando há tráfego suficiente; nenhum benchmark ou nota de qualidade é inferido.
API
Uso da API de geração de imagens GPT Image 2 e atividade
Esta seção reflete solicitações reais da Flatkey para GPT Image 2 e permanece sem dados até que haja tráfego suficiente.
Preços
Preços da API do GPT Image 2 por token de entrada de imagem
Os preços abaixo são calculados com os dados de preços da Flatkey para este modelo e os grupos visíveis retornados pela nossa API de preços.
Adicionar créditos
Use o mesmo saldo e a mesma chave de API Flatkey para modelos de imagem, vídeo, áudio e texto.
- Tipo de modelo
- Texto para imagem
- API
- /v1/images/generations
- Base de cobrança
- imagem
- Saídas
- 1–10
Capacidades de geração e edição de imagens do GPT Image 2
Os campos de integração documentados aparecem abaixo; nenhuma promessa de benchmark ou qualidade é inferida.
Texto para imagem e edição
Crie imagens novas ou edite uma existente por /v1/images/generations e /v1/images/edits.
Criação guiada por referência
Use contexto visual para preservar assunto e composição em peças de produto, editorial e campanha.
Variações visuais flexíveis
Transforme um briefing em peças quadradas, verticais ou horizontais para cada canal.
Entrega segura para produção
Prepare resultados com o comportamento documentado de qualidade, formato, fundo e moderação.
Comparar
GPT Image 2 e GPT Image 1: controles da API de imagens
Os campos de integração documentados aparecem abaixo; nenhuma promessa de benchmark ou qualidade é inferida.
| Capacidade | GPT Image 1 | GPT Image 2 |
|---|---|---|
| Endpoint da API | Desconhecido aqui | /v1/images/generations; /v1/images/edits |
| Tamanhos | Desconhecido aqui | 1024x1024, 1536x1024, 1024x1536, auto |
| Formatos | Desconhecido aqui | PNG, JPEG, WebP |
| Qualidade/fundo/moderação | Desconhecido aqui | Documentado acima |
| Ranking de qualidade da imagem | Não declarado | Não declarado |
Escolha a solicitação certa antes de integrar
Confira fluxos, custo atual e o contrato exato de parâmetros da Flatkey.
Calculada com a configuração atual e a tarifa do catálogo quando há preço unitário fixo.
- Configuração atual
- 1024x1024 · high · 1×
- Tarifa ao vivo
- $0.0088 / imagem
- Solicitação estimada
- $0.0088
A cobrança final depende da solicitação e do grupo; preços por tokens ou faixas não viram uma estimativa fixa.
Fluxos e Model ID
Use apenas fluxos disponíveis. A Flatkey mantém um ID público e o corpo escolhe a entrada.
Texto para imagem
Gere uma imagem a partir de texto.
- ID do modelo
gpt-image-2- Endpoint
/v1/images/generations
Imagem guiada por referência
Use referências para editar ou variar imagens.
- ID do modelo
gpt-image-2- Endpoint
/v1/images/generations
Compatibilidade de parâmetros
Valores da mesma configuração usada no Playground.
nImagens1–10
Padrão: 1sizeTamanho1024x1024 · 1536x1024 · 1024x1536 · auto
Padrão: 1024x1024qualityQualidadeauto · high · medium · low
Padrão: highoutput_formatFormato de saídapng · jpeg · webp
Padrão: pngbackgroundFundoopaque · auto
Padrão: opaquemoderationModeraçãoauto · low
Padrão: autoimageLimite de referências4 arquivos
Endpoint: /v1/images/generationsDo primeiro teste à produção
Use o mesmo contrato no Playground e na API e trate validação, saldo, estado e resultado.
Ciclo da solicitação
- 1
Validar modelo, prompt, referências e saída.
- 2
Enviar a solicitação de imagem.
- 3
Ler o resultado ou tarefa assíncrona.
- 4
Armazenar arquivo e metadados.
Verificações de produção
Bloqueie valores e referências incompatíveis.
Trate chave, saldo, limites e grupo separadamente.
Identifique o prompt ou mídia a substituir.
Guarde o task ID, limite tentativas e não trate timeout como sucesso.
Biblioteca de prompts
Exemplos de prompts do GPT Image 2 para produto e marketing
Comece com um sujeito, ação, câmera ou composição concretos e uma configuração de saída explícita; mantenha os campos da solicitação separados do prompt.
Gerado com Image 2
Scene: A 16:9 sci-fi game loadout screen in a blue-purple spaceship hangar. A female armored character stands centered with a glowing violet rifle; armor slots line the left, three weapon slots line the right, and the lower slot is highlighted. Keep the HUD hierarchy, pose, lighting, and equipment geometry fixed for an exact equipment-switch frame. No readable words, logos, invented stats, or watermark. Composition and rendering: Deliver one 16:9 game-interface keyframe. Separate the character, held equipment and HUD into clear depth layers; keep slot spacing and icon scale consistent, with the selected item visually dominant. Preserve the scene-specific game art style. Let material edges catch the existing light without letting interface glow obscure the silhouette. Keep all equipment fully inside the frame. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Gerado com Image 2
Scene: A rainy night soccer broadcast frame under bright stadium floodlights. A black-uniform player strikes the ball while two white-uniform defenders close in, with the goal and a blurred crowd behind. Keep the broadcast camera angle, blank blue score bars, spray, and motion blur consistent; leave overlays abstract and text-free. No real teams, athletes, logos, or watermark. Composition and rendering: Deliver one 16:9 broadcast still at the decisive instant. Use coherent perspective for the playing surface and equipment, readable separation between competitors, and a softer crowd behind the action. Keep the principal subject sharp, with directional blur limited to fast extremities and spray. Preserve the specified broadcast angle and reserve overlay areas clear of the action. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Gerado com Image 2
Scene: A closed matte-black square smartwatch with a black strap rests on a wet glossy black tabletop. Water droplets catch a cool blue rim light and a soft reflection sits beneath the watch. Use a low three-quarter product camera and preserve the case, strap, highlights, and empty dark background. No readable branding, extra products, hands, or watermark. Composition and rendering: Deliver one 16:9 commercial product still. Keep the complete product silhouette inside generous crop-safe margins. Resolve edges, seams and material transitions precisely; shape highlights to describe volume without clipping bright surfaces or crushing dark details. Align reflections with the object and light source. Retain the stated camera angle and existing negative space; add no decorative props. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Gerado com Image 2
Scene: On a rain-soaked observatory roof at sunrise, a lone figure in a long dark coat walks toward a large telescope beside an open dome. Wet stone reflects the warm doorway light; cloud-covered mountains sit beyond. Hold the screen direction and end on the telescope and figure in the same wide composition. No readable text, logos, or watermark. Composition and rendering: Deliver one 16:9 storyboard keyframe, not a collage or an action sequence. Separate foreground, character plane and background through the existing light and atmospheric depth. Keep the stated pose and eyelines readable, and retain environmental landmarks for continuity. Match the described photographic, illustrated or puppet medium; give fabric and surfaces texture appropriate to that medium. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Gerado com Image 2
Scene: In a bright home kitchen, a surprised cook in a blue shirt and cream apron reaches toward pancakes, a frying pan, flour, bowl, and whisk suspended midair. Freeze the comic cause-and-effect moment with believable weight, warm daylight, and a clear path for each object. No injury, logos, readable words, or watermark. Composition and rendering: Deliver one 16:9 physical-comedy still at the described instant. Make the cause of the gag readable through body balance, grip and object spacing. Separate the reaction from the airborne props; retain believable gravity and contact points. Keep the face and important props sharp, with only restrained directional blur on fast edges. Preserve the stated lighting and room or stage layout. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Gerado com Image 2
Scene: A rainy period harbor platform shelters people in era-appropriate coats and umbrellas as a vintage green tram arrives on wet tracks. Wooden waterfront buildings and misty mountains sit behind the reflections. Preserve the historical clothing, tram shape, rain direction, and stable wide composition; no readable signage, logos, or watermark. Composition and restoration: Deliver one 16:9 archival-style still. Preserve the original framing, period-specific construction and restrained tonal range; retain monochrome, sepia or faded color as described. Recover modest detail in faces, clothing and architecture without synthetic sharpening. Keep fine grain and authentic aging consistent across the frame; repair damage without replacing historical features. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Por que a Flatkey
Um fluxo reproduzível do GPT Image 2 para imagens de produto
Mantenha visíveis os fatos, campos da solicitação e limites de preço do GPT Image 2 ao passar do teste para a produção.
Todos os controles documentados
Defina n, size, quality, output format, background e moderation antes de enviar a solicitação ao console.
Preço explícito
Veja cada dimensão documentada de entrada, saída e cache do GPT Image 2; um preço de destaque não é universal.
Pontos de partida por fluxo
Use exemplos de produto, anúncio e storyboard como prompts editáveis.
Sem promessas não verificadas
A página não promete saída transparente, geração gratuita ou preço fixo por imagem quando esses fatos não estão verificados.
API
Use a API do GPT Image 2 para gerar e editar imagens
Envie o ID do modelo com os controles de imagem compatíveis.
curl https://router.flatkey.ai/v1/images/generations \ -H "Content-Type: application/json" \ -H "Authorization: Bearer sk-***" \ -d '{"model":"gpt-image-2","prompt":"Creative direction: Create a vertical illustrated fantasy narrative poster with a large dark-haired profile silhouette on the left, a layered mountain temple city in the center, a red torii gate, drifting cherry blossoms, and small travelers along the bottom edge. Use a cream paper texture with ink-and-wash detail, a restrained charcoal, vermilion, and muted violet palette, and reserve a clean right margin for the title and credits.\n\nComposition, material and finish: Build the silhouette and temple city as nested visual layers, with the gate as the vermilion focal accent. Keep small travelers distinct from architectural detail. Use delicate ink edges and paper grain rather than photographic skin or glossy 3D surfaces. Leave the right title margin genuinely empty; avoid added lettering, duplicated gates, muddied silhouettes and watermarks.","n":1,"size":"1024x1024","quality":"high","output_format":"png","background":"opaque","moderation":"auto"}'
Perguntas frequentes
GPT Image 2
API: preços e perguntas sobre prompts
Preços, compatibilidade, limites e detalhes de solicitação específicos do modelo.





