GPT-image-2 AI Image API
GPT Image 2 is OpenAI's image-generation model. OpenAI documents image generation and editing with flexible sizes; Flatkey exposes the image routes and request controls shown below.
Output
Preview
Request summary
- Endpoint
- POST /v1/images/generations
- Model ID
- GPT-image-2
- Aspect ratio
- 4:5
- Resolution
- 1024x1024
- Quality
- high
- Outputs
- 1
Preview only. Sign in to submit this request to Flatkey.
{
"model": "gpt-image-2",
"prompt": "Goal: Create a luxury fantasy tea advertisement poster for {argument name=\"product name\" default=\"MARINE LUCENT\"}, themed as an oceanic, character-inspired magical tea extracted from an ethereal mermaid-like heroine. The mood is translucent, premium, dreamy, and not a normal cafe menu.\n\nCanvas: Vertical 4:5 high-resolution poster, soft pastel blue, lavender, pearl white, and iridescent aqua palette. Use glossy commercial advertising composition with delicate typography and a high-fashion fantasy tea brand atmosphere.\n\nLayout: Place the main product scene on the right and lower center: a clear glass teapot in the upper-right pours glowing blue-violet tea in a sparkling stream into a transparent glass teacup on a saucer at the lower center. On the left side, create a refined editorial text column with the brand logo at top, product title in large serif letters, subtitle, flavor copy, and ingredient list. Add a pale lavender rectangular placeholder block near the upper-left/center, partially covering the background character art. Keep the whole design airy, luminous, and layered with bokeh, water reflections, and glass highlights.\n\nBackground and subject details: Behind the tea, show a barely visible ethereal anime-style marine princess figure as a soft-focus illustration: long flowing aqua-blue hair, shell and pearl accessories, pale skin, delicate floral/starfish ornaments, translucent fabric, and an underwater fantasy aura. She should feel like the source of the tea’s identity and world, but remain mostly hidden behind mist, glass, and typography.\n\nMain visual: The teapot is round, transparent, crystal-like, and filled with vivid magical tea in gradients of electric blue, cobalt, violet, and cyan. Inside the tea, show exactly 8 visible floating decorative elements: 3 lavender star-shaped flowers, 2 pale pearl-shell spheres, 1 small golden starfish, 1 purple blossom cluster, and 1 cloud of glittering bubbles. The pouring stream should look like liquid starlight, with tiny bubbles and radiant sparkles. The teacup contains the same glowing ocean tea, with suspended flowers, bubbles, shell-like pearls, and refracted light. Add a glass saucer, a small starfish charm, pearls, seashells, crystalline ice-like rocks, and a tiny ornate perfume-bottle-like vessel in the background.\n\nText content: At the top-left, create a diamond monogram logo with the letter “M”, then the brand name “MAISON DE LUMIÈRE” and small line “FINE FANTASY TEA”. Add the tagline in Japanese “彼女の世界を、あなたの一杯に。” and below it “Her world, in your teacup.” Set the product title as large elegant serif text: “{argument name=\"headline text\" default=\"MARINE LUCENT\"}”. Below it add small Japanese reading “マリン・ルーセント”. Add the flavor name “{argument name=\"flavor name\" default=\"AURORA BLUE BLOOM\"}” and Japanese reading “オーロラブルー・ブルーム”. Add a short Japanese body copy block, then English copy: “The sparkle of waves, the dance of bubbles, the gentleness of a clear blue breeze—blended into a tea that is pure, radiant, and enchantingly alive.” Add an ingredient list with exactly 6 items and small matching icons: “BLUE PEA”, “WHITE TEA”, “CORNFLOWER”, “PEARL SHELL”, “CITRUS & PEACH”, and “SEA SALT DROP”, each with smaller Japanese text underneath. At bottom-right, add handwritten script: “{argument name=\"signature phrase\" default=\"Sip the Ocean’s Whisper.\"}” with Japanese “海のささやきを、ひとくちに。” and the brand name “MAISON DE LUMIÈRE”.\n\nVisual style: Ultra-detailed glossy fantasy advertising, premium cosmetics-ad lighting, crystalline glass refractions, shimmering caustics, pearlescent highlights, soft bloom, layered transparency, pastel marine glow, elegant serif typography mixed with delicate script. Make the typography readable and balanced, with the title dominating the left lower half.\n\nConstraints: Use exactly one teapot and one teacup. Use exactly 6 listed ingredient rows. Preserve the luxury tea poster format, do not make it a cafe menu, product package, or character sheet. Avoid extra brand logos, watermarks, QR codes, or unrelated objects.",
"n": 1,
"size": "1024x1024",
"quality": "high",
"output_format": "png",
"background": "opaque",
"moderation": "auto"
}Performance
GPT Image 2 API performance and availability
Live Flatkey request telemetry appears here when enough image-generation traffic is available.
Activity
GPT Image 2 API usage and generation activity
This chart reflects live Flatkey generation requests and remains unreported until enough traffic is collected.
Pricing
GPT Image 2 API pricing by token and image input
Prices below are calculated from Flatkey pricing data for this model and the visible groups currently returned by our pricing API.
Add credits
Use the same Flatkey balance and API key across image, video, audio, and text models.
- Model Type
- Text to Image
- API
- /v1/images/generations
- Billing basis
- image
- Outputs
- 1–10
GPT Image 2 image generation and editing capabilities
Use the documented image workflows for new visuals, reference-led edits, and production-ready creative variants.
Text-to-image creation
Generate new product, editorial, and campaign visuals from a structured natural-language prompt.
Reference-based editing
Supply an image input to revise an existing visual while preserving important subject and composition details.
Flexible image composition
Create square, portrait, or landscape assets with flexible image sizes for different placements and channels.
High-fidelity visual input
Combine text and image context in one workflow for richer visual direction; audio and video inputs are not supported by this model.
Compare
GPT Image 2 and GPT Image 1: image API controls
GPT Image 1 values are not verified in this fact set, so only documented GPT Image 2 fields are shown.
| Capability | GPT Image 1 | GPT Image 2 |
|---|---|---|
| Endpoints | Unknown here | /v1/images/generations; /v1/images/edits |
| Sizes | Unknown here | 1024x1024, 1536x1024, 1024x1536, auto |
| Formats | Unknown here | PNG, JPEG, WebP |
| Quality/background/moderation | Unknown here | Documented above |
| Image quality ranking | Not asserted | Not asserted |
Choose the right request before you integrate
Review the supported workflows, live cost for the current setup, and the exact parameter contract used by Flatkey.
Calculated from the current Playground settings and the live catalog rate when a fixed unit price is available.
- Current setup
- 1024x1024 · high · 1×
- Live rate
- $0.0088 / image
- Estimated request
- $0.0088
The final charge follows the submitted request and your account group. Token-priced or tiered requests cannot be reduced to one fixed estimate.
Workflows and Model ID
Use only the workflows marked available. Flatkey keeps one public model ID while the request body selects the input workflow.
Text to Image
Generate a new image from a text prompt and output controls.
- Model ID
gpt-image-2- Endpoint
/v1/images/generations
Reference-guided Image
Use reference images for edits, variations, or identity and composition guidance.
- Model ID
gpt-image-2- Endpoint
/v1/images/generations
Parameter compatibility
These values come from the same model configuration used by the Playground and request preview.
nImages1–10
Default: 1sizeSize1024x1024 · 1536x1024 · 1024x1536 · auto
Default: 1024x1024qualityQualityauto · high · medium · low
Default: highoutput_formatOutput formatpng · jpeg · webp
Default: pngbackgroundBackgroundopaque · auto
Default: opaquemoderationModerationauto · low
Default: autoimageReference limit4 files
Endpoint: /v1/images/generationsFrom first test to production
Use the same request contract in the Playground and API, then handle validation, balance, task state, and output retrieval explicitly.
Request lifecycle
- 1
Validate the model, prompt, references, and output settings.
- 2
POST the image request with the documented fields.
- 3
Read the generated output or asynchronous task result returned by the route.
- 4
Store the generated file and request metadata in your own workflow.
Production checks
Reject unsupported sizes, durations, ratios, formats, or reference counts before submitting.
Handle invalid keys, insufficient balance, rate limits, and account-group pricing separately.
Surface the rejected prompt or asset clearly so the user can replace only the failing input.
Keep the task ID, use bounded retries, and never treat a timeout as a successful generation.
Prompt library
GPT Image 2 prompt examples for product and marketing visuals
Start with a concrete subject, composition, output size, and brand constraint; adjust the request fields separately.
Generated with Image 2
Scene: A 16:9 sci-fi game loadout screen in a blue-purple spaceship hangar. A female armored character stands centered with a glowing violet rifle; armor slots line the left, three weapon slots line the right, and the lower slot is highlighted. Keep the HUD hierarchy, pose, lighting, and equipment geometry fixed for an exact equipment-switch frame. No readable words, logos, invented stats, or watermark. Composition and rendering: Deliver one 16:9 game-interface keyframe. Separate the character, held equipment and HUD into clear depth layers; keep slot spacing and icon scale consistent, with the selected item visually dominant. Preserve the scene-specific game art style. Let material edges catch the existing light without letting interface glow obscure the silhouette. Keep all equipment fully inside the frame. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Generated with Image 2
Scene: A rainy night soccer broadcast frame under bright stadium floodlights. A black-uniform player strikes the ball while two white-uniform defenders close in, with the goal and a blurred crowd behind. Keep the broadcast camera angle, blank blue score bars, spray, and motion blur consistent; leave overlays abstract and text-free. No real teams, athletes, logos, or watermark. Composition and rendering: Deliver one 16:9 broadcast still at the decisive instant. Use coherent perspective for the playing surface and equipment, readable separation between competitors, and a softer crowd behind the action. Keep the principal subject sharp, with directional blur limited to fast extremities and spray. Preserve the specified broadcast angle and reserve overlay areas clear of the action. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Generated with Image 2
Scene: A closed matte-black square smartwatch with a black strap rests on a wet glossy black tabletop. Water droplets catch a cool blue rim light and a soft reflection sits beneath the watch. Use a low three-quarter product camera and preserve the case, strap, highlights, and empty dark background. No readable branding, extra products, hands, or watermark. Composition and rendering: Deliver one 16:9 commercial product still. Keep the complete product silhouette inside generous crop-safe margins. Resolve edges, seams and material transitions precisely; shape highlights to describe volume without clipping bright surfaces or crushing dark details. Align reflections with the object and light source. Retain the stated camera angle and existing negative space; add no decorative props. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Generated with Image 2
Scene: On a rain-soaked observatory roof at sunrise, a lone figure in a long dark coat walks toward a large telescope beside an open dome. Wet stone reflects the warm doorway light; cloud-covered mountains sit beyond. Hold the screen direction and end on the telescope and figure in the same wide composition. No readable text, logos, or watermark. Composition and rendering: Deliver one 16:9 storyboard keyframe, not a collage or an action sequence. Separate foreground, character plane and background through the existing light and atmospheric depth. Keep the stated pose and eyelines readable, and retain environmental landmarks for continuity. Match the described photographic, illustrated or puppet medium; give fabric and surfaces texture appropriate to that medium. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Generated with Image 2
Scene: In a bright home kitchen, a surprised cook in a blue shirt and cream apron reaches toward pancakes, a frying pan, flour, bowl, and whisk suspended midair. Freeze the comic cause-and-effect moment with believable weight, warm daylight, and a clear path for each object. No injury, logos, readable words, or watermark. Composition and rendering: Deliver one 16:9 physical-comedy still at the described instant. Make the cause of the gag readable through body balance, grip and object spacing. Separate the reaction from the airborne props; retain believable gravity and contact points. Keep the face and important props sharp, with only restrained directional blur on fast edges. Preserve the stated lighting and room or stage layout. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Generated with Image 2
Scene: A rainy period harbor platform shelters people in era-appropriate coats and umbrellas as a vintage green tram arrives on wet tracks. Wooden waterfront buildings and misty mountains sit behind the reflections. Preserve the historical clothing, tram shape, rain direction, and stable wide composition; no readable signage, logos, or watermark. Composition and restoration: Deliver one 16:9 archival-style still. Preserve the original framing, period-specific construction and restrained tonal range; retain monochrome, sepia or faded color as described. Recover modest detail in faces, clothing and architecture without synthetic sharpening. Keep fine grain and authentic aging consistent across the frame; repair damage without replacing historical features. Avoid: warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.
Why use Flatkey for GPT Image 2?
A repeatable GPT Image 2 workflow for product images
Keep image controls and image-input token dimensions together so a prompt test can become a reproducible API request.
All documented controls in one place
Set n, size, quality, output format, background, and moderation before you hand the request to the console.
Image input stays explicit
The catalog exposes token and image-input dimensions; this is not a separate fee for output pixel size, and the final amount depends on request and usage fields.
Marketing-ready starting points
Use product, ad, and storyboard examples as editable starting prompts rather than generic image filler.
No unsupported promises
The page does not promise transparent output, free generation, or a fixed per-image price when those facts are not verified.
API
Use the GPT Image 2 API for generation and edits
Send the model ID with the supported image controls through the documented generation or edits route.
curl https://router.flatkey.ai/v1/images/generations \ -H "Content-Type: application/json" \ -H "Authorization: Bearer sk-***" \ -d '{"model":"gpt-image-2","prompt":"A complex AI image generation mood wall photographed as one premium studio scene: a large violet-and-gold abstract floral artwork surrounded by pinned botanical sketches, macro insect study, translucent vellum flower sheets, black-and-white portrait, crystal minerals, perfume product render, architectural arch and staircase studies, film strips, fabric swatches, tape, brass pins, graphite notes, and warm spotlights on a charcoal wall.","n":1,"size":"1024x1024","quality":"high","output_format":"png","background":"opaque","moderation":"auto"}'
FAQ
GPT Image 2 API
pricing and prompt questions
Answers about GPT Image 2 API access, image-generation fields, token and image-input pricing, formats, and background behavior.





