grok-imagine-image AI 图像 API

grok-imagine-image 可通过 Flatkey 使用,适用于图像生成;Flatkey 通过 /v1/images/generations 路由此模型,供应商为 xAI。$0.016 per image。

文生图参考图引导
供应商
xAI
官方价格
$0.020 / 图像
Flatkey 价格
$0.016 / 图像-20%
生成器配置 · Playground(注册前可编辑)

输入

快捷 Prompt
参考图片
开始生成

输出

预览
图片预览

请求摘要

接口
POST /v1/images/generations
模型 ID
grok-imagine-image
分辨率
1k
质量
medium
画面比例
auto
输出格式
url
输出数量
1

仅供预览。登录后即可向 Flatkey 提交此请求。

请求预览
{
  "model": "grok-imagine-image",
  "prompt": "Creative direction: Create a 16:9 fictional open-world game livestream screenshot on a bright palm-lined coastal street. Show a pink sports car in the roadway, a third-person player avatar, a streamer facecam in the lower-left corner, a vertical chat column on the right, and a clear HUD with minimap and status panels. Use a lively neon-sunset palette, crisp game-rendered depth, and distinct overlay zones for the broadcast interface.\n\nComposition, material and finish: Keep the game world and broadcast overlays on separate visual layers. Preserve the street perspective and vehicle scale; reserve the lower-left facecam and right chat column without covering the player. Use coherent simplified HUD elements and fictional interface copy, with consistent padding and contrast. Avoid real platform logos, duplicated avatars, warped car wheels and watermarks.",
  "n": 1,
  "resolution": "1k",
  "quality": "medium",
  "aspect_ratio": "auto",
  "response_format": "url"
}

API 性能

grok-imagine-image API 可靠性与可用性

当生产流量达到足够规模后,这里会显示 grok-imagine-image 的 Flatkey 实时请求遥测数据。

供应商平均可用性
4.3%
最近 30 天
延迟
最近 30 天
请求量
208
30 天窗口
成功推理趋势

API 活动

grok-imagine-image API 使用量与请求活动

查看最近统计窗口内 grok-imagine-image 的请求量和成功推理活动。

请求量
208
30 天窗口
延迟
最近 30 天
可用性
4.3%
最近 30 天
成功推理趋势活动

grok-imagine-image 价格

grok-imagine-image API 价格与计费

下方价格来自当前模型的 Flatkey 定价数据,以及定价 API 返回的可见分组。

Flatkey 价格

价格 / 张图片

实时目录模型
$0.016 / 图像
价格 / 张图片
$0.016 / 图像$0.020 / 图像
先试 prompt
共用余额

充值余额

充值余额

图像、视频、音频和文本模型共用同一个 Flatkey 余额和 API Key。

模型目录
模型类型
文生图
API
/v1/images/generations
计费依据
图像
输出数量
1–10

grok-imagine-image AI 图像生成器与 API: 核心能力

xAI 的 grok-imagine-image 路由支持 文本 · 图像。目录分类:Marketing。

grok-imagine-image图像生成

通过 /v1/images/generations 将文字简报转换为图像生成结果。

grok-imagine-image参考工作流

当 grok-imagine-image 路由接受参考媒体或上下文时,可使用已有参考进行创作。

grok-imagine-image渠道适配版本

将一个创意调整为店铺、营销活动或内容流水线所需的不同版本。

grok-imagine-image生产交接

将确认后的结果从提示词探索带入可重复的 Flatkey 生成工作流。

对比

grok-imagine-image API 对比:能力与访问方式

在调整集成前,将 grok-imagine-image 的目录事实与上一代基线进行对比。

能力上一代grok-imagine-image
供应商此目录快照未验证xAI
模态此目录快照未验证文本 · 图像
上下文此目录快照未验证目录未列出上下文窗口
端点此目录快照未验证/v1/images/generations
计费此目录快照未验证$0.016 per image

提示词库

grok-imagine-image 图像生成提示词示例

先使用一个面向图像生成的 grok-imagine-image 提示词,再调整 Playground 中显示的请求字段。

游戏 UI 交互与装备动态切换
游戏 UI 交互与装备动态切换

A 16:9 cyberpunk game equipment screen shows a masked figure crouched on a neon rooftop above a dense night city. A black-and-lime HUD frames the character, with weapon and tool slots on the right including a glowing pickaxe. Hold the crouch, rooftop line, neon rim light, and selected slot fixed for an equipment-switch frame. No readable words, logos, invented stats, or watermark.

展开完整提示词收起提示词

Scene

A 16:9 cyberpunk game equipment screen shows a masked figure crouched on a neon rooftop above a dense night city. A black-and-lime HUD frames the character, with weapon and tool slots on the right including a glowing pickaxe. Hold the crouch, rooftop line, neon rim light, and selected slot fixed for an equipment-switch frame. No readable words, logos, invented stats, or watermark.

Composition and rendering

Deliver one 16:9 game-interface keyframe. Separate the character, held equipment and HUD into clear depth layers; keep slot spacing and icon scale consistent, with the selected item visually dominant. Preserve the scene-specific game art style. Let material edges catch the existing light without letting interface glow obscure the silhouette. Keep all equipment fully inside the frame.

Avoid

warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.

做一个类似的
高逼真电视/体育赛事直播模拟
高逼真电视/体育赛事直播模拟

A sunset beach-volleyball broadcast frame captures a player spiking over the net as sand sprays from the jump. Show the opposing court, hazy beach crowd, warm sky, and a simple teal-and-white lower-third made from abstract blocks. Keep the net, player position, ball path, and camera angle coherent; no real athletes, teams, logos, or watermark.

展开完整提示词收起提示词

Scene

A sunset beach-volleyball broadcast frame captures a player spiking over the net as sand sprays from the jump. Show the opposing court, hazy beach crowd, warm sky, and a simple teal-and-white lower-third made from abstract blocks. Keep the net, player position, ball path, and camera angle coherent; no real athletes, teams, logos, or watermark.

Composition and rendering

Deliver one 16:9 broadcast still at the decisive instant. Use coherent perspective for the playing surface and equipment, readable separation between competitors, and a softer crowd behind the action. Keep the principal subject sharp, with directional blur limited to fast extremities and spray. Preserve the specified broadcast angle and reserve overlay areas clear of the action.

Avoid

warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.

做一个类似的
品牌商业 TVC 与电商产品无缝展示
品牌商业 TVC 与电商产品无缝展示

A vintage black camera with a large textured lens rests on a black pedestal against a deep green-black background. Warm highlights describe the metal dials, leather grip, lens rings, and small empty space to the left. Preserve the camera body, lens perspective, pedestal edge, and controlled studio lighting. No readable branding, extra products, hands, or watermark.

展开完整提示词收起提示词

Scene

A vintage black camera with a large textured lens rests on a black pedestal against a deep green-black background. Warm highlights describe the metal dials, leather grip, lens rings, and small empty space to the left. Preserve the camera body, lens perspective, pedestal edge, and controlled studio lighting. No readable branding, extra products, hands, or watermark.

Composition and rendering

Deliver one 16:9 commercial product still. Keep the complete product silhouette inside generous crop-safe margins. Resolve edges, seams and material transitions precisely; shape highlights to describe volume without clipping bright surfaces or crushing dark details. Align reflections with the object and light source. Retain the stated camera angle and existing negative space; add no decorative props.

Avoid

warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.

做一个类似的
影视级剧本角色与多视角分镜演播
影视级剧本角色与多视角分镜演播

In a warmly lit theater dressing room, a curly-haired actor in a cream shirt and dark vest looks into a large aged mirror beside costumes and glowing round bulbs. The reflected face and three-quarter profile form a character continuity frame. Keep the mirror geometry, wardrobe, eye line, and amber lighting consistent. No readable signage, logos, or watermark.

展开完整提示词收起提示词

Scene

In a warmly lit theater dressing room, a curly-haired actor in a cream shirt and dark vest looks into a large aged mirror beside costumes and glowing round bulbs. The reflected face and three-quarter profile form a character continuity frame. Keep the mirror geometry, wardrobe, eye line, and amber lighting consistent. No readable signage, logos, or watermark.

Composition and rendering

Deliver one 16:9 storyboard keyframe, not a collage or an action sequence. Separate foreground, character plane and background through the existing light and atmospheric depth. Keep the stated pose and eyelines readable, and retain environmental landmarks for continuity. Match the described photographic, illustrated or puppet medium; give fabric and surfaces texture appropriate to that medium.

Avoid

warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.

做一个类似的
喜剧段子与物理剧情演播
喜剧段子与物理剧情演播

On a beach at sunset, a man in a light shirt pulls a yellow-striped picnic cloth as a sandwich and two white plates lift into the air above a wooden table. Freeze the harmless physical gag with visible cloth tension, believable object paths, warm sea light, and clear spacing. No injury, readable words, logos, or watermark.

展开完整提示词收起提示词

Scene

On a beach at sunset, a man in a light shirt pulls a yellow-striped picnic cloth as a sandwich and two white plates lift into the air above a wooden table. Freeze the harmless physical gag with visible cloth tension, believable object paths, warm sea light, and clear spacing. No injury, readable words, logos, or watermark.

Composition and rendering

Deliver one 16:9 physical-comedy still at the described instant. Make the cause of the gag readable through body balance, grip and object spacing. Separate the reaction from the airborne props; retain believable gravity and contact points. Keep the face and important props sharp, with only restrained directional blur on fast edges. Preserve the stated lighting and room or stage layout.

Avoid

warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.

做一个类似的
老旧照片修复与历史/人文动态复活
老旧照片修复与历史/人文动态复活

A warm faded archival photograph shows a family having a picnic in a grassy field: one adult pours from a thermos, another spreads food, and a child sits beside a bicycle. Preserve the period clothing, enamel cups, bicycle, hills, film grain, and soft light-leak edge. No modern branding, readable text, logos, or watermark.

展开完整提示词收起提示词

Scene

A warm faded archival photograph shows a family having a picnic in a grassy field: one adult pours from a thermos, another spreads food, and a child sits beside a bicycle. Preserve the period clothing, enamel cups, bicycle, hills, film grain, and soft light-leak edge. No modern branding, readable text, logos, or watermark.

Composition and restoration

Deliver one 16:9 archival-style still. Preserve the original framing, period-specific construction and restrained tonal range; retain monochrome, sepia or faded color as described. Recover modest detail in faces, clothing and architecture without synthetic sharpening. Keep fine grain and authentic aging consistent across the frame; repair damage without replacing historical features.

Avoid

warped perspective, malformed anatomy, merged objects, inconsistent shadows, excessive sharpening and watermarks. Follow the scene-specific text and branding restrictions above.

做一个类似的

为什么使用 grok-imagine-image API

为什么通过 Flatkey 使用 grok-imagine-image?

通过一个网关管理 grok-imagine-image、账户控制以及模型目录中的其他模型。

一个 API 使用 grok-imagine-image

明确保留 grok-imagine-image 模型 ID 和端点,同时与其他工作负载共用 Flatkey 密钥。

实时目录价格

发送请求前查看当前 grok-imagine-image 计费维度,并在账户中确认最终预估。

实用的模型交接

在公开 Playground 中测试 grok-imagine-image 提示词,再将相同设置带入经过身份验证的集成。

用量与路由控制

集中管理 grok-imagine-image 的密钥、配额和路由,无需改变应用面向供应商的工作流。

grok-imagine-image API

grok-imagine-image API 集成指南

使用上方模型 ID 调用 /v1/images/generations;为当前路由保留已记录的请求字段和计费单位。

向 /v1/images/generations 发送经过身份验证的请求,并将 model 设置为 grok-imagine-image。
在 SDK 配置中使用 grok-imagine-image,让路由和用量报告对应到预期的目录条目。
扩展 grok-imagine-image 请求前,查看 $0.016 per image 和账户限制。
先在 Playground 中开始,再在服务器或 agent 中使用 Flatkey API 密钥复用相同请求结构。
curl https://router.flatkey.ai/v1/images/generations \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer sk-..." \
  -d '{
  "model": "grok-imagine-image",
  "prompt": "Creative direction: Create a 16:9 fictional open-world game livestream screenshot on a bright palm-lined coastal street. Show a pink sports car in the roadway, a third-person player avatar, a streamer facecam in the lower-left corner, a vertical chat column on the right, and a clear HUD with minimap and status panels. Use a lively neon-sunset palette, crisp game-rendered depth, and distinct overlay zones for the broadcast interface.\n\nComposition, material and finish: Keep the game world and broadcast overlays on separate visual layers. Preserve the street perspective and vehicle scale; reserve the lower-left facecam and right chat column without covering the player. Use coherent simplified HUD elements and fictional interface copy, with consistent padding and contrast. Avoid real platform logos, duplicated avatars, warped car wheels and watermarks.",
  "n": 1,
  "resolution": "1k",
  "quality": "medium",
  "aspect_ratio": "auto",
  "response_format": "url"
}'
文档

常见问题

grok-imagine-image API
价格、功能与使用常见问题

关于 grok-imagine-image 价格、能力、端点访问和目录限制的答案。

grok-imagine-image 用于什么任务?
grok-imagine-image 由 xAI 列入图像生成目录;目录列出的模态为:文本 · 图像。
grok-imagine-image 如何计费?
grok-imagine-image 当前显示 $0.016 per image。适用费率可能因路由、账户组和请求设置而变化。
调用 grok-imagine-image 使用哪个 API 端点?
Flatkey 通过 /v1/images/generations 路由此模型。将模型字段设为 grok-imagine-image,并遵循该端点支持的字段。
grok-imagine-image 有哪些上下文或输入限制?
grok-imagine-image 标注为 目录未列出上下文窗口。其他限制取决于路由和当前账户可用性,生产使用前请确认。