Find AI models that fit your budget

Explore discounted AI models and compare their API prices against the stated reference rates. Review input and output costs, model capabilities, and applicable conditions before choosing.

Featured models include deepseek-v4-flash, gpt-5.4-mini, and gpt-5.6-terra, ranked by live weekly usage when available.

Discounted AI Models

65,770,114,004,200 weekly usage

Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding. deepseek-v4-flash by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: from $0.176 / 1M tokens
Configured referencefrom $0.22 / 1M tokens
Output
Current public price: from $0.528 / 1M tokens
Configured referencefrom $0.66 / 1M tokens
Cache read
Current public price: from $0.0056 / 1M tokens
Configured referencefrom $0.007 / 1M tokens
1.0M context
View model details
2.
3,036,362,789,400 weekly usage

Strong small GPT for coding subagents, quick tool use, and high-volume work. gpt-5.4-mini by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.6 / 1M tokens
Configured reference$0.75 / 1M tokens
Output
Current public price: $3.6 / 1M tokens
Configured reference$4.5 / 1M tokens
Cache read
Current public price: $0.06 / 1M tokens
Configured reference$0.075 / 1M tokens
400K context
View model details
1,488,445,691,500 weekly usage

Balanced GPT-5.6 model for capable, cost-efficient everyday work. gpt-5.6-terra by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $9.6 / 1M tokens
Configured reference$12 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
Cache write
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
1.0M context
View model details
4.
1,343,187,127,900 weekly usage

A frontier reasoning model with a 1M-token window, built for multi-step analysis, repository-scale code review and research workflows — at a fraction of the cost of comparable frontier models. DeepSeek V4 Pro by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: from $0.528 / 1M tokens
Configured referencefrom $0.66 / 1M tokens
Output
Current public price: from $1.584 / 1M tokens
Configured referencefrom $1.98 / 1M tokens
Cache read
Current public price: from $0.0176 / 1M tokens
Configured referencefrom $0.022 / 1M tokens
1.0M context
View model details
5.
1,224,744,420,700 weekly usage

Cost-efficient GPT-5.6 model for fast, high-volume workloads. gpt-5.6-luna by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
Output
Current public price: $0.96 / 1M tokens
Configured reference$1.2 / 1M tokens
Cache read
Current public price: $0.016 / 1M tokens
Configured reference$0.02 / 1M tokens
Cache write
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
1.0M context
View model details
6.
1,060,552,118,600 weekly usage

The dependable workhorse of the GPT line — strong general reasoning, reliable structured output and first-class tool calling, priced so you can put it on the hot path of a production app. GPT-5.6 Sol by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $3.2 / 1M tokens
Configured reference$4 / 1M tokens
Output
Current public price: $19.2 / 1M tokens
Configured reference$24 / 1M tokens
Cache read
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
Cache write
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
1.0M context
View model details
7.
claude-opus-5

Anthropic

967,117,925,000 weekly usage

Strongest Claude Opus model for coding, agents, and professional work. claude-opus-5 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $20 / 1M tokens
Configured reference$25 / 1M tokens
Cache read
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Cache write
Current public price: $5 / 1M tokens
Configured reference$6.25 / 1M tokens
1.0M context
View model details
8.
gpt-5.5

OpenAI

684,912,587,600 weekly usage

Default frontier GPT for coding, computer use, research, and knowledge work. gpt-5.5 by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $24 / 1M tokens
Configured reference$30 / 1M tokens
Cache read
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
1.0M context
View model details
572,420,181,600 weekly usage

Deepseek 4.1 Flash is built for harder reasoning, richer multimodal understanding, and long-context collaboration. It brings code, images, text, and task flow into one orbit — for work that needs sharper thinking, steadier output, and deeper memory. deepseek-v4-1-flash by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: from $0.12 / 1M tokens
Configured referencefrom $0.15 / 1M tokens
Output
Current public price: from $0.48 / 1M tokens
Configured referencefrom $0.6 / 1M tokens
Cache read
Current public price: from $0.0024 / 1M tokens
Configured referencefrom $0.003 / 1M tokens
10.
343,468,212,700 weekly usage

Claude model for creative writing, analysis, and controlled agent workflows. claude-fable-5 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $9 / 1M tokens
Configured reference$10 / 1M tokens
Output
Current public price: $45 / 1M tokens
Configured reference$50 / 1M tokens
Cache read
Current public price: $0.9 / 1M tokens
Configured reference$1 / 1M tokens
Cache write
Current public price: $11.25 / 1M tokens
Configured reference$12.5 / 1M tokens
1.0M context
View model details
11.
335,043,223,300 weekly usage

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents. claude-opus-4-8 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $20 / 1M tokens
Configured reference$25 / 1M tokens
Cache read
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Cache write
Current public price: $5 / 1M tokens
Configured reference$6.25 / 1M tokens
1.0M context
View model details
12.
308,805,481,000 weekly usage

Everyday Claude agent model for coding, planning, browsing, and general work. claude-sonnet-5 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $8 / 1M tokens
Configured reference$10 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
Cache write
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
1.0M context
View model details
13.
240,832,691,600 weekly usage

High-end Claude for difficult coding, planning, and slower expert reasoning. claude-opus-4-6 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $20 / 1M tokens
Configured reference$25 / 1M tokens
Cache read
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Cache write
Current public price: $5 / 1M tokens
Configured reference$6.25 / 1M tokens
1.0M context
View model details
14.A

apodex-1-0-deep-discover by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
15.A

apodex-1-0-deep-research by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
16.A
apodex-1-0-deep-solve

Flatkey catalog

apodex-1-0-deep-solve by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
17.A

apodex-1-1-deep-discover by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
18.A

apodex-1-1-deep-research by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
19.A
apodex-1-1-deep-solve

Flatkey catalog

apodex-1-1-deep-solve by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
20.A
apodex-1.1

Flatkey catalog

apodex-1.1 by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
21.A
apodex-1.1-mini

Flatkey catalog

apodex-1.1-mini by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
Output
Current public price: $60 / 1M tokens
Configured reference$75 / 1M tokens
22.

Stronger Opus tier for advanced software work and high-stakes reasoning. claude-opus-4-7 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $20 / 1M tokens
Configured reference$25 / 1M tokens
Cache read
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Cache write
Current public price: $5 / 1M tokens
Configured reference$6.25 / 1M tokens
1.0M context
View model details
23.

Claude workhorse for coding agents, careful analysis, and production cost control. claude-sonnet-4-6 by Anthropic is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $2.4 / 1M tokens
Configured reference$3 / 1M tokens
Output
Current public price: $12 / 1M tokens
Configured reference$15 / 1M tokens
Cache read
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Cache write
Current public price: $3 / 1M tokens
Configured reference$3.75 / 1M tokens
1.0M context
View model details
24.
deepseek-v3

DeepSeek

deepseek-v3 by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
Output
Current public price: $1.04 / 1M tokens
Configured reference$1.3 / 1M tokens
Cache read
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
128K context
View model details
25.

deepseek-v3.1 by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.168 / 1M tokens
Configured reference$0.21 / 1M tokens
Output
Current public price: $0.632 / 1M tokens
Configured reference$0.79 / 1M tokens
Cache read
Current public price: $0.104 / 1M tokens
Configured reference$0.13 / 1M tokens
128K context
View model details
26.

deepseek-v3.2 by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.2128 / 1M tokens
Configured reference$0.266 / 1M tokens
Output
Current public price: $0.3552 / 1M tokens
Configured reference$0.444 / 1M tokens
Cache read
Current public price: $0.02128 / 1M tokens
Configured reference$0.0266 / 1M tokens
128K context
View model details

deepseek-v3.2-thinking by DeepSeek is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.232 / 1M tokens
Configured reference$0.29 / 1M tokens
Output
Current public price: $0.336 / 1M tokens
Configured reference$0.42 / 1M tokens
Cache read
Current public price: $0.232 / 1M tokens
Configured reference$0.29 / 1M tokens
128K context
View model details

Fast Gemini workhorse for multimodal apps where latency and price matter. gemini-2.5-flash by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.072 / 1M tokens
Configured reference$0.09 / 1M tokens
Output
Current public price: $0.568 / 1M tokens
Configured reference$0.71 / 1M tokens
Cache read
Current public price: $0.00216 / 1M tokens
Configured reference$0.0027 / 1M tokens
1.0M context
View model details

Nano Banana image model for fast generation, edits, and character-consistent assets. gemini-2.5-flash-image by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
Cache read
Current public price: $0.06 / 1M tokens
Configured reference$0.075 / 1M tokens
Image
Current public price: $0.024 / image
Configured reference$0.03 / image

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents. gemini-2.5-flash-lite by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.072 / 1M tokens
Configured reference$0.09 / 1M tokens
Output
Current public price: $0.288 / 1M tokens
Configured reference$0.36 / 1M tokens
Cache read
Current public price: $0.018 / 1M tokens
Configured reference$0.0225 / 1M tokens
1.0M context
View model details

gemini-2.5-flash-tts by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Output
Current public price: $8 / 1M tokens
Configured reference$10 / 1M tokens

Google's proven reasoning model for coding, math, and multimodal analysis. gemini-2.5-pro by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.9 / 1M tokens
Configured reference$1.125 / 1M tokens
Output
Current public price: $7.2 / 1M tokens
Configured reference$9 / 1M tokens
Cache read
Current public price: $0.09 / 1M tokens
Configured reference$0.1125 / 1M tokens
1.0M context
View model details

gemini-2.5-pro-tts by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.8 / 1M tokens
Configured reference$1 / 1M tokens
Output
Current public price: $16 / 1M tokens
Configured reference$20 / 1M tokens

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs. gemini-3-flash-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.056 / 1M tokens
Configured reference$0.07 / 1M tokens
Output
Current public price: $0.344 / 1M tokens
Configured reference$0.43 / 1M tokens
Cache read
Current public price: $0.0056 / 1M tokens
Configured reference$0.007 / 1M tokens
1.0M context
View model details

Nano Banana Pro for higher-fidelity image generation and design-heavy edits. gemini-3-pro-image by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $9.6 / 1M tokens
Configured reference$12 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
Image
Current public price: $0.096 / image
Configured reference$0.12 / image

Image model for prompt-driven generation, editing, and visual design workflows. gemini-3.1-flash-image by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Output
Current public price: $2.4 / 1M tokens
Configured reference$3 / 1M tokens
Image
Current public price: $0.048 / image
Configured reference$0.06 / image

Low-latency Gemini model for high-volume multimodal and agent workloads. gemini-3.1-flash-lite by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
Output
Current public price: $1.2 / 1M tokens
Configured reference$1.5 / 1M tokens
Cache read
Current public price: $0.02 / 1M tokens
Configured reference$0.025 / 1M tokens
1.0M context
View model details

Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing. gemini-3.1-flash-lite-image by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
Output
Current public price: $1.2 / 1M tokens
Configured reference$1.5 / 1M tokens
Image
Current public price: $0.024 / image
Configured reference$0.03 / image

Legacy model retained for compatibility with older integrations. gemini-3.1-flash-lite-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
Output
Current public price: $1.2 / 1M tokens
Configured reference$1.5 / 1M tokens
Cache read
Current public price: $0.02 / 1M tokens
Configured reference$0.025 / 1M tokens

Low-latency speech generation with steerable prompts and expressive audio tags. gemini-3.1-flash-tts-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.8 / 1M tokens
Configured reference$1 / 1M tokens
Output
Current public price: $16 / 1M tokens
Configured reference$20 / 1M tokens

Reasoning-first Gemini preview for agentic coding and complex problem solving. gemini-3.1-pro-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $9.6 / 1M tokens
Configured reference$12 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
1.0M context
View model details

Advanced Gemini model for complex reasoning, coding, and multimodal analysis. gemini-3.1-pro-preview-customtools by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $9.6 / 1M tokens
Configured reference$12 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
1.0M context
View model details

Fast Gemini model balancing multimodal reasoning, tool use, and cost. gemini-3.5-flash by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.2 / 1M tokens
Configured reference$1.5 / 1M tokens
Output
Current public price: $7.2 / 1M tokens
Configured reference$9 / 1M tokens
Cache read
Current public price: $0.12 / 1M tokens
Configured reference$0.15 / 1M tokens
1.0M context
View model details

Fast Gemini model balancing multimodal reasoning, tool use, and cost. gemini-3.5-flash-lite by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
Cache read
Current public price: $0.024 / 1M tokens
Configured reference$0.03 / 1M tokens
1.0M context
View model details

Fast Gemini model balancing multimodal reasoning, tool use, and cost. gemini-3.6-flash by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.2 / 1M tokens
Configured reference$1.5 / 1M tokens
Output
Current public price: $6 / 1M tokens
Configured reference$7.5 / 1M tokens
Cache read
Current public price: $0.12 / 1M tokens
Configured reference$0.15 / 1M tokens
1.0M context
View model details

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning. gemini-3.7-flash by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.2 / 1M tokens
Configured reference$1.5 / 1M tokens
Output
Current public price: $6 / 1M tokens
Configured reference$7.5 / 1M tokens
Cache read
Current public price: $0.12 / 1M tokens
Configured reference$0.15 / 1M tokens
Cache write
Current public price: $0.066664 / 1M tokens
Configured reference$0.08333 / 1M tokens

gemini-3.8-flash by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $6.4 / 1M tokens
Configured reference$8 / 1M tokens

Embedding model for semantic search, retrieval, clustering, and ranking pipelines. gemini-embedding-001 by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.12 / 1M tokens
Configured reference$0.15 / 1M tokens

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning. gemini-flash-latest by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
Cache read
Current public price: $0.06 / 1M tokens
Configured reference$0.075 / 1M tokens

Fast Gemini model balancing multimodal reasoning, tool use, and cost. gemini-flash-lite-latest by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.08 / 1M tokens
Configured reference$0.1 / 1M tokens
Output
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
Cache read
Current public price: $0.02 / 1M tokens
Configured reference$0.025 / 1M tokens

gemini-pro-latest by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $9.6 / 1M tokens
Configured reference$12 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens

gemini-robotics-er-1.6-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $0.96 / 1M tokens
Configured reference$1.2 / 1M tokens
Cache read
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
128K context
View model details

Open Gemma instruction model for efficient chat and self-hosted deployments. gemma-4-26b-a4b-it by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
Output
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning. gemma-4-31b-it by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.144 / 1M tokens
Configured reference$0.18 / 1M tokens
Output
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
55.

Mature GLM model for dependable coding, reasoning, and structured agent tasks. glm-4.7 by Z.ai is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.48 / 1M tokens
Configured reference$0.6 / 1M tokens
Output
Current public price: $1.76 / 1M tokens
Configured reference$2.2 / 1M tokens
Cache read
Current public price: $0.088 / 1M tokens
Configured reference$0.11 / 1M tokens
205K context
View model details
56.
glm-5

Z.ai

General GLM flagship for coding, analysis, and tool-heavy engineering workflows. glm-5 by Z.ai is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.8 / 1M tokens
Configured reference$1 / 1M tokens
Output
Current public price: $2.56 / 1M tokens
Configured reference$3.2 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
205K context
View model details

Faster GLM-5 lane for coding agents that need lower latency. glm-5-turbo by Z.ai is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.96 / 1M tokens
Configured reference$1.2 / 1M tokens
Output
Current public price: $3.2 / 1M tokens
Configured reference$4 / 1M tokens
Cache read
Current public price: $0.192 / 1M tokens
Configured reference$0.24 / 1M tokens
205K context
View model details
58.

Strong GLM coding model for agentic engineering, terminals, and repository generation. glm-5.1 by Z.ai is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.12 / 1M tokens
Configured reference$1.4 / 1M tokens
Output
Current public price: $3.52 / 1M tokens
Configured reference$4.4 / 1M tokens
Cache read
Current public price: $0.208 / 1M tokens
Configured reference$0.26 / 1M tokens
205K context
View model details
59.

Open flagship GLM for long-horizon coding agents and million-token context work. glm-5.2 by Z.ai is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.12 / 1M tokens
Configured reference$1.4 / 1M tokens
Output
Current public price: $3.52 / 1M tokens
Configured reference$4.4 / 1M tokens
Cache read
Current public price: $0.208 / 1M tokens
Configured reference$0.26 / 1M tokens
1.0M context
View model details
60.

Flagship GLM model for long-horizon coding, agents, and complex project delivery. glm-5.3 by Z.ai is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.12 / 1M tokens
Configured reference$1.4 / 1M tokens
Output
Current public price: $3.52 / 1M tokens
Configured reference$4.4 / 1M tokens
Cache read
Current public price: $0.208 / 1M tokens
Configured reference$0.26 / 1M tokens
1.0M context
View model details
61.

Affordable GPT-4.1 lane for fast coding help and structured extraction. gpt-4.1-mini by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
Output
Current public price: $1.28 / 1M tokens
Configured reference$1.6 / 1M tokens
Cache read
Current public price: $0.08 / 1M tokens
Configured reference$0.1 / 1M tokens
1.0M context
View model details
62.
gpt-4o

OpenAI

Omni-era GPT for multimodal chat, practical coding, and general assistants. gpt-4o by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
Output
Current public price: $8 / 1M tokens
Configured reference$10 / 1M tokens
Cache read
Current public price: $1 / 1M tokens
Configured reference$1.25 / 1M tokens
128K context
View model details
63.

Small omni GPT for cheap multimodal assistance and production-scale traffic. gpt-4o-mini by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.12 / 1M tokens
Configured reference$0.15 / 1M tokens
Output
Current public price: $0.48 / 1M tokens
Configured reference$0.6 / 1M tokens
Cache read
Current public price: $0.064 / 1M tokens
Configured reference$0.08 / 1M tokens
128K context
View model details
64.

Small GPT-5 for responsive agents, coding help, and everyday automation. gpt-5-mini by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
Output
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Cache read
Current public price: $0.02 / 1M tokens
Configured reference$0.025 / 1M tokens
400K context
View model details
65.
gpt-5.4

OpenAI

Agent-ready GPT for coding and computer-use workflows at a lower cost. gpt-5.4 by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
Output
Current public price: $12 / 1M tokens
Configured reference$15 / 1M tokens
Cache read
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
1.0M context
View model details
66.

Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation. gpt-5.4-nano by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
Output
Current public price: $1 / 1M tokens
Configured reference$1.25 / 1M tokens
Cache read
Current public price: $0.016 / 1M tokens
Configured reference$0.02 / 1M tokens
400K context
View model details
67.

GPT-6 Astra is built for harder reasoning, richer multimodal understanding, and long-context collaboration. It brings code, images, text, and task flow into one orbit — for work that needs sharper thinking, steadier output, and deeper memory. gpt-6-astra by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $8 / 1M tokens
Configured reference$10 / 1M tokens
Output
Current public price: $40 / 1M tokens
Configured reference$50 / 1M tokens
Cache read
Current public price: $0.8 / 1M tokens
Configured reference$1 / 1M tokens
Cache write
Current public price: $10 / 1M tokens
Configured reference$12.5 / 1M tokens
68.

Image model for prompt-driven generation, editing, and visual design workflows. gpt-image-2 by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per request
Current public price: $0.0088 / request
Configured reference$0.011 / request

gpt-image-2.5-flare by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $24 / 1M tokens
Configured reference$30 / 1M tokens
Cache read
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Image
Current public price: $6.4 / image
Configured reference$8 / image

gpt-image-2.5-sunburst by OpenAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4 / 1M tokens
Configured reference$5 / 1M tokens
Output
Current public price: $24 / 1M tokens
Configured reference$30 / 1M tokens
71.

xAI's frontier model for long-running agents, coding, knowledge work, and visual projects. grok-4.6 by xAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $4.8 / 1M tokens
Configured reference$6 / 1M tokens
Cache read
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens

Image model for prompt-driven generation, editing, and visual design workflows. grok-imagine-image by xAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per request
Current public price: $0.016 / request
Configured reference$0.02 / request

grok-imagine-image-pro by xAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per request
Current public price: $0.04 / request
Configured reference$0.05 / request

Image model for prompt-driven generation, editing, and visual design workflows. grok-imagine-image-quality by xAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per request
Current public price: $0.04 / request
Configured reference$0.05 / request

Image model for prompt-driven generation, editing, and visual design workflows. grok-imagine-video by xAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: $0.072 / second
Configured reference$0.09 / second

Video model for image-to-video generation, editing, and extension workflows. grok-imagine-video-1.5 by xAI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: $0.088 / second
Configured reference$0.11 / second
77.
kimi-k2.5

Moonshot AI

Earlier Kimi frontier model for long-context agents, coding, and multimodal work. kimi-k2.5 by Moonshot AI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.48 / 1M tokens
Configured reference$0.6 / 1M tokens
Output
Current public price: $2.4 / 1M tokens
Configured reference$3 / 1M tokens
Cache read
Current public price: $0.08 / 1M tokens
Configured reference$0.1 / 1M tokens
262K context
View model details
78.
kimi-k2.6

Moonshot AI

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context. kimi-k2.6 by Moonshot AI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.76 / 1M tokens
Configured reference$0.95 / 1M tokens
Output
Current public price: $3.2 / 1M tokens
Configured reference$4 / 1M tokens
Cache read
Current public price: $0.128 / 1M tokens
Configured reference$0.16 / 1M tokens
262K context
View model details
79.
Kimi K3

Moonshot AI

Tuned for long-horizon agent work. It holds a full 1M-token context, chains tool calls without losing the thread, and keeps an entire codebase in view across a session. Kimi K3 by Moonshot AI is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $2.4 / 1M tokens
Configured reference$3 / 1M tokens
Output
Current public price: $12 / 1M tokens
Configured reference$15 / 1M tokens
Cache read
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
1.0M context
View model details
80.M

macaron-v1-coding-venti by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $4.16 / 1M tokens
Configured reference$5.2 / 1M tokens
Output
Current public price: $14.56 / 1M tokens
Configured reference$18.2 / 1M tokens
Cache read
Current public price: $1.04 / 1M tokens
Configured reference$1.3 / 1M tokens
Cache write
Current public price: $1.04 / 1M tokens
Configured reference$1.3 / 1M tokens
1.0M context
View model details
81.M
macaron-v1-tall

Flatkey catalog

macaron-v1-tall by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.44 / 1M tokens
Configured reference$1.8 / 1M tokens
Output
Current public price: $8 / 1M tokens
Configured reference$10 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
262K context
View model details
82.M
macaron-v1-venti

Flatkey catalog

macaron-v1-venti by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $5.6 / 1M tokens
Configured reference$7 / 1M tokens
Output
Current public price: $20 / 1M tokens
Configured reference$25 / 1M tokens
Cache read
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Cache write
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
1.0M context
View model details
83.
minimax-m2

MiniMax

minimax-m2 by MiniMax is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $0.96 / 1M tokens
Configured reference$1.2 / 1M tokens
Cache read
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
205K context
View model details
84.

minimax-m2.5 by MiniMax is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $0.96 / 1M tokens
Configured reference$1.2 / 1M tokens
Cache read
Current public price: $0.024 / 1M tokens
Configured reference$0.03 / 1M tokens
205K context
View model details
85.

minimax-m2.7 by MiniMax is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $0.96 / 1M tokens
Configured reference$1.2 / 1M tokens
Cache read
Current public price: $0.048 / 1M tokens
Configured reference$0.06 / 1M tokens
205K context
View model details

nano-banana-pro-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $9.6 / 1M tokens
Configured reference$12 / 1M tokens
Cache read
Current public price: $0.16 / 1M tokens
Configured reference$0.2 / 1M tokens
Image
Current public price: $0.096 / image
Configured reference$0.12 / image
87.

Qwen vision-language model for visual reasoning, documents, and agent tasks. qwen3.5-27b by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.24 / 1M tokens
Configured reference$0.3 / 1M tokens
Output
Current public price: $1.92 / 1M tokens
Configured reference$2.4 / 1M tokens
128K context
View model details
88.

Qwen vision-language model for visual reasoning, documents, and agent tasks. qwen3.5-35b-a3b by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.196 / 1M tokens
Configured reference$0.245 / 1M tokens
Output
Current public price: $1.568 / 1M tokens
Configured reference$1.96 / 1M tokens
128K context
View model details

Large open Qwen multimodal MoE for visual agents and long technical tasks. qwen3.5-397b-a17b by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.4704 / 1M tokens
Configured reference$0.588 / 1M tokens
Output
Current public price: $2.8224 / 1M tokens
Configured reference$3.528 / 1M tokens
128K context
View model details
90.
qwen3.5-flash

Alibaba (China)

Qwen vision-language model for visual reasoning, documents, and agent tasks. qwen3.5-flash by Alibaba (China) is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.08 / 1M tokens
Configured reference$0.1 / 1M tokens
Output
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
128K context
View model details
91.

Qwen vision-language model for visual reasoning, documents, and agent tasks. qwen3.5-plus by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
Output
Current public price: $1.92 / 1M tokens
Configured reference$2.4 / 1M tokens
Cache read
Current public price: $0.032 / 1M tokens
Configured reference$0.04 / 1M tokens
128K context
View model details

Flagship Qwen model for complex reasoning, coding, and agentic workflows. qwen3.6-max-preview by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.04 / 1M tokens
Configured reference$1.3 / 1M tokens
Output
Current public price: $6.24 / 1M tokens
Configured reference$7.8 / 1M tokens
Cache read
Current public price: $0.104 / 1M tokens
Configured reference$0.13 / 1M tokens
1.0M context
View model details
93.

Earlier Qwen multimodal workhorse for million-token agent and document tasks. qwen3.6-plus by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.4 / 1M tokens
Configured reference$0.5 / 1M tokens
Output
Current public price: $2.4 / 1M tokens
Configured reference$3 / 1M tokens
Cache read
Current public price: $0.04 / 1M tokens
Configured reference$0.05 / 1M tokens
256K context
View model details
94.

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks. qwen3.7-max by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
Output
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
1.0M context
View model details
95.

Multimodal Qwen workhorse for long-context agents, visual inputs, and coding. qwen3.7-plus by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.32 / 1M tokens
Configured reference$0.4 / 1M tokens
Output
Current public price: $1.28 / 1M tokens
Configured reference$1.6 / 1M tokens
Cache read
Current public price: $0.064 / 1M tokens
Configured reference$0.08 / 1M tokens
256K context
View model details
96.

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows. qwen3.8-max by Alibaba is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $1.6 / 1M tokens
Configured reference$2 / 1M tokens
Output
Current public price: $4.8 / 1M tokens
Configured reference$6 / 1M tokens
Cache read
Current public price: $0.2 / 1M tokens
Configured reference$0.25 / 1M tokens
Cache write
Current public price: $2 / 1M tokens
Configured reference$2.5 / 1M tokens
1.0M context
View model details
97.
seedance-2.0

ByteDance

seedance-2.0 by ByteDance is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: from $0.037181 / second
Configured referencefrom $0.041312 / second
98.

seedance-2.0-fast by ByteDance is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: from $0.028535 / second
Configured referencefrom $0.031705 / second
99.

seedance-2.0-mini by ByteDance is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: from $0.018158 / second
Configured referencefrom $0.020176 / second
100.

seedance-2.0-pro by ByteDance is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: from $0.037181 / second
Configured referencefrom $0.041312 / second

Generate production-ready music from any video with synchronized timing, optional speech preservation, configurable ducking, segment-level direction, and up to 10 audio variants. Supports MP3, M4A, and WAV output through Flatkey's asynchronous API. sonilo-video-to-music by Sonilo is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per request
Current public price: $0.0072 / request
Configured reference$0.009 / request
102.T
typesafe/jev-1.13

Flatkey catalog

jev-1.13 is a purpose-built "judgment layer" model for AI Agents. Unlike traditional generative LLMs, it does not generate text. Instead, it outputs typed, structured decisions (Choice, Score, Noul) and calibrated probabilities. Ideal for routing, scoring, and conditional branching. typesafe/jev-1.13 by Flatkey catalog is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Input
Current public price: $0.0336 / 1M tokens
Configured reference$0.042 / 1M tokens

Video model for prompt-guided generation, editing, and motion workflows. veo-3.1-fast-generate-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: from $0.08 / second
Configured referencefrom $0.1 / second

Video model for prompt-guided generation, editing, and motion workflows. veo-3.1-generate-preview by Google is available through the Flatkey unified API and is included in our Discounted AI Models collection. Review its current pricing, context, and availability before integrating it.

Per second
Current public price: from $0.32 / second
Configured referencefrom $0.4 / second