Models Directory

Discover live model availability, pricing, endpoint support, and model detail pages.

ReasoningMultimodalLong ContextCoding

gpt-6-astra

GPT-6 Astra is built for harder reasoning, richer multimodal understanding, and long-context collaboration. It brings code, images, text, and task flow into one orbit — for work that needs sharper thinking, steadier output, and deeper memory.

Learn More

All Models

96 models found
ModelOfficialOur priceDiscountContextLatencyHealth score
glm-5.3Zhipu AI · GLMNew releaseInput$1.4per 1M tokensOutput$4.4per 1M tokensCache$0.26per 1M tokensInput$1.12per 1M tokensOutput$3.52per 1M tokensCache$0.208per 1M tokens-20.0% ↓1M600ms
100%
deepseek-v4-proDeepSeek · DeepSeekHOTInput$1.32per 1M tokensOutput$3.96per 1M tokensCache$0.044per 1M tokensInput$1.056per 1M tokensOutput$3.168per 1M tokensCache$0.0352per 1M tokens-20.0% ↓1M600ms
100%
gpt-5.6-solOpenAI · GPTHOTInput$5per 1M tokensOutput$30per 1M tokensCache$0.5per 1M tokensInput$4per 1M tokensOutput$24per 1M tokensCache$0.4per 1M tokens-20.0% ↓1M600ms
100%
kimi-k3Moonshot AI · KimiHOTInput$3per 1M tokensOutput$15per 1M tokensCache$0.3per 1M tokensInput$2.4per 1M tokensOutput$12per 1M tokensCache$0.24per 1M tokens-20.0% ↓1M600ms
100%
seedance-2.5ByteDance · SeedanceHOTfrom$0.084per secondfrom$0.084per second-0.0% ↓600ms
100%
gpt-6-astraOpenAINew releaseInput$10per 1M tokensOutput$50per 1M tokensCache$1per 1M tokensInput$8per 1M tokensOutput$40per 1M tokensCache$0.8per 1M tokens-20.0% ↓600ms
100%
gpt-image-2.5-flareOpenAINew release$8per image$6.4per image-20.0% ↓600ms
100%
gpt-image-2.5-sunburstOpenAINew releaseInput$5per 1M tokensOutput$30per 1M tokensCacheper 1M tokensInput$4per 1M tokensOutput$24per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
deepseek-v4-flashDeepSeek · DeepSeekInput$0.44per 1M tokensOutput$1.32per 1M tokensCache$0.014per 1M tokensInput$0.352per 1M tokensOutput$1.056per 1M tokensCache$0.0112per 1M tokens-20.0% ↓1M600ms
100%
openai/gpt-image-2.5-flareOpenAIInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
Aapodex-1-0-deep-discoverFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemma-4-31b-itGoogleInput$0.18per 1M tokensOutput$0.5per 1M tokensCacheper 1M tokensInput$0.144per 1M tokensOutput$0.4per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
veo-3.1-generate-previewGooglefrom$0.4per secondfrom$0.32per second-20.0% ↓600ms
100%
Aapodex-1-1-deep-solveFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-3.7-flashGoogleInput$1.5per 1M tokensOutput$7.5per 1M tokensCache$0.15per 1M tokensInput$1.2per 1M tokensOutput$6per 1M tokensCache$0.12per 1M tokens-20.0% ↓600ms
100%
Aapodex-1.1-miniFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-pro-latestGoogleInput$2per 1M tokensOutput$12per 1M tokensCache$0.2per 1M tokensInput$1.6per 1M tokensOutput$9.6per 1M tokensCache$0.16per 1M tokens-20.0% ↓600ms
100%
gemini-3.8-flashGoogleInput$2per 1M tokensOutput$8per 1M tokensCacheper 1M tokensInput$1.6per 1M tokensOutput$6.4per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
Aapodex-1-1-deep-discoverFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
openai/gpt-image-2.5-sunburstOpenAIInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-3.1-flash-lite-previewGoogleInput$0.25per 1M tokensOutput$1.5per 1M tokensCache$0.025per 1M tokensInput$0.2per 1M tokensOutput$1.2per 1M tokensCache$0.02per 1M tokens-20.0% ↓600ms
100%
gemini-flash-latestGoogleInput$0.3per 1M tokensOutput$2.5per 1M tokensCache$0.075per 1M tokensInput$0.24per 1M tokensOutput$2per 1M tokensCache$0.06per 1M tokens-20.0% ↓600ms
100%
Aapodex-1-0-deep-solveFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
nano-banana-pro-previewGoogle$0.12per image$0.096per image-20.0% ↓600ms
100%
Aapodex-1-1-deep-researchFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
veo-3.1-fast-generate-previewGooglefrom$0.1per secondfrom$0.08per second-20.0% ↓600ms
100%
gemma-4-26b-a4b-itGoogleInput$0.25per 1M tokensOutput$0.5per 1M tokensCacheper 1M tokensInput$0.2per 1M tokensOutput$0.4per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-flash-lite-latestGoogleInput$0.1per 1M tokensOutput$0.4per 1M tokensCache$0.025per 1M tokensInput$0.08per 1M tokensOutput$0.32per 1M tokensCache$0.02per 1M tokens-20.0% ↓600ms
100%
Aapodex-1.1Flatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
Aapodex-1-0-deep-researchFlatkey catalogInput$75per 1M tokensOutput$75per 1M tokensCacheper 1M tokensInput$60per 1M tokensOutput$60per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
deepseek-v3.2-thinkingDeepSeek · DeepSeekInput$0.29per 1M tokensOutput$0.42per 1M tokensCache$0.29per 1M tokensInput$0.232per 1M tokensOutput$0.336per 1M tokensCache$0.232per 1M tokens-20.0% ↓128K600ms
100%
deepseek-v3.2DeepSeek · DeepSeekInput$0.266per 1M tokensOutput$0.444per 1M tokensCache$0.0266per 1M tokensInput$0.2128per 1M tokensOutput$0.3552per 1M tokensCache$0.02128per 1M tokens-20.0% ↓128K600ms
100%
deepseek-v3.1DeepSeek · DeepSeekInput$0.21per 1M tokensOutput$0.79per 1M tokensCache$0.13per 1M tokensInput$0.168per 1M tokensOutput$0.632per 1M tokensCache$0.104per 1M tokens-20.0% ↓128K600ms
100%
deepseek-v3DeepSeek · DeepSeekInput$0.4per 1M tokensOutput$1.3per 1M tokensCache$0.4per 1M tokensInput$0.32per 1M tokensOutput$1.04per 1M tokensCache$0.32per 1M tokens-20.0% ↓128K600ms
100%
gemini-3.1-pro-previewGoogle · GeminiInput$2per 1M tokensOutput$12per 1M tokensCache$0.2per 1M tokensInput$1.6per 1M tokensOutput$9.6per 1M tokensCache$0.16per 1M tokens-20.0% ↓1M600ms
100%
gemini-3.1-pro-preview-customtoolsGoogle · GeminiInput$2per 1M tokensOutput$12per 1M tokensCache$0.2per 1M tokensInput$1.6per 1M tokensOutput$9.6per 1M tokensCache$0.16per 1M tokens-20.0% ↓1M600ms
100%
gemini-3.1-flash-liteGoogle · GeminiInput$0.25per 1M tokensOutput$1.5per 1M tokensCache$0.025per 1M tokensInput$0.2per 1M tokensOutput$1.2per 1M tokensCache$0.02per 1M tokens-20.0% ↓1M600ms
100%
gemini-3.1-flash-imageGoogle · Gemini$0.06per image$0.048per image-20.0% ↓600ms
100%
gemini-3.1-flash-lite-imageGoogle · Gemini$0.03per image$0.024per image-20.0% ↓600ms
100%
gemini-3.1-flash-tts-previewGoogle · GeminiInput$1per 1M tokensOutput$20per 1M tokensCacheper 1M tokensInput$0.8per 1M tokensOutput$16per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-3.5-flashGoogle · GeminiInput$1.5per 1M tokensOutput$9per 1M tokensCache$0.15per 1M tokensInput$1.2per 1M tokensOutput$7.2per 1M tokensCache$0.12per 1M tokens-20.0% ↓1M600ms
100%
gemini-3.6-flashGoogle · GeminiInput$1.5per 1M tokensOutput$7.5per 1M tokensCache$0.15per 1M tokensInput$1.2per 1M tokensOutput$6per 1M tokensCache$0.12per 1M tokens-20.0% ↓1M600ms
100%
gemini-3-pro-imageGoogle · Gemini$0.12per image$0.096per image-20.0% ↓600ms
100%
gemini-3.5-flash-liteGoogle · GeminiInput$0.3per 1M tokensOutput$2.5per 1M tokensCache$0.03per 1M tokensInput$0.24per 1M tokensOutput$2per 1M tokensCache$0.024per 1M tokens-20.0% ↓1M600ms
100%
gemini-2.5-proGoogle · GeminiInput$1.125per 1M tokensOutput$9per 1M tokensCache$0.1125per 1M tokensInput$0.9per 1M tokensOutput$7.2per 1M tokensCache$0.09per 1M tokens-20.0% ↓1M600ms
100%
gemini-2.5-flash-imageGoogle · Gemini$0.03per image$0.024per image-20.0% ↓600ms
100%
gemini-3-flash-previewGoogle · GeminiInput$0.07per 1M tokensOutput$0.43per 1M tokensCache$0.007per 1M tokensInput$0.056per 1M tokensOutput$0.344per 1M tokensCache$0.0056per 1M tokens-20.0% ↓1M600ms
100%
gemini-2.5-flashGoogle · GeminiInput$0.09per 1M tokensOutput$0.71per 1M tokensCache$0.0027per 1M tokensInput$0.072per 1M tokensOutput$0.568per 1M tokensCache$0.00216per 1M tokens-20.0% ↓1M600ms
100%
gemini-embedding-001Google · GeminiInput$0.15per 1M tokensOutputper 1M tokensCacheper 1M tokensInput$0.12per 1M tokensOutputper 1M tokensCacheper 1M tokens-20.0% ↓8K600ms
100%
gemini-2.5-flash-liteGoogle · GeminiInput$0.09per 1M tokensOutput$0.36per 1M tokensCache$0.0225per 1M tokensInput$0.072per 1M tokensOutput$0.288per 1M tokensCache$0.018per 1M tokens-20.0% ↓1M600ms
100%
gemini-2.5-pro-ttsGoogle · GeminiInput$1per 1M tokensOutput$20per 1M tokensCacheper 1M tokensInput$0.8per 1M tokensOutput$16per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-2.5-flash-ttsGoogle · GeminiInput$0.5per 1M tokensOutput$10per 1M tokensCacheper 1M tokensInput$0.4per 1M tokensOutput$8per 1M tokensCacheper 1M tokens-20.0% ↓600ms
100%
gemini-robotics-er-1.6-previewGoogle · GeminiInput$0.3per 1M tokensOutput$1.2per 1M tokensCache$0.3per 1M tokensInput$0.24per 1M tokensOutput$0.96per 1M tokensCache$0.24per 1M tokens-20.0% ↓128K600ms
100%
glm-5.2Zhipu AI · GLMInput$1.4per 1M tokensOutput$4.4per 1M tokensCache$0.26per 1M tokensInput$1.12per 1M tokensOutput$3.52per 1M tokensCache$0.208per 1M tokens-20.0% ↓1M600ms
100%
glm-5.1Zhipu AI · GLMInput$1.4per 1M tokensOutput$4.4per 1M tokensCache$0.26per 1M tokensInput$1.12per 1M tokensOutput$3.52per 1M tokensCache$0.208per 1M tokens-20.0% ↓205K600ms
100%
glm-5Zhipu AI · GLMInput$1per 1M tokensOutput$3.2per 1M tokensCache$0.2per 1M tokensInput$0.8per 1M tokensOutput$2.56per 1M tokensCache$0.16per 1M tokens-20.0% ↓205K600ms
100%
glm-5-turboZhipu AI · GLMInput$1.2per 1M tokensOutput$4per 1M tokensCache$0.24per 1M tokensInput$0.96per 1M tokensOutput$3.2per 1M tokensCache$0.192per 1M tokens-20.0% ↓205K600ms
100%
glm-4.7Zhipu AI · GLMInput$0.6per 1M tokensOutput$2.2per 1M tokensCache$0.11per 1M tokensInput$0.48per 1M tokensOutput$1.76per 1M tokensCache$0.088per 1M tokens-20.0% ↓205K600ms
100%
gpt-5.6-lunaOpenAI · GPTInput$0.2per 1M tokensOutput$1.2per 1M tokensCache$0.02per 1M tokensInput$0.16per 1M tokensOutput$0.96per 1M tokensCache$0.016per 1M tokens-20.0% ↓1M600ms
100%
gpt-5.5OpenAI · GPTInput$5per 1M tokensOutput$30per 1M tokensCache$0.5per 1M tokensInput$4per 1M tokensOutput$24per 1M tokensCache$0.4per 1M tokens-20.0% ↓1M600ms
100%
gpt-4o-miniOpenAI · GPTInput$0.15per 1M tokensOutput$0.6per 1M tokensCache$0.08per 1M tokensInput$0.12per 1M tokensOutput$0.48per 1M tokensCache$0.064per 1M tokens-20.0% ↓128K600ms
100%
gpt-5.4-miniOpenAI · GPTInput$0.75per 1M tokensOutput$4.5per 1M tokensCache$0.075per 1M tokensInput$0.6per 1M tokensOutput$3.6per 1M tokensCache$0.06per 1M tokens-20.0% ↓400K600ms
100%
gpt-5.6-terraOpenAI · GPTInput$2per 1M tokensOutput$12per 1M tokensCache$0.2per 1M tokensInput$1.6per 1M tokensOutput$9.6per 1M tokensCache$0.16per 1M tokens-20.0% ↓1M600ms
100%
gpt-5.4OpenAI · GPTInput$2.5per 1M tokensOutput$15per 1M tokensCache$0.25per 1M tokensInput$2per 1M tokensOutput$12per 1M tokensCache$0.2per 1M tokens-20.0% ↓1M600ms
100%
gpt-image-2OpenAI · GPT$0.011per image$0.0088per image-20.0% ↓600ms
100%
gpt-5-miniOpenAI · GPTInput$0.04per 1M tokensOutput$0.32per 1M tokensCache$0.0048per 1M tokensInput$0.032per 1M tokensOutput$0.256per 1M tokensCache$0.00384per 1M tokens-20.0% ↓400K600ms
100%
gpt-5.4-nanoOpenAI · GPTInput$0.2per 1M tokensOutput$1.25per 1M tokensCache$0.02per 1M tokensInput$0.16per 1M tokensOutput$1per 1M tokensCache$0.016per 1M tokens-20.0% ↓400K600ms
100%
gpt-4oOpenAI · GPTInput$2.5per 1M tokensOutput$10per 1M tokensCache$1.25per 1M tokensInput$2per 1M tokensOutput$8per 1M tokensCache$1per 1M tokens-20.0% ↓128K600ms
100%
gpt-4.1-miniOpenAI · GPTInput$0.4per 1M tokensOutput$1.6per 1M tokensCache$0.1per 1M tokensInput$0.32per 1M tokensOutput$1.28per 1M tokensCache$0.08per 1M tokens-20.0% ↓1M600ms
100%
grok-imagine-image-qualityxAI · Grok$0.05per image$0.04per image-20.0% ↓600ms
100%
grok-imagine-image-proxAI · Grok$0.05per image$0.04per image-20.0% ↓600ms
100%
grok-imagine-video-1.5xAI · Grok$0.11per second$0.088per second-20.0% ↓600ms
100%
grok-imagine-imagexAI · Grok$0.02per image$0.016per image-20.0% ↓600ms
100%
grok-imagine-videoxAI · Grok$0.09per second$0.072per second-20.0% ↓600ms
100%
kimi-k2.6Moonshot AI · KimiInput$0.95per 1M tokensOutput$4per 1M tokensCache$0.16per 1M tokensInput$0.76per 1M tokensOutput$3.2per 1M tokensCache$0.128per 1M tokens-20.0% ↓262K600ms
100%
kimi-k2.5Moonshot AI · KimiInput$0.6per 1M tokensOutput$3per 1M tokensCache$0.1per 1M tokensInput$0.48per 1M tokensOutput$2.4per 1M tokensCache$0.08per 1M tokens-20.0% ↓262K600ms
100%
macaron-v1-coding-ventiMacaron · MacaronInput$5.2per 1M tokensOutput$18.2per 1M tokensCache$1.3per 1M tokensInput$4.16per 1M tokensOutput$14.56per 1M tokensCache$1.04per 1M tokens-20.0% ↓1M600ms
100%
macaron-v1-ventiMacaron · MacaronInput$7per 1M tokensOutput$25per 1M tokensCache$2per 1M tokensInput$5.6per 1M tokensOutput$20per 1M tokensCache$1.6per 1M tokens-20.0% ↓1M600ms
100%
macaron-v1-tallMacaron · MacaronInput$1.8per 1M tokensOutput$10per 1M tokensCache$0.2per 1M tokensInput$1.44per 1M tokensOutput$8per 1M tokensCache$0.16per 1M tokens-20.0% ↓262K600ms
100%
MiniMax-H3MiniMax · MiniMaxfrom$0.08per secondfrom$0.08per second-0.0% ↓600ms
100%
minimax-m2.7MiniMax · MiniMaxInput$0.3per 1M tokensOutput$1.2per 1M tokensCache$0.06per 1M tokensInput$0.24per 1M tokensOutput$0.96per 1M tokensCache$0.048per 1M tokens-20.0% ↓205K600ms
100%
minimax-m2.5MiniMax · MiniMaxInput$0.3per 1M tokensOutput$1.2per 1M tokensCache$0.03per 1M tokensInput$0.24per 1M tokensOutput$0.96per 1M tokensCache$0.024per 1M tokens-20.0% ↓205K600ms
100%
minimax-m2MiniMax · MiniMaxInput$0.3per 1M tokensOutput$1.2per 1M tokensCache$0.3per 1M tokensInput$0.24per 1M tokensOutput$0.96per 1M tokensCache$0.24per 1M tokens-20.0% ↓205K600ms
100%
qwen3.8-maxAlibaba · QwenInput$2per 1M tokensOutput$6per 1M tokensCache$0.25per 1M tokensInput$1.6per 1M tokensOutput$4.8per 1M tokensCache$0.2per 1M tokens-20.0% ↓1M600ms
100%
qwen3.7-maxAlibaba · QwenInput$2.5per 1M tokensOutput$2.5per 1M tokensCacheper 1M tokensInput$2per 1M tokensOutput$2per 1M tokensCacheper 1M tokens-20.0% ↓1M600ms
100%
qwen3.7-plusAlibaba · QwenInput$0.4per 1M tokensOutput$1.6per 1M tokensCache$0.08per 1M tokensInput$0.32per 1M tokensOutput$1.28per 1M tokensCache$0.064per 1M tokens-20.0% ↓256K600ms
100%
qwen3.6-max-previewAlibaba · QwenInput$1.3per 1M tokensOutput$7.8per 1M tokensCache$0.13per 1M tokensInput$1.04per 1M tokensOutput$6.24per 1M tokensCache$0.104per 1M tokens-20.0% ↓1M600ms
100%
qwen3.6-plusAlibaba · QwenInput$0.5per 1M tokensOutput$3per 1M tokensCache$0.05per 1M tokensInput$0.4per 1M tokensOutput$2.4per 1M tokensCache$0.04per 1M tokens-20.0% ↓256K600ms
100%
qwen3.5-27bAlibaba · QwenInput$0.3per 1M tokensOutput$2.4per 1M tokensCacheper 1M tokensInput$0.24per 1M tokensOutput$1.92per 1M tokensCacheper 1M tokens-20.0% ↓128K600ms
100%
qwen3.5-397b-a17bAlibaba · QwenInput$0.588per 1M tokensOutput$3.528per 1M tokensCacheper 1M tokensInput$0.4704per 1M tokensOutput$2.8224per 1M tokensCacheper 1M tokens-20.0% ↓128K600ms
100%
qwen3.5-plusAlibaba · QwenInput$0.4per 1M tokensOutput$2.4per 1M tokensCache$0.04per 1M tokensInput$0.32per 1M tokensOutput$1.92per 1M tokensCache$0.032per 1M tokens-20.0% ↓128K600ms
100%
qwen3.5-flashAlibaba · QwenInput$0.1per 1M tokensOutput$0.4per 1M tokensCacheper 1M tokensInput$0.08per 1M tokensOutput$0.32per 1M tokensCacheper 1M tokens-20.0% ↓128K600ms
100%
qwen3.5-35b-a3bAlibaba · QwenInput$0.245per 1M tokensOutput$1.96per 1M tokensCacheper 1M tokensInput$0.196per 1M tokensOutput$1.568per 1M tokensCacheper 1M tokens-20.0% ↓128K600ms
100%
seedance-2.0-proByteDance · Seedancefrom$0.041312per secondfrom$0.037181per second-10.0% ↓600ms
100%
seedance-2.0ByteDance · Seedancefrom$0.041312per secondfrom$0.037181per second-10.0% ↓600ms
100%
sonilo-video-to-musicSonilo · Sonilo$0.009per request$0.0072per request-20.0% ↓600ms
100%