Transparent Pricing

200+ models. 1 API. 0 hidden fees.

218models
$5Sign up → $5.00 free

Eastern Efficiency

Maximum value at minimum cost — top-tier capability, unbeatable pricing

51 models

DeepSeek Chat (Legacy Alias)

Cost-optimized model with strong reasoning capabilities.

Input$0.65·Output$1.30/1M tokens
Reasoning128K ctx

DeepSeek V3.2

DeepSeek's cost-optimized workhorse — reliable, affordable, production-ready.

Input$0.41·Output$1.65/1M tokens
ValueReasoning64K ctx

DeepSeek V3.2 (nf)

DeepSeek's cost-optimized workhorse — reliable, affordable, production-ready.

Input$0.42·Output$1.60/1M tokens
ValueReasoning64K ctx

DeepSeek V4 Flash

DeepSeek's fastest model — lightning inference for high-throughput applications.

Input$0.21·Output$0.42/1M tokens
FastValue128K ctx

DeepSeek V4 Flash (nf)

DeepSeek's fastest model — lightning inference for high-throughput applications.

Input$0.21·Output$0.42/1M tokens
FastValue128K ctx

DeepSeek V4 Flash (2026-06)

DeepSeek's fastest model — lightning inference for high-throughput applications.

Input$0.21·Output$0.42/1M tokens
FastValue128K ctx

DeepSeek V4 Flash (Thinking)

DeepSeek's fastest model — lightning inference for high-throughput applications.

Input$0.21·Output$0.42/1M tokens
FastValue128K ctx

DeepSeek V4 Flash (Thinking)

DeepSeek's fastest model — lightning inference for high-throughput applications.

Input$0.21·Output$0.42/1M tokens
FastValue128K ctx

DeepSeek V4 Pro

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (nf)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (128K)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (64K)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (Coder)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (Fast)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (Thinking)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (Thinking)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Input$0.65·Output$1.30/1M tokens
ReasoningFlagshipExtreme Value128K ctx

GLM-5

High-performance AI model available through Barq unified gateway.

Input$1.50·Output$6.00/1M tokens
General128K ctx

GLM-5 (nf)

High-performance AI model available through Barq unified gateway.

Input$1.20·Output$3.60/1M tokens
General128K ctx

GLM-5.1

High-performance AI model available through Barq unified gateway.

Input$2.10·Output$6.60/1M tokens
General128K ctx

GLM-5.1 (nf)

High-performance AI model available through Barq unified gateway.

Input$1.80·Output$5.40/1M tokens
General128K ctx

GLM-5.2

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.30/1M tokens
General128K ctx

GLM-5.2 (Thinking)

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.30/1M tokens
Reasoning128K ctx

Kimi K2.5

High-performance AI model available through Barq unified gateway.

Input$2.25·Output$9.00/1M tokens
General128K ctx

Kimi K2.6

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.75/1M tokens
General128K ctx

Kimi K2.6 (nf)

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.60/1M tokens
General128K ctx

Kimi K2.6 (128K)

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.60/1M tokens
General128K ctx

Kimi K2.6 (32K)

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.60/1M tokens
General128K ctx

Kimi K2.6 (Non-Thinking)

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.60/1M tokens
Reasoning128K ctx

Kimi K2.6 (Thinking)

High-performance AI model available through Barq unified gateway.

Input$0.90·Output$3.60/1M tokens
Reasoning128K ctx

MiMo V2 Pro (mixaicloud)

High-performance AI model available through Barq unified gateway. Flagship variant.

Input$1.13·Output$3.00/1M tokens
FlagshipChinese Native1M ctx

MiMo V2.5

Xiaomi's flagship reasoning model — 1M context, superior multilingual performance.

Input$0.12·Output$0.48/1M tokens
FlagshipReasoningHuge Context1M ctx

MiMo 2.5 Pro

Xiaomi's flagship reasoning model — 1M context, superior multilingual performance.

Input$1.50·Output$4.50/1M tokens
FlagshipReasoningHuge Context1M ctx

MiMo V2.5 Pro (mixaicloud)

Xiaomi's flagship reasoning model — 1M context, superior multilingual performance.

Input$1.50·Output$4.50/1M tokens
FlagshipReasoningHuge Context1M ctx

MiMo V2.5 Pro (Non-Thinking)

Xiaomi's flagship reasoning model — 1M context, superior multilingual performance.

Input$1.50·Output$4.50/1M tokens
FlagshipReasoningHuge Context1M ctx

MiMo V2.5 Pro (Thinking)

Xiaomi's flagship reasoning model — 1M context, superior multilingual performance.

Input$1.50·Output$4.50/1M tokens
FlagshipReasoningHuge Context1M ctx

MiniMax M2.5

MiniMax model — strong multilingual and creative capabilities.

Input$1.50·Output$6.00/1M tokens
FastFlagship128K ctx

MiniMax M2.7

MiniMax model — strong multilingual and creative capabilities.

Input$2.25·Output$9.00/1M tokens
FastFlagship128K ctx

Qwen Max (Legacy)

Alibaba's flagship — native Chinese mastery with unbeatable cost-performance ratio.

Input$4.50·Output$18.00/1M tokens
Chinese NativeFlagshipValue32K ctx

Qwen3 32B

Alibaba model — native Chinese optimization, excellent value.

Input$0.12·Output$0.36/1M tokens
Chinese Native32K ctx

Qwen3.6 27B

Alibaba model — native Chinese optimization, excellent value.

Input$0.12·Output$0.36/1M tokens
Chinese Native32K ctx

Qwen 3.5 Flash

Alibaba model — native Chinese optimization, excellent value. Speed-optimized variant.

Input$0.22·Output$0.90/1M tokens
FastChinese Native32K ctx

Qwen 3.5 Plus

Alibaba model — native Chinese optimization, excellent value.

Input$1.20·Output$4.80/1M tokens
Chinese Native32K ctx

Qwen 3.6 Flash

Alibaba model — native Chinese optimization, excellent value. Speed-optimized variant.

Input$0.30·Output$1.20/1M tokens
FastChinese Native32K ctx

Qwen 3.6 Plus

Alibaba model — native Chinese optimization, excellent value.

Input$1.20·Output$4.80/1M tokens
Chinese Native32K ctx

Qwen 3.6 Plus (nf)

Alibaba model — native Chinese optimization, excellent value.

Input$1.20·Output$4.20/1M tokens
Chinese Native32K ctx

Qwen 3.7 Max

Alibaba model — native Chinese optimization, excellent value.

Input$3.75·Output$15.00/1M tokens
FlagshipChinese Native32K ctx

Qwen 3.7 Max (nf)

Alibaba model — native Chinese optimization, excellent value.

Input$4.50·Output$18.00/1M tokens
FlagshipChinese Native32K ctx

Qwen 3.7 Max (16K)

Alibaba model — native Chinese optimization, excellent value.

Input$4.50·Output$18.00/1M tokens
FlagshipChinese Native32K ctx

Qwen 3.7 Max (32K)

Alibaba model — native Chinese optimization, excellent value.

Input$4.50·Output$18.00/1M tokens
FlagshipChinese Native32K ctx

Qwen 3.7 Max (Non-Thinking)

Alibaba model — native Chinese optimization, excellent value.

Input$4.50·Output$18.00/1M tokens
FlagshipReasoningChinese Native32K ctx

Qwen 3.7 Max (Thinking)

Alibaba model — native Chinese optimization, excellent value.

Input$4.50·Output$18.00/1M tokens
FlagshipReasoningChinese Native32K ctx

Global Benchmarks

Industry-leading models — the reference standard for capability and pricing

135 models

Agnes 2.0 Flash (Free)

Agnes AI model — cost-efficient, production-optimized inference. Speed-optimized variant.

Input$0.06·Output$0.18/1M tokens
Fast128K ctx

Agnes Image 2.0 Flash

Agnes AI model — cost-efficient, production-optimized inference. Speed-optimized variant.

Input$0.0040·Output$0.0040/1M tokens
FastMultimodal128K ctx

Agnes Image 2.1 Flash

Agnes AI model — cost-efficient, production-optimized inference. Speed-optimized variant.

Input$0.0060·Output$0.0060/1M tokens
FastMultimodal128K ctx

Agnes Video V2.0

Agnes AI model — cost-efficient, production-optimized inference.

Input$0.05·Output$0.05/1M tokens
Multimodal128K ctx

ALLaM 2 7B (Groq)

High-performance AI model available through Barq unified gateway.

Input$0.03·Output$0.09/1M tokens
Arabic Native128K ctx

Orpheus Arabic Saudi

High-performance AI model available through Barq unified gateway.

Input$0.03·Output$0.09/1M tokens
Arabic Native128K ctx

Claude 2 (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$9.60·Output$28.80/1M tokens
General128K ctx

Claude 2.1 (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$9.60·Output$28.80/1M tokens
General128K ctx

Claude 3.5 Sonnet (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
General200K ctx

Claude 3 Haiku (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$0.30·Output$1.50/1M tokens
General128K ctx

Claude 3 Opus (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$18.00·Output$90.00/1M tokens
Flagship200K ctx

Claude 3 Sonnet (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
General200K ctx

Claude Fable 5

Anthropic model — reliable reasoning, safety-first design.

Input$12.00·Output$60.00/1M tokens
General128K ctx

Claude Haiku 4.5

Anthropic model — reliable reasoning, safety-first design.

Input$1.20·Output$6.00/1M tokens
General128K ctx

Claude Haiku 4.5 (mixaicloud)

Anthropic model — reliable reasoning, safety-first design.

Input$0.96·Output$4.80/1M tokens
General128K ctx

Claude Instant (Legacy)

Anthropic model — reliable reasoning, safety-first design.

Input$0.96·Output$2.88/1M tokens
General128K ctx

Claude Opus 4.5

Anthropic's heavy-duty reasoning engine — unmatched for deep analysis and complex logic.

Input$6.00·Output$30.00/1M tokens
ReasoningFlagshipCoding200K ctx

Claude Opus 4.5 (mixaicloud)

Anthropic's heavy-duty reasoning engine — unmatched for deep analysis and complex logic.

Input$6.00·Output$30.00/1M tokens
ReasoningFlagshipCoding200K ctx

Claude Opus 4.5 (Thinking)

Anthropic model — reliable reasoning, safety-first design.

Input$6.00·Output$30.00/1M tokens
FlagshipReasoning200K ctx

Claude Opus 4.5 (Vision)

Anthropic model — reliable reasoning, safety-first design.

Input$6.00·Output$30.00/1M tokens
FlagshipMultimodal200K ctx

Claude Opus 4.6

Anthropic model — reliable reasoning, safety-first design.

Input$6.00·Output$30.00/1M tokens
Flagship200K ctx

Claude Opus 4.7

Anthropic model — reliable reasoning, safety-first design.

Input$6.00·Output$30.00/1M tokens
Flagship200K ctx

Claude Opus 4.7 (Thinking)

Anthropic model — reliable reasoning, safety-first design.

Input$6.00·Output$30.00/1M tokens
FlagshipReasoning200K ctx

Claude Opus 4.8

Anthropic model — reliable reasoning, safety-first design.

Input$6.00·Output$30.00/1M tokens
Flagship200K ctx

Claude Sonnet 4.5

Anthropic's speed-intelligence hybrid — blazing coding with elite reasoning balance.

Input$3.60·Output$18.00/1M tokens
FastCodingReasoning200K ctx

Claude Sonnet 4.6

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
General200K ctx

Claude Sonnet 4.6 (mixaicloud)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
General200K ctx

Claude Sonnet 4.6 (128K)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
General200K ctx

Claude Sonnet 4.6 (200K)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
General200K ctx

Claude Sonnet 4.6 (Coder)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
Coding200K ctx

Claude Sonnet 4.6 (Non-Thinking)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
Reasoning200K ctx

Claude Sonnet 4.6 (Thinking)

Anthropic model — reliable reasoning, safety-first design.

Input$3.60·Output$18.00/1M tokens
Reasoning200K ctx

Gemini 1.0 Pro (Legacy)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$0.60·Output$1.80/1M tokens
FastFlagship1M ctx

Gemini 1.0 Pro Vision (Legacy)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$0.60·Output$1.80/1M tokens
FastFlagshipMultimodal1M ctx

Gemini 1.0 Ultra (Legacy)

Google model — multimodal, massive context, production-ready.

Input$2.40·Output$9.60/1M tokens
FastFlagship128K ctx

Gemini 1.5 Flash (Legacy)

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.09·Output$0.36/1M tokens
Fast1M ctx

Gemini 1.5 Pro (Legacy)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$6.00/1M tokens
FastFlagship1M ctx

Gemini 2.0 Flash (Legacy)

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.12·Output$0.48/1M tokens
Fast1M ctx

Gemini 2.5 Flash

Google's speed-optimized workhorse — fast, affordable, with vision capabilities.

Input$0.36·Output$3.00/1M tokens
FastMultimodalValue1M ctx

Gemini 2.5 Flash Image

Google's speed-optimized workhorse — fast, affordable, with vision capabilities.

Input$0.01·Output$0.01/1M tokens
FastMultimodalValue1M ctx

Gemini 2.5 Flash Lite

Google's speed-optimized workhorse — fast, affordable, with vision capabilities.

Input$0.10·Output$0.36/1M tokens
FastMultimodalValue1M ctx

Gemini 2.5 Pro

Google's production flagship — powerful reasoning with massive context window.

Input$1.50·Output$12.00/1M tokens
FlagshipReasoningHuge Context2M ctx

Gemini 3 Flash

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.36·Output$1.44/1M tokens
Fast1M ctx

Gemini 3 Pro Image

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$0.01·Output$0.01/1M tokens
FastFlagshipMultimodal1M ctx

Gemini 3 Pro Image

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$2.40·Output$14.40/1M tokens
FastFlagshipMultimodal1M ctx

Gemini 3 Pro

Google's latest frontier model — massive context, native multimodal reasoning.

Input$2.40·Output$9.60/1M tokens
FlagshipMultimodalHuge Context2M ctx

Gemini 3.1 Flash Image

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.36·Output$3.00/1M tokens
FastMultimodal1M ctx

Gemini 3.1 Flash Image

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$1.80·Output$10.80/1M tokens
FastMultimodal1M ctx

Gemini 3.1 Flash Lite

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.12·Output$0.48/1M tokens
Fast1M ctx

Gemini 3.1 Flash Lite (nf)

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.23·Output$0.90/1M tokens
Fast1M ctx

Gemini 3.1 Flash Lite Preview

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$0.12·Output$0.48/1M tokens
Fast1M ctx

Gemini 3.1 Pro (1M)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$12.00/1M tokens
FastFlagship1M ctx

Gemini 3.1 Pro (500K)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$12.00/1M tokens
FastFlagship1M ctx

Gemini 3.1 Pro

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$2.40·Output$14.40/1M tokens
FastFlagship1M ctx

Gemini 3.1 Pro (nf)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$12.00/1M tokens
FastFlagship1M ctx

Gemini 3.1 Pro Preview Customtools

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$6.00/1M tokens
FastFlagship1M ctx

Gemini 3.1 Pro (Non-Thinking)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$12.00/1M tokens
FastFlagshipReasoning1M ctx

Gemini 3.1 Pro (Thinking)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$12.00/1M tokens
FastFlagshipReasoning1M ctx

Gemini 3.1 Pro (Vision)

Google model — multimodal, massive context, production-ready. Flagship variant.

Input$1.50·Output$12.00/1M tokens
FastFlagshipMultimodal1M ctx

Gemini 3.5 Flash

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Input$1.80·Output$10.80/1M tokens
Fast1M ctx

GPT-3.5 Turbo (Legacy)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.60·Output$1.80/1M tokens
General128K ctx

GPT-3.5 Turbo 16K (Legacy)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$3.60·Output$4.80/1M tokens
General128K ctx

GPT-3.5 Turbo Instruct (Legacy)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$1.80·Output$2.40/1M tokens
General128K ctx

GPT-4 (Legacy)

OpenAI's flagship multimodal model — fast, versatile, with vision and reasoning.

Input$36.00·Output$72.00/1M tokens
MultimodalFastFlagship128K ctx

GPT-4 0314 (Legacy)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$36.00·Output$72.00/1M tokens
General128K ctx

GPT-4 0613 (Legacy)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$36.00·Output$72.00/1M tokens
General128K ctx

GPT-4 32K (Legacy)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$72.00·Output$144.00/1M tokens
General128K ctx

GPT-4 Turbo (Legacy)

128K extended context — ideal for complex reasoning and enterprise code generation.

Input$12.00·Output$36.00/1M tokens
ReasoningCodingHuge Context128K ctx

GPT-4.1

OpenAI model — versatile, fast, industry-leading capabilities.

Input$3.00·Output$12.00/1M tokens
General128K ctx

GPT-4o

OpenAI's flagship multimodal model — fast, versatile, with vision and reasoning.

Input$3.00·Output$12.00/1M tokens
MultimodalFastFlagship128K ctx

GPT-4o 2024-05-13 (Legacy)

OpenAI's flagship multimodal model — fast, versatile, with vision and reasoning.

Input$6.00·Output$18.00/1M tokens
MultimodalFastFlagship128K ctx

GPT-4o Mini (Legacy)

OpenAI's flagship multimodal model — fast, versatile, with vision and reasoning.

Input$0.18·Output$0.72/1M tokens
MultimodalFastFlagship128K ctx

GPT-5

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$1.50·Output$12.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5 (mixaicloud)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$1.50·Output$12.00/1M tokens
General128K ctx

GPT-5 Mini

OpenAI's cost-efficient model — fast, capable, great value for everyday tasks.

Input$0.30·Output$2.40/1M tokens
FastValue128K ctx

GPT-5 (Non-Thinking)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$1.50·Output$12.00/1M tokens
Reasoning128K ctx

GPT-5 (Thinking)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$1.50·Output$12.00/1M tokens
Reasoning128K ctx

GPT-5.1

OpenAI model — versatile, fast, industry-leading capabilities.

Input$1.50·Output$12.00/1M tokens
General128K ctx

GPT-5.2

OpenAI model — versatile, fast, industry-leading capabilities.

Input$2.10·Output$16.80/1M tokens
General128K ctx

GPT-5.2 (Function Calling)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$2.10·Output$16.80/1M tokens
General128K ctx

GPT-5.2 (Non-Thinking)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$2.10·Output$16.80/1M tokens
Reasoning128K ctx

GPT-5.2 (Thinking)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$2.10·Output$16.80/1M tokens
Reasoning128K ctx

GPT-5.3 Codex

OpenAI model — versatile, fast, industry-leading capabilities.

Input$2.40·Output$19.20/1M tokens
Coding128K ctx

GPT-5.3 Codex (nf)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.96·Output$3.84/1M tokens
Coding128K ctx

GPT-5.4

OpenAI model — versatile, fast, industry-leading capabilities.

Input$3.00·Output$18.00/1M tokens
General128K ctx

GPT 5.4 Mini

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.72·Output$2.88/1M tokens
Fast128K ctx

GPT-5.4 Mini (nf)

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.18·Output$0.72/1M tokens
Fast128K ctx

GPT-5.4 Pro

OpenAI model — versatile, fast, industry-leading capabilities. Flagship variant.

Input$3.00·Output$18.00/1M tokens
Flagship128K ctx

GPT-5.4 Pro (nf)

OpenAI model — versatile, fast, industry-leading capabilities. Flagship variant.

Input$0.66·Output$1.32/1M tokens
Flagship128K ctx

GPT-5.5

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$6.00·Output$36.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (128K)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (32K)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (Coder)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (Function Calling)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (Non-Thinking)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (Thinking)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

GPT-5.5 (Vision)

OpenAI's most advanced model — superior reasoning, coding, and agentic capabilities.

Input$3.00·Output$24.00/1M tokens
FlagshipReasoningCoding256K ctx

Gpt Image 2

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.02·Output$0.02/1M tokens
Multimodal128K ctx

Grok 4

xAI's latest model — real-time knowledge, unfiltered reasoning, massive context.

Input$6.00·Output$18.00/1M tokens
FlagshipReasoningHuge Context1M ctx

Grok 4 (mixaicloud)

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.96·Output$3.84/1M tokens
General1M ctx

Grok 4 1 Fast Non Reasoning

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.24·Output$0.60/1M tokens
FastReasoning1M ctx

Grok 4 1 Fast Reasoning

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.24·Output$0.60/1M tokens
FastReasoning1M ctx

Grok 4 Fast Non Reasoning

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.48·Output$2.40/1M tokens
FastReasoning1M ctx

Grok 4 Fast Reasoning

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.48·Output$2.40/1M tokens
FastReasoning1M ctx

Grok 4 Fast (Non-Reasoning)

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.48·Output$2.40/1M tokens
FastReasoning1M ctx

Grok 4 Fast (Reasoning)

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.48·Output$2.40/1M tokens
FastReasoning1M ctx

Grok 4.1

xAI model — real-time knowledge, bold reasoning, large context.

Input$3.60·Output$18.00/1M tokens
General1M ctx

Grok 4.1 (mixaicloud)

xAI model — real-time knowledge, bold reasoning, large context.

Input$1.20·Output$4.80/1M tokens
General1M ctx

Grok 4.1 Fast

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.24·Output$0.60/1M tokens
Fast1M ctx

Grok 4.1 Fast (Non-Reasoning)

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.24·Output$0.60/1M tokens
FastReasoning1M ctx

Grok 4.1 Fast (Reasoning)

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.24·Output$0.60/1M tokens
FastReasoning1M ctx

Grok 4.2

xAI model — real-time knowledge, bold reasoning, large context.

Input$3.60·Output$18.00/1M tokens
General1M ctx

Grok 4.2 (mixaicloud)

xAI model — real-time knowledge, bold reasoning, large context.

Input$1.80·Output$7.20/1M tokens
General1M ctx

Grok 4.20 Non-Reasoning

xAI model — real-time knowledge, bold reasoning, large context.

Input$1.50·Output$3.00/1M tokens
Reasoning1M ctx

Grok 4.20 Reasoning

xAI model — real-time knowledge, bold reasoning, large context.

Input$1.50·Output$3.00/1M tokens
Reasoning1M ctx

Grok 4.20 Multi-Agent

xAI model — real-time knowledge, bold reasoning, large context.

Input$1.50·Output$3.00/1M tokens
General1M ctx

Grok 4.3 (mixaicloud)

xAI's latest model — real-time knowledge, unfiltered reasoning, massive context.

Input$2.40·Output$9.60/1M tokens
FlagshipReasoningHuge Context1M ctx

Grok 4.3 (1M)

xAI's latest model — real-time knowledge, unfiltered reasoning, massive context.

Input$2.40·Output$9.60/1M tokens
FlagshipReasoningHuge Context1M ctx

Grok 4.3 (256K)

xAI's latest model — real-time knowledge, unfiltered reasoning, massive context.

Input$2.40·Output$9.60/1M tokens
FlagshipReasoningHuge Context1M ctx

Grok 4.3 (Non-Thinking)

xAI's latest model — real-time knowledge, unfiltered reasoning, massive context.

Input$2.40·Output$9.60/1M tokens
FlagshipReasoningHuge Context1M ctx

Grok 4.3 (Thinking)

xAI's latest model — real-time knowledge, unfiltered reasoning, massive context.

Input$2.40·Output$9.60/1M tokens
FlagshipReasoningHuge Context1M ctx

Grok Imagine Image

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.08·Output$0.08/1M tokens
Multimodal1M ctx

Grok Imagine Image Quality

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.18·Output$0.18/1M tokens
Multimodal1M ctx

Grok Imagine Video

xAI model — real-time knowledge, bold reasoning, large context.

Input$0.30·Output$0.30/1M tokens
Multimodal1M ctx

Kling V3

High-performance AI model available through Barq unified gateway.

Input$0.04·Output$0.04/1M tokens
General128K ctx

Llama 3.1 8B Instant

Meta open-source model — versatile, community-driven, reliable.

Input$0.03·Output$0.09/1M tokens
Open Source128K ctx

Llama 3.3 70B Versatile

Meta open-source model — versatile, community-driven, reliable.

Input$0.12·Output$0.36/1M tokens
Open Source128K ctx

Llama 4 Scout 17B 16E Instruct

Meta open-source model — versatile, community-driven, reliable.

Input$0.06·Output$0.18/1M tokens
Open Source128K ctx

Mistral Large (Legacy)

Mistral model — efficient, open-weight, strong multilingual support.

Input$4.80·Output$14.40/1M tokens
General128K ctx

Mistral Medium (Legacy)

Mistral model — efficient, open-weight, strong multilingual support.

Input$3.24·Output$9.72/1M tokens
General128K ctx

Mistral Small (Legacy)

Mistral model — efficient, open-weight, strong multilingual support.

Input$1.20·Output$3.60/1M tokens
General128K ctx

Nemotron 3.5 Content Safety

NVIDIA reasoning model — MoE architecture, high throughput.

Input$0.50·Output$0.50/1M tokens
General128K ctx

Gpt Oss 120B

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.48·Output$1.44/1M tokens
Open Source128K ctx

Gpt Oss 20B

OpenAI model — versatile, fast, industry-leading capabilities.

Input$0.12·Output$0.36/1M tokens
Open Source128K ctx

PaLM 2 (Legacy)

High-performance AI model available through Barq unified gateway.

Input$0.60·Output$2.40/1M tokens
General128K ctx

Free Models

Zero-cost AI exploration — start building immediately

32 models

Agnes 1.5 Flash

Agnes AI model — cost-efficient, production-optimized inference. Speed-optimized variant.

Free
Fast128K ctx

ALLaM 2 7B (Groq Free)

High-performance AI model available through Barq unified gateway.

Free
Arabic Native128K ctx

Orpheus V1 English

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

DeepSeek V4 Flash (Groq Free)

DeepSeek's fastest model — lightning inference for high-throughput applications.

Free
FastValue128K ctx

DeepSeek V4 Pro (Groq Free)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Free
ReasoningFlagshipExtreme Value128K ctx

DeepSeek V4 Pro (Free Tier)

DeepSeek's flagship — superior multilingual reasoning at extreme value pricing.

Free
ReasoningFlagshipExtreme Value128K ctx

Gemini 3 Flash Preview (Free)

Google model — multimodal, massive context, production-ready. Speed-optimized variant.

Free
FastFree1M ctx

Gemma 2 9B

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Gemma 3 12B

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Gemma 3 27B

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

GPT-OSS 120B (Groq Free)

OpenAI model — versatile, fast, industry-leading capabilities.

Free
General128K ctx

GPT-OSS 20B (Groq Free)

OpenAI model — versatile, fast, industry-leading capabilities.

Free
General128K ctx

Groq Compound

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Groq Compound Mini

High-performance AI model available through Barq unified gateway.

Free
Fast128K ctx

Llama 2 13B (Legacy)

Meta open-source model — versatile, community-driven, reliable.

Free
Open Source128K ctx

Llama 2 70B (Legacy)

Meta open-source model — versatile, community-driven, reliable.

Free
Open Source128K ctx

Llama 2 7B (Legacy)

Meta open-source model — versatile, community-driven, reliable.

Free
Open Source128K ctx

Llama 3.1 8B (Groq Free)

Meta open-source model — versatile, community-driven, reliable.

Free
Open Source128K ctx

Llama 3.1 8B (Free)

Meta open-source model — versatile, community-driven, reliable.

Free
FreeOpen Source128K ctx

Llama 3.3 70B (Groq Free)

Meta open-source model — versatile, community-driven, reliable.

Free
Open Source128K ctx

Llama 4 Scout (Groq Free)

Meta open-source model — versatile, community-driven, reliable.

Free
Open Source128K ctx

Llama Prompt Guard 2 (22M)

Meta open-source model — versatile, community-driven, reliable. Flagship variant.

Free
FlagshipOpen Source128K ctx

Llama Prompt Guard 2 (86M)

Meta open-source model — versatile, community-driven, reliable. Flagship variant.

Free
FlagshipOpen Source128K ctx

Mixtral 8x22B (Legacy)

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Mixtral 8x7B (Legacy)

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Nemotron 3 Ultra (Thinking)

NVIDIA reasoning model — MoE architecture, high throughput.

Free
FlagshipReasoning1M ctx

GPT-OSS Safeguard 20B

OpenAI model — versatile, fast, industry-leading capabilities.

Free
Open Source128K ctx

Orpheus Arabic Saudi (Groq Free)

High-performance AI model available through Barq unified gateway.

Free
Arabic Native128K ctx

Qwen 3 32B (Groq Free)

Alibaba model — native Chinese optimization, excellent value.

Free
Chinese Native32K ctx

Qwen 3.6 27B (Groq Free)

Alibaba model — native Chinese optimization, excellent value.

Free
Chinese Native32K ctx

Whisper Large v3

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Whisper Large v3 Turbo

High-performance AI model available through Barq unified gateway.

Free
General128K ctx

Invite a friend, both get free tokens!

💡 Login to get your personal referral link

Read our referral program policy