AI Models & Pricing

Explore 854 AI models with transparent per-token pricing — ChatGPT, Claude, Gemini, DeepSeek, Qwen and more, all through one unified API.

Popular models

Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview

Browse by model author

OpenAI (133) · Anthropic (28) · Google (83) · Grok (26) · Qwen (141) · DeepSeek (38) · Z.AI (65) · ByteDance (46) · Llama (50) · AI21 (2) · Microsoft (14) · Cohere (19) · Mistral (10) · Yi (6) · Moonshot AI (34) · StepFun (4) · Nvidia (17) · Minimax (31) · Perplexity (4) · Baichuan (5) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (5) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · BAAI (5) · Agnes (4) · Meta (3) · KLing (2) · Poolside (2) · Dots Studio (1) · Liquid (1) · Stealth (1)

Auto

by OpenAI

AIHubMix Smart Router: Fill in the model name as auto, and the gateway will automatically…

$2/1M in · $2/1M out
1,000,000 tokens context

GLM 5.3

by Z.AI

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software…

$1.127 $1.014/1M in · $3.944 $3.549/1M out
10% off · 00:00–23:59 UTC
1,000,000 tokens context

Coding GLM 5.3

by Z.AI

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex…

$0.06/1M in · $0.22/1M out

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Ox Alpha

by Stealth

Developed by Stealth, Ox Alpha is a reasoning model designed for coding, sustained…

1,048,576 tokens context

Dots 3 Note Preview (free)

by Dots Studio

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio…

512,000 tokens context

Gemini 3.7 Flash (free)

by Google

Gemini 3.7 Flash free version: Free model resources are limited and provided only for…

1,000,000 tokens context

GLM 5.2

by Z.AI

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable…

$1.127/1M in · $3.944/1M out
1,000,000 tokens context

DeepSeek V4 Flash 0731

by DeepSeek

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model…

$0.142/1M in · $0.284/1M out
1,000,000 tokens context

DeepSeek V4 Flash Vision Exp

by DeepSeek

DeepSeek’s officially released new multimodal visual-understanding model…

$0.142/1M in · $0.284/1M out
1,000,000 tokens context

DeepSeek V4 Pro 0813

by DeepSeek

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent…

$0.692/1M in · $2.075/1M out
1,000,000 tokens context

Grok 4.6

by Grok

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running…

$2/1M in · $6/1M out
500,000 tokens context

DeepSeek V4 Flash 0731 Fast

by DeepSeek

DeepSeek V4 Flash 0731 Fast is a high-speed deployment of DeepSeek’s agentic model…

$0.28/1M in · $0.56/1M out
1,000,000 tokens context

Mai Thinking 1

by Microsoft

MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for…

$2/1M in · $8/1M out
256,000 tokens context

GPT 5.6 Sol Disc

by OpenAI

GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex…

$5 $2.5/1M in · $30 $15/1M out
50% off · 00:00–23:59 UTC
1,050,000 tokens context

Doubao Seedance 2.5 260628

by ByteDance

Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video…

$2/1M in

GPT 5.6 Luna

by OpenAI

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly…

$0.2/1M in · $1.2/1M out
1,050,000 tokens context

GPT 5.6 Sol

by OpenAI

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving…

$5/1M in · $30/1M out
1,050,000 tokens context

GPT 5.6 Terra

by OpenAI

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly…

$2/1M in · $12/1M out
1,050,000 tokens context

Agnes 2.5 Flash

by Agnes

Agnes 2.5 Flash is Agnes AI’s fast and efficient language model, an upgraded, fully…

$0.03/1M in · $0.15/1M out
512,000 tokens context

Agnes 2.5 Pro

by Agnes

Agnes 2.5 Pro is Agnes AI’s paid inference model and the commercially stable version of…

$0.45/1M in · $0.9/1M out
1,000,000 tokens context

Agnes 2.5 Pro Alpha

by Agnes

Agnes 2.5 Pro Alpha is Agnes AI’s paid inference model, suitable for advanced coding…

$0.45/1M in · $0.9/1M out
1,000,000 tokens context

Agnes Image 2.1 Flash

by Agnes

Agnes Image 2.1 Flash is Agnes AI’s high-performance image generation and image editing…

$2/1M in · $2/1M out

Grok 4.5

by Grok

Grok 4.5 was trained on datasets spanning knowledge in coding, science, engineering, and…

$2/1M in · $6/1M out
500,000 tokens context

Lfm 2.5 2.6b (free)

by Liquid

LFM-2.5-2.6B is a compact reasoning model developed by Liquid, featuring a generous…

128,000 tokens context

MiniMax H3

by Minimax

MiniMax H3 (minimax-h3) is a general-purpose omni-modal generation model developed by the…

$2/1M in · $2/1M out

Muse Glimmer 30B

by Meta

Muse Glimmer 30B is a dense causal language model distilled from Muse Spark, designed for…

$0.35/1M in · $1.5/1M out
131,000 tokens context

Nemotron Lightning 3.5 30B A3B

by Nvidia

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring…

$0.05/1M in · $0.2/1M out
262,000 tokens context

Qwen3.8 2.4t A95B

by Qwen

Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…

$2/1M in · $6/1M out
262,000 tokens context

Claude Opus 5

by Anthropic

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and…

$5/1M in · $25/1M out
1,000,000 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Ling 3.0 Tiny (free)

by Inclusionai

Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, featuring 1.3B active…

262,144 tokens context

Nemotron 3.5 Lightning (free)

by Nvidia

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring…

1,000,000 tokens context

Qwen Image 3.0

by Qwen

Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…

$2/1M in

Qwen Image 3.0 Pro

by Qwen

Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation…

$2/1M in

Qwen3.8 Max

by Qwen

Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built…

$1.69/1M in · $5.07/1M out
991,000 tokens context

Claude Sonnet 5

by Anthropic

Claude Sonnet 5 is the next generation of Anthropic's Sonnet model family. It is a…

$2/1M in · $10/1M out
1,000,000 tokens context

Kimi K3

by Moonshot AI

Kimi K3 is Kimi’s flagship model for long-horizon coding and end-to-end knowledge work…

$3/1M in · $15/1M out
1,048,576 tokens context

Ling 3.0 Flash (free)

by InclusionAI

Developed by Inclusionai, ling-3.0-flash-free is a 124B-parameter Mixture-of-Experts…

262,144 tokens context

Muse Spark 1.2

by Meta

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It…

$1.375/1M in · $4.675/1M out
1,000,000 tokens context

Qwen3.8 Max Preview

by Qwen

Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in…

$0.338/1M in · $1.014/1M out
983,616 tokens context

Gemini 3.1 Flash Lite Image

by Google

Google's newest, most compact, and most cost-effective image generation and editing…

$0.25/1M in · $1.5/1M out

Gemini 3.5 Flash Lite

by Google

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for…

$0.3/1M in · $2.5/1M out
1,048,576 tokens context

Gemini 3.5 Flash Lite (free)

by Google

Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…

1,048,576 tokens context

Gemini 3.6 Flash (free)

by Google

Gemini 3.6 Flash free version: Free model resources are limited and provided only for…

1,000,000 tokens context

GLM 5.2 Fast Preview

by Z.AI

GLM-5.2-Fast-Preview is the high-speed version of Zhipu AI’s flagship model GLM-5.2…

$2.254/1M in · $7.889/1M out
1,000,000 tokens context

Muse Spark 1.1

by Meta

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It…

$1.375/1M in · $4.675/1M out
1,000,000 tokens context

Qwen Audio 3.0 Tts Flash

by Qwen

qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for…

$14.2/1M in · $14.2/1M out

Qwen Audio 3.0 Tts Plus

by Qwen

qwen-audio-3.0-tts-plus is a high-performance speech synthesis large model designed for…

$15/1M in · $15/1M out

Claude Fable 5

by Anthropic

Anthropic's most capable widely released model, for the most demanding reasoning and…

$11/1M in · $55/1M out
1,000,000 tokens context

Jina Reranker V3.5

by Jina AI

jina-reranker-v3.5 is a 0.6B-parameter multilingual listwise document reranker and a…

$0.05/1M in · $0.05/1M out
131,000 tokens context

Claude Opus 4.8

by Anthropic

Claude Opus 4.8 is Anthropic’s newest and most powerful publicly available model. It is…

$5/1M in · $25/1M out
200,000 tokens context

Hy3

by Hunyuan

The Hy3 official version is honed for real-world business scenarios, using a…

$0.156/1M in · $0.625/1M out
256,000 tokens context

Doubao Seed 2.1 Pro

by ByteDance

A new generation of large models moving toward production-grade intelligence…

$0.929/1M in · $4.647/1M out
256,000 tokens context

Doubao Seed 2.1 Turbo

by ByteDance

Balancing performance and cost, comprehensively upgrading coding, agent, and multimodal…

$0.465/1M in · $2.324/1M out
256,000 tokens context

Mai Image 2.5 Pro

by Microsoft

MAI-Image-2.5 is Microsoft's flagship AI image generation and editing model. With…

$5/1M in · $5/1M out

Gemini 3.5 Flash

by Google

Gemini 3.5 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $9/1M out
1,000,000 tokens context

Grok Build 0.1

by Grok

Fast coding model trained specifically for agentic coding workflows.

$1/1M in · $2/1M out
256,000 tokens context

Mai Image 2.5

by Microsoft

MAI-Image-2.5 is Microsoft's flagship AI image generation and editing model. With…

$5/1M in · $5/1M out

Mai Image 2.5 Flash

by Microsoft

MAI-Image-2.5 is Microsoft's flagship AI image generation and editing model. With…

$1.75/1M in · $1.75/1M out

Coding Kimi K3

by Moonshot AI
$0.44/1M in · $1.613/1M out
1,048,576 tokens context

Happyhorse 1.1 I2v

by Qwen

HappyHorse-1.1-I2V supports image-to-video generation, further enhancing visual texture…

$2/1M in

Happyhorse 1.1 R2v

by Qwen

HappyHorse-1.1-R2V supports reference-based video generation, further improving the…

$2/1M in

Happyhorse 1.1 T2v

by Qwen

HappyHorse-1.1-T2V supports text-to-video generation, further enhancing text semantic…

$2/1M in

Coding GLM 5.2 (free)

by Z.AI

coding glm 5.2 free (coding-glm-5.2-free) is a free coding-focused API route offered by…

Coding Kimi K3 (free)

by Moonshot AI

coding-kimi-k3-free is the open and free version of coding-kimi-k3. To maintain reliable…

1,048,576 tokens context

Gemini 3.1 Flash Image

by Google

gemini-3.1-flash-image (Nano Banana 2) features professional-grade visual intelligence…

$0.5/1M in · $3/1M out

GPT Oss 20B (free)

by Openai

Developed by OpenAI, gpt-oss-20b-free is an open-weight 21B parameter model released…

131,072 tokens context

Kimi K2.7 Code

by Moonshot AI

Kimi K2.7 Code is Kimi’s most intelligent Coding model, capable of completing programming…

$0.95/1M in · $3.999/1M out
262,144 tokens context

Kimi K2.7 Code Highspeed

by Moonshot AI

High-Speed version of Kimi K2.7 Code model, with output speed of approximately 180…

$1.9/1M in · $7.999/1M out
262,144 tokens context

Gemini 3 Pro Image

by Google

Gemini-3-Pro-Image (Nano Banana Pro) is a high-performance image generation and editing…

$2/1M in · $12/1M out

GPT 4o Transcribe Diarize

by OpenAI

GPT-4o Transcribe Diarize is an automatic speech recognition (ASR) model with built-in…

$2.5/1M in · $10/1M out
16,000 tokens context

GPT Audio 1.5

by OpenAI

The gpt-audio model is OpenAI's first officially released (generally available) audio…

$2.5/1M in · $10/1M out
128,000 tokens context

Hy 3d 3.1

by Hunyuan

Using the Hunyuan Sheng 3D 3.1 model, it can generate higher-precision and higher-quality…

$2/1M in · $2/1M out

Kling V3 Omni

by KLing

VIDEO 3.0 Omni: All-in-One Multimodal Input, Voice-Driven Characters, Direct Audio-Visual…

$2/1M in

Kling Video O1

by KLing

Kling Video O1 is a major unified multimodal video model launched by Kuaishou. It…

$2/1M in

Longcat 2.0

by Meituan

Designed for agent development scenarios, it natively supports tool invocation…

$0.775/1M in · $3.098/1M out

Nemotron Nano 9B V2 (free)

by Nvidia

NVIDIA-Nemotron-Nano-9B-v2-free is a large language model trained from scratch by NVIDIA…

128,000 tokens context

Hy3 Preview

by Hunyuan

Hunyuan Hy3 preview is designed for agent workloads, adopting a MoE architecture with…

$0.17/1M in · $0.567/1M out
256,000 tokens context

MiniMax M3

by Minimax

The MiniMax M3 is a flagship programming model built for real-world productivity. As a…

$0.288/1M in · $1.152/1M out
204,800 tokens context

Nemotron Nano 12B V2 VL (free)

by Nvidia

Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open…

128,000 tokens context

Qwen3.7 Flash

by Qwen

The Qwen 3.7 series' mid-to-high cost-performance "Plus" model builds on strong text…

$0.028/1M in · $0.113/1M out
991,000 tokens context

Qwen3.7 Plus

by Qwen

The Qwen 3.7 series' mid-to-high cost-performance "Plus" model builds on strong text…

$0.282/1M in · $1.128/1M out
991,000 tokens context

Step 3.7 Flash

by StepFun

step-3.7-flash is stepfun's flagship inference model, designed for high-complexity tasks…

$0.22/1M in · $1.32/1M out
256,000 tokens context

Claude Opus 4.8 Thinking

by Anthropic

The claude-opus-4-8-think model has adaptive thinking mode pre-enabled; the default…

$5/1M in · $25/1M out
200,000 tokens context

Nemotron 3 Super 120B A12B (free)

by Nvidia

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid…

262,144 tokens context

Nemotron 3 Nano Omni 30B A3B (reasoning) (free)

by Nvidia

Developed by Nvidia, NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model…

256,000 tokens context

Nemotron 3 Ultra 550B A55B (free)

by Nvidia

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model featuring…

1,000,000 tokens context

Qwen3.7 Max

by Qwen

The Max model, the largest and most capable in the Qwen3.7 series, is currently offering…

$1.69/1M in · $5.07/1M out
991,000 tokens context

GPT Image 2

by OpenAI

GPT-image-2 is OpenAI's latest cutting-edge image generation model. Key value adds…

$5/1M in · $30/1M out

Nemotron 3.5 Content Safety (free)

by Nvidia

Developed by NVIDIA, Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal…

128,000 tokens context

Coding GLM 5.2

by Z.AI

Currently, the special resources for this model are limited, but due to its popularity…

$0.06/1M in · $0.22/1M out

ERNIE 5.1

by Baidu

ERNIE 5.1 is the latest model in the Wenxin series, with comprehensive upgrades to its…

$0.563/1M in · $2.535/1M out
119,000 tokens context

Gemini 3.1 Flash Lite

by Google

gemini-3.1-flash-lite is currently Google's latest and most cost-effective model…

$0.25/1M in · $1.5/1M out
1,000,000 tokens context

gemini-3.1-flash-lite-nothink

by Google

gemini-3.1-flash-lite is currently Google's latest and most cost-effective model…

$0.25/1M in · $1.5/1M out
1,000,000 tokens context

Grok 4.3

by Grok

Grok 4.3 is amongst the leading models in intelligence and well priced when comparing to…

$1.25/1M in · $2.5/1M out
1,000,000 tokens context

Happyhorse 1.0 I2v

by Qwen

HappyHorse-1.0-I2V supports image-to-video generation, featuring highly faithful dynamic…

$2/1M in

Happyhorse 1.0 R2v

by Qwen

HappyHorse-1.0-R2V supports reference-guided video generation, offering more stable…

$2/1M in

Happyhorse 1.0 T2v

by Qwen

HappyHorse-1.0-T2V supports text-to-video generation, featuring highly faithful dynamic…

$2/1M in

Happyhorse 1.0 Video Edit

by Qwen

HappyHorse-1.0-Video-Edit supports video editing, allows editing videos via natural…

$2/1M in

North Mini Code (free)

by Cohere

Developed by Cohere, north-mini-code-free is the debut model of the North family and…

256,000 tokens context

GPT 5.5

by OpenAI

GPT-5.5 raises the baseline for complex production workflows. It’s a strong fit for…

$5/1M in · $30/1M out
1,050,000 tokens context

GPT 5.5 Pro

by OpenAI

Please note: this model is extremely expensive and very slow. If a request fails due to…

$30/1M in · $180/1M out
1,050,000 tokens context

Laguna Xs 2.1 (free)

by Poolside

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from Poolside…

262,144 tokens context

DeepSeek V4 Flash

by DeepSeek

(This model currently points to the older 0423 version; if you need to request the latest…

$0.142/1M in · $0.284/1M out
1,000,000 tokens context

DeepSeek V4 Pro

by DeepSeek

(This model currently points to the older 0423 version; if you need to request the latest…

$1.69/1M in · $3.38/1M out
1,000,000 tokens context

Gemma 4 31B It (free)

by Google

Gemma 4 31B Instruct is a 30.7B dense multimodal model developed by Google DeepMind that…

262,144 tokens context

Command A Plus 05 2026

by Cohere

Cohere's stronger command model for multilingual agents and enterprise workflows

$2.5/1M in · $10/1M out
128,000 tokens context

Doubao Seedream 5.0 Pro

by ByteDance

Seedream-5.0-pro is the latest image-creation model released by ByteDance. The model…

$2/1M in

ERNIE 5.0

by Baidu

ERNIE 5.0 is the next-generation natively multimodal foundation model in the ERNIE…

$0.822/1M in · $3.288/1M out
119,000 tokens context

Kimi K2.6

by Moonshot AI

Kimi K2.6 is Kimi's latest and most intelligent model, with stronger and more stable…

$0.95/1M in · $3.999/1M out
262,144 tokens context

Laguna S 2.1 (free)

by Poolside

Laguna S 2.1 is the latest coding agent model from Poolside, featuring an impressive…

262,144 tokens context

Qwen3.6 Max Preview

by Qwen

The Max model Preview version, the largest and most capable model in the Qwen3.6 series…

$1.268/1M in · $7.608/1M out
240,000 tokens context

Xiaomi Mimo V2.5

by Xiaomi

MiMo-V2.5 is a native, fully multimodal large model designed for agent scenarios; it can…

$0.155/1M in · $0.31/1M out
256,000 tokens context

Xiaomi Mimo V2.5 Pro

by Xiaomi

MiMo-V2.5-Pro is Xiaomi's most powerful model to date. In areas such as general agent…

$0.48/1M in · $0.96/1M out
1,000,000 tokens context

Claude Opus 4.7

by Anthropic

Claude Opus 4.7 is Anthropic’s latest and most powerful publicly available model. It has…

$5/1M in · $25/1M out
200,000 tokens context

Claude Opus 4.7 Thinking

by Anthropic

The claude-opus-4-7-think model has adaptive thinking mode pre-enabled; the default…

$5/1M in · $25/1M out
200,000 tokens context

GPT Chat

by OpenAI

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to…

$5/1M in · $30/1M out
1,050,000 tokens context

Nemotron 3 Nano 30B A3B (free)

by Nvidia

NVIDIA Nemotron 3 Nano 30B A3B is a highly efficient small language Mixture of Experts…

256,000 tokens context

Qwen3.6 27B

by Qwen

The Qwen3.6 series 27B native vision-language Dense model. Compared with the 3.5-27B, the…

$0.422/1M in · $2.532/1M out
254,000 tokens context

Qwen3.6 35B A3B

by Qwen

Qwen 3.6, the native vision-language Plus series model, demonstrates outstanding…

$0.254/1M in · $1.524/1M out
254,000 tokens context

Qwen3.6 Flash

by Qwen

Qwen 3.6, the native vision-language Plus series model, demonstrates outstanding…

$0.169/1M in · $1.014/1M out
991,000 tokens context

Cohere Rerank V4.0 Fast

by Cohere

Rerank 4 is the most advanced set of reranker models available today, purpose-built to…

$0.068/1M in

Cohere Rerank V4.0 Pro

by Cohere

Rerank 4 is the most advanced set of reranker models available today, purpose-built to…

$0.068/1M in

Gemma 4 26B A4B It (free)

by Google

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model developed by…

262,144 tokens context

grok-4-20-non-reasoning

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

Grok 4 20 (reasoning)

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

Qwen Image 2.0

by Qwen

The Qwen-Image-2.0 series accelerated models integrate image generation and image…

$2/1M in

Qwen Image 2.0 Pro

by Qwen

The Qwen-Image-2.0 full-powered models achieve the integration of image generation and…

$2/1M in

Coding MiniMax M3 (free)

by Minimax

coding-minimax-m3-free is a free and open version offered by AIHubMix specifically for…

204,800 tokens context

Doubao Seedance 2.0 260128

by ByteDance

The Doubao large-model team has launched a new-generation professional-grade multimodal…

$2/1M in

Doubao Seedance 2.0 Fast 260128

by Doubao

Seedance 2.0 fast is a next-generation multimodal video-creation model launched by the…

$2/1M in

Doubao Seedance 2.0 Mini 260615

by ByteDance

Seedance 2.0 mini is a next-generation, cost-effective video generation model launched to…

$2/1M in

GLM 5.1

by Z.AI

GLM-5.1 is Zhipu's latest flagship model, with greatly enhanced coding capabilities and…

$0.845/1M in · $3.38/1M out
200,000 tokens context

GLM Image

by Z.AI

GLM-Image is Zhipu AI's new flagship image generation model. The model is trained…

$2/1M in · $2/1M out

Qwen3.6 Plus

by Qwen

Qwen 3.6, the native vision-language Plus series model, demonstrates outstanding…

$0.282/1M in · $1.692/1M out
991,000 tokens context

Wan2.7 I2v

by Qwen

Wanxiang 2.7 — image-to-video: performance capabilities comprehensively upgraded…

$2/1M in

Wan2.7 R2v

by Qwen

Wanxiang 2.7 — reference-driven video generation: more stable references for characters…

$2/1M in

Wan2.7 T2v

by Qwen

Wanxiang 2.7 — text-to-video: performance capabilities comprehensively upgraded…

$2/1M in

Wan2.7 Videoedit

by Qwen

Wanxiang 2.7 — video editing: edit videos using natural-language commands, supporting…

$2/1M in

CC K2.6 Code Preview

by Moonshot AI

for claude code

$0.2/1M in · $0.2/1M out

Gemma 4 26B A4B It

by Google

A Mixture-of-Experts model that activates only 4B parameters per inference,delivering…

$0.14/1M in · $0.4/1M out
262,100 tokens context

Gemma 4 31B It

by Google

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text…

$0.14/1M in · $0.4/1M out
262,100 tokens context

GPT 5.4

by OpenAI

GPT-5.4 is our frontier model for complex professional work.Reasoning.effort supports…

$2.5/1M in · $15/1M out
400,000 tokens context

Wan2.7 Image

by Qwen

Wanxiang 2.7 — image generation and editing: supports text-to-image, text-to-multi-image…

$2/1M in · $2/1M out

Wan2.7 Image Pro

by Qwen

Wanxiang 2.7 — image generation and editing: supports text-to-image, text-to-multi-image…

$2/1M in · $2/1M out

Claude Sonnet 4.6

by Anthropic

Claude Sonnet 4.6 delivers frontier intelligence at scale—built for coding, agents, and…

$3/1M in · $15/1M out
1,000,000 tokens context

Coding Xiaomi Mimo V2.5

by Xiaomi

Only supports OpenAI-compatible formats.

$0.08/1M in · $0.16/1M out

Coding Xiaomi Mimo V2.5 Pro

by Xiaomi

Only supports OpenAI-compatible formats.

$0.2/1M in · $0.4/1M out

Doubao Seed 2.0 Lite 260428

by Doubao

Doubao Coding model optimized for real-world programming environments that can reliably…

$0.09/1M in · $0.542/1M out
256,000 tokens context

Doubao Seed 2.0 Mini 260428

by Doubao

Doubao Coding model optimized for real-world programming environments that can reliably…

$0.028/1M in · $0.282/1M out
256,000 tokens context

Gemini 3.1 Flash Image Preview

by Google

gemini-3.1-flash-image-preview (Nano Banana 2) features professional-grade visual…

$0.5/1M in · $3/1M out

Gemini 3.1 Pro Preview

by Google

Gemini 3.1 Pro Preview is designed to further optimize the performance and reliability of…

$2/1M in · $12/1M out
1,000,000 tokens context

Gemini 3.1 Pro Preview Customtools

by Google

gemini-3.1-pro-preview-customtools For users who build applications mixing bash and…

$2/1M in · $12/1M out
1,000,000 tokens context

Gemini 3.1 Pro Preview Search

by Google

Gemini-3.1-pro-preview-search integrates Google's official search functionality; the…

$2/1M in · $12/1M out

GPT 5.4 Mini

by OpenAI

GPT-5.4 mini is a faster, more efficient model that inherits the advantages of GPT-5.4…

$0.75/1M in · $4.5/1M out
400,000 tokens context

GPT 5.4 Nano

by OpenAI

GPT-5.4 nano is designed for tasks where speed and cost are most important, such as…

$0.2/1M in · $1.25/1M out
400,000 tokens context

GPT 5.5 (free)

by OpenAI

This free model API comes from the OpenAI model deployed on Azure. To prevent abuse, the…

1,050,000 tokens context

Qwen3.5 Plus

by Qwen

The Qwen 3.5 native vision-language Plus model is built on a hybrid architecture that…

$0.11/1M in · $0.658/1M out
991,000 tokens context

Claude Sonnet 4.6 Thinking

by Anthropic

Claude sonnet 4.6 does not enable reasoning mode by default. To access its deep reasoning…

$3/1M in · $15/1M out
200,000 tokens context

Coding Xiaomi Mimo V2 Omni

by Xiaomi

Only supports OpenAI-compatible formats.

$0.08/1M in · $0.4/1M out

Coding Xiaomi Mimo V2 Pro

by Xiaomi

Only supports OpenAI-compatible formats.

$0.2/1M in · $0.6/1M out

GPT 5.3 Chat

by OpenAI

GPT-5.3Chat refers to the GPT-5.3 snapshot currently used in ChatGPT and is optimized for…

$1.75/1M in · $14/1M out
128,000 tokens context

GPT-5.3-Codex

by OpenAI

GPT-5.3-Codex is optimized for agentic coding tasks in Codex or similar environments…

$1.75/1M in · $14/1M out
400,000 tokens context

GPT Image 2 (free)

by OpenAI

This free model API comes from the OpenAI model deployed on Azure. To prevent abuse, the…

Qwen3.5 122B A10B

by Qwen

The Qwen 3.5 native vision-language Plus model is built on a hybrid architecture that…

$0.113/1M in · $0.901/1M out
991,000 tokens context

Qwen3.5 27B

by Qwen

The Qwen 3.5 native vision-language Plus model is built on a hybrid architecture that…

$0.085/1M in · $0.677/1M out
991,000 tokens context

Qwen3.5 35B A3B

by Qwen

The Qwen 3.5 native vision-language Plus model is built on a hybrid architecture that…

$0.056/1M in · $0.451/1M out
991,000 tokens context

Qwen3.5 397B A17B

by Qwen

The Qwen 3.5 native vision-language Plus model is built on a hybrid architecture that…

$0.164/1M in · $0.986/1M out
991,000 tokens context

Qwen3.5 Flash

by Qwen

The Qwen3.5 native vision-language Flash series models are designed with a hybrid…

$0.028/1M in · $0.282/1M out
991,000 tokens context

Coding GLM 5.1

by Z.AI

Only supports OpenAI-compatible formats.

$0.06/1M in · $0.22/1M out

Doubao Seed 2.0 Pro

by Doubao

Doubao flagship all-purpose general model, targeting complex reasoning and long-chain…

$0.482/1M in · $2.411/1M out
256,000 tokens context

GPT 5.4 High

by OpenAI

GPT-5.4 supports configurable reasoning effort only through the /responses endpoint. To…

$2.5/1M in · $15/1M out
400,000 tokens context

GPT 5.4 Low

by OpenAI

GPT-5.4 supports configuring reasoning strength only through the /responses endpoint. To…

$2.5/1M in · $15/1M out
400,000 tokens context

GPT 5.4 Pro

by OpenAI

Please note: this model is extremely expensive and very slow. If a request fails due to…

$30/1M in · $180/1M out
1,050,000 tokens context

Qwen3 Coder Next

by Qwen

The Qwen3 series is a next-generation code-generation model with results close to…

$0.137/1M in · $0.548/1M out
2,000,000 tokens context

Xiaomi Mimo V2 Omni (free)

by Xiaomi

xiaomi-mimo-v2-omni-free is the open free version of xiaomi-mimo-v2-omni. To maintain…

256,000 tokens context

Xiaomi Mimo V2 Pro (free)

by Xiaomi

xiaomi-mimo-v2-pro-free is the open free version of xiaomi-mimo-v2-pro. To maintain…

256,000 tokens context

Xiaomi Mimo V2.5 (free)

by Xiaomi

xiaomi-mimo-v2.5-free is the open free version of xiaomi-mimo-v2.5. To maintain reliable…

256,000 tokens context

Xiaomi Mimo V2.5 Pro (free)

by Xiaomi

xiaomi-mimo-v2.5-pro-free is the open free version of xiaomi-mimo-v2.5-pro5. To maintain…

256,000 tokens context

Claude Opus 4.6

by Anthropic

Claude Opus 4.6 is Anthropic’s latest state-of-the-art reasoning model. It features an…

$5/1M in · $25/1M out
200,000 tokens context

Coding GLM 5.1 (free)

by Z.AI

coding-glm-5.1-free is the open and free version of coding-glm-5.1. To maintain reliable…

Coding MiniMax M2.7 (free)

by Minimax

coding-minimax-m2.7-free is a free and open version offered by AIHubMix specifically for…

204,800 tokens context

GLM 5

by Z.AI

GLM-5 is an advanced, open-source large language model designed for developers tackling…

$0.88/1M in · $2.816/1M out
202,752 tokens context

GLM 5 Vision Turbo

by Z.AI

GLM-5V-Turbo is Zhipu's first multimodal coding foundation model, built for visual…

$0.704/1M in · $3.098/1M out
200,000 tokens context

MiniMax M2.7

by Minimax

MiniMax M2.7 can autonomously build complex Agent Harnesses and, leveraging capabilities…

$0.296/1M in · $1.183/1M out
200,000 tokens context

Claude Opus 4.6 Thinking

by Anthropic

Claude Opus 4.6 does not enable reasoning mode by default. To access its deep reasoning…

$5/1M in · $25/1M out
200,000 tokens context

Coding GLM 5 (free)

by Z.AI

coding-glm-5-free is the open and free version of coding-glm-5. To ensure stable service…

Coding GLM 5 Turbo (free)

by Z.AI

coding-glm-5-turbo-free is the open and free version of coding-glm-5-turbo. To ensure…

Coding MiniMax M2.5 (free)

by Minimax

coding-minimax-m2.5-free is a free and open version offered by AIHubMix specifically for…

204,800 tokens context

Doubao Seed 2.0 Code Preview

by Doubao

The Doubao 2.0 series is a coding model optimized for real programming environments…

$0.482/1M in · $2.411/1M out
256,000 tokens context

Doubao Seed 2.0 Lite 260215

by Doubao

Doubao Coding model optimized for real-world programming environments that can reliably…

$0.09/1M in · $0.542/1M out
256,000 tokens context

Doubao Seed 2.0 Mini

by Doubao

Doubao 2.0 series is designed for low-latency, high-concurrency, and cost-sensitive…

$0.03/1M in · $0.301/1M out
256,000 tokens context

Gemini 3 Flash Preview

by Google

gemini-3-flash-preview is Google's latest released, most balanced model, excelling in…

$0.5/1M in · $3/1M out
1,048,576 tokens context

Gemini 3 Flash Preview Search

by Google

Gemini-3-flash-preview-search integrates Google's official search functionality; the…

$0.5/1M in · $3/1M out
1,048,576 tokens context

GLM 5 Turbo

by Z.AI

GLM-5-Turbo is a foundational model deeply optimized for the OpenClaw scenario. From the…

$1.2/1M in · $4/1M out
202,752 tokens context

CC GLM 5.1

by Z.AI

Supports Claude native interface, can be directly requested in Claude Code.

$0.06/1M in · $0.22/1M out

Claude Opus 4.5

by Anthropic

Claude Opus 4.5 is Anthropic’s latest frontier reasoning model, optimized for complex…

$5/1M in · $25/1M out
200,000 tokens context

Claude Opus 4.5 Thinking

by Anthropic

Claude Opus 4.5 does not enable reasoning mode by default. To access its deep reasoning…

$5/1M in · $25/1M out
200,000 tokens context

Embed V 4.0

by Cohere

Cohere’s Embed 4 is a multilingual multimodal embedding model. It is capable of…

$0.12/1M in
128,000 tokens context

ERNIE Image Turbo

by Baidu

The Ernie-image-Turbo model is an 8-step distilled version of the Ernie-image model, also…

$2/1M in

Gemini 3.1 Flash Image Preview (free)

by Google

This model is the free trial version of gemini-3.1-flash-image-preview (officially…

MiMo V2 Omni

by Xiaomi

MiMo-V2-Omni is designed for complex real-world multimodal interaction and execution…

$0.44/1M in · $2.2/1M out
256,000 tokens context

MiMo V2 Pro

by Xiaomi

Xiaomi MiMo-V2-Pro is built for high-intensity agent work scenarios in the real world. It…

$1.1/1M in · $3.3/1M out
1,000,000 tokens context

Cohere Command A

by Cohere

Command A is Cohere most performant model to date, excelling at tool use, agents…

$2.5/1M in · $10/1M out

Gemini 3 Flash Preview (free)

by Google

gemini-3-flash-preview-free is the free, publicly available version of…

1,048,576 tokens context

CC MiniMax M3

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

Coding MiniMax M3

by Minimax
$0.2/1M in · $0.2/1M out
204,800 tokens context

GPT 4.1 (free)

by OpenAI

This free model API comes from the OpenAI model deployed on Azure. To prevent abuse, the…

1,047,576 tokens context

GPT 4.1 Mini (free)

by OpenAI

This free model API comes from the OpenAI model deployed on Azure. To prevent abuse, the…

1,047,576 tokens context

GPT 4.1 Nano (free)

by OpenAI

This free model API comes from the OpenAI model deployed on Azure. To prevent abuse, the…

1,047,576 tokens context

GPT 4o (free)

by OpenAI

This free model API comes from the OpenAI model deployed on Azure. To prevent abuse, the…

1,047,576 tokens context

Coding GLM 5

by Z.AI

Only supports OpenAI-compatible formats.

$0.06/1M in · $0.22/1M out

Coding GLM 5 Turbo

by Z.AI

Only supports OpenAI-compatible formats.

$0.06/1M in · $0.22/1M out

GLM 4.7

by Z.AI

GLM-4.7 is Zhiyuan's latest flagship model. GLM-4.7 enhances coding capabilities…

$0.274/1M in · $1.096/1M out
200,000 tokens context

Veo 3.1 Lite Generate Preview

by Google

Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p…

$2/1M in

GLM 4.7 Flash (free)

by Z.AI

The glm-4.7-flash free model has usage restrictions to ensure stable service operation: a…

Coding GLM 4.7 (free)

by Z.AI

coding-glm-4.7-free is the open and free version of coding-glm-4.7. To ensure stable…

Doubao Seedance 1.5 Pro 251215

by Doubao

The Doubao video generation model Seedance 1.5 Pro, as a world-leading video generation…

$2/1M in

Doubao Seedance 1.0 Pro 250528

by Doubao

Seedance 1.0 Pro is a foundational video-generation model that supports multi-shot…

$2/1M in

Doubao Seedance 1.0 Pro Fast 251015

by Doubao

Seedance 1.0 Pro Fast is a comprehensive model that delivers rock-bottom prices and peak…

$2/1M in

Gemini 3 Pro Image Preview

by Google

Gemini-3-Pro-Image-Preview (Nano Banana Pro) is a high-performance image generation and…

$2/1M in · $12/1M out

Gemini Embedding 2

by Google

Google's first multimodal embedding model .

$0.2/1M in

Deepinfra Gemma 4 26B A4B It

by Google

A Mixture-of-Experts model that activates only 4B parameters per inference,delivering…

$0.088/1M in · $0.385/1M out
262,100 tokens context

GPT-5.2-Codex

by OpenAI

GPT-5.2-Codex is an upgraded version of GPT-5.2, optimized for agentic coding tasks in…

$1.75/1M in · $14/1M out
400,000 tokens context

Doubao Seedream 5.0 Lite

by Doubao

Doubao-Seedream-5.0-lite is the latest image-creation model released by ByteDance. For…

$2/1M in

GPT Image 1.5

by OpenAI

GPT Image 1.5 is a new image generation model powered by OpenAI’s flagship visual…

$5/1M in · $10/1M out

GPT 5.2

by OpenAI

GPT-5.2 is an advanced general-purpose model that improves on GPT-5.1 with more reliable…

$1.75/1M in · $14/1M out
400,000 tokens context

GPT 5.2 Chat

by OpenAI

GPT-5.2Chat refers to the GPT-5.2 snapshot currently used in ChatGPT and is optimized for…

$1.75/1M in · $14/1M out
128,000 tokens context

GPT 5.2 High

by OpenAI

GPT-5.2 supports configurable reasoning effort only through the /responses endpoint. To…

$1.75/1M in · $14/1M out
400,000 tokens context

GPT 5.2 Low

by OpenAI

GPT-5.2 supports configuring reasoning strength only through the /responses endpoint. To…

$1.75/1M in · $14/1M out
400,000 tokens context

GPT 5.2 Pro

by OpenAI

GPT-5.2 pro is available in the Responses API only to enable support for multi-turn model…

$21/1M in · $168/1M out
400,000 tokens context

GPT 5.1

by OpenAI

GPT-5 is OpenAI’s most advanced language model, designed for complex tasks that require…

$1.25/1M in · $10/1M out
400,000 tokens context

GPT-5.1-Codex Max

by OpenAI

GPT-5.1-Codex-Max is a frontier programming model built for the agent-driven era. Powered…

$1.25/1M in · $10/1M out
400,000 tokens context

Doubao Seed 1.8

by Doubao

Doubao's strongest multimodal Agent model Seed1.8 has powerful multimodal capabilities…

$0.11/1M in · $0.274/1M out
256,000 tokens context

GPT 5.1 Chat

by OpenAI

GPT-5.1 Chat refers to the GPT-5.1 snapshot currently used in ChatGPT and is optimized…

$1.25/1M in · $10/1M out
128,000 tokens context

GPT-5.1-Codex

by OpenAI

GPT-5.1-Codex is a version of GPT-5 optimized for agentic coding tasks in Codex or…

$1.25/1M in · $10/1M out
400,000 tokens context

GPT-5.1-Codex Mini

by OpenAI

GPT-5.1 Codex mini is a smaller, more cost-effective, less-capable version of…

$0.25/1M in · $2/1M out
400,000 tokens context

Claude Haiku 4.5

by Anthropic

Claude Haiku 4.5 is a fast, affordable, and highly capable AI model, excelling at coding…

$1.1/1M in · $5.5/1M out
204,800 tokens context

Claude Sonnet 4.5

by Anthropic

Sonnet 4.5 is the best model in the world for agents, coding, and computer usage. It is…

$3.3/1M in · $16.5/1M out
1,000,000 tokens context

Claude Sonnet 4.5 Thinking

by Anthropic

Claude Sonnet 4.5 does not enable reasoning mode by default. To access its deep reasoning…

$3.3/1M in · $16.5/1M out
1,000,000 tokens context

Grok 4.20 Multi Agent 0309

by Grok

Grok 4.20 is our newest flagship model with industry-leading speed and agentic tool…

$2/1M in · $6/1M out
2,000,000 tokens context

Mistral Large 3

by Mistral

Mistral Large 3 is a MoE model with 67.5B total parameters and 41B active parameters…

$0.5/1M in · $1.5/1M out
256,000 tokens context

CC GLM 5

by Z.AI

Supports Claude native interface, can be directly requested in Claude Code.

$0.06/1M in · $0.22/1M out

CC GLM 5 Turbo

by Z.AI

Supports Claude native interface, can be directly requested in Claude Code.

$0.06/1M in · $0.22/1M out

Cloudflare Glm 5.2

by Z.AI
$1.4/1M in · $4.4/1M out

Gemini 2.5 Flash Image

by Google

Gemini 2.5 Flash Image (Nano-Banana) is a state-of-the-art image generation and editing…

$0.3/1M in · $2.499/1M out
32,800 tokens context

grok-4-1-fast-non-reasoning

by Grok

Grok 4.1 is a new conversational model with significant improvements in real-world…

$0.2/1M in · $0.5/1M out
2,000,000 tokens context

Grok 4.1 Fast (reasoning)

by Grok

Grok 4.1 is a new conversational model with significant improvements in real-world…

$0.2/1M in · $0.5/1M out
2,000,000 tokens context

Grok Code Fast 1

by Grok

Grok 4.1 is a new conversational model with significant improvements in real-world…

$0.2/1M in · $0.5/1M out
256,000 tokens context

K2.6 Code Preview (free)

by Moonshot AI

kimi-for-coding-free is a free and open version offered by AIHubMix specifically for Kimi…

256,000 tokens context

MiMo V2 Flash

by Xiaomi

MiMo-V2-Flash is a mixture of experts (MoE) language model with a total of 309 billion…

$0.192/1M in · $0.575/1M out

Musesteamer Air Image

by Baidu

musesteamer-air-image is a text-to-image model developed by the Baidu Search team aimed…

$2/1M in

Qwen3.6 Plus Preview (free)

by Qwen

This model has been removed from the platform.

1,000,000 tokens context

Zai Glm 5 Turbo

by Z.AI
$1.2/1M in · $4/1M out

GPT 5

by OpenAI

GPT-5 is OpenAI’s most advanced general-purpose model, delivering major improvements in…

$1.25/1M in · $10/1M out
400,000 tokens context

DeepSeek V3.2

by DeepSeek

DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…

$0.302/1M in · $0.453/1M out
128,000 tokens context

DeepSeek V3.2 Thinking

by DeepSeek

DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…

$0.302/1M in · $0.453/1M out
128,000 tokens context

GPT-5-Codex

by OpenAI

GPT-5-Codex is a version of GPT-5 optimized for autonomous coding tasks in Codex or…

$1.25/1M in · $10/1M out
400,000 tokens context

DeepSeek V3.1 Terminus

by DeepSeek

DeepSeek-V3.1 non-thinking mode has now been updated to the DeepSeek-V3.1-Terminus…

$0.56/1M in · $1.68/1M out
160,000 tokens context

DeepSeek V3.1 Thinking

by DeepSeek

Thinking mode of DeepSeek-V3.1; DeepSeek V3.1 is a text generation model provided by…

$0.56/1M in · $1.68/1M out
128,000 tokens context

GPT 5 Pro

by OpenAI

GPT-5 pro uses more compute to think harder and provide consistently better…

$15/1M in · $120/1M out
400,000 tokens context

GPT 5 Mini

by OpenAI

GPT-5 mini is a faster, more cost-efficient version of GPT-5. It's great for well-defined…

$0.25/1M in · $2/1M out
400,000 tokens context

GPT 5 Nano

by OpenAI

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, designed specifically…

$0.05/1M in · $0.4/1M out
400,000 tokens context

GPT 5 Chat

by OpenAI

GPT-5 Chat points to the GPT-5 snapshot currently used in ChatGPT. GPT-5 is our…

$1.25/1M in · $10/1M out
400,000 tokens context

Claude Opus 4.1

by Anthropic

Opus 4.1 is an upgraded version of Claude Opus 4, with improvements mainly in agent…

$16.5/1M in · $82.5/1M out
200,000 tokens context

O3 Deep Research

by OpenAI

Only supported through requests to the v1/responses interface. o3-deep-research is OpenAI…

$10/1M in · $40/1M out

Kimi K2.5

by Moonshot AI

Kimi K2.5 is the smartest model of Kimi to date, achieving open-source state-of-the-art…

$0.6/1M in · $3/1M out
256,000 tokens context

Qwen3 Max 2026 01-23

by Qwen

The snapshot version of the Tongyi Qianwen 3 series Max model is from January 23, 2026…

$0.451/1M in · $1.803/1M out
252,000 tokens context

Qwen3 VL Flash

by Qwen

The Qwen3 series of compact visual-understanding models achieves an effective fusion of…

$0.021/1M in · $0.206/1M out
254,000 tokens context

Qwen3 VL Flash 2026 01-22

by Qwen

The Qwen3 series of compact visual-understanding models achieves an effective fusion of…

$0.021/1M in · $0.206/1M out
254,000 tokens context

Qwen3 VL Plus

by Qwen

The Qwen3 series visual understanding model achieves an effective fusion of thinking and…

$0.137/1M in · $1.37/1M out
256,000 tokens context

CC MiniMax M2.7

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

CC MiniMax M2.7 Highspeed

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

MiniMax M2.5

by Minimax

The MiniMax M2.5 is a flagship programming model built for real-world productivity. As a…

$0.288/1M in · $1.152/1M out
204,800 tokens context

MiniMax M2.5 Highspeed

by Minimax

• Same performance as minimax-m2.5 • Significantly faster inference

$0.288/1M in · $1.152/1M out
204,800 tokens context

Mm Minimax M2.7 Highspeed

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

Coding MiniMax M2.7

by Minimax
$0.2/1M in · $0.2/1M out
204,800 tokens context

Coding MiniMax M2.7 Highspeed

by Minimax
$0.2/1M in · $0.2/1M out
204,800 tokens context

CC MiniMax M2.5

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

CC MiniMax M2.5 Highspeed

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

Coding MiniMax M2.5

by Minimax
$0.2/1M in · $0.2/1M out
204,800 tokens context

Coding MiniMax M2.5 Highspeed

by Minimax
$0.2/1M in · $0.2/1M out
204,800 tokens context

Doubao Seedream 4.5

by Doubao

Seedream 4.5 is ByteDance's latest multimodal image model, integrating capabilities such…

$2/1M in

Sora 2

by OpenAI

Sora-2 is the next-generation text-to-video model evolved from Sora, optimized for higher…

$2/1M in · $2/1M out

Sora 2 Pro

by OpenAI

OpenAI video model Sora2-pro official API.

$2/1M in · $2/1M out

CC GLM 4.7

by Z.AI

Supports Claude native interface, can be directly requested in Claude Code.

$0.06/1M in · $0.22/1M out

CC MiniMax M2.1

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

Coding GLM 4.7

by Z.AI

Only supports OpenAI-compatible formats.

$0.06/1M in · $0.22/1M out

Coding MiniMax M2.1

by Minimax
$0.2/1M in · $0.2/1M out
204,800 tokens context

Coding MiniMax M2.1 (free)

by Minimax

coding-minimax-m2.1-free is a free and open version offered by AIHubMix specifically for…

204,800 tokens context

GPT 4o Audio Preview

by OpenAI

OpenAI voice input and output model, with prices consistent with the official ones. For…

$2.5/1M in · $10/1M out
128,000 tokens context

GPT 4o Mini Audio Preview

by OpenAI

openai声音输入输出模型,价格和官方一致,暂时只展示文字部分价格,声音价格见openai官网;后台扣费和官方一致

$0.15/1M in · $0.6/1M out

MiniMax M2.1

by Minimax

MiniMax-M2.1 redefines efficiency for intelligent agents. It is a compact, fast, and…

$0.288/1M in · $1.152/1M out
204,800 tokens context

O3

by OpenAI

OpenAI o3 is a powerful model across multiple domains, setting a new standard for coding…

$2/1M in · $8/1M out
200,000 tokens context

Wan2.6 I2v

by Qwen

Wan 2.6 - Text-to-Video generation features intelligent storyboard scheduling supporting…

$2/1M in

Wan2.6 T2v

by Qwen

Wan 2.6 - Text-to-Video generation features intelligent storyboard scheduling supporting…

$2/1M in

CC GLM 4.6

by Z.AI

for claude code

$0.06/1M in · $0.22/1M out

Coding GLM 4.6

by Z.AI
$0.06/1M in · $0.22/1M out

Coding GLM 4.6 (free)

by Z.AI

coding-glm-4.6-free is the open and free version of coding-glm-4.6. To ensure stable…

200,000 tokens context

Coding MiniMax M2

by Minimax

coding-minimax-m2 is a free and open version offered by AIHubMix specifically for MiniMax…

$0.2/1M in · $0.2/1M out
204,800 tokens context

Coding MiniMax M2 (free)

by Minimax

coding-minimax-m2-free is a free and open version offered by AIHubMix specifically for…

204,800 tokens context

Flux 2 Flex

by Flux

FLUX.2 is purpose-built for real-world creative production workflows. It delivers…

$2/1M in

Flux 2 Pro

by Flux

FLUX.2 is purpose-built for real-world creative production workflows. It delivers…

$2/1M in

Gemini 2.5 Pro

by Google

Gemini 2.5 Pro is an advanced reasoning model developed by Google, optimized for solving…

$1.25/1M in · $10/1M out
1,048,576 tokens context

GLM 4.6

by Z.AI

GLM-4.6 is Zhipu’s latest flagship model (total parameters 355B, activation parameters…

$0.274/1M in · $1.096/1M out
204,800 tokens context

GLM 4.6 Vision

by Z.AI

Zhipu's latest visual reasoning model achieves state-of-the-art visual understanding…

$0.137/1M in · $0.411/1M out
128,000 tokens context

GLM Ocr

by Z.AI

GLM-OCR is a lightweight professional OCR model with only 0.9B parameters, yet multiple…

$0.028/1M in · $0.028/1M out
32,000 tokens context

Kimi For Coding (free)

by Moonshot AI

kimi-for-coding-free is a free and open version offered by AIHubMix specifically for Kimi…

256,000 tokens context

O3 Pro

by OpenAI

o3-pro This model only supports Requests API interface requests.The model's thinking time…

$20/1M in · $80/1M out
200,000 tokens context

Qianfan Ocr

by Baidu

Qianfan-OCR-Fast is a multimodal large model specialized for OCR, trained primarily on…

$0.062/1M in · $0.248/1M out
32,000 tokens context

Qianfan Ocr Fast

by Baidu

Qianfan-OCR-Fast is a multimodal large model specialized for OCR, trained primarily on…

$0.664/1M in · $2.738/1M out
32,000 tokens context

Step 3.5 Flash

by StepFun

step-3.5-flash is stepfun's flagship inference model, designed for high-complexity tasks…

$0.11/1M in · $0.33/1M out
256,000 tokens context

Wan2.2 I2v Plus

by Qwen

The newly upgraded Tongyi Wanxiang 2.2 text-to-video offers higher video quality. It…

$2/1M in

Wan2.5 I2v Preview

by Qwen

Tongyi Wanxiang 2.5 - Text-to-Video Preview features a newly upgraded technical…

$2/1M in

Wan2.5 T2v Preview

by Qwen

Tongyi Wanxiang 2.5 - Text-to-Video Preview, newly upgraded model architecture, supports…

$2/1M in

Gemini 2.5 Pro Search

by Google

gemini-2.5-pro-search integrates Google's official search functionality; the search…

$1.25/1M in · $10/1M out
1,048,576 tokens context

Kimi K2 Thinking

by Moonshot AI

Kimi K2 Thinking is Moonshot AI's most advanced open-source inference model to date…

$0.548/1M in · $2.192/1M out
262,144 tokens context

Gemini 2.5 Flash

by Google

Gemini 2.5 Flash is Google’s best model in terms of both performance and cost efficiency…

$0.3/1M in · $2.499/1M out
1,048,576 tokens context

Gemini 2.5 Flash Preview 09 2025

by Google

This latest 2.5 Flash model comes with improvements in two key areas we heard consistent…

$0.3/1M in · $2.499/1M out
1,048,576 tokens context

GLM 4.5 Vision

by Z.AI

GLM-4.5V is a vision-language foundational model designed for multimodal agent…

$0.274/1M in · $0.822/1M out
64,000 tokens context

Gemini 2.5 Flash Lite

by Google

Gemini 2.5 Flash-Lite is a balanced model from Google, optimized for applications that…

$0.1/1M in · $0.4/1M out
1,048,576 tokens context

gemini-2.5-flash-lite-nothink

by Google

Gemini 2.5 Flash-Lite is a balanced model from Google, optimized for applications that…

$0.1/1M in · $0.4/1M out
1,048,576 tokens context

Gemini 2.5 Flash Lite Preview 09 2025

by Google

gemini-2.5-flash-lite latest preview version

$0.1/1M in · $0.4/1M out
1,048,576 tokens context

gemini-2.5-flash-lite-preview-09-2025-nothink

by Google

gemini-2.5-flash-lite latest preview version

$0.1/1M in · $0.4/1M out
1,048,576 tokens context

gemini-2.5-flash-nothink

by Google

Gemini-2.5-flash defaults to thinking enabled; to disable thinking, request the name…

$0.3/1M in · $2.499/1M out
1,047,576 tokens context

Gemini 2.5 Flash Search

by Google

gemini-2.5-flash-search integrates Google's official search functionality; the search…

$0.3/1M in · $2.499/1M out
1,048,576 tokens context

gemini-2.5-flash-preview-05-20-nothink

by Google

Gemini-2.5-flash-preview-05-20 is enabled by default for thinking; to disable it, request…

$0.3/1M in · $2.499/1M out
1,048,576 tokens context

Gemini 2.5 Flash Preview 05-20 Search

by Google

Gemini-2.5 Flash Preview 05-20 Search integrates Google's official search functionality…

$0.3/1M in · $2.499/1M out
1,048,576 tokens context

DeepSeek V3 Fast

by DeepSeek

V3 Ultra-Fast Version,The current price is a limited-time 50% discount and will return to…

$0.56/1M in · $2.24/1M out
32,000 tokens context

Imagen 4.0

by Google

Imagen 4 is a high-quality text-to-image model developed by Google, designed for strong…

$2/1M in · $2/1M out

Imagen 4.0 Fast Generate 001

by Google

Imagen 4 is a new-generation image generation model designed to balance high-quality…

$2/1M in · $2/1M out

Imagen 4.0 Generate 001

by Google

Imagen 4 is a new-generation image generation model designed to balance high-quality…

$2/1M in · $2/1M out

Imagen 4.0 Ultra Generate 001

by Google

Imagen-4.0 最新正式版本

$2/1M in · $2/1M out

Imagen 4.0 Ultra

by Google

Imagen-4.0 Ultra 最新预览版

$2/1M in · $2/1M out

GPT Image 1

by OpenAI

Azure OpenAI’s gpt-image-1 image generation API offers both text-to-image generation and…

$5/1M in · $40/1M out

GPT Image 1 Mini

by OpenAI

OpenAI image generation model gpt-image-1-mini Before use, please run pip install -U…

$5/1M in · $40/1M out

O4 Mini

by OpenAI

o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to…

$1.1/1M in · $4.4/1M out
200,000 tokens context

DeepSeek-OCR

by DeepSeek

DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical…

$0.02/1M in · $0.02/1M out
8,000 tokens context

Alicloud Kimi K2 Instruct

by Moonshot AI

Kimi-K2 is a MoE architecture foundational model with extremely powerful coding and agent…

$0.548/1M in · $2.192/1M out

DeepSeek Ocr

by DeepSeek

DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical…

$0.02/1M in · $0.02/1M out
8,000 tokens context

ERNIE 5.0 Thinking Exp

by Baidu

ERNIE 5.0 is the next-generation natively multimodal foundation model in the ERNIE…

$0.822/1M in · $3.288/1M out
119,000 tokens context

Flux Kontext Max

by Flux
$2/1M in

Gemini 2.5 Flash Image Preview

by Google

Aihubmix supports the gemini-2.5-flash-image-preview model; you can add extra parameters…

$0.3/1M in · $1.2/1M out
32,800 tokens context

GLM 4.5

by ChatGLM

GLM-4.5

$0.4/1M in · $1.6/1M out
131,072 tokens context

GPT 4.1

by OpenAI

The latest flagship multimodal model supports million-token context, with encoding…

$2/1M in · $8/1M out
1,047,576 tokens context

Grok 4

by Grok

Grok, their latest and greatest flagship model, offers unparalleled performance in…

$3.3/1M in · $16.5/1M out
256,000 tokens context

grok-4-fast-non-reasoning

by Grok

Grok-4-fast is a cost-effective inference model developed by xAI that delivers…

$0.2/1M in · $0.5/1M out
2,000,000 tokens context

Grok 4 Fast (reasoning)

by Grok

Grok-4-fast is a cost-effective inference model developed by xAI that delivers…

$0.2/1M in · $0.5/1M out
2,000,000 tokens context

Kimi K2 0711

by Moonshot AI

Kimi-K2 is a MoE architecture foundational model with extremely powerful coding and agent…

$0.54/1M in · $2.16/1M out
131,000 tokens context

Kimi K2 Instruct

by Moonshot AI

Kimi-K2 is a MoE architecture foundational model with extremely powerful coding and agent…

$0.54/1M in · $2.16/1M out

Kimi K2 Turbo Preview

by Moonshot AI

The kimi-k2-turbo-preview model is a high-speed version of kimi-k2, with the same model…

$1.2/1M in · $4.8/1M out
262,144 tokens context

Paddleocr VL 0.9b

by Baidu

PaddleOCR-VL is an advanced and efficient document parsing model specifically designed…

$2/1M in

Pp Structurev3

by Baidu

PP-StructureV3 is an efficient and comprehensive document parsing solution that can…

$2/1M in

Qwen3 VL 235B A22B Instruct

by Qwen

The Qwen3 series open-source models include hybrid models, thinking models, and…

$0.274/1M in · $1.096/1M out
131,000 tokens context

Qwen3 VL 235B A22B Thinking

by Qwen

The Qwen3 series open-source models include hybrid models, thinking models, and…

$0.274/1M in · $2.74/1M out
131,000 tokens context

Qwen3 VL 30B A3B Instruct

by Qwen

The Qwen3-VL series’ second-largest MoE model Instruct version offers fast response speed…

$0.103/1M in · $0.411/1M out
128,000 tokens context

Qwen3 VL 30B A3B Thinking

by Qwen

The Qwen3-VL series’ second-largest MoE model Thinking version offers fast response…

$0.103/1M in · $1.028/1M out
128,000 tokens context

Veo 3.0 Generate Preview

by Google

Veo 3.0 Generate Preview is an advanced AI video generation model that supports…

$2/1M in · $2/1M out

Veo 3.1 Fast Generate Preview

by Google

Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p…

$2/1M in

Veo 3.1 Generate Preview

by Google

Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p…

$2/1M in · $2/1M out

Aihubmix Router

by OpenAI

New model routing capability; request aihubmix-router to automatically route models based…

$0.4/1M in · $1.6/1M out

GPT 4.1 Mini

by OpenAI

Lightweight, high-performance model with million-token context and near-flagship-level…

$0.4/1M in · $1.6/1M out
1,047,576 tokens context

GPT 4.1 Nano

by OpenAI

Ultra-lightweight model with million-token context, optimized for speed and low latency…

$0.1/1M in · $0.4/1M out
1,047,576 tokens context

Gemini 2.5 Pro Preview 05-06

by Google

gemini-2.5-pro latest model

$1.25/1M in · $10/1M out
1,048,576 tokens context

Gemini 2.5 Pro Preview 03-25

by Google

Supports high concurrency. The Gemini 2.5 Pro preview version is here, with higher…

$1.25/1M in · $10/1M out

Gemini 2.5 Pro Preview 05-06 Search

by Google

Integrated with Google's official search function.

$1.25/1M in · $10/1M out

Gemini 2.5 Pro Preview 03-25 Search

by Google

Integrated with Google's official search function.

$1.25/1M in · $10/1M out

Qwen3 Max Preview

by Qwen

Qwen3-Max-Preview is the latest preview model in the Qwen3 series. This version is…

$0.846/1M in · $3.384/1M out

Qwen3 Max

by Qwen

The Tongyi Qianwen 3 series Max model has undergone special upgrades in intelligent agent…

$0.451/1M in · $1.803/1M out
262,144 tokens context

Qwen3 Next 80B A3B Instruct

by Qwen

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned model in the Qwen3-Next series…

$0.138/1M in · $0.552/1M out
256,000 tokens context

Qwen3 Next 80B A3B Thinking

by Qwen

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that…

$0.142/1M in · $1.42/1M out
256,000 tokens context

Qwen3 235B A22B Instruct 2507

by Qwen

Qwen3-235B-A22B-Instruct-2507

$0.28/1M in · $1.12/1M out
262,144 tokens context

Qwen3 235B A22B Thinking 2507

by Qwen

The open-source thinking model based on Qwen3 has significantly improved in logical…

$0.28/1M in · $2.8/1M out
262,144 tokens context

Qwen3 Coder 30B A3B Instruct

by Qwen

The code generation model based on Qwen3 has powerful Coding Agent capabilities…

$0.2/1M in · $0.8/1M out
2,000,000 tokens context

Qwen3 Coder 480B A35B Instruct

by Qwen

The code generation model based on Qwen3 has powerful Coding Agent capabilities…

$0.82/1M in · $3.28/1M out
262,000 tokens context

DeepSeek V3

by DeepSeek

It has been automatically upgraded to the latest released version, 250324. Automatically…

$0.272/1M in · $1.088/1M out
1,638,000 tokens context

LongCat-Flash-Chat

by Meituan

Meituan has officially released and open-sourced LongCat-Flash-Chat, which utilizes an…

$0.14/1M in · $0.7/1M out

Gemini 2.5 Pro Preview 06-05 Search

by Google

Integrated with Google's official search function.

$1.25/1M in · $10/1M out

Jina Embeddings V5 Text Nano

by Jina AI

A 3.8-billion-parameter general vector model (embedding model) for state-of-the-art…

$0.05/1M in

Jina Embeddings V5 Text Small

by Jina AI

A 3.8-billion-parameter general vector (embedding) model providing state-of-the-art…

$0.05/1M in

Qwen3 235B A22B

by Qwen

Qwen3-235B-A22B is a massive 235B parameter Mixture-of-Experts (MoE) model that operates…

$0.28/1M in · $1.12/1M out
131,100 tokens context

Qwen3 Coder Flash

by Qwen

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3…

$0.136/1M in · $0.544/1M out
256,000 tokens context

Qwen3 Coder Plus

by Qwen

The code generation model based on Qwen3 has powerful Coding Agent capabilities, excels…

$0.54/1M in · $2.16/1M out
1,048,576 tokens context

Qwen3 Coder Plus 2025 07-22

by Qwen

The code generation model based on Qwen3 has powerful Coding Agent capabilities, excels…

$0.54/1M in · $2.16/1M out
128,000 tokens context

Qwen2.5 VL 72B Instruct

by Qwen

The model provider is the Sophon platform. Qwen2.5-VL-72B-Instruct is the latest…

$0.62/1M in · $0.62/1M out

ERNIE 5.0 Thinking Preview

by Baidu

The new generation Wenxin model, Wenxin 5.0, is a native full-modal large model that…

$0.822/1M in · $3.288/1M out
183,000 tokens context

inclusionAI/Ling-1T

by InclusionAI

Ling-1T is the first flagship non-thinking model in the “Ling 2.0” series, featuring 1…

$0.548/1M in · $2.192/1M out

inclusionAI/Ring-1T

by InclusionAI

Ring-1T is an open-source idea model with a trillion parameters released by the Bailing…

$0.548/1M in · $2.192/1M out

Bce Reranker Base

by Qwen

Based on the dense foundational model of the Qwen3 series, it is specifically designed…

$0.068/1M in

Codex Mini

by OpenAI

Only supports v1/responses API calls.https://docs.aihubmix.com/en/api/Responses-API codex-…

$1.5/1M in · $6/1M out

Doubao Seedream 4.0

by Doubao

Seedream 4.0 is a SOTA-level multimodal image creation model based on leading…

$2/1M in

Embedding V1

by Baidu

Embedding-V1 is a text representation model based on Baidu's Wenxin large model…

$0.068/1M in

ERNIE 4.5 Turbo

by Baidu

Wenxin 4.5 Turbo also has significant improvements in hallucination reduction, logical…

$0.11/1M in · $0.44/1M out
135,000 tokens context

GLM 4.5 X

by ChatGLM

GLM-4.5-X is the high-speed version of GLM-4.5, offering powerful performance with a…

$2.2/1M in · $8.91/1M out

Gme Qwen2 VL 2B Instruct

by Qwen

The GME-Qwen2VL series is a unified multimodal Embedding model trained based on the…

$0.138/1M in · $0.138/1M out

Gte Rerank V2

by Qwen

gte-rerank-v2 is a multilingual unified text ranking model developed by Tongyi Lab…

$0.11/1M in · $0.11/1M out

inclusionAI/Ling-flash-2.0

by InclusionAI

Ling-flash-2.0 is a language model from inclusionAI with a total of 100 billion…

$0.136/1M in · $0.544/1M out

inclusionAI/Ling-mini-2.0

by InclusionAI

Ling-mini-2.0 is a small-sized, high-performance large language model based on the MoE…

$0.068/1M in · $0.272/1M out

inclusionAI/Ring-flash-2.0

by InclusionAI

Ring-flash-2.0 is a high-performance thinking model deeply optimized based on the…

$0.136/1M in · $0.544/1M out

Jina Deepsearch V1

by Jina AI

DeepSearch combines search, reading, and reasoning capabilities to pursue the best…

$0.05/1M in · $0.05/1M out
1,000,000 tokens context

Jina Embeddings V4

by Jina AI

A general-purpose vector model with 3.8 billion parameters, used for multimodal and…

$0.05/1M in · $0.05/1M out

Jina Reranker V3

by Jina AI

Multimodal multilingual document reranker, 131K context, 0.6B parameters, for visual…

$0.05/1M in · $0.05/1M out
131,000 tokens context

Llama 4 Maverick

by Llama

Llama 4 Maverick is a high-capacity Mixture-of-Experts (MoE) model from Meta, featuring…

$0.2/1M in · $0.2/1M out
1,048,576 tokens context

Llama 4 Scout

by Llama

Llama 4 Scout is a highly efficient Mixture-of-Experts (MoE) model from Meta, activating…

$0.2/1M in · $0.2/1M out
131,000 tokens context

Qwen Image

by Qwen

Qwen-Image is a foundational image generation model in the Qwen series, achieving…

$2/1M in

Qwen Image Edit

by Qwen

Qwen-Image-Edit is the image editing version of Qwen-Image. Based on the 20B Qwen-Image…

$2/1M in

Qwen Image Max

by Qwen

Qwen-Image-Edit is the image editing version of Qwen-Image. Based on the 20B Qwen-Image…

$2/1M in

Qwen Mt Plus

by Qwen

Based on the comprehensive upgrade of Qwen3, this flagship translation large model…

$0.492/1M in · $1.476/1M out
16,000 tokens context

Qwen Mt Turbo

by Qwen

Based on the comprehensive upgrade of Qwen3, this flagship translation large model…

$0.192/1M in · $0.535/1M out
16,000 tokens context

Qwen3 Embedding 0.6b

by Qwen

The Qwen3 Embedding model series is the latest proprietary model family from Qwen…

$0.068/1M in

Qwen3 Embedding 4B

by Qwen

The Qwen3 Embedding model series is the latest proprietary model family from Qwen…

$0.068/1M in · $0.068/1M out

Qwen3 Embedding 8B

by Qwen

The Qwen3 Embedding model series is the latest proprietary model family from Qwen…

$0.068/1M in

Qwen3 Reranker 0.6b

by Qwen

Based on the dense foundational model of the Qwen3 series, it is specifically designed…

$0.11/1M in · $0.11/1M out
16,000 tokens context

Qwen3 Reranker 4B

by Qwen

Based on the dense foundational model of the Qwen3 series, it is specifically designed…

$0.11/1M in · $0.11/1M out

Qwen3 Reranker 8B

by Qwen

Based on the dense foundational model of the Qwen3 series, it is specifically designed…

$0.11/1M in · $0.11/1M out

Tao 8K

by Baidu

tao-8k是由Huggingface开发者amu研发并开源的长文本向量表示模型,支持8k上下文长度,模型效果在C-MTEB上居前列,是当前最优的中文长文本embeddings模型…

$0.068/1M in · $0.068/1M out

Jina Clip V2

by Jina AI

Multi-modal Embeddings Model, multilingual, 1024-dimensional, 865M parameters.

$0.05/1M in · $0.05/1M out

Jina Reranker M0

by Jina AI

Multimodal multilingual document reranker, 10K context, 2.4B parameters, for visual…

$0.05/1M in · $0.05/1M out

Jina Colbert V2

by Jina AI

Multi-language ColBERT embeddings model, 560M parameters, used for embedding and…

$0.05/1M in · $0.05/1M out

DeepSeek R1

by DeepSeek

DeepSeek R1 is a new open-source model with performance on par with OpenAI's o1 and…

$0.4/1M in · $2/1M out
1,638,000 tokens context

GPT 4o Search Preview

by OpenAI

Using the Chat Completions API, you can directly access the fine-tuned models and tool…

$2.5/1M in · $10/1M out
128,000 tokens context

GPT 4o Mini Search Preview

by OpenAI

Using the Chat Completions API, you can directly access the fine-tuned models and tool…

$0.15/1M in · $0.6/1M out
128,000 tokens context

Jina Embeddings V3

by Jina AI

Text Embeddings Model, multilingual, 1024-dimensional, 570M parameters.

$0.05/1M in · $0.05/1M out

Claude 3.7 Sonnet

by Anthropic

Support for the thinking parameter through the original Claude SDK.

$3.3/1M in · $16.5/1M out
200,000 tokens context

ERNIE 4.5

by Baidu

Wenxin Large Model 4.5 is a next-generation native multimodal foundational model…

$0.068/1M in · $0.272/1M out
160,000 tokens context

ERNIE 4.5 Turbo VL

by Baidu

The new version of the Wenxin Yiyan large model significantly improves capabilities in…

$0.4/1M in · $1.2/1M out
139,000 tokens context

MiMo V2 Flash (free)

by Xiaomi

MiMo-V2-Flash is an open-source foundation language model developed by Xiaomi. It adopts…

256,000 tokens context

FLUX-1.1-pro

by Flux

FLUX-1.1-pro is an AI image generation tool for professional creators and content…

$2/1M in · $2/1M out

O3 Mini

by OpenAI

OpenAI's latest fast inference model excels at STEAM tasks and offers exceptional…

$1.1/1M in · $4.4/1M out
200,000 tokens context

Doubao Seed 1.6

by Doubao

Doubao-Seed-1.6 is a brand new multimodal deep reasoning model that supports four types…

$0.18/1M in · $1.8/1M out
256,000 tokens context

Doubao Seed 1.6 Flash

by Doubao

Doubao-Seed-1.6-flash is an extremely fast multimodal deep thinking model, with TPOT…

$0.044/1M in · $0.44/1M out
256,000 tokens context

Doubao Seed 1.6 Lite

by Doubao

Doubao-Seed-1.6-lite is a brand new multimodal deep reasoning model that supports…

$0.082/1M in · $0.656/1M out
256,000 tokens context

Doubao Seed 1.6 Thinking

by Doubao

The Doubao-Seed-1.6-thinking model has significantly enhanced reasoning capabilities…

$0.18/1M in · $1.8/1M out
256,000 tokens context

Qwen3 30B A3B Instruct 2507

by Qwen

Significantly improved performance on reasoning tasks, including logical reasoning…

$0.103/1M in · $0.411/1M out

Qwen3 30B A3B Thinking 2507

by Qwen

Significantly improved performance on reasoning tasks, including logical reasoning…

$0.12/1M in · $1.2/1M out

Qwen2 VL 72B Instruct

by Qwen

The model provider is the Sophnet platform. Qwen2-VL-72B-Instruct is the latest iteration…

$2.18/1M in · $6.54/1M out

Qwen2 VL 7B Instruct

by Qwen

The model provider is the Sophnet platform. Qwen2-VL-7B-Instruct is the latest…

$0.28/1M in · $0.7/1M out

CC Kimi For Coding

by Moonshot AI

for claude code

$0.2/1M in · $0.2/1M out

Gemini Embedding 001

by Google

Latest version

$0.15/1M in · $0.15/1M out

gpt-oss-120b

by OpenAI

gpt-oss-120b is a 117B-parameter open-weight Mixture-of-Experts (MoE) language model from…

$0.18/1M in · $0.9/1M out
131,072 tokens context

Qwen 3 235B A22B Thinking 2507

by Qwen

cerebras

$0.28/1M in · $2.8/1M out

Qwen/Qwen3-30B-A3B

by Qwen

Provided by chutes.ai

$1/1M in · $1/1M out

Qwen/Qwen3-32B

by Qwen
$0.4/1M in · $0.8/1M out

Qwen3 32B

by Qwen
$0.16/1M in · $0.64/1M out

Qwen/Qwen3-14B

by Qwen

Provided by chutes.ai

$0.5/1M in · $0.5/1M out

Qwen/Qwen3-8B

by Qwen

Provided by chutes.ai

$0.2/1M in · $0.2/1M out

Embedding 2

by 智谱 ChatGLM

A text vector model that converts input text information into vector representations so…

$0.069/1M in · $0.069/1M out
8,000 tokens context

Embedding 3

by 智谱 ChatGLM

A text vector model that converts input text into vector representations to work with a…

$0.069/1M in · $0.069/1M out
8,000 tokens context

Gemini 2.5 Pro Preview 06-05

by Google

Google’s latest multimodal flagship model, combining exceptional coding and reasoning…

$1.25/1M in · $10/1M out
1,048,576 tokens context

Qwen/Qwen2.5-VL-72B-Instruct

by Qwen

Qwen2.5-VL is a visual language model from the Qwen2.5 series, equipped with strong…

$0.5/1M in · $0.5/1M out

O1

by OpenAI

OpenAI's most powerful O-series model supports official cache hits that halve the input…

$15/1M in · $60/1M out

O1 Pro

by OpenAI

The o1 series of models are trained with reinforcement learning to think before they…

$170/1M in · $680/1M out

ByteDance-Seed/Seed-OSS-36B-Instruct

by Doubao

Seed-OSS is a series of open-source large language models developed by ByteDance's Seed…

$0.2/1M in · $0.534/1M out
256,000 tokens context

Doubao Seed 1.6 250615

by Doubao

Doubao-Seed-1.6 is a brand new multimodal deep reasoning model that supports four types…

$0.18/1M in · $2.52/1M out

Doubao Seed 1.6 Flash 250615

by Doubao

Doubao-Seed-1.6-flash is an extremely fast multimodal deep thinking model, with TPOT…

$0.044/1M in · $0.44/1M out

Doubao Seed 1.6 Thinking 250615

by Doubao

The Doubao-Seed-1.6-thinking model has significantly enhanced reasoning capabilities…

$0.18/1M in · $2.52/1M out

Doubao Seed 1.6 Vision 250815

by Doubao

Doubao-Seed-1.6-vision is a visual deep-thinking model that demonstrates stronger general…

$0.11/1M in · $1.096/1M out

Doubao 1.5 Thinking Pro

by Doubao

Doubao-1.5 is a brand-new deep thinking model that excels in specialized fields such as…

$0.62/1M in · $2.48/1M out

CC MiniMax M2

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

deepseek-ai/DeepSeek-Prover-V2-671B

by DeepSeek

Provided by chutes.ai DeepSeek Prover V2 is a 671B parameter model, speculated to be…

$0.1/1M in · $0.1/1M out

Gemini 2.5 Flash Preview Tts

by Google

Gemini 2.5 Flash Preview TTS is a lightweight, low-latency text-to-speech model designed…

$0.5/1M in · $0.5/1M out

Gemini 2.5 Pro Preview Tts

by Google

Gemini 2.5 Pro Preview TTS is a high-fidelity text-to-speech model designed for premium…

$1/1M in · $1/1M out

Gemma 3 12B It

by Google

Gemma 3 models are multimodal, handling text and image input and generating text output…

$0.2/1M in · $0.2/1M out

Gemma 3 27B It

by Google

Gemma 3 models are multimodal, handling text and image input and generating text output…

$0.2/1M in · $0.2/1M out

Gemma 3 4B It

by Google

Gemma 3 models are multimodal, handling text and image input and generating text output…

$0.2/1M in · $0.2/1M out

Gemma 3n E4B It

by Google

Gemma 3n is a generative AI model optimized for use in everyday devices, such as phones…

$0.2/1M in · $0.2/1M out

Gemma 3 1B It

by Google

Gemma 3 models are multimodal, handling text and image input and generating text output…

$0.2/1M in · $0.2/1M out

DeepSeek R1 Distill Llama 70B

by DeepSeek

Provided by Groq, the DeepSeek-R1-Distill model is fine-tuned based on an open-source…

$0.8/1M in · $1.6/1M out

GPT 4o Mini Tts

by OpenAI

OpenAI’s latest TTS model, gpt-4o-mini-tts, uses the same API endpoint (/v1/audio/speech)…

$0.6/1M in · $12/1M out

tngtech/DeepSeek-R1T-Chimera

by DeepSeek

Provided by chutes.ai DeepSeek-R1T-Chimera merges DeepSeek-R1’s reasoning strengths with…

$0.02/1M in · $0.02/1M out

Veo 2.0 Generate 001

by Google

Veo 2.0 is an advanced video generation model capable of producing high-quality videos…

$2/1M in · $2/1M out

O1 Preview

by OpenAI

The latest and most powerful inference model from OpenAI; AiHubMix uses both OpenAI and…

$15/1M in · $60/1M out

O1 Mini

by OpenAI

o1-mini is faster and 80% cheaper, and is competitive with o1-preview on coding tasks…

$3/1M in · $12/1M out

GPT 4o 2024 11-20

by OpenAI

The latest version of the GPT-4o model; it is recommended to use this version, as it is…

$2.5/1M in · $10/1M out
128,000 tokens context

GPT 4o

by OpenAI

GPT-4o (“o” stands for “omni”) is a new-generation multimodal model designed for more…

$2.5/1M in · $10/1M out
128,000 tokens context

GPT 4o Mini

by OpenAI

The lightweight version of GPT-4o, which is affordable and fast, suitable for handling…

$0.15/1M in · $0.6/1M out
128,000 tokens context

AiHubmix-mistral-medium

by Mistral

Mistral Medium 3 is a SOTA & versatile model designed for a wide range of tasks…

$0.4/1M in · $2/1M out

ERNIE X1.1 Preview

by Baidu

The Wenxin large model X1.1 has made significant improvements in question answering, tool…

$0.136/1M in · $0.544/1M out
119,000 tokens context

Qwen/QwQ-32B

by Qwen

Silicon-based flow provision

$0.14/1M in · $0.56/1M out

chutesai/Mistral-Small-3.1-24B-Instruct-2503

by Mistral

Mistral's latest open-source small model; provided by chutes.ai.

$0.2/1M in · $0.8/1M out

ERNIE X1.1 Preview

by Baidu

The Wenxin large model X1.1 has made significant improvements in question answering, tool…

$0.136/1M in · $0.544/1M out

MiniMax M2

by Minimax

MiniMax-M2 redefines efficiency for intelligent agents. It is a compact, fast, and…

$0.288/1M in · $1.152/1M out
204,800 tokens context

MiniMaxAI/MiniMax-M1-80k

by Minimax

MiniMax-M1 is an open-source large-scale hybrid attention model with 456B total…

$0.6/1M in · $2.4/1M out

Qwen/Qwen2.5-VL-32B-Instruct

by Qwen

Qwen2.5-VL-32B-Instruct is an advanced multimodal model from the Tongyi Qianwen team that…

$0.24/1M in · $0.24/1M out

baidu/ERNIE-4.5-300B-A47B

by Baidu

ERNIE-4.5-300B-A47B is a large language model developed by Baidu based on a Mixture of…

$0.32/1M in · $1.28/1M out

Bge Large En

by BAAI

bge-large-en, open-sourced by the Beijing Academy of Artificial Intelligence (BAAI), is…

$0.068/1M in · $0.068/1M out

Bge Large Zh

by BAAI

bge-large-zh, open-sourced by the Beijing Academy of Artificial Intelligence (BAAI), is…

$0.068/1M in · $0.068/1M out

Codestral

by Mistral

Mistral has launched a new code model - Codestral 25.01…

$0.4/1M in · $1.2/1M out

ERNIE 4.5 0.3b

by Baidu

Wenxin Large Model 4.5 is a next-generation native multimodal foundational large model…

$0.014/1M in · $0.054/1M out

ERNIE 4.5 Turbo 128K Preview

by Baidu

Wenxin 4.5 Turbo also shows significant enhancements in reducing hallucinations, logical…

$0.108/1M in · $0.432/1M out

ERNIE X1 Turbo

by Baidu

Wenxin Large Model X1 possesses enhanced abilities in understanding, planning…

$0.136/1M in · $0.544/1M out
50,500 tokens context

Kat Dev

by Qwen

KAT-Dev (32B) is an open-source 32B parameter model specifically designed for software…

$0.137/1M in · $0.548/1M out
128,000 tokens context

Llama 3.3 70B

by Llama

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and…

$0.6/1M in · $0.6/1M out
65,536 tokens context

moonshotai/Kimi-Dev-72B

by Moonshot AI

Kimi-Dev-72B is a new generation open-source programming large model that achieved a…

$0.32/1M in · $1.28/1M out

moonshotai/Moonlight-16B-A3B-Instruct

by Moonshot AI

Provided by chutes.ai.

$0.2/1M in · $0.2/1M out

Nvidia Nemotron 3 Super 120B A12B

by Nvidia

An open-source, efficient hybrid Mamba-Transformer MoE model that supports a context…

$0.11/1M in · $0.55/1M out
1,000,000 tokens context

O1 Global

by OpenAI

OpenAI new model

$15/1M in · $60/1M out

Qianfan Qi VL

by Baidu

The Qianfan-QI-VL model is a proprietary image quality inspection and visual…

$0.2/1M in · $0.6/1M out

Qwen2.5 VL 72B Instruct

by Qwen

Strong capability in Chinese domain recognition, comparable to ChatGPT-4.0.

$2.4/1M in · $7.2/1M out

tencent/Hunyuan-A13B-Instruct

by Hunyuan

Hunyuan-A13B-Instruct has 8 billion parameters and can match larger models by activating…

$0.14/1M in · $0.56/1M out

unsloth/gemma-3-27b-it

by Google

Google's latest open-source model; provided by chutes.ai

$0.22/1M in · $0.22/1M out

gemini-exp-1206

by Google

Google's latest experimental model, currently Google's most powerful model.

$1.25/1M in · $5/1M out

GPT 4o Zh

by OpenAI

输入任何语言自动翻译为英文给模型,模型输出内容自动翻译为中文返回;测试阶段不支持高,仅支持文本输入;

$2.5/1M in · $10/1M out

Qwen Qwq 32B

by Qwen
$0.4/1M in · $0.8/1M out

unsloth/gemma-3-12b-it

by Google

Provided by chutes.ai.

$0.2/1M in · $0.8/1M out

Qwen Max 0125

by Qwen

Qwen 2.5-Max latest model

$0.38/1M in · $1.52/1M out

BAAI/bge-large-en-v1.5

by BAAI

BAAI/bge-large-en-v1.5 is a large English text embedding model and part of the BGE (BAAI…

$0.034/1M in · $0.034/1M out

BAAI/bge-large-zh-v1.5

by BAAI

BAAI/bge-large-zh-v1.5 is a large Chinese text embedding model and part of the BGE (BAAI…

$0.034/1M in · $0.034/1M out

BAAI/bge-reranker-v2-m3

by BAAI

BAAI/bge-reranker-v2-m3 is a lightweight multilingual reranking model. It is developed…

$0.034/1M in · $0.034/1M out

tencent/Hunyuan-MT-7B

by Hunyuan

Hunyuan-MT-7B is a lightweight translation model with 7 billion parameters, designed to…

$0.2/1M in · $0.2/1M out

V3

by Ideogram

Fast and high-quality — top image quality in just 11 seconds per piece, with almost no…

$2/1M in · $2/1M out

V_2

by Ideogram

The Ideogram AI drawing interface is now live. This model boasts powerful text-to-image…

$2/1M in · $2/1M out

V_2_TURBO

by Ideogram

The Ideogram AI drawing interface is now live. This model boasts powerful text-to-image…

$2/1M in · $2/1M out

V_2A

by Ideogram

The Ideogram AI drawing interface is now live. This model boasts powerful text-to-image…

$2/1M in · $2/1M out

V_2A_TURBO

by Ideogram

The Ideogram AI drawing interface is now live. This model boasts powerful text-to-image…

$2/1M in · $2/1M out

V_1

by Ideogram

V_1 is a text-to-image model in the Ideogram series. It delivers strong text rendering…

$2/1M in · $2/1M out

V_1_TURBO

by Ideogram

The Ideogram AI drawing interface is now live. This model boasts powerful text-to-image…

$2/1M in · $2/1M out

Doubao Embedding Large Text 240915

by Doubao

doubao-embedding-large-text-240915 Doubao Embedding is a semantic vectorization model…

$0.1/1M in · $0.1/1M out

Kimi Thinking Preview

by Moonshot AI

The latest kimi model.

$30/1M in · $30/1M out

GPT 4o 2024 08-06

by OpenAI

Supports caching, with automatic halving of charges upon a cache hit.

$2.5/1M in · $10/1M out

Qwen Plus 2025 07-28

by Qwen

The Tongyi Qianwen series balanced capability model has inference performance and speed…

$0.113/1M in · $1.126/1M out

Qwen Plus

by Qwen

The Qwen series models with balanced capabilities have inference performance and speed…

$0.113/1M in · $1.126/1M out

Sonar

by Perplexity

Latest Perplexity Model

$1.6/1M in · $1.6/1M out

stepfun-ai/step3

by StepFun

Step3 is a multimodal reasoning model released by StepFun. It uses a Mixture‑of‑Experts…

$1.1/1M in · $2.75/1M out

Text Embedding V4

by Qwen

This is the Tongyi Laboratory's multilingual unified text vector model trained based on…

$0.08/1M in · $0.08/1M out

Aihubmix Phi 4 Mini (reasoning)

by Microsoft

Phi-4-mini-reasoning is a lightweight open model designed for advanced mathematical…

$0.12/1M in · $0.12/1M out
128,000 tokens context

Qwen Turbo

by Qwen

The Qwen series model with the fastest speed and lowest cost, suitable for simple tasks…

$0.046/1M in · $0.092/1M out

Aihub Phi 4 Multimodal Instruct

by Microsoft

Microsoft's latest model

$0.12/1M in · $0.48/1M out
128,000 tokens context

Qwen3 30B A3B

by Qwen

Achieves effective integration of thinking and non-thinking modes, allowing mode…

$0.12/1M in · $1.2/1M out

Aihub Phi 4 Mini Instruct

by Microsoft

Microsoft's latest model

$0.12/1M in · $0.48/1M out
128,000 tokens context

Grok 3

by Grok

Grok's latest model

$3/1M in · $15/1M out

Aihub Phi 4

by Microsoft

Phi-4 is a state-of-the-art open model based on a combination of synthetic datasets…

$0.12/1M in · $0.48/1M out
16,400 tokens context

Claude 3 Opus 20240229

by Anthropic

Claude’s previous generation strongest model

$16.5/1M in · $82.5/1M out

Dall E 3

by OpenAI

dall-e-3 is an AI image generation model that converts natural language prompts into…

$40/1M in · $40/1M out

Doubao Embedding Text 240715

by Doubao

doubao-embedding-text-240715 Doubao Embedding is a semantic vectorization model developed…

$0.7/1M in · $0.7/1M out

Grok 3 Beta

by Grok

Grok's latest model This model ID with beta has been officially taken offline. Using this…

$3/1M in · $15/1M out

Qwen3 14B

by Qwen

Achieves effective integration of thinking and non-thinking modes, enabling mode…

$0.16/1M in · $1.6/1M out

Grok 3 Fast

by Grok
$5.5/1M in · $27.5/1M out

Qwen3 8B

by Qwen

Achieves effective integration of thinking and non-thinking modes, enabling mode…

$0.08/1M in · $0.8/1M out

deepseek-ai/DeepSeek-R1-Zero

by DeepSeek

Openly deployed by chutes.ai; inference with FP8; zero is the initial preliminary version…

$2.2/1M in · $2.2/1M out

Grok 3 Fast Beta

by Grok
$5.5/1M in · $27.5/1M out

Grok 3 Mini

by Grok
$0.3/1M in · $0.501/1M out

Qwen3 4B

by Qwen

Achieves effective integration of thinking and non-thinking modes, allowing mode…

$0.046/1M in · $0.46/1M out

Grok 3 Mini Beta

by Grok

This model ID with beta has been officially taken offline. Using this model…

$0.33/1M in · $0.551/1M out

Qwen3 1.7b

by Qwen

Effectively integrates thinking and non-thinking modes, allowing mode switching during…

$0.046/1M in · $0.46/1M out

Qwen3 0.6b

by Qwen

Effectively integrates thinking and non-thinking modes, allowing mode switching during…

$0.046/1M in · $0.46/1M out

Alicloud Glm 5

by Z.AI

GLM-5 is an advanced, open-source large language model designed for developers tackling…

$0.563/1M in · $2.535/1M out

Command A 03 2025

by Cohere

Command A is Cohere most performant model to date, excelling at tool use, agents…

$2.5/1M in · $10/1M out

Grok 3 Mini Fast Beta

by Grok
$0.33/1M in · $2.2/1M out

Qwen 3 32B

by Qwen

cerebras

$0.4/1M in · $1.6/1M out

Qwen Turbo 2025 04-28

by Qwen

The Qwen3 series Turbo model effectively integrates thinking and non-thinking modes…

$0.046/1M in · $0.092/1M out

Qwen Plus 2025 04-28

by Qwen

The Qwen3 series Plus model effectively integrates thinking and non-thinking modes…

$0.113/1M in · $1.126/1M out

THUDM/GLM-Z1-32B-0414

by ChatGLM

GLM-Z1-32B-0414 is a reasoning-focused AI model built on GLM-4-32B-0414. It has been…

$0.08/1M in · $0.08/1M out

THUDM/GLM-4.1V-9B-Thinking

by ChatGLM

GLM-4.1V-9B-Thinking is an open-source Vision Language Model (VLM) jointly released by…

$0.1/1M in · $0.1/1M out

Text Embedding 004

by Google
$0.02/1M in · $0.02/1M out

THUDM/GLM-4-32B-0414

by ChatGLM

GLM-4-32B-0414 is a next-generation open-source model with 32 billion parameters…

$0.08/1M in · $0.08/1M out

THUDM/GLM-Z1-9B-0414

by ChatGLM

GLM-Z1-9B-0414 is a small but powerful model in the GLM series, with only 9 billion…

$0.05/1M in · $0.05/1M out

THUDM/GLM-4-9B-0414

by ChatGLM

GLM-4-9B-0414 is a lightweight model in the GLM family, with 9 billion parameters. It…

$0.05/1M in · $0.05/1M out

CC Doubao Seed Code Preview

by Doubao

claude code

$0.2/1M in · $0.2/1M out

Doubao Seed Code Preview

by Doubao

chat

$0.2/1M in · $0.2/1M out

deepseek-ai/Janus-Pro-7B

by DeepSeek

Janus-Pro deepseek最新发布图片生成模型;是一个新颖的自回归框架,统一了多模态理解和生成。它通过将视觉编码解耦为独立路径来解决以往方法的局限性,同时仍然利用单一的统…

$2/1M in · $2/1M out

GLM Zero Preview

by ChatGLM

Simply put, it is the intelligent enhanced version of O1.

$2/1M in · $2/1M out

Qwen 3 235B A22B Instruct 2507

by Qwen

cerebras

$0.28/1M in · $1.4/1M out

Coding GLM 4.5 Air

by ChatGLM
$0.014/1M in · $0.084/1M out

Deepinfra Nvidia Nemotron 3 Nano 30B A3b2

by Nvidia
$0.066/1M in · $0.264/1M out

GLM 4.5 Air

by ChatGLM
$0.14/1M in · $0.84/1M out
131,072 tokens context

GPT 4 32K

by OpenAI

The smartest version of GPT-4; OpenAI no longer offers it officially. All the 32k…

$60/1M in · $120/1M out

Nvidia Llama 3.1 Nemotron 70B Instruct

by Nvidia
$1.32/1M in · $1.32/1M out

Nvidia Llama 3.3 Nemotron Super 49B V1.5

by Nvidia
$0.11/1M in · $0.44/1M out

Nvidia Nemotron 3 Nano 30B A3B

by Nvidia
$0.066/1M in · $0.264/1M out

Nvidia Nemotron Nano 12B V2 VL

by Nvidia
$0.22/1M in · $0.66/1M out

Nvidia Nemotron Nano 9B V2

by Nvidia
$0.044/1M in · $0.176/1M out

O1 Preview 2024 09-12

by OpenAI
$15/1M in · $60/1M out

Qwen/QVQ-72B-Preview

by Qwen
$1.2/1M in · $1.2/1M out

Qwen/QwQ-32B-Preview

by Qwen
$0.16/1M in · $0.16/1M out

Llama 3.1 Sonar Huge 128K Online

by Perplexity

On February 22, 2025, this model will be officially discontinued. The Perplexity AI…

$5.6/1M in · $5.6/1M out

Aihubmix Mistral Large 2411

by Mistral

The latest Mistral Large 2 model is deployed on Azure.

$2/1M in · $6/1M out

Llama 3.1 Sonar Large 128K Online

by Perplexity

On February 22, 2025, this model will be officially discontinued; Perplexity AI's…

$1.2/1M in · $1.2/1M out

Aihubmix Mistral Large 2407

by Mistral
$3/1M in · $9/1M out

Grok 2 1212

by Grok

马斯克的xai的最新模型,aihubmix价格低于官网10%

$1.8/1M in · $9/1M out

Llama 3.1 70B

by Llama
$0.44/1M in · $0.44/1M out

Wan2.6 T2i

by Qwen
$2/1M in

DESCRIBE

by Ideogram

This endpoint is used to describe an image. Supported image formats include JPEG, PNG…

$2/1M in · $2/1M out

UPSCALE

by Ideogram

The super-resolution upscale interface of the Ideogram AI drawing model is designed to…

$2/1M in · $2/1M out

Bai Qwen3 VL 235B A22B Instruct

by Qwen

The Qwen3 series open-source models include hybrid models, thinking models, and…

$0.274/1M in · $1.096/1M out

CC MiniMax M2

by Minimax

For Claude Code only

$0.1/1M in · $0.1/1M out

CC DeepSeek V3

by DeepSeek

For Claude code only

$0.3/1M in · $0.3/1M out

CC DeepSeek V3.1

by DeepSeek

For Claude code only

$0.56/1M in · $1.68/1M out

CC ERNIE 4.5 300B A47B

by Baidu

For Claude code only

$0.32/1M in · $1.28/1M out

CC Kimi Dev 72B

by Moonshot AI

For Claude code only

$0.32/1M in · $1.28/1M out

CC Kimi K2 Instruct

by Moonshot AI

For Claude code only

$1.1/1M in · $3.3/1M out

CC Kimi K2 Instruct 0905

by Moonshot AI

For Claude code only

$1.1/1M in · $3.3/1M out

CC Kimi K2 Thinking

by Moonshot AI

Dedicated for Claude Code

$0.548/1M in · $2.192/1M out

Computer Use Preview

by OpenAI
$3/1M in · $12/1M out

GPT Image Test

by OpenAI
$5/1M in · $40/1M out

grok-4.20-beta-0309-non-reasoning

by Grok

Grok 4.20 Beta is our latest flagship model, offering industry-leading speed and agent…

$2/1M in · $6/1M out
2,000,000 tokens context

Grok 4.20 Beta 0309 (reasoning)

by Grok

Grok 4.20 Beta is our latest flagship model, offering industry-leading speed and agent…

$2/1M in · $6/1M out
2,000,000 tokens context

Grok 4.20 Multi Agent Beta 0309

by Grok

Grok 4.20 Beta is our latest flagship model, offering industry-leading speed and agent…

$2/1M in · $6/1M out
2,000,000 tokens context

Jina Reader

by Jina AI
$0.05/1M in · $0.05/1M out

Jina Search

by Jina AI
$0.05/1M in · $0.05/1M out

Llama3.1 8B

by Llama

cerebras

$0.3/1M in · $0.6/1M out

O1 2024 12-17

by OpenAI
$15/1M in · $60/1M out

Sf Kimi K2 Thinking

by Moonshot AI
$0.548/1M in · $2.192/1M out

Baichuan3 Turbo

by Baichuan
$1.9/1M in · $1.9/1M out

Baichuan3 Turbo 128K

by Baichuan
$3.8/1M in · $3.8/1M out

Baichuan4

by Baichuan
$16/1M in · $16/1M out

Baichuan4 Air

by Baichuan
$0.16/1M in · $0.16/1M out

Baichuan4 Turbo

by Baichuan
$2.4/1M in · $2.4/1M out

DeepSeek V3

by DeepSeek
$0.272/1M in · $1.088/1M out

Doubao 1.5 Lite 32K

by Doubao

Doubao-1.5-lite, a brand-new generation of lightweight model, offers exceptional response…

$0.05/1M in · $0.1/1M out

Doubao 1.5 Pro 256K

by Doubao

Doubao-1.5-pro-256k, a fully upgraded version based on Doubao-1.5-Pro, delivers an…

$0.8/1M in · $1.44/1M out

Doubao 1.5 Pro 32K

by Doubao

Doubao-1.5-pro, a brand-new generation of flagship model, features comprehensive…

$0.134/1M in · $0.335/1M out

Doubao 1.5 Vision Pro 32K

by Doubao

Doubao-1.5-vision-pro is a newly upgraded multimodal large model that supports image…

$0.46/1M in · $1.38/1M out

Doubao Lite 128K

by 字节跳动
$0.14/1M in · $0.28/1M out

Doubao Lite 32K

by 字节跳动
$0.06/1M in · $0.12/1M out

Doubao Lite 4K

by 字节跳动
$0.06/1M in · $0.12/1M out

Doubao Pro 128K

by 字节跳动
$0.8/1M in · $1.44/1M out

Doubao Pro 256K

by 字节跳动
$0.8/1M in · $1.44/1M out

Doubao Pro 32K

by 字节跳动
$0.14/1M in · $0.35/1M out

Doubao Pro 4K

by 字节跳动
$0.14/1M in · $0.35/1M out

GPT-OSS-20B

by OpenAI
$0.11/1M in · $0.55/1M out

Gryphe/MythoMax-L2-13b

by Meta
$0.4/1M in · $0.4/1M out

MiniMax Text 01

by Minimax
$0.14/1M in · $1.12/1M out

Mistral Large 2407

by Mistral
$3/1M in · $9/1M out

Qwen/Qwen2-1.5B-Instruct

by Qwen
$0.2/1M in · $0.2/1M out

Qwen/Qwen2-57B-A14B-Instruct

by Qwen
$0.24/1M in · $0.24/1M out

Qwen/Qwen2-72B-Instruct

by Qwen
$0.8/1M in · $0.8/1M out

Qwen/Qwen2-7B-Instruct

by Qwen
$0.08/1M in · $0.08/1M out

Qwen/Qwen2.5-32B-Instruct

by Qwen
$0.6/1M in · $0.6/1M out

Qwen/Qwen2.5-72B-Instruct

by Qwen
$0.8/1M in · $0.8/1M out

Qwen/Qwen2.5-72B-Instruct-128K

by Qwen
$0.8/1M in · $0.8/1M out

Qwen/Qwen2.5-7B-Instruct

by Qwen
$0.4/1M in · $0.4/1M out

Qwen/Qwen2.5-Coder-32B-Instruct

by Qwen
$0.16/1M in · $0.16/1M out

Qwen3 235B A22B Thinking 2507

by Qwen

算能提供

$0.28/1M in · $2.8/1M out

Stable Diffusion 3.5 Large

by Stable diffusion

Stable Diffusion 3.5 Large, developed by Stability AI, is a text-to-image generation…

$4/1M in · $4/1M out

WizardLM/WizardCoder-Python-34B-V1.0

by Meta
$0.9/1M in · $0.9/1M out

Ahm Phi 3.5 Moe Instruct

by Microsoft

Phi-3.5-MoE 是一个轻量级的最先进开放模型,基于用于 Phi-3 的数据集构建——合成数据和经过筛选的公开可用文档,重点关注高质量、推理密集的数据。该模型支持多语言,并具…

$0.4/1M in · $1.6/1M out

Ahm Phi 3.5 Mini Instruct

by Microsoft

Phi-3.5-mini is a lightweight, state-of-the-art open model built upon the dataset used…

$1/1M in · $3/1M out

Ahm Phi 3.5 Vision Instruct

by Microsoft
$0.4/1M in · $1.6/1M out

Ahm Phi 3 Medium 128K

by Microsoft
$6/1M in · $18/1M out

Ahm Phi 3 Medium 4K

by Microsoft
$1/1M in · $3/1M out

Ahm Phi 3 Small 128K

by Microsoft
$1/1M in · $3/1M out

Aihubmix Codestral 2501

by Mistral

azure部署

$0.4/1M in · $1.2/1M out

Aihubmix Cohere Command R

by Cohere
$0.64/1M in · $1.92/1M out

Aihubmix Jamba 1.5 Large

by AI21
$2.2/1M in · $8.8/1M out

Aihubmix Llama 3.1 405B Instruct

by Meta
$5/1M in · $15/1M out

Aihubmix Llama 3.1 70B Instruct

by Meta
$0.6/1M in · $0.78/1M out

Aihubmix Llama 3.1 8B Instruct

by Meta
$0.3/1M in · $0.6/1M out

Aihubmix Llama 3.2 11B Vision

by Meta
$0.4/1M in · $0.4/1M out

Aihubmix Llama 3.2 90B Vision

by Meta
$2.4/1M in · $2.4/1M out

Aihubmix Llama 3 70B Instruct

by Meta
$0.7/1M in · $0.7/1M out

Aihubmix Mistral Large

by Mistral
$4/1M in · $12/1M out

Aihubmix Command R 08 2024

by Cohere
$0.2/1M in · $0.8/1M out

Aihubmix Command R Plus

by Cohere
$3.84/1M in · $19.2/1M out

Aihubmix Command R Plus 08 2024

by Cohere
$2.8/1M in · $11.2/1M out

Alicloud Deepseek V3.2

by DeepSeek
$0.274/1M in · $0.411/1M out

Alicloud Glm 4.7

by Z.AI
$0.411/1M in · $1.918/1M out

Alicloud Kimi K2 Thinking

by Moonshot AI
$0.548/1M in · $2.192/1M out

Alicloud Kimi K2.5

by Moonshot AI
$0.548/1M in · $2.877/1M out
256,000 tokens context

Alicloud Minimax M2.5

by Minimax
$0.288/1M in · $1.15/1M out

Anthropic Opus 4.6

by Anthropic

Claude Opus 4.6 is Anthropic’s latest state-of-the-art reasoning model. It features an…

$5/1M in · $25/1M out
200,000 tokens context

Azure Deepseek V3.2

by DeepSeek
$0.58/1M in · $1.68/1M out

Azure Deepseek V3.2 Speciale

by DeepSeek
$0.58/1M in · $1.68/1M out

Azure Kimi K2.5

by Moonshot AI
$0.6/1M in · $3/1M out
256,000 tokens context

Cbs Glm 4.7

by Z.AI
$2.25/1M in · $2.75/1M out

Cerebras Llama 3.3 70B

by Llama
$0.6/1M in · $0.6/1M out

Chatglm_lite

by 智谱 ChatGLM
$0.286/1M in · $0.286/1M out

Chatglm_pro

by 智谱 ChatGLM
$1.429/1M in · $1.429/1M out

Chatglm_std

by 智谱 ChatGLM
$0.714/1M in · $0.714/1M out

Chatglm_turbo

by 智谱 ChatGLM
$0.714/1M in · $0.714/1M out

Claude 2

by Anthropic
$8.8/1M in · $8.8/1M out

Claude 2.0

by Anthropic
$8.8/1M in · $39.6/1M out

Claude 2.1

by Anthropic
$8.8/1M in · $39.6/1M out

Claude 3 Haiku 20240229

by Anthropic
$0.275/1M in · $0.275/1M out

Claude 3 Haiku 20240307

by Anthropic
$0.275/1M in · $1.375/1M out

Claude 3 Sonnet 20240229

by Anthropic
$3.3/1M in · $16.5/1M out

Claude Instant 1

by Anthropic
$1.793/1M in · $1.793/1M out

Claude Instant 1.2

by Anthropic
$0.88/1M in · $3.96/1M out

Code Davinci Edit 001

by 智谱 ChatGLM
$20/1M in · $20/1M out

Cogview 3

by 智谱 ChatGLM
$35.5/1M in · $35.5/1M out

Cogview 3 Plus

by 智谱 ChatGLM
$10/1M in · $10/1M out

Command

by Cohere
$1/1M in · $2/1M out

Command Light

by Cohere
$1/1M in · $2/1M out

Command Light Nightly

by Cohere
$1/1M in · $2/1M out

Command Nightly

by Cohere
$1/1M in · $2/1M out

Command R

by Cohere
$0.64/1M in · $1.92/1M out

Command R 08 2024

by Cohere
$0.2/1M in · $0.8/1M out

Command R Plus

by Cohere
$3.84/1M in · $19.2/1M out

Command R Plus 08 2024

by Cohere
$2.8/1M in · $11.2/1M out

Dall E 2

by OpenAI
$16/1M in · $16/1M out

Davinci

by OpenAI
$20/1M in · $20/1M out

Davinci 002

by OpenAI
$2/1M in · $2/1M out

Deepinfra Llama 3.1 8B Instant

by Llama
$0.033/1M in · $0.055/1M out

Deepinfra Llama 3.3 70B Instant Turbo

by Llama
$0.11/1M in · $0.352/1M out

Deepinfra Llama 4 Maverick 17B 128e Instruct

by Llama
$0.33/1M in · $1.32/1M out

Deepinfra Llama 4 Scout 17B 16e Instruct

by Llama
$0.088/1M in · $0.33/1M out

deepseek-ai/DeepSeek-Coder-V2-Instruct

by DeepSeek
$0.16/1M in · $0.32/1M out

deepseek-ai/DeepSeek-R1-Distill-Llama-70B

by DeepSeek

来自siliconflow开源部署,模型本身通过知识蒸馏得到的模型

$0.6/1M in · $0.6/1M out

deepseek-ai/DeepSeek-R1-Distill-Llama-8B

by DeepSeek

来自siliconflow开源部署,模型本身通过知识蒸馏得到的模型

$0.01/1M in · $0.01/1M out

deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B

by DeepSeek

来自siliconflow开源部署,模型本身通过知识蒸馏得到的模型

$0.01/1M in · $0.01/1M out

deepseek-ai/DeepSeek-R1-Distill-Qwen-14B

by DeepSeek

Open source deployment from SiliconFlow, the model itself is obtained through knowledge…

$0.1/1M in · $0.1/1M out

deepseek-ai/DeepSeek-R1-Distill-Qwen-32B

by DeepSeek

Open source deployment from SiliconFlow, the model itself is obtained through knowledge…

$0.2/1M in · $0.2/1M out

deepseek-ai/DeepSeek-R1-Distill-Qwen-7B

by DeepSeek

Open source deployment from SiliconFlow, the model itself is obtained through knowledge…

$0.01/1M in · $0.01/1M out

deepseek-ai/DeepSeek-V2-Chat

by DeepSeek
$0.16/1M in · $0.32/1M out

deepseek-ai/DeepSeek-V2.5

by DeepSeek
$0.16/1M in · $0.32/1M out

deepseek-ai/deepseek-llm-67b-chat

by DeepSeek
$0.16/1M in · $0.16/1M out

deepseek-ai/deepseek-vl2

by DeepSeek
$0.16/1M in · $0.16/1M out

DeepSeek V3

by DeepSeek
$0.272/1M in · $1.088/1M out

Distil Whisper Large V3 En

by OpenAI

Groq开源部署非

$5.556/1M in · $5.556/1M out

Doubao 1.5 Thinking Vision Pro 250428

by Doubao

Deep Thinking Image Understanding Visual Localization Video Understanding Tool…

$2/1M in · $2/1M out

Fx Flux 2 Pro

by Flux
$2/1M in

gemini-2.5-pro-exp-03-25

by Google

Google’s latest experimental model, highly unstable, for experience only. It boasts…

$1.25/1M in · $5/1M out

gemini-embedding-exp-03-07

by Google
$0.02/1M in · $0.02/1M out

gemini-exp-1114

by Google
$1.25/1M in · $5/1M out

gemini-exp-1121

by Google
$1.25/1M in · $5/1M out

Gemini Pro

by Google

gemini最初版不推荐

$0.2/1M in · $0.6/1M out

Gemini Pro Vision

by Google

gemini最初版不推荐

$1/1M in · $1/1M out

Gemma 7B It

by Google
$0.1/1M in · $0.1/1M out

GLM 3 Turbo

by 智谱 ChatGLM
$0.71/1M in · $0.71/1M out

GLM 4

by 智谱 ChatGLM
$14.2/1M in · $14.2/1M out

GLM 4 Flash

by 智谱 ChatGLM
$0.1/1M in · $0.1/1M out

GLM 4 Plus

by 智谱 ChatGLM
$8/1M in · $8/1M out

GLM 4.5 Airx

by ChatGLM

GLM-4.5-AirX is the high-speed version of GLM-4.5-Air, with faster response times…

$1.1/1M in · $4.51/1M out

GLM 4 Vision

by 智谱 ChatGLM
$14.2/1M in · $14.2/1M out

GLM 4 Vision Plus

by 智谱 ChatGLM
$2/1M in · $2/1M out

Google Gemma 3 12B It

by Google
$0.2/1M in · $0.2/1M out

Google Gemma 3 27B It

by Google
$0.2/1M in · $0.2/1M out

Google Gemma 3 4B It

by Google
$0.2/1M in · $0.2/1M out

google/gemini-exp-1114

by Google
$1.25/1M in · $5/1M out

google/gemma-2-27b-it

by Google
$0.8/1M in · $0.8/1M out

google/gemma-2-9b-it:free

by Google
$0.02/1M in · $0.02/1M out

GPT 3.5 Turbo

by OpenAI

Since the GPT-3.5-turbo model has been officially deprecated, all requests targeting this…

$0.5/1M in · $1.5/1M out

GPT 3.5 Turbo 0301

by OpenAI
$1.5/1M in · $1.5/1M out

GPT 3.5 Turbo 0613

by OpenAI
$1.5/1M in · $2/1M out

GPT 3.5 Turbo 1106

by OpenAI
$1/1M in · $2/1M out

GPT 3.5 Turbo 16K

by OpenAI
$3/1M in · $4/1M out

GPT 3.5 Turbo 16K 0613

by OpenAI
$3/1M in · $4/1M out

GPT 3.5 Turbo Instruct

by OpenAI
$1.5/1M in · $2/1M out

GPT 4

by OpenAI
$30/1M in · $60/1M out

GPT 4 0125 Preview

by OpenAI
$10/1M in · $30/1M out

GPT 4 0314

by OpenAI
$30/1M in · $60/1M out

GPT 4 0613

by OpenAI
$30/1M in · $60/1M out

GPT 4 1106 Preview

by OpenAI
$10/1M in · $30/1M out

GPT 4 32K 0314

by OpenAI
$60/1M in · $120/1M out

GPT 4 32K 0613

by OpenAI
$60/1M in · $120/1M out

GPT 4 Turbo

by OpenAI
$10/1M in · $30/1M out

GPT 4 Turbo 2024 04-09

by OpenAI
$10/1M in · $30/1M out

GPT 4 Turbo Preview

by OpenAI
$10/1M in · $30/1M out

GPT 4 Vision Preview

by OpenAI
$10/1M in · $30/1M out

GPT 4o 2024 05-13

by OpenAI
$5/1M in · $15/1M out
128,000 tokens context

GPT 4o Mini 2024 07-18

by OpenAI
$0.15/1M in · $0.6/1M out

gpt-oss-20b

by OpenAI

gpt-oss-20b is a 21-billion parameter open-weight model released by OpenAI under the…

$0.11/1M in · $0.55/1M out
128,000 tokens context

Grok 2 Vision 1212

by Grok

grok-2-vision-1212 is the latest vision model in the Grok family, delivering outstanding…

$1.8/1M in · $9/1M out

Grok Vision Beta

by Grok
$5.6/1M in · $16.8/1M out

Groq Llama 3.1 8B Instant

by Llama
$0.055/1M in · $0.088/1M out

Groq Llama 3.3 70B Versatile

by Llama
$0.649/1M in · $0.869/1M out

Groq Llama 4 Maverick 17B 128e Instruct

by Llama
$0.22/1M in · $0.66/1M out

Groq Llama 4 Scout 17B 16e Instruct

by Llama
$0.122/1M in · $0.366/1M out

Jina Embeddings V2 Base Code

by Jina AI

Model optimized for code and document search, 768-dimensional, 137M parameters.

$0.05/1M in · $0.05/1M out

Learnlm 1.5 Pro Experimental

by Google
$1.25/1M in · $5/1M out

Llama 3.1 405B Instruct

by Meta
$4/1M in · $4/1M out

Llama 3.1 405B (reasoning)

by Meta
$4/1M in · $4/1M out

Llama 3.1 70B Versatile

by Meta
$0.6/1M in · $0.6/1M out

Llama 3.1 8B Instant

by Llama
$0.3/1M in · $0.6/1M out

Llama 3.1 Sonar Small 128K Online

by Perplexity

On February 22, 2025, this model will be officially discontinued. The Perplexity AI…

$0.3/1M in · $0.3/1M out

Llama 3.2 11B Vision Preview

by Meta
$0.2/1M in · $0.2/1M out

Llama 3.2 1B Preview

by Meta
$0.2/1M in · $0.2/1M out

Llama 3.2 3B Preview

by Meta
$0.2/1M in · $0.2/1M out

Llama 3.2 90B Vision Preview

by Meta
$2.4/1M in · $2.4/1M out

Llama2 70B 4096

by Llama
$0.5/1M in · $0.5/1M out

Llama2 70B 40960

by Llama
$0.5/1M in · $0.5/1M out

Llama2 7B 2048

by Meta
$0.1/1M in · $0.1/1M out

Llama3 70B 8192

by Meta
$0.7/1M in · $0.937/1M out

Llama3 8B 8192

by Llama
$0.06/1M in · $0.12/1M out

Llama3 Groq 70B 8192 Tool Use Preview

by Meta
$0.00089/1M in · $0.00089/1M out

Llama3 Groq 8B 8192 Tool Use Preview

by Meta
$0.00019/1M in · $0.00019/1M out

meta-llama/Llama-3.2-90B-Vision-Instruct

by Meta
$0.5/1M in · $0.5/1M out

meta-llama/llama-3.1-405b-instruct:free

by Meta
$0.02/1M in · $0.02/1M out

meta-llama/llama-3.1-70b-instruct:free

by Meta
$0.02/1M in · $0.02/1M out

meta-llama/llama-3.1-8b-instruct:free

by Meta
$0.02/1M in · $0.02/1M out

meta-llama/llama-3.2-11b-vision-instruct:free

by Meta
$0.02/1M in · $0.02/1M out

meta-llama/llama-3.2-3b-instruct:free

by Meta
$0.02/1M in · $0.02/1M out

meta/llama-3.1-405b-instruct

by Meta
$5/1M in · $5/1M out

meta/llama3-8B-chat

by Meta
$0.3/1M in · $0.3/1M out

mistralai/mistral-7b-instruct:free

by Mistral
$0.002/1M in · $0.002/1M out

Moonshot Kimi K2.5

by Moonshot AI
$0.6/1M in · $3/1M out

Moonshot V1 128K

by Moonshot AI
$10/1M in · $10/1M out

Moonshot V1 128K Vision Preview

by Moonshot AI
$10/1M in · $10/1M out

Moonshot V1 32K

by Moonshot AI
$4/1M in · $4/1M out

Moonshot V1 32K Vision Preview

by Moonshot AI
$4/1M in · $4/1M out

Moonshot V1 8K

by Moonshot AI
$2/1M in · $2/1M out

Moonshot V1 8K Vision Preview

by Moonshot AI
$2/1M in · $2/1M out

nvidia/Llama-3_1-Nemotron-Ultra-253B-v1

by Nvidia

Llama-3.1-Nemotron-Ultra-253B is a 253 billion parameter reasoning-focused language model…

$0.5/1M in · $0.5/1M out

O1 Mini 2024 09-12

by OpenAI
$3/1M in · $12/1M out

Omni Moderation

by OpenAI
$0.02/1M in · $0.02/1M out

Qwen Flash

by Qwen

The model adopts tiered pricing.

$0.02/1M in · $0.2/1M out

Qwen Flash 2025 07-28

by Qwen

The model adopts tiered pricing.

$0.02/1M in · $0.2/1M out

Qwen Long

by Qwen
$0.1/1M in · $0.4/1M out

Qwen Max

by Qwen
$0.38/1M in · $1.52/1M out

Qwen Max Longcontext

by Qwen
$7/1M in · $21/1M out

Qwen Plus

by Qwen
$0.113/1M in · $1.126/1M out

Qwen Turbo

by Qwen
$0.046/1M in · $0.092/1M out

Qwen Turbo 2024 11-01

by Qwen
$0.046/1M in · $0.092/1M out

Qwen2.5 14B Instruct

by Qwen
$0.4/1M in · $1.2/1M out

Qwen2.5 32B Instruct

by Qwen
$0.6/1M in · $1.2/1M out

Qwen2.5 3B Instruct

by Qwen
$0.4/1M in · $0.8/1M out

Qwen2.5 72B Instruct

by Qwen
$0.8/1M in · $2.4/1M out

Qwen2.5 7B Instruct

by Qwen
$0.4/1M in · $0.8/1M out

Qwen2.5 Coder 1.5b Instruct

by Qwen
$0.2/1M in · $0.4/1M out

Qwen2.5 Coder 7B Instruct

by Qwen
$0.2/1M in · $0.4/1M out

Qwen2.5 Math 1.5b Instruct

by Qwen
$0.2/1M in · $0.2/1M out

Qwen2.5 Math 72B Instruct

by Qwen
$0.8/1M in · $2.4/1M out

Qwen2.5 Math 7B Instruct

by Qwen
$0.2/1M in · $0.4/1M out

Step 2 16K

by 阶跃星辰
$2/1M in · $2/1M out

Text Ada 001

by OpenAI
$0.4/1M in · $0.4/1M out

Text Babbage 001

by OpenAI
$0.5/1M in · $0.5/1M out

Text Curie 001

by OpenAI
$2/1M in · $2/1M out

Text Davinci 002

by OpenAI
$20/1M in · $20/1M out

Text Davinci 003

by OpenAI
$20/1M in · $20/1M out

Text Davinci Edit 001

by OpenAI
$20/1M in · $20/1M out

Text Embedding 3 Large

by OpenAI
$0.13/1M in · $0.13/1M out

Text Embedding 3 Small

by OpenAI
$0.02/1M in · $0.02/1M out

Text Embedding Ada 002

by OpenAI
$0.1/1M in · $0.1/1M out

Text Embedding V1

by OpenAI
$0.1/1M in · $0.1/1M out

Text Moderation 007

by OpenAI
$0.2/1M in · $0.2/1M out

Text Moderation

by OpenAI
$0.2/1M in · $0.2/1M out

Text Moderation Stable

by OpenAI
$0.2/1M in · $0.2/1M out

Text Search Ada Doc 001

by OpenAI
$20/1M in · $20/1M out

Tts 1

by OpenAI
$15/1M in · $15/1M out

Tts 1 1106

by OpenAI
$15/1M in · $15/1M out

Tts 1 Hd

by OpenAI
$30/1M in · $30/1M out

Tts 1 Hd 1106

by OpenAI
$30/1M in · $30/1M out

Whisper 1

by OpenAI

Ignore the displayed price on the page; the actual charge for this model request is…

$100/1M in · $100/1M out

Whisper Large V3

by OpenAI

Groq开源部署

$30.834/1M in · $30.834/1M out

Whisper Large V3 Turbo

by OpenAI

Groq开源部署

$5.556/1M in · $5.556/1M out

Yi Large

by 零一万物
$3/1M in · $3/1M out

Yi Large Rag

by 零一万物
$4/1M in · $4/1M out

Yi Large Turbo

by 零一万物
$1.8/1M in · $1.8/1M out

Yi Lightning

by 零一万物
$0.2/1M in · $0.2/1M out

Yi Medium

by 零一万物
$0.4/1M in · $0.4/1M out

Yi VL Plus

by 零一万物
$0.00085/1M in · $0.00085/1M out

DeepSeek R1 Distill Qianfan Llama 8B

by DeepSeek
$0.137/1M in · $0.548/1M out

Doubao 1.5 Pro 256K 250115

by 字节跳动豆包
$0.684/1M in · $1.231/1M out

Doubao 1.5 Pro 32K 250115

by 字节跳动豆包
$0.108/1M in · $0.27/1M out

GPT 4o 2024 08-06 Global

by OpenAI
$2.5/1M in · $10/1M out

GPT 4o Mini Global

by OpenAI
$0.15/1M in · $0.6/1M out

Meta Llama 3 70B

by Meta
$4.795/1M in · $4.795/1M out

Meta Llama 3 8B

by Meta
$0.548/1M in · $0.548/1M out

O3 Global

by OpenAI
$2/1M in · $8/1M out

O3 Mini Global

by OpenAI
$1.1/1M in · $4.4/1M out

O3 Pro Global

by OpenAI
$20/1M in · $80/1M out

Qianfan Chinese Llama 2 13B

by Baidu
$0.822/1M in · $0.822/1M out

Qianfan Llama VL 8B

by Baidu
$0.274/1M in · $0.685/1M out