Black Friday Recharge Offer, offer ends on November 30
One API, every model

Explore the CometAPI model catalog

Search and compare text, image, video and audio models from leading providers. Review capabilities and pricing, then integrate with one API key.

256+ modelsUnified billingProduction-ready API

Filters

Input PricingFREE$5+
FREE$0.5$1$5+
256 models available
Filter by capability, provider or price.
https://www

Auto

CometAPI
Popular
https://www
Text Generation

Auto

auto
Text modelPopularauto-router

自动路由:根据请求内容自动选择最合适的底层模型。

From
$75/1M tokens
View model
O

GPT Image 2.5

OpenAI
Popular
O
Image Generation

GPT Image 2.5

gpt-image-2.5-sunburst
Image modelPopularimage-editingtext-to-image

Explore the GPT Image 2.5 API.

From
$5/1M tokens
View model
Claude Fable 5.1
Anthropic
A
Text Generation

Claude Fable 5.1

claude-fable-5-1
Text modelimage-to-textpdf-to-textvideo-to-texttext-to-text

claude-fable-5-1 Next generation intelligence for long-running agents

From
$10/1M tokens
View model
GLM 5.3
Zhipu AI
Popular
Z
AI Model

GLM 5.3

GLM 5.3
AI modelPopulartext-to-text

Explore the GLM 5.3 API.

From
$75/1M tokens
View model
Claude Fable 5
Anthropic
Popular
C
Text Generation

Claude Fable 5

claude-fable-5
Text modelPopulartext-to-textimage-to-textpdf-to-text

Access to Claude Fable 5 has been restored. It brings 5th-generation intelligence to your most ambitious coding and professional work.

From
$10/1M tokens
View model
X

Grok 4.3

xAI
Popular
X
Text Generation

Grok 4.3

grok-4.3
Text modelPopulartext-to-text

Excels at agentic reasoning, knowledge work, and tool use.

From
$1.25/1M tokens
View model
O

Sora 2

OpenAI
Popular
O
Video Generation

Sora 2

sora-2
Video modelPopulartext-to-videoimage-to-video

Super powerful video generation model, with sound effects, supports chat format.

From
$0.1/s
View model
O

GPT 5.5 Pro

OpenAI
Popular
O
Text Generation

GPT 5.5 Pro

gpt-5.5-pro
Text modelPopulartext-to-text

OpenAI's most capable model, engineered for the hardest tasks and long-running agentic workflows. GPT-5.5 Pro excels at complex coding, computer use, deep research, data analysis, and scientific reasoning — delivering frontier-level intelligence at GPT-5.4 latency with greater token efficiency. Ideal for enterprise and professional use cases demanding the highest standard of accuracy and autonomous task execution.

From
$75/1M tokens
View model
D

DeepSeek V4 Flash

DeepSeek
Popular
D
Text Generation

DeepSeek V4 Flash

deepseek-v4-flash
Text modelPopulartext-to-text

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.

From
$75/1M tokens
View model
O

GPT Image 2

OpenAI
Popular
O
Image Generation

GPT Image 2

gpt-image-2
Image modelPopulartext-to-image

OpenAI's most capable image generation model, featuring near-perfect text rendering across multiple languages, up to 4K resolution, and reasoning-powered Thinking Mode. Built for production workflows that demand accuracy, speed, and on-brand visual output.

From
$5/1M tokens
View model
X

Grok Imagine Video 1.5

xAI
Coming soon
X
Video Generation

Grok Imagine Video 1.5

grok-imagine-video-1.5
Video modelComing soonimage-to-video

xAI's latest image-to-video model features native synchronized audio generation — video and sound are produced in a single inference pass. Supports 480p/720p output, clips up to 15 seconds, and ranked #1 on the Image-to-Video Arena leaderboard.

From
Coming soon
View model
O

GPT 5.5

OpenAI
Popular
10% off
O
Text Generation

GPT 5.5

gpt-5.5
Text modelPopulartext-to-text

OpenAI's smartest and most intuitive flagship model, designed for complex coding, agentic workflows, computer use, data analysis, and in-depth research. Delivers frontier-level intelligence with the same low latency as GPT-5.4, while completing tasks with greater token efficiency. The go-to model for demanding professional and enterprise workloads.

From
$83.333333$75/1M tokens
View model
O

GPT Image 2 ALL

OpenAI
Popular
O
Text Generation

GPT Image 2 ALL

gpt-image-2-all
Text modelPopulartext-to-image

GPT Image 2 is openai state-of-the-art image generation model for fast, high-quality image generation and editing. It supports flexible image sizes and high-fidelity image inputs.

From
$0.05/request
View model
GPT-5.2 Chat
OpenAI
Popular
O
Text Generation

GPT-5.2 Chat

gpt-5.2-chat-latest
Text modelPopulartext-to-textimage-to-text128,000 context

gpt-5.2-chat-latest is the Chat-optimized snapshot of OpenAI’s GPT-5.2 family (branded in ChatGPT as GPT-5.2 Instant). It is the model for interactive/chat use cases that need a blend of speed, long-context handling, multimodal inputs and reliable conversational behaviour.

From
$1.75/1M tokens
View model
M

MiniMax-M2.7

MiniMax
Popular
M
Text Generation

MiniMax-M2.7

minimax-m2.7
Text modelPopulartext-to-text

MiniMax-M2.7 offers the same top-tier intelligence as the standard version—including recursive self-evolution and expert-level office productivity—but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

From
$0.3/1M tokens
View model
?

MiMo-V2.5

Test provider Q1
Coming soon
?
Text Generation

MiMo-V2.5

mimo-v2.5
Text modelComing soontext-to-text

MiMo-V2.5 is Xiaomi's native full-modal model. It achieves professional-grade agent performance at about half the cost of inference, while outperforming MiMo-V2-Omni in multimodal perception in image and video understanding tasks.

From
Coming soon
View model
?

MiMo-V2.5-Pro

Test provider Q1
Coming soon
?
Text Generation

MiMo-V2.5-Pro

mimo-v2.5-pro
Text modelComing soontext-to-text

MiMo-V2.5-Pro is Xiaomi's flagship model, excelling in general-purpose agent capabilities and complex software engineering.

From
Coming soon
View model
Q

Qwen3.6-Plus

Aliyun
Popular
Q
Text Generation

Qwen3.6-Plus

qwen3.6-plus
Text modelPopulartext-to-text

Qwen 3.6-Plus is now available, featuring enhanced code development capabilities and improved efficiency in multimodal recognition and inference, making the Vibe Coding experience even better.

From
$0.5/1M tokens
View model
Z

GLM 5.1

Zhipu AI
Popular
Z
Text Generation

GLM 5.1

glm-5.1
Text modelPopulartext-to-text

GLM-5.1 (released April 2026), purpose-built for long-horizon autonomous tasks. Unlike traditional models optimized for short interactions, GLM-5.1 excels at maintaining goal alignment, reducing strategy drift, and delivering production-grade results over extended periods — up to 8 hours of continuous autonomous work on a single complex task. It represents a major leap in agentic engineering, shifting evaluation from single-turn intelligence to real-world sustained execution.

From
$1.4/1M tokens
View model
M

Kimi K2.6

Moonshot AI
Popular
M
Text Generation

Kimi K2.6

kimi-k2.6
Text modelPopulartext-to-text

Kimi K2.6 is Kimi's latest and most intelligent model, possessing stronger and more stable long-term code writing capabilities, significantly improved instruction compliance and self-correction abilities, and supporting text, image, and video input, thinking and non-thinking modes, and dialogue and agent tasks.

From
$0.95/1M tokens
View model
A

Claude Mythos Preview

Anthropic
Coming soon
A
Text Generation

Claude Mythos Preview

Claude Mythos Preview
Text modelComing soontext-to-text

Claude Mythos Preview is our most capable frontier model to date, and shows a striking leap in scores on many evaluation benchmarks compared to our previous frontier model, Claude Opus 4.6.

From
Coming soon
View model
?

mimo-v2-omni

Test provider Q1
Popular
?
Text Generation

mimo-v2-omni

mimo-v2-omni
Text modelPopulartext-to-textimage-to-textvideo-to-textspeech-to-text

MiMo-V2-Omni is a frontier omni-modal model that natively processes image, video, and audio inputs within a unified architecture. It combines strong multimodal perception with agentic capability - visual grounding, multi-step planning, tool use, and code execution - making it well-suited for complex real-world tasks that span modalities. 256K context window.

From
$0.4/1M tokens
View model
X

mimo-v2-pro

Test provider Q1
Popular
X
Text Generation

mimo-v2-pro

mimo-v2-pro
Text modelPopulartext-to-text

MiMo-V2-Pro is Xiaomi's flagship foundation model, featuring over 1T total parameters and a 1M context length, deeply optimized for agentic scenarios. It is highly adaptable to general agent frameworks like OpenClaw. It ranks among the global top tier in the standard PinchBench and ClawBench benchmarks, with perceived performance approaching that of Opus 4.6. MiMo-V2-Pro is designed to serve as the brain of agent systems, orchestrating complex workflows, driving production engineering tasks, and delivering results reliably.

From
$1/1M tokens
View model
Z

GLM 5 Turbo

Zhipu AI
Popular
Z
Text Generation

GLM 5 Turbo

glm-5-turbo
Text modelPopulartext-to-text200k context

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios.

From
$1.2/1M tokens
View model
O

GPT-5.3 Chat

OpenAI
Popular
O
Text Generation

GPT-5.3 Chat

gpt-5.3-chat-latest
Text modelPopulartext-to-textimage-to-text

GPT-5.3 Instant model used in ChatGPT

From
$1.75/1M tokens
View model
O

Sora 2 Pro

OpenAI
Popular
O
Video Generation

Sora 2 Pro

sora-2-pro
Video modelPopulartext-to-videoimage-to-video

Sora 2 Pro is our most advanced and powerful media generation model, capable of generating videos with synchronized Audio. It can create detailed, dynamic video clips from natural language or images.

From
$0.3/s
View model
M

mj_fast_video

Midjourney
Popular
M
Video Generation

mj_fast_video

mj_fast_video
Video modelPopulartext-to-video

Midjourney video generation

From
$1.125/request
View model
M

mj_turbo_imagine

Midjourney
Popular
M
Image Generation

mj_turbo_imagine

mj_turbo_imagine
Image modelPopulartext-to-image

Explore the mj_turbo_imagine API.

From
$0.315/request
View model
M

mj_fast_imagine

Midjourney
Popular
M
Image Generation

mj_fast_imagine

mj_fast_imagine
Image modelPopulartext-to-image

Midjourney drawing

From
$0.105/request
View model
X

Grok Imagine Video

xAI
Popular
X
Video Generation

Grok Imagine Video

grok-imagine-video
Video modelPopularimage-to-videovideo-editingtext-to-video

Generate videos from text prompts, animate still images, or edit existing videos with natural language. The API supports configurable duration, aspect ratio, and resolution for generated videos — with the SDK handling the asynchronous polling automatically.

From
$0.05/s
View model
O

gpt-audio-1.5

OpenAI
Popular
O
Audio Generation

gpt-audio-1.5

gpt-audio-1.5
Audio modelPopulartext-to-speechaudio

The best voice model for audio in, audio out with Chat Completions.

From
$2.5/1M tokens
View model
O

gpt-realtime-1.5

OpenAI
Popular
O
Audio Generation

gpt-realtime-1.5

gpt-realtime-1.5
Audio modelPopulartext-to-speechaudio32,000 context

The best voice model for audio in, audio out.

From
$4/1M tokens
View model