GPT-5.6 Luna price down 80%, Terra down 20% →
One API, every model

Explore the CometAPI model catalog

Search and compare text, image, video and audio models from leading providers. Review capabilities and pricing, then integrate with one API key.

273+ modelsUnified billingProduction-ready API

Filters

Input PricingFREE $5+
FREE$0.5$1$5+
273 models available
Filter by capability, provider or price.
Qwen3.8-Max
Aliyun
Q
Text Generation

Qwen3.8-Max

qwen3.8-max-preview
Text modeltext-to-textimage-to-textpdf-to-textvideo-to-text

Qwen3.8-Max is Alibaba Qwen’s flagship large language model designed for advanced reasoning, agentic workflows, multimodal understanding, and enterprise-scale AI applications.. It has 2.4T parameters, adopts the MoE architecture, supports switching between thinking and fast inference modes, can handle various content formats, and performs excellently in scenarios such as code engineering, professional office work, and complex logical reasoning, second only to Anthropic's Fable 5.

From
$60/1M tokens
View model
Claude Opus 5
Anthropic
Popular
C
Text Generation

Claude Opus 5

claude-opus-5
Text modelPopulartext-to-textimage-to-textpdf-to-text

Claude Opus 5 is available today. It’s a thoughtful and proactive model that comes close to the frontier intelligence of Claude Fable 5 at half the price.

From
$4/1M tokens
View model
Flux 3
Flux
Coming soon
F
Video Generation

Flux 3

flux-3
Video modelComing soontext-to-video

coming soon

From
Coming soon
View model
GPT 5.6
OpenAI
O
Text Generation

GPT 5.6

gpt-5.6
Text modeltext-to-textimage-to-text

Start with GPT-5.6 Sol for complex reasoning and coding, choose GPT-5.6 Terra to balance intelligence and cost, or use GPT-5.6 Luna for cost-sensitive, high-volume workloads.

From
View pricing
View model
Gemini 3.6 Flash
Google
G
Text Generation

Gemini 3.6 Flash

gemini-3.6-flash
Text modeltext-to-text

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at code generation, agentic execution, and spatial reasoning. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

From
$1.2/1M tokens
View model
Nano Banana 2 lite
Google
Popular
G
Image Generation

Nano Banana 2 lite

gemini-3.1-flash-lite-image
Image modelPopulartext-to-image

Gemini 3.1 Flash Lite Image model is an efficiency expert in the image generation family, designed for ultra-low latency and cost-effective image generation and modification.

From
$0.2/1M tokens
View model
Claude Sonnet 5
Anthropic
Popular
C
Text Generation

Claude Sonnet 5

claude-sonnet-5
Text modelPopulartext-to-textpdf-to-textimage-to-text1M context

Claude Sonnet 5 API is live on CometAPI at $1.6 per million input tokens and $8 per million output tokens, 20 percent below Anthropic list price. One key gives you Claude Sonnet 5 plus 500+ models from OpenAI, Google, and ByteDance under all in one pricing and a single invoice. No seat fees. No monthly minimum. Pay only for what you use. Start free and make your first call in under five minutes.

From
$1.6/1M tokens
View model
Seedance-2-5
Bytedance
Coming soon
B
Video Generation

Seedance-2-5

doubao-seedance-2-5
Video modelComing soontext-to-videoimage-to-videovideo-editing

coming soon

From
Coming soon
View model
Happy Horse 1.1
Aliyun
Popular
Q
Video Generation

Happy Horse 1.1

happyhorse-1.1
Video modelPopulartext-to-videoimage-to-video

HappyHorse 1.1 is a multimodal video-generation model designed for professional content creation, advertising, short films, social media production, and storytelling. It extends the capabilities of HappyHorse 1.0—which gained significant attention after ranking highly in independent video-generation evaluations—with stronger scene coherence and improved visual fidelity.

From
$0.112/s
View model
Claude Fable 5
Anthropic
Popular
C
Text Generation

Claude Fable 5

claude-fable-5
Text modelPopulartext-to-textpdf-to-textimage-to-text

Access to Claude Fable 5 has been restored. It brings 5th-generation intelligence to your most ambitious coding and professional work.

From
$8/1M tokens
View model
GPT Image 2
OpenAI
Popular
O
Image Generation

GPT Image 2

gpt-image-2
Image modelPopulartext-to-image

GPT Image 2 is openai state-of-the-art image generation model for fast, high-quality image generation and editing. It supports flexible image sizes and high-fidelity image inputs.

From
$4/1M tokens
View model
Seedance 2-0
Bytedance
Popular
B
Video Generation

Seedance 2-0

doubao-seedance-2-0
Video modelPopulartext-to-videoimage-to-video

Seedance 2.0 is ByteDance’s next-generation multimodal video foundation model focused on cinematic, multi-shot narrative video generation. Unlike single-shot text-to-video demos, Seedance 2.0 emphasizes reference-based control (images, short clips, audio), coherent character/style consistency across shots, and native audio/video synchronization — aiming to make AI video useful for professional creative and previsualization workflows.

From
$0.056/s
View model
Claude Opus 4.8
Anthropic
Popular
C
Text Generation

Claude Opus 4.8

claude-opus-4-8
Text modelPopulartext-to-textimage-to-textpdf-to-text200K tokens context

Claude Opus 4.8 is a premium AI model designed for advanced reasoning, deep analysis, and high-quality content generation. It excels at handling complex instructions, long-context understanding, and sophisticated problem-solving across professional and technical domains.

From
$4/1M tokens
View model
Gemini 3.5 Flash
Google
Popular
G
Text Generation

Gemini 3.5 Flash

gemini-3.5-flash
Text modelPopulartext-to-textimage-to-textvideo-to-textspeech-to-textpdf-to-text

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

From
$1.2/1M tokens
View model
Gemini 3.1 Pro
Google
G
Text Generation

Gemini 3.1 Pro

gemini-3.1-pro-preview
Text modeltext-to-textimage-to-textvideo-to-textspeech-to-textpdf-to-text

Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Google’s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories

From
$1.6/1M tokens
View model
Kimi K3
Moonshot AI
Popular
M
Text Generation

Kimi K3

kimi-k3
Text modelPopulartext-to-text1,000k tokens context

Kimi K3 is Kimi's flagship model, designed for long-range programming and end-to-end knowledge work, featuring 1M token context and leading-edge comprehensive intelligence.

From
$2.4/1M tokens
View model
Kimi K2.7 Code
Moonshot AI
M
Text Generation

Kimi K2.7 Code

kimi-k2.7-code
Text modeltext-to-textimage-to-textvideo-to-text

Kimi K2.7 Code is Kimi's most intelligent coding model to date, reliably following instructions in long contexts and completing programming tasks with a higher success rate. It supports text, image, and video input, and only supports thought mode, dialogue, and agent tasks.

From
$0.76/1M tokens
View model
Happy Horse 1.0
Aliyun
Q
Video Generation

Happy Horse 1.0

happyhorse-1.0
Video modeltext-to-videoimage-to-video

Happy Horse 1.0 — A high-quality audio-video generation model that supports text-to-video and image-to-video creation. It can generate synchronized visuals, audio, and lip movements, making it suitable for short films, advertising creatives, and product showcases.

From
$0.112/s
View model
Claude Mythos 5
Anthropic
Coming soon
C
Text Generation

Claude Mythos 5

claude-mythos-5
Text modelComing soontext-to-textpdf-to-textimage-to-text

Anthropic's most capable, widely released model, for the most demanding reasoning and long-horizon agentic work

From
Coming soon
View model
Claude Opus 4.7
Anthropic
Popular
C
Text Generation

Claude Opus 4.7

claude-opus-4-7
Text modelPopulartext-to-text

Claude Opus 4.7 is a hybrid reasoning model designed specifically for frontier-level coding, AI agents, and complex multi-step professional work. Unlike lighter models (e.g., Sonnet or Haiku variants), Opus 4.7 prioritizes depth, consistency, and autonomy on the hardest tasks.

From
$4/1M tokens
View model
MiniMax-M3
MiniMax
M
Text Generation

MiniMax-M3

minimax-m3
Text modeltext-to-text

Minimax-m3 is a multimodal AI model designed for strong reasoning, natural conversation, and creative content generation. It provides balanced performance across text and visual understanding tasks, making it suitable for general-purpose AI applications.

From
$0.48/1M tokens
View model
GPT 5.5 Pro
OpenAI
O
Text Generation

GPT 5.5 Pro

gpt-5.5-pro
Text modeltext-to-text

GPT-5.5 Pro combines state-of-the-art intelligence, precision, and efficiency to tackle sophisticated challenges. From software development and data analysis to research and decision support, it delivers expert-level assistance with speed and consistency.

From
View pricing
View model
DeepSeek V4 Pro
DeepSeek
Popular
D
Text Generation

DeepSeek V4 Pro

deepseek-v4-pro
Text modelPopulartext-to-text

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding, and long-horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks.

From
$0.416/1M tokens
View model
GPT 5.5
OpenAI
O
Text Generation

GPT 5.5

gpt-5.5
Text modeltext-to-text

Model 5.5 is a next-generation AI model designed for stronger reasoning, faster responses, and improved accuracy across a wide range of tasks. It excels at understanding complex instructions, generating high-quality content, and assisting with coding, analysis, and problem-solving.

From
View pricing
View model
MiniMax-M2.7
MiniMax
M
Text Generation

MiniMax-M2.7

minimax-m2.7
Text modeltext-to-text

MiniMax-M2.7 offers the same top-tier intelligence as the standard version—including recursive self-evolution and expert-level office productivity—but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

From
$0.24/1M tokens
View model
DeepSeek V4 Flash
DeepSeek
Popular
D
Text Generation

DeepSeek V4 Flash

deepseek-v4-flash
Text modelPopulartext-to-text

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.

From
$0.12/1M tokens
View model
GPT-5.4 nano
OpenAI
O
Text Generation

GPT-5.4 nano

gpt-5.4-nano
Text modeltext-to-text400,000 context

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.

From
View pricing
View model
GPT-5.4 mini
OpenAI
O
Text Generation

GPT-5.4 mini

gpt-5.4-mini
Text modeltext-to-text400,000 context

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

From
View pricing
View model
GPT-5.4 pro
OpenAI
O
Text Generation

GPT-5.4 pro

gpt-5.4-pro
Text modeltext-to-textimage-to-text1,050,000 context

GPT-5.4 Pro is a high-performance AI model designed for professional and business applications. It offers strong reasoning, reliable accuracy, and efficient execution across tasks such as content creation, coding, research, and data analysis.

From
View pricing
View model
Nano Banana 2
Google
Popular
G
Image Generation

Nano Banana 2

gemini-3.1-flash-image-preview
Image modelPopulartext-to-imageimage-editing

Core Capabilities Overview: Resolution: Up to 4K (4096×4096), on par with Pro. Reference Image Consistency: Up to 14 reference images (10 objects + 4 characters), maintaining style/character consistency. Extreme Aspect Ratios: New 1:4, 4:1, 1:8, 8:1 ratios added, suitable for long images, posters, and banners. Text Rendering: Advanced text generation, suitable for infographics and marketing poster layouts. Search Enhancement: Integrated Google Search + Image Search. Grounding: Built-in thinking process; complex prompts are reasoned before generation.

From
$0.4/1M tokens
View model
Claude Sonnet 4.6
Anthropic
Popular
C
Text Generation

Claude Sonnet 4.6

claude-sonnet-4-6
Text modelPopulartext-to-textimage-to-text

Claude Sonnet 4.6 is our most capable Sonnet model yet. It’s a full upgrade of the model’s skills across coding, computer use, long-context reasoning, agent planning, knowledge work, and design. Sonnet 4.6 also features a 1M token context window in beta.

From
$2.4/1M tokens
View model
GLM 5.2
Zhipu AI
Popular
Z
Text Generation

GLM 5.2

glm-5.2
Text modelPopulartext-to-text

GLM-5.2 is a significant update from Zhipu in the areas of open-source large models and AI coding.

From
$1.12/1M tokens
View model