DeepSeek v4
DeepSeek-V4 Preview (Flash/Pro variants) is a Mixture-of-Experts model series with 1M token context, top-tier reasoning, and strong agent capabilities. Fully OpenAI/Anthropic compatible.
Input / 1M
$0.140
Output / 1M
$0.280
Pricing database
Compare official API pricing and technical limits across providers.
20 models
DeepSeek-V4 Preview (Flash/Pro variants) is a Mixture-of-Experts model series with 1M token context, top-tier reasoning, and strong agent capabilities. Fully OpenAI/Anthropic compatible.
Input / 1M
$0.140
Output / 1M
$0.280
Luma’s first unified understanding and generation model, a decoder-only autoregressive transformer that interleaves text and images for multimodal reasoning, visual editing, and image generation
Input / 1M
N/A
Output / 1M
N/A
Next-generation foundational image generation model unifying text-to-image generation and image editing with professional typography rendering, native 2K resolution support, and 1k-token prompt instructions.
Input / 1M
N/A
Output / 1M
N/A
State-of-the-art image generation and editing model (Gemini 3.1 Flash Image) combining Pro-level quality, advanced world knowledge, real-time web search grounding, subject consistency, and precise text rendering with Flash-level speed.
Input / 1M
N/A
Output / 1M
N/A
Most capable and efficient frontier model for professional work, with native computer-use capabilities, tool search, and extreme reasoning
Input / 1M
$2.50
Output / 1M
$15.00
An efficient Mixture-of-Experts language model with 671B total parameters, featuring DeepSeek Sparse Attention (DSA) for enhanced reasoning, agentic performance, and long-context efficiency, comparable to frontier models like GPT-5.
Input / 1M
$0.280
Output / 1M
$0.420
Grok 4.20 Beta: xAI's frontier multimodal model with native 4-agent multi-agent collaboration system (Grok, Harper, Benjamin, Lucas) for real-time debate, fact-checking, and reduced hallucinations
Input / 1M
N/A
Output / 1M
N/A
A 1-trillion parameter Mixture-of-Experts (MoE) multimodal model featuring native vision capabilities, complex reasoning, and Agent Swarm support for parallel sub-agent coordination.
Input / 1M
$0.600
Output / 1M
$3.00
Gemini 3.1 Pro is the next iteration in the Gemini 3 series of models, a suite of highly capable, natively multimodal reasoning models.
Input / 1M
$2.00
Output / 1M
$12.00
Lightweight 9B-parameter open-source multimodal vision-language model optimized for local deployment, low-latency inference, and edge/consumer hardware; part of the GLM-4.6V series with native multimodal function calling, strong visual understanding, and long-context capabilities
Input / 1M
$0.0000
Output / 1M
$0.0000
Paid, enhanced version of GLM-4.6V-Flash multimodal model with higher capacity and stability; supports native multimodal tool calling, vision-language tasks, and long-context processing
Input / 1M
$0.0000
Output / 1M
$0.400
Open-source multimodal vision-language model with native function calling, state-of-the-art visual understanding and reasoning at its scale, long-context multimodal processing, and support for interleaved image-text generation and agentic workflows
Input / 1M
$0.300
Output / 1M
$0.900
Specialized coding variant of the GLM-5 flagship model, optimized for advanced programming, complex code generation, agentic development workflows, and superior performance in software engineering tasks
Input / 1M
$1.20
Output / 1M
$5.00
Flagship Mixture-of-Experts foundation model designed for Agentic Engineering, complex systems reasoning, long-horizon agent tasks, advanced coding, and low hallucination rates
Input / 1M
$1.00
Output / 1M
$3.20
Open-weight, general-purpose, flagship multimodal and multilingual model.
Input / 1M
$0.500
Output / 1M
$1.50
Frontier-class multimodal model optimized for enterprise-grade performance, high-speed reasoning, and long-context document understanding.
Input / 1M
$0.400
Output / 1M
$2.00
A proprietary Optical Character Recognition (OCR) model specialized in complex tables, forms, handwritten content, and multi-page PDFs, outputting high-fidelity Markdown or HTML structures.
Input / 1M
N/A
Output / 1M
N/A
Released in Feb 2026, Opus 4.6 is Anthropic's frontier model featuring 'Adaptive Thinking' and 'Agent Teams' capabilities. It offers a 1M token context window (in beta) and is optimized for complex reasoning, deep research, and autonomous coding tasks.
Input / 1M
$5.00
Output / 1M
$25.00
Anthropic's intelligent flagship model optimized for complex agents, software engineering, and computer use. It features 'extended thinking' for deep reasoning and achieves state-of-the-art performance on coding benchmarks like SWE-bench Verified.
Input / 1M
$3.00
Output / 1M
$15.00
GPT-5.2 is a more reliable and capable AI model with stronger reasoning, longer context understanding, and improved real-world usability.
Input / 1M
$1.75
Output / 1M
$14.00