AI Model Registry - Compare LLM Costs & Providers | Helicone
Helicone Joins Mintlify
Filters
Providers
- Anthropic
- AWS Bedrock
- Azure OpenAI
- Baseten
- Canopy Wave
- Cerebras
- Chutes
- DeepInfra
- DeepSeek
- Fireworks
- Google AI Studio
- Groq
- Helicone
- Mistral AI
- Nebius Token Factory
- Novita
- OpenAI
- OpenRouter
- Perplexity
- Vertex AI
- xAI
Price Range
- $0.00 - $200/M tokens
Context Size
- Minimum: 0 tokens
- 01.0M
Special Capabilities
- Caching
- Web Search
Input Modalities
- Text
- Image
- Audio
- Video
Output Modalities
- Text
- Image
- Audio
- Video
Filters
- Credits
- Newest First
Request Model Compare
Showing 111 of 111 models
Models
| Model | Description |
|---|---|
| OpenAI GPT-5.4 | GPT-5.4 is our frontier model for complex professional work. Reasoning effort supports: none (default), low, medium, high and xhigh. Features a 1.05M context. by openai • 1.1M context • $2.5/M in, $15.0/M out |
| OpenAI GPT-5.4 Pinned Version | GPT-5.4 is our frontier model for complex professional work. Reasoning effort supports: none (default), low, medium, high and xhigh. Features a 1.05M context. by openai • 1.1M context • $2.5/M in, $15.0/M out |
| Google Gemini 3.1 Flash-Lite Preview | Gemini 3.1 Flash-Lite Preview is Google's most cost-efficient model, optimized for high-volume agentic tasks, translation, and simple data processing. by google • 1.0M context • $0.25/M in, $1.5/M out |
| Claude Sonnet 4.6 | Claude Sonnet 4.6 is Anthropic's most capable Sonnet model, released February 2026. Features near-Opus-level intelligence at Sonnet pricing. by anthropic • 1.0M context • $3.0/M in, $15.0/M out |
| Google Gemini 3.1 Pro Preview | Gemini 3.1 Pro Preview is Google's most advanced reasoning model, released February 2026. It uses extended thinking/chain-of-thought reasoning to work. by google • 1.0M context • $2.0/M in, $12.0/M out |
| Claude Opus 4.6 | Claude Opus 4.6 is Anthropic's most capable model to date, released February 2026. Building on the intelligence of Opus 4.5. by anthropic • 1.0M context • $5.0/M in, $25.0/M out |
| Google Gemini 3 Flash Preview | Gemini 3 Flash Preview is Google's latest fast and efficient AI model optimized for quick response times while maintaining high quality. by google • 1.0M context • $0.50/M in, $3.0/M out |
| OpenAI GPT-5.2 | GPT-5.2 is our best general-purpose model, part of the GPT-5 flagship model family. by openai • 400K context • $1.8/M in, $14.0/M out |
| GPT-5.2 Pro | Tough problems that may take longer to solve but require harder thinking. by openai • 400K context • $21.0/M in, $168.0/M out |
| OpenAI GPT-5.2 Chat | GPT-5.2 Chat is continuously updated, optimized for conversational interactions. by openai • 128K context • $1.8/M in, $14.0/M out |
| OpenAI GPT Image 1.5 | GPT Image 1.5 is OpenAI's state-of-the-art image generation model with faster generation and cheaper image tokens. by openai • 8K context • $5.0/M in, $10.0/M out |
| Claude Opus 4.5 | Claude Opus 4.5 is Anthropic's flagship model released November 2025. by anthropic • 200K context • $5.0/M in, $25.0/M out |
| Google Gemini 3 Pro Image Preview | Gemini 3 Pro Image is Google's native image generation model with state-of-the-art reasoning capabilities. by google • 66K context • $2.0/M in, $12.0/M out |
| OpenAI GPT-5.1 | GPT-5.1 is an enhanced version of GPT-5 with improved performance. by openai • 400K context • $1.3/M in, $10.0/M out |
| Kimi K2 Thinking | Kimi K2 Thinking is a powerful AI model designed for complex reasoning. by moonshotai • 256K context • $0.48/M in, $2.0/M out |
| DeepSeek V3.2 | DeepSeek-V3.2-Exp introduces the groundbreaking DeepSeek Sparse Attention mechanism for long-context processing. by deepseek • 164K context • $0.26/M in, $0.40/M out |
| Qwen3 Next 80B A3B Instruct | This model is instruction-optimized for chat and agent applications. by qwen • 262K context • $0.14/M in, $1.4/M out |