Skip to content
-
  • Facebook
  • X
  • Telegram
  • DEVDEV
AI Feed AI Feed AI Feed

AI news, tools, comparisons and practical guides

Subscribe
AI Feed AI Feed AI Feed

AI news, tools, comparisons and practical guides

  • AI News
  • Radar
  • AI Comparisons
  • About
  • AI API Prices
  • Local AI
  • Jobs

Sections

  • AI Comparisons
  • AI Features
  • AI Guides
  • AI News
  • Blog
  • Jobs
  • Uncategorized

Latest stories

  • Which Work Tasks Will AI Automate First—and How to Audit Your Own Job
  • Google Launches Guided Vision in Gemini Live for Android
  • Google unveils Gemini 4 Argon with a 1M-token output limit
  • Barclays Expands Claude Across Software and Bank Operations
  • Sber Details HG SDLC, an Orchestrator for Coding Agents
  • AI News
  • Radar
  • AI Comparisons
  • About
  • AI API Prices
  • Local AI
  • Jobs
Subscribe
Close

Search

Home/AI API Token Prices and Request Calculator
AI API Token Prices and Request Calculator
AI MODEL PRICE INDEX

AI API token prices
and request calculator

Compare input and output token prices, calculate a real workload and track provider changes.

LAST UPDATEOct 2, 2026 01:48Automatic hourly check
FLAGSHIP INDEX$5.98average for 1M mixed tokens
MODELS TRACKED464hourly catalogue scan
AVERAGE INPUT$1.85per 1M tokens
AVERAGE OUTPUT$8.65per 1M tokens
90-DAY PRICE HISTORYFlagship model price index
History collection has started. New real data points will appear every hour.
DeepSeek: DeepSeek V4 Pro 0813$1.320.00%Google: Gemini 3.8 Flash$2.250.00%OpenAI: GPT-6 Astra$30.000.00%Anthropic: Claude Fable 5.1$30.000.00%

Choose a typical workload or enter your own values. Prices in the table are recalculated instantly.

PERSONAL COMPARISONSelect up to 4 models in the table

Use the + buttons next to model names.

Showing 60 of 290 modelsOpen full catalogue →
ModelInput / 1MOutput / 1MChangeYour request
OpenAI: o1-pro⌄OpenaiLegacyVery expensive$150$6000%—
MODEL PROFILE

OpenAI: o1-pro

API profile based on provider metadata: image input, file input, coding, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingComplex analysisStructured outputLong contextLong responses
Context window
200K tokens
Maximum output
100K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
6 supported
Model ID
openai/o1-pro
Published
Mar 20, 2025
Open model documentation ↗
OpenAI: GPT-5.5 Pro⌄OpenaiVery expensive$30$1800%—
MODEL PROFILE

OpenAI: GPT-5.5 Pro

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1.1M tokens
Maximum output
128K tokens
Accepts
Files, Images, Text
Produces
Text
Safety moderation
Enabled
API features
10 supported
Model ID
openai/gpt-5.5-pro
Published
Apr 24, 2026
Open model documentation ↗
OpenAI: GPT-5.4 Pro⌄OpenaiVery expensive$30$1800%—
MODEL PROFILE

OpenAI: GPT-5.4 Pro

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1.1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
openai/gpt-5.4-pro
Published
Mar 5, 2026
Open model documentation ↗
OpenAI: GPT-5.2 Pro⌄OpenaiVery expensive$21$1680%—
MODEL PROFILE

OpenAI: GPT-5.2 Pro

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
400K tokens
Maximum output
128K tokens
Accepts
Images, Text, Files
Produces
Text
Safety moderation
Enabled
API features
10 supported
Model ID
openai/gpt-5.2-pro
Published
Dec 10, 2025
Open model documentation ↗
OpenAI: GPT-5 Pro⌄OpenaiVery expensive$15$1200%—
MODEL PROFILE

OpenAI: GPT-5 Pro

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
400K tokens
Maximum output
128K tokens
Accepts
Images, Text, Files
Produces
Text
Safety moderation
Enabled
API features
10 supported
Model ID
openai/gpt-5-pro
Published
Oct 6, 2025
Open model documentation ↗
OpenAI: o3 Pro⌄OpenaiVery expensive$20$800%—
MODEL PROFILE

OpenAI: o3 Pro

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
200K tokens
Maximum output
100K tokens
Accepts
Text, Files, Images
Produces
Text
Safety moderation
Enabled
API features
8 supported
Model ID
openai/o3-pro
Published
Jun 11, 2025
Open model documentation ↗
Anthropic: Claude Opus 4.1⌄AnthropicVery expensive$15$750%—
MODEL PROFILE

Anthropic: Claude Opus 4.1

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisLong contextLong responses
Context window
200K tokens
Maximum output
32K tokens
Accepts
Images, Text, Files
Produces
Text
Safety moderation
Enabled
API features
9 supported
Model ID
anthropic/claude-opus-4.1
Published
Aug 5, 2025
Open model documentation ↗
OpenAI: GPT-4⌄OpenaiVery expensive$30$600%—
MODEL PROFILE

OpenAI: GPT-4

API profile based on provider metadata: ai agents & tools, structured output.

AI agents & toolsStructured output
Context window
8.2K tokens
Maximum output
4.1K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
15 supported
Model ID
openai/gpt-4
Published
May 28, 2023
Open model documentation ↗
OpenAI: o1⌄OpenaiVery expensive$15$600%—
MODEL PROFILE

OpenAI: o1

API profile based on provider metadata: image input, file input, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
200K tokens
Maximum output
100K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
8 supported
Model ID
openai/o1
Published
Dec 17, 2024
Open model documentation ↗
OpenAI: GPT-6 Astra⌄OpenaiFlagshipVery expensive$10$500%—
MODEL PROFILE

OpenAI: GPT-6 Astra

API profile based on provider metadata: image input, file input, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1.1M tokens
Maximum output
128K tokens
Accepts
Files, Images, Text
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
openai/gpt-6-astra
Published
Sep 4, 2026
Open model documentation ↗
OpenAI: GPT-6 Astra Pro⌄OpenaiFlagshipVery expensive$10$500%—
MODEL PROFILE

OpenAI: GPT-6 Astra Pro

API profile based on provider metadata: image input, file input, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1.1M tokens
Maximum output
128K tokens
Accepts
Files, Images, Text
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
openai/gpt-6-astra-pro
Published
Sep 4, 2026
Open model documentation ↗
Anthropic: Claude Fable 5.1⌄AnthropicFlagshipVery expensive$10$500%—
MODEL PROFILE

Anthropic: Claude Fable 5.1

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
anthropic/claude-fable-5.1
Published
Sep 1, 2026
Open model documentation ↗
Anthropic: Claude Fable 5⌄AnthropicFlagshipVery expensive$10$500%—
MODEL PROFILE

Anthropic: Claude Fable 5

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
anthropic/claude-fable-5
Published
Jun 9, 2026
Open model documentation ↗
OpenAI: GPT-4 Turbo⌄Openai$10$300%—
MODEL PROFILE

OpenAI: GPT-4 Turbo

API profile based on provider metadata: image input, ai agents & tools, structured output.

Image inputAI agents & toolsStructured output
Context window
128K tokens
Maximum output
4.1K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Enabled
API features
14 supported
Model ID
openai/gpt-4-turbo
Published
Apr 9, 2024
Open model documentation ↗
OpenAI: GPT Chat Latest⌄Openai$5$300%—
MODEL PROFILE

OpenAI: GPT Chat Latest

API profile based on provider metadata: image input, file input, ai agents & tools, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputAI agents & toolsStructured outputLong contextLong responses
Context window
400K tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
6 supported
Model ID
openai/gpt-chat-latest
Published
May 5, 2026
Open model documentation ↗
OpenAI: GPT-5.5⌄Openai$5$300%—
MODEL PROFILE

OpenAI: GPT-5.5

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1.1M tokens
Maximum output
128K tokens
Accepts
Files, Images, Text
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
openai/gpt-5.5
Published
Apr 24, 2026
Open model documentation ↗
Anthropic: Claude Opus 5⌄AnthropicFlagship$5$250%—
MODEL PROFILE

Anthropic: Claude Opus 5

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
12 supported
Model ID
anthropic/claude-opus-5
Published
Jul 24, 2026
Open model documentation ↗
Anthropic: Claude Opus 4.8⌄Anthropic$5$250%—
MODEL PROFILE

Anthropic: Claude Opus 4.8

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
12 supported
Model ID
anthropic/claude-opus-4.8
Published
May 27, 2026
Open model documentation ↗
Anthropic: Claude Opus 4.7⌄Anthropic$5$250%—
MODEL PROFILE

Anthropic: Claude Opus 4.7

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
11 supported
Model ID
anthropic/claude-opus-4.7
Published
Apr 16, 2026
Open model documentation ↗
Anthropic: Claude Opus 4.6⌄Anthropic$5$250%—
MODEL PROFILE

Anthropic: Claude Opus 4.6

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
14 supported
Model ID
anthropic/claude-opus-4.6
Published
Feb 4, 2026
Open model documentation ↗
Anthropic: Claude Opus 4.5⌄Anthropic$5$250%—
MODEL PROFILE

Anthropic: Claude Opus 4.5

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
200K tokens
Maximum output
64K tokens
Accepts
Files, Images, Text
Produces
Text
Safety moderation
Enabled
API features
12 supported
Model ID
anthropic/claude-opus-4.5
Published
Nov 24, 2025
Open model documentation ↗
Anthropic: Claude Opus 5.5⌄AnthropicFlagship$4$200%—
MODEL PROFILE

Anthropic: Claude Opus 5.5

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
12 supported
Model ID
anthropic/claude-opus-5.5
Published
Sep 22, 2026
Open model documentation ↗
OpenAI: GPT-5.6 Sol Pro⌄Openai$4$200%—
MODEL PROFILE

OpenAI: GPT-5.6 Sol Pro

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1.1M tokens
Maximum output
128K tokens
Accepts
Files, Images, Text
Produces
Text
Safety moderation
Not declared
API features
11 supported
Model ID
openai/gpt-5.6-sol-pro
Published
Jul 9, 2026
Open model documentation ↗
OpenAI: GPT-5.4 Image 2⌄Openai$8$150%—
MODEL PROFILE

OpenAI: GPT-5.4 Image 2

API profile based on provider metadata: image input, file input, image generation, coding, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputImage generationCodingComplex analysisStructured outputLong contextLong responses
Context window
272K tokens
Maximum output
128K tokens
Accepts
Images, Text, Files
Produces
Images, Text
Safety moderation
Enabled
API features
14 supported
Model ID
openai/gpt-5.4-image-2
Published
Apr 21, 2026
Open model documentation ↗
OpenAI: GPT-5 Image⌄Openai$10$100%—
MODEL PROFILE

OpenAI: GPT-5 Image

API profile based on provider metadata: image input, file input, image generation, coding, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputImage generationCodingComplex analysisStructured outputLong contextLong responses
Context window
400K tokens
Maximum output
128K tokens
Accepts
Images, Text, Files
Produces
Images, Text
Safety moderation
Enabled
API features
15 supported
Model ID
openai/gpt-5-image
Published
Oct 14, 2025
Open model documentation ↗
OpenAI: GPT-4o (2024-05-13)⌄Openai$5$150%—
MODEL PROFILE

OpenAI: GPT-4o (2024-05-13)

API profile based on provider metadata: image input, file input, ai agents & tools, structured output.

Image inputFile inputAI agents & toolsStructured output
Context window
128K tokens
Maximum output
4.1K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
17 supported
Model ID
openai/gpt-4o-2024-05-13
Published
May 13, 2024
Open model documentation ↗
Qwen: Qwen3.8 Max Prime⌄Qwen$4$120%—
MODEL PROFILE

Qwen: Qwen3.8 Max Prime

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
17 supported
Model ID
qwen/qwen3.8-max-prime
Published
Sep 23, 2026
Open model documentation ↗
Anthropic: Claude Sonnet 5.5⌄Anthropic$2$100%—
MODEL PROFILE

Anthropic: Claude Sonnet 5.5

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
128K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Enabled
API features
12 supported
Model ID
anthropic/claude-sonnet-5.5
Published
Sep 28, 2026
Open model documentation ↗
SpaceXAI: Grok 4.7⌄X-ai$2$60%—
MODEL PROFILE

SpaceXAI: Grok 4.7

API profile based on provider metadata: image input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
500K tokens
Maximum output
450K tokens
Accepts
Text, Images, Files
Produces
Text
Safety moderation
Not declared
API features
13 supported
Model ID
x-ai/grok-4.7
Published
Sep 21, 2026
Open model documentation ↗
Qwen: Qwen3.8 Max (0902)⌄Qwen$2$60%—
MODEL PROFILE

Qwen: Qwen3.8 Max (0902)

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
17 supported
Model ID
qwen/qwen3.8-max-0902
Published
Sep 4, 2026
Open model documentation ↗
Qwen: Qwen3.8 2.4T A95B⌄Qwen$2$60%—
MODEL PROFILE

Qwen: Qwen3.8 2.4T A95B

API profile based on provider metadata: coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

CodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
20 supported
Model ID
qwen/qwen3.8-2.4t-a95b
Published
Aug 12, 2026
Open model documentation ↗
Google: Gemini 3.8 Flash⌄Google$0.75$3.750%—
MODEL PROFILE

Google: Gemini 3.8 Flash

API profile based on provider metadata: image input, audio, video input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputAudioVideo inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
65.5K tokens
Accepts
Text, Images, Video, Files, Audio
Produces
Text
Safety moderation
Not declared
API features
12 supported
Model ID
google/gemini-3.8-flash
Published
Sep 2, 2026
Open model documentation ↗
Google: Gemini 3.7 Flash⌄Google$0.75$3.750%—
MODEL PROFILE

Google: Gemini 3.7 Flash

API profile based on provider metadata: image input, audio, video input, file input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputAudioVideo inputFile inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
65.5K tokens
Accepts
Text, Images, Video, Files, Audio
Produces
Text
Safety moderation
Not declared
API features
12 supported
Model ID
google/gemini-3.7-flash
Published
Aug 13, 2026
Open model documentation ↗
Qwen: Qwen3.8 27B⌄Qwen$0.42$30%—
MODEL PROFILE

Qwen: Qwen3.8 27B

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
21 supported
Model ID
qwen/qwen3.8-27b
Published
Aug 14, 2026
Open model documentation ↗
DeepSeek: DeepSeek V4 Pro 0813⌄Deepseek$0.66$1.980%—
MODEL PROFILE

DeepSeek: DeepSeek V4 Pro 0813

API profile based on provider metadata: coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

CodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
393.2K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
20 supported
Model ID
deepseek/deepseek-v4-pro-0813
Published
Aug 12, 2026
Open model documentation ↗
DeepSeek: DeepSeek V4 Flash Vision Exp⌄DeepseekGood value$0.2156$0.64680%—
MODEL PROFILE

DeepSeek: DeepSeek V4 Flash Vision Exp

API profile based on provider metadata: image input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
262.1K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Not declared
API features
20 supported
Model ID
deepseek/deepseek-v4-flash-vision-exp
Published
Aug 21, 2026
Open model documentation ↗
Qwen: Qwen3.8 Omni Flash⌄QwenGood value$0.15$0.470%—
MODEL PROFILE

Qwen: Qwen3.8 Omni Flash

API profile based on provider metadata: image input, audio, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputAudioVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text, Images, Audio, Video
Produces
Text
Safety moderation
Not declared
API features
16 supported
Model ID
qwen/qwen3.8-omni-flash
Published
Sep 21, 2026
Open model documentation ↗
Qwen: Qwen3.8 Flash⌄QwenGood value$0.15$0.470%—
MODEL PROFILE

Qwen: Qwen3.8 Flash

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
16 supported
Model ID
qwen/qwen3.8-flash
Published
Aug 26, 2026
Open model documentation ↗
DeepSeek: DeepSeek V4.1 Flash⌄DeepseekGood value$0.03$0.50%—
MODEL PROFILE

DeepSeek: DeepSeek V4.1 Flash

API profile based on provider metadata: image input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
943.7K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Not declared
API features
20 supported
Model ID
deepseek/deepseek-v4.1-flash
Published
Sep 10, 2026
Open model documentation ↗
Qwen: Qwen3.5-Flash⌄QwenGood value$0.065$0.260%—
MODEL PROFILE

Qwen: Qwen3.5-Flash

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
65.5K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
14 supported
Model ID
qwen/qwen3.5-flash-02-23
Published
Feb 25, 2026
Open model documentation ↗
Qwen: Qwen2.5 7B Instruct⌄QwenGood value$0.1$0.20%—
MODEL PROFILE

Qwen: Qwen2.5 7B Instruct

API profile based on provider metadata: coding, ai agents & tools, structured output.

CodingAI agents & toolsStructured output
Context window
32.8K tokens
Maximum output
29.5K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
14 supported
Model ID
qwen/qwen-2.5-7b-instruct
Published
Oct 16, 2024
Open model documentation ↗
Mistral: Ministral 3 8B 2512⌄MistralaiGood value$0.15$0.150%—
MODEL PROFILE

Mistral: Ministral 3 8B 2512

API profile based on provider metadata: image input, ai agents & tools, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputAI agents & toolsStructured outputLong contextLong responses
Context window
262.1K tokens
Maximum output
209.7K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Not declared
API features
11 supported
Model ID
mistralai/ministral-8b-2512
Published
Dec 2, 2025
Open model documentation ↗
Qwen: Qwen3.5-9B⌄QwenGood value$0.1$0.150%—
MODEL PROFILE

Qwen: Qwen3.5-9B

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
262.1K tokens
Maximum output
32.8K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
19 supported
Model ID
qwen/qwen3.5-9b
Published
Mar 10, 2026
Open model documentation ↗
Qwen: Qwen3 30B A3B Instruct 2507⌄QwenGood value$0.0482$0.19310%—
MODEL PROFILE

Qwen: Qwen3 30B A3B Instruct 2507

API profile based on provider metadata: coding, ai agents & tools, structured output, long context, long responses. Suitable for workloads with a large context window.

CodingAI agents & toolsStructured outputLong contextLong responses
Context window
262.1K tokens
Maximum output
32K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
16 supported
Model ID
qwen/qwen3-30b-a3b-instruct-2507
Published
Jul 29, 2025
Open model documentation ↗
Meta: Llama 3.2 1B Instruct⌄Meta-llamaGood value$0.027$0.2010%—
MODEL PROFILE

Meta: Llama 3.2 1B Instruct

API profile based on provider metadata: long responses.

Long responses
Context window
60K tokens
Maximum output
54K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
11 supported
Model ID
meta-llama/llama-3.2-1b-instruct
Published
Sep 25, 2024
Open model documentation ↗
OpenAI: gpt-oss-120b⌄OpenaiGood value$0.037$0.170%—
MODEL PROFILE

OpenAI: gpt-oss-120b

API profile based on provider metadata: ai agents & tools, complex analysis, structured output, long responses.

AI agents & toolsComplex analysisStructured outputLong responses
Context window
131.1K tokens
Maximum output
118K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
21 supported
Model ID
openai/gpt-oss-120b
Published
Aug 5, 2025
Open model documentation ↗
Google: Gemma 3 12B⌄GoogleGood value$0.05$0.150%—
MODEL PROFILE

Google: Gemma 3 12B

API profile based on provider metadata: image input, ai agents & tools, structured output.

Image inputAI agents & toolsStructured output
Context window
131.1K tokens
Maximum output
16.4K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Not declared
API features
15 supported
Model ID
google/gemma-3-12b-it
Published
Mar 13, 2025
Open model documentation ↗
Mistral: Ministral 3 3B 2512⌄MistralaiGood value$0.1$0.10%—
MODEL PROFILE

Mistral: Ministral 3 3B 2512

API profile based on provider metadata: image input, ai agents & tools, structured output, long responses.

Image inputAI agents & toolsStructured outputLong responses
Context window
131.1K tokens
Maximum output
104.9K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Not declared
API features
11 supported
Model ID
mistralai/ministral-3b-2512
Published
Dec 2, 2025
Open model documentation ↗
Qwen: Qwen3.7 Flash⌄QwenGood value$0.03$0.130%—
MODEL PROFILE

Qwen: Qwen3.7 Flash

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
65.5K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
12 supported
Model ID
qwen/qwen3.7-flash
Published
Jul 28, 2026
Open model documentation ↗
Google: Gemma 3 4B⌄GoogleGood value$0.05$0.10%—
MODEL PROFILE

Google: Gemma 3 4B

API profile based on provider metadata: image input, structured output.

Image inputStructured output
Context window
131.1K tokens
Maximum output
16.4K tokens
Accepts
Text, Images
Produces
Text
Safety moderation
Not declared
API features
13 supported
Model ID
google/gemma-3-4b-it
Published
Mar 14, 2025
Open model documentation ↗
Meta: Llama 3.1 8B Instruct⌄Meta-llamaGood value$0.05$0.080%—
MODEL PROFILE

Meta: Llama 3.1 8B Instruct

API profile based on provider metadata: ai agents & tools, structured output, long responses.

AI agents & toolsStructured outputLong responses
Context window
131.1K tokens
Maximum output
118K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
17 supported
Model ID
meta-llama/llama-3.1-8b-instruct
Published
Jul 23, 2024
Open model documentation ↗
Mistral: Mistral Small 3⌄MistralaiGood value$0.05$0.080%—
MODEL PROFILE

Mistral: Mistral Small 3

API profile based on provider metadata: structured output.

Structured output
Context window
32.8K tokens
Maximum output
16.4K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
13 supported
Model ID
mistralai/mistral-small-24b-instruct-2501
Published
Jan 30, 2025
Open model documentation ↗
DeepSeek: DeepSeek V4 Flash 0423⌄DeepseekGood value$0.0419$0.08370%—
MODEL PROFILE

DeepSeek: DeepSeek V4 Flash 0423

API profile based on provider metadata: coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

CodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
131.1K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
22 supported
Model ID
deepseek/deepseek-v4-flash
Published
Apr 24, 2026
Open model documentation ↗
OpenAI: gpt-oss-20b⌄OpenaiGood value$0.018$0.090%—
MODEL PROFILE

OpenAI: gpt-oss-20b

API profile based on provider metadata: ai agents & tools, complex analysis, structured output, long responses.

AI agents & toolsComplex analysisStructured outputLong responses
Context window
131.1K tokens
Maximum output
32.8K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
20 supported
Model ID
openai/gpt-oss-20b
Published
Aug 5, 2025
Open model documentation ↗
Mistral: Mistral Nemo⌄MistralaiGood value$0.019$0.030%—
MODEL PROFILE

Mistral: Mistral Nemo

API profile based on provider metadata: ai agents & tools, structured output.

AI agents & toolsStructured output
Context window
131.1K tokens
Maximum output
16.4K tokens
Accepts
Text
Produces
Text
Safety moderation
Not declared
API features
17 supported
Model ID
mistralai/mistral-nemo
Published
Jul 19, 2024
Open model documentation ↗
Google: Lyria 3 Clip Preview⌄GoogleFree$0$0——
MODEL PROFILE

Google: Lyria 3 Clip Preview

API profile based on provider metadata: image input, audio, coding, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputAudioCodingStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
65.5K tokens
Accepts
Text, Images
Produces
Text, Audio
Safety moderation
Not declared
API features
5 supported
Model ID
google/lyria-3-clip-preview
Published
Mar 31, 2026
Open model documentation ↗
Google: Lyria 3 Pro Preview⌄GoogleFree$0$0——
MODEL PROFILE

Google: Lyria 3 Pro Preview

API profile based on provider metadata: image input, audio, coding, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputAudioCodingStructured outputLong contextLong responses
Context window
1M tokens
Maximum output
65.5K tokens
Accepts
Text, Images
Produces
Text, Audio
Safety moderation
Not declared
API features
5 supported
Model ID
google/lyria-3-pro-preview
Published
Mar 31, 2026
Open model documentation ↗
Google: Gemma 4 31B (free)⌄GoogleFree$0$0——
MODEL PROFILE

Google: Gemma 4 31B (free)

API profile based on provider metadata: image input, video input, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
262.1K tokens
Maximum output
32.8K tokens
Accepts
Images, Text, Video
Produces
Text
Safety moderation
Not declared
API features
9 supported
Model ID
google/gemma-4-31b-it:free
Published
Apr 2, 2026
Open model documentation ↗
Google: Gemma 4 26B A4B (free)⌄GoogleFree$0$0——
MODEL PROFILE

Google: Gemma 4 26B A4B (free)

API profile based on provider metadata: image input, video input, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
262.1K tokens
Maximum output
32.8K tokens
Accepts
Images, Text, Video
Produces
Text
Safety moderation
Not declared
API features
9 supported
Model ID
google/gemma-4-26b-a4b-it:free
Published
Apr 3, 2026
Open model documentation ↗
Qwen: Qwen3.8 27B (free)⌄QwenFree$0$0——
MODEL PROFILE

Qwen: Qwen3.8 27B (free)

API profile based on provider metadata: image input, video input, coding, ai agents & tools, complex analysis, structured output, long context, long responses. Suitable for workloads with a large context window.

Image inputVideo inputCodingAI agents & toolsComplex analysisStructured outputLong contextLong responses
Context window
262.1K tokens
Maximum output
235.9K tokens
Accepts
Text, Images, Video
Produces
Text
Safety moderation
Not declared
API features
13 supported
Model ID
qwen/qwen3.8-27b:free
Published
Aug 14, 2026
Open model documentation ↗

Estimates use currently published token rates. Caching, reasoning, tools, images and provider fees may change the final bill. Source: OpenRouter model API; official provider pages are used for editorial verification.

PRICE GUIDES

Compare providers and real workloads

PROVIDEROpenAI API pricing→PROVIDERAnthropic Claude API pricing→PROVIDERGoogle Gemini API pricing→PROVIDERDeepSeek API pricing→PROVIDERxAI Grok API pricing→CALCULATIONOpenAI API vs Claude pricing→CALCULATIONGemini API vs OpenAI pricing→CALCULATIONCheapest AI APIs→CALCULATIONHow much does an AI chatbot cost?→CALCULATIONAI content generation API cost→

How to compare AI API token prices

AI API pricing usually has separate rates for input tokens sent to a model and output tokens generated in the response. Enter your typical prompt, answer length and number of requests in the calculator to compare GPT, Claude, Gemini, DeepSeek, Llama, Mistral and other models on the same workload.

Price per 1M tokens

The standard unit used by providers. The table keeps input and output rates separate so expensive generations do not disappear behind a low prompt price.

Estimated request cost

The calculator multiplies your input and output volume by the current rates. It is useful for prototypes, chatbots, agents, content pipelines and production forecasts.

Hourly price monitoring

The catalogue is checked every hour. The update time shown above tells you when the current snapshot was collected. Historical points accumulate from actual checks.

Frequently asked questions

What is an AI token?

A token is a text unit processed by a model. One token may be a word, part of a word, punctuation mark or other fragment, depending on the tokenizer.

Why are input and output prices different?

Generating new tokens generally requires more computation than reading a prompt, so providers often charge more for output.

Does the calculator show the final provider bill?

It provides an estimate based on published token rates. Caching, reasoning tokens, images, tools, minimum charges, routing and taxes may affect the final amount.

How often are AI model prices updated?

AI Feed checks the catalogue hourly. The timestamp at the top of the page shows the latest successful data snapshot.

Recent posts

  • Which Work Tasks Will AI Automate First—and How to Audit Your Own Job
  • Google Launches Guided Vision in Gemini Live for Android
  • Google unveils Gemini 4 Argon with a 1M-token output limit
  • Barclays Expands Claude Across Software and Bank Operations
  • Sber Details HG SDLC, an Orchestrator for Coding Agents

Recent comments

No comments to show.

Archives

  • October 2026
  • September 2026
  • May 2026

Sections

  • AI Comparisons
  • AI Features
  • AI Guides
  • AI News
  • Blog
  • Jobs
  • Uncategorized

    © 2026 AI Feed. All rights reserved.
    RUEN
    AboutEditorial PolicySources & methodologyCorrectionsContactPrivacyAnalytics settings
    AI Feed analytics

    Helps us understand which pages are useful. Advertising tracking is disabled.