Skip to content
Pricing
MODEL DIRECTORY

LLM models & pricing

Choose LLMs by native protocol, context and input/output pricing. Inspect declared capabilities and API examples.

View all models & pricing →
12 models
ModelInput → OutputActions
Input$2.00Output$10.00per 1M tokens80% OFF1.05M / $0.20per 1M tokens
Model details
API ID
gpt-6-astra
Context
1.05M
Cache write / read
/ $0.20 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Claude Opus 5Anthropic
Input$0.50Output$2.50per 1M tokens90% OFF1M$0.625 / $0.05per 1M tokens
Model details
API ID
claude-opus-5
Context
1M
Cache write / read
$0.625 / $0.05 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$0.04Output$0.24per 1M tokens80% OFF1.05M / $0.004per 1M tokens
Model details
API ID
gpt-5.6-luna
Context
1.05M
Cache write / read
/ $0.004 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$1.00Output$6.00per 1M tokens72% OFF1.05M / $0.10per 1M tokens
Model details
API ID
gpt-5.6-sol
Context
1.05M
Cache write / read
/ $0.10 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$0.40Output$2.40per 1M tokens80% OFF1.05M / $0.04per 1M tokens
Model details
API ID
gpt-5.6-terra
Context
1.05M
Cache write / read
/ $0.04 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$1.00Output$5.00per 1M tokens90% OFF1M$1.25 / $0.10per 1M tokens
Model details
API ID
claude-fable-5
Context
1M
Cache write / read
$1.25 / $0.10 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$0.50Output$2.50per 1M tokens90% OFF1M$0.625 / $0.05per 1M tokens
Model details
API ID
claude-opus-4-8
Context
1M
Cache write / read
$0.625 / $0.05 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
GPT-5.5OpenAI
Input$1.00Output$6.00per 1M tokens80% OFF1.05M / $0.10per 1M tokens
Model details
API ID
gpt-5.5
Context
1.05M
Cache write / read
/ $0.10 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$0.50Output$2.50per 1M tokens90% OFF1M$0.625 / $0.05per 1M tokens
Model details
API ID
claude-opus-4-7
Context
1M
Cache write / read
$0.625 / $0.05 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$0.50Output$2.50per 1M tokens90% OFF1M$0.625 / $0.05per 1M tokens
Model details
API ID
claude-opus-4-6
Context
1M
Cache write / read
$0.625 / $0.05 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
Input$0.20Output$1.00per 1M tokens90% OFF128K$0.25 / $0.02per 1M tokens
Model details
API ID
claude-sonnet-5
Context
128K
Cache write / read
$0.25 / $0.02 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
gpt-5.4OpenAI
Input$0.50Output$3.00per 1M tokens80% OFF1.05M / $0.05per 1M tokens
Model details
API ID
gpt-5.4
Context
1.05M
Cache write / read
/ $0.05 per 1M tokens
Input → Output
texttext
Discount
Based on equally weighted official standard input and output rates, excluding cache and other charges.
12 modelsUSD · Public pricing

Choose and integrate

Native protocols in this catalog: openai · claude. Choose a protocol that matches your app, then inspect the model’s declared capabilities.

Context and maximum output constrain request size and generated length separately. Check each model’s declarations for tools, streaming and structured output.

The table shows input and output prices per million tokens, with cache rates separately. Price sorting uses the public input starting rate; inspect the model page for all billing items.

Docs: native API protocols → · Docs: your first request →