Models

Call Portal M2-25 or Portal M3-55 and pay per token in USD.

Available models
Portal M2-25 Non-fast 200k

A frontier agentic coding model. Built for long-context reasoning, precise edits, and tool use - exposed through a single stable model id.

Get an API key →
200K context window Frontier class Streaming Tool / function calling Prompt caching
Model id
portal/m2-25
Context window
200,000 tokens
Billing
Per token, USD
Endpoint
/v1/chat/completions
Per 1M tokens · USD
InputTokens you send in the request
$0.12
/ 1M
Cached inputCache read - repeated context served from cache
$0.05
/ 1M
Cache writeWriting new context into the cache
$0.12
/ 1M
OutputTokens the model generates
$0.59
/ 1M

Every call is metered per token in USD and billed against your wallet - monthly plan credit first, then your prepaid balance. See your real spend per call, with full token breakdowns, in Usage and Logs.

Portal M2-25 Fast 200k

Priority tier - same capabilities as Portal M2-25 Non-fast, higher throughput and on-demand billing.

Get an API key →
200K context window Priority tier Frontier class Streaming Tool / function calling
Model id
portal/m2-25-fast
Context window
200,000 tokens
Billing
Per token, USD
Endpoint
/v1/chat/completions
Per 1M tokens · USD
Input
$0.70
/ 1M
Cached input
$0.30
/ 1M
Cache write
$0.70
/ 1M
Output
$3.50
/ 1M
Portal M3-55 XH 1M

Top-tier general-purpose frontier model at Extra-High reasoning effort with a 1M-token context window. For complex agents, analysis and tool-heavy workflows.

Get an API key →
1M context window Extra-High reasoning effort Frontier class Streaming Tool / function calling Prompt caching
Model id
portal/m3-55-xh-1m
Context window
1,000,000 tokens
Reasoning effort
Extra-High
Billing
Per token, USD
Per 1M tokens · USD
Input
$3.50
/ 1M
Cached input
$0.35
/ 1M
Cache write
$3.50
/ 1M
Output
$21.00
/ 1M

Reasoning effort and context length don't change the per-token rate. Deeper reasoning spends more tokens, so hard problems cost more in total, never more per token. Metered per token in USD and billed against your wallet.

Portal M3 500K Fast New

Flagship-speed frontier tier: near-top reasoning at priority speed with a 500K-class context window. The default when you want strong and fast at once.

Get an API key →
500K context window Priority speed Frontier class Streaming Tool / function calling Prompt caching
Model id
portal/m3-500k-fast
Context window
500,000 tokens
Billing
Per token, USD
Endpoint
/v1/chat/completions
Per 1M tokens · USD
Input
$2.80
/ 1M
Cached input
$0.70
/ 1M
Cache write
$2.80
/ 1M
Output
$8.40
/ 1M

This catalog contains Portal-branded customer routes. Underlying routing and availability may change.

Start building

Mint a key in the dashboard and point any OpenAI-compatible client at https://api.portal.ai/v1.