Model pricesheet

Every model.
One price list.

These are the rates you pay on AETHER, platform fee included. No per-seat licence and no minimum: credits cover exactly the tokens your agents and developers use.

23 models 9 providers 6 kinds of work
ZAR at R17.00 per $1

OpenAI

9 models

Chat and agents
OpenAI Chat and agents rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
GPT-5.4
gpt-5.4 128K context Images
$3.00R51.00 $0.30R5.10 $18.00R306.00
GPT-5.4 mini
gpt-5.4-mini 128K context Images
$0.90R15.30 $0.09R1.53 $5.40R91.80
GPT-5.4 nano
gpt-5.4-nano 128K context Images
$0.24R4.08 $0.024R0.41 $1.50R25.50
GPT-5.6 Luna
gpt-5.6-luna 1M context Images
$0.24R4.08 $0.024R0.41 $1.44R24.48
GPT-5.6 Sol
gpt-5.6-sol 1M context Images
$6.00R102.00 $0.60R10.20 $36.00R612.00
GPT-5.6 Terra
gpt-5.6-terra 1M context Images
$3.00R51.00 $0.30R5.10 $18.00R306.00
Coding
OpenAI Coding rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
GPT-5.3 Codex
gpt-5.3-codex 128K context Images
$2.10R35.70 $0.21R3.57 $16.80R285.60
Embeddings
OpenAI Embeddings rates
Model Embeddings per 1M tokens
Text Embedding 3 Large
text-embedding-3-large
$0.156R2.65
Text Embedding 3 Small
text-embedding-3-small
$0.024R0.41

Anthropic

4 models

Coding
Anthropic Coding rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
Claude Fable 5
claude-fable-5 1M context Images
$12.00R204.00 $1.20R20.40 $60.00R1,020.00
Claude Haiku 4.5
claude-haiku-4-5 200K context Images
$1.20R20.40 $0.12R2.04 $6.00R102.00
Claude Opus 4.8
claude-opus-4-8 1M context Images
$6.00R102.00 $0.60R10.20 $30.00R510.00
Claude Sonnet 5
claude-sonnet-5 1M context Images
$2.40R40.80 $0.24R4.08 $12.00R204.00

Google Gemini

4 models

Chat and agents
Google Gemini Chat and agents rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
Gemini 3.6 Flash
gemini-3.6-flash 1M context Images
$1.80R30.60 $0.18R3.06 $9.00R153.00
Gemini 3.7 Flash
gemini-3.7-flash 1M context Images
$0.90R15.30 $0.09R1.53 $4.50R76.50
Voice
Google Gemini Voice rates
Model Input per 1M tokens Output per 1M tokens Audio input per 1M tokens Audio output per 1M tokens
Gemini 3.1 Flash Live
gemini-3.1-flash-live-preview
$0.90R15.30 $5.40R91.80 $3.60R61.20 $14.40R244.80
Video
Google Gemini Video rates
Model Input per 1M tokens Output per 1M tokens Video per second
Gemini Omni Flash (Preview)
gemini-omni-flash-preview
$1.80R30.60 $10.80R183.60 $0.1216R2.07

DeepSeek

1 model

Chat and agents
DeepSeek Chat and agents rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
DeepSeek V4.1 Flash
deepseek-flash 1M context Text only
$0.18R3.06 $0.0036R0.06 $0.72R12.24

Z.ai

1 model

Chat and agents
Z.ai Chat and agents rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
GLM 5.3 Flash
glm-5.3-flash 1M context Images
$0.18R3.06 $0.036R0.61 $0.60R10.20

AWS Bedrock

1 model

Chat and agents
AWS Bedrock Chat and agents rates
Model Input per 1M tokens Cached input per 1M tokens Output per 1M tokens
Claude Haiku 4.5 (AWS Bedrock)
global.anthropic.claude-haiku-4-5-20251001-v1:0 200K context Images
$1.20R20.40 $0.12R2.04 $6.00R102.00

Deepgram

1 model

Voice
Deepgram Voice rates
Model Audio per minute
Deepgram Nova-2 (STT)
deepgram-nova-2
$0.0072R0.12

ElevenLabs

1 model

Voice
ElevenLabs Voice rates
Model Audio per minute
ElevenLabs v3 (TTS)
eleven_v3
$0.144R2.45

AETHER

1 model

Compute
AETHER Compute rates
Model Compute per hour
AETHER Server Runner
aether-server-runner
$0.1598R2.72
Nothing in the catalog matches that filter.

What the numbers include. Every rate is the price billed to your credits on AETHER, with the platform fee already applied. There is no separate provider invoice and nothing added at checkout.

Currencies. Billing is in USD. The ZAR view converts at the fixed R17.00 per $1 rate used for credit top-ups, rounded to the cent.

Units. Token rates are per million tokens. Voice is per minute of audio, video per second generated, and compute per hour of the AETHER Server Runner at the small tier.

Cached input is the rate for prompt-cache hits on chat and coding models. Writing a prompt into the cache bills at 1.25x the input rate on providers that charge for it.

Changes. This page reads the live catalog, so it always shows the current rate. When a provider changes a list price, the row here changes with it.

Ready to put these models to work?

Create an agent on your own knowledge and tools, pick any model on this sheet, and see every token attributed to a business unit.