Every model.
One price list.
These are the rates you pay on AETHER, platform fee included. No per-seat licence and no minimum: credits cover exactly the tokens your agents and developers use.
OpenAI
9 models
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
GPT-5.4
|
$3.00R51.00 | $0.30R5.10 | $18.00R306.00 |
|
GPT-5.4 mini
|
$0.90R15.30 | $0.09R1.53 | $5.40R91.80 |
|
GPT-5.4 nano
|
$0.24R4.08 | $0.024R0.41 | $1.50R25.50 |
|
GPT-5.6 Luna
|
$0.24R4.08 | $0.024R0.41 | $1.44R24.48 |
|
GPT-5.6 Sol
|
$6.00R102.00 | $0.60R10.20 | $36.00R612.00 |
|
GPT-5.6 Terra
|
$3.00R51.00 | $0.30R5.10 | $18.00R306.00 |
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
GPT-5.3 Codex
|
$2.10R35.70 | $0.21R3.57 | $16.80R285.60 |
| Model | Embeddings per 1M tokens |
|---|---|
|
Text Embedding 3 Large
|
$0.156R2.65 |
|
Text Embedding 3 Small
|
$0.024R0.41 |
Anthropic
4 models
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
Claude Fable 5
|
$12.00R204.00 | $1.20R20.40 | $60.00R1,020.00 |
|
Claude Haiku 4.5
|
$1.20R20.40 | $0.12R2.04 | $6.00R102.00 |
|
Claude Opus 4.8
|
$6.00R102.00 | $0.60R10.20 | $30.00R510.00 |
|
Claude Sonnet 5
|
$2.40R40.80 | $0.24R4.08 | $12.00R204.00 |
Google Gemini
4 models
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
Gemini 3.6 Flash
|
$1.80R30.60 | $0.18R3.06 | $9.00R153.00 |
|
Gemini 3.7 Flash
|
$0.90R15.30 | $0.09R1.53 | $4.50R76.50 |
| Model | Input per 1M tokens | Output per 1M tokens | Audio input per 1M tokens | Audio output per 1M tokens |
|---|---|---|---|---|
|
Gemini 3.1 Flash Live
|
$0.90R15.30 | $5.40R91.80 | $3.60R61.20 | $14.40R244.80 |
| Model | Input per 1M tokens | Output per 1M tokens | Video per second |
|---|---|---|---|
|
Gemini Omni Flash (Preview)
|
$1.80R30.60 | $10.80R183.60 | $0.1216R2.07 |
DeepSeek
1 model
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
DeepSeek V4.1 Flash
|
$0.18R3.06 | $0.0036R0.06 | $0.72R12.24 |
Z.ai
1 model
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
GLM 5.3 Flash
|
$0.18R3.06 | $0.036R0.61 | $0.60R10.20 |
AWS Bedrock
1 model
| Model | Input per 1M tokens | Cached input per 1M tokens | Output per 1M tokens |
|---|---|---|---|
|
Claude Haiku 4.5 (AWS Bedrock)
|
$1.20R20.40 | $0.12R2.04 | $6.00R102.00 |
Deepgram
1 model
| Model | Audio per minute |
|---|---|
|
Deepgram Nova-2 (STT)
|
$0.0072R0.12 |
ElevenLabs
1 model
| Model | Audio per minute |
|---|---|
|
ElevenLabs v3 (TTS)
|
$0.144R2.45 |
AETHER
1 model
| Model | Compute per hour |
|---|---|
|
AETHER Server Runner
|
$0.1598R2.72 |
What the numbers include. Every rate is the price billed to your credits on AETHER, with the platform fee already applied. There is no separate provider invoice and nothing added at checkout.
Currencies. Billing is in USD. The ZAR view converts at the fixed R17.00 per $1 rate used for credit top-ups, rounded to the cent.
Units. Token rates are per million tokens. Voice is per minute of audio, video per second generated, and compute per hour of the AETHER Server Runner at the small tier.
Cached input is the rate for prompt-cache hits on chat and coding models. Writing a prompt into the cache bills at 1.25x the input rate on providers that charge for it.
Changes. This page reads the live catalog, so it always shows the current rate. When a provider changes a list price, the row here changes with it.
Ready to put these models to work?
Create an agent on your own knowledge and tools, pick any model on this sheet, and see every token attributed to a business unit.