Pricing calculator
Pick a model, set your volume, and see the bill per day, week or month — with what Fireworks, Together AI and Nebius would charge for the same work. Every rate is the committed rate card; rival figures are dated snapshots.
Save $131,000/mo
on deepseek-v4-pro vs Fireworks
12%
vs Fireworks
12%
vs Together AI
Pick the model you use most
Choose a period, then a volume — or type your own
What share of your tokens are input rather than output
Based on 500 billion tokens/mo · 350 billion input, 150 billion output
| Input / 1M | Output / 1M | Monthly | |
|---|---|---|---|
| Run BiOS | $1.40 | $3.40 | $1,000,000 |
| Fireworks | $1.74 | $3.48 | $1,131,000 |
| Together AI | $1.74 | $3.48 | $1,131,000 |
Fireworks Standard tier, as published 26 July 2026 · Together AI serverless, as published 26 July 2026.
| Model | Context | Input / 1M | Cached / 1M | Output / 1M |
|---|---|---|---|---|
| bios-adaptive | 1M | $0.14 – $1.25 | $0.14 | $0.28 – $3.95 |
| claude-opus-5 | 1M | $5.00 | $0.50 | $25.00 |
| claude-sonnet-5 | 1M | $2.00 | $0.20 | $10.00 |
| deepseek-v4-flash | 128K | $0.10 | $0.01 | $0.25 |
| deepseek-v4-pro | 128K | $1.40 | $0.14 | $3.40 |
| glm-5.2 | 1M | $1.40 | $0.14 | $3.40 |
| kimi-k2.7-code | 128K | $0.80 | $0.08 | $3.40 |
| minimax-m3 | 128K | $0.30 | $0.03 | $1.20 |
| qwen3.5-397b-a17b | 256K | $0.50 | $0.05 | $3.40 |
USD per 1M tokens, as of 17 August 2026. Run BiOS Adaptive shows a range because its rate follows the model each request lands on; pinned models show their floor. Prompt caching is billed separately. The exact rate for the model you are about to call is shown in the dashboard before you send a request — treat that as authoritative over this table.
Neither product can run up a debt. Both draw from one pre-paid wallet. If the balance reaches zero you get a grace warning, and a running endpoint pauses rather than accruing charges. Top up and resume.