Home / Pricing / Providers / Fireworks

Fireworks API pricing

Published Fireworks rates in this catalog. Example cost uses 1000 input tokens and 500 output tokens. Open any model in the calculator for your own prompt.

Fireworks rates

#ModelProviderContextInput / 1MOutput / 1MExample10k requests
1Llama 3.2 3B (Fireworks)Fireworks128K$0.100/1M$0.100/1M$0.00015$1.5Cost
2Llama 3.1 8B (Fireworks)Fireworks128K$0.200/1M$0.200/1M$0.0003$3.00Cost
3MythoMax L2 13B (Fireworks)Fireworks4K$0.200/1M$0.200/1M$0.0003$3.00Cost
4Llama 4 Maverick (Fireworks)Fireworks1M$0.220/1M$0.880/1M$0.00066$6.6Cost
5DeepSeek R1 (Fireworks)Fireworks128K$0.900/1M$0.900/1M$0.00135$13.5Cost
6DeepSeek V3 (Fireworks)Fireworks128K$0.900/1M$0.900/1M$0.00135$13.5Cost
7Llama 3.1 70B (Fireworks)Fireworks128K$0.900/1M$0.900/1M$0.00135$13.5Cost
8Llama 3.3 70B (Fireworks)Fireworks128K$0.900/1M$0.900/1M$0.00135$13.5Cost
9Mixtral 8x22B (Fireworks)Fireworks66K$0.900/1M$0.900/1M$0.00135$13.5Cost
10Qwen2.5 32B (Fireworks)Fireworks33K$0.900/1M$0.900/1M$0.00135$13.5Cost
11Qwen2.5 72B (Fireworks)Fireworks33K$0.900/1M$0.900/1M$0.00135$13.5Cost

Use the calculator

Open a model in the cost calculator or tokenizer. Rates checked on each model page.

Related tools

LLM cost calculator · Cheapest LLM API · Compare model prices · Cache savings · Batch pricing

Related guides

How LLM API pricing works · How prompt cost is calculated · Prompt caching explained · Exact vs Approx

Sources and references

Official documentation used for definitions, counting methods, or rate cards. Always confirm critical budgets on the provider page.