Catalog
Model catalog
Every model AgentCost can bill, with live per-token pricing and announced retirement dates.
Total Models
2,460
Providers
61
Announced Retirements
480
Last Updated
Sep 4, 2026
Announced retirements
Dates published by the providers, earliest first. Plan migrations before a model disappears from your bill.
| Model | Provider | Retires on | When |
|---|---|---|---|
| aws | Sep 10, 2026 | in 4 days | |
| aws | Sep 10, 2026 | in 4 days | |
| aws | Sep 10, 2026 | in 4 days | |
| aws | Sep 10, 2026 | in 4 days | |
| aws | Sep 10, 2026 | in 4 days | |
| aws | Sep 10, 2026 | in 4 days |
Showing 2,460 of 2,460 models
| Model Name | Provider | Input / 1K | Output / 1K | Cached In / 1K | Type |
|---|---|---|---|---|---|
@cf/aisingapore/gemma-sea-lion-v4-27b-it | $0.000351 | $0.000555 | — | chat | |
@cf/deepseek-ai/deepseek-r1-distill-qwen-32b | $0.000497 | $0.004881 | — | chat | |
@cf/google/gemma-4-26b-a4b-it | $1.00e-4 | $0.000300 | — | chat | |
@cf/ibm-granite/granite-4.0-h-micro | $1.70e-5 | $0.000112 | — | chat | |
@cf/meta/llama-2-7b-chat-fp16 | $0.001923 | $0.001923 | — | chat | |
@cf/meta/llama-2-7b-chat-int8 | $0.001923 | $0.001923 | — | chat | |
@cf/meta/llama-3.1-8b-instruct-fp8 | $0.000152 | $0.000287 | — | chat | |
@cf/meta/llama-3.2-11b-vision-instruct | $4.85e-5 | $0.000676 | — | chat | |
@cf/meta/llama-3.2-1b-instruct | $2.70e-5 | $0.000201 | — | chat | |
@cf/meta/llama-3.2-3b-instruct | $5.09e-5 | $0.000335 | — | chat | |
@cf/meta/llama-3.3-70b-instruct-fp8-fast | $0.000293 | $0.002253 | — | chat | |
@cf/meta/llama-4-scout-17b-16e-instruct | $0.000270 | $0.000850 | — | chat | |
@cf/meta/llama-guard-3-8b | $0.000484 | $3.00e-5 | — | chat | |
@cf/mistral/mistral-7b-instruct-v0.1 | $0.001923 | $0.001923 | — | chat | |
@cf/mistralai/mistral-small-3.1-24b-instruct | $0.000351 | $0.000555 | — | chat | |
@cf/moonshotai/kimi-k2.6 | $0.000950 | $0.004000 | $0.000160 | chat | |
@cf/moonshotai/kimi-k2.7-code | $0.000950 | $0.004000 | $0.000190 | chat | |
@cf/nvidia/nemotron-3-120b-a12b | $0.000500 | $0.001500 | — | chat | |
@cf/openai/gpt-oss-120b | $0.000350 | $0.000750 | — | chat | |
@cf/openai/gpt-oss-20b | $0.000200 | $0.000300 | — | chat | |
@cf/qwen/qwen2.5-coder-32b-instruct | $0.000660 | $0.001000 | — | chat | |
@cf/qwen/qwen3-30b-a3b-fp8 | $5.09e-5 | $0.000335 | — | chat | |
@cf/qwen/qwq-32b | $0.000660 | $0.001000 | — | chat | |
@cf/zai-org/glm-4.7-flash | $6.05e-5 | $0.000400 | — | chat | |
@cf/zai-org/glm-5.2 | $0.001400 | $0.004400 | $0.000260 | chat | |
@hf/thebloke/codellama-7b-instruct-awq | $0.001923 | $0.001923 | — | chat | |
accounts/fireworks/models/ | $1.00e-4 | Free | — | embedding | |
accounts/fireworks/models/chronos-hermes-13b-v2 | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-llama-13b | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-llama-13b-instruct | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-llama-13b-python | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-llama-34b | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/code-llama-34b-instruct | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/code-llama-34b-python | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/code-llama-70b | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/code-llama-70b-instruct | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/code-llama-70b-python | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/code-llama-7b | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-llama-7b-instruct | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-llama-7b-python | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/code-qwen-1p5-7b | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/codegemma-2b | $1.00e-4 | $1.00e-4 | — | chat | |
accounts/fireworks/models/codegemma-7b | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/cogito-671b-v2-p1 | $0.001200 | $0.001200 | — | chat | |
accounts/fireworks/models/cogito-v1-preview-llama-3b | $1.00e-4 | $1.00e-4 | — | chat | |
accounts/fireworks/models/cogito-v1-preview-llama-70b | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/cogito-v1-preview-llama-8b | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/cogito-v1-preview-qwen-14b | $0.000200 | $0.000200 | — | chat | |
accounts/fireworks/models/cogito-v1-preview-qwen-32b | $0.000900 | $0.000900 | — | chat | |
accounts/fireworks/models/dbrx-instruct | $0.001200 | $0.001200 | — | chat |
Page 1 of 50 (2,460 results)
1 / 50
Pricing data sourced from LiteLLM. Prices are per 1,000 tokens in USD.