Llama 4 Scout
Meta's Llama 4 Scout MoE (17B active / 109B total, 16 experts) for multimodal chat and long-context work, hosted on Together AI and AWS Bedrock.
Source: Llama 4 Scout on Together AI, Llama 4 Maverick on Together AI, Llama 4 Scout on Hugging Face, AWS Bedrock Llama pricing
Specs
Details that do not fit the rate card above.
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate |
|---|---|
| Input | $0.18 / M |
| Output | $0.59 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Compare providers
The same model, priced across every platform that serves it. Lowest combined in/out rate is flagged.
| Platform | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|
Together AI | $0.18 | $0.59 | — | — | Direct |
DeepInfraLowest | $0.1 | $0.3 | — | — | via OpenRouter |
Groq | $0.11 | $0.34 | $0.055 | — | via OpenRouter |
Novita | $0.18 | $0.59 | — | — | via OpenRouter |
AWS Bedrock | $0.17 | $0.66 | — | ✓ | Direct |
AWS Bedrock · Regional | from$0.17 | $0.66 | — | ✓ | Direct |
Google | $0.25 | $0.7 | — | — | via OpenRouter |
Regional deployments are priced per region; the lowest available region rate is shown.
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| May 6, 2026 – July 10, 2026 | $0.18 / $0.35 | past |
| July 10, 2026 | $0.18 / $0.59 | current |
More from Meta
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Llama 4 Scout call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.