Seminal AI
§1

Llama 4 Maverick

Higher-quality sibling to Scout with the same 17B activated cost per token but 400B total weights — better reasoning and image understanding, at the price of multi-GPU hosting and a shorter 1M window.

Data checked 2026-09-06

Meta ga open weights meta-llama/Llama-4-Maverick-17B-128E-Instruct

Context window
1M tokens in
Max output
Input
per 1M tokens
Output
per 1M tokens
Knowledge cutoff
2024-08
Released
2025-04-05
Parameters
17B active / 400B total, 128 experts (MoE)
Licence
Llama 4 Community License Agreement

A What it is for

Weights-only from Meta: both prices -1.

B Capabilities

chat multilingual image-understanding early-fusion-multimodal long-context fine-tuning distillation
Input modalitiestext, image
Output modalitiestext

C Others from Meta

Model Context Max out In $/M Out $/M Blended Status
Muse Spark 1.3 muse-spark-1.3 1M $1.25 $4.25 $2.00 ga
Muse Spark 1.3 (Contributor tier) muse-spark-1.3-contributor 1M $0.10 $0.20 $0.13 ga
Muse Spark 1.2 muse-spark-1.2 1M $1.25 $4.25 $2.00 ga
Muse Spark 1.2 (Contributor tier) muse-spark-1.2-contributor 1M $0.10 $0.20 $0.13 ga
Muse Spark 1.1 muse-spark-1.1 1M $1.25 $4.25 $2.00 ga
Muse Image 1.0 muse-image-1.0 ga
Muse Voice Transcribe 1.0 muse-voice-transcribe-1.0 ga
Muse Glimmer 30B meta-models/Muse-Glimmer-30B 131.1K ga open
Llama 4 Scout meta-llama/Llama-4-Scout-17B-16E-Instruct 10M ga open
Llama 3.3 70B Instruct meta-llama/Llama-3.3-70B-Instruct 128K ga open
Llama 3.2 90B Vision Instruct meta-llama/Llama-3.2-90B-Vision-Instruct 128K ga open
Llama 3.2 11B Vision Instruct meta-llama/Llama-3.2-11B-Vision-Instruct 128K ga open

D Verify before you commit

Model pricing changes without notice and this page is a snapshot. Confirm against the vendor's own page before you build a budget on it.

Source: github.com · checked 2026-09-06 · Meta pricing · API docs