Llama 4 Maverick
Open weightVisionToolsMeta · Released April 2025 · Knowledge cutoff August 2024
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
It belongs to Meta's open-weight Llama family: Llama 4 Maverick is the capability lead, Scout the efficient long-context sibling, and the 3.x line remains a dependable workhorse.
Facts and pricing
- Provider
- Meta
- Context window
- 128K tokens
- Input price
- 2.2 credits / 1M tokens ($0.20 raw)
- Output price
- 7.7 credits / 1M tokens ($0.70 raw)
- Vision (image input)
- Yes
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Llama 4 Maverick now
What Llama 4 Maverick is best for
- Everyday chat and drafting on an open-weight model with transparent lineage
- Working with images: screenshots, charts, photos and scanned documents alongside text
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a open weight model like Llama 4 Maverick:
- Draft a project update from these rough notes
- Explain how HTTPS works to a curious teenager
- Turn this list of features into a changelog entry
- Brainstorm objections to this proposal and how to answer them
Meta family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Llama Guard 4 12B | 164K | 2.0 | 2.0 | April 2025 |
| Llama 4 Maverick | 128K | 2.2 | 7.7 | April 2025 |
| Llama 4 Scout | 328K | 1.1 | 3.3 | April 2025 |
| Llama 3.3 70B Instruct | 131K | 1.1 | 3.5 | December 2024 |
| Llama 3.2 3B Instruct | 131K | 0.55 | 3.6 | September 2024 |
| Llama 3.2 1B Instruct | 60K | 0.30 | 2.2 | September 2024 |
| Llama 3.1 8B Instruct | 131K | 0.55 | 0.88 | July 2024 |
| Llama 3.1 70B Instruct | 131K | 4.4 | 4.4 | July 2024 |
Compare Llama 4 Maverick
- Llama 4 Maverick vs DeepSeek V4 Pro 0423
- Llama 4 Maverick vs Qwen3.7 Max
- Llama 4 Maverick vs Mistral Medium 3.5
- Llama 4 Maverick vs Llama 4 Scout
- Llama 4 Maverick vs Llama 3.3 70B Instruct
Frequently asked questions
- What is Llama 4 Maverick?
- Llama 4 Maverick is an AI model by Meta. Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
- How much does Llama 4 Maverick cost on Zeplik?
- Input tokens cost 2.2 credits per million and output tokens 7.7 credits per million (1 credit = $0.10; the raw provider rates are $0.20 and $0.70 per million). New accounts start with free credits.
- How long can a conversation with Llama 4 Maverick be?
- Llama 4 Maverick has a 128K-token context window (128,000 tokens), which covers the conversation plus any documents you attach.
- Does Llama 4 Maverick support images and tools?
- Llama 4 Maverick accepts image input and supports tool calling.
Related models
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
More you can do with Llama 4 Maverick
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Llama 4 Maverick or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.