
MODELS
by Meta
Open-weight, natively multimodal mixture-of-experts Llama 4 model with 17B active parameters, 128 experts, and a 1M-token context window.
Llama 4 Maverick (17B-128E) is Meta’s open-weight, natively multimodal mixture-of-experts model. Meta describes it as having 17B active parameters, 128 experts, and 400B total parameters. The official model card describes multilingual text-and-image input, text-and-code output, and a 1M-token context window. Hosted documentation from AWS describes multimodal inference, streaming, guardrails, and tool calling for the Maverick Instruct model.
Use for applications that need to process text and images and return text or code. Use for workflows that benefit from a long up-to-1M-token context window. Evaluate for agentic applications requiring tool calling streaming or structured interactions through a supported hosted provider.
Last updated Aug 11, 2026