Skip to main content
Microsoft Foundry
Llama-4-Scout-17B-16E

Llama-4-Scout-17B-16E

Llama 4 Scout 17B 16E is great at multi-document summarization, parsing extensive user activity for personalized tasks, and reasoning over vast codebases.
Meta
Version: 2

Key capabilities

About this model

The Llama 4 collection of models are natively multimodal AI models that enable text and multimodal experiences. These models leverage a mixture-of-experts architecture to offer industry-leading performance in text and image understanding. For vision, Llama 4 models are also optimized for visual recognition, image reasoning, captioning, and answering general questions about an image.

Key model capabilities

  • Multilingual text and image processing
  • Visual recognition, image reasoning, captioning, and answering general questions about an image
  • Natural language generation
  • Assistant-like chat and visual reasoning tasks
  • Code generation and understanding
  • Synthetic data generation and distillation
  • Long context processing (up to 10M tokens for Scout, 1M tokens for Maverick)

See Responsible AI for additional considerations for responsible use.

Key use cases

Llama 4 is intended for commercial and research use in multiple languages. Instruction tuned models are intended for assistant-like chat and visual reasoning tasks, whereas pretrained models can be adapted for natural language generation. For vision, Llama 4 models are also optimized for visual recognition, image reasoning, captioning, and answering general questions about an image. The Llama 4 model collection also supports the ability to leverage the outputs of its models to improve other models including synthetic data generation and distillation. The Llama 4 Community License allows for these use cases.

Out of scope use cases

Use in any manner that violates applicable laws or regulations (including trade compliance laws). Use in any other way that is prohibited by the Acceptable Use Policy and Llama 4 Community License. Use in languages or capabilities beyond those explicitly referenced as supported in this model card.

Pricing is based on a number of factors, including deployment type and tokens used. See pricing details here.

Quick facts

PublisherMeta
AuthorMeta
TypeChat completion
LifecyclePreview
Input typetext, image
Output typetext
Context window10000k
Token limits4096 output