Skip to main content
Microsoft Foundry
Llama-4-Scout-17B-16E-Instruct

Llama-4-Scout-17B-16E-Instruct

Llama 4 Scout 17B 16E Instruct is great at multi-document summarization, parsing extensive user activity for personalized tasks, and reasoning over vast codebases.
Meta
Version: 4

Models from Microsoft, Partners, and Community models are a select portfolio of curated models both general-purpose and niche models across diverse scenarios by developed by Microsoft teams, partners, and community contributors

  • Managed by Microsoft: Purchase and manage models directly through Azure with a single license, world class support and enterprise grade Azure infrastructure
  • Validated by providers: Each model is validated and maintained by its respective provider, with Azure offering integration and deployment guidance.
  • Innovation and agility: Combines Microsoft research models with rapid, community-driven advancements.
  • Seamless Azure integration: Standard Microsoft Foundry experience, with support managed by the model provider.
  • Flexible deployment: Deployable as Managed Compute or Serverless API, based on provider preference.

Learn more about models from Microsoft, Partners, and Community

About this model

Llama 4 is intended for commercial and research use in multiple languages. Instruction tuned models are intended for assistant-like chat and visual reasoning tasks, whereas pretrained models can be adapted for natural language generation. For vision, Llama 4 models are also optimized for visual recognition, image reasoning, captioning, and answering general questions about an image. The Llama 4 model collection also supports the ability to leverage the outputs of its models to improve other models including synthetic data generation and distillation.

Key model capabilities

  • Multilingual text processing in Arabic, English, French, German, Hindi, Indonesian, Italian, Portuguese, Spanish, Tagalog, Thai, and Vietnamese
  • Visual recognition and image reasoning
  • Image captioning and answering general questions about an image
  • Assistant-like chat capabilities
  • Natural language generation
  • Code generation
  • Synthetic data generation and distillation
  • Long context processing (up to 10M tokens for Scout, 1M tokens for Maverick)
  • Multi-image understanding (tested up to 5 input images)

Quick facts

Model providerMeta
TypeChat completion
LifecyclePreview
Input typetext, image
Output typetext
Context window10000k
Token limits4096 output