Skip to main content
Microsoft Foundry
Cohere-parse-v5

Cohere-parse-v5

Cohere Parse 5 is a high-performance document vision parser.
Cohere
Direct from Azure
Version: 1

Direct from Azure models are a select portfolio curated for their market-differentiated capabilities:

  • Secure and managed by Microsoft: Purchase and manage models directly through Azure with a single license, consistent support, and no third-party dependencies, backed by Azure's enterprise-grade infrastructure.
  • Streamlined operations: Benefit from unified billing, governance, and seamless PTU portability across models hosted on Azure - all as part of one Azure AI Foundry platform.
  • Future-ready flexibility: Access the latest models as they become available, and easily test, deploy, or switch between them within Azure AI Foundry; reducing integration effort.
  • Cost control and optimization: Scale on demand with pay-as-you-go flexibility or reserve PTUs for predictable performance and savings.

Learn more about Direct from Azure models .

About this model

Cohere’s Parse 5 is a high-performance document vision parser that delivers market-leading document understanding at its size and price range. The gateway to enterprise document intelligence, Parse 5 transforms complex, multilingual documents into structured, machine-readable content for high-throughput AI workloads. More than text recognition, Parse 5 detects and understands key visual elements - such as tables and embedded images - and returns descriptions of visual elements in addition to bounding boxes.

Key model capabilities

  • Best-in-class value: Outperforms leading document parsers and hyperscaler services while remaining cost-effective at enterprise scale.
  • Beyond OCR: Understands tables, forms, diagrams, and images to extract richer semantic context across major global commercial languages.
  • Structured table output: Automatically detects tables and outputs structured HTML, including complicated cases like rotated tables, merged cells, and more.
  • Spatially aware: Returns bounding boxes alongside extracted visual elements, preserving document structure for retrieval, grounding, and downstream automation.

Quick facts

Model providerCohere
TypeText classification
LifecyclePreview
Hosted onAzure
Input typeimage
Output typetext
Context window8192
Token limits64000 output