Deepseek

FW-DeepSeek-V3.1

DeepSeek V3.1 is a 685B-parameter Mixture-of-Experts model with dual-mode thinking and non-thinking chat, featuring two-phase long context extension up to 163K tokens.
Fireworks
Version: 1

Models available for use with Fireworks on Foundry deliver optimized, best-in-class performance on the Fireworks Inference Cloud. Fireworks on Foundry is a Non-Microsoft Product. The following terms apply to a Customer's use of Fireworks on Foundry: When you use Fireworks on Foundry, data is shared between Microsoft and Fireworks AI, Customer Data will be sent outside of Microsoft systems, Customer Data will not be processed pursuant to any Foundry data residency documentation, and different compliance and data handling rules will apply. See Trust Center - Fireworks AI for details. Customers are responsible for evaluating whether data sharing between Microsoft and Fireworks is appropriate for their organization's compliance requirements.

About this model

DeepSeek-V3.1 is post-trained on the top of DeepSeek-V3.1-Base, which is built upon the original V3 base checkpoint through a two-phase long context extension approach, following the methodology outlined in the original DeepSeek-V3 report. We have expanded our dataset by collecting additional long documents and substantially extending both training phases. The 32K extension phase has been increased 10-fold to 630B tokens, while the 128K extension phase has been extended by 3.3x to 209B tokens. Additionally, DeepSeek-V3.1 is trained using the UE8M0 FP8 scale data format to ensure compatibility with microscaling data formats.

Key model capabilities

  • Dual-mode architecture with "thinking" and "non-thinking" chat modes for both fast inference and complex agentic behaviors
  • Two-phase long context extension with 32K extension (630B tokens) and 128K extension (209B tokens)
  • Trained using the UE8M0 FP8 scale data format for compatibility with microscaling data formats
  • Function calling support including custom tools, code agents, search agents, and multi-turn tool use

Quick facts

Model providerFireworks
TypeChat completion
LifecycleGenerally available (GA)
Input typetext
Output typetext
Context window163.84k