Phi-4-reasoning-plus

State-of-the-art open-weight reasoning model.

Microsoft

Version: 1

Phi-4-reasoning-plus is a state-of-the-art open-weight reasoning model finetuned from Phi-4 using supervised fine-tuning on a dataset of chain-of-thought traces and reinforcement learning. The supervised fine-tuning dataset includes a blend of synthetic prompts and high-quality filtered data from public domain websites, focused on math, science, and coding skills as well as alignment data for safety and Responsible AI. The goal of this approach was to ensure that small capable models were trained with data focused on high quality and advanced reasoning. Phi-4-reasoning-plus has been trained additionally with Reinforcement Learning, hence, it has higher accuracy and higher latency since it generates on average 50% more tokens than Phi-4-reasoning.

Quick facts

Model providerMicrosoft

TypeChat completion

LifecyclePreview

Input typetext

Output typetext

Context window32768

Token limits4096 output

Phi-4-reasoning-plus

Quick facts

Quick start