microsoft/phi-4
#11 of 24 by Downloads / 30d among Model (Hugging Face)s
About
Our training data is an extension of the data used for Phi-3 and includes a wide variety of sources from:
Multilingual data constitutes about 8% of our overall data. We are focusing on the quality of data that could potentially improve the reasoning ability for the model, and we filter the publicly available documents to contain the correct level of knowledge.
We evaluated phi-4 using OpenAI’s SimpleEval and our own internal benchmarks to understand the model’s capabilities, more specifically:
phi-4 has adopted a robust safety post-training approach. This approach leverages a variety of both open-source and in-house generated synthetic datasets. The overall technique employed to do the…
Excerpted from huggingface.co/microsoft/phi-4
Latest metrics
| Downloads / 30d | 708.1k | 2026-09-12 |
|---|---|---|
| Likes | 2.3k | 2026-09-12 |