Models / phi-4-mini-instruct

Microsoft

phi-4-mini-instruct

2 published results from 1 source. Each card shows where the number comes from and what it does not measure. The overall leaderboard combines them; here each stands alone.

Provider
Microsoft
Sources
1
Our benchmarks
0
Price per million tokens
Not listed on OpenRouter

Reported by others

92.5%
Answer rate · rank 99 of 108
Unit
% of documents, higher is better
Configuration
phi-4-mini-instruct
Measured
22 Sep 2026
Not shown
Not a quality score: a low rate usually means content filters were triggered, and hallucination rates are measured on answered documents only.
23.5%
Hallucination rate · rank 107 of 108
Unit
% of summaries, lower is better
Configuration
phi-4-mini-instruct
Measured
22 Sep 2026
Not shown
Not errors in open questions or other tasks: only summarisation, judged by Vectara's own model (HHEM-2.3), not by people, on news-style documents rather than your data.