Models / Mistral Medium 3.1

Mistral

Mistral Medium 3.1

12 published results from 3 sources. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
Mistral
Sources
3
Our benchmarks
0
Price
Not yet published

Reported by others

1,429
Overall · rank 95 of 177
Unit
Arena rating, higher is better
Range
1,426 to 1,431
Sample
68578 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not an error rate: claims that cannot be checked on the web are skipped, and preference still carries most of the weight.
1,407
Business, management and finance · rank 117 of 402
Unit
Arena rating, higher is better
Range
1,402 to 1,412
Sample
17840 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,378
Creative writing · rank 114 of 407
Unit
Arena rating, higher is better
Range
1,372 to 1,384
Sample
13967 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,410
Expert prompts · rank 134 of 359
Unit
Arena rating, higher is better
Range
1,403 to 1,418
Sample
6528 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,395
Instruction following · rank 127 of 409
Unit
Arena rating, higher is better
Range
1,391 to 1,399
Sample
27541 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,408
Overall · rank 133 of 409
Unit
Arena rating, higher is better
Range
1,405 to 1,411
Sample
95010 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,386
Writing, literature and language · rank 124 of 408
Unit
Arena rating, higher is better
Range
1,382 to 1,391
Sample
21836 votes
Configuration
Mistral Medium 3.1
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
Reported by UGI Leaderboard
27.0%
Requested-length error · rank 239 of 370
Unit
% off the requested word count, lower is better
Configuration
Mistral Medium 3.1
Measured
1 Oct 2025
Not shown
Not other format limits such as character counts or bullet counts.
Reported by UGI Leaderboard
0.39
Style adherence · rank 48 of 370
Unit
score from 0 to 1, higher is better
Configuration
Mistral Medium 3.1
Measured
1 Oct 2025
Not shown
Not brand voice on your own examples: UGI's prompts are private and lean towards creative writing.
Reported by UGI Leaderboard
39.5
Writing score · rank 219 of 370
Unit
score out of 100, higher is better
Configuration
Mistral Medium 3.1
Measured
1 Oct 2025
Not shown
Not business copy quality: UGI's prompts are private and lean towards creative writing, and models that often refuse get no score.
99.7%
Answer rate · rank 44 of 108
Unit
% of documents, higher is better
Configuration
Mistral Medium 3.1
Measured
22 Sep 2026
Not shown
Not a quality score: a low rate usually means content filters were triggered, and hallucination rates are measured on answered documents only.
22.7%
Hallucination rate · rank 105 of 108
Unit
% of summaries, lower is better
Configuration
Mistral Medium 3.1
Measured
22 Sep 2026
Not shown
Not errors in open questions or other tasks: only summarisation, judged by Vectara's own model (HHEM-2.3), not by people, on news-style documents rather than your data.

Compare Mistral Medium 3.1 with