Models / Qwen3 Next 80B A3B Thinking

Alibaba

Qwen3 Next 80B A3B Thinking

11 published results from 3 sources. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
Alibaba
Sources
3
Our benchmarks
0
Price
Not yet published

Reported by others

1,367
Business, management and finance · rank 168 of 402
Unit
Arena rating, higher is better
Range
1,355 to 1,378
Sample
2478 votes
Configuration
Qwen3 Next 80B A3B Thinking
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,326
Creative writing · rank 170 of 407
Unit
Arena rating, higher is better
Range
1,312 to 1,341
Sample
1751 votes
Configuration
Qwen3 Next 80B A3B Thinking
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,394
Expert prompts · rank 137 of 359
Unit
Arena rating, higher is better
Range
1,371 to 1,417
Sample
635 votes
Configuration
Qwen3 Next 80B A3B Thinking
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,359
Instruction following · rank 174 of 409
Unit
Arena rating, higher is better
Range
1,349 to 1,369
Sample
3491 votes
Configuration
Qwen3 Next 80B A3B Thinking
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,369
Overall · rank 183 of 409
Unit
Arena rating, higher is better
Range
1,363 to 1,375
Sample
13373 votes
Configuration
Qwen3 Next 80B A3B Thinking
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,341
Writing, literature and language · rank 174 of 408
Unit
Arena rating, higher is better
Range
1,330 to 1,352
Sample
3028 votes
Configuration
Qwen3 Next 80B A3B Thinking
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
Reported by UGI Leaderboard
67.0%
Requested-length error · rank 329 of 370
Unit
% off the requested word count, lower is better
Configuration
Qwen3 Next 80B A3B Thinking
Measured
29 Nov 2025
Not shown
Not other format limits such as character counts or bullet counts.
Reported by UGI Leaderboard
0.28
Style adherence · rank 354 of 370
Unit
score from 0 to 1, higher is better
Configuration
Qwen3 Next 80B A3B Thinking
Measured
29 Nov 2025
Not shown
Not brand voice on your own examples: UGI's prompts are private and lean towards creative writing.
Reported by UGI Leaderboard
27.2
Writing score · rank 291 of 370
Unit
score out of 100, higher is better
Configuration
Qwen3 Next 80B A3B Thinking
Measured
29 Nov 2025
Not shown
Not business copy quality: UGI's prompts are private and lean towards creative writing, and models that often refuse get no score.
94.4%
Answer rate · rank 95 of 108
Unit
% of documents, higher is better
Configuration
Qwen3 Next 80B A3B Thinking
Measured
22 Sep 2026
Not shown
Not a quality score: a low rate usually means content filters were triggered, and hallucination rates are measured on answered documents only.
9.3%
Hallucination rate · rank 47 of 108
Unit
% of summaries, lower is better
Configuration
Qwen3 Next 80B A3B Thinking
Measured
22 Sep 2026
Not shown
Not errors in open questions or other tasks: only summarisation, judged by Vectara's own model (HHEM-2.3), not by people, on news-style documents rather than your data.

Compare Qwen3 Next 80B A3B Thinking with