Models / qwen1.5-7b-chat

Alibaba

qwen1.5-7b-chat

6 published results from 1 source. Each card shows where the number comes from and what it does not measure. The overall leaderboard combines them; here each stands alone.

Provider
Alibaba
Sources
1
Our benchmarks
0
Price per million tokens
Not listed on OpenRouter

Reported by others

1,148
Business, management and finance · rank 336 of 402
Unit
Arena rating, higher is better
Range
1,120 to 1,176
Sample
447 votes
Configuration
qwen1.5-7b-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,081
Creative writing · rank 367 of 407
Unit
Arena rating, higher is better
Range
1,058 to 1,104
Sample
701 votes
Configuration
qwen1.5-7b-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,148
Expert prompts · rank 324 of 359
Unit
Arena rating, higher is better
Range
1,112 to 1,185
Sample
231 votes
Configuration
qwen1.5-7b-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,124
Instruction following · rank 357 of 409
Unit
Arena rating, higher is better
Range
1,110 to 1,138
Sample
1715 votes
Configuration
qwen1.5-7b-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,144
Overall · rank 363 of 409
Unit
Arena rating, higher is better
Range
1,134 to 1,154
Sample
4737 votes
Configuration
qwen1.5-7b-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,113
Writing, literature and language · rank 358 of 408
Unit
Arena rating, higher is better
Range
1,095 to 1,131
Sample
1093 votes
Configuration
qwen1.5-7b-chat
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.