Models / llama-3-8b-instruct

Meta

llama-3-8b-instruct

6 published results from 1 source. Each card shows where the number comes from and what it does not measure. The overall leaderboard combines them; here each stands alone.

Provider
Meta
Sources
1
Our benchmarks
0
Price
Not yet published

Reported by others

1,205
Business, management and finance · rank 314 of 402
Unit
Arena rating, higher is better
Range
1,196 to 1,214
Sample
10764 votes
Configuration
llama-3-8b-instruct
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,197
Creative writing · rank 307 of 407
Unit
Arena rating, higher is better
Range
1,188 to 1,205
Sample
15365 votes
Configuration
llama-3-8b-instruct
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,211
Expert prompts · rank 301 of 359
Unit
Arena rating, higher is better
Range
1,199 to 1,223
Sample
5360 votes
Configuration
llama-3-8b-instruct
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,193
Instruction following · rank 326 of 409
Unit
Arena rating, higher is better
Range
1,187 to 1,198
Sample
37733 votes
Configuration
llama-3-8b-instruct
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,224
Overall · rank 323 of 409
Unit
Arena rating, higher is better
Range
1,220 to 1,227
Sample
104642 votes
Configuration
llama-3-8b-instruct
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,192
Writing, literature and language · rank 326 of 408
Unit
Arena rating, higher is better
Range
1,185 to 1,199
Sample
25481 votes
Configuration
llama-3-8b-instruct
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.