Models / olmo-7b-instruct
Ai2olmo-7b-instruct
5 published results from 1 source. Each card shows where the number comes from and what it does not measure. The overall leaderboard combines them; here each stands alone.
- Provider
- Ai2
- Sources
- 1
- Our benchmarks
- 0
- Price per million tokens
- Not listed on OpenRouter
Reported by others
1,086
Business, management and finance · rank 367 of 402
- Unit
- Arena rating, higher is better
- Range
- 1,060 to 1,113
- Sample
- 548 votes
- Configuration
- olmo-7b-instruct
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,002
Creative writing · rank 398 of 407
- Unit
- Arena rating, higher is better
- Range
- 981 to 1,022
- Sample
- 1019 votes
- Configuration
- olmo-7b-instruct
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,029
Instruction following · rank 395 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,013 to 1,045
- Sample
- 1908 votes
- Configuration
- olmo-7b-instruct
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,074
Overall · rank 394 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,062 to 1,085
- Sample
- 6328 votes
- Configuration
- olmo-7b-instruct
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,023
Writing, literature and language · rank 394 of 408
- Unit
- Arena rating, higher is better
- Range
- 1,004 to 1,041
- Sample
- 1511 votes
- Configuration
- olmo-7b-instruct
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.