Models / gpt-3.5-turbo-1106
OpenAIgpt-3.5-turbo-1106
6 published results from 1 source. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.
- Provider
- OpenAI
- Sources
- 1
- Our benchmarks
- 0
- Price
- Not yet published
Reported by others
1,150
Business, management and finance · rank 337 of 402
- Unit
- Arena rating, higher is better
- Range
- 1,131 to 1,169
- Sample
- 1382 votes
- Configuration
- gpt-3.5-turbo-1106
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,157
Creative writing · rank 330 of 407
- Unit
- Arena rating, higher is better
- Range
- 1,141 to 1,172
- Sample
- 2828 votes
- Configuration
- gpt-3.5-turbo-1106
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,215
Expert prompts · rank 289 of 359
- Unit
- Arena rating, higher is better
- Range
- 1,187 to 1,243
- Sample
- 437 votes
- Configuration
- gpt-3.5-turbo-1106
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,198
Instruction following · rank 319 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,186 to 1,210
- Sample
- 5238 votes
- Configuration
- gpt-3.5-turbo-1106
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,204
Overall · rank 332 of 409
- Unit
- Arena rating, higher is better
- Range
- 1,195 to 1,213
- Sample
- 16619 votes
- Configuration
- gpt-3.5-turbo-1106
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,193
Writing, literature and language · rank 320 of 408
- Unit
- Arena rating, higher is better
- Range
- 1,179 to 1,206
- Sample
- 3932 votes
- Configuration
- gpt-3.5-turbo-1106
- Measured
- 25 Sep 2026
- Not shown
- Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.