Models / Nova Micro 1.0

Amazon

Nova Micro 1.0

16 published results from 3 sources. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
Amazon
Sources
3
Our benchmarks
0
Price
Not yet published

Reported by others

1,233
Business, management and finance · rank 300 of 402
Unit
Arena rating, higher is better
Range
1,220 to 1,247
Sample
2061 votes
Configuration
Nova Micro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,197
Creative writing · rank 306 of 407
Unit
Arena rating, higher is better
Range
1,187 to 1,208
Sample
3061 votes
Configuration
Nova Micro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,244
Expert prompts · rank 273 of 359
Unit
Arena rating, higher is better
Range
1,227 to 1,261
Sample
1094 votes
Configuration
Nova Micro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,215
Instruction following · rank 317 of 409
Unit
Arena rating, higher is better
Range
1,208 to 1,222
Sample
7716 votes
Configuration
Nova Micro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,241
Overall · rank 314 of 409
Unit
Arena rating, higher is better
Range
1,236 to 1,246
Sample
19364 votes
Configuration
Nova Micro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,223
Writing, literature and language · rank 307 of 408
Unit
Arena rating, higher is better
Range
1,214 to 1,231
Sample
5246 votes
Configuration
Nova Micro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
70.7%
Irrelevance detection · rank 82 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not general refusal or safety behaviour.
2.4%
Memory · rank 100 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not long-term personal memory in a product: sessions are BFCL's scripted ones.
1.4%
Multi-turn tasks · rank 98 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not open-ended agent work: the tools and tasks are BFCL's simulated APIs.
22.3%
Overall accuracy · rank 95 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not a neutral average: the weighting is BFCL's. Not reliability on your own tools: BFCL's functions and queries are a fixed test set, and a correct call is judged by its form, not by what it achieved.
81.2%
Relevance detection · rank 41 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not whether the call itself was right; only that one was attempted.
74.1%
Single-turn calls (curated) · rank 80 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not reliability on your own tools: BFCL's functions and queries are a fixed test set, and a correct call is judged by its form, not by what it achieved.
66.3%
Single-turn calls (user-contributed) · rank 75 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not reliability on your own tools: BFCL's functions and queries are a fixed test set, and a correct call is judged by its form, not by what it achieved.
1.5%
Web search · rank 77 of 109
Unit
% correct, higher is better
Configuration
nova-micro-v1
Measured
16 Dec 2025
Not shown
Not general research quality: questions have short, checkable answers.
100.0%
Answer rate · rank 1 of 108
Unit
% of documents, higher is better
Configuration
Nova Micro 1.0
Measured
22 Sep 2026
Not shown
Not a quality score: a low rate usually means content filters were triggered, and hallucination rates are measured on answered documents only.
5.5%
Hallucination rate · rank 18 of 108
Unit
% of summaries, lower is better
Configuration
Nova Micro 1.0
Measured
22 Sep 2026
Not shown
Not errors in open questions or other tasks: only summarisation, judged by Vectara's own model (HHEM-2.3), not by people, on news-style documents rather than your data.

Compare Nova Micro 1.0 with