Models / Nova Pro 1.0

Amazon

Nova Pro 1.0

16 published results from 3 sources. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
Amazon
Sources
3
Our benchmarks
0
Price
Not yet published

Reported by others

1,279
Business, management and finance · rank 260 of 402
Unit
Arena rating, higher is better
Range
1,268 to 1,290
Sample
2591 votes
Configuration
Nova Pro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,237
Creative writing · rank 280 of 407
Unit
Arena rating, higher is better
Range
1,227 to 1,246
Sample
3944 votes
Configuration
Nova Pro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,273
Expert prompts · rank 245 of 359
Unit
Arena rating, higher is better
Range
1,257 to 1,289
Sample
1387 votes
Configuration
Nova Pro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,277
Instruction following · rank 269 of 409
Unit
Arena rating, higher is better
Range
1,270 to 1,283
Sample
9525 votes
Configuration
Nova Pro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,290
Overall · rank 277 of 409
Unit
Arena rating, higher is better
Range
1,286 to 1,295
Sample
24745 votes
Configuration
Nova Pro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1,274
Writing, literature and language · rank 263 of 408
Unit
Arena rating, higher is better
Range
1,267 to 1,281
Sample
6609 votes
Configuration
Nova Pro 1.0
Measured
25 Sep 2026
Not shown
Not accuracy or correctness: it ranks which answer voters preferred, with answer length and formatting controlled for.
1.9%
Memory · rank 102 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not long-term personal memory in a product: sessions are BFCL's scripted ones.
1.9%
Multi-turn tasks · rank 95 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not open-ended agent work: the tools and tasks are BFCL's simulated APIs.
25.0%
Overall accuracy · rank 88 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not a neutral average: the weighting is BFCL's. Not reliability on your own tools: BFCL's functions and queries are a fixed test set, and a correct call is judged by its form, not by what it achieved.
93.8%
Relevance detection · rank 8 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not whether the call itself was right; only that one was attempted.
86.6%
Single-turn calls (curated) · rank 33 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not reliability on your own tools: BFCL's functions and queries are a fixed test set, and a correct call is judged by its form, not by what it achieved.
78.5%
Single-turn calls (user-contributed) · rank 23 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not reliability on your own tools: BFCL's functions and queries are a fixed test set, and a correct call is judged by its form, not by what it achieved.
2.5%
Web search · rank 71 of 109
Unit
% correct, higher is better
Configuration
nova-pro-v1
Measured
16 Dec 2025
Not shown
Not general research quality: questions have short, checkable answers.
99.3%
Answer rate · rank 61 of 108
Unit
% of documents, higher is better
Configuration
Nova Pro 1.0
Measured
22 Sep 2026
Not shown
Not a quality score: a low rate usually means content filters were triggered, and hallucination rates are measured on answered documents only.
5.1%
Hallucination rate · rank 10 of 108
Unit
% of summaries, lower is better
Configuration
Nova Pro 1.0
Measured
22 Sep 2026
Not shown
Not errors in open questions or other tasks: only summarisation, judged by Vectara's own model (HHEM-2.3), not by people, on news-style documents rather than your data.

Compare Nova Pro 1.0 with