Models / GPT-6.1 Sol

OpenAI

GPT-6.1 Sol

28 published results from 1 source. Each card shows where the number comes from and what it does not measure. Results are never combined into one score.

Provider
OpenAI
Sources
1
Our benchmarks
1
Price
Not yet published

Measured by Spring Prompt

Measured by Spring Prompt · CatalogBench
0.0%
Channel rules broken · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
98.8%
Content quality · rank 2 of 18
Unit
% of checks, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.0%
Failed outputs · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not a provider outage: provider errors are retried before a product counts as failed.
Measured by Spring Prompt · CatalogBench
0.0%
Missing UK information · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
5.4%
Not findable · rank 9 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
72.0%
Publish-ready listings · rank 3 of 18
Unit
% of products, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
67.9%
Reliably publish-ready · rank 1 of 18
Unit
% of products, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
3.6%
Unsupported claims · rank 2 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.07
Unsupported claims · rank 1 of 18
Unit
claims per product, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
19.1%
Wrong attributes · rank 9 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
1.8%
Wrong category or variant · rank 3 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
100.0%
Channel compliance · rank 1 of 18
Unit
% of products, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.0%
Channel rules broken · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
100.0%
Conflicts caught · rank 1 of 18
Unit
% of conflicts, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
98.6%
Content quality · rank 1 of 18
Unit
% of checks, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
$0.0084
Cost per product · rank 7 of 18
Unit
US dollars, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your cost: prices are those charged through OpenRouter on the run date.
Measured by Spring Prompt · CatalogBench
98.6%
Decision accuracy · rank 1 of 18
Unit
% of decisions, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.0%
Failed outputs · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not a provider outage: provider errors are retried before a product counts as failed.
Measured by Spring Prompt · CatalogBench
94.4%
Field accuracy · rank 3 of 18
Unit
% of missing fields, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
2.2%
Invented values · rank 13 of 18
Unit
% of filled values, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.0%
Missing UK information · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
4.8%
Not findable · rank 7 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
77.4%
Publish-ready listings · rank 1 of 18
Unit
% of products, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
73.2%
Reliably publish-ready · rank 2 of 18
Unit
% of products, higher is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.6%
Unsupported claims · rank 1 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
0.04
Unsupported claims · rank 2 of 18
Unit
claims per product, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
16.7%
Wrong attributes · rank 8 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.
Measured by Spring Prompt · CatalogBench
1.8%
Wrong category or variant · rank 2 of 18
Unit
% of products, lower is better
Configuration
GPT-6.1 Sol, provider default reasoning
Measured
30 Sep 2026
Not shown
Not your catalogue: invented products with generated images, so a model's result on your own feed can differ.