Compare / GPT-6.1 Sol vs Inkling Small

GPT-6.1 Sol vs Inkling Small

Which is cheaper, which is more capable, and which does better at each kind of work, from 29 results on 4 sources that measured both models.

OpenAIGPT-6.1 Sol Released 29 Sep 2026 · $2.00 in · $10.00 out
Thinking MachinesInkling Small Released 15 Jul 2026 · $0.45 in · $1.20 out
GPT-6.1 Sol wins 211 too close to callInkling Small wins 7

Of the 29 results both models have, the better value on each. Results whose reported ranges overlap are too close to call.

Results as of 9 October 2026 (catalogue release 2026-10-09-dd620e778ac5).

GPT-6.1 Sol or Inkling Small? The short answer

Pick GPT-6.1 Sol if quality matters most: it scores higher (Intelligence Index 51.8 against 25.7). Pick Inkling Small if cost matters more: it costs 2.8 times less per game of bullet chess on BulletBench. Inkling Small is also faster.

  • Inkling Small is cheaper to run: $0.007 against $0.021 per game of bullet chess on BulletBench (2.8 times less).
  • GPT-6.1 Sol scores higher on the Artificial Analysis Intelligence Index: 51.8 against 25.7.
  • Inkling Small writes faster: 173 tokens a second against 55.
  • GPT-6.1 Sol is newer: released 29 Sep 2026, against 15 Jul 2026 for Inkling Small.
  • On our own benchmarks, on the results both models have: GPT-6.1 Sol does better for speed under pressure.
  • On third-party benchmarks, on the results both models have: GPT-6.1 Sol does better for agents and tool use, reasoning and knowledge, writing and sticking to the facts.

Head to head

51.8
Intelligence Index
25.7
$4.00
Price per million tokens, blendedlower is better
$0.64
55
Output speed, tokens a second
173
1,050,000
Context window, tokens
524,288

Every fact side by side

GPT-6.1 SolInkling Small
Price per million tokens$2.00 in · $10.00 out$0.45 in · $1.20 out
Blended (3 in : 1 out)$4.00$0.64
Cheapest host, blended$4.00 (Azure)–
Cost per game of bullet chess (BulletBench)$0.021$0.007
Cached input, per million$0.10 cached input–
Intelligence Index51.825.7
Spring Prompt overall81–
Output speed, tokens a second55173
Context window1,050,000 tokens524,288 tokens
Released29 Sep 2026 (newer)15 Jul 2026
WeightsClosedOpen
Developer based inthe USthe US

Prices are the developer's list price where we have read it, otherwise the typical host on OpenRouter; see every model's price. Intelligence and speed from Artificial Analysis.

Which is better for what

Each model's rank among the models on sale today, for each kind of work, from our own benchmarks and licensed sources. Further right is better; the better rank is in green.

Product listingsTurning a sparse product feed and photos into listings that can go live 2 of 20 not measured
Decks from an analysisTurning a finished analysis into a deck you could present as it is 3 of 19 not measured
User surveysPlanning a user survey and reading its results without being misled 1 of 8 not measured
Agents and tool useMulti-step tasks with tools: support desks, coding agents, function calls 3 of 164 71 of 164
Professional workReal tasks from banking, consulting and law, business documents and freelance projects 9 of 51 not measured
Reasoning and knowledgeHard questions across science, maths and general knowledge 5 of 198 61 of 198
WritingWhat people prefer in blind comparisons, and judged writing quality 15 of 188 103 of 188
Sticking to the factsSummarising without inventing things, and factual answers 18 of 152 127 of 152
Speed under pressureGood decisions against a real clock (fast chess) 3 of 23 7 of 23

Every shared result

Each model's best result on every measure both have, grouped by benchmark, with the setting that produced it. Under each measure, every model that benchmark has measured is a grey tick, best to the right, so you can see where the two sit in the field. The headline measures of each benchmark come first; the rest are under "Show all". Green marks the better value; where a source reports ranges that overlap, neither is marked.

MeasureGPT-6.1 SolInkling Small
BulletBench measured by Spring Prompt · 3 results
Cost per game · Bullet 60sUS dollars, lower is better 22 · 15 of 28 measured $0.0210 $0.0075no reasoning
Ladder Elo · Bullet 60sladder Elo, higher is better 2 · 11 of 28 measured 928Ultrafast (low reasoning) 529no reasoning
Ladder Elo · Lightning 10+1ladder Elo, higher is better 5 · 7 of 23 measured 584Ultrafast (low reasoning) 571no reasoning
Arena (formerly LMArena) reported by Arena (formerly LMArena) · 3 results
Overall · TextArena rating, higher is better 18 · 131 of 387 measured 1,483max reasoning 1,405
Overall · Text factualityArena rating, higher is better 37 · 104 of 163 measured 1,463max reasoning 1,426
Business, management and finance · TextArena rating, higher is better 27 · 125 of 380 measured 1,475max reasoning 1,409
Artificial Analysis reported by Artificial Analysis · 3 results
Artificial Analysis Intelligence Indexindex score, higher is better 6 · 70 of 243 measured 51.8max reasoning 25.7
Humanity's Last Exam% of questions, higher is better 8 · 67 of 242 measured 52.9%max reasoning 33.3%
Terminal-Bench 4.0% of tasks, higher is better 5 · 52 of 112 measured 56.1%max reasoning 1.0%
SimpleQA Verified (Epoch AI) reported by SimpleQA Verified (Epoch AI) · 1 result
Correct answers% of questions, higher is better 2 · 66 of 76 measured 73.9%max reasoning 19.1%xhigh reasoning
Show all 29 results (19 more)
MeasureGPT-6.1 SolInkling Small
BulletBench measured by Spring Prompt · 7 results
Cost per game · Lightning 10+1US dollars, lower is better 23 · 16 of 23 measured $0.16Ultrafast $0.0077no reasoning
Games lost on time · Bullet 60s% of games, lower is better 13 · 1 of 28 measured 25.0%Ultrafast (low reasoning) 0.0%no reasoning
Games lost on time · Lightning 10+1% of games, lower is better 16 · 1 of 23 measured 50.0%Ultrafast (low reasoning) 0.0%no reasoning
Invalid moves · Bullet 60s% of moves, lower is better 1 · 28 of 28 measured 0.0% 7.2%no reasoning
Invalid moves · Lightning 10+1% of moves, lower is better 1 · 23 of 23 measured 0.0%Ultrafast 5.3%no reasoning
Median move time · Bullet 60smilliseconds, lower is better 18 · 7 of 28 measured 1.4 sUltrafast (low reasoning) 0.6 sno reasoning
Median move time · Lightning 10+1milliseconds, lower is better 19 · 8 of 23 measured 1.2 sUltrafast (low reasoning) 0.6 sno reasoning
Arena (formerly LMArena) reported by Arena (formerly LMArena) · 8 results
Confirmed task success · AgentIPS effect estimate, higher is better 2 · 46 of 48 measured 0.15max reasoning, arena agent -0.20arena agent
Creative writing · TextArena rating, higher is better 21 · 192 of 385 measured 1,460max reasoning 1,311
Expert prompts · TextArena rating, higher is better 4 · 106 of 338 measured 1,543max reasoning 1,445
Instruction following · TextArena rating, higher is better 10 · 134 of 387 measured 1,490max reasoning 1,394
Praise over complaint · AgentIPS effect estimate, higher is better 6 · 46 of 48 measured 0.26max reasoning, arena agent -0.22arena agent
Steerability · AgentIPS effect estimate, higher is better 3 · 48 of 48 measured 0.11max reasoning, arena agent -0.14arena agent
Tool grounding · AgentIPS effect estimate, higher is better 1 · 40 of 48 measured 0.00max reasoning, arena agent -0.00arena agent
Writing, literature and language · TextArena rating, higher is better 17 · 175 of 386 measured 1,472max reasoning 1,342
Artificial Analysis reported by Artificial Analysis · 4 results
Long-context reasoning (AA-LCR)% of questions, higher is better 6 · 76 of 233 measured 84.0%low reasoning 75.7%
Output speedtokens per second, higher is better 58 · 15 of 74 measured 55.2max reasoning 173
SciCode% of problems, higher is better 22 · 50 of 113 measured 55.8%high reasoning 49.7%
Time to first answer tokenseconds, lower is better 36 · 49 of 74 measured 1.69low reasoning 13.9

Quick answers

Is GPT-6.1 Sol better than Inkling Small?

Pick GPT-6.1 Sol if quality matters most: it scores higher (Intelligence Index 51.8 against 25.7). Pick Inkling Small if cost matters more: it costs 2.8 times less per game of bullet chess on BulletBench. Inkling Small is also faster.

Which is cheaper, GPT-6.1 Sol or Inkling Small?

Inkling Small is cheaper to run: $0.007 against $0.021 per game of bullet chess on BulletBench (2.8 times less).

Which is faster, GPT-6.1 Sol or Inkling Small?

Inkling Small writes about 173 tokens a second on its usual API, against 55 for GPT-6.1 Sol (Artificial Analysis).

Which is better for agents and tool use, GPT-6.1 Sol or Inkling Small?

GPT-6.1 Sol ranks 3 of 164 models on sale for multi-step tasks with tools: support desks, coding agents, function calls, against 71 for Inkling Small.

Which is better for reasoning and knowledge, GPT-6.1 Sol or Inkling Small?

GPT-6.1 Sol ranks 5 of 198 models on sale for hard questions across science, maths and general knowledge, against 61 for Inkling Small.

Which is better for writing, GPT-6.1 Sol or Inkling Small?

GPT-6.1 Sol ranks 15 of 188 models on sale for what people prefer in blind comparisons, and judged writing quality, against 103 for Inkling Small.

Which is better for sticking to the facts, GPT-6.1 Sol or Inkling Small?

GPT-6.1 Sol ranks 18 of 152 models on sale for summarising without inventing things, and factual answers, against 127 for Inkling Small.

Which has the bigger context window, GPT-6.1 Sol or Inkling Small?

GPT-6.1 Sol: 1,050,000 tokens, against 524,288 for Inkling Small.

Run GPT-6.1 Sol and Inkling Small on your own prompt

Benchmarks aren't your data. Try both side by side in the playground, with the cost of every answer.