Models / Gemini 2.5 Pro Preview 06-05
Gemini 2.5 Pro Preview 06-05
Among models available today, in the bottom quarter for professional work.
- Price per million tokens
- $1.25 in · $10.00 out
Typical host on OpenRouter, 7 Oct 2026 · compare prices - Speed
- Not measured
- Intelligence Index
- Not measured
- Spring Prompt overall
- Not ranked yet
needs our benchmarks and two groups of results - Developer
- Google, based in the US
Weights not published - Released
- Not recorded
1,048,576 tokens of context
Where it stands among the models you could choose
Ranked only among models available today (those you can call through OpenRouter), in the groups people choose between: the same price band, similar intelligence, the same speed, developers based in the same place. Retired models are left out.
| Among | Intelligence Index | Spring Prompt overall | Price | Speed |
|---|---|---|---|---|
| Models available today | – | – | 179th of 221Mistral Nemo | – |
| Models from developers based in the US | – | – | 64th of 103Llama 3.1 8B Instruct | – |
The name under each rank is the leader of that group. Price is the developer's list price (or the typical OpenRouter host where we haven't read one) for three input tokens to one output; intelligence and speed are from Artificial Analysis (data sourced from Artificial Analysis); "similar intelligence" means within 4 points on its Intelligence Index. Developer locations are where each company is based, not where a model is served. Groups under 3 models are not ranked.
How good is it, and for what?
Each use case ranks the available models its benchmarks measured, then those priced $3–10 per million tokens. The bar shows where it falls in that field, best to the right. Open a row for the results behind it.
Product listingsTurning a sparse product feed and photos into listings that can go live
- CatalogBench: has not measured this model.
Decks from an analysisTurning a finished analysis into a deck you could present as it is
- DeckBench: has not measured this model.
Marketing planningPlanning a year of ad spend without overspending
- ROASBench: has not measured this model.
Agents and tool useMulti-step tasks with tools: support desks, coding agents, function calls
- tau2-bench: has not measured this model.
- Berkeley Function Calling Leaderboard (BFCL) V4: has not measured this model.
- OpenHands Index: has not measured this model.
- Microsoft STATE-Bench: has not measured this model.
- Artificial Analysis: has not measured this model.
- Vending-Bench 2: has not measured this model.
Professional workReal tasks from banking, consulting and law, business documents and freelance projects
| Benchmark | Gemini 2.5 Pro Preview 06-05 | Rank among available | Best available |
|---|---|---|---|
| Remote Labor IndexRemote Labor Index: projects done to client standard · reported | 0.8% | 12th of 12 | GPT-6 Astra 20.8% |
- APEX-Agents: has not measured this model.
- GDP.pdf: has not measured this model.
Reasoning and knowledgeHard questions across science, maths and general knowledge
- Artificial Analysis: has not measured this model.
WritingWhat people prefer in blind comparisons, and judged writing quality
- Arena (formerly LMArena): has not measured this model.
- UGI Leaderboard: has not measured this model.
Sticking to the factsSummarising without inventing things, and factual answers
- Vectara Hallucination Leaderboard: has not measured this model.
- Arena (formerly LMArena): has not measured this model.
- SimpleQA Verified (Epoch AI): has not measured this model.
Speed under pressureGood decisions against a real clock (fast chess)
- BulletBench: has not measured this model.
Against the alternatives
The models you would most likely weigh it against: the leaders of the groups above.
| Gemini 2.5 Pro Preview 06-05 | GPT-6.1 Sol leads overall | Gemini 3.8 Flash Google's best other model | Claude Opus 5.5 near the top overall | GPT-6 Astra near the top overall | |
|---|---|---|---|---|---|
| Price per million tokens | $3.44 | $4.00 | $1.50 | $8.00 | $20.00 |
| Tokens a second | – | 57 | 187 | 97 | 52 |
| Intelligence Index | – | 51.8 | 40.9 | 57.6 | 52.7 |
| Spring Prompt overall | – | 88 | 61 | 82 | 78 |
| Rank among available models, by use case | |||||
| Product listings | – | 2nd | 11th | 6th | 1st |
| Decks from an analysis | – | 3rd | 11th | 5th | 1st |
| Marketing planning | – | – | 7th | 5th | 1st |
| Agents and tool use | – | 3rd | 30th | 2nd | 4th |
| Professional work | 49th of 51 | 9th | 11th | 3rd | 5th |
| Reasoning and knowledge | – | 5th | 12th | 1st | 3rd |
| Writing | – | – | 2nd | 4th | 17th |
| Sticking to the facts | – | 2nd | 13th | 6th | 27th |
| Speed under pressure | – | – | 6th | – | 14th |
Shaded figures are better than Gemini 2.5 Pro Preview 06-05's. A dash means no result.
Where it has been measured
- Remote Labor Index1 result
- BulletBenchNot measured
- CatalogBenchNot measured
- DeckBenchNot measured
- ROASBenchNot measured
- APEX-AgentsNot measured
- Arena (formerly LMArena)Not measured
- Artificial AnalysisNot measured
- Berkeley Function Calling Leaderboard (BFCL) V4Not measured
- GDP.pdfNot measured
- Microsoft STATE-BenchNot measured
- OpenHands IndexNot measured
- SimpleQA Verified (Epoch AI)Not measured
- tau2-benchNot measured
- UGI LeaderboardNot measured
- Vectara Hallucination LeaderboardNot measured
- Vending-Bench 2Not measured
Every result
1 published results from 1 source, each in the source's own units, with the configuration that produced it.
Remote Labor Index · reported by Remote Labor Index · 1 result
| Measure | Value | Rank | Configuration | Dated |
|---|---|---|---|---|
| Remote Labor Index: projects done to client standard% of projects, higher is better | 0.8% | 14 of 14 | Gemini 2.5 Pro Preview 06-05 | 1 Oct 2026 |
Not shown: Not speed or cost; a small set of projects, so a few points either way are noise.