Search Intelligence Score
66.7
+35.3 lift from search
050100
- Lift from search
- +35.3
- 31.3 without search
- Cost per 1K tasks
- $217
- #8 in Search Efficiency
- Time per task
- 52.9s
- Search plus inference time to complete one task
Accuracy by benchmark
Accuracy on each eval, with search (solid) and without (light).
- DSQA60.0 vs 19.0
- HLE68.0 vs 57.0
- WISER72.0 vs 18.0
Cost per 1K tasks
How the cost splits between model inference, search calls, and page extraction.
- Inference
- $207
- 96%
- Search
- $5.83
- 3%
- Extract
- $3.80
- 2%
InferenceSearchExtract
Compared with
Other models from the same provider alongside models that score closest.
| Model | Score | Lift | Cost per 1K tasks | Time per task |
|---|---|---|---|---|
| GPT-5.6 Sol | 66.7 | +35.3 | $217 | 52.9s |
| GPT-6 (Astra)GPT-6 (Astra) | 70.3 | +35.3 | $301 | 34.7s |
| GPT-5.6 LunaGPT-5.6 Luna | 54.7 | +32.3 | $23.0 | 59.4s |
| DS Flash 4.1DeepSeek V4.1 Flash | 65.3 | +45.0 | $44.8 | 124s |
| Muse Spark 1.3Muse Spark 1.3 | 65.3 | +34.3 | $98.2 | 66.2s |
Search Intelligence Score averages accuracy on DSQA, HLE, and WISER with equal weight; cost and time are averaged the same way, with cost shown per 1,000 tasks. Latest update September 14, 2026. Full methodology.
# GPT-5.6 Sol · Search Capability Leaderboard
- Search Intelligence Score: 66.7 (#2 in Search Intelligence)
- Without search: 31.3 · Lift from search +35.3
- Cost per 1K tasks: $217 (#8 in Search Efficiency)
- Time per task: 52.9s
- Medals: Silver, Search Intelligence
- Lab: OpenAI · Model id: gpt-5.6-sol
## Accuracy by benchmark
| Suite | With search | Without search |
|---|---|---|
| DSQA | 60.0 | 19.0 |
| HLE | 68.0 | 57.0 |
| WISER | 72.0 | 18.0 |
## Compared with
| Model | Score | Lift | Cost per 1K tasks | Time per task |
|---|---|---|---|---|
| GPT-5.6 Sol | 66.7 | +35.3 | $217 | 52.9s |
| GPT-6 (Astra) | 70.3 | +35.3 | $301 | 34.7s |
| GPT-5.6 Luna | 54.7 | +32.3 | $23.0 | 59.4s |
| DeepSeek V4.1 Flash | 65.3 | +45.0 | $44.8 | 124s |
| Muse Spark 1.3 | 65.3 | +34.3 | $98.2 | 66.2s |
Latest update September 14, 2026. Methodology: /leaderboard#methodology