HealthBench Professional

OpenAI logoGPT-5.6 Sol on HealthBench Professional

rank 2 of 9 · updated August 16, 2026

GPT-5.6 Sol scores 0.605 on HealthBench Professional, rank 2 of 9 evaluated models. The flagship tier of OpenAI's GPT-5.6 family, aimed at its hardest reasoning and agentic workloads. HealthBench Professional scores models on 525 tasks drawn from real clinician conversations, graded criterion by criterion against physician-written rubrics on a 0 to 1 scale.

Result and API facts

rank2 of 9
score0.605
labOpenAI
context window1.1M tokens
API price per 1M tokens$5.00 in / $30.00 out
licenseproprietary
released2026-07-09

Position in the field

The gap to the leader, Claude Fable 5 at 0.660, is 0.055. Directly below sits Claude Opus 5 at 0.598. Scores on this page come from the same evaluation run, so differences between models are differences on identical tasks, not across configurations.

What does GPT-5.6 Sol score on HealthBench Professional?

GPT-5.6 Sol scores 0.605 on HealthBench Professional, which places it at rank 2 of 9 evaluated models as of August 16, 2026.

How much does GPT-5.6 Sol cost to run?

GPT-5.6 Sol is priced at $5.00 per million input tokens and $30.00 per million output tokens through OpenAI's API.

Head to head

Pairings with a dedicated comparison page are linked; every other difference is in the score-difference matrix.

How tasks are selected and graded is on the methodology page. The full ranking is on the leaderboard.