HealthBench Professional

OpenAI logoGPT-5.6 Luna on HealthBench Professional

rank 7 of 9 · updated August 16, 2026

GPT-5.6 Luna scores 0.557 on HealthBench Professional, rank 7 of 9 evaluated models. The small, high-volume tier of the GPT-5.6 family, repriced to $0.20 per million input tokens on July 30, 2026. HealthBench Professional scores models on 525 tasks drawn from real clinician conversations, graded criterion by criterion against physician-written rubrics on a 0 to 1 scale.

Result and API facts

rank7 of 9
score0.557
labOpenAI
context window1.1M tokens
API price per 1M tokens$0.20 in / $1.20 out
licenseproprietary
released2026-07-09

Position in the field

The gap to the leader, Claude Fable 5 at 0.660, is 0.103. Directly above sits Claude Opus 4.8 at 0.558. Directly below sits GPT-5.5 Instant at 0.384. Scores on this page come from the same evaluation run, so differences between models are differences on identical tasks, not across configurations.

What does GPT-5.6 Luna score on HealthBench Professional?

GPT-5.6 Luna scores 0.557 on HealthBench Professional, which places it at rank 7 of 9 evaluated models as of August 16, 2026.

How much does GPT-5.6 Luna cost to run?

GPT-5.6 Luna is priced at $0.20 per million input tokens and $1.20 per million output tokens through OpenAI's API.

Head to head

Pairings with a dedicated comparison page are linked; every other difference is in the score-difference matrix.

How tasks are selected and graded is on the methodology page. The full ranking is on the leaderboard.