HealthBench Professional

OpenAI logoGPT-5.6 Terra on HealthBench Professional

rank 5 of 9 · updated August 16, 2026

GPT-5.6 Terra scores 0.577 on HealthBench Professional, rank 5 of 9 evaluated models. The balanced middle tier of the GPT-5.6 family, between Sol and Luna. HealthBench Professional scores models on 525 tasks drawn from real clinician conversations, graded criterion by criterion against physician-written rubrics on a 0 to 1 scale.

Result and API facts

rank5 of 9
score0.577
labOpenAI
context window1.1M tokens
API price per 1M tokens$2.00 in / $12.00 out
licenseproprietary
released2026-07-09

Position in the field

The gap to the leader, Claude Fable 5 at 0.660, is 0.083. Directly above sits Claude Sonnet 5 at 0.578. Directly below sits Claude Opus 4.8 at 0.558. Scores on this page come from the same evaluation run, so differences between models are differences on identical tasks, not across configurations.

What does GPT-5.6 Terra score on HealthBench Professional?

GPT-5.6 Terra scores 0.577 on HealthBench Professional, which places it at rank 5 of 9 evaluated models as of August 16, 2026.

How much does GPT-5.6 Terra cost to run?

GPT-5.6 Terra is priced at $2.00 per million input tokens and $12.00 per million output tokens through OpenAI's API.

Head to head

Pairings with a dedicated comparison page are linked; every other difference is in the score-difference matrix.

How tasks are selected and graded is on the methodology page. The full ranking is on the leaderboard.