Corollary

Research

  • Ask

Library

  • Catalog

Account

  • Overview
  • Jobs
  • Usage
  • Billing
Settings
Corollary
  1. Catalog
  2. Datasets
  3. openai/healthbench

openai

healthbench

Realistic multi-turn health conversations graded against physician-written rubrics across multiple axes (accuracy, completeness, communication) — an open evaluation benchmark for AI assistants in medicine.

Original source
Rows
0
On disk
—
Downloads
4.4k

Last 30 days

Updated
Aug 27

Explore

Read the real rows without downloading anything

Reading rows…

LiveRead from openai/healthbench at the moment you asked. Nothing is cached or stored — every row above came from that request.

Splits

1
  • default/test

Query support

  • Row preview
  • Paginated browse
  • Full-text search
  • SQL filter
  • Column statistics

Provenance

Licence
mit
Likes
166

Fields

MedicineBenchmarkScientific Reasoning