Corollary

Research

  • Ask

Library

  • Catalog

Account

  • Overview
  • Jobs
  • Usage
  • Billing
Settings
Corollary
  1. Catalog
  2. Datasets
  3. openai/frontierscience

openai

frontierscience

Frontier science evaluation benchmark probing model capabilities on expert-level reasoning across natural sciences — designed to surface what AI systems can and cannot do at the research frontier.

Original source
Rows
160
On disk
503 kB
Downloads
73k

Last 30 days

Updated
Dec 16

Explore

Read the real rows without downloading anything

Reading rows…

LiveRead from openai/frontierscience at the moment you asked. Nothing is cached or stored — every row above came from that request.

Column statistics

How the columns are distributed

Over 160 rows of default/test

answer

string_text
min
1
median
134.50
mean
1471.29
max
21,458

problem

string_text

Splits

1
  • default/test

Query support

  • Row preview
  • Paginated browse
  • Full-text search
  • SQL filter
  • Column statistics

Provenance

Licence
apache-2.0
Likes
171

Fields

Scientific ReasoningBenchmark
min
408
median
1242.50
mean
1509.76
max
6,589

subject

string_label
  • physics44%
  • chemistry38%
  • biology19%

task_group_id

string_text
min
36
median
36
mean
36
max
36