openai
Frontier science evaluation benchmark probing model capabilities on expert-level reasoning across natural sciences — designed to surface what AI systems can and cannot do at the research frontier.
Last 30 days
Explore
Read the real rows without downloading anything
Reading rows…
LiveRead from openai/frontierscience at the moment you asked. Nothing is cached or stored — every row above came from that request.
Column statistics
Over 160 rows of default/test