OpenAI Releases LifeSciBench, a 750-Task Benchmark Grading AI Models on Real Life-Science Research With Expert-Written Rubric
Most biology benchmarks ask slender, fact-based questions with clear solutions. Scientists weigh imperfect proof and make selections. OpenAI launched LifeSciBench and it targets that hole instantly. Even the strongest mannequin passes roughly one process in three. The benchmark is much from saturated. What is LifeSciBench LifeSciBench incorporates 750 expert-authored duties. They span seven workflows and…
