Benchmark Radar
AI BENCHMARK PROFILE

onepot-Bench

Science & ResearchKnowledge & Reasoning

onepot-Bench 0 is a proprietary benchmark suite evaluating language models on synthetic chemistry capabilities, including cheminformatics literacy, safety behavior, and reaction outcome prediction.

Released
2026-08-03
Readiness
Paper only
Primary field
Science & Research

Why it matters

Targets skills needed for reliable laboratory decisions, addressing limitations of public data benchmarks.

Motivation

Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc analysis.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.