AI BENCHMARK PROFILE
onepot-Bench
onepot-Bench 0 is a proprietary benchmark suite evaluating language models on synthetic chemistry capabilities, including cheminformatics literacy, safety behavior, and reaction outcome prediction.
- Released
- 2026-08-03
- Readiness
- Paper only
- Primary field
- Science & Research
Why it matters
Targets skills needed for reliable laboratory decisions, addressing limitations of public data benchmarks.
Motivation
Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc analysis.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.