Can LLMs Discover Scientific Laws in Real and Parallel Worlds?
Examines core capabilities of LLMs for scientific discovery with rigorous, paper-grounded evaluation methods.
AI Summary
Researchers introduce SCILAWS-BENCH, a with 118 problems across six disciplines to evaluate LLMs' ability to discover scientific laws from real data and synthetic hidden laws.
Excerpt
Scientific equation discovery has long been central to scientific progress, proceeding through iterative cycles of hypothesis generation, observational testing, and refinement under scientific constraints. As LLM capabilities advance and their role in AI for Science expands, it remains an open problem whether they can genuinely discover scientific laws and how this ability should be evaluated. Existing evaluations, however, often either simplify discovery through synthetic settings or reuse publ
