Fetching the paper…

NewtonBench: Benchmarking Generalizable Scientific Law Discovery in LLM Agents · Around