OpenProblemBench: 82 unresolved problems benchmark for testing AI on foundational theoretical science questions
Read the original at arxiv.org→arXiv:2610.11118v1 Announce Type: new Abstract: The next frontier for artificial general intelligence is tackling unresolved scientific problems, calling for benchmarks that assess progress beyond established...
Original headline: "OpenProblemBench: Benchmarking AI on Open Problems in the Foundational Theoretical Sciences"
Coverage timeline
- Oct 9, 04:00 UTC arXiv cs.AI lead source OpenProblemBench: Benchmarking AI on Open Problems in the Foundational Theoretical Sciences