PhysElite: evaluating how far LLMs are from solving olympiad-level physics problems
Read the original at arxiv.org→arXiv:2608.25097v1 Announce Type: new Abstract: Understanding how (multimodal) large language models perform on physics problems requires benchmarks that reflect the difficulty and breadth of expert-level physical...
Original headline: "PhysElite: How Far Are LLMs from Solving Olympiad-Level Physics Problems?"
Coverage timeline
- Aug 27, 04:00 UTC arXiv cs.AI lead source PhysElite: How Far Are LLMs from Solving Olympiad-Level Physics Problems?