Benchmarking language models for statistical problem formulation
Read the original at arxiv.org→arXiv:2609.01982v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as assistants for statistical and data science work, yet existing evaluations largely assume the analysis target is...
Original headline: "Benchmarking Language Models for Statistical Problem Formulation"
Coverage timeline
- Sep 3, 04:00 UTC arXiv cs.AI lead source Benchmarking Language Models for Statistical Problem Formulation