SemPlan benchmarks structured semantic planning for LLM-based queries over enterprise data with a bilingual synthetic benchmark and a frozen scientific evaluation subset
Read the original at arxiv.org→arXiv:2608.13612v1 Announce Type: new Abstract: Natural-language interfaces to enterprise data must translate underspecified requests into governed, executable behavior while controlling invalid queries, policy...
Original headline: "SemPlan: Benchmarking Structured Semantic Planning for LLM-Based Queries over Enterprise Data"
Coverage timeline
- Aug 17, 04:00 UTC arXiv cs.AI lead source SemPlan: Benchmarking Structured Semantic Planning for LLM-Based Queries over Enterprise Data