From isolated tasks to structured capabilities: a multilayer taxonomy for large language models
Read the original at arxiv.org→arXiv:2607.22182v1 Announce Type: new Abstract: Large language model (LLM) evaluation spans diverse tasks and benchmarks, yet evidence remains organized around tasks rather than the capabilities they probe. This...
Original headline: "From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models"
Coverage timeline
- Jul 27, 04:00 UTC arXiv cs.CL lead source From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models