General decision models benchmark JEVal for 11,257 instances across 36 datasets and 10 domains; 25 model configurations evaluated.
Read the original at arxiv.org→arXiv:2610.03935v1 Announce Type: new Abstract: General decision models, such as Jev, have recently emerged as efficient alternatives to LLMs for structured judgment and selection. But what kinds of decisions can...
Original headline: "General Decision Models: Benchmarking and Insights Beyond Jev"
Coverage timeline
- Oct 5, 20:01 UTC r/LocalLLaMA nokia-applied-research/AnyJev: Turn any LLM into a Jev-style decision model
- Oct 6, 04:00 UTC arXiv cs.CL lead source General Decision Models: Benchmarking and Insights Beyond Jev