Measuring the microtask eligibility gap: when is an off-the-shelf SLM enough for an agent harness?
Read the original at arxiv.org→arXiv:2610.00025v1 Announce Type: new Abstract: Agent harnesses increasingly want to run small language models (SLMs) on the microtasks around a frontier large language model (LLM) planner: auto-approving shell...
Original headline: "Measuring the Microtask Eligibility Gap: When Is an Off-the-Shelf SLM Enough for an Agent Harness?"
Coverage timeline
- Oct 2, 04:00 UTC arXiv cs.AI lead source Measuring the Microtask Eligibility Gap: When Is an Off-the-Shelf SLM Enough for an Agent Harness?