Online learning with LLM experts from limited feedback
Read the original at arxiv.org→arXiv:2609.05820v1 Announce Type: new Abstract: We study adaptive routing of prompts to large language model (LLM) experts to maximize response quality in an online setting with limited feedback. We formulate it as...
Original headline: "Online Learning with LLM Experts from Limited Feedback"
Coverage timeline
- Sep 9, 04:00 UTC arXiv cs.LG lead source Online Learning with LLM Experts from Limited Feedback