FineServe: a fine-grained dataset and characterization of global LLM serving workloads
Read the original at arxiv.org→arXiv:2607.19349v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as always-on online services, making efficient LLM serving a critical systems challenge. Achieving low latency...
Original headline: "FineServe: A Fine-Grained Dataset and Characterization of Global LLM Serving Workloads"
Coverage timeline
- Jul 23, 04:00 UTC arXiv cs.AI lead source FineServe: A Fine-Grained Dataset and Characterization of Global LLM Serving Workloads