Running an LLM-driven town with 800+ persistent agents: concurrency, context caching, and inference costs
Read the original at www.reddit.com→I spent the past year independently building Slow Vale, an LLM-driven life simulation. The Chinese server now has 800+ AI residents sharing one continuously running city. This is an engineering write-up about...
Coverage timeline
- Oct 8, 09:49 UTC r/LocalLLaMA lead source Running an LLM-driven town with 800+ persistent agents: concurrency, context caching, and inference costs