Trying to run Kimi K3 locally: can I amortize the cost of analyzing a 2.8T teacher model across multiple MoE compression experiments?
Read the original at old.reddit.com→Hey r/LocalLLaMA, I’m a retired engineer with a background in distributed computing, currently running a 1-person startup. Like many people here, I’d love to experiment with 2T+ MoE models locally. The problem is...
Original headline: "I want to run Kimi K3 at home, so I’m trying to make 2.8T-scale experimentation cheaper"
Coverage timeline
- Jul 27, 08:26 UTC r/LocalLLaMA lead source I want to run Kimi K3 at home, so I’m trying to make 2.8T-scale experimentation cheaper