ggml-llama.cpp adds -sm tensor for tensor parallelism across multiple computers
Read the original at www.reddit.com→This is big. Imagine running your model across multiple computers, and now imagine doing tensor parallelism across all of them. submitted by /u/jacek2023 [link] [comments]
Original headline: "RPC: add `-sm tensor` by am17an · Pull Request #26610 · ggml-org/llama.cpp"
Coverage timeline
- Oct 6, 15:45 UTC r/LocalLLaMA lead source RPC: add `-sm tensor` by am17an · Pull Request #26610 · ggml-org/llama.cpp