Optimal configuration for 4x3090s using tensor parallelism and pipelining to approximate RTX 6000-like performance at 25% of the cost
Read the original at old.reddit.com→Aiming for RTX 6000 like performance at 25% of the cost. https://preview.redd.it/mi6fqpdd5hih1.png?width=631&format=png&auto=webp&s=9a639bcc79c7eb0a3834a88317220c289b7a52b6 The top 2x3090s are connected via tensor...
Original headline: "Optimal Configuration for 4x3090s"
Coverage timeline
- Aug 10, 04:24 UTC r/LocalLLaMA lead source Optimal Configuration for 4x3090s