NVIDIA Nemotron Omni loads only its text half on Mac; author writes vision and audio towers in mlx to support full runtime
Read the original at old.reddit.com→nvidias nemotron omni is open weights and it sees, hears and reasons. theres already a 4bit mlx quant on hugging face but only the text backbone loads with standard mlx tooling. the model card says it plainly, the...
Original headline: "nvidias nemotron omni only loads its text half on a mac, so i wrote the vision and audio towers in mlx"
Coverage timeline
- Aug 6, 17:41 UTC r/LocalLLaMA lead source nvidias nemotron omni only loads its text half on a mac, so i wrote the vision and audio towers in mlx