LIVE · refreshes every 20 min
updated Aug 25, 18:22 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
NVIDIA
Aug 11
14d ago
Nvidia building 1T-parameter Nemotron 4 to rival OpenAI models
Hacker News (AI)
→ story
14d ago
NVIDIA Nemo Switchyard introduces an open source LLM router; alternative to OpenRouter Fusion and Sakana Fugu
r/LocalLLaMA
→ story
14d ago
Unsloth Desktop app released for Mac, Windows, and Linux to run and train models locally, with support for MLX, diffusion models, audio models, and GGUF, plus local Claude/Code integration and sandboxed execution.
r/LocalLLaMA
→ story
14d ago
NVIDIA Nemotron 3.5 Lightning 30B A3B BF16 model on Hugging Face
r/LocalLLaMA
→ story
14d ago
NVIDIA Nemotron 3.5 Lightning delivers fast, accurate specialized task execution for long-running agents
NVIDIA Developer
→ story
14d ago
NVIDIA and local AI communities fuel open source models and intelligent agents
NVIDIA Blog
→ story
14d ago
NVIDIA NeMo Switchyard routes AI agent workloads across models
NVIDIA Developer
→ story
14d ago
Nvidia tests lower memory configurations of Rubin Ultra as memory shortage bites back, including as little as 192 GB and stepping back to HBM4
r/LocalLLaMA
→ story
Aug 10
15d ago
Open Weights and full deployment control with NVIDIA Magpie TTS for building low-latency multilingual voice agents
Hugging Face
→ story
15d ago
Retired engineer develops a ML programming language to teach and demonstrate ML concepts with in-browser training and CUDA/NVIDIA and MLX/Apple Silicon support
r/LocalLLaMA
→ story
Aug 8
16d ago
Enabling PCIe peer-to-peer for consumer Nvidia cards yields more performance than expected
r/LocalLLaMA
→ story
17d ago
M225 for 80–100€: worth it?
r/LocalLLaMA
→ story
17d ago
TPUs for inference discussed; user experiments and scale considerations highlighted
r/LocalLLaMA
→ story
17d ago
Firebird launches the CIS region’s largest AI factory in Armenia.
NVIDIA Blog
→ story
Aug 7
17d ago
Serving Deepseek v4 Flash 0731 on 2x DGX Spark; how to lower VRAM usage and increase OS headroom for RAM?
r/LocalLLaMA
→ story
18d ago
Parakeet.wgsl enables fast, accurate ASR in the browser using raw WebGPU compute shaders and SIMD WASM with NVIDIA’s Parakeet TDT 0.6B V2 English transcription model.
r/LocalLLaMA
→ story
18d ago
RTX 5090 owner builds open-source tool that shuts down PC if 12VHPWR cable draws too much power; works only on specific GPUs
r/LocalLLaMA
→ story
18d ago
I made a simple local voice input extension for pi using NVIDIA Nemtron 3.5 ASR 0.6B on CPU
r/LocalLLaMA
→ story
Aug 6
18d ago
NVIDIA brings outsourcing? No. The rewrite: "NVIDIA's speech stack runs locally on-device with NeMo-Speech.cpp, including ASR, TTS, and codec quantized to GGUF" but should be single sentence, factual. Also sentence case. Keep names exact: NeMo-Speech.cpp, GGUF, ASR, TTS, Magpie TTS Multilingual? The primary source headline says "NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp". We'll produce: "NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp." Ensure sentence case: Only first word capitalized and proper nouns; "GGUF" is acronym, capital. "ASR" "TTS" are acronyms. "NeMo-Speech.cpp" includes hyphen and extension; treat as proper noun capitalization; first word "NVIDIA's" capitalized; rest lowercase except acronyms and model names. We'll deliver: "NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp" That's good. Under 200 chars. Count roughly. Let's output. NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp
r/LocalLLaMA
→ story
18d ago
NVIDIA CMP 40HX 8GB could receive updates alongside 170HX variants; analysis explores current trends
r/LocalLLaMA
→ story
19d ago
NVIDIA Nemotron Parse 2.0 converts document images into structured representations with text, layout classes, bounding boxes, and reading order; outputs formatted text and spatial annotations for document elements.
r/LocalLLaMA
→ story
19d ago
Open World Models push the frontier of physical AI; NVIDIA signs open letter urging open weights and American AI leadership across sectors
NVIDIA Blog
→ story
19d ago
Frontier narrows pricing gap as performance catches up to price concerns; users report downgraded free version after new models release
r/LocalLLaMA
→ story
19d ago
The Calibration Floor: format repair can masquerade as self-correction at small-to-mid scale
arXiv cs.CL
→ story
Aug 5
19d ago
Using an Nvidia RTX 5090 with 64 GB RAM and an older AMD GPU with 12 GB VRAM to run a local Llama.cpp model; a trade business asks if both GPUs can be utilized for two different AI models.
r/LocalLLaMA
→ story
20d ago
NVIDIA V100 32GB suitable for DeepSeek V4 Flash for 30-50 users, per user inquiry
r/LocalLLaMA
→ story
20d ago
Mark Cuban and Michael Burry warn on Nvidia and AI stocks
Hacker News (AI)
→ story
Aug 4
20d ago
GPT-OSS turns one year old; user shares favorable comparisons to Qwen 3.5 122B and Nemotron 3 Super in local models
r/LocalLLaMA
→ story
21d ago
NVIDIA joins NSF State and Regional AI Hubs program to expand AI research and education across the US.
NVIDIA Blog
→ story
21d ago
Full 1M context on a single RTX5090 with DDR5 desktop setup using vLLM CPU/RAM offloading; ~800 tps per prompt and 15+ tps decode
r/LocalLLaMA
→ story
21d ago
Llama.cpp adds GPU-based sampling for MTP, claiming about an 8% speedup in tok/s on Qwen3.6-35B with a 5090; observed ~4% speedup on Nvidia P40 in tests.
r/LocalLLaMA
→ story
Aug 3
21d ago
NVIDIA NemotronLabs VoiceChat 11B model released on Hugging Face with full duplex capability
r/LocalLLaMA
→ story
22d ago
NVIDIA Vera Storage benchmarks show faster encryption, compression, integrity checking, and recovery for AI-native storage.
NVIDIA Developer
→ story
22d ago
Nvidia's CUDA faces new threats from AI coding agents.
Hacker News (AI)
→ story
22d ago
NVidia VRAM stagnation: 70-class GPUs stuck at 12GB for a decade despite improved cores; 24GB would enable capable local model machines for AI workloads
r/LocalLLaMA
→ story
Aug 2
22d ago
China’s DFSX offers 2x the memory bandwidth of NVIDIA’s GB200
r/LocalLLaMA
→ story
23d ago
Open Weights and American AI Leadership signed by 235 AI-adjacent companies including OpenAI and Microsoft to counter government calls for limits on AI development
Simon Willison
→ story
Jul 31
24d ago
AI boom spurs insider selling by Nvidia, CoreWeave billionaires
Hacker News (AI)
→ story
25d ago
Hygon reveals 512-thread CPU and AI GPU to rival Intel Xeon and Nvidia
Hacker News (AI)
→ story
25d ago
Nvidia's $750B AI bet deepens fears of a circular tech bubble
Hacker News (AI)
→ story
←
1
2
3
4
→