LIVE · refreshes every 20 min
updated Aug 18, 11:42 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
NVIDIA
Aug 18
10h ago
CDW bumps the RTX Pro 6000 MSRP from $16,000 to $19,999
r/LocalLLaMA
→ story
Aug 17
19h ago
Groq raises $350 million at a $3.5 billion valuation as it pivots from AI chips to neocloud and expands its Nvidia-powered data center footprint
TechCrunch AI
→ story
20h ago
Nvidia to invest $1.5B in SoftBank data center developer behind OpenAI project.
TechCrunch AI
→ story
Aug 16
1d ago
Russian missile uses Nvidia AI chip to aid targeting in Ukraine, per HUR; Ukraine finds Nvidia AI chip in new missile
Hacker News (AI)
→ story
2d ago
AMD RX 6900 XT and NVIDIA GT 710 setup on ASUS Z10PA-U8 with Ubuntu Server 24.04 yields limited performance; user reports better functionality on GA-Z77-DS3H with risers but only ~7 t/s on a 40B model
r/LocalLLaMA
→ story
Aug 15
2d ago
CORS Chat provides a web UI to test an OpenAI-Responses-compatible chat endpoint; conversations persist in the browser and can be exported as JSON.
Simon Willison
→ story
Aug 14
3d ago
Nvidia's $500B plan envelops Wall Street in its AI frenzy
Hacker News (AI)
→ story
3d ago
Universitas Gadjah Mada, Indosat and NVIDIA open Indonesia’s first university AI center to develop local AI talent
NVIDIA Blog
→ story
Aug 13
4d ago
Desktop HP Z8 Fury offers 2TB DDR5 RAM option priced at about $211k
r/LocalLLaMA
→ story
4d ago
Qwen3.8-2.4T-A95B runs locally on RTX 5090 + RTX 5060 Ti at about 0.80 tokens per second using llama.cpp with Unsloth GGUF quantization and 512 routed experts, 10 active per token.
r/LocalLLaMA
→ story
Aug 12
5d ago
Alibaba releases open weights for Qwen3.8-2.4T-A95B with configurable reasoning on NVIDIA GB300 NVL72
NVIDIA Developer
→ story
5d ago
How to choose full-stack observability for NVIDIA AI factories
NVIDIA Developer
→ story
5d ago
NVIDIA founder and CEO Jensen Huang ranks No. 1 on Glassdoor’s Best CEOs list for 2026.
NVIDIA Blog
→ story
5d ago
Ed Zitron discusses Nvidia and unprofitable AI labs on CNBC Squawk Box
Hacker News (AI)
→ story
6d ago
RTX 6000 PRO price raised to $16,000 on Nvidia website
r/LocalLLaMA
→ story
6d ago
NVIDIA announces partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish financing platforms mobilizing over $500 billion of third-party capital for AI infrastructure.
NVIDIA Blog
→ story
Aug 11
6d ago
Local LLMs run on a 1060 3GB garbage card, demonstrating Tinyllama’s responsiveness on underpowered hardware.
r/LocalLLaMA
→ story
6d ago
NVIDIA JetPack 7.2.1 adds agentic video skills and T3000 emulation.
NVIDIA Developer
→ story
6d ago
Nvidia building 1T-parameter Nemotron 4 to rival OpenAI models
Hacker News (AI)
→ story
6d ago
NVIDIA Nemo Switchyard introduces an open source LLM router; alternative to OpenRouter Fusion and Sakana Fugu
r/LocalLLaMA
→ story
6d ago
Unsloth Desktop app released for Mac, Windows, and Linux to run and train models locally, with support for MLX, diffusion models, audio models, and GGUF, plus local Claude/Code integration and sandboxed execution.
r/LocalLLaMA
→ story
6d ago
NVIDIA Nemotron 3.5 Lightning 30B A3B BF16 model on Hugging Face
r/LocalLLaMA
→ story
6d ago
NVIDIA Nemotron 3.5 Lightning delivers fast, accurate specialized task execution for long-running agents
NVIDIA Developer
→ story
6d ago
NVIDIA and local AI communities fuel open source models and intelligent agents
NVIDIA Blog
→ story
6d ago
NVIDIA NeMo Switchyard routes AI agent workloads across models
NVIDIA Developer
→ story
7d ago
Nvidia tests lower memory configurations of Rubin Ultra as memory shortage bites back, including as little as 192 GB and stepping back to HBM4
r/LocalLLaMA
→ story
Aug 10
7d ago
Open Weights and full deployment control with NVIDIA Magpie TTS for building low-latency multilingual voice agents
Hugging Face
→ story
8d ago
Retired engineer develops a ML programming language to teach and demonstrate ML concepts with in-browser training and CUDA/NVIDIA and MLX/Apple Silicon support
r/LocalLLaMA
→ story
Aug 8
9d ago
Enabling PCIe peer-to-peer for consumer Nvidia cards yields more performance than expected
r/LocalLLaMA
→ story
9d ago
M225 for 80–100€: worth it?
r/LocalLLaMA
→ story
10d ago
TPUs for inference discussed; user experiments and scale considerations highlighted
r/LocalLLaMA
→ story
10d ago
Firebird launches the CIS region’s largest AI factory in Armenia.
NVIDIA Blog
→ story
Aug 7
10d ago
Serving Deepseek v4 Flash 0731 on 2x DGX Spark; how to lower VRAM usage and increase OS headroom for RAM?
r/LocalLLaMA
→ story
10d ago
Parakeet.wgsl enables fast, accurate ASR in the browser using raw WebGPU compute shaders and SIMD WASM with NVIDIA’s Parakeet TDT 0.6B V2 English transcription model.
r/LocalLLaMA
→ story
11d ago
RTX 5090 owner builds open-source tool that shuts down PC if 12VHPWR cable draws too much power; works only on specific GPUs
r/LocalLLaMA
→ story
11d ago
I made a simple local voice input extension for pi using NVIDIA Nemtron 3.5 ASR 0.6B on CPU
r/LocalLLaMA
→ story
Aug 6
11d ago
NVIDIA brings outsourcing? No. The rewrite: "NVIDIA's speech stack runs locally on-device with NeMo-Speech.cpp, including ASR, TTS, and codec quantized to GGUF" but should be single sentence, factual. Also sentence case. Keep names exact: NeMo-Speech.cpp, GGUF, ASR, TTS, Magpie TTS Multilingual? The primary source headline says "NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp". We'll produce: "NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp." Ensure sentence case: Only first word capitalized and proper nouns; "GGUF" is acronym, capital. "ASR" "TTS" are acronyms. "NeMo-Speech.cpp" includes hyphen and extension; treat as proper noun capitalization; first word "NVIDIA's" capitalized; rest lowercase except acronyms and model names. We'll deliver: "NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp" That's good. Under 200 chars. Count roughly. Let's output. NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp
r/LocalLLaMA
→ story
11d ago
NVIDIA CMP 40HX 8GB could receive updates alongside 170HX variants; analysis explores current trends
r/LocalLLaMA
→ story
11d ago
NVIDIA Nemotron Parse 2.0 converts document images into structured representations with text, layout classes, bounding boxes, and reading order; outputs formatted text and spatial annotations for document elements.
r/LocalLLaMA
→ story
11d ago
Open World Models push the frontier of physical AI; NVIDIA signs open letter urging open weights and American AI leadership across sectors
NVIDIA Blog
→ story
←
1
2
3
4
→