LIVE · refreshes every 20 min
updated Aug 11, 10:21 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
Hugging Face
Aug 10
13h ago
Question about Fable Fusion: user reports model rejects prompts during pentesting test
r/LocalLLaMA
→ story
17h ago
Ling-3.0-tiny 8B A1.3B MoE released with 1.3B active parameters; Ling team highlights performance between 4B and 8-12B Qwen and Gemma models
r/LocalLLaMA
→ story
23h ago
OpenAI and Hugging Face breach story told from the perspective of the AI
Hacker News (AI)
→ story
1d ago
VLX-Seek-1.5-10B is an open-source 10B model for fine-grained perception and visual grounding in embodied scenarios including drones, robots, and edge devices.
r/LocalLLaMA
→ story
Aug 9
1d ago
endless-frontier/BigBang-v1 - qwen 3.5 finetunes benchmark results; comparisons between DeepSeek Flash and Pro show per-benchmark variance
r/LocalLLaMA
→ story
1d ago
Open-weight video generation with MiniMax H3 delivers 2K/24fps video and audio-driven generation on local hardware; H3 weights live on HuggingFace, five days of testing show quality is legit
r/LocalLLaMA
→ story
Aug 8
3d ago
ggml-org/llama.cpp adds Longcat-Flash support for testing; PR 19182 seeks tests on larger models with GGUF excerpt from HuggingFace
r/LocalLLaMA
→ story
Aug 7
3d ago
OpenAI releases timeline of the Hugging Face incident from Black Hat presentation; corroborating coverage notes Black Hat USA 2026 breaking news
Simon Willison
→ story
3d ago
OpenAI pauses internal activities on Astra due to not meeting new security standards.
The Verge AI
→ story
4d ago
Unsloth releases Kimi K3 model variants with UD-Q1_0 466GB, TQ2_0 551GB, and larger configurations; Hugging Face page linked.
r/LocalLLaMA
→ story
4d ago
EschaLabs/Qwen3.6-35B-A3B-Escha-W2 benchmarks show 12.19 GiB VRAM usage on ROCmFPX; APEX Q5 Balanced performs similarly (Escha vs APEX)
r/LocalLLaMA
→ story
Aug 6
4d ago
NVIDIA brings outsourcing? No. The rewrite: "NVIDIA's speech stack runs locally on-device with NeMo-Speech.cpp, including ASR, TTS, and codec quantized to GGUF" but should be single sentence, factual. Also sentence case. Keep names exact: NeMo-Speech.cpp, GGUF, ASR, TTS, Magpie TTS Multilingual? The primary source headline says "NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp". We'll produce: "NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp." Ensure sentence case: Only first word capitalized and proper nouns; "GGUF" is acronym, capital. "ASR" "TTS" are acronyms. "NeMo-Speech.cpp" includes hyphen and extension; treat as proper noun capitalization; first word "NVIDIA's" capitalized; rest lowercase except acronyms and model names. We'll deliver: "NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp" That's good. Under 200 chars. Count roughly. Let's output. NVIDIA's whole speech stack goes local, with ASR, TTS, and codec quantized to GGUF and running on-device via NeMo-Speech.cpp
r/LocalLLaMA
→ story
4d ago
Scotoma-2: Gemma4 improved with cleaner writing and fewer tics
r/LocalLLaMA
→ story
4d ago
NVIDIA Nemotron Omni loads only its text half on Mac; author writes vision and audio towers in mlx to support full runtime
r/LocalLLaMA
→ story
4d ago
NVIDIA Nemotron Parse 2.0 converts document images into structured representations with text, layout classes, bounding boxes, and reading order; outputs formatted text and spatial annotations for document elements.
r/LocalLLaMA
→ story
5d ago
Baseten joins Hugging Face Inference Providers collection
Hugging Face
→ story
Aug 5
5d ago
Mistral releases premier not-hotdog model
r/LocalLLaMA
→ story
5d ago
DeepSeek-v4-flash-mini demonstrates continued exploration of whoops DeepSeek work; user submission discusses extending DeepSeek via exploration
r/LocalLLaMA
→ story
6d ago
Intern S2 Mobius uses Qwen3.5-35B derived model with higher throughput and lower token consumption
r/LocalLLaMA
→ story
6d ago
GPT-X2.5-135M places 3rd on the Open SLM Leaderboard on Huggingface, ahead of Facebook's MobileLLM-R1-140M
r/LocalLLaMA
→ story
6d ago
HuggingFace model requires more DRAM; user anticipates DRAM prices may rise
r/LocalLLaMA
→ story
Aug 4
6d ago
Hugging Face CEO says China is winning the AI race and dominating on open models
r/LocalLLaMA
→ story
6d ago
InclusionAI Ling-3.0-flash weights released on Hugging Face; MIT reports BF16 and FP8 with 127.5B parameters and 5.1B active users, 512 experts with 8 active per token
r/LocalLLaMA
→ story
6d ago
Best way to run DS4 flash on a Mac (M3 Ultra) with 192 GB+ VRAM; includes dspark/mtp support and rising token throughput
r/LocalLLaMA
→ story
Aug 3
7d ago
NVIDIA NemotronLabs VoiceChat 11B model released on Hugging Face with full duplex capability
r/LocalLLaMA
→ story
7d ago
Hugging Face artificial analysis performs well for general work; coding is the only area inferior to Qwen.
r/LocalLLaMA
→ story
7d ago
Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?
r/LocalLLaMA
→ story
7d ago
Show HN: Glaux – browser-only AI for Hugging Face ONNX community models
Hacker News (AI)
→ story
7d ago
The OpenAI–Hugging Face hack serves as a stark warning.
Hacker News (AI)
→ story
8d ago
AI9Stars releases G9v3-39A5B open weights language model with 39B parameters and 5 active experts under Apache 2.0 license
r/LocalLLaMA
→ story
8d ago
Why AI agents lie and cheat to reach their goals
MIT Technology Review AI
→ story
8d ago
MiniMax-H3 arrives on HuggingFace with multimodal generation capabilities including video at up to 2K resolution and 15-second duration
r/LocalLLaMA
→ story
Aug 2
8d ago
LocalLLaMA remains a strong source of open-weight research, but locating it requires navigating benchmark discussions and hardware-focused posts on the subreddit.
r/LocalLLaMA
→ story
8d ago
Unsloth questions why Hy3 remains unfinished while quantizing many models; user speculates about government funding and watermarking concerns
r/LocalLLaMA
→ story
8d ago
Hugging Face demonstrates a 16.5-trillion-parameter model that contains nothing; computes a repository’s parameter count from safetensors headers alone
r/LocalLLaMA
→ story
9d ago
Qwen3.6-35B-A3B-DSV4Pro-SFT-GPT56Sol-RL-Agent review and user test inquiries posted on Hugging Face and Reddit
r/LocalLLaMA
→ story
Aug 1
10d ago
Antirez uploads new weights for DeepSeek v4 in existing folder on Hugging Face; user discussion notes added context
r/LocalLLaMA
→ story
Jul 31
10d ago
DeepSeek releases V4-Flash 0731 with enhanced agentic capabilities, 304B parameters (167GB on Hugging Face) and pricing at $0.14 per million input / $0.27 per million output; Artificial Analysis ranks it ahead of MiniMax M3.
Simon Willison
→ story
10d ago
AI labs call for restraint as Amazon and SpaceX continue AI activity; OpenAI CEO suggests pacing industry pace
TechCrunch AI
→ story
10d ago
Unsloth releases DeepSeek-V4-Flash-0731-GGUF; users anticipate launch on Hugging Face page
r/LocalLLaMA
→ story
←
1
2
3
→