LIVE · refreshes every 20 min
updated Sep 4, 00:22 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
DeepSeek
Sep 1
2d ago
DeepSeek-V4-Flash-Vision-Exp (Note: This input is already a specific model/version name; as a headline it remains unchanged to reflect the primary source.)
Hacker News (AI)
→ story
2d ago
The Halt Vector internalizes a causal steering intervention to halt chain-of-thought reasoning for more efficient reasoning
arXiv cs.LG
→ story
Aug 29
5d ago
Cut Claude code bill by routing to DeepSeek or Grok.
Hacker News (AI)
→ story
Aug 25
9d ago
DeepSeek V4 shows vision intelligence, performance, and price analysis
Hacker News (AI)
→ story
Aug 21
13d ago
DeepSeek-V4-Flash-Vision-Exp Release: Multimodal API Now Live
DeepSeek (via Google News)
→ story
Aug 20
14d ago
Beijing AI-themed bar offers DeepSeek tokens with pints.
Hacker News (AI)
→ story
Aug 19
15d ago
QWEN plans to reuse the 397B-A17B architecture to compete with Deepseek V4 0731 Flash
r/LocalLLaMA
→ story
15d ago
GLM-5.3 released on AA; creator disputes AA’s intelligence/cost plot and presents an alternative chart using a linear scale and DeepSeek V4 Flash 0731 pricing
r/LocalLLaMA
→ story
15d ago
Qwen 3.8 27B performs worse than Claude Code and GitHub Copilot for agentic coding, according to user experience with local models and tools
r/LocalLLaMA
→ story
15d ago
Frontier level LLM with updateable or fixed engrams: when will this appear?
r/LocalLLaMA
→ story
15d ago
Lucebox now, or wait for the new Framework Desktop (Ryzen AI Max+ PRO 495 / 192 GB) + PCIe x4-to-x16 adapter & Radeon AI PRO R9700?
r/LocalLLaMA
→ story
Aug 18
16d ago
Qwen 3.8 27B matches frontier intelligence and outperforms Google’s current frontier model without a data center; most local model experts can run it on their own hardware.
r/LocalLLaMA
→ story
16d ago
Running DeepSeek V4 Flash 0731 on Strix Halo: draft model and n_max sweep notes
r/LocalLLaMA
→ story
16d ago
Qwen 3.8 27b saved me $650+ in API costs this evening.
r/LocalLLaMA
→ story
Aug 17
17d ago
Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index, tying with GPT-5.6 Luna Max and trailing GLM-5.2 Max and DeepSeek V4 Pro 0813 Max
Simon Willison
→ story
17d ago
User pairs DeepSeek V4 Flash with Antirez's Dwarfstar 4 framework -- personal review
r/LocalLLaMA
→ story
Aug 16
18d ago
GitHub trending page excludes DeepSeek projects; analysis investigates possible causes
Hacker News (AI)
→ story
18d ago
Show HN: I shrank DeepSeek V4 Flash to 57GB; it wrote a compiler on my Mac
Hacker News (AI)
→ story
18d ago
Qwen 3.8 27b with DSH (DeepSeek Harness) experiences strong performance and reliability
r/LocalLLaMA
→ story
18d ago
Best setup for a 16 GB VRAM + 128 GB RAM system for running LLMs like Qwen 3.6 35B and related GGUF quantizations; user tests with 12700k, 5060 Ti, and 128 GB RAM explored performance.
r/LocalLLaMA
→ story
Aug 15
19d ago
RunInfra launches inference API pricing pages for DeepSeek V4 Flash (278 tok/s) and V4 Pro (207 tok/s, full 1M context)
Hacker News (AI)
→ story
Aug 14
20d ago
DeepSeek V4-Flash released with 284B parameters and 1M-token context, free to use
Hacker News (AI)
→ story
20d ago
AI models released within a month in China include Kimi K3-2.8T, Qwen3.8-2.4T, DeepSeek-V4-Pro-0813-1.6T, and GLM-5.3-743B
r/LocalLLaMA
→ story
20d ago
Model creators are urged to request and respond to VRAM, RAM, and model-size feedback from users; survey posts target major labs like Kimi, GLM, MiniMax, MiMo, Deepseek, Tencent, and Inkling
r/LocalLLaMA
→ story
20d ago
"Caveman reasoning": r/LocalLLaMA discusses finetunes and official models (Muse Glimmer, DeepSeek V4 Pro) with blunt reasoning styles
r/LocalLLaMA
→ story
20d ago
DeepSeek v4 Flash 0731 Oneshots Tetris under Windows XP release notes and demo video
r/LocalLLaMA
→ story
20d ago
Motif 3 NVFP4 performs closely to DeepSeek V4 Flash 0731 on benchmarks and may be underrated; best with custom vLLM.
r/LocalLLaMA
→ story
Aug 13
21d ago
DeepSeek API pricing update
Hacker News (AI)
→ story
21d ago
DeepSeek-V4-Pro GA release announced
DeepSeek (via Google News)
→ story
21d ago
User has DeepSeek-V4-Flash work with Muse-Glimmer for vision ability in a personal agent -- demo/experiment post
r/LocalLLaMA
→ story
Aug 12
22d ago
DeepSeek V4 Flash 0731 uncensored jailbreak guide
r/LocalLLaMA
→ story
22d ago
Benched a 124B on one DGX Spark for a week and published results; 38.7 tok/s on the fastest path found, 2.4x DeepSeek V4 Flash on the same box
r/LocalLLaMA
→ story
22d ago
Idea for a deepseek-v4-flash-0731 backed automated research workflow to be leveraged via qwen3.6/3.8 27b for difficult tasks that require highly technical, not easy to find information
r/LocalLLaMA
→ story
22d ago
DeepSeek releases Harness and unveils its invention title "What They Invented
Hacker News (AI)
→ story
Aug 11
23d ago
Quantizing DeepSeek V4 0731 and benchmarking against popular quants on 8× RTX 5090; issues found with --no-lazy and FP8 downconversion noted
r/LocalLLaMA
→ story
23d ago
Solar Open 2 (250B, 15B) model; user questions its use and comparison to DeepSeek V4 Flash
r/LocalLLaMA
→ story
23d ago
DeepSeek-V4-Flash logs into remote Linux machine, identifies syncthing syncing unwanted folder and removes it
r/LocalLLaMA
→ story
23d ago
DeepSeek: reverse engineering an AI assistant by interviewing itself
Hacker News (AI)
→ story
23d ago
Interpreting reasoning mechanisms of large language models via sparse autoencoders: separating Thinking from NoThinking in CoT-enabled models
arXiv cs.CL
→ story
23d ago
DeepSeek V4 Flash gains basic vision by training a 40.1M connector on 100K image-text examples
r/LocalLLaMA
→ story
←
1
2
3
4
→