LIVE · refreshes every 20 min
updated Sep 4, 01:23 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
DeepSeek
Aug 1
Aug 1, 2026
Show HN: DSCode – Coding agent powered by DeepSeek
Hacker News (AI)
→ story
Aug 1, 2026
DeepSeek outlines its plan for AGI as a Costco hot dog strategy.
Hacker News (AI)
→ story
Aug 1, 2026
Acquarium Panel Failure on DS4 flash 0731; Q3_K_XL Unsloth | DS4 flash 0731 issue reported with Acquarium panel failure on Q3_K_XL Unsloth
r/LocalLLaMA
→ story
Aug 1, 2026
Antirez uploads new weights for DeepSeek v4 in existing folder on Hugging Face; user discussion notes added context
r/LocalLLaMA
→ story
Aug 1, 2026
Hacker uses DeepSeek AI to autonomously attack vulnerable servers
Hacker News (AI)
→ story
Jul 31
Jul 31, 2026
DeepSeek releases V4-Flash 0731 with enhanced agentic capabilities, 304B parameters (167GB on Hugging Face) and pricing at $0.14 per million input / $0.27 per million output; Artificial Analysis ranks it ahead of MiniMax M3.
Simon Willison
→ story
Jul 31, 2026
OpenAI-style AI pricing: we cut price by 80% after DeepSeek v4 with 284B/13B active parameters outperformed on price/performance
r/LocalLLaMA
→ story
Jul 31, 2026
Deepseek flash 0731 reasoning is hilarious
r/LocalLLaMA
→ story
Jul 31, 2026
DeepSeek-V4 paper exposes a lesson about retries
r/LocalLLaMA
→ story
Jul 31, 2026
Unsloth releases DeepSeek-V4-Flash-0731-GGUF; users anticipate launch on Hugging Face page
r/LocalLLaMA
→ story
Jul 31, 2026
Deepseek weights when?
r/LocalLLaMA
→ story
Jul 31, 2026
Can we expect Deepseek v4 distills into smaller models?
r/LocalLLaMA
→ story
Jul 31, 2026
Floatboat DeepSeek Agent – AI with real browser access
Hacker News (AI)
→ story
Jul 31, 2026
Does the second strix halo add value with the deepseek flash update; how would it work with a Bosgame m5 and is USB4 networking fast enough?
r/LocalLLaMA
→ story
Jul 31, 2026
DeepSeek-V4-Flash-0731 is going to cause another market crash.
r/LocalLLaMA
→ story
Jul 30
Jul 30, 2026
DeepSeek plans gigawatt scale AI data center in Inner Mongolia
Hacker News (AI)
→ story
Jul 30, 2026
Show HN: Distilling DeepSeek into GPT-OSS does not transfer censorship; user evaluation suggests censorship behavior not preserved
Hacker News (AI)
→ story
Jul 30, 2026
China’s open-weight AI models are released for global download and local deployment, reflecting a combination of policy-backed openness and industrial strategy.
r/LocalLLaMA
→ story
Jul 28
Jul 28, 2026
LoRA over GGUF enables training DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM; 19 s/iteration on Strix Halo with vibe-coded Triton kernels for sliding attention, CSA, and HCA
r/LocalLLaMA
→ story
Jul 28, 2026
SWE-rebench multilingual update expands leaderboard to 5 languages with real-world software engineering tasks; GLM-5.2, DeepSeek-V4 Pro, Qwen3.6-27B and others evaluated
r/LocalLLaMA
→ story
Jul 28, 2026
DeepSeek V4 Flash reaches up to 32 tokens per second on AMD Ryzen AI MAX+ 395.
r/LocalLLaMA
→ story
Jul 28, 2026
Llama.cpp updates include chunked SSD matmul acceleration for Mamba-2 prefill and a ggml-metal FWHT kernel for the metal backend
r/LocalLLaMA
→ story
Jul 28, 2026
Spec: add DSpark speculative decoding in llama.cpp PR 25173 by ggml-org
r/LocalLLaMA
→ story
Jul 28, 2026
Update chat template for dsv4 in llama.cpp to override gguf’s template with new DeepSeek-V4.jinja
r/LocalLLaMA
→ story
Jul 26
Jul 26, 2026
Harness showdown: Claude Code, OpenCode, and Pi produce similar quality on DeepSeek V4 Flash benchmark; Claude Code takes longer despite faster tokens in some configurations
r/LocalLLaMA
→ story
Jul 25
Jul 25, 2026
Deepseek V4 flash – is Qwen3.6 27B still the most solid for agentic coding?
r/LocalLLaMA
→ story
Jul 25, 2026
What is good? Extracting and testing implicit theories of literary quality from LLM reasoning traces
arXiv cs.CL
→ story
Jul 23
Jul 23, 2026
LISA: Linear-Indexed Sparse Attention for efficient long-context reasoning
arXiv cs.AI
→ story
Jul 16
Jul 16, 2026
Moonshot AI unveils Kimi K3, a 2.8-trillion-parameter model described as their most capable to date, with open weights promised by July 27, 2026.
Simon Willison
→ story
Apr 24
Apr 24, 2026
DeepSeek V4 preview release
DeepSeek (via Google News)
→ story
Apr 24, 2026
DeepSeek V4 preview release
DeepSeek (via Google News)
→ story
Dec 1
Dec 1, 2025
DeepSeek-V3.2 Release - DeepSeek
DeepSeek (via Google News)
→ story
Sep 29
Sep 29, 2025
Introducing DeepSeek-V3.2-Exp - DeepSeek
DeepSeek (via Google News)
→ story
Sep 22
Sep 22, 2025
DeepSeek-V3.1-Terminus continues to improve unified search and data access across enterprise repositories
DeepSeek (via Google News)
→ story
Sep 22, 2025
DeepSeek-V3.1-Terminus released by DeepSeek
DeepSeek (via Google News)
→ story
Apr 28
Apr 28, 2025
Qwen3-235B-A22B and Qwen3-30B-A3B outperform several top models in benchmarks, the company says.
Qwen (Alibaba)
→ story
Mar 25
Mar 25, 2025
DeepSeek-R1 Architecture HiddenLayer
HiddenLayer (via Google News)
→ story
Mar 5
Mar 5, 2025
QwQ-32B explores scaling reinforcement learning to enhance model reasoning capabilities
Qwen (Alibaba)
→ story
Jan 30
Jan 30, 2025
DeepSh*t: exposing the security risks of DeepSeek-R1 HiddenLayer
HiddenLayer (via Google News)
→ story
Jan 28
Jan 28, 2025
Qwen2.5-Max: exploring the intelligence of a large-scale MoE model
Qwen (Alibaba)
→ story
←
1
2
3
4
→