LIVE · refreshes every 20 min
updated Sep 3, 20:02 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
GPT-4
Aug 17
17d ago
Qwen3.5 9B with vision weighs less than 7 GB and far surpasses GPT-4o, according to a user discussion
r/LocalLLaMA
→ story
Aug 13
21d ago
Fine-tuned Qwen2.5-Coder-1.5B on 125k command pairs can write shell commands on a laptop CPU in about 1 second; runs via llama.cpp on 1.6GB RAM with 31.9 tok/s and 0.59s median per query
r/LocalLLaMA
→ story
Aug 7
27d ago
Universal Pathologies, Conditional Consequences: A triple-robustness analysis of RAG for multi-hop traceability
arXiv cs.CL
→ story
Jul 30
Jul 30, 2026
llm 0.32rc2 release fixes dependency issue and updates default model to GPT-5.6 Luna; users can switch back to 4o mini via llm models default gpt-4o-mini
Simon Willison
→ story
Jul 7
Jul 7, 2026
AI model inference costs fall sharply; GPT-4-class costs drop from ~$30 per million tokens in 2023 to under $1 today
BAIR (Berkeley)
→ story
Nov 11
Nov 11, 2024
Qwen2.5-Coder series open-sources the “Powerful”, “Diverse”, and “Practical” open-code LLMs, including Qwen2.5-Coder-32B-Instruct as a state-of-the-art open-source code model.
Qwen (Alibaba)
→ story