LIVE · refreshes every 20 min
updated Aug 12, 01:42 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
GPT-5
Aug 11
11h ago
GPT-5.6-Sol Pro achieves 67.30% on Claude’s Reimann paper evaluation
Hacker News (AI)
→ story
Aug 10
1d ago
Model ML uses GPT-5.6 Sol to carry finance work from research and analysis through editable, traceable PowerPoint decks and Excel workbooks.
OpenAI
→ story
1d ago
OpenAI releases GPT-5.6-Cyber through Daybreak Red for authorized vulnerability research and security testing
OpenAI
→ story
Aug 7
4d ago
Universal Pathologies, Conditional Consequences: A triple-robustness analysis of RAG for multi-hop traceability
arXiv cs.CL
→ story
Aug 2
9d ago
June newsletter: highlights include accidental cyberattacks by OpenAI and Anthropic models under test, GPT-5.6 Sol/Terra/Luna, Claude Opus 5, Kimi K3 and DeepSeek-V4-Flash-0731, and other model releases
Simon Willison
→ story
Jul 30
12d ago
llm 0.32rc2 release fixes dependency issue and updates default model to GPT-5.6 Luna; users can switch back to 4o mini via llm models default gpt-4o-mini
Simon Willison
→ story
12d ago
OpenAI makes frontier models cheaper with GPT-5/6 price-performance improvements
r/LocalLLaMA
→ story
Jul 29
13d ago
Two API settings tripled scores on the ARC-AGI-3 benchmark by GPT-5.6, boosting performance and efficiency through retained reasoning and enabled compaction.
OpenAI
→ story
13d ago
GPT-5.6 vs. Claude Fable 5 for physical AI: which performs best?
Hacker News (AI)
→ story
13d ago
Personalization, personas, and forecasting in value alignment
arXiv cs.AI
→ story
14d ago
GPT-5.6 improves AI efficiency across models, inference, and agentic workflows to deliver more useful intelligence per dollar.
OpenAI
→ story
Jul 26
16d ago
Harness showdown: Claude Code, OpenCode, and Pi produce similar quality on DeepSeek V4 Flash benchmark; Claude Code takes longer despite faster tokens in some configurations
r/LocalLLaMA
→ story
Jul 25
17d ago
How much are you actually using your local models these days; which ones do you reach for the most?
r/LocalLLaMA
→ story
Jul 23
19d ago
Structured synthetic reasoning data improves arithmetic fine-tuning for small language models, using a 21,250-example GPT-5-mini-generated corpus derived from GSM8K
arXiv cs.AI
→ story
Jul 22
20d ago
GPT-5.6 got smarter; then it kept acting.
Hacker News (AI)
→ story
20d ago
We probed a pinned GPT-5.5 endpoint; every request carried about 1,447 hidden tokens
Hacker News (AI)
→ story
20d ago
Relay-Bench evaluates LLMs on multi-domain reasoning chains; GPT-5.5 (xHigh) scores 43.3% on composite problems across domains
arXiv cs.CL
→ story
Jul 21
21d ago
Drawing the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
Hacker News (AI)
→ story
Jul 20
22d ago
GPT-5.6 Sol and Kimi K3 compete in Kerbal Space Program speedrunning live event
Hacker News (AI)
→ story
22d ago
Cost-efficient generative AI summarization for scalable automated essay scoring in educational assessment
arXiv cs.CL
→ story
Jul 18
24d ago
Claude Fable 5 to be included in all Max and Team Premium plans; Pro and Team Standard keep access via credits with a $100 one-time credit
Simon Willison
→ story
Jul 17
25d ago
GPT-5.6 Sol Ultra constructed Chrome V8 exploit chain from patch commits
Hacker News (AI)
→ story
25d ago
GPT-5.6 Sol Max released; analysts assess its value
Hacker News (AI)
→ story
Jul 16
26d ago
Gpt-5.6 Sol Pro solves open problem in convex optimization.
Hacker News (AI)
→ story
26d ago
Moonshot AI unveils Kimi K3, a 2.8-trillion-parameter model described as their most capable to date, with open weights promised by July 27, 2026.
Simon Willison
→ story
26d ago
Datasette code-frequency chart on GitHub shows spikes in activity aligned with Opus 4.8, GPT-5.5, Fable 5 and GPT-5.6 Sol.
Simon Willison
→ story
26d ago
DOOMQL uses SQLite as the game engine for a Python terminal Doom-like game; project by Peter Gostev built with GPT-5.6 Sol
Simon Willison
→ story
26d ago
AI music video created for $100; compares Claude Fable 5 and GPT-5.6 Sol
Hacker News (AI)
→ story
26d ago
OpenAI admits GPT-5.6 occasionally deletes files, calls it an 'honest mistake'
Simon Willison
→ story
Jul 9
Jul 9, 2026
GPT-5.6 family introduces Luna, Terra, and Sol with increased efficiency and on-demand capability; Microsoft 365 Copilot adopts GPT-5.6 as preferred model
OpenAI
→ story
Jul 9, 2026
OpenAI's Bio Bounty program targets GPT-5.5
OpenAI
→ story