LIVE · refreshes every 20 min
updated Sep 3, 18:23 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ topic
Qwen
Sep 3
14h ago
Post-training ternarization of Qwen3-4B: effective bit budget, storage compression, and deployment
arXiv cs.AI
→ story
17h ago
Show HN: DeltaCode beats Qwen's zvec-grep on our test (92% vs 48%)
Hacker News (AI)
→ story
Sep 1
2d ago
The Halt Vector internalizes a causal steering intervention to halt chain-of-thought reasoning for more efficient reasoning
arXiv cs.LG
→ story
Aug 31
3d ago
Below the noise floor: bimodal seed collapse and distinct failure modes in small-model knowledge distillation
arXiv cs.CL
→ story
3d ago
First make it playable, then make it good: staged interaction learning for small dialogue-game agents
arXiv cs.CL
→ story
Aug 21
13d ago
Qwen image 3.0 Pro vs. GPT image 2 for production image APIs
Hacker News (AI)
→ story
Aug 19
14d ago
Ornith-1.5 9B may not be bad after all
r/LocalLLaMA
→ story
15d ago
QWEN plans to reuse the 397B-A17B architecture to compete with Deepseek V4 0731 Flash
r/LocalLLaMA
→ story
15d ago
NVFP4 on VOLTA matches RTX 5090 for Qwen 3.8 in FP4/FP8 on four 2017 Tesla V100s
r/LocalLLaMA
→ story
15d ago
Prompt extend model fine-tuned from gemma-4-12B-it for Qwen Image Edit 2511 generates enhanced editing prompts using PERL with ROLL; Kimi K2.6 serves as reward worker to evaluate results
r/LocalLLaMA
→ story
15d ago
llama.cpp adds --n-cpu-ffn option for Dense models (building on --n-cpu-moe / --cpu-moe for MOE models) via pull request 26622
r/LocalLLaMA
→ story
15d ago
Qwen 3.8 27B performs worse than Claude Code and GitHub Copilot for agentic coding, according to user experience with local models and tools
r/LocalLLaMA
→ story
Aug 18
15d ago
DFlash 2 available for Qwen 3.8 27B and Muse Glimmer
r/LocalLLaMA
→ story
15d ago
Alibaba's RISC-V CPU XuanTie C950 runs Qwen-3.8 27B at 30 tps
r/LocalLLaMA
→ story
15d ago
Qwen 3.8 27B matches frontier intelligence and outperforms Google’s current frontier model without a data center; most local model experts can run it on their own hardware.
r/LocalLLaMA
→ story
15d ago
What’s the best tool for offline Wikipedia RAG right now?
r/LocalLLaMA
→ story
16d ago
Idea: compress Qwen 3.8 KV cache by using a single bit for the token "wait
r/LocalLLaMA
→ story
16d ago
Qwen3.8 2.4T open weights replicate a Call of Duty clone in one prompt; model run uses 1.1M tokens over 5 hours on rented hardware
r/LocalLLaMA
→ story
16d ago
Qwen 4 expected for September
r/LocalLLaMA
→ story
16d ago
Agent on CPU, which to pick?
r/LocalLLaMA
→ story
16d ago
OpenCode overrides Qwen model samplers to top-p 1.0 instead of 0.95 or 0.80
r/LocalLLaMA
→ story
16d ago
Llama.cpp lacks support for changing thinking amount; user requests on-the-fly mode switching with qwen 3.8 27b referenced
r/LocalLLaMA
→ story
16d ago
MacBook becomes an AI workstation by running Qwen 3.8 locally
Hacker News (AI)
→ story
16d ago
Qwen 3.8 27b saved me $650+ in API costs this evening.
r/LocalLLaMA
→ story
16d ago
AA says Qwen3.8 27B ships with xhigh reasoning by default to maximize benchmark performance
r/LocalLLaMA
→ story
Aug 17
16d ago
Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index, tying with GPT-5.6 Luna Max and trailing GLM-5.2 Max and DeepSeek V4 Pro 0813 Max
Simon Willison
→ story
16d ago
Waiting for Qwen 3.8 35B A3B
r/LocalLLaMA
→ story
16d ago
Qwen3.8-27B Uncensored Aggressive release adds K_P quant range, Vision, native NextN, and HauhauCS FastMTP
r/LocalLLaMA
→ story
16d ago
Opus 4.8, Muse Glimmer, Gemma 4 draw criticism for reasoning behavior; user complaints discussed in thread
r/LocalLLaMA
→ story
16d ago
Qwen 3.8 27B outperforms alternatives and raises pricing concerns; user questions its reasonableness and notes continued Alibaba training after prior release
r/LocalLLaMA
→ story
17d ago
Ling 3.0 Tiny 8b runs fastest on low-end PC with 4GB VRAM, 36 tokens/sec; claims strong performance vs Qwen 3.5 9b and Gemma 12
r/LocalLLaMA
→ story
17d ago
User pairs DeepSeek V4 Flash with Antirez's Dwarfstar 4 framework -- personal review
r/LocalLLaMA
→ story
17d ago
Mimir: a 1.7B model claims to outperform Qwen 3.5 0.8B and Gemma 4 E2B on benchmarks (english and danish; builds on sapient’s hrm-text)
r/LocalLLaMA
→ story
17d ago
Qwen 3.8 runs ThinkingCap LoRA designed for Qwen 3.6; prompts show partial compatibility in early tests
r/LocalLLaMA
→ story
17d ago
Qwen3.5 9B with vision weighs less than 7 GB and far surpasses GPT-4o, according to a user discussion
r/LocalLLaMA
→ story
17d ago
Qwen 3.8 27B defeats SOL in Q4 tests on complex animated SVG tasks; prompts include drone park view, beach scene with evil cat, and AGI sandbox scenario
r/LocalLLaMA
→ story
Aug 16
17d ago
Qwen 3.8 27b performs better than 3.6 27b on Turtle graphics tasks, according to a user prompt requesting complete working Python code for a recursive tree using the Turtle library.
r/LocalLLaMA
→ story
17d ago
Qwen 3.8 27B shows strong performance but tends to overthink tasks, according to its self-reported benchmarks and related coverage.
Simon Willison
→ story
17d ago
Qwen 3.8 27B reasoning dialogue prompts internal monologue and expressions of frustration, self-doubt
r/LocalLLaMA
→ story
18d ago
Qwen 3.8 fixes fixed templates for Claude compatibility
r/LocalLLaMA
→ story
←
1
2
3
4
5
6
→