LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 27
27d ago
A consensus-based framework for relative preference evaluation of large language models
arXiv cs.CL
→ story
27d ago
A drift-stable quantum federated learning for intelligent services
arXiv cs.LG
→ story
27d ago
Adjustment speed as a safety constraint for nonstationary reinforcement learning
arXiv cs.LG
→ story
27d ago
Adversarial style optimization: enhancing VLM jailbreaks by GRPO-based stylistic triggers optimization
arXiv cs.CL
→ story
27d ago
Agentic evaluation of copyright law compliance with Copyright-Bench assessing LLM agents’ compliance with copyright law
arXiv cs.CL
→ story
27d ago
Analyzing self-harm representations in language models: a cross-architecture study
arXiv cs.CL
→ story
27d ago
Analyzing toxic behavior and its impact on the Mastodon community
arXiv cs.CL
→ story
27d ago
Benchmarking fine-tuning and retrieval strategies for a multimodal language model on the NRC Reactor Operator licensing examination
arXiv cs.CL
→ story
27d ago
Bounding the causal impact of ML-assisted decision-making via counterfactual correctness
arXiv cs.LG
→ story
27d ago
CARNet: cycle-conditioned core aggregation and redistribution for multivariate time series forecasting
arXiv cs.LG
→ story
Jul 26
28d ago
Teaching LLMs to update beliefs for efficient long-horizon interaction
BAIR (Berkeley)
→ story
Jul 25
29d ago
A three-stage neural-symbolic pipeline using an ensemble of DeBERTa-v3-base and XLM-RoBERTa-base with a Linguistically-Informed Mediator for gaming toxicity detection.
arXiv cs.CL
→ story
29d ago
Answer-then-Edit: Reasoning skeleton editing for anti-distillation with preserved utility
arXiv cs.CL
→ story
29d ago
AsymVerify achieves 2nd place on SemEval-2026 Task 6 with a confidence-gated verification approach for political evasion detection
arXiv cs.CL
→ story
29d ago
Belief propagation in LLM world models: measuring strategic information bias with prediction markets
arXiv cs.CL
→ story
29d ago
Break Through the Compression Bottleneck: From Theory to Practice
arXiv cs.CL
→ story
29d ago
Confidently deceptive: how confidence amplifies the risk of LLM deception
arXiv cs.CL
→ story
29d ago
Distinguishing artificial from authentic: evaluating LLMs for detecting LLM-generated content
arXiv cs.CL
→ story
29d ago
GLAN-QnA-KR: a 303,581-row Korean instruction-QA corpus produced by seedless taxonomy-driven GLAN synthesis using Microsoft Phi-3.5-MoE-instruct; open and redistributable under OpenRAIL
arXiv cs.CL
→ story
29d ago
Human-in-the-loop LLM framework improves detection of cutaneous immune-related adverse events in clinical notes; study reports higher accuracy and agreement vs manual review
arXiv cs.CL
→ story
29d ago
Knowledge injection exists in MoE; exploring expert-aware contrast decoding in MoE for mitigating LLMs' hallucinations
arXiv cs.CL
→ story
29d ago
LLM-INSTRUCT wins UZH Shared Task 2026 on paragraph-level argument mining with constraint-aware retrieval and selective debate
arXiv cs.CL
→ story
29d ago
MoE routing follows a Huffman code pattern, as the Frequency-Diversity Law shows that state-of-the-art models act as information-theoretic engines.
arXiv cs.CL
→ story
29d ago
Moir: Let the model direct its own story for robust cross-domain knowledge editing
arXiv cs.CL
→ story
29d ago
More Is Not More: what matters for diversity in LLM opinions?
arXiv cs.CL
→ story
29d ago
Natural language should not fully replace formal languages, argues a position paper highlighting underspecification in open-ended contexts
arXiv cs.CL
→ story
29d ago
Naver-News-KO: a Korean news summarization dataset for open-source fine-tuning of summarization models
arXiv cs.CL
→ story
29d ago
Open-source text LLM watermarks fail to withstand model merging, study finds
arXiv cs.CL
→ story
29d ago
Preference tuning as spectral update reorganization
arXiv cs.CL
→ story
29d ago
Routing Subspaces: Auditing evaluation-to-deployment mismatch in fine-tuned language models
arXiv cs.CL
→ story
29d ago
SCoPE: Shift-aware speaker-conditioned priors for emotion recognition in conversations
arXiv cs.CL
→ story
29d ago
Skill-contracted agents for evidence-aware materials literature analysis
arXiv cs.CL
→ story
29d ago
The Storyteller in the Model: Narrative pattern inheritance, escalation dynamics, and alignment governance in LLMs
arXiv cs.CL
→ story
29d ago
TopoGuard: graph theory based defenses against split-knowledge attacks on RAG
arXiv cs.CL
→ story
29d ago
What is good? Extracting and testing implicit theories of literary quality from LLM reasoning traces
arXiv cs.CL
→ story
Jul 24
Jul 24, 2026
Stochastic sampling is epistemically shallow; the dimensionality gap between temperature variation and model diversity in LLMs
arXiv cs.AI
→ story
Jul 24, 2026
The Devil is in the spectrum: mitigating representation collapse in LLMs via topologically regularized side-path
arXiv cs.AI
→ story
Jul 24, 2026
Tractable hierarchical control of autoregressive language models
arXiv cs.AI
→ story
Jul 24, 2026
Uncertainty-aware trust estimation for multi-LLM systems via structured expert judgement
arXiv cs.LG
→ story
Jul 24, 2026
When RLVR shrinks the reasoning boundary: diagnosing pass@k inversion
arXiv cs.LG
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→