LIVE · refreshes every 20 min
updated Sep 2, 17:22 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 25
Jul 25, 2026
A three-stage neural-symbolic pipeline using an ensemble of DeBERTa-v3-base and XLM-RoBERTa-base with a Linguistically-Informed Mediator for gaming toxicity detection.
arXiv cs.CL
→ story
Jul 25, 2026
Answer-then-Edit: Reasoning skeleton editing for anti-distillation with preserved utility
arXiv cs.CL
→ story
Jul 25, 2026
AsymVerify achieves 2nd place on SemEval-2026 Task 6 with a confidence-gated verification approach for political evasion detection
arXiv cs.CL
→ story
Jul 25, 2026
Belief propagation in LLM world models: measuring strategic information bias with prediction markets
arXiv cs.CL
→ story
Jul 25, 2026
Break Through the Compression Bottleneck: From Theory to Practice
arXiv cs.CL
→ story
Jul 25, 2026
Confidently deceptive: how confidence amplifies the risk of LLM deception
arXiv cs.CL
→ story
Jul 25, 2026
Distinguishing artificial from authentic: evaluating LLMs for detecting LLM-generated content
arXiv cs.CL
→ story
Jul 25, 2026
GLAN-QnA-KR: a 303,581-row Korean instruction-QA corpus produced by seedless taxonomy-driven GLAN synthesis using Microsoft Phi-3.5-MoE-instruct; open and redistributable under OpenRAIL
arXiv cs.CL
→ story
Jul 25, 2026
Human-in-the-loop LLM framework improves detection of cutaneous immune-related adverse events in clinical notes; study reports higher accuracy and agreement vs manual review
arXiv cs.CL
→ story
Jul 25, 2026
Knowledge injection exists in MoE; exploring expert-aware contrast decoding in MoE for mitigating LLMs' hallucinations
arXiv cs.CL
→ story
Jul 25, 2026
LLM-INSTRUCT wins UZH Shared Task 2026 on paragraph-level argument mining with constraint-aware retrieval and selective debate
arXiv cs.CL
→ story
Jul 25, 2026
MoE routing follows a Huffman code pattern, as the Frequency-Diversity Law shows that state-of-the-art models act as information-theoretic engines.
arXiv cs.CL
→ story
Jul 25, 2026
Moir: Let the model direct its own story for robust cross-domain knowledge editing
arXiv cs.CL
→ story
Jul 25, 2026
More Is Not More: what matters for diversity in LLM opinions?
arXiv cs.CL
→ story
Jul 24
Jul 24, 2026
InferenceBench: a benchmark for open-ended LLM inference optimization by AI agents
arXiv cs.AI
→ story
Jul 24, 2026
JAXBench: Benchmarking autonomous TPU kernel optimization
arXiv cs.AI
→ story
Jul 24, 2026
Leveraging biokinetic knowledge priors for data-scarce bioprocess modeling
arXiv cs.LG
→ story
Jul 24, 2026
Lie typology, depth, and sparsity affect deception detection in LLM outputs
arXiv cs.AI
→ story
Jul 24, 2026
Marking the wrong symptoms: evaluating LLM watermarks in medical texts
arXiv cs.AI
→ story
Jul 24, 2026
Multimodal CoLRAG-TF uses four-axis fusion—dense text embeddings, BM25, knowledge-graph triple filtering, and image similarity—to enable robust retrieval over complex PDFs.
arXiv cs.LG
→ story
Jul 24, 2026
Muon optimizer reaches grokking threshold on modular arithmetic faster than AdamW; ablation shows speedup from Newton-Schulz orthogonalization
arXiv cs.LG
→ story
Jul 24, 2026
Optimal noise allocation for diffusion training in the convex regime
arXiv cs.LG
→ story
Jul 24, 2026
OPTScientist: multi-agent discovery of typed optimizer programs for transformer pretraining
arXiv cs.AI
→ story
Jul 24, 2026
PersonaTrail benchmarks personalized web agents through browsing histories to evaluate how agents infer context from user browsing data
arXiv cs.AI
→ story
Jul 24, 2026
PhantomFill: when the form demands an answer, language models invent one
arXiv cs.LG
→ story
Jul 24, 2026
PlanE: Meta planning of data, tuning, and inference for extractive-based LLMs
arXiv cs.AI
→ story
Jul 24, 2026
Proactive test-driven AI development is proposed to replace reactive patching of models based on user feedback.
arXiv cs.LG
→ story
Jul 24, 2026
ReliableTableQA evaluates the amount of supervision needed for reliability annotation of tabular QA results
arXiv cs.LG
→ story
Jul 24, 2026
Routing Without Training: Controllable-Ratio LLM Offloading via Reliability Gating
arXiv cs.AI
→ story
Jul 24, 2026
Scaling closed-loop feature channel configuration with LLMs
arXiv cs.LG
→ story
Jul 24, 2026
Semi-supervised text-attributed graph distillation
arXiv cs.AI
→ story
Jul 24, 2026
SenCos-GEM: SENet-Calibrated and Law-of-Cosines-Constrained geometry-enhanced molecular representation for property prediction
arXiv cs.LG
→ story
Jul 24, 2026
SevDiff: Severity-conditioned diffusion for long-tail conflict trajectory generation
arXiv cs.LG
→ story
Jul 24, 2026
SOAP, Muon, and beyond: pushing LLM pretraining scales
arXiv cs.LG
→ story
Jul 24, 2026
SonicSampler: unified tile-aware kernels for LLM sampling and speculative verification
arXiv cs.AI
→ story
Jul 24, 2026
Stochastic sampling is epistemically shallow; the dimensionality gap between temperature variation and model diversity in LLMs
arXiv cs.AI
→ story
Jul 24, 2026
The Devil is in the spectrum: mitigating representation collapse in LLMs via topologically regularized side-path
arXiv cs.AI
→ story
Jul 24, 2026
Tractable hierarchical control of autoregressive language models
arXiv cs.AI
→ story
Jul 24, 2026
Uncertainty-aware trust estimation for multi-LLM systems via structured expert judgement
arXiv cs.LG
→ story
Jul 24, 2026
When RLVR shrinks the reasoning boundary: diagnosing pass@k inversion
arXiv cs.LG
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
→