LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 17
Jul 17, 2026
arXiv:2607.14112v1 paper shows language models have reliability ceilings from resolvable output uncertainty limits
arXiv cs.CL
→ story
Jul 16
Jul 16, 2026
Muon emerges as strong optimizer for deep learning, outperforming Adam and AdamW; theoretical work interprets it as steepest descent under spectral norm.
arXiv cs.LG
→ story
Jul 16, 2026
MyAG is a graph-based framework for designing and analyzing composable LLM agent systems with three graph abstractions.
arXiv cs.CL
→ story
Jul 16, 2026
Networked intelligence: active shared context graphs for human-AI team science
arXiv cs.AI
→ story
Jul 16, 2026
New paper explores federated learning and explainable AI integration for privacy-preserving transparent machine learning
arXiv cs.LG
→ story
Jul 16, 2026
OpenRCA dataset highlights challenges in achieving high accuracy for root cause analysis on real-world telemetry data.
arXiv cs.AI
→ story
Jul 16, 2026
Oracle Agent Memory as an enterprise memory substrate for long-horizon AI agents
arXiv cs.AI
→ story
Jul 16, 2026
OriginBlame introduces record- and token-level data provenance for AI training datasets
arXiv cs.AI
→ story
Jul 16, 2026
Probabilistic extension of neuro-symbolic AGI robots based on Belnap's typed intensional FOL
arXiv cs.AI
→ story
Jul 16, 2026
RAGthoven presents multi-stage LLM pipeline for SemEval-2026 Task 1's multilingual humor generation in English, Spanish, and Chinese.
arXiv cs.CL
→ story
Jul 16, 2026
Researchers introduce a benchmark and reference system for live captioning Sikh Kirtan; arXiv:2607.13457v1 presents a closed-vocabulary task requiring exact SGGS line transcription.
arXiv cs.CL
→ story
Jul 16, 2026
Researchers introduce black-box interventional grounding audits to test LLM premise dependency via predicate substitution; method replaces target predicates with fresh symbols and re-runs models to assess reasoning changes.
arXiv cs.AI
→ story
Jul 16, 2026
Researchers introduce CoDiffGRN, a method for gene regulatory network inference using BEELINE-KGC benchmark and co-evolutionary discrete diffusion.
arXiv cs.LG
→ story
Jul 16, 2026
Researchers introduce DROPJ, a human-centered method for safe agent training and deployment in safety-critical environments with unknown dynamics and no reward function.
arXiv cs.AI
→ story
Jul 16, 2026
Researchers present EZSMT Version 3, an extensible SMT-based CASP framework.
arXiv cs.AI
→ story
Jul 16, 2026
Researchers propose a meta-learning framework to address data scarcity in low-resource language alignment for multilingual LLMs; arXiv:2607.13315v1
arXiv cs.CL
→ story
Jul 16, 2026
Researchers propose lightweight training strategy for efficient transfer learning by decoupling feature extraction and classifier optimization with margin-based weighted loss.
arXiv cs.LG
→ story
Jul 16, 2026
Researchers propose Samba, a hybrid Mamba model for audio-visual navigation, addressing inadequacies of existing frameworks since 2020.
arXiv cs.LG
→ story
Jul 16, 2026
Researchers use persona vectors to audit open-weight LLMs, revealing 53 traits across four models in arXiv:2607.13162v1.
arXiv cs.CL
→ story
Jul 16, 2026
Safe-Psych, a sequential evaluation benchmark for LLMs in psychiatry, addresses incomplete information by requiring clarification.
arXiv cs.CL
→ story
Jul 16, 2026
Safety Sentry: Context-aware human intervention via execute-ask-refuse routing
arXiv cs.AI
→ story
Jul 16, 2026
Set-shifting benchmark tests how LLM agents adapt to hidden reliability shifts in tool availability
arXiv cs.AI
→ story
Jul 16, 2026
ShortOPD: recovering pruned LLMs with short-to-long on-policy distillation
arXiv cs.LG
→ story
Jul 16, 2026
Small language models for biomedical data-to-text generation: a case study on medication leaflets and post-training alignment methods
arXiv cs.CL
→ story
Jul 16, 2026
SPINE: bridging the cyber-physical gap with agentic AI
arXiv cs.AI
→ story
Jul 16, 2026
SteinGate: tail-sensitive safe reinforcement learning via Stein discrepancy
arXiv cs.LG
→ story
Jul 16, 2026
STKAN: Kolmogorov-Arnold networks for spatio-temporal forecasting
arXiv cs.LG
→ story
Jul 16, 2026
Stocktake: Measuring the gap between perception and action in LLM agents with a fair oracle
arXiv cs.AI
→ story
Jul 16, 2026
Study compares KANs and MLPs on structured data classification using twelve datasets.
arXiv cs.LG
→ story
Jul 16, 2026
Study introduces BioASQ Task 14B 2026 system using hybrid retrieval and multi-model answer combination.
arXiv cs.CL
→ story
Jul 16, 2026
Survey Examines Self-Improving Agentic Systems' Shift to Deployment
arXiv cs.AI
→ story
Jul 16, 2026
Tabular foundation models for discrete choice estimation show limited performance due to row-independence assumptions
arXiv cs.LG
→ story
Jul 16, 2026
Targeted PD identifies only the components that process specific inputs to scale parameter decomposition in neural networks.
arXiv cs.LG
→ story
Jul 16, 2026
Text2Sign, a text-to-sign language video diffusion model, runs on a single NVIDIA L4 GPU as a baseline.
arXiv cs.CL
→ story
Jul 16, 2026
Theory-Level Autoformalization: from isolated statements to unified formal knowledge bases
arXiv cs.AI
→ story
Jul 16, 2026
TSSM model enhances global station weather forecasting by addressing accuracy limitations through temporal-variable-historical modeling.
arXiv cs.LG
→ story
Jul 16, 2026
UESF-Bench benchmarks unified embodied seeking and following, addressing existing benchmarks' assumption of initial target visibility.
arXiv cs.AI
→ story
Jul 16, 2026
Uncertainty-aware sequential decision rules for event-triggered LLM invocation in streaming systems
arXiv cs.LG
→ story
Jul 16, 2026
Weight feedback computes the Jacobian transpose locally in modern deep networks
arXiv cs.LG
→ story
Jul 16, 2026
Where should RL post-training compute go? Model size, search, learning, and feedback
arXiv cs.LG
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→