LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Aug 12
10d ago
MindTopo reveals VLMs’ spatial reasoning abilities
Microsoft Research
→ story
11d ago
Recall is the bottleneck for parametric factuality in generative AI
Google Research
→ story
11d ago
Mitigating bus bunching with reinforcement learning enhanced by semantic stop embedding
arXiv cs.AI
→ story
11d ago
Multilingual quantization tax: edge SLMs suffer performance degradation from 4-bit weight quantization across languages, study finds across Gemma 4 and Qwen 3.5 using MMLU ProX Lite and GlobalPIQA
arXiv cs.CL
→ story
11d ago
Multimodal item parameter estimation using simulated response probabilities
arXiv cs.CL
→ story
11d ago
Neuroevolution Arena: nested ecological evaluation of update-and-inheritance regimes across neural architectures
arXiv cs.AI
→ story
11d ago
NL2SHACL-Bench: a benchmark suite for translating natural language to SHACL shapes
arXiv cs.AI
→ story
11d ago
Observational policy ranking for SMB financial guidance from multi-action accounting logs
arXiv cs.LG
→ story
11d ago
Off-Axis, On Purpose: where a Transformer computes concepts and why it does so
arXiv cs.CL
→ story
11d ago
PERCEPT: a corpus for POS tagging and analysis of Persian-English code-mixing
arXiv cs.CL
→ story
11d ago
Physics-informed machine learning in prognostics and health management: a systematic literature review
arXiv cs.LG
→ story
11d ago
Position encoding in transformers: from absolute and relative methods to rotary position embeddings and long-context scaling
arXiv cs.CL
→ story
11d ago
Post-hoc sparse coding of latent communication between vision-language model agents
arXiv cs.AI
→ story
11d ago
Power law graph attention: exact generalization of scaled dot-product attention; empirical collapse at inference
arXiv cs.LG
→ story
11d ago
Procedural fairness failures in RLHF from preference averaging
arXiv cs.LG
→ story
11d ago
Protecting patient privacy in clinical foundation models: technical and legal perspectives
arXiv cs.AI
→ story
11d ago
QuantumMind proposes an auditable agentic workflow to generate and conservatively screen quantum-acceleration hypotheses for speedup analysis in quantum computing
arXiv cs.AI
→ story
11d ago
REATS: LLM reasoning-based ensemble learning for adaptive time series forecasting
arXiv cs.LG
→ story
11d ago
ReCBM: uncertainty-gated relational reasoning for concept bottleneck models
arXiv cs.AI
→ story
11d ago
Relational geometry attacks on contrastive embedding manifolds emerge in new study
arXiv cs.AI
→ story
11d ago
SeFoRA uses sketch-aggregated federated LoRA to handle heterogeneous client ranks in federated parameter-efficient fine-tuning with low-rank adaptation.
arXiv cs.LG
→ story
11d ago
Sheaf-based federated representation learning
arXiv cs.LG
→ story
11d ago
Similarity gates approve reversals: a validity audit of embedding-cosine thresholds in agent systems
arXiv cs.CL
→ story
11d ago
Simplex Relaxation for Discrete Diffusion; arXiv:2608.10615 introduces Simplax, an exact Dirichlet–categorical augmentation that enriches training objectives and reverse transitions without changing the underlying corruption process
arXiv cs.CL
→ story
11d ago
SPOT: Sampling Policy Observation Tree provides lookahead explanations for deep reinforcement learning policies.
arXiv cs.AI
→ story
11d ago
STCAD: scalable trajectory clustering and anomaly detection on terabyte-scale AIS data
arXiv cs.LG
→ story
11d ago
TAF-MED measures multi-turn safety refusal collapse in LLMs when users express self-treatment intent
arXiv cs.CL
→ story
11d ago
TeXFix-Bench: an empirically grounded multi-format benchmark for LLM-based full-source document repair
arXiv cs.AI
→ story
11d ago
Toward human rights benchmarking for LLMs: a pilot methodology
arXiv cs.LG
→ story
11d ago
Towards an argumentative foundation for evaluative AI
arXiv cs.AI
→ story
11d ago
Towards researcher agents for knowledge-graph question answering
arXiv cs.AI
→ story
11d ago
Towards Sustainable AI: a comprehensive review and comparative analysis of deep learning models’ carbon footprint
arXiv cs.AI
→ story
11d ago
TRACE: Trustworthy retrieval-augmented conversational engine discusses improving reliability of public-service chatbot recommendations with retrieval augmentation
arXiv cs.AI
→ story
11d ago
Training Variable Long Sequences with Data-Centric Parallel
arXiv cs.AI
→ story
11d ago
Transformer Geometry Observatory TGO-IV: Developmental topology of representations across transformer layers explored via inductive analysis
arXiv cs.LG
→ story
11d ago
TREAT: evaluating access to formal knowledge across equivalent mathematical representations
arXiv cs.AI
→ story
11d ago
Uncertainty-aware ensemble deep randomized neural networks for classification
arXiv cs.LG
→ story
11d ago
UserToolBench tests personalized decision making in tool-use LLMs by inferring latent user preferences from interaction history
arXiv cs.LG
→ story
11d ago
When chain-of-thought helps and when it hurts: an empirical investigation of the serial-depth bottleneck in LLM reasoning
arXiv cs.CL
→ story
11d ago
When LLM agents negotiate: private information and dynamic bargaining in supply chains
arXiv cs.AI
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→