LIVE · refreshes every 20 min
updated Sep 2, 17:22 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 23
Jul 23, 2026
Adaptive capitulation: a structural failure mode of LLM responses in vulnerability contexts
arXiv cs.CL
→ story
Jul 23, 2026
AdaRoPE: not all attention heads should rotate and scale equally
arXiv cs.AI
→ story
Jul 23, 2026
AI/ML deepfake research misaligned with AI-generated non-consensual intimate imagery, according to landscape analysis of highly cited works
arXiv cs.AI
→ story
Jul 23, 2026
Air Quality Arena: a large-scale multi-region ground monitoring dataset and benchmark for air quality forecasting with time-series foundation models
arXiv cs.LG
→ story
Jul 23, 2026
BatchDAG: LLM-planned execution graphs for scalable ad-hoc analysis over enterprise data
arXiv cs.AI
→ story
Jul 23, 2026
Bayesian wind tunnels for model selection
arXiv cs.LG
→ story
Jul 23, 2026
Benchmarking confidential GPU inference on NVIDIA H100 under Intel TDX
arXiv cs.AI
→ story
Jul 23, 2026
Beyond Tracking or Shortcut: Composition-bounded predictive states in poker autoregressive models
arXiv cs.AI
→ story
Jul 23, 2026
Calibrated selective fact-checking via evidence chain evaluation.
arXiv cs.AI
→ story
Jul 23, 2026
CrackedPDFs: a controlled benchmark for hidden prompt injection in PDFs
arXiv cs.AI
→ story
Jul 23, 2026
Cross-dialect generalization without retraining: benchmarks and evaluation of schema-derived constrained decoding for MLIR
arXiv cs.AI
→ story
Jul 23, 2026
Cross-subject semantic decoding with shared-space alignment for generalized neural representation learning
arXiv cs.LG
→ story
Jul 23, 2026
CruiseBench: a real-flight-aligned N-CMAPSS benchmark for engine RUL prediction
arXiv cs.LG
→ story
Jul 23, 2026
Decodable but Not Detectable: A Leakage Fingerprint for Near-OOD Benchmarks reveals a benchmark leak where removing the OOD class and retraining 35 models raises AUROC from 0.326 to 0.911.
arXiv cs.LG
→ story
Jul 23, 2026
Embedding-based measurement of data diversity for NLP; emb-diversity tool assesses diversity using embeddings
arXiv cs.CL
→ story
Jul 23, 2026
Euclean: automated geometry problem formalization with unified verification in Lean
arXiv cs.AI
→ story
Jul 23, 2026
Explainability challenges in continual learning for time series forecasting with Experience Replay strategies
arXiv cs.LG
→ story
Jul 23, 2026
Fence: Specialized SLM guardrails for LLM applications
arXiv cs.AI
→ story
Jul 23, 2026
FindStatBench: evaluating large language models on combinatorial code synthesis
arXiv cs.AI
→ story
Jul 23, 2026
FineServe: a fine-grained dataset and characterization of global LLM serving workloads
arXiv cs.AI
→ story
Jul 23, 2026
FinMMEval 2026 Task 1 evaluates multilingual financial multiple-choice question answering in English, Chinese, Arabic, and Hindi.
arXiv cs.CL
→ story
Jul 23, 2026
FORCE-Bench: a benchmark, dataset, and evaluation harness for agentic AI in enterprise finance
arXiv cs.AI
→ story
Jul 23, 2026
FormulaSPIN: Self-play fine-tuning for natural language to spreadsheet formula generation
arXiv cs.AI
→ story
Jul 23, 2026
From Agent Failure Paths to Quantified Residual Risk: A Compositional Framework for Resilient Agentic AI
arXiv cs.AI
→ story
Jul 23, 2026
From trajectories to prefixes: reusing teacher trajectories via replayed prefixes and online continuation
arXiv cs.LG
→ story
Jul 23, 2026
Geometry-Guided Constraint Learning for LLM safety classification achieves near-perfect per-category accuracy with sparse autoencoder feature extraction, reducing constraint counts from K=4-25 to K=2 for most categories
arXiv cs.AI
→ story
Jul 23, 2026
GraphContainer provides a unified platform for comparing and debugging Graph RAG methods
arXiv cs.AI
→ story
Jul 23, 2026
Hybrid LSTM-Graph neural framework for robust financial fraud detection and adversarial resilience
arXiv cs.AI
→ story
Jul 23, 2026
HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions
arXiv cs.AI
→ story
Jul 23, 2026
Integro-differential equations in angular stabilization of drone motion by distributed feedback control
arXiv cs.AI
→ story
Jul 23, 2026
ITPEval benchmarks automated translation of formal proofs across Lean 4, Rocq, Isabelle, and HOL Light
arXiv cs.AI
→ story
Jul 22
Jul 22, 2026
SymptomAI: a conversational AI agent for everyday symptom assessment
Google Research
→ story
Jul 22, 2026
Dimitri Bertsekas, professor emeritus and influential computer scientist, dies at 83.
MIT News (AI)
→ story
Jul 22, 2026
Spectral evidence bundling for selective reliability estimation in time-series classification
arXiv cs.LG
→ story
Jul 22, 2026
Stochastic Meta-Unlearning: Bridging language backbone and multimodal unlearning
arXiv cs.CL
→ story
Jul 22, 2026
Structured Output collapses answer diversity across 44 language models
arXiv cs.CL
→ story
Jul 22, 2026
The information shadow: measuring structural limits on what language models can learn
arXiv cs.LG
→ story
Jul 22, 2026
Towards principled continual anomaly detection: a systematic framework and benchmark scenarios
arXiv cs.LG
→ story
Jul 22, 2026
Uncertainty quantification for AI-driven crash simulation surrogates: Monte Carlo dropout versus deep ensemble on an open-source bumper beam benchmark.
arXiv cs.LG
→ story
Jul 22, 2026
Using fine-tuned LLMs to identify indicators of vulnerability in UK police incident logs
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
→