LIVE · refreshes every 20 min
updated Aug 23, 12:01 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 23
Jul 23, 2026
CruiseBench: a real-flight-aligned N-CMAPSS benchmark for engine RUL prediction
arXiv cs.LG
→ story
Jul 23, 2026
Decodable but Not Detectable: A Leakage Fingerprint for Near-OOD Benchmarks reveals a benchmark leak where removing the OOD class and retraining 35 models raises AUROC from 0.326 to 0.911.
arXiv cs.LG
→ story
Jul 23, 2026
Embedding-based measurement of data diversity for NLP; emb-diversity tool assesses diversity using embeddings
arXiv cs.CL
→ story
Jul 23, 2026
Euclean: automated geometry problem formalization with unified verification in Lean
arXiv cs.AI
→ story
Jul 23, 2026
Explainability challenges in continual learning for time series forecasting with Experience Replay strategies
arXiv cs.LG
→ story
Jul 23, 2026
Fence: Specialized SLM guardrails for LLM applications
arXiv cs.AI
→ story
Jul 23, 2026
FindStatBench: evaluating large language models on combinatorial code synthesis
arXiv cs.AI
→ story
Jul 23, 2026
FineServe: a fine-grained dataset and characterization of global LLM serving workloads
arXiv cs.AI
→ story
Jul 23, 2026
FinMMEval 2026 Task 1 evaluates multilingual financial multiple-choice question answering in English, Chinese, Arabic, and Hindi.
arXiv cs.CL
→ story
Jul 23, 2026
FORCE-Bench: a benchmark, dataset, and evaluation harness for agentic AI in enterprise finance
arXiv cs.AI
→ story
Jul 23, 2026
FormulaSPIN: Self-play fine-tuning for natural language to spreadsheet formula generation
arXiv cs.AI
→ story
Jul 23, 2026
From Agent Failure Paths to Quantified Residual Risk: A Compositional Framework for Resilient Agentic AI
arXiv cs.AI
→ story
Jul 23, 2026
From trajectories to prefixes: reusing teacher trajectories via replayed prefixes and online continuation
arXiv cs.LG
→ story
Jul 23, 2026
Geometry-Guided Constraint Learning for LLM safety classification achieves near-perfect per-category accuracy with sparse autoencoder feature extraction, reducing constraint counts from K=4-25 to K=2 for most categories
arXiv cs.AI
→ story
Jul 23, 2026
GraphContainer provides a unified platform for comparing and debugging Graph RAG methods
arXiv cs.AI
→ story
Jul 23, 2026
Hybrid LSTM-Graph neural framework for robust financial fraud detection and adversarial resilience
arXiv cs.AI
→ story
Jul 23, 2026
HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions
arXiv cs.AI
→ story
Jul 23, 2026
Integro-differential equations in angular stabilization of drone motion by distributed feedback control
arXiv cs.AI
→ story
Jul 23, 2026
ITPEval benchmarks automated translation of formal proofs across Lean 4, Rocq, Isabelle, and HOL Light
arXiv cs.AI
→ story
Jul 23, 2026
LAARA: Layer-Aware Adaptive Rank Allocation for parameter-efficient fine-tuning
arXiv cs.LG
→ story
Jul 23, 2026
Language-specific versus cross-lingual knowledge graphs for implicit aspect identification in Arabic: a comparative study of reasoning and adaptation strategies
arXiv cs.CL
→ story
Jul 23, 2026
Latency-aware LLM query routing for dynamic workloads improves by considering generation latency alongside accuracy and cost
arXiv cs.AI
→ story
Jul 23, 2026
Learning the Arabic dialect continuum as a continuous space: a regression approach to speaker origin prediction
arXiv cs.CL
→ story
Jul 23, 2026
Lightweight system for person–place relation extraction in historical newspapers using dependency graphs and proximity features
arXiv cs.CL
→ story
Jul 23, 2026
LISA: Linear-Indexed Sparse Attention for efficient long-context reasoning
arXiv cs.AI
→ story
Jul 23, 2026
LLMs update memory through shared latent structures rather than isolated facts, supporting the lifted representation hypothesis
arXiv cs.AI
→ story
Jul 23, 2026
Memory Merge DQN: sensitivity weighted target updates for stable value learning
arXiv cs.LG
→ story
Jul 23, 2026
MILP-Evo enables closed-loop, fully automatic design of MILP solvers
arXiv cs.AI
→ story
Jul 23, 2026
Mitigating scaffolding collapse in Socratic tutors via representation alignment
arXiv cs.AI
→ story
Jul 23, 2026
Multi-dimensional evaluation of explainability in media bias detection using the BABE dataset
arXiv cs.CL
→ story
Jul 23, 2026
MUX: continuous reasoning via multiplexed tokens
arXiv cs.AI
→ story
Jul 23, 2026
Native multi-dimensional subquadratic operators via input dependent long convolutions
arXiv cs.LG
→ story
Jul 23, 2026
Neural operator surrogates for two-dimensional neutron flux estimation.
arXiv cs.LG
→ story
Jul 23, 2026
NEXUS: Structured runtime safety for tool-using LLM agents
arXiv cs.AI
→ story
Jul 23, 2026
NMR elucidation treated as an agentic search problem; an autonomous agent using a frozen LLM matches graduate-level chemistry students in structure determination
arXiv cs.LG
→ story
Jul 23, 2026
On the computational complexity of structural generalization
arXiv cs.CL
→ story
Jul 23, 2026
OpenEvoShield: Dual non-stationary continual defense for open-world multi-agent system attacks
arXiv cs.AI
→ story
Jul 23, 2026
Orthogonalized read improves noisy associative recall in mLSTM memory by reconditioning the learning problem during training plateau (arXiv:2607.19390)
arXiv cs.LG
→ story
Jul 23, 2026
PEARL: solver-in-the-loop interactive optimization modeling from natural language
arXiv cs.AI
→ story
Jul 23, 2026
Phionyx introduces a deterministic AI runtime architecture with structured state management and governance-first state evolution; LLM outputs are treated as noisy sensor measurements rather than decisions.
arXiv cs.AI
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→