LIVE · refreshes every 20 min
updated Sep 2, 17:22 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 16
Jul 16, 2026
Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations
arXiv cs.CL
→ story
Jul 16, 2026
Discourse-aware policy analysis with argumentation: a hybrid LLM-symbolic framework for disaster governance
arXiv cs.CL
→ story
Jul 16, 2026
Disentangling knowledge states with ability and proficiency modeling for knowledge tracing
arXiv cs.LG
→ story
Jul 16, 2026
Do LLMs need architectural changes for simultaneous speech translation? A prefix-to-prefix data driven approach
arXiv cs.CL
→ story
Jul 16, 2026
EMAGN uses learned clustering for scalable traffic forecasting.
arXiv cs.LG
→ story
Jul 16, 2026
Evaluation ability does not imply optimization utility; LLM-as-a-judge signals in closed-loop table recognition show weak judge signals on FinTabNet and OmniDocBench
arXiv cs.CL
→ story
Jul 16, 2026
Explaining reinforcement learning agents via inductive logic programming
arXiv cs.AI
→ story
Jul 16, 2026
Finding the right tables and columns: a benchmark and corpus-adaptive embeddings for SQL schema retrieval
arXiv cs.CL
→ story
Jul 16, 2026
FixItFlow automates troubleshooting guide generation from historical cloud incidents using large language models
arXiv cs.CL
→ story
Jul 16, 2026
GFlowRL: scaling distribution-matching RL to large language models
arXiv cs.CL
→ story
Jul 16, 2026
GIFT teaches vision-language generative AI models to produce CAD programs for simulating and testing 3D objects.
MIT News (AI)
→ story
Jul 16, 2026
Graded entity-familiarity readouts in language models: Polish adaptation, cross-language robustness, and refusal steering
arXiv cs.CL
→ story
Jul 16, 2026
Gsm-Plus-Bn Introduces Perturbation-Based Benchmark for Bangla Mathematical Reasoning in LLMs; Addresses Lack of Evaluation in Bengali
arXiv cs.CL
→ story
Jul 16, 2026
Harness Handbook: making evolving agent harnesses readable, navigable, and editable
arXiv cs.AI
→ story
Jul 16, 2026
Hedgehog: hierarchical evaluation of drug generators through rigorous filtration
arXiv cs.LG
→ story
Jul 16, 2026
Improving molecular property prediction in small language models using graph-based tools.
arXiv cs.AI
→ story
Jul 16, 2026
LAPO uses leave-one-turn attribution for multi-turn search reasoning as self-generated process supervision.
arXiv cs.AI
→ story
Jul 16, 2026
LLM-powered agentic system enables automatic discovery of ordinary differential equations in biological systems
arXiv cs.AI
→ story
Jul 16, 2026
Masking, fingerprinting, and privacy from discarded geometry: what your model threw away and why you'll want it back
arXiv cs.LG
→ story
Jul 16, 2026
Memory as a controlled process: learned adaptive memory management for LLM agents
arXiv cs.CL
→ story
Jul 16, 2026
Multi-Agent Collaborative Reasoning with Tool-Augmented Evidence for Urban Region Profiling
arXiv cs.AI
→ story
Jul 16, 2026
Muon emerges as strong optimizer for deep learning, outperforming Adam and AdamW; theoretical work interprets it as steepest descent under spectral norm.
arXiv cs.LG
→ story
Jul 16, 2026
MyAG is a graph-based framework for designing and analyzing composable LLM agent systems with three graph abstractions.
arXiv cs.CL
→ story
Jul 16, 2026
Networked intelligence: active shared context graphs for human-AI team science
arXiv cs.AI
→ story
Jul 16, 2026
New paper explores federated learning and explainable AI integration for privacy-preserving transparent machine learning
arXiv cs.LG
→ story
Jul 16, 2026
OpenRCA dataset highlights challenges in achieving high accuracy for root cause analysis on real-world telemetry data.
arXiv cs.AI
→ story
Jul 16, 2026
Oracle Agent Memory as an enterprise memory substrate for long-horizon AI agents
arXiv cs.AI
→ story
Jul 16, 2026
OriginBlame introduces record- and token-level data provenance for AI training datasets
arXiv cs.AI
→ story
Jul 16, 2026
Probabilistic extension of neuro-symbolic AGI robots based on Belnap's typed intensional FOL
arXiv cs.AI
→ story
Jul 16, 2026
RAGthoven presents multi-stage LLM pipeline for SemEval-2026 Task 1's multilingual humor generation in English, Spanish, and Chinese.
arXiv cs.CL
→ story
Jul 16, 2026
Researchers introduce a benchmark and reference system for live captioning Sikh Kirtan; arXiv:2607.13457v1 presents a closed-vocabulary task requiring exact SGGS line transcription.
arXiv cs.CL
→ story
Jul 16, 2026
Researchers introduce black-box interventional grounding audits to test LLM premise dependency via predicate substitution; method replaces target predicates with fresh symbols and re-runs models to assess reasoning changes.
arXiv cs.AI
→ story
Jul 16, 2026
Researchers introduce CoDiffGRN, a method for gene regulatory network inference using BEELINE-KGC benchmark and co-evolutionary discrete diffusion.
arXiv cs.LG
→ story
Jul 16, 2026
Researchers introduce DROPJ, a human-centered method for safe agent training and deployment in safety-critical environments with unknown dynamics and no reward function.
arXiv cs.AI
→ story
Jul 16, 2026
Researchers present EZSMT Version 3, an extensible SMT-based CASP framework.
arXiv cs.AI
→ story
Jul 16, 2026
Researchers propose a meta-learning framework to address data scarcity in low-resource language alignment for multilingual LLMs; arXiv:2607.13315v1
arXiv cs.CL
→ story
Jul 16, 2026
Researchers propose lightweight training strategy for efficient transfer learning by decoupling feature extraction and classifier optimization with margin-based weighted loss.
arXiv cs.LG
→ story
Jul 16, 2026
Researchers propose Samba, a hybrid Mamba model for audio-visual navigation, addressing inadequacies of existing frameworks since 2020.
arXiv cs.LG
→ story
Jul 16, 2026
Researchers use persona vectors to audit open-weight LLMs, revealing 53 traits across four models in arXiv:2607.13162v1.
arXiv cs.CL
→ story
Jul 16, 2026
Safe-Psych, a sequential evaluation benchmark for LLMs in psychiatry, addresses incomplete information by requiring clarification.
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
→