LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 31
23d ago
Adam converges under heavy-tailed noise in the plain vector-form setting, with guarantees for stochastic gradients having bounded p-th moment (p in (1,2]).
arXiv cs.LG
→ story
23d ago
AHA-Memes: a fine-grained multimodal benchmark for understanding hate in Arabic memes
arXiv cs.CL
→ story
23d ago
AI agents discover statistical mechanical mappings from a raw partition function to a tractable representation in six Ising-type problems; StatMechBench-v0 benchmark introduced
arXiv cs.AI
→ story
23d ago
AI-assisted pre-review of open-source submissions used at BOSC 2026 to help volunteer reviewers by pre-reviewing abstracts for certain criteria
arXiv cs.CL
→ story
23d ago
AlphaSchema: exploring the space of trading semantics for LLM-based alpha mining
arXiv cs.AI
→ story
23d ago
AWARE-FX: an auditable AI/NLP decision-support system converts corporate annual report text into traceable foreign-exchange hedging disclosure measures
arXiv cs.CL
→ story
23d ago
B1ade: 335M embedding model and 1B parameter small language models for minimalist RAG
arXiv cs.CL
→ story
23d ago
Belief-guided decision making with uncertainty gating in the game of Go
arXiv cs.AI
→ story
23d ago
Benchmarking LLM competence on logical inference over probability operators
arXiv cs.CL
→ story
23d ago
Benchmarking the residual: long-horizon evaluations reveal degradation beyond short-task performance
arXiv cs.LG
→ story
23d ago
BridgeAlign: Bridging preference alignment for humanities and social sciences
arXiv cs.CL
→ story
23d ago
CaM-Wolf: Causal-Aware Multimodal Agents for Social Deduction Games
arXiv cs.AI
→ story
23d ago
CG-World: a large-scale world-state dataset and protocol for world models
arXiv cs.AI
→ story
23d ago
ClinLens: a benchmark of 200 executable tasks over five linked MIMIC resources for longitudinal multimodal clinical data science
arXiv cs.AI
→ story
23d ago
Compression-based behavioral similarity enables open-world Sybil discovery on Ethereum
arXiv cs.LG
→ story
23d ago
Context-informed ship trajectory prediction via conditional attention
arXiv cs.LG
→ story
23d ago
DoTime: a synthetic benchmark generator for interventional and counterfactual time series
arXiv cs.LG
→ story
23d ago
DualAnchor: preserving language priors and improving lexical fidelity in gloss-free sign language translation
arXiv cs.CL
→ story
23d ago
ECG-InterpBench benchmarks the interpretability of ECG foundation-model representations using matched-scale sparse autoencoders
arXiv cs.LG
→ story
23d ago
Eco3S: a socio-economic system simulation framework for agent-based modeling and policy analysis
arXiv cs.AI
→ story
23d ago
Evaluation scores are perishable knowledge claims
arXiv cs.AI
→ story
23d ago
Evidence-ledger adjudication for claim-evidence traceability in AI agents assesses support relations and routes unsupported or contradicted claims back to the author
arXiv cs.AI
→ story
23d ago
EvoPINN: agentic discovery of executable algorithms for physics-informed neural networks
arXiv cs.AI
→ story
23d ago
Explorative modeling: unlocking a third pretraining axis and end-to-end generation
arXiv cs.LG
→ story
23d ago
Fewer clarifications, better code: benchmarking cross-session personalized ambiguity adaptation in coding assistants
arXiv cs.AI
→ story
23d ago
Flat Score, Amplified Failures: Quantization to 4-bit weights hides damage in multi-turn tool-calling LLM agents
arXiv cs.LG
→ story
23d ago
From single- to cross-document: benchmarking multi-granularity event analysis of large language models
arXiv cs.CL
→ story
23d ago
FunL2O: LLM-guided feature function design for learning to optimize
arXiv cs.LG
→ story
23d ago
GoGoTB: agentic RTL verification with specification-grounded coverage closure
arXiv cs.AI
→ story
23d ago
Good rankers, bad objectives: bilinear contrastive critics under expressive policy search
arXiv cs.LG
→ story
23d ago
Gradient-free task-conditioned retrieval for on-device in-context learning
arXiv cs.CL
→ story
23d ago
Grounded agentic extraction and expert-adjudicated evaluation of intertextuality in classical Chinese histories
arXiv cs.CL
→ story
Jul 30
23d ago
Daniela Rus receives the Bavarian Minister-President's High-Tech Prize.
MIT News (AI)
→ story
23d ago
Capitol Hill hosts Congressional Visit Days as researchers connect with policy makers; participants pose on the Capitol steps with Senator Alex Padilla.
MIT News (AI)
→ story
23d ago
Echoverse: deep, evolving environments for computer-use agents
Microsoft Research
→ story
24d ago
Voice Memory enables agentic speech recognition with a frozen corrector deciding per utterance to act on the hypothesis or abstain, and an asynchronous optimizer revising a per-domain memory.md through bounded edits
arXiv cs.CL
→ story
24d ago
Weak-to-Strong on-policy distillation improves alignment when no larger teacher exists
arXiv cs.LG
→ story
24d ago
When synthetic users fail: a cross-domain benchmark of LLM-simulated human survey responses
arXiv cs.CL
→ story
24d ago
Where detectors fail: expert-guided mutual distillation improves robustness to tail-domain gaps
arXiv cs.CL
→ story
24d ago
WikiLoop jointly learns to build and navigate an agent-native wiki with downstream feedback
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→