LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Aug 4
19d ago
A Constitution-Grid instrument for data-efficient RL alignment (C-Guard)
arXiv cs.CL
→ story
19d ago
A Few Neurons Reveal When LLMs Misuse Tools: Sparse Detection and Selective Steering for Reliable Tool Use
arXiv cs.CL
→ story
19d ago
Abstention as an action can kill both the reward gradient and the KL anchor: collapse law and repair for error-penalized reinforcement learning
arXiv cs.LG
→ story
19d ago
Agentic Bayesian Optimization through surrogate-augmented autoresearch
arXiv cs.LG
→ story
19d ago
Agentic Graph Token Reasoning
arXiv cs.LG
→ story
19d ago
AutoCause is an open-source Python workflow that automates expert decisions in environmental time-series causal discovery.
arXiv cs.LG
→ story
19d ago
Averaging bias: human faithfulness annotations are not locally faithful
arXiv cs.CL
→ story
19d ago
Boosting diversity in text diffusion models via entropy-based guidance
arXiv cs.CL
→ story
19d ago
Comparing and modeling argumentation in German political communication across arenas
arXiv cs.CL
→ story
19d ago
Cost-effective automated judging of natural-language mathematical proofs using open-weight models aligns with human grading decisions on IMO-GradingBench instances
arXiv cs.CL
→ story
19d ago
CoSynFlow: conformal symplectic neural flows for cross-system prediction of dissipative Hamiltonian dynamics
arXiv cs.LG
→ story
19d ago
Cross-task dissociation in frontier vision-language model theory of mind
arXiv cs.CL
→ story
19d ago
CurveShift: Is Agent Progress Scalar? Separating Level from Shape
arXiv cs.CL
→ story
Aug 3
19d ago
Alexander Rakhlin named director of the MIT Statistics and Data Science Center.
MIT News (AI)
→ story
19d ago
Orchard: an open framework for scalable agentic AI
Microsoft Research
→ story
20d ago
Preference-Optimized LLM counselors trade goal persistence for relational attunement in motivational interviewing
arXiv cs.CL
→ story
20d ago
Reasoning in real world clinical care: why large language models are not yet safe for autonomous clinical decision support; corroborating coverage notes real-world credibility tests of LLM reasoning
arXiv cs.AI
→ story
20d ago
Reflection or re-generation? Why LLM revision fails where human revision succeeds
arXiv cs.LG
→ story
20d ago
Representations from pretrained machine-learning interatomic potentials as coarse coordinates for material generation and evaluation
arXiv cs.LG
→ story
20d ago
Safety, or just capability? A validity audit of agent-safety benchmarks
arXiv cs.AI
→ story
20d ago
Scaling scientific discovery environments for turn-level agentic RL
arXiv cs.AI
→ story
20d ago
SciToolAgent-Evo: an ontology-aware self-evolving agent for open-world scientific tool acquisition
arXiv cs.AI
→ story
20d ago
SEDR-Seq2P: a lightweight dilated residual sequence-to-point network for multi-task industrial NILM
arXiv cs.LG
→ story
20d ago
Self-Supervised Skill Optimization learns reusable agent skills from unlabeled task instances alone
arXiv cs.CL
→ story
20d ago
Sensitivity analysis of GRU, LSTM and Transformer encoder in classification of automated driving systems
arXiv cs.LG
→ story
20d ago
Stateful knowledge learning enables LLM agents to move from trajectory-level reflection to predictive foresight
arXiv cs.CL
→ story
20d ago
TAGTorch: a PyTorch library for geometry, topology, and symmetry-aware machine learning
arXiv cs.LG
→ story
20d ago
TAPR: a task-aware prompt rewriter improves downstream LLM performance by reformulating user prompts into task-optimized prompts using reinforcement learning with GRPO
arXiv cs.AI
→ story
20d ago
Technological advances in detecting and managing cognitive impairment in older adults: trends, challenges, and future directions
arXiv cs.LG
→ story
20d ago
TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking
arXiv cs.CL
→ story
20d ago
TextCloak defends against unauthorized LLM exploitation by RL-driven unlearnable text
arXiv cs.CL
→ story
20d ago
The Formalism Trap: LLM-as-a-Judge evaluators conflating proceduralism with semantic truth under social load
arXiv cs.CL
→ story
20d ago
The Morphological core of Dungan: a two-dialect finite-state model and a multi-genre evaluation
arXiv cs.CL
→ story
20d ago
ThinkReset: Learnable intermediate interface construction for bounded-context long-horizon reasoning
arXiv cs.AI
→ story
20d ago
Token-Level diagnosis of sycophancy in LLMs with attribution-guided steering
arXiv cs.CL
→ story
20d ago
TokenSwap benchmarks and reduces the modality gap in multimodal LLMs
arXiv cs.CL
→ story
20d ago
Topology-aware data movement for disaggregated GPU inference
arXiv cs.LG
→ story
20d ago
ViSAGE: constructing self-correcting memories for long-form video understanding
arXiv cs.AI
→ story
20d ago
What must be true before AI ships in a regulated firm
arXiv cs.CL
→ story
20d ago
ZeroR@CHiPSAL 2026: two-stage vision-language adaptation with contrastive learning for Nepali meme classification
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→