LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Aug 15
8d ago
Agreement is not alignment: divergent moral grounds in human and LLM ethical judgments
arXiv cs.AI
→ story
8d ago
Alignment methods can be misused for censorship and manipulation, as a position paper argues that modern AI alignment is dual-use and may provide malicious actors with better tools.
arXiv cs.AI
→ story
8d ago
Asking an LLM in Japanese can change its willingness to advise a nuclear strike, study finds
arXiv cs.AI
→ story
8d ago
AstraZeneca develops Research Assistant, an internal LLM-based system to explore biomedical questions across literature, knowledge graphs, chemistry, and clinical data
arXiv cs.AI
→ story
8d ago
Attention is all you have; arXiv:2608.12610v1 reports that 56,804 public agent skills exist and installation bundles content, persistence, and trigger management, limiting long-tail usage.
arXiv cs.AI
→ story
8d ago
Auditable agentic AI coordinates diagnostic tools and stores outputs as auditable case-level evidence for thyroid ultrasound diagnosis and reporting
arXiv cs.AI
→ story
8d ago
CAS: a causal attribution score for local and global explainable AI
arXiv cs.AI
→ story
8d ago
Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues
arXiv cs.AI
→ story
8d ago
Designing AI pipelines for decision-ready ITSM intelligence
arXiv cs.AI
→ story
8d ago
DiG-bench: discovery in games
arXiv cs.AI
→ story
8d ago
Dual-flow transformers decouple the primary prefill path from additional decode computation
arXiv cs.AI
→ story
8d ago
Epsilon-MemEvo enables cross-task memory transfer in LLM program evolution by storing prior experience as task-agnostic tactic memories
arXiv cs.AI
→ story
8d ago
Governed Persistent Memory introduces a source-bound, auditable bitemporal state-transition model with fail-closed release for long-horizon agents
arXiv cs.AI
→ story
8d ago
IntegrityBench benchmarks LLMs on misconduct classification, ethical action reasoning, and artifact-grounded decision making across 36 paired tasks under pressure; evaluates 18 frontier model variants
arXiv cs.AI
→ story
8d ago
Jagged judges: Epistemic stability under silence, pressure, and persistence
arXiv cs.AI
→ story
8d ago
Large language models can follow instructions, but struggle with many constraints simultaneously; phase transitions in compositional constraint satisfaction
arXiv cs.AI
→ story
8d ago
Learning to adapt cross-domain preferences via meta-LoRA for LLM personalization
arXiv cs.AI
→ story
8d ago
MindMemOS: a portable and self-evolving memory operating layer for AI agents
arXiv cs.AI
→ story
8d ago
Multi-Agent Scheduling with LLM-assisted contract net negotiation for stream processing in mobile edge computing
arXiv cs.AI
→ story
8d ago
Probabilities of causation with causal knowledge tightened for binary PNS bounds by incorporating additional information
arXiv cs.AI
→ story
8d ago
Reasoning is a learnable rule-based process
arXiv cs.AI
→ story
8d ago
Reasoning Jury: multi-model consensus for evaluating reasoning traces
arXiv cs.AI
→ story
Aug 14
9d ago
MaSRead addresses the read to content in replicated latent stores; content-addressed reading improves access in a conflict-free replicated cache
arXiv cs.AI
→ story
9d ago
Multi-AUV ad-hoc network-based target tracking via a value gradient guidance multi-agent diffusion reinforcement learning approach
arXiv cs.LG
→ story
9d ago
Off-support barrier: semantic safety constraints are not learning-problem invariants and implications for design, containment, and verification
arXiv cs.AI
→ story
9d ago
Personalized scorer modeling: a learning-based framework for deriving robust sleep stage labels from multiple experts
arXiv cs.LG
→ story
9d ago
Poor Man's Agentic Modeling: simulating large LLM-agent societies on a laptop
arXiv cs.AI
→ story
9d ago
Predicting when random low-dimensional reparameterizations train neural networks
arXiv cs.LG
→ story
9d ago
Prob-K: probabilistic one-pass filtering for efficient top-k selection
arXiv cs.LG
→ story
9d ago
RecSys Factory: bounding LLM agent autonomy to decision points in the industrial recommender lifecycle
arXiv cs.AI
→ story
9d ago
Represent, then generate: multimodal-conditioned time-series generation under irregular missingness
arXiv cs.LG
→ story
9d ago
Scaling recurrent memory with content-routed state anchors
arXiv cs.LG
→ story
9d ago
Structure-preserving uncertainty quantification for GENERIC dynamics
arXiv cs.LG
→ story
9d ago
Synchronizing beliefs with second-order theory-of-mind in human-autonomy teams (extended version)
arXiv cs.AI
→ story
9d ago
The Boolean power of ReLU: ReLU-MPLang expresses more Boolean queries than Σ-MPLang with eventually constant activations on finite graphs, resolving an open question
arXiv cs.LG
→ story
9d ago
Towards query-agnostic RAG evaluation via query coverage and claim verifiability
arXiv cs.AI
→ story
9d ago
Training Under Challenge: Executable certificates and challenge-closed optimality for neural networks
arXiv cs.LG
→ story
9d ago
Unifying generative models with path integrals
arXiv cs.LG
→ story
9d ago
VQ-bench: a composable vector quantization framework
arXiv cs.AI
→ story
9d ago
When can you trust offline evaluation of equal-cost top-k allocation? a controlled, reproducible benchmark and practitioner's guide
arXiv cs.LG
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→