LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Aug 3
20d ago
Adaptivity via a parallel architecture for stochastic gradient methods
arXiv cs.LG
→ story
20d ago
Benchmarks are not monolithic: sample-level auditing and orchestration for LLM evaluation show dataset-centric meta-evaluation along five latent dimensions
arXiv cs.CL
→ story
20d ago
Best Friends, Not Forever: Evaluating long-horizon persona collapse and behavioral drift in AI companions
arXiv cs.AI
→ story
20d ago
BLADE: Boundary-expanded and layer-adaptive dynamic exit for efficient LLM reasoning
arXiv cs.CL
→ story
20d ago
Can AI evaluate AI scientists? A benchmarking study of autonomous research generation systems using automated multi-model review
arXiv cs.AI
→ story
20d ago
Chain-of-Models: cross-model auditing for bias-robust LLM judges
arXiv cs.CL
→ story
20d ago
Context-preserving organization of inline notes and collected commentaries in classical Chinese texts for exegetical knowledge; NLP framework proposes preserving contextual dependencies while automated compilation
arXiv cs.CL
→ story
20d ago
Distilling knowledge from large language models into lightweight reinforcement learning agents for autonomous cyber operations
arXiv cs.LG
→ story
Jul 31
23d ago
GuideSkill: evolving executable LLM agent skills for guideline-grounded clinical reasoning
arXiv cs.AI
→ story
23d ago
Harness-G: a graph-structured harness for search agents
arXiv cs.CL
→ story
23d ago
ICLE++: modeling fine-grained traits for holistic essay scoring
arXiv cs.CL
→ story
23d ago
Kinetics of training describe a driven-nucleation rate law for emergence, plasticity loss, and circuit control in language models
arXiv cs.LG
→ story
23d ago
Latent channels in multi-agent LLMs are evaluated for actual communication through a causal audit of latent messages.
arXiv cs.AI
→ story
23d ago
LVLMs evaluated on perception and reasoning jointly to uncover truth behind visual illusions
arXiv cs.CL
→ story
23d ago
Modeling decisions in blockchain analytics: a leakage-aware evaluation of tree-based vs. sequential models
arXiv cs.LG
→ story
23d ago
MultivationBench: a benchmark for multimodal sequential motivation reasoning
arXiv cs.AI
→ story
23d ago
Objective misalignment in mixed-motive LLM multi-agent systems evaluated with Werewolf game
arXiv cs.AI
→ story
23d ago
Passive video to editable experience: Pegasus translates human demonstrations into robot-learnable data through structured knowledge transfer
arXiv cs.AI
→ story
23d ago
Prompt chaining in practice: a case study in automated scholarly report generation
arXiv cs.CL
→ story
23d ago
Property-driven causal abstractions for Markov decision processes
arXiv cs.AI
→ story
23d ago
Recall Before You Rank: Similarity-Guided Top-K Reuse for Efficient Long-Context Attention
arXiv cs.CL
→ story
23d ago
Recursive transformers for semiconductor thermo-mechanical reliability
arXiv cs.LG
→ story
23d ago
Representational quality accounts for RL models’ superior mathematical reasoning performance over SFT fine-tuned models
arXiv cs.AI
→ story
23d ago
Rethinking EEG-based disease diagnosis: decoupling instance representation learning from subject-level supervision
arXiv cs.LG
→ story
23d ago
Rethinking self-evolution: a constrained exploration-exploitation process for mitigating skill overfitting
arXiv cs.AI
→ story
23d ago
RLPF: reinforcement learning from performance feedback for code generation
arXiv cs.LG
→ story
23d ago
Same facts, different diagnosis: measuring and mitigating narrative anchoring in clinical language models
arXiv cs.CL
→ story
23d ago
SE(3)-MeanFlow: few-step protein backbone generation on Lie groups
arXiv cs.LG
→ story
23d ago
SkillSmith: learning to compose parametric skills and textual knowledge
arXiv cs.CL
→ story
23d ago
Structure-aware data organization for efficient LLM post-training
arXiv cs.LG
→ story
23d ago
Sympathetic framing: evaluating AI alignment across sociodemographic groups
arXiv cs.CL
→ story
23d ago
Systematic evaluation of 41 open-weight language models for zero-shot intent classification across eight datasets (135M–9B parameters) and 15 model families
arXiv cs.CL
→ story
23d ago
THGFM: Dual-Branch Temporal Heterogeneous Graph Fusion Model
arXiv cs.LG
→ story
23d ago
TIER-MoE: trust-informed expert routing via conditional modality risk for multimodal fusion in biomedical classification
arXiv cs.LG
→ story
23d ago
TraceCoder enables explainable and auditable code generation with position-key snippet versioning
arXiv cs.AI
→ story
23d ago
Training skills like parameters via self-supervised semantic diffusion
arXiv cs.CL
→ story
23d ago
UrbanDS: a graph-guided LLM multi-agent system for data-intensive urban tasks
arXiv cs.AI
→ story
23d ago
What Does It Take to Detect an AI Agent? Minimal feature sets for behavioral detection under browser automation
arXiv cs.AI
→ story
23d ago
When benchmark inferences do not compose: projectibility in AI evaluation
arXiv cs.AI
→ story
23d ago
ZUNA1.1: a 380M-parameter diffusion autoencoder for flexible EEG reconstruction across variable lengths, channels, and intervals
arXiv cs.LG
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→