LIVE · refreshes every 20 min
updated Aug 23, 12:41 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
110010
.art
/ archive
Research
Jul 28
26d ago
Activation Oracles learn not to read: concept-specific blind spots in fine-tuned oracles
arXiv cs.CL
→ story
26d ago
ADAGE: a language-agnostic pipeline for analogical reasoning evaluation
arXiv cs.CL
→ story
26d ago
AgentKVShift: Efficient KV cache reuse for agentic memory systems
arXiv cs.AI
→ story
26d ago
Aligning educational LLMs as Socratic guides via heuristic reinforcement learning; study presents HeuristicEdu for Qwen2.5-7B using SocraticEdu data and GRPO training
arXiv cs.CL
→ story
26d ago
Attention-guided layer selection for contrastive decoding in large language models
arXiv cs.CL
→ story
26d ago
Automated detection of documentation inconsistencies in electronic health records using a two-stage LLM pipeline
arXiv cs.CL
→ story
26d ago
AutoThinkSQL enables selective reasoning for text-to-SQL by combining SFT and DPO guidance
arXiv cs.CL
→ story
26d ago
BERT-based models vs. large language models for low-resource named entity recognition: a comparative study on Marathi
arXiv cs.CL
→ story
26d ago
Beyond a global norm: personalizing toxicity sensitivity in language models without retraining
arXiv cs.CL
→ story
26d ago
Beyond Shapley: an influence-based data auditing pipeline for LLM alignment and evaluation
arXiv cs.LG
→ story
26d ago
Bharati: morphology-aware tokenizers for classical Indian languages with subword fertility analysis
arXiv cs.CL
→ story
26d ago
CausalGate: causal importance distillation for transformer module pruning
arXiv cs.LG
→ story
26d ago
CC-AOS: Cost- and horizon-conditioned amortized backward induction for finite-horizon optimal stopping
arXiv cs.LG
→ story
26d ago
CHiPS: character histograms and positional signals for lightweight authorship attribution in Romanian texts
arXiv cs.CL
→ story
26d ago
cMoLLM at scale: horizontal scaling laws for mixture-of-LLMs
arXiv cs.AI
→ story
26d ago
Co-evolving graph and text memory for training-free multi-hop question answering
arXiv cs.CL
→ story
26d ago
Codifying the Judge: scalable evaluation via program distillation
arXiv cs.AI
→ story
26d ago
Concept-based visual counterfactual explanations with diffusion models
arXiv cs.AI
→ story
26d ago
Context anxiety causes frontier reasoning models to doubt their solutions despite capability; a study analyzes how misestimation of token difficulty affects performance.
arXiv cs.AI
→ story
26d ago
Corvus: context optimization and reduction via underlying synchronization for LLM coding agents
arXiv cs.LG
→ story
26d ago
Coupled hierarchical search over topology and execution for agentic workflow synthesis
arXiv cs.AI
→ story
26d ago
DeepLens Diagnosis Agent uses a five-stage reasoning pipeline to enable a small medical reasoning model to compete with frontier LLMs.
arXiv cs.AI
→ story
26d ago
Defining AI-native systems: autonomy as revision authority
arXiv cs.AI
→ story
26d ago
Dementia etiology diagnosis via collaborative meta knowledge enhancement
arXiv cs.LG
→ story
26d ago
Discrete action space as a prerequisite for GRPO convergence in small-model continuous control
arXiv cs.AI
→ story
26d ago
Do modules stay in their lane? Role drift in compound LLM systems
arXiv cs.AI
→ story
26d ago
Do VLMs read or rewrite? Evidence of transcription rewriting in vision-language models using FaithC4 benchmark
arXiv cs.AI
→ story
26d ago
DomainPilot: Domain-Level loss-guided two-stage data mixture optimization for efficient language model fine-tuning
arXiv cs.LG
→ story
26d ago
DSTFView: multi-view workload forecasting for cloud-edge platforms using dual-input spatio-temporal-frequency modeling
arXiv cs.AI
→ story
26d ago
Evaluating LLM reliability beyond accuracy: how model answers vary with meaning-preserving paraphrases across tasks
arXiv cs.AI
→ story
26d ago
Evaluating narrative unlearning with LENS: a level-based evaluation protocol for suppressing disinformation-aligned narrative reproduction in LLMs
arXiv cs.CL
→ story
26d ago
Execution-Grounded security testing for coding agents in software engineering pipelines
arXiv cs.AI
→ story
26d ago
FBLayout: Optimizing memory layout for efficient LLM finetuning on mobile GPUs
arXiv cs.AI
→ story
26d ago
FlowEvo: self-evolving agents through the co-evolution of workflows and executable skills
arXiv cs.AI
→ story
26d ago
FMOPF: Latent Flow Matching with Constraint-Aware Interaction Priors for AC Optimal Power Flow
arXiv cs.LG
→ story
26d ago
FrED: External data influence estimation via domain knowledge graph grounding
arXiv cs.AI
→ story
26d ago
From hybrid mechanistic–data-driven modeling toward neuro-symbolic AI: what, why, and how
arXiv cs.LG
→ story
26d ago
From profiles to steering vectors: Global Sparse Priors and Local Semantic Calibration for personalized text generation
arXiv cs.AI
→ story
26d ago
GAND: a resource on gender-ambiguous natural data and contrastive attribution for evaluating gender bias in machine translation
arXiv cs.CL
→ story
26d ago
HDL emerges as a hard decision layer in transformers, causing abrupt stabilization of answer option rankings during inference
arXiv cs.AI
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
→