How fragile is on-device language model safety? Localizing safety-critical parameters for sparse fault analysis
Read the original at arxiv.org→arXiv:2610.09000v1 Announce Type: new Abstract: As small language models (SLMs) are increasingly deployed on resource-constrained and on-device platforms, including as components of agentic systems, the integrity of...
Original headline: "How Fragile Is On-Device Language Model Safety? Localizing Safety-Critical Parameters for Sparse Fault Analysis"
Coverage timeline
- Oct 8, 04:00 UTC arXiv cs.AI lead source How Fragile Is On-Device Language Model Safety? Localizing Safety-Critical Parameters for Sparse Fault Analysis
- Oct 8, 04:00 UTC arXiv cs.CL Quad-State Safety Evaluation of Open-Weight Large Language Models on Non-Canonical Inputs