Semalith v1.4, a 184M DeBERTa-v3-base classifier, achieves state-of-the-art prompt-injection detection with 44x fewer parameters than Llama-Guard-3-8B
Read the original at arxiv.org→arXiv:2607.22545v1 Announce Type: new Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, regulatory...
Original headline: "Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B"
Coverage timeline
- Jul 28, 04:00 UTC arXiv cs.LG lead source Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B