Probabilistic concept-aware steering for trustworthy LLM inference
Read the original at arxiv.org→arXiv:2607.18259v1 Announce Type: new Abstract: Steering vectors (SVs), an inference-time intervention technique for large language models (LLMs), guide the generation process by adding a concept-specific direction...
Original headline: "Probabilistic Concept-Aware Steering for Trustworthy LLM Inference"