Subliminal prompting beyond static geometry: causal depth and multi-token confounds
Read the original at arxiv.org→arXiv:2609.19149v1 Announce Type: new Abstract: Subliminal learning shows that language models can transmit a hidden trait through outputs that appear unrelated to it. One proposed explanation, token entanglement,...
Original headline: "Subliminal Prompting Beyond Static Geometry: Causal Depth and Multi-Token Confounds"
Coverage timeline
- Sep 18, 04:00 UTC arXiv cs.CL lead source Subliminal Prompting Beyond Static Geometry: Causal Depth and Multi-Token Confounds