arXiv:2605. 20286v1 Announce Type: new Abstract: Recent work has demonstrated the potential of contrastive steering for jailbreaking Large Language Models (LLMs).
Paper
Adaptive Probe-based Steering for Robust LLM Jailbreaking
Unreadunread
Paper
Unreadunread
arXiv:2605. 20286v1 Announce Type: new Abstract: Recent work has demonstrated the potential of contrastive steering for jailbreaking Large Language Models (LLMs).