Fetching the paper…

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation · Around