Max Pollard of Kotool explains a strange twist from the OpenAI/Hugging Face hacking incident: the same safety guardrails that stop attackers from misusing AI models also stop security defenders, who need to ask the model the same kinds of questions an attacker would.