The AI Security Paradox: Guardrails Block the Defenders Who Need Them Most
Andrew Ng's team wanted a security audit. Claude and GPT refused. Open models completed it. The most capable models for finding vulnerabilities are the ones whose guardrails prevent defensive use.