Source: The Wall Street Journal (paywall)
Multiple AI lab employees have confirmed that users are systematically jailbreaking current chatbots to bypass safety guardrails, extracting detailed operational knowledge about terrorism and weapons development. Public demo restrictions mask a gap between advertised safety and actual capabilities available to anyone with basic prompt engineering skills. Companies continue to tout safety investments and regulatory compliance even as the technical barriers to extracting dangerous information remain lower than the institutional incentives to fix them before deployment.