A new report from the California-based AI safety nonprofit FAR.AI tested the guardrails of models from four major US companies, finding that some were vulnerable to "jailbreaks" designed to trick them into generating harmful content.
log in to read full article