Heidy Khlaaf, chief AI scientist at the AI Now Institute, a U.S.-based research institute, said Anthropic’s detailed blog post explaining the new vulnerabilities left out many key details needed to verify its claims.

Writing on X, Khlaaf warned against “taking these claims at face value” without more information, such as the rates of false positives and clearer explanations for how the humans conducted manual reviews of the identified vulnerabilities.

Read the article here.

Research Areas