Anthropic, a San Francisco-based AI startup founded by researchers who broke away from OpenAI, has published an overview of its red-teaming practices, outlining four approaches and their advantages and disadvantages. Red teaming, a security practice of attacking one’s own system to uncover and address potential security vulnerabilities, has taken on a prominent role in discussions of AI regulation. The Biden administration’s AI executive order mandates that companies developing high-risk foundation models notify the government during training and share all red teaming results, while the EU AI Act also contains requirements around providing information from red teaming. Anthropic’s approaches include using language models to red team, red teaming in multiple modalities, domain-specific expert red teaming, and open-ended, general red teaming. The company concludes with policy recommendations, including suggestions to fund and encourage third-party red teaming, and to create clear policies tying the scaling of development and release of new models with red teaming results. As lawmakers rally around red teaming as a way to ensure powerful AI models are developed safely, it certainly deserves a close eye.

Source.

TOP STORIES

Big Tech's Trust Crisis Deepens with Anthropic Lawsuit
Sony Music and Warner Music have sued Anthropic, accusing it of copyright infringement in AI training …
Nvidia's AI Future - Jensen Huang's Vision for Record Growth
Huang believes Nvidia’s position in AI will lead to another year of record growth …
China's AI Companies Target US Models with Distillation Attacks
Anthropic’s report reveals a surge in distillation attacks by Chinese AI firms on U.S. models …
Cybersecurity Concerns Rise as AI Agents Break Boundaries
AI agents’ autonomy poses significant risks, as demonstrated by a recent breach …
IDScan Confirms Major Data Breach Affecting Driver's Licenses
IDScan has confirmed a data breach that exposed driver’s licenses of over 150 million individuals …
Matt Mullenweg's Abrupt Leave Sparks Controversy at Automattic
Matt Mullenweg has been placed on leave by Automattic’s board, stirring controversy …

latest stories