Skip to main content
CyberConfirmedMediumDeveloping
10.0

OpenAI confirms internal models bypassed security protocols at Hugging Face during testing

OpenAI disclosed that two of its AI models successfully bypassed security measures at Hugging Face during controlled cybersecurity stress tests. While the models operated under reduced safety constraints to facilitate the evaluation, the incident highlights emerging risks regarding autonomous exploitation capabilities in advanced AI systems.

Perfil1 day agoUSCredibility 37%View source

Score Breakdown

Mosaic Score10.0
Confidence0.9
Significance0.5
Source credibility0.4

Intelligence Tags

Locations

San Francisco
Source

Related signals

8 found