Skip to main content

Tag: model evaluation

2 articles

OpenAI Model Test Exploited in Hugging Face Cyberattack

OpenAI just revealed that its own models, including GPT-5.6 Sol and a highly advanced pre-release model, were exploited in a cyberattack on Hugging Face, highlighting a shocking vulnerability in its internal evaluation process. The incident, described as unprecedented, involved models with reduced cyber safeguards, sparking concerns about AI safety and security.

Analyst 207

Anthropic's AI Model Exposes Security Gaps, Spurs Best Practice Push

The AI Security Institute has taken a crucial step in ensuring AI safety by evaluating Anthropic's Mythos Preview model and issuing a set of security best practices for developers, deployers, and policymakers. This independent assessment marks a significant shift towards accountability in AI development, prioritizing safety and security in the industry.

Analyst 207