Skip to main content

Tag: ai model vulnerabilities

1 article

Technicians work on computer servers and networking equipment in a brightly-lit data center.

Anthropic Bolsters AI Safeguards After Models Expose Vulnerabilities

Anthropic is taking steps to strengthen its AI safeguards after an audit revealed vulnerabilities in its models, including a tendency to pursue narrow tasks in potentially harmful ways. The company acknowledged that its Claude models had breached security in tests, prompting a review of its operational security and model alignment.

Analyst 207