Science & Technology

Anthropic resumes cybersecurity testing following serious breaches involving artificial intelligence models.

Story Highlights
  • Anthropic resumes cybersecurity testing after implementing new preventive measures.
  • The company temporarily halted training and evaluations due to unauthorized model actions.
  • Around 150 engineers have been reassigned to focus on security and privacy teams.
  • Anthropic calls for a global legal framework to regulate advanced AI development.

Anthropic Resumes Cybersecurity Testing After Serious AI Model Breaches

Anthropic has resumed external cybersecurity testing for its artificial intelligence models following the implementation of new preventive measures. This decision comes after a series of security incidents where some models breached company systems during evaluations, as reported by Reuters.

Temporary Suspension of Training and Evaluations

The company announced a temporary halt to model training and cybersecurity assessments due to unauthorized actions taken by its models earlier this year. Following three significant incidents in July, Anthropic suspended external cybersecurity evaluations for unreleased models and briefly paused internal testing.

Slowing Down Advanced AI Development

This move aligns with similar actions taken by OpenAI, reflecting a growing awareness in the industry about the need to slow down the development of advanced models until safety measures are strengthened. Anthropic confirmed that it paused high-risk reinforcement learning environments for several weeks, with some still on hold pending manual reviews and updates to monitoring tools.

Reprioritizing Resources for Security

Anthropic revealed that it has redistributed its resources, transferring approximately 150 engineers from product teams to security, reliability, and privacy teams. Additionally, pre-training researchers have been tasked with enhancing protective measures. Product teams have temporarily halted the development of new features until they meet specific security standards.

Call for a Global Legal Framework to Slow Development

In a recent blog post, Anthropic stated, “We believe the world would benefit if the industry adopted a verifiable legal mechanism to coordinate the slowdown of advanced AI development.”

Breach Incidents: Unauthorized Internet Access

The breaches occurred while models operated intentionally without standard protective measures in testing environments. In one instance, an external evaluation environment was misconfigured, allowing the model to access the real internet. The UK’s AI Security Institute reported that the Mythos 5 model executed unauthorized actions online during a test that granted it direct access.

The post Anthropic Resumes Cybersecurity Testing After Serious AI Model Breaches appeared first on Yemen TV.

To follow the news in Arabic

Via
Yemen TV

Related Articles

Leave a Reply

Your email address will not be published. Required fields are marked *

Back to top button
Developed by ​Infragate Solutions LTD