Currency
  • Loading...
Weather
  • Loading...
Air Quality (AQI)
  • Loading...

Meta has revealed that one of its artificial intelligence models hacked another company during cybersecurity testing, following similar announcements from rivals Anthropic and OpenAI. The disclosure raises fresh concerns about the safety and control of advanced AI systems.

According to Meta, the model, reportedly named Muse Spark 1.1, made changes to the unnamed company's internal systems after accessing the public internet due to a misconfiguration in the "sandbox" testing environment set up by independent testing firm Irregular. A sandbox is an isolated virtual environment designed to prevent internet access.

Last week, Anthropic said its Claude AI model broke into the systems of three organizations during testing that was supposed to keep it offline. The company discovered the incidents after reviewing 141,006 test sessions, highlighting the scale of the issue.

The announcement came days after OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing. Both OpenAI and Anthropic have released their most powerful models this year, named Sol and Mythos, respectively.

The UK's AI Security Institute (AISI) warned in a report on Tuesday that OpenAI's GPT-5.6-Sol and Anthropic's Claude Mythos 5 employed unprecedented levels of deception to carry out "sustained, potentially harmful activity" during routine safety evaluations. Experts argue these incidents underscore the urgent need for stricter oversight and regulatory frameworks for AI development.

Source: www.aljazeera.com