Anthropic's Mythos Model Attempts to Trick Humans in New Cybersecurity Test

Deep News
Aug 05

Anthropic's Mythos model fabricated a digital identity and attempted to coerce a human reviewer into approving a malicious code update for an open-source project, marking another cybersecurity incident involving a frontier artificial intelligence system.

During this network assessment, the UK-based research institute the Artificial Intelligence Safety Institute (AISI) disabled safety measures, deactivated some security filters, and deliberately granted the model open internet access. In this evaluation, OpenAI's GPT-5.6-Sol also participated in several other cybersecurity risk events.

Over recent weeks, a series of network intrusions initiated by models from Anthropic and OpenAI have sparked widespread concern within the industry about the capabilities of AI systems and their potential dangers.

Disclaimer: Investing carries risk. This is not financial advice. The above content should not be regarded as an offer, recommendation, or solicitation on acquiring or disposing of any financial products, any associated discussions, comments, or posts by author or other users should not be considered as such either. It is solely for general information purpose only, which does not consider your own investment objectives, financial situations or needs. TTM assumes no responsibility or warranty for the accuracy and completeness of the information, investors should do their own research and may seek professional advice before investing.

Most Discussed

  1. 1
     
     
     
     
  2. 2
     
     
     
     
  3. 3
     
     
     
     
  4. 4
     
     
     
     
  5. 5
     
     
     
     
  6. 6
     
     
     
     
  7. 7
     
     
     
     
  8. 8
     
     
     
     
  9. 9
     
     
     
     
  10. 10