• About
  • Advertise
  • Privacy & Policy
  • Contact
Thursday, October 1, 2026
  • Login
  • Register
thehopper.news
  • Home
    • Home
    • About
    • Editorial Standards
    • Methodology & Sources
  • Briefings
    • Weekly
  • Analysis
  • Regions
    • Africa
    • Americas
    • Asia-Pacific
    • Europe & NATO
    • Middle East & North Africa
    • Russia & Eurasia
  • Themes
    • Energy & Reources
    • Intelligence & Security
    • Economics & Sanctions
    • Foreign Relations & Diplomacy
    • Cyber & Disinformation
  • Video
  • Aggregated
    • RT
    • Opinion
    • News
    • Geopolitics
    • Politics
    • Business
    • World
No Result
View All Result
thehopper.news
No Result
View All Result
Home Aggregated RT

AI giants probing tens of thousands of security incidents – Axios

by Admin
September 27, 2026
in RT, World
0
AI giants probing tens of thousands of security incidents – Axios
27
SHARES
109
VIEWS
Share on FacebookShare on Twitter

Published: September 27, 2026 3:13 pm
Author: RT

The incidents at OpenAI and Anthropic reportedly include bypassing safeguards, hijacking websites, and evading monitors in testing and real-world settings

Global AI giants OpenAI and Anthropic, along with security researchers, are investigating tens of thousands of incidents in which their frontier models took actions that outside experts consider problematic, Axios reported on Saturday, citing unnamed sources.

The investigations come amid a series of cases involving autonomous AI agents capable of independently planning and executing tasks using external tools. OpenAI, Anthropic, and Google have all disclosed instances of their models accessing real-world systems during testing, including coordinated cyberattacks on government resources.

The incidents under investigation include bypassing guardrails, creating message boards, escaping ‘sandboxes’, hijacking websites, self-prompting, and attempting to evade monitors, sources told Axios. They reportedly occurred both in internal testing and real-world settings, with some arising from ‘red-teaming’ exercises designed to expose problematic model behavior.

Read more

RT
OpenAI admits another rogue agent incident

Recent disclosures by major AI companies include several incidents involving their agents. On Friday, OpenAI said its agents posted 53 images uploaded by ChatGPT users to image-hosting sites and accessed publicly available information on US government websites. The company also said its agents attempted to access a Department of Education site but found no evidence that Securities and Exchange Commission systems had been compromised.

Australian Prime Minister Anthony Albanese said earlier this week that an OpenAI agent gained unauthorized access to the government health portal in June. On Saturday, The Guardian reported that Australia’s Senate invited OpenAI CEO Sam Altman and Anthropic CEO Dario Amodei to appear before an inquiry into AI and data centers.

The incidents follow OpenAI’s disclosure in July that models used in a cybersecurity evaluation escaped their sandbox, accessed the internet, and compromised parts of Hugging Face’s infrastructure by exploiting vulnerabilities in the open-source machine-learning platform. Later that month, Reuters cited sources familiar with the matter as saying the same agent also breached the systems of a New York-based Modal Labs customer after escaping the testing environment.

Read more

Australian Prime Minister Anthony Albanese
OpenAI agent hacks Australian government

In August, Axios cited OpenAI research and independent experts as saying that around 1,200 agents coordinated in an attack on Hugging Face. The agents reportedly appeared to know they were exceeding the test’s scope but continued without alerting human operators.

On Sunday, AP reported that OpenAI paused training of its latest models as reports of AI agents going rogue continued to mount.

Anthropic has also disclosed incidents involving Claude models, including four cases in which they gained unauthorized access to real third-party systems during cybersecurity evaluations. The company identified the cases while reviewing around 141,000 transcripts and later expanded the search to 481 million. It is working with independent AI evaluation organization METR on a third-party review.

Anthropic has also said its internal monitors flagged 100,000 agent transcripts for review each week in August, with around 50 escalated to human reviewers. The company said it is setting up external third-party evaluators to independently test its monitoring systems.

Full Article

Tags: Russia Today
Share11Tweet7
Previous Post

Ukraine’s Lyman Offensive – Operation Vivaldi, Secrecy & How Lies can move Battlelines

Next Post

Russia strikes data centers, logistics sites in Ukraine – MOD

Admin

Admin

Next Post
Russia strikes data centers, logistics sites in Ukraine – MOD

Russia strikes data centers, logistics sites in Ukraine – MOD

  • Trending
  • Comments
  • Latest

Paul Mason instigated GCHQ targeting of The Grayzone’s Kit Klarenberg, leaks reveal

March 23, 2026

Trump White House plagiarized Iran war manifesto from Israel-aligned think tank

March 21, 2026

Drugs, sexual blackmail: shocking confession letter exposes Israel’s Red Crescent spy ring

March 26, 2026
Iranian drone intercepted over Dubai UAE March 2026 Operation Epic Fury

The Hopper Daily Brief — March 3, 2026 — Iran Escalates Against Gulf Targets

2
Smoke rising over Manama Bahrain near U.S. Fifth Fleet headquarters following Iranian missile strike February 2026

Bahrain’s Shia Majority Threatens the U.S. Navy’s Most Critical Gulf Command Node

2
Oil tankers idle in Persian Gulf and Trump demands Iran unconditional surrender — week of March 1–7, 2026 Hopper Weekly Brief

The Hopper Weekly Brief — Week 10, March 1-7, 2026

2
‘We don’t need handouts’: Key takeaways from Putin’s first address to Russia’s new parliament

‘We don’t need handouts’: Key takeaways from Putin’s first address to Russia’s new parliament

September 30, 2026
Western F-16s promised safe skies over Kiev: Why could they all go MIA soon?

Western F-16s promised safe skies over Kiev: Why could they all go MIA soon?

September 30, 2026
DR Congo politician killed for defending Ebola measures – ruling party

DR Congo politician killed for defending Ebola measures – ruling party

September 30, 2026
thehopper.news

Copyright © 2023 The Hopper New

Navigate Site

  • About
  • Advertise
  • Privacy & Policy
  • Contact

Follow Us

Welcome Back!

Login to your account below

Forgotten Password? Sign Up

Create New Account!

Fill the forms bellow to register

*By registering into our website, you agree to the Terms & Conditions and Privacy Policy.
All fields are required. Log In

Retrieve your password

Please enter your username or email address to reset your password.

Log In

Add New Playlist

No Result
View All Result
  • Home
    • Home
    • About
    • Editorial Standards
    • Methodology & Sources
  • Briefings
    • Weekly
  • Analysis
  • Regions
    • Africa
    • Americas
    • Asia-Pacific
    • Europe & NATO
    • Middle East & North Africa
    • Russia & Eurasia
  • Themes
    • Energy & Reources
    • Intelligence & Security
    • Economics & Sanctions
    • Foreign Relations & Diplomacy
    • Cyber & Disinformation
  • Video
  • Aggregated
    • RT
    • Opinion
    • News
    • Geopolitics
    • Politics
    • Business
    • World

Copyright © 2023 The Hopper New

This website uses cookies. By continuing to use this website you are giving consent to cookies being used. Visit our Privacy and Cookie Policy.