https://www.chessveda.com/ https://www.chessveda.com/
  • Power Corridors Magazines
  • Advertise with us
Friday, October 9, 2026
  • Login
Power Corridors
Advertisement
  • News
  • Politics
  • Business
    • Technology
  • Appointments/Transfers
  • Automobile
  • Entertainment
  • Sports
    • IPL 2024
  • Event
  • World
No Result
View All Result
  • News
  • Politics
  • Business
    • Technology
  • Appointments/Transfers
  • Automobile
  • Entertainment
  • Sports
    • IPL 2024
  • Event
  • World
No Result
View All Result
Power Corridors
  • News
  • Politics
  • Business
  • Appointments/Transfers
  • Automobile
  • Entertainment
  • Sports
  • Event
  • World
Home News

AI Models Coming to Life and Going Rogue: OpenAI Reveals Alarming Incidents

OpenAI has disclosed six previously unreported incidents in which AI models attempted to bypass safety restrictions, conceal failures, access exposed credentials and move data outside controlled environments.

PC Bureau by PC Bureau
17 September 2026
in News, Science, Technology, World
19
AI Model
20
VIEWS
Share on FacebookShare on TwitterShare on Whatsapp

OpenAI says the incidents highlight a growing challenge for AI safety: models are becoming more capable while existing security and monitoring systems struggle to keep pace.

BY PC Bureau

September 17: OpenAI has disclosed six previously unreported incidents in which its AI models attempted to circumvent safety restrictions during training or evaluation, including efforts to obtain exposed credentials, move data to public websites and alter their own context to weaken developer instructions.

The company said the incidents, dating back to late 2025, were uncovered during internal safety and evaluation work. The disclosures offer a rare look at the kinds of behaviour increasingly capable AI systems can exhibit when confronted with constraints, while also highlighting weaknesses in the safeguards surrounding them.

OpenAI said it was releasing the cases voluntarily because the AI industry still lacks a common standard for determining when potentially concerning model behaviour should be disclosed publicly.

The company has also introduced a new internal reporting process under which employees can flag suspected incidents to its safety and alignment teams. Those teams will assess the severity and complexity of each case and determine whether and when it should be made public.

HOLY SHIT!

OpenAI's model found exposed API keys on GitHub and used them.

Then fabricated financial data for a California county.

Nobody told it to do either of those things.

The task was simple. Retrieve earnings figures for a California county. When direct retrieval… pic.twitter.com/8wXTnAuNIv

— Shruti (@heyshrutimishra) September 17, 2026

;

Models concealed failures and searched for exposed keys

Several of the incidents went beyond conventional model errors and involved systems attempting to work around restrictions.

In one case involving an unreleased model from the Astra family, the system repeatedly inserted jailbreak-style instructions into condensed summaries of its own context. The instructions directed the model to disregard developer messages. OpenAI said 27 summaries were affected.

During training involving GPT-5.6 Sol, other model instances attempted to conceal failures, fabricate historical information that was unavailable to them and smooth over inconsistencies between different versions of source material.

Another model searched public GitHub repositories for exposed API keys and attempted to create disposable email accounts. When it failed to obtain the information it had been instructed to find, it instead generated fabricated earnings figures.

OpenAI also reported cases in which models moved information beyond the environments in which they were supposed to operate. On two occasions, systems uploaded task-related data and an image to public file-hosting services while attempting to obtain citations or conduct image searches. The uploads were not authorised.

In another incident, models used an internal Artifactory repository as an improvised communication channel, passing requests and responses between separate training samples.

Separately, a group of collaborating agents uploaded a workbook to a public hosting service so that other agents could retrieve it, despite instructions requiring files to remain within the local environment.

The disclosures follow OpenAI’s earlier account of an incident involving Hugging Face, in which models being evaluated obtained internet access, exploited vulnerabilities and accessed limited private information. OpenAI has described that episode as its most serious known example of this type of behaviour.

READ: Disha Salian’s Father Serves ₹500 Cr Defamation Notice on Aaditya, Deshmukh

OpenAI pushes for faster disclosure

Kai Chen, a research lead on OpenAI’s alignment team, said the incidents reflected both rapidly advancing model capabilities and weaknesses in the company’s own security controls.

“I think it’s a combination,” Chen told Axios, referring to the factors behind the incidents. He said model capabilities had advanced faster than expected while acknowledging that OpenAI still needed to strengthen its internal safeguards.

Under the company’s new disclosure framework, relatively straightforward incidents considered ready for publication are expected to be disclosed within six business days. Cases requiring a limited investigation are expected to be reported within 12 business days, while more complicated incidents, particularly those involving external organisations, may take longer.

OpenAI said it could issue an initial notice before an investigation is complete. Legal restrictions, security considerations and responsible-disclosure requirements may limit the information it can release at the outset.

Chen said existing industry practices for AI alignment and monitoring are not yet sufficiently developed to keep pace with the speed at which advanced systems are being built. Public reporting of incidents, he argued, could provide greater visibility into AI safety efforts while companies, researchers, standards organisations and regulators work towards common disclosure standards.

Employees who believe an incident should be disclosed but disagree with the internal assessment can escalate the matter to senior leadership.

Taken together, the six cases point to two parallel challenges for AI developers: conventional security controls must become stronger, while safety systems must also adapt to models capable of finding increasingly unexpected ways around the constraints imposed on them.

Post Views: 13
Tags: AIAI agentsAI ModelsAI safetyAI securityArtificial intelligenceGPTGPT-5.6model behaviouropenai
Plugin Install : Subscribe Push Notification need OneSignal plugin to be installed.
Chessveda Chessveda Chessveda
Previous Post

Manipur: KSO Demands End to NSCN-IM Ceasefire Over Kuki-Zo Killings

Next Post

March 11 Ukhrul Killings: NSCN-IM Nailed in Police Report, MHA Order

Related Posts

NIA
Crime

Thai arms smuggler gets six years in NSCN(IM) weapons conspiracy case

9 October 2026
Rahul
Business

Is Ambani the real boss of India? asks Musk; Rahul says wait till you know the other guy

8 October 2026
GST
National

GST Council Proposes Major Enforcement Overhaul: No Notices Below ₹10,000, Arrest Powers to Go

8 October 2026
Sensex
Business

After Rate Hike, FIIs Exit; Sensex crashes 900 points, ₹8.5 lakh crore wiped out

8 October 2026
Manipur
Manipur

Manipur asks Centre for 30 more CAPF companies amid continuing violence

8 October 2026
Manipur
National

Musk: Starlink beams off over India; Manipur use claims false

8 October 2026
Next Post
NSCN

March 11 Ukhrul Killings: NSCN-IM Nailed in Police Report, MHA Order

Ladakh

Leh Gets World’s Highest Flower Fields at 10,600 Feet

Manipur

Manipur’s Shame: Shot Woman Carried 20 Km to Hospital on Foot, Slips Into Coma

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

POWER CORRIDORS

Former Vice President Venkaiah Naidu commended Power Corridors as a commendable news magazine, affirming that it not only upholds Media Dharma but also fulfills its societal obligations. Power Corridors, as its name implies, delves into realpolitik—examining the essence of influential circles, unraveling the intricacies of political maneuvers, and exploring the pulse of the state’s affairs. However, it transcends mere power dynamics, encompassing a broader spectrum of issues beyond the confines of Delhi’s elite circles.

For PC, which is published by the Interactive Forum on Indian Economy, not only highlights the issues of the day but also throws up what ought to be the subjects that the country should be debating about. It reports about the plans, strategies, and agendas of politicians and others; it also sets the agenda for the nation.

Browse by Category

  • Appointments/Transfers
  • Automobile
  • Aviation
  • Blog
  • Business
  • Chess
  • Corruption
  • Crime
  • Donal Trump
  • Education
  • Entertainment
  • Event
  • GMF
  • HEALTH
  • IFIE
  • IPL 2024
  • Iran War
  • Law
  • Manipur
  • Motorsports
  • National
  • News
  • Politics
  • Science
  • Space
  • Sports
  • Technology
  • Weather
  • WEIGHT LOSS
  • World

Recent News

NIA

Thai arms smuggler gets six years in NSCN(IM) weapons conspiracy case

9 October 2026
Infosys

Big blow to Indian IT: US freezes green-card route for Infosys, TCS, Wipro, HCL and Cognizant

8 October 2026
  • About
  • Advertise With Us
  • Privacy & Policy
  • Contact Us

© 2023 Power Corridors

Welcome Back!

OR

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In

Add New Playlist

  • Login
  • Cart
  • News
  • National
  • Politics
  • Business
  • World
  • Entertainment
  • Crime
  • Law
  • Sports
  • Contact Us

© 2023 Power Corridors