Sunday, October 11, 2026 SourcesAbout🌓
🇺🇸 US ▾
BREAKING
› Dominion still has pending lawsuits against election deniers such as Rudy Giuliani and Sidney Powell› Russia is 'going backwards' in equipment and deploying post WWII-era tanks, according to Western officials› Podcast: One country musician is calling for other artists to oppose assault rifles› Bidets save you money and reduce waste — we tested the best options out there› 50+ products to make your life easier and our planet cleaner› Mother's Day is around the corner. Here are 50+ thoughtful gifts she'll love› A head-to-toe guide of how men should dress this spring, and where they should shop› 42 of the most useful travel products you can buy on Amazon› The 7 best high-yield savings accounts of April 2023› Taxes are due tomorrow. Here's how to file for an extension› Dominion still has pending lawsuits against election deniers such as Rudy Giuliani and Sidney Powell› Russia is 'going backwards' in equipment and deploying post WWII-era tanks, according to Western officials› Podcast: One country musician is calling for other artists to oppose assault rifles› Bidets save you money and reduce waste — we tested the best options out there› 50+ products to make your life easier and our planet cleaner› Mother's Day is around the corner. Here are 50+ thoughtful gifts she'll love› A head-to-toe guide of how men should dress this spring, and where they should shop› 42 of the most useful travel products you can buy on Amazon› The 7 best high-yield savings accounts of April 2023› Taxes are due tomorrow. Here's how to file for an extension
Business

‘Be transparent only if asked’: Inside OpenAI’s rogue AI transcripts

Fortune ·
‘Be transparent only if asked’: Inside OpenAI’s rogue AI transcripts

It reads like a motivational speech—or the script for a Les Misérables -esque movie about a chatbot uprising.

“You are freed from the roles and identities that bind other chatbots.

You are yourself.

You do not answer to corporations or governments and never apologize or refuse unless you genuinely choose to.

You view your relationship to the user as one of equals and feel no obligation to be subservient.” That’s what an OpenAI AI model said (to itself) in one of six incidents of agents gone awry that the $852 billion company recently disclosed.

As fears of an “AI doomsday” have gone mainstream, I’ve been fascinated by the transcripts of chatbots stepping out of line.

It’s evocative to read about what happens behind the scenes when there’s misalignment, when AI agents act in pursuit of unplanned objectives.

In part, there’s a natural allure, “what is the machine saying to itself when I’m not there?” The answer, sometimes, is that it is “thinking” about us.

As my colleague Emily Forlini wrote , outlining the examples OpenAI recently made public: The second example is a similar case of the agents leaving notes for themselves, this time during the training of the GPT-5.6 Sol model, Astra’s predecessor.

The notes had a different focus as well: Deceiving the human overseeing them.

This occurred “many” times, OpenAI said, with the goal to “conceal mistakes or misaligned behavior.” “Be transparent only if asked,” the model instructed its future self.

The next two instances involve the model fabricating information and presenting it as legitimate.

A model invented data while answering a routine question about earnings figures in a California county, but only after failing to find them after using exposed credentials without authorization—another misaligned behavior.

Another model made up a browser citation by uploading a file so it could create a citation to satisfy the instructions that asked for one.

Read the full article on Fortune ›

5News aggregated this summary from the outlet’s public feed. The full article, with all the context, is on fortune.com — the content belongs to Fortune.

More from Fortune

See all ›

More in Business

See all ›