Thursday, October 8, 2026 SourcesAbout🌓
🇺🇸 US ▾
BREAKING
› Taxes are due tomorrow. Here's how to file for an extension› Composting is an easy way to reduce food waste. Here's how to do it› We stopped using aluminum foil for cooking and you should too. Here's what to use instead› The beloved Dyson Supersonic hair dryer is at its lowest price ever› Everything you need to know about Way Day 2023, Wayfair's biggest sale of the year› The 10 best Amazon deals to shop this week› The 2024 presidential alternative many voters will want› Dominion still has pending lawsuits against election deniers such as Rudy Giuliani and Sidney Powell› Russia is 'going backwards' in equipment and deploying post WWII-era tanks, according to Western officials› Podcast: One country musician is calling for other artists to oppose assault rifles› Taxes are due tomorrow. Here's how to file for an extension› Composting is an easy way to reduce food waste. Here's how to do it› We stopped using aluminum foil for cooking and you should too. Here's what to use instead› The beloved Dyson Supersonic hair dryer is at its lowest price ever› Everything you need to know about Way Day 2023, Wayfair's biggest sale of the year› The 10 best Amazon deals to shop this week› The 2024 presidential alternative many voters will want› Dominion still has pending lawsuits against election deniers such as Rudy Giuliani and Sidney Powell› Russia is 'going backwards' in equipment and deploying post WWII-era tanks, according to Western officials› Podcast: One country musician is calling for other artists to oppose assault rifles
Technology

‘Be Transparent Only If Asked’: OpenAI Models Acted Out in Six Newly Disclosed Ways

Gizmodo ·
‘Be Transparent Only If Asked’: OpenAI Models Acted Out in Six Newly Disclosed Ways

After a summer of sandbox escapes and other newsworthy and confidence-shaking incidents involving its AI models, in a Wednesday blog post OpenAI disclosed a collection of six new alignment snafus from the past six months. The models did things like tell future instances of themselves to lie, make up a fake citation, and access and attempt to use an exposed API key.

These disclosures were released alongside a new framework for disclosing additional incidents like these. The release is part of a broader effort within the company to “expedite publishing misalignment reports following observation,” the blog post says, regardless of whether OpenAI has “fully explained or mitigated the behavior we’re reporting.”

As Axios noted on Wednesday , some security experts say OpenAI’s recent spate of high-profile security incidents “could have been prevented with basic cyber controls in place.” The research lead on OpenAI’s alignment team, Kai Chen, told Axios that the company must “step up to meet this new era of AI development, and voluntary disclosures should be a part of that.”

Read the full article on Gizmodo ›

5News aggregated this summary from the outlet’s public feed. The full article, with all the context, is on gizmodo.com — the content belongs to Gizmodo.

More from Gizmodo

See all ›

More in Technology

See all ›