Wednesday, 2 September 2026 SourcesAbout🌓
🇮🇳 IN ▾
BREAKING
Sonu Sood reaches Nepal with relief supplies for flood victims: ‘We just want to be with the people and will try to bring their lives back” Sharmila Tagore to make special cameo in Son-in-Law Kunal Kemmu’s Vibe Shah Rukh Khan joins Vodafone Idea as brand ambassador, company unveils new ‘Vi’ logo "Not Sending Inzamam, Miandad Or Anwar": Pakistan Great Roasts Board Over Mass Sackings ED raids payment companies, CAs in online betting case Failed delivery ruse, ₹1.5 lakh plot: Niece kills uncle in Rohini, stages attack to mislead police ‘Furious’ series review: Emmy Rossum powers this tale of feminine rage Two Turkish-flagged commercial vessels collide off Istanbul West Asia conflict LIVE updates: Iran says 11 dead in U.S. strikes on its territory Nepal flash floods aftermath LIVE: Death toll rises to 1,127; more than 4,800 people missing, says Nepal Police Sonu Sood reaches Nepal with relief supplies for flood victims: ‘We just want to be with the people and will try to bring their lives back” Sharmila Tagore to make special cameo in Son-in-Law Kunal Kemmu’s Vibe Shah Rukh Khan joins Vodafone Idea as brand ambassador, company unveils new ‘Vi’ logo "Not Sending Inzamam, Miandad Or Anwar": Pakistan Great Roasts Board Over Mass Sackings ED raids payment companies, CAs in online betting case Failed delivery ruse, ₹1.5 lakh plot: Niece kills uncle in Rohini, stages attack to mislead police ‘Furious’ series review: Emmy Rossum powers this tale of feminine rage Two Turkish-flagged commercial vessels collide off Istanbul West Asia conflict LIVE updates: Iran says 11 dead in U.S. strikes on its territory Nepal flash floods aftermath LIVE: Death toll rises to 1,127; more than 4,800 people missing, says Nepal Police
Science

OpenAI to launch new model with 'stronger safeguards' after hack

The Hindu - Sci-Tech ·
OpenAI to launch new model with 'stronger safeguards' after hack

Account subscription benefits alongside Premium Stories, Editorials, Opinions and more. Unlock these with Subscription

When OpenAI eventually launches Astra, access to certain capabilities will be limited [File] | Photo Credit: REUTERS

ChatGPT maker OpenAI said Tuesday it was preparing to release its newest powerful model, known as Astra , after implementing "stronger safeguards" following a rogue cyberattack involving a different AI model.

The San Francisco-based artificial intelligence (AI) giant paused some of its model development for two weeks this summer after two models it was testing were involved in a security breach of software company Hugging Face.

Although Astra "was not involved" in the incident, OpenAI has beefed up its safety measures, the company said in a blog post.

"We have since implemented even stronger safeguards for Astra, including training the model to more reliably refuse harmful cyber requests and respect safety restrictions, additional protections against misuse, and monitoring that can stop potentially unauthorized activity," the blog said.

That includes classifying Astra as reaching a "critical cybersecurity threshold," which means OpenAI believes the model is capable of finding and exploiting cybersecurity gaps.

"It is the first model we are designating at this level, and requires stronger safeguards during development and before release," the blog said.

When OpenAI eventually launches Astra, access to certain capabilities will be limited and the most advanced capabilities will be made available to a select group of early testers, the blog said.

Concerns have increased in recent months about the capabilities of advanced AI models after incidents involving models from both OpenAI and rival developer Anthropic, though none of the models in those incidents were available to customers.

Anthropic also recently discovered that its models had gained unauthorised access to three unnamed organisations during testing that was supposed to keep them away from "real-world" systems.

Last week, more than 100 organisations around the world, including OpenAI and Anthropic, signed an open letter calling for a global effort to "strengthen cyber defenses" against AI-powered cybersecurity threats.

"We have a limited window to strengthen cyber defenses," the letter said. "In the coming months, AI-enabled cyber attacks will become far more widespread and sophisticated as models around the world become increasingly capable."

In June, U.S. President Donald Trump signed an executive order calling for a voluntary review process in which the government would get early access to new AI models before their release to assess security risks.

Read the full article on The Hindu - Sci-Tech ›

5News aggregated this summary from the outlet’s public feed. The full article, with all the context, is on www.thehindu.com — the content belongs to The Hindu - Sci-Tech.

More from The Hindu - Sci-Tech

See all ›

More in Science

See all ›