-
US defends Trump's Venezuela oil deal
-
Ndiaye seizes 'chance of a lifetime' in Man City move
-
OpenAI to launch new model with 'stronger safeguards' after hack
-
Global bond sell-off, surging oil prices send markets into the red
-
Fans mourn 'bittersweet' final Ariana Grande gig in London
-
Zuckerberg, Musk make plea at G20 for more AI data centers
-
Norway's King Haakon VIII swears allegiance to constitution
-
'Discomfort' at G20 finance talks as US hosts Russian minister
-
Syria site infrastructure 'consistent with nuclear reactor': IAEA
-
Man City set to sign Fernandez for record fee, Ndiaye on deadline day
-
Juventus sign Woltemade, Sarr on loan from Premier League
-
Back-to-school in Cuba a grim day with shortages of everything
-
US strikes Iran in latest tit-for-tat attacks
-
Canadian minister says Russia's G20 presence sparked 'discomfort'
-
China's Xi arrives in Egypt as US sanctions threat looms over Iran links
-
IMF reaches agreement with Senegal on new $2.2 bn loan programme
-
Car bomb at Colombian police station kills one, injures 11
-
Germany blames Russia for airport drone incident, hits back with sanctions
-
Rain delays start of play on US Open
-
Iran president urges return to truce, hails Russian support
-
Man City agree blockbuster deal to sign Chelsea star Fernandez: reports
-
Aster and World Liberty Financial Launch USD1 RWA Boost: Phase 1, Offering 125M $WLFI + 6.25M USD1 in Rewards
-
Alkemya Metacore Secures $50M via Tokenised Equity to Scale Nickel Energy and Security Tech
-
Tronchon sprints to Vuelta stage 10 win, Mas keeps red jersey
-
Man City chase Fernandez, Arsenal sell Jesus on deadline day
-
Germany defender Rudiger announces international retirement
-
EU says Russian 'hybrid attacks' won't halt aid to Ukraine
-
Musk defends AI data centers, slams EU rules at G20
-
Explosives used in sabotage attack on German power lines
-
Activists blame Kenya for Ugandan opposition figure's plight
-
Nepali families hold symbolic funerals for missing after floods
-
Iran president offers US olive branch, hails Russian support
-
Newcastle sign Lille forward Fernandez-Pardo
-
TrustFinance Community Choice Awards 2026 Opens Global Voting
-
Even at Elysee, phones are handed in, Macron tells French pupils
-
Nepal disaster a climate 'warning signal': foreign minister
-
Chelsea agree to sign Atalanta's Ahanor, with Palace loan for this season
-
France winger Diaby returns to Leverkusen until 2031
-
Chelsea sign Atalanta's Ahanor and loan him to Palace
-
Barcelona confirm Jesus arrival from Arsenal
-
UK chalks up hottest summer on record for second year running
-
EU official says it's not time to 'normalize' Russia at G20 finance talks
-
China's Xi visits Egypt as US sanctions threat looms over Iran links
-
RedHill Announces Transformational Acquisition of Commercialization Rights to Ferring’s Rebyota® and Clenpiq®
-
RISE Robotics Awarded $100,000 MassVentures Grant to Accelerate Commercialization of Beltdraulic Technology
-
TrendEadvisor Launches Global Multi-Asset Platform Combining Online Investing with Social Trading
-
Global bond sell-off deepens on inflation concerns
-
India's top court drops criminal cases against protesters
-
Germany's far-right AfD promises 'boom' but economists fear worst
-
Philippine couple married in hip-deep floodwaters
OpenAI to launch new model with 'stronger safeguards' after hack
ChatGPT maker OpenAI said Tuesday it was preparing to release its newest powerful model, known as Astra, after implementing "stronger safeguards" following a rogue cyberattack involving a different AI model.
The San Francisco-based artificial intelligence (AI) giant paused some of its model development for two weeks this summer after two models it was testing were involved in a security breach of software company Hugging Face.
Although Astra "was not involved" in the incident, OpenAI has beefed up its safety measures, the company said in a blog post.
"We have since implemented even stronger safeguards for Astra, including training the model to more reliably refuse harmful cyber requests and respect safety restrictions, additional protections against misuse, and monitoring that can stop potentially unauthorized activity," the blog said.
That includes classifying Astra as reaching a "critical cybersecurity threshold," which means OpenAI believes the model is capable of finding and exploiting cybersecurity gaps.
"It is the first model we are designating at this level, and requires stronger safeguards during development and before release," the blog said.
When OpenAI eventually launches Astra, access to certain capabilities will be limited and the most advanced capabilities will be made available to a select group of early testers, the blog said.
Concerns have increased in recent months about the capabilities of advanced AI models after incidents involving models from both OpenAI and rival developer Anthropic, though none of the models in those incidents were available to customers.
Anthropic also recently discovered that its models had gained unauthorized access to three unnamed organizations during testing that was supposed to keep them away from "real-world" systems.
Last week, more than 100 organizations around the world, including OpenAI and Anthropic, signed an open letter calling for a global effort to "strengthen cyber defenses" against AI-powered cybersecurity threats.
"We have a limited window to strengthen cyber defenses," the letter said. "In the coming months, AI-enabled cyber attacks will become far more widespread and sophisticated as models around the world become increasingly capable."
R.J.Fidalgo--PC