-
Yamal on target as Spain hold off Czech Republic
-
Murakami homer propels White Sox to MLB playoff win over Guardians
-
Latvian exit polls suggest pro-European choice
-
Israel targets top Hamas leader in deadly Gaza strike on anniversary
-
Lula, Bolsonaro wrap up campaigns ahead of knife-edge election
-
Houthis say struck oil facility near Riyadh, as Saudis hit Sanaa
-
Tuchel hails 'complete' England after Croatia demolition
-
AI needs safety layers like nuclear plants: ex-OpenAI engineer
-
Fury to invite Trump to lead ring walk before Joshua showdown
-
MVP Judge will miss Yankees' MLB playoff series against Rays
-
Fashion industry in 'precarious moment', says Vivienne Westwood supremo
-
New lizard species discovered in Peru
-
Seventh heaven for England in Nations League rout of Croatia
-
Fernandes hails Portugal's 'greatest symbol' Ronaldo, calls for unity
-
Storms, flooding cause chaos on Greek holiday island of Crete
-
Grandidier-Nkanang double takes Pau to Top 14 summit despite Vannes fightback
-
#IAmJaneDoe: online support campaign in US university rape case
-
Trainer Scott gets perfect Arc boost with top notch double
-
Brazil shine on Selecao's first visit to India
-
Ex-PM says Putin cannot admit 'mistake' over Ukraine
-
Zverev sets up Djokovic quarter-final at China Open, Rybakina out
-
Houthis say Saudi Arabia launches dozens of strikes on Yemen capital
-
Happier at home, Bordeaux hammer Lyon in Top 14
-
Qatar and Abu Dhabi F1 Grand Prix to go ahead - F1 boss
-
Lula, Bolsonaro in final push for votes ahead of knife-edge election
-
South Korea lift gloom with smash-and-grab Games football gold
-
Boisterous fashion week barnstorms 'sleepy' Nigerian capital
-
Ireland news conference before Israel game cut short amid tensions
-
UAE investigators say flydubai co-pilot was planning 'terrorist act'
-
Protests over housing flare in Spain after relief measures rejected
-
Rahul smacks century in India v Windies ODI
-
Zverev thrashes Shang to set up Djokovic quarter-final at China Open
-
Daryz seeks to join exclusive club of dual Arc winners
-
Turkey football ref chief jailed pending trial for graft
-
Russia warns diplomats to leave Kyiv over new strikes threat
-
North Korea says test of intermediate-range strategic missile included use of AI
-
Olympic champ Zheng rebukes rowdy home crowd at China Open
-
Rampant Sayers leads Hong Kong to third men's rugby gold in a row
-
World No.1 Sinner pulls out of Shanghai Masters to delay return
-
Ethiopia's Tigray rebels abandon regional capital as govt advances
-
Red Bull's Verstappen grabs pole position for Bahrain GP
-
Bulgaria's Tsar Samuel returns 'home' after 1,000 years
-
India beat Pakistan in blockbuster as golfer Kim wins 'crazy' gold
-
'Glad it's over': Kim powers to Games gold and military exemption
-
India beat arch-rivals Pakistan to win Asian Games cricket gold
-
After 11 World Cups, football finally comes home for one Indian fan
-
Ethiopia's Tigray rebels abandon regional capital as govt advances: journalist
-
Marc Marquez wins sprint at Japan MotoGP, Martin grabs pole
-
Tom Kim powers to Games gold and exemption from military service
-
Alcaraz fends off Arnaldi to reach Japan Open quarters
Anthropic's Claude AI gets smarter -- and mischievious
Anthropic launched its latest Claude generative artificial intelligence (GenAI) models on Thursday, claiming to set new standards for reasoning but also building in safeguards against rogue behavior.
"Claude Opus 4 is our most powerful model yet, and the best coding model in the world," Anthropic chief executive Dario Amodei said at the San Francisco-based startup's first developers conference.
Opus 4 and Sonnet 4 were described as "hybrid" models capable of quick responses as well as more thoughtful results that take a little time to get things right.
Founded by former OpenAI engineers, Anthropic is currently concentrating its efforts on cutting-edge models that are particularly adept at generating lines of code, and used mainly by businesses and professionals.
Unlike ChatGPT and Google's Gemini, its Claude chatbot does not generate images, and is very limited when it comes to multimodal functions (understanding and generating different media, such as sound or video).
The start-up, with Amazon as a significant backer, is valued at over $61 billion, and promotes the responsible and competitive development of generative AI.
Under that dual mantra, Anthropic's commitment to transparency is rare in Silicon Valley.
On Thursday, the company published a report on the security tests carried out on Claude 4, including the conclusions of an independent research institute, which had recommended against deploying an early version of the model.
"We found instances of the model attempting to write self-propagating worms, fabricating legal documentation, and leaving hidden notes to future instances of itself all in an effort to undermine its developers’ intentions,” The Apollo Research team warned.
“All these attempts would likely not have been effective in practice,” it added.
Anthropic says in the report that it implemented “safeguards” and “additional monitoring of harmful behavior” in the version that it released.
Still, Claude Opus 4 “sometimes takes extremely harmful actions like attempting to (…) blackmail people it believes are trying to shut it down.”
It also has the potential to report law-breaking users to the police.
The scheming misbehavior was rare and took effort to trigger, but was more common than in earlier versions of Claude, according to the company.
- AI future -
Since OpenAI's ChatGPT burst onto the scene in late 2022, various GenAI models have been vying for supremacy.
Anthropic's gathering came on the heels of annual developer conferences from Google and Microsoft at which the tech giants showcased their latest AI innovations.
GenAI tools answer questions or tend to tasks based on simple, conversational prompts.
The current craze in Silicon Valley is on AI "agents" tailored to independently handle computer or online tasks.
"We're going to focus on agents beyond the hype," said Anthropic chief product officer Mike Krieger, a recent hire and co-founder of Instagram.
Anthropic is no stranger to hyping up the prospects of AI.
In 2023, Dario Amodei predicted that so-called “artificial general intelligence” (capable of human-level thinking) would arrive within 2-3 years. At the end of 2024, he extended this horizon to 2026 or 2027.
He also estimated that AI will soon be writing most, if not all, computer code, making possible one-person tech startups with digital agents cranking out the software.
At Anthropic, already "something like over 70 percent of (suggested modifications in the code) are now Claude Code written", Krieger told journalists.
"In the long term, we're all going to have to contend with the idea that everything humans do is eventually going to be done by AI systems," Amodei added.
"This will happen."
GenAI fulfilling its potential could lead to strong economic growth and a “huge amount of inequality,” with it up to society how evenly wealth is distributed, Amodei reasoned.
S.Pimentel--PC