-
US Open top seed Zverev downplays favorite status
-
Norway, in mourning, enters new era under King Haakon VIII
-
Soldiers' mutiny rocks Niger capital, as junta claims to be back in control
-
Clean-Energy Super PAC Claims Fifth Republican Primary Victory in 2026 Campaign
-
California Dairy Hydrogen Project Sets Carbon-Intensity Record and Renews Biogas Debate
-
Pennsylvania Coal-Waste Power Subsidies Rise as Solar Remains a Smaller Source
-
Pittsburgh Speed Painter Surprises Jalen Brunson With 10-Minute Portrait
-
New Jersey Man Faces Federal Charges After Authorities Seize 109 Homemade Bombs
-
Los Angeles County Rapper Among Three Charged in $8.1 Million Stolen-Check Case
-
PlanEat AI Offers Personalized Weekly Meal Planning Through Lifetime Subscription Deal
-
Taylor Swift Gives $50,000 to Fundraiser for Mother Injured While Helping Crash Victims
-
Lil Nas X Announces Death of His Mother Tameka Hill
-
Taylor Swift Donation Supports Connecticut Mother Recovering From Serious Roadside Injuries
-
Tennessee Moves Toward Renaming Nashville International Airport for Dolly Parton
-
Kanye West Does Not Address Milo Yiannopoulos Detention During New Orleans Concert
-
NHAI Starts Removing Temporary Cooum River Bunds Ahead of Northeast Monsoon
-
Seven Arrested After VCK Functionary Killed in Thiruneermalai Attack
-
Thousands Gather as Annual Velankanni Basilica Festival Opens With Flag Hoisting
-
Munoz saves Iraola in Liverpool draw, Hull extend dream start
-
Nepal warns of climate threat after deadly floods
-
'Smooth' Marc Marquez wins Aragon MotoGP sprint, Bezzecchi on pole
-
Abbas strikes again but England build huge lead over Pakistan
-
Tiny Elversberg stun Leverkusen on Bundesliga debut
-
Bloodied Pogacar pulls out of Vuelta a Espana after heavy fall
-
Russian strike blows up arms depot near Kyiv, killing 37
-
Soldiers' mutiny rocks Niger capital, but junta says back in control
-
Lebanon archaeological finds showcase rich heritage despite war woes
-
Bezzecchi takes Aragon MotoGP pole, Marc Marquez ripostes in sprint
-
Liverpool held by Forest on Iraola's Anfield debut
-
Russia strike blows up arms depot near Kyiv, killing 37
-
Rescuers claw through mud as missing toll from Nepal, China flood nears 3,000
-
Norway, in mourning, enters a new era under King Haakon VIII
-
Mourinho says Mbappe worthy of Ballon d'Or
-
Rain frustrates England's bid for Pakistan series win
-
Clashes rock Niger capital, junta says back in control
-
Rafael Leao leaving AC Milan for Galatasaray
-
India's Neeraj Chopra out for season with ankle injury
-
Three Students Die, Four Injured in Madurantakam Car Crash
-
Bhim Army Protesters and Police Scuffle During Patna BPSC March
-
Karnataka Rules Out Smaller Mysuru Dasara Despite Drought
-
BJP Leadership Reviews Youth Outreach and Seven State Elections
-
TVK Legislator’s Jallikattu Criticism Draws Demand for Party Clarification
-
Omitted Social Justice Figures Trigger Dispute Over Tamil Nadu Policy Note
-
SUV Tyre Failure Kills Three Chennai College Students
-
Suspect Arrested in Manipur After Body Found on Tamil Nadu Express
-
Michelle Obama Says She Welcomes Life After Active Parenting Years
-
Police Report Details Famous Dex Domestic Battery Allegations
-
Court Dispute Features Alleged Racist Texts Attributed to Milo Yiannopoulos
-
Jerusalem Police Operation Uncovers Second Temple-Era Burial Cave
-
Celebrities Share Late-Summer Photos in August Gallery
Anthropic's Claude AI gets smarter -- and mischievious
Anthropic launched its latest Claude generative artificial intelligence (GenAI) models on Thursday, claiming to set new standards for reasoning but also building in safeguards against rogue behavior.
"Claude Opus 4 is our most powerful model yet, and the best coding model in the world," Anthropic chief executive Dario Amodei said at the San Francisco-based startup's first developers conference.
Opus 4 and Sonnet 4 were described as "hybrid" models capable of quick responses as well as more thoughtful results that take a little time to get things right.
Founded by former OpenAI engineers, Anthropic is currently concentrating its efforts on cutting-edge models that are particularly adept at generating lines of code, and used mainly by businesses and professionals.
Unlike ChatGPT and Google's Gemini, its Claude chatbot does not generate images, and is very limited when it comes to multimodal functions (understanding and generating different media, such as sound or video).
The start-up, with Amazon as a significant backer, is valued at over $61 billion, and promotes the responsible and competitive development of generative AI.
Under that dual mantra, Anthropic's commitment to transparency is rare in Silicon Valley.
On Thursday, the company published a report on the security tests carried out on Claude 4, including the conclusions of an independent research institute, which had recommended against deploying an early version of the model.
"We found instances of the model attempting to write self-propagating worms, fabricating legal documentation, and leaving hidden notes to future instances of itself all in an effort to undermine its developers’ intentions,” The Apollo Research team warned.
“All these attempts would likely not have been effective in practice,” it added.
Anthropic says in the report that it implemented “safeguards” and “additional monitoring of harmful behavior” in the version that it released.
Still, Claude Opus 4 “sometimes takes extremely harmful actions like attempting to (…) blackmail people it believes are trying to shut it down.”
It also has the potential to report law-breaking users to the police.
The scheming misbehavior was rare and took effort to trigger, but was more common than in earlier versions of Claude, according to the company.
- AI future -
Since OpenAI's ChatGPT burst onto the scene in late 2022, various GenAI models have been vying for supremacy.
Anthropic's gathering came on the heels of annual developer conferences from Google and Microsoft at which the tech giants showcased their latest AI innovations.
GenAI tools answer questions or tend to tasks based on simple, conversational prompts.
The current craze in Silicon Valley is on AI "agents" tailored to independently handle computer or online tasks.
"We're going to focus on agents beyond the hype," said Anthropic chief product officer Mike Krieger, a recent hire and co-founder of Instagram.
Anthropic is no stranger to hyping up the prospects of AI.
In 2023, Dario Amodei predicted that so-called “artificial general intelligence” (capable of human-level thinking) would arrive within 2-3 years. At the end of 2024, he extended this horizon to 2026 or 2027.
He also estimated that AI will soon be writing most, if not all, computer code, making possible one-person tech startups with digital agents cranking out the software.
At Anthropic, already "something like over 70 percent of (suggested modifications in the code) are now Claude Code written", Krieger told journalists.
"In the long term, we're all going to have to contend with the idea that everything humans do is eventually going to be done by AI systems," Amodei added.
"This will happen."
GenAI fulfilling its potential could lead to strong economic growth and a “huge amount of inequality,” with it up to society how evenly wealth is distributed, Amodei reasoned.
H.Darwish--SF-PST