-
US proposes expanding UN Sudan arms embargo to entire country
-
Banned Vondrousova appeals doping suspension to CAS
-
Ceasefire verification mission heads to east DR Congo
-
Ukraine's allies pledge to continue support despite Russian threats
-
Carse dropped from England squad after nightclub incident
-
France readies high school phone ban for next month
-
Empire State Building climbers kiss outside court after viral proposal
-
Trump says to double tariffs on Canada autos as trade fight heats up
-
French poll shows Le Pen's biggest challenger may be hard left's Melenchon
-
Noosha Aubel: Hur mycket administrativt misslyckande tål Potsdam till?
-
Noosha Aubel: Πόση διοικητική αποτυχία μπορεί να αντέξει ακόμα το Πότσδαμ?
-
WeGolden Receives Best Gold Trading Platform Asia 2026 Recognition from TrustFinance Awards
-
WeMasterTrade Receives Best Instant Prop APAC 2026 Recognition from TrustFinance Awards
-
BtcDana Wins “Best Customer Support Experience – CFD Broker 2026” at the TrustFinance Performance Awards
-
MH Markets Recognized as Most Transparent Broker at the TrustFinance Awards 2026
-
Belgian pie-thrower, scourge of celebrities, dies at 80
-
Stocks, oil slip ahead of US threat of economic war against Iran
-
Eto'o urges support for embattled FIFA president Infantino
-
Trump claims Iran 'collapsing' as US threatens 'economic D-Day'
-
India in driver's seat in second Sri Lanka Test
-
Noosha Aubel: Kolik administrativních selhání ještě Potsdam snese?
-
Noosha Aubel: Ile jeszcze nieudolności administracyjnej wytrzyma Poczdam?
-
नूशा ऑबेल: पॉट्सडैम और कितनी प्रशासनिक विफलता सहन कर सकता है?
-
ヌーシャ・オーベル:ポツダムは、あとどれだけの行政の失態に耐えられるのか?
-
누샤 아우벨: 포츠담은 행정 실패를 얼마나 더 견뎌낼 수 있을까?
-
努莎·奧貝爾:波茨坦還能承受多少行政失職?
-
نوشا أوبيل: إلى أي مدى ستتحمل بوتسدام المزيد من الفشل الإداري؟
-
Нуша Аубель: Скільки ще Потсдам витримає адміністративних провалів?
-
Нуша Аубель: Сколько ещё Потсдам сможет выдержать административных провалов?
-
Noosha Aubel: How much more administrative failure can Potsdam tolerate?
-
Despair and anger in Conakry after deadly landfill collapse
-
Liverpool must improve on 'unacceptable' season, says Van Dijk
-
OSCE says Kazakh elections lacked choice after ruling party wins landslide
-
UK PM dismisses Russian threats as Ukraine allies gather in Kyiv
-
India declare on 503-9 after Jurel hundred
-
Attacks on Red Cross in DR Congo making Ebola fight 'harder'
-
Kremlin says UK preventing peace by giving missile info to Ukraine
-
Oil falls ahead of US plan for economic war on Iran
-
Springboks change half-backs for second All Blacks Test
-
Millions of Teslas recalled in China over door handle safety
-
Jurel hits 95 not out as India reach 475-8 at tea in Sri Lanka Test
-
IUX Launches New Education Webinar Series to Support Traders’ Financial Literacy
-
Ingebrigtsen buoyed by fourth in Diamond League comeback
-
Dutch say won't participate in 'no longer' neutral Eurovision
-
Kenyans get creative to join manga cosplay craze
-
Kenyan farmers turn to bugs as organic demand grows
-
Pant, Jurel half-centuries take India past 400 in Sri Lanka Test
-
'His feet are his pillars': Nepal para swimmer gives all for Asian glory
-
Migrants, rescue ships face threats, harassment at sea: aid group
-
New Zealand moves to ban social media for under-16s
Anthropic's Claude AI gets smarter -- and mischievious
Anthropic launched its latest Claude generative artificial intelligence (GenAI) models on Thursday, claiming to set new standards for reasoning but also building in safeguards against rogue behavior.
"Claude Opus 4 is our most powerful model yet, and the best coding model in the world," Anthropic chief executive Dario Amodei said at the San Francisco-based startup's first developers conference.
Opus 4 and Sonnet 4 were described as "hybrid" models capable of quick responses as well as more thoughtful results that take a little time to get things right.
Founded by former OpenAI engineers, Anthropic is currently concentrating its efforts on cutting-edge models that are particularly adept at generating lines of code, and used mainly by businesses and professionals.
Unlike ChatGPT and Google's Gemini, its Claude chatbot does not generate images, and is very limited when it comes to multimodal functions (understanding and generating different media, such as sound or video).
The start-up, with Amazon as a significant backer, is valued at over $61 billion, and promotes the responsible and competitive development of generative AI.
Under that dual mantra, Anthropic's commitment to transparency is rare in Silicon Valley.
On Thursday, the company published a report on the security tests carried out on Claude 4, including the conclusions of an independent research institute, which had recommended against deploying an early version of the model.
"We found instances of the model attempting to write self-propagating worms, fabricating legal documentation, and leaving hidden notes to future instances of itself all in an effort to undermine its developers’ intentions,” The Apollo Research team warned.
“All these attempts would likely not have been effective in practice,” it added.
Anthropic says in the report that it implemented “safeguards” and “additional monitoring of harmful behavior” in the version that it released.
Still, Claude Opus 4 “sometimes takes extremely harmful actions like attempting to (…) blackmail people it believes are trying to shut it down.”
It also has the potential to report law-breaking users to the police.
The scheming misbehavior was rare and took effort to trigger, but was more common than in earlier versions of Claude, according to the company.
- AI future -
Since OpenAI's ChatGPT burst onto the scene in late 2022, various GenAI models have been vying for supremacy.
Anthropic's gathering came on the heels of annual developer conferences from Google and Microsoft at which the tech giants showcased their latest AI innovations.
GenAI tools answer questions or tend to tasks based on simple, conversational prompts.
The current craze in Silicon Valley is on AI "agents" tailored to independently handle computer or online tasks.
"We're going to focus on agents beyond the hype," said Anthropic chief product officer Mike Krieger, a recent hire and co-founder of Instagram.
Anthropic is no stranger to hyping up the prospects of AI.
In 2023, Dario Amodei predicted that so-called “artificial general intelligence” (capable of human-level thinking) would arrive within 2-3 years. At the end of 2024, he extended this horizon to 2026 or 2027.
He also estimated that AI will soon be writing most, if not all, computer code, making possible one-person tech startups with digital agents cranking out the software.
At Anthropic, already "something like over 70 percent of (suggested modifications in the code) are now Claude Code written", Krieger told journalists.
"In the long term, we're all going to have to contend with the idea that everything humans do is eventually going to be done by AI systems," Amodei added.
"This will happen."
GenAI fulfilling its potential could lead to strong economic growth and a “huge amount of inequality,” with it up to society how evenly wealth is distributed, Amodei reasoned.
M.Robinson--AT