-
Meloni government becomes Italy's longest-lasting
-
Venice to roll out red carpet for Clooney on opening night
-
Funerals in Nepal for missing, a week after deadly floods
-
Hong Kong democracy activist Wong pleads guilty to foreign collusion: AFP
-
US aircraft carrier arrives in Thailand after months at sea
-
Stocks sink, yields at multi-decade highs and oil spikes on Mideast flare-up
-
US aircraft carrier approaches Thai port after months at sea
-
US defends Trump Venezuela oil deal as energy secretary lands in Caracas
-
Nepal's young PM faces gravest test after flood catastrophe
-
Jubilant Monfils celebrates birthday with US Open win
-
What rainbow? LGBTQ Indonesians on edge as hostility mounts
-
The Japanese idealist behind the last deradicalisation NGO in Somalia
-
O'Connor eyes Wallabies World Cup swansong after Force return
-
Rybakina, Gauff advance as Zverev launches US Open campaign
-
Syria UXO causing death, injury every six hours: charity
-
Macron in Britain to help unveil tapestry, hold Burnham talks
-
Tiger Woods set to change not guilty plea in DUI case: court documents
-
China's Xi to hold talks with Sisi in rare visit to Egypt
-
Zelensky warns airlines Russian skies not safe due to Ukraine drones
-
Confident Gauff reaches US Open 2nd round with straight-set win
-
UK vows anti-settler measures as Israel warns of retaliation
-
US postal whistleblower warns ballot system could disrupt midterms
-
New pro Koivun among US Presidents Cup picks
-
John Ternus takes over as Apple enters era of AI, foldable phones
-
G20 finance leaders close talks marked by US welcoming Russia
-
Second-ranked Rybakina eases into US Open 2nd round
-
From Hollywood to big time tennis: teen makes US Open debut
-
Man City sign Fernandez, Ndiaye as Premier League clubs set new spending record
-
Man City secure Premier League record deal to sign Fernandez
-
Alcaraz aims to step it up in round two of US Open title defense
-
Kaepernick says NFL 'blackballing me' decade after protest
-
Honduran judge drops corruption charges against ex-leader Hernandez
-
Atletico loan David from Juve, send Lemar to Elche
-
Mudryk joins Spurs on loan from Chelsea after doping ban ends
-
Chelsea 'keeper Sanchez makes deadline-day move to Como
-
What's next after judge strikes down New York's landmark climate fund law
-
Kroenke to buy baseball's Los Angeles Angels: statement
-
Cobolli claws out five-set victory at rain-hit US Open
-
US defends Trump's Venezuela oil deal
-
Ndiaye seizes 'chance of a lifetime' in Man City move
-
OpenAI to launch new model with 'stronger safeguards' after hack
-
Global bond sell-off, surging oil prices send markets into the red
-
Fans mourn 'bittersweet' final Ariana Grande gig in London
-
Zuckerberg, Musk make plea at G20 for more AI data centers
-
Norway's King Haakon VIII swears allegiance to constitution
-
'Discomfort' at G20 finance talks as US hosts Russian minister
-
Syria site infrastructure 'consistent with nuclear reactor': IAEA
-
Man City set to sign Fernandez for record fee, Ndiaye on deadline day
-
Juventus sign Woltemade, Sarr on loan from Premier League
-
Back-to-school in Cuba a grim day with shortages of everything
Inner workings of AI an enigma - even to its creators
Even the greatest human minds building generative artificial intelligence that is poised to change the world admit they do not comprehend how digital minds think.
"People outside the field are often surprised and alarmed to learn that we do not understand how our own AI creations work," Anthropic co-founder Dario Amodei wrote in an essay posted online in April.
"This lack of understanding is essentially unprecedented in the history of technology."
Unlike traditional software programs that follow pre-ordained paths of logic dictated by programmers, generative AI (gen AI) models are trained to find their own way to success once prompted.
In a recent podcast Chris Olah, who was part of ChatGPT-maker OpenAI before joining Anthropic, described gen AI as "scaffolding" on which circuits grow.
Olah is considered an authority in so-called mechanistic interpretability, a method of reverse engineering AI models to figure out how they work.
This science, born about a decade ago, seeks to determine exactly how AI gets from a query to an answer.
"Grasping the entirety of a large language model is an incredibly ambitious task," said Neel Nanda, a senior research scientist at the Google DeepMind AI lab.
It was "somewhat analogous to trying to fully understand the human brain," Nanda added to AFP, noting neuroscientists have yet to succeed on that front.
Delving into digital minds to understand their inner workings has gone from a little-known field just a few years ago to being a hot area of academic study.
"Students are very much attracted to it because they perceive the impact that it can have," said Boston University computer science professor Mark Crovella.
The area of study is also gaining traction due to its potential to make gen AI even more powerful, and because peering into digital brains can be intellectually exciting, the professor added.
- Keeping AI honest -
Mechanistic interpretability involves studying not just results served up by gen AI but scrutinizing calculations performed while the technology mulls queries, according to Crovella.
"You could look into the model...observe the computations that are being performed and try to understand those," the professor explained.
Startup Goodfire uses AI software capable of representing data in the form of reasoning steps to better understand gen AI processing and correct errors.
The tool is also intended to prevent gen AI models from being used maliciously or from deciding on their own to deceive humans about what they are up to.
"It does feel like a race against time to get there before we implement extremely intelligent AI models into the world with no understanding of how they work," said Goodfire chief executive Eric Ho.
In his essay, Amodei said recent progress has made him optimistic that the key to fully deciphering AI will be found within two years.
"I agree that by 2027, we could have interpretability that reliably detects model biases and harmful intentions," said Auburn University associate professor Anh Nguyen.
According to Boston University's Crovella, researchers can already access representations of every digital neuron in AI brains.
"Unlike the human brain, we actually have the equivalent of every neuron instrumented inside these models", the academic said. "Everything that happens inside the model is fully known to us. It's a question of discovering the right way to interrogate that."
Harnessing the inner workings of gen AI minds could clear the way for its adoption in areas where tiny errors can have dramatic consequences, like national security, Amodei said.
For Nanda, better understanding what gen AI is doing could also catapult human discoveries, much like DeepMind's chess-playing AI, AlphaZero, revealed entirely new chess moves that none of the grand masters had ever thought about.
Properly understood, a gen AI model with a stamp of reliability would grab competitive advantage in the market.
Such a breakthrough by a US company would also be a win for the nation in its technology rivalry with China.
"Powerful AI will shape humanity's destiny," Amodei wrote.
"We deserve to understand our own creations before they radically transform our economy, our lives, and our future."
P.A.Mendoza--AT