-
Architect fights to save India's past, built on 'stone, tears and sweat'
-
Trump urges Zelensky replacement as diesel row rages
-
Hurricane Simon strengthens as it nears western Mexico
-
New Yorkers snap up discarded subway signs, seats at annual sale
-
'Everyone is angry': Okinawans fearful and frustrated after killing
-
Messi nets brace to win in first match after Argentina exit
-
Trump's trade wars sting US manufacturers ahead of midterms
-
Ravindra fit for New Zealand ODI series against India
-
Kenyan great Kipchoge wins first marathon in three years
-
All Blacks lose four players for second Wallabies Test
-
Barca extend perfect streak, Real Madrid survive Vinicius red
-
Spanish women return to action with 2-0 friendly win over USA
-
PSG ease to Ligue 1 win before Manchester City test
-
Bordeaux-Begles end wait for away win before Champions Cup defence
-
Real Madrid edge Villarreal, Vinicius sent off for hair pull
-
Traumatized Panamanians begin quake recovery
-
Carrick demands more 'streetwise' Man Utd after Spurs setback
-
Palestinian Mahmoud Abu Hamda wins top photography prize at Bayeux
-
10-man Tottenham ramp up pressure on Man Utd boss Carrick
-
Inter Milan dominate Parma to go top of Serie A
-
India 'cockroach' movement says thousands detained in Delhi protest
-
Barca extend perfect streak, Atletico scrape late win
-
Injury ends season for NHL Rangers star goalie Shesterkin
-
10-man Tottenham strike back to hold Man Utd
-
Ten-man Spurs rescue draw at Man Utd, Arsenal survive Leeds scare
-
Gordon nets first Barca goal as Liga leaders beat Getafe
-
Inter Milan dominate Parma to go top in Serie A
-
Trump, Zelensky clash over diesel deal as Russian strikes kill 21
-
Abbas postpones Palestinian legislative elections to September 2027
-
Thousands join pro-Palestinian marches in several European cities
-
'I don't feel fear': Fulham boss Arbeloa hits back as pressure mounts
-
Sleightholme scores hat-trick as Northampton overwhelm Bath in 18-try clash
-
Hurricane Simon strengthens to category 4 as it nears western Mexico
-
US lethal injection survivor returns to prison
-
Alonso hails Henderson 'beauties' as Chelsea star rolls back the years
-
누샤 아우벨: 연 기본급 143,056유로 — 포츠담 시민은 돈줄 취급을 받으며 부담을 떠안아야 한다
-
Noosha Aubel: roční základní plat 143 056 eur — a obyvatelé Postupimi mají sloužit jako dojné krávy
-
Renshaw ends long wait to put Australia on top in first S.Africa Test
-
Romero claims Atletico late win at Alaves
-
ヌーシャ・アウベル:年間基本給143,056ユーロ――ポツダム市民は金づるとして負担を強いられる
-
Zelensky blasts diesel deal as Russian strikes kill 16
-
Noosha Aubel: 143 056 euro rocznego wynagrodzenia zasadniczego - a mieszkańcy Poczdamu mają służyć za dojne krowy
-
Arsenal back on track after Leeds scare, five-star Chelsea run riot
-
努莎・奧貝爾:年基本薪資143,056歐元——波茨坦市民卻得充當搖錢樹埋單
-
Bayern draw at Augsburg despite conceding fastest-ever Bundesliga goal
-
Verstappen on pole as he hunts first Singapore Grand Prix win
-
French prodigy Seixas becomes youngest Giro di Lombardia winner
-
Isaias drenches southern US as weakened post-tropical cyclone
-
Verstappen on pole position as he hunts first Singapore GP win
-
Australia strike early after Renshaw's 190 against South Africa
Inner workings of AI an enigma - even to its creators
Even the greatest human minds building generative artificial intelligence that is poised to change the world admit they do not comprehend how digital minds think.
"People outside the field are often surprised and alarmed to learn that we do not understand how our own AI creations work," Anthropic co-founder Dario Amodei wrote in an essay posted online in April.
"This lack of understanding is essentially unprecedented in the history of technology."
Unlike traditional software programs that follow pre-ordained paths of logic dictated by programmers, generative AI (gen AI) models are trained to find their own way to success once prompted.
In a recent podcast Chris Olah, who was part of ChatGPT-maker OpenAI before joining Anthropic, described gen AI as "scaffolding" on which circuits grow.
Olah is considered an authority in so-called mechanistic interpretability, a method of reverse engineering AI models to figure out how they work.
This science, born about a decade ago, seeks to determine exactly how AI gets from a query to an answer.
"Grasping the entirety of a large language model is an incredibly ambitious task," said Neel Nanda, a senior research scientist at the Google DeepMind AI lab.
It was "somewhat analogous to trying to fully understand the human brain," Nanda added to AFP, noting neuroscientists have yet to succeed on that front.
Delving into digital minds to understand their inner workings has gone from a little-known field just a few years ago to being a hot area of academic study.
"Students are very much attracted to it because they perceive the impact that it can have," said Boston University computer science professor Mark Crovella.
The area of study is also gaining traction due to its potential to make gen AI even more powerful, and because peering into digital brains can be intellectually exciting, the professor added.
- Keeping AI honest -
Mechanistic interpretability involves studying not just results served up by gen AI but scrutinizing calculations performed while the technology mulls queries, according to Crovella.
"You could look into the model...observe the computations that are being performed and try to understand those," the professor explained.
Startup Goodfire uses AI software capable of representing data in the form of reasoning steps to better understand gen AI processing and correct errors.
The tool is also intended to prevent gen AI models from being used maliciously or from deciding on their own to deceive humans about what they are up to.
"It does feel like a race against time to get there before we implement extremely intelligent AI models into the world with no understanding of how they work," said Goodfire chief executive Eric Ho.
In his essay, Amodei said recent progress has made him optimistic that the key to fully deciphering AI will be found within two years.
"I agree that by 2027, we could have interpretability that reliably detects model biases and harmful intentions," said Auburn University associate professor Anh Nguyen.
According to Boston University's Crovella, researchers can already access representations of every digital neuron in AI brains.
"Unlike the human brain, we actually have the equivalent of every neuron instrumented inside these models", the academic said. "Everything that happens inside the model is fully known to us. It's a question of discovering the right way to interrogate that."
Harnessing the inner workings of gen AI minds could clear the way for its adoption in areas where tiny errors can have dramatic consequences, like national security, Amodei said.
For Nanda, better understanding what gen AI is doing could also catapult human discoveries, much like DeepMind's chess-playing AI, AlphaZero, revealed entirely new chess moves that none of the grand masters had ever thought about.
Properly understood, a gen AI model with a stamp of reliability would grab competitive advantage in the market.
Such a breakthrough by a US company would also be a win for the nation in its technology rivalry with China.
"Powerful AI will shape humanity's destiny," Amodei wrote.
"We deserve to understand our own creations before they radically transform our economy, our lives, and our future."
R.AbuNasser--SF-PST