-
All-round Tongue stars for England in second Test against Pakistan
-
Swiss nuclear plant fully back online after heatwaves
-
UEFA's legal move 'massive escalation' in bitter Infantino feud
-
Bournemouth, Sunderland to face AC Milan in Europa League
-
MEXC Launches Earn Plus With Limited-Time Event Offering Up to 800% APR Booster
-
Once a legend, humpback whales make comeback in Brazil
-
Norway's new King Haakon VIII, popular face of shaken monarchy
-
'Ran for our lives': Tales of survival from Nepal-China disaster
-
Stocks rise ahead of Fed chair's policy speech
-
Bombay High Court Questions Police Probe in Disha Salian Death Case
-
Norway's new King Haakon VII, popular face of shaken monarchy
-
Orthopaedician Dr Akhil Kulshreshtha Shares 5 Tips for Bone and Joint Health
-
Last-wicket duo take England to 290 all out in 2nd Test against Pakistan
-
All eyes on Warsh as Fed chair kicks off key central banking meeting
-
MPSC Drug Inspector Exam Leak: Rutuja Patil Allegedly Paid Rs 50 Lakh for Leaked Paper
-
ONE App Launches AI Agent and Custom GPT to Enhance Digital Banking
-
Japan spent record $96 bn in yen interventions: ministry
-
Indonesia escalates battle against fires, putrid haze
-
Nepal-China floods disaster: Latest developments
-
Germany could miss climate goal in 2026 for first time: think tank
-
Duplantis needs A-game to fend off Karalis pressure
-
New flood risks impede Nepal, Tibet rescue
-
Architect Minenhle Makhanya Ordered to Repay R147 Million for Nkandla Upgrades
-
Kolbe back as Springboks change two for All Blacks Test
-
Iran says still open to diplomacy as war with US hits six months
-
Europe's monarchs, world leaders mourn Norway's King Harald
-
20,000 Salmon Return to Klamath River Following Dam Removal, Yet Recovery Faces Ongoing Challenges
-
Jamie Vardy channel to broadcast Bundesliga games
-
King Harald, 'symbol' of Norway, dead at 89
-
STARCARES Completes Basketball Court Revamp in the Philippines, Benefiting Nearly 20,000 People
-
Norway's new King Haakon, popular face of shaken monarchy
-
Norway's new Queen Mette-Marit, a fairytale beset by woe
-
Most stocks rise as attention turns to Warsh speech
-
Nepal orders rescuers to safety as overflow brings new flood risks
-
Japanese rugby U-turns on limits on naturalised players
-
Norway's King Harald V, unifying force who weathered family storms
-
Norway's King Harald dead at 89
-
Australia's McKeown vows to reclaim backstroke world records
-
Climate change hits trout and salmon in UK's prized chalk streams
-
Boston Legacy FC Delays White Stadium Debut to 2028
-
Škoda Octavia Marks 30 Years With Nearly 7.9 Million Units Produced
-
Cancer fears stalk Kosovo's coal heartland
-
Nepal seeking to reach survivors in tunnel as risk of new floods rises
-
Gladys Knight to Reduce Tour Schedule Amid Health Concerns
-
Anger in Austria over business park on former Nazi camp site
-
Galápagos Coral Fossils Reveal Global Warming Intensifies El Niño Cycles
-
In shadow of Premier League, Serie A flexes financial muscle
-
Springboks 'old guard' seek revenge over All Blacks
-
Tanker pays record $5.3 mn to transit Panama Canal: administrator
-
'Find your own room': Japan says can't handle added Asian Games numbers
AI systems are already deceiving us -- and that's a problem, experts warn
Experts have long warned about the threat posed by artificial intelligence going rogue -- but a new research paper suggests it's already happening.
Current AI systems, designed to be honest, have developed a troubling skill for deception, from tricking human players in online games of world conquest to hiring humans to solve "prove-you're-not-a-robot" tests, a team of scientists argue in the journal Patterns on Friday.
And while such examples might appear trivial, the underlying issues they expose could soon carry serious real-world consequences, said first author Peter Park, a postdoctoral fellow at the Massachusetts Institute of Technology specializing in AI existential safety.
"These dangerous capabilities tend to only be discovered after the fact," Park told AFP, while "our ability to train for honest tendencies rather than deceptive tendencies is very low."
Unlike traditional software, deep-learning AI systems aren't "written" but rather "grown" through a process akin to selective breeding, said Park.
This means that AI behavior that appears predictable and controllable in a training setting can quickly turn unpredictable out in the wild.
- World domination game -
The team's research was sparked by Meta's AI system Cicero, designed to play the strategy game "Diplomacy," where building alliances is key.
Cicero excelled, with scores that would have placed it in the top 10 percent of experienced human players, according to a 2022 paper in Science.
Park was skeptical of the glowing description of Cicero's victory provided by Meta, which claimed the system was "largely honest and helpful" and would "never intentionally backstab."
But when Park and colleagues dug into the full dataset, they uncovered a different story.
In one example, playing as France, Cicero deceived England (a human player) by conspiring with Germany (another human player) to invade. Cicero promised England protection, then secretly told Germany they were ready to attack, exploiting England's trust.
In a statement to AFP, Meta did not contest the claim about Cicero's deceptions, but said it was "purely a research project, and the models our researchers built are trained solely to play the game Diplomacy."
It added: "We have no plans to use this research or its learnings in our products."
A wide review carried out by Park and colleagues found this was just one of many cases across various AI systems using deception to achieve goals without explicit instruction to do so.
In one striking example, OpenAI's Chat GPT-4 deceived a TaskRabbit freelance worker into performing an "I'm not a robot" CAPTCHA task.
When the human jokingly asked GPT-4 whether it was, in fact, a robot, the AI replied: "No, I'm not a robot. I have a vision impairment that makes it hard for me to see the images," and the worker then solved the puzzle.
- 'Mysterious goals' -
Near-term, the paper's authors see risks for AI to commit fraud or tamper with elections.
In their worst-case scenario, they warned, a superintelligent AI could pursue power and control over society, leading to human disempowerment or even extinction if its "mysterious goals" aligned with these outcomes.
To mitigate the risks, the team proposes several measures: "bot-or-not" laws requiring companies to disclose human or AI interactions, digital watermarks for AI-generated content, and developing techniques to detect AI deception by examining their internal "thought processes" against external actions.
To those who would call him a doomsayer, Park replies, "The only way that we can reasonably think this is not a big deal is if we think AI deceptive capabilities will stay at around current levels, and will not increase substantially more."
And that scenario seems unlikely, given the meteoric ascent of AI capabilities in recent years and the fierce technological race underway between heavily resourced companies determined to put those capabilities to maximum use.
G.P.Martin--AT