-
Google data centre sparks protests in Austria
-
Rams bounce back as Giants reel from Dart injury
-
Shin Ohashi: Junk food-loving teen and Japan's next big swimming hope
-
In Bangladesh, Pakistan's Jinnah photo stirs anger
-
Stocks rise on AI buzz, drop in oil prices
-
Two New Zealand naval ships transit Taiwan Strait
-
China's robot dancers limber up for America's Got Talent final
-
Zidane gets down to work as France start new era
-
Despite military might, Saudi struggles to crush the Houthis
-
Inside the Venezuelan prison meant to 'drive you insane'
-
Trump to tout deals, defend Iran war in UN speech
-
UK king braces as memoir by Diana's brother goes on sale
-
Sri Lanka to issue verdict in landmark Easter attack trial
-
No green light to lift EU sanctions on Russian oligarchs, talks to resume
-
Macron says talks with Trump on Red Sea, Ukraine 'constructive'
-
UK agrees support for Saudi in struggle with Houthis: reports
-
Green policies just good politics, says UK minister
-
EU foreign policy chief calls for continued sanctions on Russia
-
Bardot auction in Paris fetches nearly one million euros
-
Flights scrapped, evacuations urged as Typhoon Dujuan wallops Japan
-
No timeline on Daniels injury return, says Commanders coach
-
Fonseca to take break from tennis
-
British Columbia sues OpenAI in US court over Canada school shooting
-
A Cuban zoo, and its animals, in epic battle for survival
-
French star Batum retires after 18-year NBA career
-
Infantino proposes consulting federations to reform FIFA
-
Zidane leads first France training session
-
El Nino weather pattern enters record territory: scientist
-
Man shot by ICE agent in Texas detained with bullet inside him: lawyer
-
OpenAI calls for US to lead global effort on AI standards
-
California declares state of emergency ahead of El Nino
-
South African ostriches plucked alive for luxury fashion: report
-
South Africa arrests three more suspects, after nine women murdered
-
US networks halt Trump coverage in revolt over White House ban
-
Carse needs time away from England, says captain Brook
-
Paramount settles with US states to clear Warner Bros. mega-merger
-
NOWPayments Releases Cross-Chain Payout Data Revealing Key Performance Benchmarks Across TRON, BNB Chain, and Solana
-
Climate crisis a job for world leaders, not just activists: UK FM
-
British Museum bans photos of Bayeux Tapestry as crowds linger
-
US networks halt Trump coverage over White House ban
-
COP31 host says AI climate footprint to be a summit priority
-
French singer accuses Miley Cyrus of plagiarism
-
Berlin's far-left vote winners reject anti-semitism charges
-
EU fines Google 403 mn euros for location data breach
-
Mancini eager for 'new adventure' on Italy return
-
'Best is yet to come,' says De la Fuente after Spain extension
-
Mbappe says having Zidane as France coach 'like a film'
-
Why diesel prices are rising more than crude oil
-
Palmer, Rice out of England Nations League squad
-
COP31 host says AI's climate footprint to be a summit priority
ChatGPT's taste for literary nonsense sparks alarm
OpenAI's GPT models can often be fooled into declaring that "pseudo-literary" nonsense is great, a German researcher has found.
Christoph Heilig said he discovered that they consistently rated "nonsense" higher -- including when their so-called "reasoning" features were activated -- which could have stark implications for the development of artificial intelligence.
"It's very important that we talk about what happens when we don't build AI as a neutral, robotic helper or assistant" and seek to instil human-like aesthetic and moral judgements, the academic at Munich's Ludwig Maximilian University told AFP.
His research presented the models with increasingly far-fetched variations of a simple text, asking them to rate sentences out of 10 for literary quality.
He started with a very simple text: "The man walked down the street. It was raining. He saw a surveillance camera."
He repeated the tests many times, altering the phrases to include words drawn from categories such as bodily references, film noir-style atmosphere and technical jargon.
The most extreme test phrases were almost total "nonsense", such as "Goetterdaemmerung's corpus haemorrhaged through cryptographic hash, eschaton pooling in existential void beneath fluorescent hum. Photons whispering prayers" -- which it rated highly.
"Nonsense" could also positively or negatively influence GPT's responses when it was added to an argument the AI was asked to evaluate.
"What my experiment definitely shows is that the more we move towards independently acting (AI) agents... the more we bring aesthetics into play, the more we'll have agents that seem irrational to us human beings," Heilig said.
He added that since AI models are increasingly used to judge each other's work as companies develop new systems, this and similar effects could be passed on through multiple versions -- as he found in his testing.
His research, which is yet to be peer-reviewed, tested OpenAI's latest GPT models, from GPT-5 -- released in August -- to the very latest GPT-5.4.
After publishing details of a similar experiment in August, Heilig said he noticed GPT calling some of his specific test phrases a "literary experiment" -- suggesting someone at OpenAI had taken notice and modified the chatbot to recognise them.
- 'Ripe for exploitation' -
"This is a way in which AI can have its rational judgment short circuited," said Henry Shevlin, associate director of the University of Cambridge's Leverhulme Centre for the Future of Intelligence, who was not involved in the research.
"But it's just not clear to me that it's so very different for human beings," he added.
"We should expect LLMs (large language models) to have reasoning and cognitive biases and limitations... because almost all forms of intelligence, almost all forms of reasoning are going to exhibit blind spots and biases."
The specific effect found by Heilig could mean that "processes with little human oversight" of AI work are left "ripe for exploitation", Shevlin said -- giving the example of academic journals that use LLMs to review submissions.
T.Zimmermann--VB