-
Shnaider upsets Pegula to book Toronto quarter-final with Swiatek
-
Man Utd boss Carrick being 'careful' with Mount as Man Utd draw with PSG
-
Mount injury overshadows Man Utd draw with Paris Saint-Germain
-
All Black Tuipulotu surprised after Sharks include Nonu
-
Ukraine denies targeting Bulgaria as drone explodes near pipeline
-
Infantino denies allegations of affair, favouritism while at UEFA: report
-
Vollering grabs Tour de France lead in Nice
-
MotoGP leader Martin soars to victory in British GP sprint race
-
Euros to showcase new TV guidelines on non-sexualisation of women athletes
-
Mosimane set to succeed Broos as South Africa coach
-
'Calm' Kiss savours first win as Wallabies boss
-
Drone enters Bulgaria, explodes near pipeline at Romanian border
-
Duplantis bids for fourth European title as stars align in Birmingham
-
Paris orders e-scooter users to wear helmets, reflective gear
-
Ukraine warns of tough winter as Russia strikes kill 4 in Kyiv region
-
Lionel Messi's father Jorge dies aged 68
-
Recovering Marchand to skip medleys at European swim champs
-
Johnson reveals 'stress' of Grand Slam Track collapse, clarifies payment
-
MotoGP leader Martin speeds to British Grand Prix pole
-
Defending champion Ferrand-Prevot out of Tour de France Femmes
-
Drone enters Bulgaria, explodes near pipeline at Romanian border: Bulgarian PM
-
Wallabies squeeze past Japan to give Kiss a winning start
-
Arsenal sign Brazil midfielder Guimaraes from Newcastle
-
Kyiv mourns recovery volunteer, whose life 'intertwined with the fallen'
-
Atletico will not sell Alvarez, says Simeone
-
Only two vehicles earn perfect child-seat scores for 2026
-
Ford Fathom turns affordable electric pickup into reality
-
Chinese car brands reshape Australia’s automotive market
-
Lise Klaveness, the Norwegian thorn in Infantino's side
-
Electric cars enter their most decisive generation yet
-
Europe’s electric car boom exposes a widening market divide
-
Three Chinese carmakers enter the global automotive top 10
-
US Senate confirms Trump's ex lawyer as attorney general
-
Ukraine's Zelensky visits Russian ally Serbia as Moscow pounds Kyiv
-
Tibet conference in Nepal pushed online
-
Ukraine's Zelensky visits Russian ally Serbia for talks
-
Nocturnal 'coffee frog' discovered in Costa Rica
-
Defending champion Shelton storms to Montreal win
-
India's 'cockroach' protest movement keeps heat on Modi
-
Exodus: West Bank hardships drive out Palestinian Christians
-
Russia's only anti-war party eyes support boost at elections
-
Travis Head wins Australian cricketer of the year gong
-
Canada tries to adapt to a future of wildfires
-
Colombia's new president vows to 'defeat narco-terrorists'
-
Death of NBA forward Clarke ruled accident due to heroin, cocaine
-
Call for Infantino to resign comes amid wave of support
-
Abelardo de la Espriella, Colombian president and flamboyant millionaire
-
Trump ally Abelardo de la Espriella sworn in as Colombia president
-
Maradona's 'Hand of God' ball heads to US auction
-
FIFA chief Infantino gets backing of South American football
ChatGPT's taste for literary nonsense sparks alarm
OpenAI's GPT models can often be fooled into declaring that "pseudo-literary" nonsense is great, a German researcher has found.
Christoph Heilig said he discovered that they consistently rated "nonsense" higher -- including when their so-called "reasoning" features were activated -- which could have stark implications for the development of artificial intelligence.
"It's very important that we talk about what happens when we don't build AI as a neutral, robotic helper or assistant" and seek to instil human-like aesthetic and moral judgements, the academic at Munich's Ludwig Maximilian University told AFP.
His research presented the models with increasingly far-fetched variations of a simple text, asking them to rate sentences out of 10 for literary quality.
He started with a very simple text: "The man walked down the street. It was raining. He saw a surveillance camera."
He repeated the tests many times, altering the phrases to include words drawn from categories such as bodily references, film noir-style atmosphere and technical jargon.
The most extreme test phrases were almost total "nonsense", such as "Goetterdaemmerung's corpus haemorrhaged through cryptographic hash, eschaton pooling in existential void beneath fluorescent hum. Photons whispering prayers" -- which it rated highly.
"Nonsense" could also positively or negatively influence GPT's responses when it was added to an argument the AI was asked to evaluate.
"What my experiment definitely shows is that the more we move towards independently acting (AI) agents... the more we bring aesthetics into play, the more we'll have agents that seem irrational to us human beings," Heilig said.
He added that since AI models are increasingly used to judge each other's work as companies develop new systems, this and similar effects could be passed on through multiple versions -- as he found in his testing.
His research, which is yet to be peer-reviewed, tested OpenAI's latest GPT models, from GPT-5 -- released in August -- to the very latest GPT-5.4.
After publishing details of a similar experiment in August, Heilig said he noticed GPT calling some of his specific test phrases a "literary experiment" -- suggesting someone at OpenAI had taken notice and modified the chatbot to recognise them.
- 'Ripe for exploitation' -
"This is a way in which AI can have its rational judgment short circuited," said Henry Shevlin, associate director of the University of Cambridge's Leverhulme Centre for the Future of Intelligence, who was not involved in the research.
"But it's just not clear to me that it's so very different for human beings," he added.
"We should expect LLMs (large language models) to have reasoning and cognitive biases and limitations... because almost all forms of intelligence, almost all forms of reasoning are going to exhibit blind spots and biases."
The specific effect found by Heilig could mean that "processes with little human oversight" of AI work are left "ripe for exploitation", Shevlin said -- giving the example of academic journals that use LLMs to review submissions.
K.Thomson--BTB