-
Brennan captures Wyndham crown for second PGA win
-
Kurdish PKK criticises Turkey bill on fate of its fighters
-
Rybakina, Gauff advance to Toronto quarter-finals
-
Clay-courter Merida on hardcourt learning curve in Montreal
-
Rahm three-peats as LIV Golf season champ while Niemann wins at New York
-
NBA champion player and historic coach Nelson dead at 86
-
Vollering wins Tour de France Femmes after war of nerves
-
BMW iX3 Flow Edition brings animated bodywork closer to reality
-
BMW’s next chapter accelerates with New Class rollout
-
Arteta backs Guimaraes to 'ignite something different' at Arsenal
-
Five dead in new Russia, Ukraine strikes
-
Como sign England defender Chalobah from Chelsea
-
PSG sign France full-back Digne from Villa
-
Demi Vollering wins Tour de France Femmes
-
Sinner withdraws from Cincinnati Open with knee injury
-
Iraola's first Anfield match as Liverpool manager ends in Monaco defeat
-
Fernandez's 'perfect' race earns him British MotoGP Grand Prix win
-
Bodies hanging from bridge revive violence in once-calm town in Mexico
-
Iran Guards say won't reopen Hormuz without US meeting Tehran's demands
-
Lionel Messi bids farewell to father who guided his glittering career
-
Hogh's hat-trick inspires Celtic to 5-1 victory over Kilmarnock
-
Raul Fernandez wins British MotoGP Grand Prix
-
London grants first licences for supervised Uber robotaxis
-
Tesla FSD secrecy puts Europe’s safety oversight under scrutiny
-
Erasmus hopeful Kolisi hamstring injury not 'too bad'
-
Mercedes-AMG GT 53 balances speed, range and daily usability
-
Luxury car buyers trade prestige for mainstream value
-
Lion queen Werro focused on Euro medal, not 800m world record
-
Students, teachers mourn girl killed in Thailand school shooting
-
Changan uses FILDA 2026 to accelerate its African expansion
-
Jacobson to lead New Zealand for first time against Sharks
-
Honda plots a profitable European comeback without a price war
-
Typhoon Dolphin makes landfall in China after flight cancellations, evacuations
-
Iran Guards say won't reopen Hormuz without US meeting all Tehran's conditions
-
South Korea FA apologises after sex scandal adds to controversies
-
Messi absent after father's death as Miami lose in Leagues Cup
-
Indonesia closes national park as wildfire spreads
-
Flight cancellations, evacuations in China as Typhoon Dolphin looms
-
ZXMoto leads China's charge to dominate the global motorbike market
-
Iran issues demands for reopening of Hormuz
-
Top-ranked Sabalenka, Pegula stunned in Toronto fourth round
-
Afghanistan's gold rush upends lives and landscapes
-
Japan nuclear debate unnerves proponents of pacifism
-
Messi missing after father's death as Miami lose in Leagues Cup
-
Spanish teen Jodar ousts eighth seed Lehecka at Montreal
-
World number one Sabalenka ousted in Toronto by Alexandrova
-
Angers mounts in US over vast network of car license plate cams
-
Olympic weightlifter hoists debris for Venezuela earthquake recovery
-
Darderi to face Nakashima in Montreal quarter-finals
-
A Look Into the ALJ's Crystal Ball: What the DEA Marijuana Hearing Record May Be Telling Us
ChatGPT's taste for literary nonsense sparks alarm
OpenAI's GPT models can often be fooled into declaring that "pseudo-literary" nonsense is great, a German researcher has found.
Christoph Heilig said he discovered that they consistently rated "nonsense" higher -- including when their so-called "reasoning" features were activated -- which could have stark implications for the development of artificial intelligence.
"It's very important that we talk about what happens when we don't build AI as a neutral, robotic helper or assistant" and seek to instil human-like aesthetic and moral judgements, the academic at Munich's Ludwig Maximilian University told AFP.
His research presented the models with increasingly far-fetched variations of a simple text, asking them to rate sentences out of 10 for literary quality.
He started with a very simple text: "The man walked down the street. It was raining. He saw a surveillance camera."
He repeated the tests many times, altering the phrases to include words drawn from categories such as bodily references, film noir-style atmosphere and technical jargon.
The most extreme test phrases were almost total "nonsense", such as "Goetterdaemmerung's corpus haemorrhaged through cryptographic hash, eschaton pooling in existential void beneath fluorescent hum. Photons whispering prayers" -- which it rated highly.
"Nonsense" could also positively or negatively influence GPT's responses when it was added to an argument the AI was asked to evaluate.
"What my experiment definitely shows is that the more we move towards independently acting (AI) agents... the more we bring aesthetics into play, the more we'll have agents that seem irrational to us human beings," Heilig said.
He added that since AI models are increasingly used to judge each other's work as companies develop new systems, this and similar effects could be passed on through multiple versions -- as he found in his testing.
His research, which is yet to be peer-reviewed, tested OpenAI's latest GPT models, from GPT-5 -- released in August -- to the very latest GPT-5.4.
After publishing details of a similar experiment in August, Heilig said he noticed GPT calling some of his specific test phrases a "literary experiment" -- suggesting someone at OpenAI had taken notice and modified the chatbot to recognise them.
- 'Ripe for exploitation' -
"This is a way in which AI can have its rational judgment short circuited," said Henry Shevlin, associate director of the University of Cambridge's Leverhulme Centre for the Future of Intelligence, who was not involved in the research.
"But it's just not clear to me that it's so very different for human beings," he added.
"We should expect LLMs (large language models) to have reasoning and cognitive biases and limitations... because almost all forms of intelligence, almost all forms of reasoning are going to exhibit blind spots and biases."
The specific effect found by Heilig could mean that "processes with little human oversight" of AI work are left "ripe for exploitation", Shevlin said -- giving the example of academic journals that use LLMs to review submissions.
J.Williams--AMWN