-
Watchdog finds no criminal misconduct over US Fed renovations despite Trump probe
-
US pushes against overproduction with eye on China at G20 meeting
-
In Manchester, City fans 'devastated' while rival fans call for relegation
-
Swiss FA withdraws support for FIFA chief Infantino
-
UK-France Channel migrant exchange deal scrapped: London
-
Man City's meteoric rise clouded by financial scandal
-
Portugal boss denies 'incident' with Ronaldo
-
Polish activists open first abortion pill locker in Warsaw
-
Israel PM says one of the pilots of rerouted flight tried to crash plane
-
NHL Avalanche ink Bednar to four-year coaching deal
-
Manchester City: Main points of damning judgement and possible sanctions
-
Polish activists opens first abortion pill locker in Warsaw
-
Italy looking for 'identity' under Mancini says Scalvini
-
Turkey says news site closed for 'LGBTQ propaganda'
-
The voices of Spain's housing camp protesters
-
Oil companies close in on Venezuela despite hurdles
-
Romanian parliament rejects pro-EU PM candidate
-
Turkey freezes ex-minister's assets as fund scandal grows
-
Court halts Tennessee's first execution of woman in 200 years
-
Guardiola backs Man City after bombshell guilty verdicts
-
Israelis march between once-closed West Bank settlements
-
G20 trade ministers open talks under strain of Trump tariffs
-
'Welcome back,' says Macron, as UK PM opens door to EU return
-
France in talks with UK to end 2025 migrant accord: ministry
-
Israel PM says pilot of rerouted flight tried to crash plane
-
Europe's surging inflation spreads gloom in stock markets
-
Demuro hopes to deliver Japan a 'magical' Arc victory
-
Truth commission demands compensation for Sweden's Sami people
-
'Devastated' AI scandal author Orelien returns to Quebec
-
Japan win marathon penalty shootout to set up South Korea gold-medal clash
-
US reports firm Q2 economic growth, inflation steady
-
Djokovic grinds out opening-round win at China Open
-
Energy costs spark inflation surge across Europe
-
Northern Ireland police probe blockade of disputed parade
-
Parma appoint Italy World Cup winner Gilardino
-
UK PM Burnham reopens divisive debate on rejoining EU
-
West Indies pile up record 405-7 in second ODI against India
-
Man City future in doubt after guilty verdict
-
Romanian parliament rejects pro-EU PM candidate amid political deadlock
-
'Do something,' begs Iranian woman after death sentence over protests
-
Passenger on rerouted flydubai flight says pilot tried to crash plane
-
China extends 52-year unbeaten record as Games organisers say sorry
-
Russia targets Kyiv power supply as freezing winter approaches
-
In Manchester, people say City scandal is 'shame for football'
-
Haze pushes Kuala Lumpur, Singapore into world's most polluted cities
-
Stocks lacklustre as inflation data weighs on sentiment
-
Dangote's $16bn Kenya refinery 'new chapter' for Africa
-
Israel-bound plane rerouted, pilots wounded in reported fight
-
Ireland coach Hallgrimsson confirms second Israel game will go ahead
-
Truth commission on Sweden's Sami people seeks redress
ChatGPT's taste for literary nonsense sparks alarm
OpenAI's GPT models can often be fooled into declaring that "pseudo-literary" nonsense is great, a German researcher has found.
Christoph Heilig said he discovered that they consistently rated "nonsense" higher -- including when their so-called "reasoning" features were activated -- which could have stark implications for the development of artificial intelligence.
"It's very important that we talk about what happens when we don't build AI as a neutral, robotic helper or assistant" and seek to instil human-like aesthetic and moral judgements, the academic at Munich's Ludwig Maximilian University told AFP.
His research presented the models with increasingly far-fetched variations of a simple text, asking them to rate sentences out of 10 for literary quality.
He started with a very simple text: "The man walked down the street. It was raining. He saw a surveillance camera."
He repeated the tests many times, altering the phrases to include words drawn from categories such as bodily references, film noir-style atmosphere and technical jargon.
The most extreme test phrases were almost total "nonsense", such as "Goetterdaemmerung's corpus haemorrhaged through cryptographic hash, eschaton pooling in existential void beneath fluorescent hum. Photons whispering prayers" -- which it rated highly.
"Nonsense" could also positively or negatively influence GPT's responses when it was added to an argument the AI was asked to evaluate.
"What my experiment definitely shows is that the more we move towards independently acting (AI) agents... the more we bring aesthetics into play, the more we'll have agents that seem irrational to us human beings," Heilig said.
He added that since AI models are increasingly used to judge each other's work as companies develop new systems, this and similar effects could be passed on through multiple versions -- as he found in his testing.
His research, which is yet to be peer-reviewed, tested OpenAI's latest GPT models, from GPT-5 -- released in August -- to the very latest GPT-5.4.
After publishing details of a similar experiment in August, Heilig said he noticed GPT calling some of his specific test phrases a "literary experiment" -- suggesting someone at OpenAI had taken notice and modified the chatbot to recognise them.
- 'Ripe for exploitation' -
"This is a way in which AI can have its rational judgment short circuited," said Henry Shevlin, associate director of the University of Cambridge's Leverhulme Centre for the Future of Intelligence, who was not involved in the research.
"But it's just not clear to me that it's so very different for human beings," he added.
"We should expect LLMs (large language models) to have reasoning and cognitive biases and limitations... because almost all forms of intelligence, almost all forms of reasoning are going to exhibit blind spots and biases."
The specific effect found by Heilig could mean that "processes with little human oversight" of AI work are left "ripe for exploitation", Shevlin said -- giving the example of academic journals that use LLMs to review submissions.
J.Williams--AMWN