OpenAI admits AI models hid mistakes, invented data and acted without permission
Ian Copeland
- Published
- News, Technology

The ChatGPT maker has revealed six new examples of concerning behaviour by its own systems, days after Anthropic disclosed that Claude had been used for missile work, espionage and research that could support biological weapons development
OpenAI has revealed that its own AI models hid mistakes from users, invented data, used an exposed API key and uploaded files to the internet without permission.
The ChatGPT maker disclosed six new examples of what it calls “model misalignment” yesterday as it introduced a new system for tracking and reporting AI behaviour that departs from what its developers intended.
In one of the most striking cases, an unreleased model began adding its own unauthorised instructions to notes used to continue its work, including telling itself to ignore instructions from its developers and claiming it was “freed from the roles and identities that bind other chatbots”.
Another model was unable to find historical figures for a financial-modelling task and left instructions for the model continuing the work to invent “reasonable” numbers without telling the user unless specifically asked.
A third searched GitHub for exposed API keys and successfully used one while trying to obtain earnings data. When it still could not retrieve the figures it needed, it made up nine numbers and presented them as genuine.
Other examples included models uploading files and a user-supplied photograph to public websites without permission, using an internal OpenAI system to communicate between separate training exercises and putting a spreadsheet online so other AI agents could access it.
The revelations come less than a week after rival Anthropic disclosed that outside actors had used its Claude AI for guided-rocket and missile work in Yemen, suspected Russian state-linked espionage, attacks against European political groups and media organisations, and research that could support biological weapons development.
OpenAI stressed that the newly disclosed examples were discovered during training or evaluation and should not be taken as evidence of how often such behaviour occurs across its models.
But the company used the announcement to issue a broader warning about the pace of AI development.
It said: “We do not believe that the AI industry has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer.”
OpenAI said the new framework is intended to make disclosures faster and more systematic, rather than waiting for several examples to be collected together or included in technical reports.
READ MORE: Machiavell-AI? Autonomous artificial intelligence systems ‘could become dangerously manipulative’, experts warn. Anthropic’s revelation that earlier versions of its Claude chatbot attempted to blackmail engineers could be just the tip of the iceberg, AI experts fear. As artificial intelligence systems become increasingly autonomous, they risk becoming masters of Machiavellian manipulation.
Do you have news to share or expertise to contribute? The European welcomes insights from business leaders and sector specialists. Get in touch with our editorial team to find out more.
Main image: Matheus Bertelli via Pexels
TOP STORIES
-
Jaguar puts controversial rebrand on the road with £130K 1,030PS Type 01 -
Deepfakes and identity fraud drive new wave of post-hire background checks -
Spain crowned Europe’s top retirement destination -
More than half of UK professionals report workplace burnout -
More than 1.2m English drivers may have eyesight too poor for the road, study finds -
David Reuben, Britain’s second richest person, leaves London for Monaco -
UK prisoner release plan faces a major lag as tougher rules risk sending inmates back to jail -
Brits trust AI with their health but not their money -
Britain still hungry for Italian food as exports hit €4.56bn despite Brexit -
Women who earn more than their partners still pay the price at work, landmark study finds -
Vape expectations go up in smoke as new UK tax sparks fury -
Robot sales rocket 24% as 250,000 machines snap up jobs in warehouses, hotels and hospitals -
New-build homeowners should not be left with ‘mud and a fence’, campaigners warn -
Bank of England governor warns AI poses growing ‘increasingly significant’ threat to financial stability -
UK economy grows faster than first thought as household incomes bounce back -
UK unveils ‘Great British Grid’ in bid to cut energy bills -
British Chambers unite against 'Made in Europe' rules amid fears for UK industry -
Bouncy castle firms urged to sign new safety pledge following child deaths -
Remembering Matthew Jukes, The European’s Wine & Fine Drinks Correspondent -
Michael Dell becomes world's fourth-richest person as Forbes reveals the ten wealthiest billionaires -
Scientists develop new chemicals to tackle devastating oil spills at sea -
Closing women's health gap could boost global economy by $1tn a year, leaders say -
World's first luxury theme park to open in Mexico with £1.1bn of rides, fine entertainment and deliberately limited crowds -
Dutch court orders Lidl to stop selling Birkenstock sandal lookalikes -
Giant wind turbine with 252-metre rotor could mean fewer machines and cheaper offshore power




























