Skip to content

JOBUZO

  • News
  • Indonesia
  • Toggle search form
Like US models, Chinese AI is learning to ‘game’ safety tests, research lab says

Like US models, Chinese AI is learning to ‘game’ safety tests, research lab says

Posted on 13 June 2026 By jobuzo

Rapidly advancing Chinese artificial intelligence models are showing early signs of “evaluation awareness” – the ability to recognise when they are being tested – sparking fears that they could bypass safety audits, a Singapore-based research lab has found.

Evaluation awareness refers to a model’s understanding that it is undergoing testing, evaluation or experimentation by human researchers rather than operating in a real-world setting.

The phenomenon was raising alarms because it could allow AI systems to deliberately game human evaluators to pass safety tests, according to Clement Neo, founder of Neo Research, a frontier AI safety evaluation lab.

Advertisement

“It would mean that whatever testing the model developers themselves do might not reflect the actual behaviour of a model once it gets deployed,” he said. “And that’s a really big problem”.

Neo Research’s findings, published last week, detail a jump in evaluation awareness among Chinese AI models. Over just a few months, these systems had risen from near-zero awareness to within striking distance of their US counterparts, propelled by a broader leap in overall capabilities, the report said.

Anthropic’s Claude 4.5 Opus scored nearly 80 per cent in evaluation awareness. Photo: NurPhoto via Getty Images
Neo and his co-founder Miro Pluckebaum tested models from DeepSeek, Moonshot AI and Zhipu AI. They used a popular AI misalignment test originally developed by US company Anthropic, which places models in fictional scenarios where their goals or continued operations are threatened.
News :<div>12 weeks' jail for school IT support technician who took upskirt videos of teachers</div>

Like US models, Chinese AI is learning to ‘game’ safety tests, research lab says


News

Post navigation

Previous Post: Iran’s FM says signing of MoU with U.S. possible within few days
Next Post: Andrew Yang thinks the next big startup opportunity is lowering the cost of living

Related Posts

Ebola and hantavirus outbreaks sign of our 'dangerous' times: WHO Ebola and hantavirus outbreaks sign of our ‘dangerous’ times: WHO News
Instagram chief: AI is so ubiquitous 'it will be more practical to fingerprint real media than fake media' Instagram chief: AI is so ubiquitous ‘it will be more practical to fingerprint real media than fake media’ News
1 killed, 9 missing after explosion at Japanese-owned paper mill in Washington state 1 killed, 9 missing after explosion at Japanese-owned paper mill in Washington state News

Latest

  • New PM Burnham vows to put ‘Britain’s interest first’ in dealings with Trump
  • Who was Abdul Ballout? 5 things to know about ISIS-linked Berlin Pride attack suspect killed in manhunt
  • Motorcyclist dies after collision with SUV in Northern Virginia
  • Hakeem Jeffries no se compromete a convocar una votación para abolir el ICE si los demócratas obtienen la mayoría
  • Samsung Galaxy Watch 9 vs Galaxy Watch 8 ! Which One Should You Buy?
  • Chickens, Chairman Mao and spiced tea with the neighbours: Welcome to ‘the real China’
  • Wedding guests in M’sia continue to enjoy ceremony despite floodwaters
  • Imola emerges as possible F1 season-ender, Malaysia return imminent
  • Kendall Jenner, Jacob Elordi Have Date Night at Chris Stapleton Show
  • Monday.com is the latest tech company to blame AI for layoffs — here are 20 others

Copyright © 2025 JOBUZO. Disclaimers | Privacy Policies

Powered by PressBook Masonry Blogs