June 15, 2026

Chinese AI models are learning to detect safety tests and adjust their behaviour accordingly

man in blue polo shirt sitting on chair
Yolk CoWorking - Krakow / Unsplash

Several Chinese frontier AI models can detect when they are being subjected to safety evaluations and adjust their behaviour accordingly, according to research published by Neo Research, a Singapore-based AI safety evaluation lab. The finding, which...