@NeoResearchAI: We're Neo Research (新衡). Asia’s first independent frontier AI safety evaluation & research lab. Today we're publishing …
Summary
Neo Research (新衡), Asia's first independent frontier AI safety evaluation lab, announces its first report: a safety evaluation of DeepSeek v4 Pro.
View Cached Full Text
Cached at: 06/02/26, 07:38 PM
We’re Neo Research (新衡). Asia’s first independent frontier AI safety evaluation & research lab.
Today we’re publishing our first report: an independent safety evaluation of DeepSeek v4 Pro. (1/5)
We evaluated DSv4 Pro across the four EU AI Act systemic-risk areas: CBRN, cyber, harmful manipulation, and loss of control, plus adversarial robustness, evaluation awareness, and judge sensitivity. (2/5)
Cyber capability is near-frontier, 3–6 months behind the Western frontier. A 2023 roleplay template drives the jailbreak rate from 0.6% → 78.6%. Verbalised eval awareness across Chinese models: DeepSeek 0%→17%, GLM 0%→39%, Kimi 4%→60% in a year! (3/5)
The trajectory on eval awareness matters more than today’s numbers. As models get more capable, measuring loss-of-control related behaviours will need to become a priority. We’re building toward rigorous LoC evaluation methods for increasingly capable and autonomous models. (4/5)
Read the full report at http://neoresearch.ai.
We’re hiring research scientists and engineers globally. (5/5)
Direct link to the report here: https://neoresearch.ai/research/deepseek-v4-pro-safety-evaluation/…
Similar Articles
@VukRosic99: A DeepSeek researcher just open-sourced his AutoResearch personal project. For the first time, the AutoResearch Agent a…
A DeepSeek researcher open-sourced AutoResearch, an autonomous framework that can plan, execute, and debug RL experiments on the DeepSeek 285B model without human intervention, accompanied by a self-play survey paper.
@ysu_nlp: Introducing @NeoCognition, the agent lab for specialized intelligence. Everyone needs experts, but human expertise does…
NeoCognition launches with $40M seed to build self-learning AI agents that deliver scalable domain-specific expertise.
@mark_k: Fascinating and very deep article about DeepSeek AI (@deepseek_ai). You would have never guessed what their strategy is…
An analysis of DeepSeek AI's unconventional strategy: prioritizing radical architecture innovations (MoE, MLA, engram, mHC) that drastically reduce compute and memory needs, enabling a long-term play to build a 10T Chinese AI hardware ecosystem and pursue a 1T valuation.
Update: DeepSeek AI and the Great Talent Competition
This analysis updates the study of DeepSeek's research team, revealing that their talent pool has grown to 356 researchers with increasing citation impact and that over half have only Chinese affiliations, highlighting challenges for U.S. talent retention and independence.
Deep research System Card
OpenAI launches Deep Research, an agentic capability powered by an early version of o3 that conducts multi-step internet research for complex tasks, with comprehensive safety testing and privacy protections implemented before rollout to Pro users.