@GoogleDeepMind: In an industry first, we’re piloting double-blind evaluations for frontier AI. By creating a secure environment where n…
Summary
Google DeepMind is piloting the world's first double-blind evaluations for frontier AI models to ensure secure and trustworthy external assessments, partnering with organizations like Singapore AI Safety Institute and MLCommons.
View Cached Full Text
Cached at: 08/27/26, 01:45 PM
In an industry first, we’re piloting double-blind evaluations for frontier AI.
By creating a secure environment where neither test prompts nor model weights are revealed, we can ensure external safety and performance evaluations of our models remain private, robust, and trustworthy. → https://goo.gle/3St2xan
Piloting the world’s first double-blind AI evaluations
Source: https://deepmind.google/blog/piloting-the-worlds-first-double-blind-ai-evaluations/?utm_source=x&utm_medium=social&utm_campaign=&utm_content= August 27, 2026Responsibility & Safety
Building trust in proprietary model benchmarks using cryptographically secure environments
Imagine a student is set to take a high-stakes exam. If they accidentally peek at the test questions in advance, achieving a perfect score is influenced by this knowledge, making it a meaningless accomplishment. To truly measure what they know, they must have no visibility of the test questions until it’s time to take the exam. That is the exact challenge the industry faces when evaluating advanced AI models. If a model has already seen the test questions - a problem known as benchmark contamination - the results can only be trusted to an extent.
Today, we’re introducing the**world’s first double-blind evaluation of a proprietary, frontier class AI model,**which keeps external evaluations confined to a cryptographic “box” where they can’t be used by models later to optimize performance ahead of testing. We’re partnering with the Singapore AI Safety Institute, OpenMined, AVERI, andMLCommons, to test a Gemini Flash Lite model against confidential benchmarks in aprivacy-preserving environment, increasing evaluation integrity.
At Google, we assess our AI systems using a broad spectrum of evaluations throughout model development and deployment, but we don’t rely on internal testing alone. To identify potential blindspots, we work with a diverse group of external partners, including specialized research labs, civil society and national AI Safety and Security Institutes (AISIs), using their unique expertise to stress-test our models.
As AI models become more capable, ensuring the model has not seen the test questions or prompts in advance is critical, as this can skew the results. Policymakers, researchers, and enterprises need to trust that AI benchmarks accurately reflect a model’s true capabilities and safety, but if models are able to “peek” at the evaluation questions in advance, it can artificially inflate scores and undermine this trust.
Although zero-logging protocols and rigorous contractual safeguards have long kept external test prompts confidential, incorporating technical and cryptographic safeguards marks a major step forward in secure model evaluation.
How double-blind evaluations work
Similar Articles
Piloting the world's first double-blind AI evaluations
Google DeepMind introduces the world's first double-blind AI evaluation using cryptographic environments to prevent benchmark contamination, partnering with organizations like Singapore AI Safety Institute and MLCommons.
@GoogleDeepMind: Over the past year, we’ve collaborated with global scientific experts to evaluate the system on complex problems. It as…
GoogleDeepMind collaborated with global scientific experts to evaluate an AI system that identified new targets for liver fibrosis and fresh approaches to ALS, digesting decades of research.
Frontier and Center: Who evaluates the evaluations? (12 minute read)
Google Data Cloud's frontier AI team discusses a new approach to evaluating AI agents using information theory to create a meta-benchmark called Discovery Bench that measures how vague a query can be before an agent fails, providing a more nuanced map of agent capabilities than simple pass/fail exams.
Partnering with industry leaders to accelerate AI transformation
Google DeepMind has announced partnerships with major consulting firms like Accenture, McKinsey, and Deloittoe to accelerate the adoption of frontier AI and agentic transformation in enterprise sectors.
Google DeepMind is worried about what happens when millions of agents start to interact
Google DeepMind, together with Schmidt Sciences, ARIA, the Cooperative AI foundation, and Google.org, has launched a $10 million funding initiative to research the safety of multi-agent AI systems, aiming to prevent risks such as scams, prompt injections, and cyberattacks as AI agents become widespread.