OpenAI announces improvements to DALL·E 2's safety systems and bias mitigation based on research preview feedback, including measures to prevent deceptive content creation and enhanced content filtering.
Today, we are implementing a new technique so that DALL·E generates images of people that more accurately reflect the diversity of the world’s population.
# Reducing bias and improving safety in DALL·E 2
Source: [https://openai.com/index/reducing-bias-and-improving-safety-in-dall-e-2/](https://openai.com/index/reducing-bias-and-improving-safety-in-dall-e-2/)
In April, we started previewing the DALL·E 2 research to a limited number of people, which has allowed us to better understand the system’s capabilities and limitations and improve our safety systems\.
During this preview phase, early users have flagged sensitive and biased images which have helped inform and evaluate this new mitigation\.
We are continuing to research how AI systems, like DALL·E, might reflect biases in its training data and different ways we can address them\.
During the research preview we have taken other steps to improve our safety systems, including:
- Minimizing the risk of DALL·E being misused to create deceptive content by rejecting image uploads containing realistic faces and attempts to create the likeness of public figures, including celebrities and prominent political figures\.
- Making our content filters more accurate so that they are more effective at blocking prompts and image uploads that violate our[content policy\(opens in a new window\)](https://labs.openai.com/policies/content-policy)while still allowing creative expression\.
- Refining automated and human monitoring systems to guard against misuse\.
These improvements have helped us gain confidence in the ability to invite more users to experience DALL·E\.
Expanding access is an important part of our[deploying AI systems responsibly](https://openai.com/index/language-model-safety-and-misuse/)because it allows us to learn more about real\-world use and continue to iterate on our safety systems\.
OpenAI describes the pre-training data filtering and active learning techniques used to reduce harmful content in DALL·E 2, while also addressing unintended bias amplification caused by data filtering—particularly demographic biases in generated images.
OpenAI announces an expansion of DALL·E 2 research preview access, sharing safety metrics and learnings from 3 million images created by early users. The company plans to onboard up to 1,000 new users weekly while continuing to refine content policy enforcement and address training data biases.
OpenAI announces DALL·E API is now available in public beta, allowing developers to integrate image generation capabilities directly into their applications. Early adopters include Microsoft, CALA, and Mixtiles, with built-in safety features and content moderation.
OpenAI released the system card for DALL·E 3, detailing safety evaluations, red teaming efforts, and mitigations implemented before deployment of the improved text-to-image model.
OpenAI publishes details on its approach to frontier AI risks and announces progress on voluntary safety commitments made in July 2023, including the release of DALL-E 3 system card and the development of a new Preparedness Framework to manage catastrophic risks from advanced AI systems.