OpenAI o3 and o4-mini System Card

OpenAI Blog Models

Summary

OpenAI released system cards for o3 and o4-mini models, which feature advanced reasoning capabilities combined with tool integration (web browsing, Python, image analysis, etc.) and are evaluated under OpenAI's Preparedness Framework v2 for safety in biological, cybersecurity, and AI self-improvement domains.

OpenAI o3 and OpenAI o4-mini combine state-of-the-art reasoning with full tool capabilities—web browsing, Python, image and file analysis, image generation, canvas, automations, file search, and memory.
Original Article
View Cached Full Text

Cached at: 04/20/26, 02:48 PM

# OpenAI o3 and o4-mini System Card Source: [https://openai.com/index/o3-o4-mini-system-card/](https://openai.com/index/o3-o4-mini-system-card/) OpenAI o3 and OpenAI o4\-mini combine state\-of\-the\-art reasoning with full tool capabilities—web browsing, Python, image and file analysis, image generation, canvas, automations, file search, and memory\. These models excel at solving complex math, coding, and scientific challenges while demonstrating strong visual perception and analysis\. The models use tools in their chains of thought to augment their capabilities; for example, cropping or transforming images, searching the web, or using Python to analyze data during their thought process\. The OpenAI o\-series models are trained with large\-scale reinforcement learning on chains of thought\. These advanced reasoning capabilities provide new avenues for improving the safety and robustness of our models\. In particular, our models can reason about our safety policies in context when responding to potentially unsafe prompts, through deliberative alignment\. This is the first launch and system card to be released under Version 2 of our[Preparedness Framework⁠](https://openai.com/index/updating-our-preparedness-framework/)\. OpenAI’s Safety Advisory Group \(SAG\) reviewed the results of our Preparedness evaluations and determined that OpenAI o3 and o4\-mini do not reach the High threshold in any of our three Tracked Categories: Biological and Chemical Capability, Cybersecurity, and AI Self\-improvement\. We describe these evaluations below, and provide an update on our work to mitigate risks in these areas\. Read the addendum to o3 and o4\-mini system card: Codex[here](https://openai.com/index/o3-o4-mini-codex-system-card-addendum/)\. Read the addendum to o3 and o4\-mini system card: OpenAI o3 Operator[here](https://openai.com/index/o3-o4-mini-system-card-addendum-operator-o3/)\.

Similar Articles

OpenAI o3-mini System Card

OpenAI Blog

OpenAI releases the o3-mini System Card, documenting safety evaluations and risk assessments for their advanced reasoning model trained with reinforcement learning. The model achieves state-of-the-art safety performance on certain benchmarks and is classified as Medium risk overall under OpenAI's Preparedness Framework.

OpenAI o1 System Card

OpenAI Blog

OpenAI releases the o1 System Card detailing safety evaluations and preparedness framework assessments for the o1 and o1-mini models, which use chain-of-thought reasoning trained with large-scale reinforcement learning to improve safety and robustness.

Introducing OpenAI o3 and o4-mini

OpenAI Blog

OpenAI releases o3 and o4-mini, its latest reasoning models that can agentically access and combine all ChatGPT tools (web search, code execution, image analysis, image generation). o3 achieves state-of-the-art performance on coding, math, and science benchmarks with 20% fewer major errors than o1, while o4-mini offers efficient reasoning optimized for cost and speed.

OpenAI o3-mini

OpenAI Blog

OpenAI releases o3-mini, a cost-efficient reasoning model with strong STEM capabilities, available in ChatGPT and API with support for function calling, structured outputs, and three reasoning effort levels. The model matches o1 performance in math and coding while being faster and cheaper, with free plan users gaining access to a reasoning model for the first time.