Superalignment Fast Grants

OpenAI Blog News

Summary

OpenAI announces Superalignment Fast Grants to fund research on aligning superintelligent AI systems, addressing the fundamental challenge of how humans can steer and trust AI systems more capable than themselves. The initiative seeks to rally top researchers to tackle this critical technical problem, which OpenAI believes superintelligence could pose within the next decade.

We’re launching $10M in grants to support technical research towards the alignment and safety of superhuman AI systems, including weak-to-strong generalization, interpretability, scalable oversight, and more.
Original Article
View Cached Full Text

Cached at: 04/20/26, 02:54 PM

# Superalignment Fast Grants Source: [https://openai.com/index/superalignment-fast-grants/](https://openai.com/index/superalignment-fast-grants/) We believe superintelligence could arrive within the next 10 years\. These AI systems would have vast capabilities—they could be hugely beneficial, but also potentially pose large risks\. Today, we[align AI systems⁠](https://openai.com/index/instruction-following/)to ensure they are safe using reinforcement learning from human feedback \(RLHF\)\. However, aligning future superhuman AI systems will pose fundamentally new and qualitatively different technical challenges\. Superhuman AI systems will be capable of complex and creative behaviors that humans cannot fully understand\. For example, if a superhuman model generates a million lines of extremely complicated code, humans will not be able to reliably evaluate whether the code is safe or dangerous to execute\. Existing alignment techniques like RLHF that rely on human supervision may no longer be sufficient\.**This leads to the fundamental challenge: how can humans steer and trust AI systems much smarter than them?** This is one of the most important unsolved technical problems in the world\. But we think it is solvable with a concerted effort\. There are many promising approaches and exciting directions, with lots of low\-hanging fruit\. We think there is an enormous opportunity for the ML research community and individual researchers to make major progress on this problem today\. As part of our[Superalignment⁠](https://openai.com/superalignment/)project, we want to rally the best researchers and engineers in the world to meet this challenge—and we’re especially excited to bring new people into the field\.

Similar Articles

Weak-to-strong generalization

OpenAI Blog

OpenAI's Superalignment team introduces weak-to-strong generalization, a new research direction for empirically aligning superhuman AI models by addressing the fundamental challenge of how weak human supervisors can reliably control and steer AI systems vastly smarter than themselves.

Advancing independent research on AI alignment

OpenAI Blog

OpenAI is contributing $7.5 million to The Alignment Project, a global independent alignment research fund created by the UK AI Security Institute, helping make it one of the largest dedicated funding efforts for independent alignment research to date. The total fund exceeds £27 million and will support a broad portfolio of alignment research projects worldwide.

Governance of superintelligence

OpenAI Blog

OpenAI outlines a framework for superintelligence governance emphasizing three key pillars: coordination among leading AI development efforts, an international authority (akin to the IAEA) to oversee systems above certain capability thresholds, and technical progress on AI safety with democratic public oversight of the most powerful systems.

Announcing the OpenAI Safety Fellowship

OpenAI Blog

OpenAI announces a new Safety Fellowship program for external researchers to conduct rigorous safety and alignment research on advanced AI systems, running September 2026 through February 2027. The program offers mentorship, compute support, stipends, and workspace at Constellation in Berkeley, with applications open until May 3.