@yuetai12575: Excited to see the gains further survive as model scaling!
Summary
The article reports initial results from OpenRSI's Marin-Scaling-Ladder experiment, where AI agents autonomously propose, implement, and evaluate new optimizers across model scales from 550M to 2.5B, discovering PSPR with Codex (GPT-5.6).
View Cached Full Text
Cached at: 09/24/26, 12:19 AM
Excited to see the gains further survive as model scaling!
To take ASI seriously is to accept a weak-to-strong premise: human intelligence can build a process that eventually produces intelligence beyond itself. RSI is the last piece of that puzzle of ASI. It serves as the key to scaling on insights: human insights are limited by bandwidth, while agents could scale the generation of ideas and sweep a far larger region of method-space.
As the early project initiator, we are fully aware that today, everyone is excited about RSI, yet the barriers to participation also keep everyone outside the door. We believe that a model’s RSI capabilities intrinsically come from researchers, independent labs, and domain teams alike, and we think the benefits should ultimately return to everyone who builds models.
This is why we take our name as OpenRSI-Index. We aim to keep RSI open through shared platforms and tools, so more people can participate and benefit. We could shape RSI together, set RSI’s standards, and challenge frontier models with our own related real research work.
For this goal, we build RSI-Anything, a Human-AI collaboration pipeline. Through about one hour of conversation, the pipeline could help turn a real research question into a runnable autoresearch environment. We want researchers and agents to solve these problems together, sharing new insights, methods, and results with the community.
Here is what we already put on the table: • Task from real research projects: Marin-Scaling-Ladder, GPIC-Leaderboard, Open-Jev-Training, Molmo2, Isaac Lab… • Ultra-long-horizon: 60+ hour agent trajectories; 100K+ H100-hours built for the preview • Everything open, including traces: tasks, harnesses, verifiers, full agent trajectories
We believe research will never end. A game turns zero-sum only when the pie is too small to share. However, research is definitely the field with the highest ceiling there is. What shifts is the mindset: it frees researchers to find and formulate the crazier, more valuable, more exciting problems in the world.
Similar Articles
Scaling laws for reward model overoptimization
OpenAI researchers empirically study how reward model overoptimization affects performance, establishing scaling laws that show the relationship between proxy reward optimization and ground truth performance varies by optimization method and scales predictably with model size.
Model Size Scaling in 2023-2031 (21 minute read)
An analysis of AI model size scaling trends from 2023 to 2031, published on LessWrong.
Scaling domain expertise in complex, regulated domains
Blue J demonstrates how to scale AI expertise in complex regulated domains by combining GPT-4.1 with retrieval-augmented generation over curated tax documents, achieving <0.14% error rates and 70% weekly user engagement through rigorous feedback loops and domain-specific optimization.
Interesting post from Adam Majmudar (research at OpenAI) — full text in body
An analysis by Adam Majmudar on the rapid progression of AI capabilities, highlighting how new scaling laws are driving leaps beyond external expectations.
Scaling Participation in Modular AI Systems
This paper introduces scaling participation, a new paradigm for building modular AI systems through contributions from diverse stakeholders, where small models collaborate to outperform monolithic LLMs by up to 15.4% across various tasks, demonstrating emergent capabilities and improved diversity benefits.