@yuetai12575: Excited to see the gains further survive as model scaling!

X AI KOLs Timeline Papers

Summary

The article reports initial results from OpenRSI's Marin-Scaling-Ladder experiment, where AI agents autonomously propose, implement, and evaluate new optimizers across model scales from 550M to 2.5B, discovering PSPR with Codex (GPT-5.6).

Excited to see the gains further survive as model scaling!
Original Article
View Cached Full Text

Cached at: 09/24/26, 12:19 AM

Excited to see the gains further survive as model scaling!

To take ASI seriously is to accept a weak-to-strong premise: human intelligence can build a process that eventually produces intelligence beyond itself. RSI is the last piece of that puzzle of ASI. It serves as the key to scaling on insights: human insights are limited by bandwidth, while agents could scale the generation of ideas and sweep a far larger region of method-space.

As the early project initiator, we are fully aware that today, everyone is excited about RSI, yet the barriers to participation also keep everyone outside the door. We believe that a model’s RSI capabilities intrinsically come from researchers, independent labs, and domain teams alike, and we think the benefits should ultimately return to everyone who builds models.

This is why we take our name as OpenRSI-Index. We aim to keep RSI open through shared platforms and tools, so more people can participate and benefit. We could shape RSI together, set RSI’s standards, and challenge frontier models with our own related real research work.

For this goal, we build RSI-Anything, a Human-AI collaboration pipeline. Through about one hour of conversation, the pipeline could help turn a real research question into a runnable autoresearch environment. We want researchers and agents to solve these problems together, sharing new insights, methods, and results with the community.

Here is what we already put on the table: • Task from real research projects: Marin-Scaling-Ladder, GPIC-Leaderboard, Open-Jev-Training, Molmo2, Isaac Lab… • Ultra-long-horizon: 60+ hour agent trajectories; 100K+ H100-hours built for the preview • Everything open, including traces: tasks, harnesses, verifiers, full agent trajectories

We believe research will never end. A game turns zero-sum only when the pie is too small to share. However, research is definitely the field with the highest ceiling there is. What shifts is the mindset: it frees researchers to find and formulate the crazier, more valuable, more exciting problems in the world.

Similar Articles

Scaling laws for reward model overoptimization

OpenAI Blog

OpenAI researchers empirically study how reward model overoptimization affects performance, establishing scaling laws that show the relationship between proxy reward optimization and ground truth performance varies by optimization method and scales predictably with model size.

Scaling domain expertise in complex, regulated domains

OpenAI Blog

Blue J demonstrates how to scale AI expertise in complex regulated domains by combining GPT-4.1 with retrieval-augmented generation over curated tax documents, achieving <0.14% error rates and 70% weekly user engagement through rigorous feedback loops and domain-specific optimization.

Scaling Participation in Modular AI Systems

arXiv cs.AI

This paper introduces scaling participation, a new paradigm for building modular AI systems through contributions from diverse stakeholders, where small models collaborate to outperform monolithic LLMs by up to 15.4% across various tasks, demonstrating emergent capabilities and improved diversity benefits.