header-langage
简体中文
繁體中文
English
Tiếng Việt
한국어
日本語
ภาษาไทย
Türkçe
Scan to Download the APP

How Far is AI from "Self-Improvement"? Scale Introduces RSI Bench

Dynamic Beating AI News: Scale AI has launched RSI Bench, specifically designed to test whether AI can autonomously conduct AI research like a researcher.

The AI will directly access papers, code, and computational resources, then conduct its own experiments, training runs, method modifications, and assess whether it can improve upon the original AI.

The initial testing involved Claude Opus 5 and GPT-5.6 Sol. Both models were able to conduct experiments, tune parameters, iterate on solutions autonomously, and some tasks did surpass the given baselines.

However, current AIs are not yet proficient at "inventing." When challenged with the latest research, most attempts still rely on existing methods and parameter tuning, with few proposing truly novel ideas.

举报 Correction/Report
Correction/Report
Submit
Add Library
Visible to myself only
Public
Save
Choose Library
Add Library
Cancel
Finish