MEGA Hub

ASI-Bench: At the Dawn of Artificial Superintelligence

Authors

Do you know Junwei Zhou?You can claim authorship or link another user.Do you know Zhen Sun?You can claim authorship or link another user.Do you know Binyu Li?You can claim authorship or link another user.Do you know Jiangyu Zhou?You can claim authorship or link another user.Do you know Yuexi Pan?You can claim authorship or link another user.Do you know Hengyu Wang?You can claim authorship or link another user.Do you know Honghe Ren?You can claim authorship or link another user.Do you know Xiaohan Jia?You can claim authorship or link another user.Do you know Xueyang Zhou?You can claim authorship or link another user.Do you know Xiaoyu Cao?You can claim authorship or link another user.Do you know Yongchao Chen?You can claim authorship or link another user.Do you know Yuanning Feng?You can claim authorship or link another user.Do you know Junhao Wu?You can claim authorship or link another user.Do you know Cheng Zhang?You can claim authorship or link another user.Do you know Sijia Chen?You can claim authorship or link another user.Do you know Haoyu Xue?You can claim authorship or link another user.Do you know Chengsong You?You can claim authorship or link another user.Do you know Huan Wang?You can claim authorship or link another user.Do you know Koutian Wu?You can claim authorship or link another user.Do you know Peigan Gao?You can claim authorship or link another user.Do you know Jiakun Wu?You can claim authorship or link another user.Do you know Wenzhe Li?You can claim authorship or link another user.Do you know Ergan Shang?You can claim authorship or link another user.Do you know Qingyuan Zheng?You can claim authorship or link another user.Do you know Jingjing Zhou?You can claim authorship or link another user.Do you know Ruixuan Jia?You can claim authorship or link another user.Do you know Yan Xu?You can claim authorship or link another user.Do you know Hongrui Zhang?You can claim authorship or link another user.Do you know Xiao-Han Ma?You can claim authorship or link another user.Do you know Zhengxiang Cheng?You can claim authorship or link another user.Do you know Yuexing Hao?You can claim authorship or link another user.Do you know Liting Mai?You can claim authorship or link another user.Do you know Xianglin Ji?You can claim authorship or link another user.Do you know Wenjun Zhang?You can claim authorship or link another user.Do you know Zhuofan Chen?You can claim authorship or link another user.Do you know Yixiao Huang?You can claim authorship or link another user.Do you know Chi Wang?You can claim authorship or link another user.Do you know Wenyue Hua?You can claim authorship or link another user.Do you know Yilun Hao?You can claim authorship or link another user.Do you know Yuantao Zhai?You can claim authorship or link another user.Do you know Ziyan Zhao?You can claim authorship or link another user.Do you know Jingyan Xie?You can claim authorship or link another user.

Abstract

Artificial superintelligence (ASI) requires AI to move beyond mastering existing knowledge toward exploring the unknown, creating new knowledge, and turning new ideas into verifiable results. However, the capabilities of today's AI systems are still largely built on learning, compressing, and applying existing human knowledge. Accordingly, existing benchmarks primarily test whether AI can produce correct answers based on learned knowledge, or whether it can complete tasks under extensive human guidance. We therefore introduce ASI-Bench, the first benchmark to jointly evaluate AI systems' capabilities of innovative exploration and autonomous scientific execution across general research domains, and the first to progressively withdraw human methodological guidance within the same research project to test how far AI can proceed on its own. Built by over 40 experts with the cost of 31,000+ human hours, ASI-Bench contains 60 project-level research tasks across 11 scientific domains and progressively reduces methodological guidance to test whether AI can independently select methods, conduct research, and produce verifiable results. All tasks undergo expert review, AI-assisted auditing, sandbox execution, and scorer validation. Across 18 state-of-the-art agent--model configurations, the average score drops from 50.91 with full methodological guidance to 29.10 with only the method specified and 26.62 when agents must determine the method themselves. This sharp decline shows that current systems remain heavily dependent on human guidance and are still far from autonomously conducting end-to-end, project-level scientific research. ASI-Bench is open to the world. We invite researchers and builders everywhere to contribute new tasks, challenge the limits of today's AI, and help accelerate humanity's collective path toward artificial superintelligence at https://asibench.apexin.ai/submit.

Community

00