Research author

Bowen Zhou

10 AI research papers in the One9Founders library, with summaries and links to original sources.

Papers by Bowen Zhou

Sep 10, 2026
cs.LG

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

Yi Duan, Ying Liu, Zirui Tang, et al.

Aug 14, 2026
cs.AI

Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning

Kai Chen, Jifeng Ding, Ning Ding, et al.

31
Aug 13, 2026
cs.LG

Intern-S2-Preview: Scientific Agentic Foundation Model

Lei Bai, Jiaqi Cao, Chiyu Chen, et al.

56
Jul 30, 2026
cs.CL

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

Junlin Yang, Che Jiang, Yu Fu, et al.

169
Jul 8, 2026
cs.CL

Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning

Chen Tang, Yizhou Wang, Jianyu Wu, et al.

84
Jul 7, 2026
cs.AI

A Definition and Roadmap for World Models

Xinyuan Chen, Haoyu Guo, Shi Guo, et al.

Jun 29, 2026
cs.CL

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

Lei Bai, Zongsheng Cao, Yang Chen, et al.

35
Jun 23, 2026
cs.CL

NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?

Yuru Wang, Lejun Cheng, Yuxin Zuo, et al.

52
May 18, 2026
cs.LG

Post-Trained MoE Can Skip Half Experts via Self-Distillation

Xingtai Lv, Li Sheng, Kaiyan Zhang, et al.

19
Mar 19, 2026
cs.AI

OS-Themis: A Scalable Critic Framework for Generalist GUI Rewards

Zehao Li, Zhenyu Wu, Yibo Zhao, et al.

OS-Themis is a new framework that helps train AI agents to better interact with graphical user interfaces (like phone apps) by providing more reliable feedback on whether the agent is performing tasks correctly. Instead of using a single judge, it breaks down agent actions into verifiable milestones and cross-checks the evidence before making a decision, similar to how a court system works. When tested on smartphone tasks, this approach improved performance by about 10% during training and 7% when filtering practice data.

reinforcement learningGUI agentsreward functionsmulti-agent systems
Bowen Zhou — AI Research Papers | One9Founders