返回新闻
创新AI Understanding 简报

中国研究人员绘制了自我改进人工智能的五个阶段

来自字节跳动、清华大学和上海人工智能实验室的研究人员概述了人工智能系统的五阶段路线图,最终可以改善自身的开发流程。

4 min readRead the original reporting
Source-provided image accompanying Chinese researchers map five stages toward self-improving AI
归因报告来源记录
出版商
scmp.com
来源链接
scmp.comhttps://www.scmp.com/tech/tech-trends/article/3367486/chinese-researchers-chart-five-stage-path-toward-last-ai-built-humans
来源类型
新闻媒体的报道——不是第一方文件。

我们无法独立确认的内容: 此声明归因于指定的商店。我们没有根据第一方文件对其进行验证。 (scmp.com)

背景60 秒内了解这一点

从这里开始

关键术语

人工智能(AI)
构建执行需要模式识别、推理、语言或决策的任务的系统的广泛领域。
内存(代理内存)
AI 代理跨步骤或会话使用存储的上下文来提高连续性。
微调
对特定领域的数据进行持续训练,以使预先训练的模型适应特定任务。
测试一下自己AI 模型解释测验

发生了什么

The South China Morning Post reports that researchers from Chinese universities and technology companies published a paper describing five stages of recursive self-improvement, from executing human-designed upgrades to persistently refining the methods used to improve AI systems. The paper presents this as a research direction, not a demonstrated product.

The South China Morning Post reports that researchers from ByteDance, Tsinghua University, the Shanghai Artificial Intelligence Laboratory and other institutions published “The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement.” The paper sets out five progressive stages for recursive self-improvement, or RSI. At first, an AI would execute improvement procedures designed by human engineers. It would then select upgrade strategies, determine what information or experience it needs, adapt after deployment, and ultimately refine the methods used to improve AI itself.

The report distinguishes RSI from a chatbot correcting a single response. For an improvement to qualify as RSI, the change would need to persist beyond one task and be inherited by successor systems. The researchers argue that automating parts of training, evaluation and could shorten development cycles and reduce labor and computing costs, but these are the authors’ claims rather than independently demonstrated results in the supplied report.

The article also describes related efforts by Chinese companies. Z.ai reportedly plans to direct about 60 percent of proceeds from a US$5 billion fundraising round toward next-generation GLM models and a self-training system. Researchers associated with MiniMax 2.7 reportedly described memory updates and reinforcement-learning experiments, while DeepSeek reportedly developed an agentic harness for multi-step tasks, code execution and external software interaction. These examples do not establish that any company has achieved the paper’s final RSI stage.

来源详情: scmp.com ↗

为什么这很重要

Automating parts of AI research could affect how quickly and cheaply developers train future models, while shifting more control over model improvement from people to AI systems. The proposal also highlights a central safety challenge: systems that change their own improvement processes could require reliable testing and oversight before deployment. The report provides no evidence that the final stages have been achieved.

If reliable, systems that automate parts of AI research could become a competitive advantage by allowing developers to run more experiments and improve models with less direct human labor. That could influence the pace and cost of foundation-model development, although the report supplies no independent measurements of those effects.

The proposal is also consequential for AI safety. The researchers say genuine RSI would require strict safeguards and verified testing environments so that updates are shown to be safe and beneficial before deployment. The supplied report does not independently verify the roadmap, the related company claims, or the effectiveness of any proposed safeguards.

Interactive Mechanism

互动机制:它实际上是如何运作的

以交互方式探索这一发展背后的基础技术。

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
交互式概念检查+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

接下来看什么

The paper gives no timetable for reaching genuine recursive self-improvement. There is no documented product access, pricing, or general availability. Key unknowns include whether the proposed stages can work reliably outside software engineering, how much compute they require, and whether safeguards can detect harmful or ineffective updates.

The researchers provide no timetable for achieving genuine RSI, and the South China Morning Post does not report a working system that has reached the final stage. Access to the research described is not specified, and no product pricing or general availability is documented.

Progress may differ substantially by field. The authors reportedly see software engineering as a clearer path, while robotics and scientific discovery present harder technical challenges. Compute availability is another limitation: the report quotes an expert saying US companies remain several months ahead and have greater deployment compute, while Chinese researchers continue to optimize under hardware constraints.

Future reporting should establish whether claimed self-training systems produce persistent, reproducible improvements, how updates are evaluated, and whether human approval remains required. No independent test results are provided in the supplied source.

相关指南和测验

人工智能模型解释人工智能培训人工智能代理测试你所知道的——尝试免费的人工智能测验在我们的词汇表中查找人工智能术语关注 AI 模型发布跟踪器
觉得这有用吗?