返回新聞
創新AI Understanding 簡報

中國研究人員繪製了自我改進人工智慧的五個階段

來自位元組跳動、清華大學和上海人工智慧實驗室的研究人員概述了人工智慧系統的五階段路線圖,最終可以改善自身的開發流程。

4 min readRead the original reporting
Source-provided image accompanying Chinese researchers map five stages toward self-improving AI
歸因報告來源記錄
出版商
scmp.com
來源連結
scmp.comhttps://www.scmp.com/tech/tech-trends/article/3367486/chinese-researchers-chart-five-stage-path-toward-last-ai-built-humans
來源類型
新聞媒體的報道-不是第一方文件。

我們無法獨立確認的內容: 此聲明歸因於指定的商店。我們沒有根據第一方文件對其進行驗證。 (scmp.com)

背景60 秒內了解這一點

從這裡開始

關鍵術語

人工智慧(AI)
建構執行需要模式識別、推理、語言或決策的任務的系統的廣泛領域。
記憶體(代理記憶體)
AI 代理程式跨步驟或會話使用儲存的上下文來提高連續性。
微調
對特定領域的資料進行持續訓練,以使預先訓練的模型適應特定任務。
測試一下自己AI 模型解釋測驗

發生了什麼事

The South China Morning Post reports that researchers from Chinese universities and technology companies published a paper describing five stages of recursive self-improvement, from executing human-designed upgrades to persistently refining the methods used to improve AI systems. The paper presents this as a research direction, not a demonstrated product.

The South China Morning Post reports that researchers from ByteDance, Tsinghua University, the Shanghai Artificial Intelligence Laboratory and other institutions published “The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement.” The paper sets out five progressive stages for recursive self-improvement, or RSI. At first, an AI would execute improvement procedures designed by human engineers. It would then select upgrade strategies, determine what information or experience it needs, adapt after deployment, and ultimately refine the methods used to improve AI itself.

The report distinguishes RSI from a chatbot correcting a single response. For an improvement to qualify as RSI, the change would need to persist beyond one task and be inherited by successor systems. The researchers argue that automating parts of training, evaluation and could shorten development cycles and reduce labor and computing costs, but these are the authors’ claims rather than independently demonstrated results in the supplied report.

The article also describes related efforts by Chinese companies. Z.ai reportedly plans to direct about 60 percent of proceeds from a US$5 billion fundraising round toward next-generation GLM models and a self-training system. Researchers associated with MiniMax 2.7 reportedly described memory updates and reinforcement-learning experiments, while DeepSeek reportedly developed an agentic harness for multi-step tasks, code execution and external software interaction. These examples do not establish that any company has achieved the paper’s final RSI stage.

來源詳情: scmp.com ↗

為什麼這很重要

Automating parts of AI research could affect how quickly and cheaply developers train future models, while shifting more control over model improvement from people to AI systems. The proposal also highlights a central safety challenge: systems that change their own improvement processes could require reliable testing and oversight before deployment. The report provides no evidence that the final stages have been achieved.

If reliable, systems that automate parts of AI research could become a competitive advantage by allowing developers to run more experiments and improve models with less direct human labor. That could influence the pace and cost of foundation-model development, although the report supplies no independent measurements of those effects.

The proposal is also consequential for AI safety. The researchers say genuine RSI would require strict safeguards and verified testing environments so that updates are shown to be safe and beneficial before deployment. The supplied report does not independently verify the roadmap, the related company claims, or the effectiveness of any proposed safeguards.

Interactive Mechanism

互動機制:它實際上是如何運作的

以互動方式探索這項發展背後的基礎技術。

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
互動式概念檢查+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

接下來看什麼

The paper gives no timetable for reaching genuine recursive self-improvement. There is no documented product access, pricing, or general availability. Key unknowns include whether the proposed stages can work reliably outside software engineering, how much compute they require, and whether safeguards can detect harmful or ineffective updates.

The researchers provide no timetable for achieving genuine RSI, and the South China Morning Post does not report a working system that has reached the final stage. Access to the research described is not specified, and no product pricing or general availability is documented.

Progress may differ substantially by field. The authors reportedly see software engineering as a clearer path, while robotics and scientific discovery present harder technical challenges. Compute availability is another limitation: the report quotes an expert saying US companies remain several months ahead and have greater deployment compute, while Chinese researchers continue to optimize under hardware constraints.

Future reporting should establish whether claimed self-training systems produce persistent, reproducible improvements, how updates are evaluated, and whether human approval remains required. No independent test results are provided in the supplied source.

相關指引和測驗

人工智慧模型解釋人工智慧培訓人工智慧代理測試你所知道的—嘗試免費的人工智慧測驗在我們的詞彙表中尋找人工智慧術語關注 AI 模型發布追蹤器
覺得有用嗎?