Quay lại Tin tức
Đổi mớiAI Understanding tóm tắt

Các nhà nghiên cứu Trung Quốc vạch ra 5 giai đoạn hướng tới việc tự cải thiện AI

Các nhà nghiên cứu từ ByteDance, Đại học Thanh Hoa và Phòng thí nghiệm trí tuệ nhân tạo Thượng Hải đã vạch ra lộ trình gồm 5 giai đoạn cho các hệ thống AI mà cuối cùng có thể cải thiện quy trình phát triển của chính chúng.

4 min readRead the original reporting
Source-provided image accompanying Chinese researchers map five stages toward self-improving AI
Báo cáo phân bổNguồn đã ghi
Nhà xuất bản
scmp.com
Liên kết nguồn
scmp.comhttps://www.scmp.com/tech/tech-trends/article/3367486/chinese-researchers-chart-five-stage-path-toward-last-ai-built-humans
Loại nguồn
Báo cáo của một cơ quan báo chí — không phải tài liệu của bên thứ nhất.

Những gì chúng tôi không thể xác nhận độc lập: Khiếu nại này được quy cho ổ cắm được đặt tên. Chúng tôi đã không xác minh nó dựa trên tài liệu của bên thứ nhất. (scmp.com)

Bối cảnhHiểu điều này trong 60 giây

Bắt đầu ở đây

Thuật ngữ chính

Trí tuệ nhân tạo (AI)
Lĩnh vực rộng lớn của việc xây dựng các hệ thống thực hiện các nhiệm vụ yêu cầu nhận dạng mẫu, lý luận, ngôn ngữ hoặc ra quyết định.
Bộ nhớ (Bộ nhớ tác nhân)
Bối cảnh được lưu trữ mà tác nhân AI sử dụng qua các bước hoặc phiên để cải thiện tính liên tục.
Tinh chỉnh
Tiếp tục đào tạo về dữ liệu theo miền cụ thể để điều chỉnh mô hình được đào tạo trước cho phù hợp với một nhiệm vụ cụ thể.
Tự kiểm traCâu đố giải thích về mô hình AI

Chuyện gì đã xảy ra

The South China Morning Post reports that researchers from Chinese universities and technology companies published a paper describing five stages of recursive self-improvement, from executing human-designed upgrades to persistently refining the methods used to improve AI systems. The paper presents this as a research direction, not a demonstrated product.

The South China Morning Post reports that researchers from ByteDance, Tsinghua University, the Shanghai Artificial Intelligence Laboratory and other institutions published “The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement.” The paper sets out five progressive stages for recursive self-improvement, or RSI. At first, an AI would execute improvement procedures designed by human engineers. It would then select upgrade strategies, determine what information or experience it needs, adapt after deployment, and ultimately refine the methods used to improve AI itself.

The report distinguishes RSI from a chatbot correcting a single response. For an improvement to qualify as RSI, the change would need to persist beyond one task and be inherited by successor systems. The researchers argue that automating parts of training, evaluation and could shorten development cycles and reduce labor and computing costs, but these are the authors’ claims rather than independently demonstrated results in the supplied report.

The article also describes related efforts by Chinese companies. Z.ai reportedly plans to direct about 60 percent of proceeds from a US$5 billion fundraising round toward next-generation GLM models and a self-training system. Researchers associated with MiniMax 2.7 reportedly described memory updates and reinforcement-learning experiments, while DeepSeek reportedly developed an agentic harness for multi-step tasks, code execution and external software interaction. These examples do not establish that any company has achieved the paper’s final RSI stage.

Chi tiết nguồn: scmp.com ↗

Tại sao nó quan trọng

Automating parts of AI research could affect how quickly and cheaply developers train future models, while shifting more control over model improvement from people to AI systems. The proposal also highlights a central safety challenge: systems that change their own improvement processes could require reliable testing and oversight before deployment. The report provides no evidence that the final stages have been achieved.

If reliable, systems that automate parts of AI research could become a competitive advantage by allowing developers to run more experiments and improve models with less direct human labor. That could influence the pace and cost of foundation-model development, although the report supplies no independent measurements of those effects.

The proposal is also consequential for AI safety. The researchers say genuine RSI would require strict safeguards and verified testing environments so that updates are shown to be safe and beneficial before deployment. The supplied report does not independently verify the roadmap, the related company claims, or the effectiveness of any proposed safeguards.

Interactive Mechanism

Cơ chế tương tác: Nó thực sự hoạt động như thế nào

Khám phá công nghệ cơ bản đằng sau sự phát triển này một cách tương tác.

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
Kiểm tra khái niệm tương tác+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

Xem gì tiếp theo

The paper gives no timetable for reaching genuine recursive self-improvement. There is no documented product access, pricing, or general availability. Key unknowns include whether the proposed stages can work reliably outside software engineering, how much compute they require, and whether safeguards can detect harmful or ineffective updates.

The researchers provide no timetable for achieving genuine RSI, and the South China Morning Post does not report a working system that has reached the final stage. Access to the research described is not specified, and no product pricing or general availability is documented.

Progress may differ substantially by field. The authors reportedly see software engineering as a clearer path, while robotics and scientific discovery present harder technical challenges. Compute availability is another limitation: the report quotes an expert saying US companies remain several months ahead and have greater deployment compute, while Chinese researchers continue to optimize under hardware constraints.

Future reporting should establish whether claimed self-training systems produce persistent, reproducible improvements, how updates are evaluated, and whether human approval remains required. No independent test results are provided in the supplied source.

Hướng dẫn và câu hỏi liên quan

Giải thích về mô hình AIĐào tạo AIĐại lý AIKiểm tra những gì bạn biết — thử một bài kiểm tra AI miễn phíTra cứu một thuật ngữ AI trong bảng thuật ngữ của chúng tôiTheo dõi trình theo dõi phát hành mô hình AI
Tìm thấy điều này hữu ích?