กลับไปที่ข่าว
นวัตกรรมAI Understanding บรรยายสรุป

นักวิจัยชาวจีนวางแผน 5 ขั้นตอนเพื่อพัฒนา AI ด้วยตนเอง

นักวิจัยจาก ByteDance, มหาวิทยาลัย Tsinghua และห้องปฏิบัติการปัญญาประดิษฐ์เซี่ยงไฮ้ ได้สรุปแผนงานห้าขั้นตอนสำหรับระบบ AI ที่สามารถปรับปรุงกระบวนการพัฒนาของตนเองได้ในที่สุด

4 min readRead the original reporting
Source-provided image accompanying Chinese researchers map five stages toward self-improving AI
การรายงานที่มีการระบุแหล่งที่มาแหล่งที่มาบันทึกไว้
สำนักพิมพ์
scmp.com
ลิงค์แหล่งที่มา
scmp.comhttps://www.scmp.com/tech/tech-trends/article/3367486/chinese-researchers-chart-five-stage-path-toward-last-ai-built-humans
ประเภทแหล่งที่มา
การรายงานโดยสำนักข่าว — ไม่ใช่เอกสารของบุคคลที่หนึ่ง

สิ่งที่เราไม่สามารถยืนยันได้อย่างอิสระ: การอ้างสิทธิ์นี้มาจากร้านที่มีชื่อ เราไม่ได้ตรวจสอบกับเอกสารของบุคคลที่หนึ่ง (scmp.com)

บริบทเข้าใจสิ่งนี้ใน 60 วินาที

เริ่มที่นี่

เงื่อนไขสำคัญ

ปัญญาประดิษฐ์ (AI)
ระบบอาคารในสาขากว้างๆ ที่ทำงานซึ่งต้องใช้การจดจำรูปแบบ การใช้เหตุผล ภาษา หรือการตัดสินใจ
หน่วยความจำ (หน่วยความจำตัวแทน)
บริบทที่จัดเก็บไว้ซึ่งตัวแทน AI ใช้ในขั้นตอนหรือเซสชันเพื่อปรับปรุงความต่อเนื่อง
การปรับแต่งแบบละเอียด
การฝึกอบรมอย่างต่อเนื่องเกี่ยวกับข้อมูลเฉพาะโดเมนเพื่อปรับโมเดลที่ได้รับการฝึกอบรมล่วงหน้าให้เข้ากับงานเฉพาะ
ทดสอบตัวเองแบบทดสอบอธิบายโมเดล AI

เกิดอะไรขึ้น

The South China Morning Post reports that researchers from Chinese universities and technology companies published a paper describing five stages of recursive self-improvement, from executing human-designed upgrades to persistently refining the methods used to improve AI systems. The paper presents this as a research direction, not a demonstrated product.

The South China Morning Post reports that researchers from ByteDance, Tsinghua University, the Shanghai Artificial Intelligence Laboratory and other institutions published “The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement.” The paper sets out five progressive stages for recursive self-improvement, or RSI. At first, an AI would execute improvement procedures designed by human engineers. It would then select upgrade strategies, determine what information or experience it needs, adapt after deployment, and ultimately refine the methods used to improve AI itself.

The report distinguishes RSI from a chatbot correcting a single response. For an improvement to qualify as RSI, the change would need to persist beyond one task and be inherited by successor systems. The researchers argue that automating parts of training, evaluation and could shorten development cycles and reduce labor and computing costs, but these are the authors’ claims rather than independently demonstrated results in the supplied report.

The article also describes related efforts by Chinese companies. Z.ai reportedly plans to direct about 60 percent of proceeds from a US$5 billion fundraising round toward next-generation GLM models and a self-training system. Researchers associated with MiniMax 2.7 reportedly described memory updates and reinforcement-learning experiments, while DeepSeek reportedly developed an agentic harness for multi-step tasks, code execution and external software interaction. These examples do not establish that any company has achieved the paper’s final RSI stage.

รายละเอียดที่มา: scmp.com ↗

ทำไมมันถึงสำคัญ

Automating parts of AI research could affect how quickly and cheaply developers train future models, while shifting more control over model improvement from people to AI systems. The proposal also highlights a central safety challenge: systems that change their own improvement processes could require reliable testing and oversight before deployment. The report provides no evidence that the final stages have been achieved.

If reliable, systems that automate parts of AI research could become a competitive advantage by allowing developers to run more experiments and improve models with less direct human labor. That could influence the pace and cost of foundation-model development, although the report supplies no independent measurements of those effects.

The proposal is also consequential for AI safety. The researchers say genuine RSI would require strict safeguards and verified testing environments so that updates are shown to be safe and beneficial before deployment. The supplied report does not independently verify the roadmap, the related company claims, or the effectiveness of any proposed safeguards.

Interactive Mechanism

กลไกเชิงโต้ตอบ: มันทำงานอย่างไร

สำรวจเทคโนโลยีเบื้องหลังการพัฒนานี้แบบโต้ตอบ

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
การตรวจสอบแนวคิดแบบโต้ตอบ+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

จะดูอะไรต่อไป.

The paper gives no timetable for reaching genuine recursive self-improvement. There is no documented product access, pricing, or general availability. Key unknowns include whether the proposed stages can work reliably outside software engineering, how much compute they require, and whether safeguards can detect harmful or ineffective updates.

The researchers provide no timetable for achieving genuine RSI, and the South China Morning Post does not report a working system that has reached the final stage. Access to the research described is not specified, and no product pricing or general availability is documented.

Progress may differ substantially by field. The authors reportedly see software engineering as a clearer path, while robotics and scientific discovery present harder technical challenges. Compute availability is another limitation: the report quotes an expert saying US companies remain several months ahead and have greater deployment compute, while Chinese researchers continue to optimize under hardware constraints.

Future reporting should establish whether claimed self-training systems produce persistent, reproducible improvements, how updates are evaluated, and whether human approval remains required. No independent test results are provided in the supplied source.

คำแนะนำและแบบทดสอบที่เกี่ยวข้อง

อธิบายโมเดล AIการฝึกอบรมเอไอตัวแทนเอไอทดสอบสิ่งที่คุณรู้ — ลองแบบทดสอบ AI ฟรีค้นหาคำศัพท์ AI ในอภิธานศัพท์ของเราติดตามตัวติดตามการเปิดตัวโมเดล AI
พบว่าสิ่งนี้มีประโยชน์หรือไม่?