返回新聞
產品展示AI Understanding 簡報

Simate-beta 實體人工智慧模型首次登頂 RoboDojo 排行榜

一家成立僅三個月的新創公司推出了 Simate-beta,這是一款通用實體人工智慧系統,在 RoboDojo 排行榜上名列第一,展示了記憶體、長期規劃和精細操作能力。

4 min readRead the linked source
Source-provided image accompanying Simate-beta physical AI model tops RoboDojo leaderboard in debut
來源參考來源記錄
出版商
eu.36kr.com
來源連結
eu.36kr.comhttps://eu.36kr.com/en/p/3999916051157129
來源類型
連結來源-主要來源狀態尚未確定。
背景60 秒內了解這一點

從這裡開始

關鍵術語

記憶體(代理記憶體)
AI 代理程式跨步驟或會話使用儲存的上下文來提高連續性。
推理
經過訓練的模型產生預測或輸出的運行時階段。
管道
預處理、模型步驟和後處理階段的有序工作流程。
測試一下自己AI 模型解釋測驗

發生了什麼事

Simate, a startup founded three months ago, released its first physical AI model called Simate‑beta. According to 36kr, the model demonstrated capabilities in memory, long‑horizon tasks, fine manipulation, and task adaptation, and achieved the top spot on the RoboDojo leaderboard with an average score of 33.95 (SR 27.96%). The company says the base model was not specially optimized for the leaderboard and that the result reflects its AI‑native research platform, AutoResearch, which automates hypothesis testing, training, and evaluation. The startup also announced that its research tools have been used by external researchers from MIT, Caltech, Tsinghua University and Peking University in a beta test, and that it has completed multiple financing rounds worth hundreds of millions of RMB. The team plans to release technical papers, open‑source components in phases, and to roll out further results by the end of 2026.

Simate announced Simate‑beta, describing it as a "general‑purpose physical fast system" that integrates a high‑level planning component (the "slow system") with a rapid perception‑action component (the "fast system"). The company claims the model can remember recent events, understand complex instructions, and execute millisecond‑level adjustments without relying on task‑by‑task retraining.

The company attributes its rapid progress to an AI‑native research called AutoResearch, which automates hypothesis generation, experiment execution, and result analysis across simulation and real‑robot environments. According to the report, this pipeline allowed Simate to top the RoboDojo leaderboard within three months of its founding.

External researchers from top universities have participated in a beta test of the AutoResearch platform, and the startup has secured several financing rounds totaling hundreds of millions of RMB. The team plans to publish three research papers and release parts of the platform as open source later in the year.

Simate‑beta’s leaderboard performance was reported on September 23, 2026, with an average score of 33.95 and a success rate of 27.96%. The company notes that the leaderboard evaluation does not fully capture real‑world robot performance, which will be disclosed separately.

來源詳情: eu.36kr.com ↗

為什麼這很重要

The launch marks a notable step toward general‑purpose physical AI that can operate in real‑world environments without task‑specific retraining. If the claimed performance holds, Simate‑beta could accelerate robot deployment in industries such as logistics, manufacturing, and home assistance, where high‑speed perception and long‑term planning are critical. The company’s AI‑native R&D workflow—combining a “SiPAI” model framework, an AutoResearch engine, and custom infrastructure—promises to dramatically increase research throughput, potentially reshaping how robotics labs iterate on models. However, independent verification of the leaderboard scores and real‑world performance is lacking, and details on model size, hardware requirements, pricing, and commercial availability remain undisclosed. These unknowns limit immediate practical impact but highlight a potentially transformative approach to robot AI development.

Physical AI that can generalize across tasks without extensive retraining is a long‑standing challenge in robotics. Simate‑beta’s claimed ability to combine long‑term memory with fast, precise motor control could reduce the engineering effort required to deploy robots in new settings.

The AutoResearch workflow, if effective, could serve as a template for other AI labs seeking to accelerate robot research, potentially lowering the barrier to entry for smaller teams and academic groups.

The involvement of researchers from institutions such as MIT and Tsinghua suggests early academic interest, which could foster collaborations and independent validation of the technology.

The lack of disclosed model size, compute requirements, and pricing means that the system’s accessibility to developers and enterprises remains uncertain. Without independent verification, the leaderboard claim alone does not guarantee real‑world efficacy.

Interactive Mechanism

互動機制:它實際上是如何運作的

以互動方式探索這項發展背後的基礎技術。

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
互動式概念檢查+10 Points
AI Models Explained Quiz

In AI, what are a model's "parameters"?

接下來看什麼

Key areas to monitor include: (1) independent benchmarking of Simate‑beta on external robot testbeds to confirm the RoboDojo results; (2) announcements of model specifications, hardware requirements, and pricing that would indicate commercial readiness; (3) the rollout of the AutoResearch platform to external partners and any open‑source releases that could enable broader adoption; and (4) regulatory or safety assessments as the system moves toward more complex, real‑world tasks.

Independent third‑party evaluations of Simate‑beta on standard robot benchmarks (e.g., OpenAI Gym Robotics, Real‑World RL Suite) to confirm the leaderboard scores.

Public release of technical specifications, including model parameters, latency, and hardware platforms needed for deployment.

Announcements regarding the open‑source components of the AutoResearch platform, which could enable broader community adoption and scrutiny.

Regulatory scrutiny or safety certifications as the system moves toward more complex manipulation tasks in uncontrolled environments.

相關指引和測驗

人工智慧模型解釋人工智慧代理AI 的未來測試你所知道的—嘗試免費的人工智慧測驗在我們的詞彙表中尋找人工智慧術語關注 AI 模型發布追蹤器
覺得有用嗎?