返回新聞
創新AI Understanding 簡報

Cohere 地圖顯示人工智慧代理工具正在影響人類工作

Cohere Labs 表示,其新的 Agentic Task Ecosystem 資料集發現,在近 70 萬個公共 MCP 工具中,只有 2.6% 端到端地執行公認的職業任務。

4 min readRead the primary source
Source-provided image accompanying Cohere maps where AI agent tools are reaching human work
主要來源文件來源記錄
出版商
cohere.com
來源連結
cohere.comhttps://cohere.com/blog/automations-early-footprint
來源類型
主要文件-我們直接閱讀的官方公告、文件、文件或第一方頁面。
背景60 秒內了解這一點

從這裡開始

關鍵術語

人工智慧代理
一種可以觀察、推理並採取行動來實現目標的軟體系統,通常使用工具和記憶體。
MCP(模型上下文協定)
一種開放協議,允許人工智慧應用程式以標準方式連接到外部工具、資料來源和上下文提供者。
數據集
用於訓練、驗證或測試的結構化或非結構化範例的集合。
測試一下自己AI 代理測驗
Source video from cohere.com · shown with attribution.

發生了什麼事

Cohere Labs released the Agentic Task Ecosystem, a built from 696,291 tools listed across 123,069 public Model Context Protocol servers. The accompanying analysis examines which occupational tasks developers are building AI agents to perform and finds that agentic tooling is concentrated, unevenly distributed and not yet evidence of widespread adoption.

Cohere Labs says it collected 696,291 tools from 123,069 public MCP server listings across seven directories in May 2026 and deduplicated servers listed in multiple places. The company describes the resulting Agentic Task Ecosystem as the largest open of its kind, but the source does not provide a download location, license or access conditions. Pricing is not documented.

Using a language model to match tools with O*NET task statements, the researchers applied a strict test: a tool had to execute a recognized occupational task from end to end, rather than provide information or complete only one step that a person still coordinates. Under that test, 2.6% of tools qualified. Cohere reports that 419 of 923 occupations had no represented agentic tool activity.

The analysis groups most unmatched tools into smaller pieces of existing work, combinations of multiple recorded tasks, or infrastructure needed to operate agents. Cohere identifies 35 categories—about 3% of the categories examined—as apparently new work with no plausible occupational counterpart, mostly involving agent management such as selecting synthetic voices, switching AI personas and assessing whether an agent can be trusted.

Cohere reports that realized tooling correlates with theoretical AI exposure across 178 occupations, but exposure does not predict whether tools reach routine tasks or specialized work. The source says tools extend toward specialized work in healthcare and computing, while they cluster toward routine edges in legal, production and sales occupations. These are findings reported by the authors; the source does not establish downstream economic effects.

來源詳情: cohere.com ↗

為什麼這很重要

The research offers an early supply-side view of AI automation: what developers have packaged for agents to do, rather than what companies have deployed or workers currently use. Cohere’s analysis suggests that overall exposure measures do not reveal which parts of jobs are affected. That distinction could matter for wages, hiring, training and how workers gain expertise, especially when AI reaches specialized software-based work.

The measures supply, not adoption, reliability or economic impact. A public tool shows that a developer considered a task concrete enough to package for an agent, but it does not show that a company uses the tool, that it works reliably in production or that workers have been displaced.

The source argues that the location of automation inside a job may be more consequential than the total share of work exposed. Removing routine work could leave employees with more specialized responsibilities, while automating specialized work could reduce the expertise required for the remaining job. Either pattern could affect wages and employment differently.

Cohere also reports that expert judgments of technical feasibility predicted which occupations received tools, while workers’ stated preferences about what they wanted automated did not. That finding raises a practical governance question: whether AI development will follow what workers and affected communities value, or primarily what developers judge technically tractable.

The research has a significant visibility limitation. It covers public directories and omits bespoke internal MCP servers, which Cohere says may be concentrated in back-office processes. The reported 2.6% share is therefore presented by the authors as a floor, not a complete measure of automation.

Interactive Mechanism

互動機制:它實際上是如何運作的

以互動方式探索這項發展背後的基礎技術。

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
互動式概念檢查+10 Points
AI Agents Quiz

An agent must create a draft calendar event for Tuesday at 2 p.m. Which evidence would establish the requested result?

接下來看什麼

Researchers, employers and policymakers will need to compare the with private enterprise deployments, actual usage, employment and wage data. The most important unresolved issue is whether tools that remove routine entry-level tasks also reduce opportunities for workers to develop expertise.

The ’s usefulness will depend on whether other researchers can access and inspect it, reproduce the classifications and test how quickly the public tool landscape changes. The source does not state the release’s license, documentation beyond the article or maintenance schedule.

Future evidence should connect public tooling with private deployments, worker use and measurable labor-market outcomes. Without those links, the ATE findings cannot show whether an occupation is economically threatened or whether a tool improves productivity.

The source highlights a particular risk around entry-level work: if routine tasks are automated, workers may have fewer opportunities to learn through those tasks before taking on specialized responsibilities. Whether organizations redesign training and career paths is an open question.

The analysis also warrants scrutiny because the occupational matches and category groupings rely partly on language-model judgments. Cohere reports a blind human check of 120 categories that pointed in the same direction, but the source does not provide enough detail here to assess the full validation process or error rate.

相關指引和測驗

人工智慧代理AI 的未來人工智慧模型解釋測試你所知道的—嘗試免費的人工智慧測驗在我們的詞彙表中尋找人工智慧術語關注 AI 模型發布追蹤器
覺得有用嗎?