समाचार पर वापस जाएँ
उत्पादAI Understanding ब्रीफिंग

कॉग्निशन ने कम लागत वाले एआई एजेंट प्रदर्शन के लिए फ्यूजन हार्नेस जारी किया

GIGAZINE की रिपोर्ट है कि कॉग्निशन ने GPT-6 एस्ट्रा और Claude फैबल के लिए फ्यूजन जारी किया है, जो बेंचमार्क में प्रतिस्पर्धी एजेंट सिस्टम से मेल खाता है और रिपोर्ट की गई लागत को 39% तक कम करता है।

4 min readRead the linked source
Source-provided image accompanying Cognition releases Fusion harness for lower-cost AI agent performance
स्रोत संदर्भस्रोत रिकार्ड किया गया
प्रकाशक
gigazine.net
स्रोत लिंक
gigazine.nethttps://gigazine.net/gsc_news/en/20260914-cognition-fusion/
स्रोत प्रकार
लिंक्ड स्रोत - प्राथमिक-स्रोत स्थिति स्थापित नहीं की गई है।
प्रसंगइसे 60 सेकंड में समझें

यहां से प्रारंभ करें

प्रमुख शर्तें

एआई एजेंट
एक सॉफ़्टवेयर सिस्टम जो किसी लक्ष्य को प्राप्त करने के लिए अक्सर टूल और मेमोरी का उपयोग करके निरीक्षण कर सकता है, तर्क कर सकता है और कार्रवाई कर सकता है।
बेंचमार्क
मॉडल प्रदर्शन को मापने और तुलना करने के लिए उपयोग किया जाने वाला एक मानकीकृत परीक्षण या डेटासेट।
अनुमान
रनटाइम चरण जहां एक प्रशिक्षित मॉडल भविष्यवाणियां या आउटपुट उत्पन्न करता है।
स्वयं की जांच करोएआई एजेंट प्रश्नोत्तरी

क्या हुआ?

GIGAZINE reports that Cognition, the company behind Devin, released Fusion, a harness designed to improve how advanced AI models plan, execute, review, and recover from stalled tasks. The system pairs a high-end model for planning and review with a less expensive model for execution, while running two agents in parallel with separate contexts and tools.

GIGAZINE reports that Cognition developed and released Fusion as a harness for OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable. The article describes a harness as the layer that supplies models with operating instructions, tool-use rules, and conditional branching, potentially affecting whether an agent completes a task or becomes stuck in a loop.

According to GIGAZINE, Fusion combines a state-of-the-art model for planning and review with a cost-effective model for execution. Cognition recommended GPT-6 Astra and Claude Fable as the advanced models, and identified Cognition’s SWE-2 coding model as the cheaper execution option. The article says the combination performed best with Claude Fable 5.1.

GIGAZINE reports a score of 62.2 for Claude Fable 5.1 with Claude Code, compared with 61.7 for Claude Fable 5.1 combined with the lower-cost model and Fusion. The latter combination was reported as 36% cheaper. For GPT-6 Astra, the article says Fusion achieved performance equivalent to Codex while reducing costs by 39%.

The article attributes evaluations to AI analytics companies Artificial Analysis and Vals AI. It says Fusion runs two agents in parallel, each with independent context and tools, and exchanges only summaries, results, and feedback rather than the full conversation. GIGAZINE says this approach makes greater use of prompt caching. The source does not provide the full methodology, absolute prices, or availability terms.

स्रोत विवरण: gigazine.net ↗

यह क्यों मायने रखता है?

If the reported results hold, Fusion could reduce the cost of using frontier AI agents without a comparable loss in coding performance. That matters for developers and organizations running tool-using systems at scale, where model and repeated agent attempts can become major expenses. The results are reported by GIGAZINE and evaluated by Artificial Analysis and Vals AI, but the source does not provide enough methodology to independently assess the comparisons.

The reported results suggest that agent architecture can materially affect the economics of frontier models. A system that delegates execution to a cheaper model while retaining a stronger model for planning and review could lower the cost of software agents without requiring users to abandon higher-capability models.

Cost reductions are particularly relevant for coding agents, which may make many tool calls, repeat failed attempts, or maintain long contexts. If independently reproducible, the reported 36% and 39% reductions could influence how companies design production agent workflows and choose between proprietary harnesses.

The evidence remains limited in the supplied report. GIGAZINE provides selected scores and percentage savings but not the task set, run count, pricing assumptions, latency, failure rates, or comparison conditions. The claimed equivalence to Codex and the stated efficiency gains are therefore not independently confirmed here.

Interactive Mechanism

इंटरैक्टिव तंत्र: यह वास्तव में कैसे काम करता है

इस विकास के पीछे अंतर्निहित प्रौद्योगिकी का अंतःक्रियात्मक रूप से अन्वेषण करें।

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
इंटरएक्टिव कॉन्सेप्ट चेक+10 Points
AI Agents Quiz

An agent must create a draft calendar event for Tuesday at 2 p.m. Which evidence would establish the requested result?

आगे क्या देखना है

The key questions are who can access Fusion, which models and tools it supports, whether it is generally available, and how its costs are calculated. Further reporting should clarify the tasks, baselines, sample sizes, prompt-cache assumptions, and whether the results transfer beyond coding. The source also uses both “SWE-2” and “SWR-2” for the lower-cost model, an inconsistency that should be resolved.

Access conditions are not documented in the source. It is unknown whether Fusion is publicly downloadable, available through Cognition’s products, restricted to selected users, or offered through another commercial arrangement. Pricing and any required model subscriptions are also unknown.

Follow-up reporting should establish whether Fusion supports models beyond GPT-6 Astra and Claude Fable, whether users can supply their own tools and prompts, and how much of the savings depends on prompt caching or specific provider pricing.

The naming contains a source-level inconsistency: the article identifies the cheaper coding model as SWE-2 but later refers to “SWR-2.” That designation should be verified before treating the comparison as a precise product claim.

Independent testing should examine coding success rates, latency, reliability, context handling, and performance on tasks other than the reported . Parallel agents may also increase orchestration complexity or total tool activity even when model costs fall.

संबंधित मार्गदर्शिकाएँ एवं प्रश्नोत्तरी

एआई एजेंटएआई मॉडल की व्याख्याPrompt Engineeringआप जो जानते हैं उसका परीक्षण करें - निःशुल्क AI प्रश्नोत्तरी आज़माएँहमारी शब्दावली में एआई शब्द देखेंएआई मॉडल रिलीज ट्रैकर का पालन करें
क्या यह उपयोगी पाया गया?