社会ガイド

AI Audits and Third-Party Assurance

An AI audit is a structured evaluation of an AI system, or the organization running it, against defined criteria such as accuracy, fairness, safety, legal requirements or governance standards, carried out internally, by an independent third party, or under a regulator's mandate.

  • 4 分で読めます
  • 最終更新日
このページでは4 分で読めます
  1. 概要
  2. ディープダイブ
  3. 戦略的影響
  4. The Future of AI Audits and Third-Party Assurance
  5. 現実世界の実装
  6. リスクとガードレール
  7. 実装ロードマップ
  8. 探検を続けましょう
  9. よくある質問

概要

Third-party assurance matters because claims about AI systems are hard for outsiders to verify, and audits are becoming a main way regulators, buyers and the public check them.

ディープダイブ

AI audits differ in who does them and what they examine. Internal audits are run by the organization itself, ideally by a team separate from the builders; Raji and colleagues' 2020 framework proposed an end-to-end internal audit process across a system's development. External audits come from outsiders. Some are cooperative, with access to data and code; others are adversarial, like journalists or researchers testing a system through its public interface. Regulatory audits are required or performed under law. Methods fall into three broad groups. Governance or process audits check policies, documentation, roles and risk management. Technical audits test the model: performance across subgroups, robustness, security, and behaviour on edge cases, sometimes including red teaming. Outcome audits look at real-world effects, for example comparing decisions across groups or using sock-puppet accounts to probe a recommendation system. Conformity assessment is a related legal concept. Under the EU AI Act, providers of high-risk systems must show compliance before placing them on the market. For most standalone high-risk uses this is an internal control procedure, while certain biometric systems and AI in products already covered by EU safety laws may involve a notified body. New York City's Local Law 144 requires independent bias audits for automated employment decision tools, and the EU Digital Services Act requires independent audits of very large online platforms. Independence is the persistent weakness. Auditors are usually paid by the company being audited, may receive limited access, and often work without agreed standards for what counts as passing. Critics warn of audit-washing, where a narrow audit is presented as a clean bill of health. A common misconception is that an audit certifies a system as safe or fair in general; it only speaks to the criteria, scope and time period examined.

戦略的影響

リスクと安全性

AI による壊滅的な被害も日常的な被害も、誰がリスクを理解し、誰が行動できるかにかかっています。

より明確な判決

国民と専門家のリテラシーは、強力な安全政策が政治的に可能かどうかを左右します。

誇大広告を打ち破る

明確な説明は、誇大広告、研究室の PR、曖昧な倫理劇場に囚われることを減らします。

The Future of AI Audits and Third-Party Assurance

An AI assurance industry is forming, including accounting firms, specialist startups, testing labs and certification bodies. Standards work, such as ISO/IEC 42001 for management systems and related standards for bodies that certify them, aims to make audits more comparable. Governments including the UK have published plans to grow AI assurance as a market. Key unresolved issues are auditor accreditation, access to model internals for external researchers, and who pays without creating conflicts of interest. Audits of general-purpose models are especially immature, since their uses are open-ended. Expect gradual professionalisation rather than a single agreed method in the near term.

現実世界の実装

An employer in New York City using an automated resume screening tool commissions an independent bias audit that reports selection rate impact ratios by sex and race or ethnicity categories, as the city's Local Law 144 requires.

Researchers query commercial face analysis services with a balanced set of faces and publish error rates by skin type and gender, an external audit carried out without the vendors' cooperation, like the 2018 Gender Shades study.

A software company seeks certification of its AI management system against ISO/IEC 42001 from an accredited certification body so enterprise customers can see its governance controls have been checked.

A very large online platform in the EU undergoes an annual independent audit of its risk management, including its recommender systems, under the Digital Services Act.

リスクとガードレール

  • 能力が複雑になる一方で、実存的なリスクを SF として扱います。

  • 高度な自律性の下での調整による表面製品の安全性を混乱させる。

  • 英語以外や専門家ではない聴衆には、低品質の情報源しか提供されません。

実装ロードマップ

  1. 製品の危害、誤使用、制御不能/調整不良のリスクを分離します。

  2. どのような証拠がタイムラインと重大度についてのあなたの見方を変えるかを尋ねてください。

  3. マーケティング上の主張よりも、一次情報源と具体的な評価を優先します。

  4. 意識だけでなく、キャリア、政策、資金、スキルなど、行動経路を 1 つ特定します。

探検を続けましょう

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the AI Audits and Third-Party Assurance quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

クイズを開始する

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

よくある質問

What is AI Audits and Third-Party Assurance?

An AI audit is a structured evaluation of an AI system, or the organization running it, against defined criteria such as accuracy, fairness, safety, legal requirements or governance standards, carried out internally, by an independent third party, or under a regulator's mandate. Third-party assurance matters because claims about AI systems are hard for outsiders to verify, and audits are becoming a main way regulators, buyers and the public check them.

Which describes an adversarial external audit?

Adversarial external audits are done by outsiders, such as journalists or researchers, probing a system without cooperation or privileged access.

What does a governance or process audit mainly examine?

Process audits check how the organization manages AI, rather than testing the model directly.

What does NYC Local Law 144 require for automated employment decision tools?

Local Law 144 requires independent bias audits, including impact ratios by demographic categories.

How is the impact ratio used in NYC bias audits calculated?

The impact ratio compares each category's selection rate with that of the most selected category.

Under the EU AI Act, how do most standalone high-risk systems undergo conformity assessment?

Most standalone high-risk uses follow internal control, while certain biometric systems and regulated products may involve a notified body.