Tiếp theoHướng dẫn tiếp theo
Kiểm tra sắc thái giới tính và nhận dạng khuôn mặt
xã hội
HƯỚNG DẪN xã hội
An AI audit is a structured evaluation of an AI system, or the organization running it, against defined criteria such as accuracy, fairness, safety, legal requirements or governance standards, carried out internally, by an independent third party, or under a regulator's mandate.
Third-party assurance matters because claims about AI systems are hard for outsiders to verify, and audits are becoming a main way regulators, buyers and the public check them.
AI audits differ in who does them and what they examine. Internal audits are run by the organization itself, ideally by a team separate from the builders; Raji and colleagues' 2020 framework proposed an end-to-end internal audit process across a system's development. External audits come from outsiders. Some are cooperative, with access to data and code; others are adversarial, like journalists or researchers testing a system through its public interface. Regulatory audits are required or performed under law. Methods fall into three broad groups. Governance or process audits check policies, documentation, roles and risk management. Technical audits test the model: performance across subgroups, robustness, security, and behaviour on edge cases, sometimes including red teaming. Outcome audits look at real-world effects, for example comparing decisions across groups or using sock-puppet accounts to probe a recommendation system. Conformity assessment is a related legal concept. Under the EU AI Act, providers of high-risk systems must show compliance before placing them on the market. For most standalone high-risk uses this is an internal control procedure, while certain biometric systems and AI in products already covered by EU safety laws may involve a notified body. New York City's Local Law 144 requires independent bias audits for automated employment decision tools, and the EU Digital Services Act requires independent audits of very large online platforms. Independence is the persistent weakness. Auditors are usually paid by the company being audited, may receive limited access, and often work without agreed standards for what counts as passing. Critics warn of audit-washing, where a narrow audit is presented as a clean bill of health. A common misconception is that an audit certifies a system as safe or fair in general; it only speaks to the criteria, scope and time period examined.
Những tác hại thảm khốc và thường ngày của AI đều phụ thuộc vào việc ai hiểu được rủi ro và ai có thể hành động.
Kiến thức công cộng và chuyên môn định hình liệu chính sách an toàn mạnh mẽ có khả thi về mặt chính trị hay không.
Những lời giải thích rõ ràng làm giảm sự thu hút bởi sự cường điệu, PR trong phòng thí nghiệm và sân khấu đạo đức mơ hồ.
An AI assurance industry is forming, including accounting firms, specialist startups, testing labs and certification bodies. Standards work, such as ISO/IEC 42001 for management systems and related standards for bodies that certify them, aims to make audits more comparable. Governments including the UK have published plans to grow AI assurance as a market. Key unresolved issues are auditor accreditation, access to model internals for external researchers, and who pays without creating conflicts of interest. Audits of general-purpose models are especially immature, since their uses are open-ended. Expect gradual professionalisation rather than a single agreed method in the near term.
An employer in New York City using an automated resume screening tool commissions an independent bias audit that reports selection rate impact ratios by sex and race or ethnicity categories, as the city's Local Law 144 requires.
Researchers query commercial face analysis services with a balanced set of faces and publish error rates by skin type and gender, an external audit carried out without the vendors' cooperation, like the 2018 Gender Shades study.
A software company seeks certification of its AI management system against ISO/IEC 42001 from an accredited certification body so enterprise customers can see its governance controls have been checked.
A very large online platform in the EU undergoes an annual independent audit of its risk management, including its recommender systems, under the Digital Services Act.
Xử lý rủi ro hiện hữu như khoa học viễn tưởng trong khi khả năng lại phức tạp.
Nhầm lẫn giữa an toàn sản phẩm bề mặt với sự liên kết dưới quyền tự chủ cao.
Chỉ để lại những khán giả không phải người Anh và không có chuyên môn với những nguồn chất lượng thấp.
Tách biệt các tác hại của sản phẩm, sử dụng sai và rủi ro mất kiểm soát/sai lệch.
Hỏi bằng chứng nào sẽ thay đổi quan điểm của bạn về thời gian và mức độ nghiêm trọng.
Ưu tiên các nguồn chính và đánh giá cụ thể hơn các tuyên bố tiếp thị.
Xác định một lộ trình hành động: sự nghiệp, chính sách, nguồn tài trợ hoặc kỹ năng - không chỉ là nhận thức.
Free newsletter
Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.
One email each weekday. Unsubscribe in one click. We never sell or share your address.
Test yourself
Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.
Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation
An AI audit is a structured evaluation of an AI system, or the organization running it, against defined criteria such as accuracy, fairness, safety, legal requirements or governance standards, carried out internally, by an independent third party, or under a regulator's mandate. Third-party assurance matters because claims about AI systems are hard for outsiders to verify, and audits are becoming a main way regulators, buyers and the public check them.
Adversarial external audits are done by outsiders, such as journalists or researchers, probing a system without cooperation or privileged access.
Process audits check how the organization manages AI, rather than testing the model directly.
Local Law 144 requires independent bias audits, including impact ratios by demographic categories.
The impact ratio compares each category's selection rate with that of the most selected category.
Most standalone high-risk uses follow internal control, while certain biometric systems and regulated products may involve a notified body.
Tiếp tục học hỏi
Đã chọn thêm hướng dẫn cho chủ đề này
Tiếp theoHướng dẫn tiếp theo
Kiểm tra sắc thái giới tính và nhận dạng khuôn mặt
xã hội