A seguirPróximo guia
Gender Shades and Facial Recognition Bias Audits
Sociedade
GUIA DA SOCIEDADE
An AI audit is a structured evaluation of an AI system, or the organization running it, against defined criteria such as accuracy, fairness, safety, legal requirements or governance standards, carried out internally, by an independent third party, or under a regulator's mandate.
Third-party assurance matters because claims about AI systems are hard for outsiders to verify, and audits are becoming a main way regulators, buyers and the public check them.
AI audits differ in who does them and what they examine. Internal audits are run by the organization itself, ideally by a team separate from the builders; Raji and colleagues' 2020 framework proposed an end-to-end internal audit process across a system's development. External audits come from outsiders. Some are cooperative, with access to data and code; others are adversarial, like journalists or researchers testing a system through its public interface. Regulatory audits are required or performed under law. Methods fall into three broad groups. Governance or process audits check policies, documentation, roles and risk management. Technical audits test the model: performance across subgroups, robustness, security, and behaviour on edge cases, sometimes including red teaming. Outcome audits look at real-world effects, for example comparing decisions across groups or using sock-puppet accounts to probe a recommendation system. Conformity assessment is a related legal concept. Under the EU AI Act, providers of high-risk systems must show compliance before placing them on the market. For most standalone high-risk uses this is an internal control procedure, while certain biometric systems and AI in products already covered by EU safety laws may involve a notified body. New York City's Local Law 144 requires independent bias audits for automated employment decision tools, and the EU Digital Services Act requires independent audits of very large online platforms. Independence is the persistent weakness. Auditors are usually paid by the company being audited, may receive limited access, and often work without agreed standards for what counts as passing. Critics warn of audit-washing, where a narrow audit is presented as a clean bill of health. A common misconception is that an audit certifies a system as safe or fair in general; it only speaks to the criteria, scope and time period examined.
Os danos catastróficos e diários da IA dependem de quem entende os riscos e de quem pode agir.
A literacia pública e profissional determina se uma política de segurança forte é politicamente possível.
Explicações claras reduzem a captura por exageros, relações públicas de laboratório e teatro de ética vaga.
An AI assurance industry is forming, including accounting firms, specialist startups, testing labs and certification bodies. Standards work, such as ISO/IEC 42001 for management systems and related standards for bodies that certify them, aims to make audits more comparable. Governments including the UK have published plans to grow AI assurance as a market. Key unresolved issues are auditor accreditation, access to model internals for external researchers, and who pays without creating conflicts of interest. Audits of general-purpose models are especially immature, since their uses are open-ended. Expect gradual professionalisation rather than a single agreed method in the near term.
An employer in New York City using an automated resume screening tool commissions an independent bias audit that reports selection rate impact ratios by sex and race or ethnicity categories, as the city's Local Law 144 requires.
Researchers query commercial face analysis services with a balanced set of faces and publish error rates by skin type and gender, an external audit carried out without the vendors' cooperation, like the 2018 Gender Shades study.
A software company seeks certification of its AI management system against ISO/IEC 42001 from an accredited certification body so enterprise customers can see its governance controls have been checked.
A very large online platform in the EU undergoes an annual independent audit of its risk management, including its recommender systems, under the Digital Services Act.
Tratar o risco existencial como ficção científica enquanto aumenta a capacidade.
Confundir segurança do produto de superfície com alinhamento sob alta autonomia.
Deixando o público não-inglês e não especializado com apenas fontes de baixa qualidade.
Separe os riscos de danos ao produto, uso indevido e perda de controle/desalinhamento.
Pergunte quais evidências mudariam sua visão sobre prazos e gravidade.
Prefira fontes primárias e avaliações concretas em vez de afirmações de marketing.
Identifique um caminho de ação: carreira, política, financiamento ou habilidades – não apenas conscientização.
Free newsletter
Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.
One email each weekday. Unsubscribe in one click. We never sell or share your address.
Test yourself
Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.
Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation
An AI audit is a structured evaluation of an AI system, or the organization running it, against defined criteria such as accuracy, fairness, safety, legal requirements or governance standards, carried out internally, by an independent third party, or under a regulator's mandate. Third-party assurance matters because claims about AI systems are hard for outsiders to verify, and audits are becoming a main way regulators, buyers and the public check them.
Adversarial external audits are done by outsiders, such as journalists or researchers, probing a system without cooperation or privileged access.
Process audits check how the organization manages AI, rather than testing the model directly.
Local Law 144 requires independent bias audits, including impact ratios by demographic categories.
The impact ratio compares each category's selection rate with that of the most selected category.
Most standalone high-risk uses follow internal control, while certain biometric systems and regulated products may involve a notified body.
Continue aprendendo
Mais guias escolhidos para este tópico
A seguirPróximo guia
Gender Shades and Facial Recognition Bias Audits
Sociedade