คู่มือสังคม

Amazon's Scrapped AI Recruiting Tool

Amazon's AI recruiting tool was an experimental system, begun around 2014, that rated job applicants' résumés from one to five stars by learning from ten years of past résumés.

  • อ่าน 4 นาที
  • อัปเดตล่าสุด
บนหน้านี้อ่าน 4 นาที
  1. ภาพรวม
  2. เจาะลึก
  3. ผลกระทบเชิงกลยุทธ์
  4. The Future of Amazon's Scrapped AI Recruiting Tool
  5. การใช้งานจริงในโลกแห่งความเป็นจริง
  6. ความเสี่ยงและรั้ว
  7. แผนงานการดำเนินงาน
  8. สำรวจต่อไป
  9. คำถามที่พบบ่อย

ภาพรวม

Amazon abandoned it after finding that it penalized résumés associated with women. It is now the standard example of how models trained on historical decisions reproduce past discrimination.

เจาะลึก

Reuters reported in October 2018 that Amazon had built, and then quietly scrapped, a machine learning tool for ranking job candidates. A team began the work around 2014. The goal was to feed in a stack of résumés and get back the top few candidates to hire. The models were trained on résumés submitted to the company over ten years and learned which patterns tended to go with people who were hired. Most applicants for technical roles were men, so the patterns the model learned were male-coded. By 2015, the team saw that the system was not rating candidates for software developer and other technical jobs in a gender-neutral way. According to Reuters, it penalized résumés containing the word "women's," as in "women's chess club captain," and downgraded graduates of two all-women's colleges. It also favored verbs such as "executed" and "captured," which appeared more often on men's résumés. Engineers edited the models to be neutral to those particular terms. That gave no guarantee the system would not find other ways to sort candidates that were just as discriminatory. The project lost support and the team was disbanded by early 2017. Amazon said the tool was never used by its recruiters to evaluate candidates. Reuters reported that recruiters looked at its recommendations but did not rely on them alone. The key lesson is that the training label was the problem. The model accurately predicted whom Amazon had hired before, and those past hiring decisions contained bias. A common misconception is that the algorithm "became sexist" on its own. In fact it learned exactly what the data rewarded. Another is that removing gender fixes the problem. Text contains many correlated proxies, and a flexible model will find them. The case now comes up in regulatory debates, including New York City's bias audit law for automated hiring tools and the EU AI Act's classification of employment systems as high risk.

ผลกระทบเชิงกลยุทธ์

ความเสี่ยงและความปลอดภัย

ความเสียหายที่เกิดจาก AI ที่เป็นหายนะและเกิดขึ้นทุกวันนั้นขึ้นอยู่กับว่าใครเข้าใจความเสี่ยงและใครสามารถดำเนินการได้

การตัดสินใจที่ชัดเจนยิ่งขึ้น

ความรู้สาธารณะและวิชาชีพเป็นตัวกำหนดว่านโยบายความปลอดภัยที่เข้มงวดจะเป็นไปได้ทางการเมืองหรือไม่

ตัดผ่านกระแสโฆษณาชวนเชื่อ

คำอธิบายที่ชัดเจนช่วยลดการจับภาพโดยการโฆษณาเกินจริง การประชาสัมพันธ์ในห้องปฏิบัติการ และการแสดงจริยธรรมที่คลุมเครือ

The Future of Amazon's Scrapped AI Recruiting Tool

Hiring is moving from keyword filters to large language models that summarize and rank candidates, and this raises the same risk at larger scale. Regulators are responding. New York City requires bias audits for automated employment decision tools, and the EU AI Act classifies recruitment systems as high risk, with obligations being phased in. In the US, existing anti-discrimination law such as Title VII still applies to algorithmic tools, even as federal agency guidance has shifted. Audit methods are still maturing, and a passing audit on one dataset does not guarantee fairness elsewhere. The lasting lesson of the Amazon case is to question what a model is trained to predict before trusting how well it predicts it.

การใช้งานจริงในโลกแห่งความเป็นจริง

A résumé listing "captain of the women's chess club" would score lower, because in the training data the word "women's" rarely appeared on the résumés of people who were hired.

An employer in New York City using an automated tool to screen candidates must commission an independent bias audit and publish a summary under Local Law 144, a requirement shaped by cases like Amazon's.

A vendor strips gender fields and gendered words from a screening model. A disparate impact test then shows women still pass at under four-fifths the rate of men, because of proxies such as hobbies or college names.

A company retrains a ranking model to predict "successful employees" and finds that the label itself reflects biased promotion decisions. It switches to structured, job-related skills assessments as the training target.

ความเสี่ยงและรั้ว

  • การรักษาความเสี่ยงที่มีอยู่เป็นไซไฟในขณะที่สารประกอบความสามารถ

  • ความปลอดภัยของผลิตภัณฑ์พื้นผิวที่สับสนด้วยการจัดตำแหน่งภายใต้ความเป็นอิสระสูง

  • ปล่อยให้ผู้ชมที่ไม่ใช่ภาษาอังกฤษและไม่ใช่ผู้เชี่ยวชาญเหลือเพียงแหล่งข้อมูลคุณภาพต่ำ

แผนงานการดำเนินงาน

  1. แยกอันตรายของผลิตภัณฑ์ การใช้ในทางที่ผิด และความเสี่ยงในการสูญเสียการควบคุม/การวางแนวที่ไม่ถูกต้อง

  2. ถามว่าหลักฐานใดที่จะเปลี่ยนมุมมองของคุณเกี่ยวกับลำดับเวลาและความรุนแรง

  3. ชอบแหล่งที่มาหลักและการประเมินที่เป็นรูปธรรมมากกว่าคำกล่าวอ้างทางการตลาด

  4. ระบุเส้นทางการดำเนินการเส้นทางเดียว: อาชีพ นโยบาย เงินทุน หรือทักษะ ไม่ใช่แค่ความตระหนักรู้เท่านั้น

สำรวจต่อไป

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the Amazon's Scrapped AI Recruiting Tool quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

เริ่มแบบทดสอบ

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

คำถามที่พบบ่อย

What is Amazon's Scrapped AI Recruiting Tool?

Amazon's AI recruiting tool was an experimental system, begun around 2014, that rated job applicants' résumés from one to five stars by learning from ten years of past résumés. Amazon abandoned it after finding that it penalized résumés associated with women. It is now the standard example of how models trained on historical decisions reproduce past discrimination.

What data was Amazon's recruiting model trained on?

The model learned from about ten years of résumés submitted to Amazon, and most applicants for technical roles were men.

Which word did Reuters report the model penalized?

Résumés containing "women's," as in "women's chess club captain," were downgraded. "Executed" and "captured" were favored, not penalized.

Why did editing the model to ignore specific gendered terms fail to solve the problem?

Blocking a few words does not remove the correlated signals spread through the text, so the model could still discriminate through other proxies.

What is the central lesson about the training label in this case?

The model predicted who had been hired before. Because past hiring was skewed, the label encoded that skew and the model reproduced it.

Which verbs did the model reportedly favor because they appeared more often on men's résumés?

Reuters reported that the model favored "executed" and "captured," which were more common on men's résumés, an example of subtle proxy features.