사회 가이드
What Research Says About AI and Workplace Productivity
Controlled field experiments show generative AI can make workers substantially faster and better at specific tasks such as customer support, writing and some coding, with the largest gains usually going to less experienced workers.
이 페이지에서4분 읽기
개요
The gains are uneven, though. AI can make people worse on tasks outside its abilities, and broad studies of whole workforces and economies so far find much smaller effects than lab and single-task studies do.
심층 분석
The best-known evidence comes from field and controlled experiments on well-defined tasks. Erik Brynjolfsson, Danielle Li and Lindsey Raymond studied more than 5,000 customer support agents who were given an AI assistant. On average they resolved about 14 percent more issues per hour. Novice and lower-skilled agents gained around 34 percent, and the most experienced agents gained little. The researchers argue the AI spread the know-how of top performers to everyone else. Shakked Noy and Whitney Zhang, publishing in Science in 2023, found that ChatGPT cut the time college-educated professionals spent on writing tasks by about 40 percent and raised rated quality by about 18 percent. A GitHub Copilot experiment found developers finished a set programming task about 55 percent faster. The 2023 Boston Consulting Group study with Harvard and other researchers added an important caveat, which it called the 'jagged technological frontier'. On tasks inside the AI's abilities, consultants with GPT-4 finished more tasks, faster and at higher quality. On a task designed to sit outside those abilities, they were less likely to get the right answer than consultants without AI, because they trusted plausible but wrong output. More recent findings are more modest. In a 2025 randomized trial by METR, experienced developers working in their own large codebases were about 19 percent slower with AI tools while believing they had been faster. Studies of whole labor markets, such as research on Danish workers, have found small effects on earnings and hours so far. Why the gap? Experiments isolate tasks that suit AI. Real jobs mix many tasks, time saved is not always reused productively, and organizations need to redesign workflows before gains show up. Economists saw a similar lag in measured productivity after electricity and computers arrived. The common misconception is that one headline percentage applies to every job.
전략적 영향
위험과 안전
치명적인 AI 피해와 일상적인 AI 피해는 누가 위험을 이해하고 누가 조치를 취할 수 있는지에 따라 달라집니다.
더 명확한 결정들
공공 및 전문 지식은 강력한 안전 정책이 정치적으로 가능한지 여부를 결정합니다.
과장된 과장을 뚫고 나가기
명확한 설명은 과대광고, 연구실 홍보, 모호한 윤리 연극에 의한 포착을 줄입니다.
The Future of What Research Says About AI and Workplace Productivity
The evidence base is growing quickly, and the models studied in 2023 are already outdated, so results may shift as tools, training and workflows improve. The key open questions are whether task-level gains add up to firm-level and national productivity growth, whether benefits keep concentrating among less experienced workers, and how deskilling or over-reliance affects people in the long run. Careful researchers expect a lag like the one seen with earlier general-purpose technologies. Treat both very large and near-zero headline numbers with caution until more long-term, economy-wide data exists.
실제 구현
In a large customer support study, agents given an AI assistant resolved more issues per hour on average, and the newest, least experienced agents improved the most.
In an experiment with professional writing tasks, participants using ChatGPT finished faster and produced work that graders rated higher in quality.
Consultants at Boston Consulting Group did better with GPT-4 on tasks inside the AI's abilities, but were more likely to reach wrong answers on a task deliberately designed to fall outside them.
Experienced open-source developers in a 2025 randomized trial by METR took longer to finish tasks with AI tools, even though they believed the tools had made them faster.
위험 및 가드레일
실존적 위험을 공상과학처럼 다루면서 능력을 합성합니다.
높은 자율성 하에서 정렬과 표면 제품 안전성을 혼동합니다.
영어가 아니거나 전문가가 아닌 청중에게는 품질이 낮은 소스만 남겨 둡니다.
구현 로드맵
제품 손상, 오용, 통제력 상실/잘못 정렬 위험을 분리합니다.
일정과 심각도에 대한 귀하의 견해를 바꿀 수 있는 증거가 무엇인지 물어보십시오.
마케팅 주장보다 기본 소스와 구체적인 평가를 선호하세요.
인식뿐만 아니라 경력, 정책, 자금 조달 또는 기술 등 하나의 행동 경로를 식별하십시오.
계속 탐색하세요
Free newsletter
Get the daily AI briefing
Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.
One email each weekday. Unsubscribe in one click. We never sell or share your address.
Test yourself
Take the What Research Says About AI and Workplace Productivity quiz
Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.
Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation
자주 묻는 질문
What is What Research Says About AI and Workplace Productivity?
Controlled field experiments show generative AI can make workers substantially faster and better at specific tasks such as customer support, writing and some coding, with the largest gains usually going to less experienced workers. The gains are uneven, though. AI can make people worse on tasks outside its abilities, and broad studies of whole workforces and economies so far find much smaller effects than lab and single-task studies do.
In the Brynjolfsson, Li and Raymond customer support study, which agents gained the most from AI?
Novice and lower-skilled agents improved by around 34 percent, while the most experienced agents gained little. The average gain was about 14 percent.
What did Noy and Zhang's 2023 Science study find about ChatGPT on writing tasks?
Professionals finished writing tasks much faster, and graders rated the quality higher.
What does the 'jagged technological frontier' describe?
AI's abilities are uneven. Inside the frontier it helps, and outside it people who trust it can do worse.
In the BCG study, what happened on the task designed to fall outside AI's abilities?
Consultants using GPT-4 on that task were less likely to be correct, because they trusted plausible but wrong output.
What did METR's 2025 randomized trial with experienced developers find?
Measured completion times were slower with AI, yet the developers believed it had sped them up. This shows how unreliable self-reports can be.
계속 학습하세요
관련 가이드
이 주제에 대해 선택된 추가 가이드