Societate GHID

Institutele de siguranță AI din întreaga lume

AI safety institutes are government bodies that test and study advanced AI models.

  • 3 minute de citit
  • Ultima actualizare
Pe această pagină3 minute de citit
  1. Prezentare generală
  2. Scufundare în profunzime
  3. Impact strategic
  4. The Future of AI Safety Institutes Around the World
  5. Implementare în lumea reală
  6. Riscuri și balustrade
  7. Foaia de parcurs de implementare
  8. Continuați să explorați
  9. Întrebări frecvente

Prezentare generală

They focus on risks such as cyberattacks, chemical and biological misuse, and loss of human control, often evaluating models before public release. They matter because they give governments independent technical expertise on frontier AI, although most depend on voluntary cooperation from companies rather than legal powers.

Scufundare în profunzime

The UK set up its AI Safety Institute in November 2023, around the Bletchley Park AI Safety Summit. It grew out of the Frontier AI Taskforce and became one of the best-funded government AI evaluation teams, hiring researchers from leading labs. In February 2025 it was renamed the AI Security Institute. The new name reflected a narrower focus on national-security harms such as cyber misuse, chemical and biological risks, and crime, rather than broader issues such as bias. The US AI Safety Institute was created within NIST in late 2023, after President Biden's executive order on AI. In June 2025 the Commerce Department reorganized it as the Center for AI Standards and Innovation (CAISI). The mandate shifted toward voluntary standards, national-security-relevant evaluations, and assessing foreign AI systems, while dropping the "safety" branding. Japan, Singapore, Canada, South Korea, France and others have set up similar bodies. The EU's AI Office has some parallel functions, but it is a regulator. In November 2024 the International Network of AI Safety Institutes held its first meeting in San Francisco. Its founding members were Australia, Canada, the EU, France, Japan, Kenya, South Korea, Singapore, the UK and the US. A common misconception is that these institutes can block model releases. The UK and US bodies have no general statutory power to compel access or stop deployment. Their access comes from voluntary agreements, and their findings shape policy and company decisions without being binding. The EU AI Office is different: under the AI Act it can request information from providers of general-purpose AI models with systemic risk and enforce obligations on them.

Impact strategic

Risc și siguranță

Daunele catastrofale și cotidiene ale IA depind de cine înțelege riscurile și cine poate acționa.

Decizii mai clare

Educația publică și profesională influențează dacă o politică puternică de siguranță este posibilă din punct de vedere politic.

Tăierea hype-ului

Explicațiile clare reduc captarea de hype, PR de laborator și teatrul vag de etică.

The Future of AI Safety Institutes Around the World

The institutes' role is still changing. Several have moved from broad safety research toward national security, standards and industrial competitiveness, and their priorities may keep shifting with governments. Key open questions are whether any of them will get statutory powers, how their work will connect to regulators such as the EU AI Office, and whether international collaboration will deepen or split along geopolitical lines. Their long-term value may depend on shared, reproducible evaluation methods and enough funding to keep up with frontier labs.

Implementare în lumea reală

The UK and US institutes jointly tested an upgraded Claude 3.5 Sonnet and OpenAI's o1 before deployment and published summaries of the results in late 2024.

The UK institute released Inspect, an open-source framework that researchers and other governments use to build and run AI evaluations.

In 2024, the US institute signed agreements with OpenAI and Anthropic that gave it access to major new models before and after public release.

Institutes from several countries have run joint testing exercises to compare methods, for example evaluating agent behavior in different languages.

Riscuri și balustrade

  • Tratarea riscului existențial ca SF în timp ce capacitatea se agravează.

  • Confuză siguranța produsului de suprafață cu alinierea sub autonomie ridicată.

  • Lăsând audiențe non-engleze și neexperte doar surse de calitate scăzută.

Foaia de parcurs de implementare

  1. Separați riscurile de deteriorare a produsului, utilizare greșită și pierderea controlului / dezaliniere.

  2. Întrebați ce dovezi v-ar schimba punctul de vedere cu privire la termene și severitate.

  3. Preferați sursele primare și evaluările concrete față de afirmațiile de marketing.

  4. Identificați o singură cale de acțiune: carieră, politică, finanțare sau abilități - nu numai conștientizare.

Continuați să explorați

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the AI Safety Institutes Around the World quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

Quiz Start

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

Întrebări frecvente

What is AI Safety Institutes Around the World?

AI safety institutes are government bodies that test and study advanced AI models. They focus on risks such as cyberattacks, chemical and biological misuse, and loss of human control, often evaluating models before public release. They matter because they give governments independent technical expertise on frontier AI, although most depend on voluntary cooperation from companies rather than legal powers.

What was the UK AI Safety Institute renamed in February 2025?

The new name, AI Security Institute, reflected a narrower focus on national-security risks such as cyber, chemical and biological misuse, and crime.

The US AI Safety Institute was reorganized into which body in 2025?

The Commerce Department turned it into CAISI, which focuses on standards, national-security evaluations and assessing foreign systems.

Where did the International Network of AI Safety Institutes hold its first meeting?

The first meeting was held in San Francisco in November 2024.

What legal power do the UK and US institutes generally have over model releases?

Their access comes from voluntary agreements, and their findings are advisory rather than binding.

What is Inspect?

Inspect organizes evaluations into datasets, solvers and scorers, and supports sandboxed agent testing.