Gesellschaftsführer

KI-Sicherheitsinstitute auf der ganzen Welt

AI safety institutes are government bodies that test and study advanced AI models.

  • 3 Minuten gelesen
  • Zuletzt aktualisiert
Auf dieser Seite3 Minuten gelesen
  1. Übersicht
  2. Tiefer Einblick
  3. Strategische Auswirkungen
  4. The Future of AI Safety Institutes Around the World
  5. Reale Umsetzung
  6. Risiken und Leitplanken
  7. Implementierungs-Roadmap
  8. Entdecken Sie weiter
  9. Häufig gestellte Fragen

Übersicht

They focus on risks such as cyberattacks, chemical and biological misuse, and loss of human control, often evaluating models before public release. They matter because they give governments independent technical expertise on frontier AI, although most depend on voluntary cooperation from companies rather than legal powers.

Tiefer Einblick

The UK set up its AI Safety Institute in November 2023, around the Bletchley Park AI Safety Summit. It grew out of the Frontier AI Taskforce and became one of the best-funded government AI evaluation teams, hiring researchers from leading labs. In February 2025 it was renamed the AI Security Institute. The new name reflected a narrower focus on national-security harms such as cyber misuse, chemical and biological risks, and crime, rather than broader issues such as bias. The US AI Safety Institute was created within NIST in late 2023, after President Biden's executive order on AI. In June 2025 the Commerce Department reorganized it as the Center for AI Standards and Innovation (CAISI). The mandate shifted toward voluntary standards, national-security-relevant evaluations, and assessing foreign AI systems, while dropping the "safety" branding. Japan, Singapore, Canada, South Korea, France and others have set up similar bodies. The EU's AI Office has some parallel functions, but it is a regulator. In November 2024 the International Network of AI Safety Institutes held its first meeting in San Francisco. Its founding members were Australia, Canada, the EU, France, Japan, Kenya, South Korea, Singapore, the UK and the US. A common misconception is that these institutes can block model releases. The UK and US bodies have no general statutory power to compel access or stop deployment. Their access comes from voluntary agreements, and their findings shape policy and company decisions without being binding. The EU AI Office is different: under the AI Act it can request information from providers of general-purpose AI models with systemic risk and enforce obligations on them.

Strategische Auswirkungen

Risiko und Sicherheit

Sowohl katastrophale als auch alltägliche Schäden durch KI hängen davon ab, wer die Risiken versteht und wer handeln kann.

Klarere Entscheidungen

Die öffentliche und berufliche Bildung bestimmt, ob eine starke Sicherheitspolitik politisch möglich ist.

Sich durch den Hype schneiden

Klare Erklärungen reduzieren die Vereinnahmung durch Hype, Labor-PR und vages Ethik-Theater.

The Future of AI Safety Institutes Around the World

The institutes' role is still changing. Several have moved from broad safety research toward national security, standards and industrial competitiveness, and their priorities may keep shifting with governments. Key open questions are whether any of them will get statutory powers, how their work will connect to regulators such as the EU AI Office, and whether international collaboration will deepen or split along geopolitical lines. Their long-term value may depend on shared, reproducible evaluation methods and enough funding to keep up with frontier labs.

Reale Umsetzung

The UK and US institutes jointly tested an upgraded Claude 3.5 Sonnet and OpenAI's o1 before deployment and published summaries of the results in late 2024.

The UK institute released Inspect, an open-source framework that researchers and other governments use to build and run AI evaluations.

In 2024, the US institute signed agreements with OpenAI and Anthropic that gave it access to major new models before and after public release.

Institutes from several countries have run joint testing exercises to compare methods, for example evaluating agent behavior in different languages.

Risiken und Leitplanken

  • Das existentielle Risiko wird als Science-Fiction behandelt, während sich die Fähigkeiten verstärken.

  • Verwechslung von Oberflächenproduktsicherheit mit Ausrichtung unter hoher Autonomie.

  • Nicht-englischsprachigen und nicht fachkundigen Zielgruppen stehen nur Quellen von geringer Qualität zur Verfügung.

Implementierungs-Roadmap

  1. Separate Risiken für Produktschäden, Missbrauch und Kontrollverlust/Fehlausrichtung.

  2. Fragen Sie, welche Beweise Ihre Sicht auf Zeitpläne und Schweregrad ändern würden.

  3. Bevorzugen Sie Primärquellen und konkrete Bewertungen gegenüber Marketingaussagen.

  4. Identifizieren Sie einen Aktionspfad: Karriere, Politik, Finanzierung oder Fähigkeiten – nicht nur Bewusstsein.

Entdecken Sie weiter

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the AI Safety Institutes Around the World quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

Quiz starten

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

Häufig gestellte Fragen

What is AI Safety Institutes Around the World?

AI safety institutes are government bodies that test and study advanced AI models. They focus on risks such as cyberattacks, chemical and biological misuse, and loss of human control, often evaluating models before public release. They matter because they give governments independent technical expertise on frontier AI, although most depend on voluntary cooperation from companies rather than legal powers.

What was the UK AI Safety Institute renamed in February 2025?

The new name, AI Security Institute, reflected a narrower focus on national-security risks such as cyber, chemical and biological misuse, and crime.

The US AI Safety Institute was reorganized into which body in 2025?

The Commerce Department turned it into CAISI, which focuses on standards, national-security evaluations and assessing foreign systems.

Where did the International Network of AI Safety Institutes hold its first meeting?

The first meeting was held in San Francisco in November 2024.

What legal power do the UK and US institutes generally have over model releases?

Their access comes from voluntary agreements, and their findings are advisory rather than binding.

What is Inspect?

Inspect organizes evaluations into datasets, solvers and scorers, and supports sandboxed agent testing.