AI Foundations
Understand what AI is, how systems learn, where they fail, and how to judge claims without hype.
Free AI library
117 plain-English guides, structured learning paths, and an open library — built by an independent 501(c)(3) nonprofit so anyone can understand modern AI.
Start here
Each course includes explicit outcomes, mapped competencies, practice activities, and an applied capstone.
Understand what AI is, how systems learn, where they fail, and how to judge claims without hype.
Use AI productively while protecting privacy, checking outputs, and preserving human accountability.
Evaluate workplace use cases, run safe pilots, measure value, and communicate changes responsibly.
Analyze AI systems through rights, equity, governance, safety, and public-interest outcomes.
Understand language models, retrieval, agents, evaluation, cost, and deployment safeguards through practical system design.
Topic tracks
Jump into the area you care about. Every track has multiple plain-English guides.
Full library
117 of 1019 guides shown. Filter by track or search above.
Cover song identification detects when two very different-sounding recordings are actually the same underlying song — a live acoustic version, a remix…
Audio AIDDSP (Differentiable Digital Signal Processing) fuses classic synthesizer building blocks with neural networks, so deep learning can control oscillators…
Audio AIVITS is a text-to-speech model that turns text directly into raw audio waveforms in a single trained system, skipping the usual two-stage pipeline.
Audio AIFastSpeech generates an entire speech spectrogram in parallel rather than one frame at a time, making synthesis dramatically faster and more stable.
Audio AINaturalSpeech is a line of Microsoft TTS research aiming for human-level speech quality, with later versions using latent diffusion to generate rich, natural…
Audio AISpeech Emotion Recognition (SER) is AI that detects a speaker's emotional state — anger, joy, sadness, frustration — from the sound of their voice, not just…
Audio AIMoshi is an open-source, real-time voice AI from Kyutai that talks and listens at the same time — full-duplex — instead of taking strict turns.
Audio AISpeech-to-Speech Translation (S2ST) takes spoken words in one language and produces spoken words in another — ideally preserving the speaker's voice, tone…
Audio AIHiFi-GAN is a generative-adversarial vocoder that turns a mel-spectrogram into a raw audio waveform almost instantly, producing studio-quality speech far…
Audio AIGrapheme-to-phoneme (G2P) conversion translates written letters into the sounds a speech system should actually pronounce.
Audio AIText normalization is the front-end step that rewrites raw written text into fully spoken-out words before a speech system says it.
Audio AISoundStream is Google's end-to-end neural audio codec that compresses speech and music to extremely low bitrates while preserving quality.
Check what you learned with topic quizzes, then explore our structured courses and current certification requirements. Core guides remain free to read.