Free AI library

Visual AI guidesFree forever.

133 plain-English guides, structured learning paths, and an open library — built by an independent 501(c)(3) nonprofit so anyone can understand modern AI.

133Free guides
1Topic tracks
~2 minPer guide
~4hReading time

Start here

Five outcome-based courses

Each course includes explicit outcomes, mapped competencies, practice activities, and an applied capstone.

Topic tracks

Browse by track

Jump into the area you care about. Every track has multiple plain-English guides.

Full library

All guides

133 of 1019 guides shown. Filter by track or search above.

Visual AI

Video Understanding

Video understanding analyzes visual and sometimes audio information across time.

2 min readRead
Visual AI

Multimodal Search

Multimodal search retrieves information across forms such as text, images, audio, and video.

2 min readRead
Visual AI

Document AI

Document AI extracts and interprets information from files such as forms, reports, invoices, and scanned pages.

2 min readRead
Visual AI

Wasserstein GAN

Wasserstein GAN (WGAN) is a redesign of the GAN training objective that uses the Wasserstein distance instead of the original min-max loss.

2 min readRead
Visual AI

Conditional GANs

Conditional GANs (cGANs) extend ordinary GANs by feeding extra information, like a class label or text, into both the generator and discriminator.

2 min readRead
Visual AI

Pix2Pix Image-to-Image Translation

Pix2Pix is a conditional GAN that learns to translate one type of image into another, such as turning a sketch into a photo or a map into a satellite view.

2 min readRead
Visual AI

Image Colorization

Image colorization uses AI to add plausible, realistic color to black-and-white photos and film.

2 min readRead
Visual AI

Structure from Motion

Structure from Motion (SfM) reconstructs 3D scene geometry and camera positions from a set of overlapping 2D photos taken from different viewpoints.

2 min readRead
Visual AI

Multi-View Stereo

Multi-View Stereo (MVS) takes many calibrated photos of a scene and produces a dense 3D reconstruction by estimating depth at nearly every pixel.

2 min readRead
Visual AI

DDPM and DDIM Samplers

DDPM and DDIM are two ways to run the reverse process of a diffusion model, turning random noise into an image step by step.

2 min readRead
Visual AI

Score-Based Generative Models

Score-based generative models create data by learning the gradient of the data distribution — the direction that makes any noisy sample look more like real…

2 min readRead
Visual AI

Fréchet Inception Distance

Fréchet Inception Distance (FID) is the standard metric for judging how realistic and varied a set of generated images is.

2 min readRead

Finished reading? Prove it.

Check what you learned with topic quizzes, then explore our structured courses and current certification requirements. Core guides remain free to read.