Free AI library

Visual AI guidesFree forever.

133 plain-English guides, structured learning paths, and an open library — built by an independent 501(c)(3) nonprofit so anyone can understand modern AI.

133Free guides
1Topic tracks
~2 minPer guide
~4hReading time

Start here

Five outcome-based courses

Each course includes explicit outcomes, mapped competencies, practice activities, and an applied capstone.

Topic tracks

Browse by track

Jump into the area you care about. Every track has multiple plain-English guides.

Full library

All guides

133 of 1019 guides shown. Filter by track or search above.

Visual AI

Optical Character Recognition

Optical Character Recognition (OCR) turns images of text — scanned documents, photos of signs, PDFs — into machine-readable, editable text.

2 min readRead
Visual AI

Optical Flow

Optical flow estimates how each pixel moves between consecutive video frames, producing a dense map of motion vectors.

2 min readRead
Visual AI

Monocular Depth Estimation

Monocular depth estimation predicts how far away every pixel is from a single ordinary photo — no stereo camera, lidar, or depth sensor required.

2 min readRead
Visual AI

Latent Diffusion Models

Latent diffusion models generate images by running the diffusion process in a compressed latent space instead of raw pixels, slashing compute costs.

2 min readRead
Visual AI

ControlNet

ControlNet is an add-on that gives image-generation models precise structural control, letting you steer output with edges, poses, depth maps, or scribbles.

2 min readRead
Visual AI

Classifier-Free Guidance

Classifier-free guidance is the technique that makes diffusion models actually follow your prompt, trading some diversity for much stronger adherence.

2 min readRead
Visual AI

Visual SLAM

Visual SLAM lets a moving camera build a map of an unknown space while simultaneously tracking its own position inside that map.

2 min readRead
Visual AI

Sora and Text-to-Video

Sora is OpenAI's text-to-video model that turns a written prompt into a short, high-resolution video clip.

2 min readRead
Visual AI

Video Diffusion Models

Video diffusion models generate moving images by gradually turning random noise into coherent frames, extending the diffusion idea from pictures to time.

2 min readRead
Visual AI

Text-to-3D Generation

Text-to-3D generation turns a written prompt like 'a vintage leather armchair' into a full 3D model you can rotate, light, and drop into a game or scene.

2 min readRead
Visual AI

Differentiable Rendering

Differentiable rendering makes the process of turning a 3D scene into a 2D image fully differentiable, so you can compute gradients from the rendered pixels…

2 min readRead
Visual AI

U-Net Architecture

U-Net is a convolutional neural network shaped like a 'U' that excels at producing pixel-precise outputs, originally for biomedical image segmentation.

2 min readRead

Finished reading? Prove it.

Check what you learned with topic quizzes, then explore our structured courses and current certification requirements. Core guides remain free to read.