Free AI library

Visual AI guidesFree forever.

133 plain-English guides, structured learning paths, and an open library — built by an independent 501(c)(3) nonprofit so anyone can understand modern AI.

133Free guides
1Topic tracks
~2 minPer guide
~4hReading time

Start here

Five outcome-based courses

Each course includes explicit outcomes, mapped competencies, practice activities, and an applied capstone.

Topic tracks

Browse by track

Jump into the area you care about. Every track has multiple plain-English guides.

Full library

All guides

133 of 1019 guides shown. Filter by track or search above.

Visual AI

Neural Radiance Fields

Neural Radiance Fields (NeRF) reconstruct a full 3D scene from a handful of ordinary photos, letting you fly the camera to brand-new viewpoints.

2 min readRead
Visual AI

Gaussian Splatting

Gaussian Splatting represents a 3D scene as millions of tiny, colored, semi-transparent blobs that can be rendered in real time.

2 min readRead
Visual AI

Masked Autoencoders

Masked Autoencoders (MAE) are a self-supervised method that teaches a vision model to reconstruct images after most of the picture has been hidden.

2 min readRead
Visual AI

Image Captioning

Image captioning is the task of automatically generating a natural-language sentence that describes what is in a picture.

2 min readRead
Visual AI

Visual Question Answering

Visual Question Answering (VQA) lets a system answer free-form natural-language questions about an image, such as 'How many people are wearing hats?

2 min readRead
Visual AI

YOLO Real-Time Detection

YOLO (You Only Look Once) is a family of object detection models that find and label every object in an image with a single neural network pass, fast enough…

2 min readRead
Visual AI

Panoptic Segmentation

Panoptic segmentation gives every single pixel in an image a label, unifying 'what is this region' with 'which specific object is this.

2 min readRead
Visual AI

Human Pose Estimation

Human pose estimation detects the positions of body joints, such as elbows, knees, and shoulders, to build a digital skeleton of a person from images…

2 min readRead
Visual AI

FLUX Image Models

FLUX is a family of open text-to-image models from Black Forest Labs known for sharp detail, strong prompt-following, and surprisingly accurate rendered text.

2 min readRead
Visual AI

Residual Networks

Residual Networks (ResNets) are deep neural networks that add 'skip connections' letting layers learn small adjustments instead of full transformations.

2 min readRead
Visual AI

Region-Based CNNs

Region-Based CNNs (R-CNNs) are a family of object detectors that first propose candidate regions in an image, then use a CNN to classify and precisely box…

2 min readRead
Visual AI

Swin Transformer

The Swin Transformer is a vision Transformer that processes images in shifted, hierarchical windows, making attention efficient enough to scale across…

2 min readRead

Finished reading? Prove it.

Check what you learned with topic quizzes, then explore our structured courses and current certification requirements. Core guides remain free to read.