Free AI library

Visual AI guidesFree forever.

133 plain-English guides, structured learning paths, and an open library — built by an independent 501(c)(3) nonprofit so anyone can understand modern AI.

133Free guides
1Topic tracks
~2 minPer guide
~4hReading time

Start here

Five outcome-based courses

Each course includes explicit outcomes, mapped competencies, practice activities, and an applied capstone.

Topic tracks

Browse by track

Jump into the area you care about. Every track has multiple plain-English guides.

Full library

All guides

133 of 1019 guides shown. Filter by track or search above.

Visual AI

LoRA Sliders for Image Editing

LoRA sliders are tiny add-on modules that give you a continuous dial to push a single attribute of an image up or down, like age, smile, or rust, without…

2 min readRead
Visual AI

InstructPix2Pix Instruction Editing

InstructPix2Pix lets you edit a photo by typing a plain command like 'make it winter' or 'turn the cat into a dog', no masks or selection tools required.

2 min readRead
Visual AI

Prompt-to-Prompt Cross-Attention Editing

Prompt-to-Prompt edits a generated image by tweaking its text prompt while reusing the model's internal attention maps, so changing one word swaps…

2 min readRead
Visual AI

ESRGAN and GAN Super-Resolution

ESRGAN uses a generator-versus-discriminator contest to invent realistic detail when upscaling images, going beyond blurry interpolation.

2 min readRead
Visual AI

Real-ESRGAN Practical Restoration

Real-ESRGAN extends ESRGAN to handle the messy, unknown degradations of real-world photos rather than clean synthetic blur.

2 min readRead
Visual AI

SwinIR Transformer Restoration

SwinIR applies the Swin Transformer's shifted-window attention to image restoration tasks like super-resolution, denoising, and JPEG artifact removal.

2 min readRead
Visual AI

Latent Consistency Models

Latent Consistency Models (LCMs) are a technique that lets diffusion image generators produce high-quality pictures in just one to four steps instead…

2 min readRead
Visual AI

Vision Transformers

Vision Transformers (ViTs) apply the transformer architecture that powers ChatGPT to images, treating a picture as a sequence of patches instead of a grid…

2 min readRead
Visual AI

Stable Diffusion

Stable Diffusion is an open-source text-to-image model, released by Stability AI in 2022, that generates pictures by gradually removing noise from a random…

2 min readRead
Visual AI

Midjourney

Midjourney is a popular commercial text-to-image service known for its striking, highly aesthetic results and its origins as a Discord bot.

2 min readRead
Visual AI

DALL-E

DALL-E is OpenAI's family of text-to-image models that turn a written description into an original picture.

2 min readRead
Visual AI

CLIP and Vision-Language Models

CLIP is a model from OpenAI that learns to connect images and text by placing both in the same mathematical space.

2 min readRead

Finished reading? Prove it.

Check what you learned with topic quizzes, then explore our structured courses and current certification requirements. Core guides remain free to read.