Free AI library

Visual AI guidesFree forever.

133 plain-English guides, structured learning paths, and an open library — built by an independent 501(c)(3) nonprofit so anyone can understand modern AI.

133Free guides
1Topic tracks
~2 minPer guide
~4hReading time

Start here

Five outcome-based courses

Each course includes explicit outcomes, mapped competencies, practice activities, and an applied capstone.

Topic tracks

Browse by track

Jump into the area you care about. Every track has multiple plain-English guides.

Full library

All guides

133 of 1019 guides shown. Filter by track or search above.

Visual AI

Magic3D Text-to-3D Pipeline

Magic3D is NVIDIA's two-stage answer to DreamFusion, producing higher-resolution, more detailed 3D content faster.

2 min readRead
Visual AI

Denoising and Deblurring Networks

Denoising and deblurring networks are neural models that clean up noisy or blurry images, recovering sharp detail from messy inputs.

2 min readRead
Visual AI

GFPGAN Face Restoration

GFPGAN is a specialized model that restores low-quality, blurry, or old face photos into sharp, realistic portraits.

2 min readRead
Visual AI

CodeFormer Robust Face Recovery

CodeFormer is a face restoration model built to handle extreme degradation, recovering recognizable faces from heavily damaged, tiny, or blurry inputs.

2 min readRead
Visual AI

GigaGAN Scaled Generators

GigaGAN is a billion-parameter GAN that proves generative adversarial networks can scale to text-to-image generation, rivaling diffusion models…

2 min readRead
Visual AI

VQGAN and Codebook Image Synthesis

VQGAN compresses images into a grid of discrete tokens drawn from a learned codebook, letting a transformer generate images the same way language models…

2 min readRead
Visual AI

MaskGIT Parallel Token Decoding

MaskGIT generates images by predicting many tokens at once and filling in the most confident ones first, replacing slow left-to-right generation…

2 min readRead
Visual AI

Zero-1-to-3 Novel View Diffusion

Zero-1-to-3 turns a single photo of an object into images of that same object seen from any new angle, using a diffusion model conditioned on the camera…

2 min readRead
Visual AI

DMTet Hybrid 3D Representation

DMTet (Deep Marching Tetrahedra) is a hybrid 3D shape representation that combines a deformable tetrahedral grid with a signed distance field so neural…

2 min readRead
Visual AI

Instant-NGP Hash Encoding

Instant-NGP is NVIDIA's technique that trains Neural Radiance Fields and other neural graphics primitives in seconds instead of hours by storing learnable…

2 min readRead
Visual AI

DepthAnything Monocular Depth

DepthAnything is a foundation model that estimates how far away every pixel is from a single ordinary photo, with no special hardware.

2 min readRead
Visual AI

Marigold Diffusion Depth Estimation

Marigold repurposes a pretrained image-generation diffusion model (Stable Diffusion) to predict highly detailed depth maps.

2 min readRead

Finished reading? Prove it.

Check what you learned with topic quizzes, then explore our structured courses and current certification requirements. Core guides remain free to read.