Tiếp theoHướng dẫn tiếp theo
Writing Effective Image Generation Prompts
AI trực quan
HƯỚNG DẪN AI trực quan
A negative prompt is text describing what you do not want in an AI-generated image.
In diffusion models such as Stable Diffusion, it replaces the empty prompt used in classifier-free guidance, so every denoising step is pushed away from that description. Negative prompts give users a direct lever over unwanted elements, but they only work on models that run true classifier-free guidance, and they can backfire when overused.
Diffusion models create an image by starting from random noise and removing it over many steps. At each step, the model predicts the noise twice: once conditioned on your prompt and once on an 'unconditional' input, normally an empty prompt. Classifier-free guidance, introduced by Jonathan Ho and Tim Salimans, combines the two. It starts from the unconditional prediction and moves further toward the conditional one, scaled by the guidance scale (often labeled CFG). The result follows the prompt more closely than the raw model would. A negative prompt simply replaces the empty prompt with your negative text. The formula now pushes away from the negative description as it moves toward the positive one. That is why a negative prompt is not a filter or a post-processing step: it shapes every denoising step. The technique spread through community tools such as the AUTOMATIC1111 web interface in 2022. It became especially popular with Stable Diffusion 2.x, whose results many users found improved noticeably with negatives. Negative prompts work best on concrete visual concepts the model knows well: 'blurry', 'text', 'watermark', 'monochrome', or a specific color or object. They work less well as abstract fixes. On many general-purpose models, terms like 'bad anatomy' or 'extra fingers' often do little, because those models rarely saw captions describing such failures. They can also make results worse. A long list pushes the image away from many directions at once. That can reduce variety, wash out or oversaturate colors, and produce stiff compositions, and high guidance scales amplify the effect. Some models, including guidance-distilled and few-step models, skip the unconditional pass entirely. On those, negative prompts are ignored unless the tool turns true CFG back on, which slows generation. A common misconception is that writing 'no cats' in the main prompt excludes cats. Text encoders handle negation poorly, so the phrase can add cats instead.
Visual AI có thể tự động hóa các nhiệm vụ kiểm tra, phát hiện và gắn thẻ trên quy mô lớn.
Các nhóm sáng tạo có thể tạo nguyên mẫu nhanh hơn với ít sửa đổi thủ công hơn.
Các hoạt động có thể sử dụng tín hiệu hình ảnh và video mà trước đây khó xử lý.
Negative prompts are becoming less central as models improve. Newer systems with stronger text encoders follow detailed positive prompts more faithfully, and guidance-distilled models give up negative prompt support in exchange for speed. Meanwhile, research continues on guidance methods that give finer control without the side effects of plain CFG, and some interfaces offer separate controls for suppressing things like text or a style. The lasting lesson for practitioners is mechanical: check whether your model actually runs a negative branch, keep negatives short and concrete, and test each term against a fixed seed so you can see what it really changes.
A product photographer generating a studio shot of a watch adds 'text, watermark, logo' as a negative prompt to discourage fake branding in the background.
An illustrator who wants a clean line drawing puts 'shading, color, photorealistic' in the negative prompt. Writing 'no shading' in the main prompt instead can accidentally add shading.
A user pastes a 60-term negative prompt copied from a forum and finds the images turn flat and repetitive. Trimming it to a few targeted terms brings the variety back.
Someone switches to a guidance-distilled model and notices the negative prompt field does nothing, because by default the model never computes the second prediction a negative prompt would feed.
Quyền và sự đồng ý về hình ảnh có thể trở thành rủi ro pháp lý nếu nguồn gốc xuất xứ không rõ ràng.
Hiệu suất của mô hình có thể khác nhau tùy theo ánh sáng, nhân khẩu học và môi trường.
Kết quả dương tính giả có thể không được chú ý trừ khi ngưỡng tin cậy được theo dõi.
Xác định tiêu chí chấp nhận về độ chính xác, thu hồi và chi phí lỗi.
Kiểm tra với dữ liệu phù hợp với điều kiện sản xuất thực tế.
Thêm đánh giá của con người đối với những dự đoán có độ tin cậy thấp hoặc tác động cao.
Theo dõi sự trôi dạt của mô hình và xác nhận lại sau khi thay đổi máy ảnh hoặc tập dữ liệu.
Free newsletter
Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.
One email each weekday. Unsubscribe in one click. We never sell or share your address.
Test yourself
Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.
Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation
A negative prompt is text describing what you do not want in an AI-generated image. In diffusion models such as Stable Diffusion, it replaces the empty prompt used in classifier-free guidance, so every denoising step is pushed away from that description. Negative prompts give users a direct lever over unwanted elements, but they only work on models that run true classifier-free guidance, and they can backfire when overused.
CFG normally compares the prompted prediction with a prediction from an empty prompt. A negative prompt takes the empty prompt's place, so guidance pushes away from that text.
The encoder registers 'cats' as a strong concept and largely misses the 'no'. Moving the word to the negative prompt reverses the direction of guidance.
Each step needs two noise predictions, one for each prompt, and they are combined using the guidance scale.
Negatives work best on concrete visual concepts the model knows well. Watermarks appear in many training images, while failure descriptions like 'bad anatomy' rarely appear in general captions.
With no unconditional or negative branch to compute, the negative text has nowhere to go. Turning true CFG back on restores it at the cost of speed.
Tiếp tục học hỏi
Đã chọn thêm hướng dẫn cho chủ đề này
Tiếp theoHướng dẫn tiếp theo
Writing Effective Image Generation Prompts
AI trực quan