VIP 👤
🏠 Home
Benchmark Hub
📊 All Benchmarks 🦖 Dinosaur v1 🦖 Dinosaur v2 ✅ To-Do List Applications 🎨 Creative Free Pages 🎯 FSACB - Ultimate Showcase 🌍 Translation Benchmark
Models
🏆 Top 10 Models 🆓 Free Models 📋 All Models ⚙️ Kilo Code
Resources
💬 Prompts Library 📖 AI Glossary 🔗 Useful Links 🔌 API & Routers

AI Glossary

The complete dictionary of Artificial Intelligence

162
categories
2,032
subcategories
23,060
terms
📖
terms

Classifier-Free Guidance

Guidance technique that eliminates the need for an external classifier by using the model itself to follow or ignore a conditioning signal, thus improving prompt fidelity.

📖
terms

IP-Adapter (Image Prompt Adapter)

Module that allows using a reference image as a prompt, encoding its visual characteristics to guide the generation process without modifying the diffusion model's weights.

📖
terms

GLIGEN (Grounded Language-to-Image Generation)

Framework that anchors bounding boxes to text concepts, enabling precise spatial control over the position and size of generated objects in an image.

📖
terms

Cross-Attention Guidance

Mechanism that leverages cross-attention maps between text and image to strengthen or weaken the influence of specific parts of the prompt on areas of the generated image.

📖
terms

LoRA (Low-Rank Adaptation)

Efficient fine-tuning technique that injects small low-rank matrices into the model's attention layers, allowing learning of new styles or concepts with minimal parameters.

📖
terms

T2I-Adapter (Text-to-Image Adapter)

Lightweight module that guides a pre-trained diffusion model using additional reference conditions (such as sketches or segmentation maps) without requiring full retraining.

📖
terms

UNet Guidance

Guidance process that operates directly on the denoising predictions of the UNet architecture, modifying the gradient direction to align generation with desired conditioning.

📖
terms

Attention Refocusing

Method that modifies attention maps to redirect the importance of a prompt token to another region of the image, allowing reorganization of generated elements' composition.

📖
terms

Semantic Guidance

Approach that uses an external semantic model (e.g., CLIP) to compute a guidance loss, ensuring that the generated image remains aligned with the prompt's meaning at a conceptual level.

📖
terms

Multi-Modal Guidance

Strategy that combines multiple types of guidance (text, image, mask, depth map) simultaneously for hierarchical and multi-faceted control over the generation process.

📖
terms

Null-Text Inversion

Image editing technique that inverts the generation process to find an optimized 'null prompt', enabling precise modifications on real images without artifacts.

📖
terms

SDEdit (Stochastic Differential Equation Edit)

Method that adds noise to an existing image then denoises it with guidance, allowing editing of real images while preserving their underlying structure.

🔍

No results found