Local AI
Local AI
Foundations
Local Image Generation Landscape
Understand models, UIs, and hardware requirements
61 of 66
What is diffusion?
Diffusion models learn to turn random noise into images matching a text description. You provide a prompt, the model denoises step by step, and an image appears.
Popular local tools
| Tool | Best for | Complexity |
|---|---|---|
| ComfyUI | Control, workflows, automation | High |
| Automatic1111 / Forge | Quick prototyping | Medium |
| Fooocus | Simple, good defaults | Low |
| InvokeAI | Clean UX, artists | Medium |
Model families
- SD 1.5: old, fast, lots of community models
- SDXL: better quality, needs more VRAM
- Flux: current open-source leader, very capable
- Stable Diffusion 3: strong text rendering
Hardware primer
| Task | Minimum VRAM | Comfortable VRAM |
|---|---|---|
| SD 1.5 | 4 GB | 6 GB |
| SDXL | 6 GB | 8 GB |
| Flux | 12 GB | 16+ GB |
Pick your first stack
- Want maximum control? Start with ComfyUI.
- Want one-click good results? Start with Fooocus.
- Want to script generation? Use ComfyUI in server mode.