Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Run compatible language models locally; explicitly distinguish local models from optional cloud models.
Platforms
Windows · macOS · Linux
Guidance & sources
Start with a small local model and check whether cloud functionality is enabled.
Disable optional cloud features when you require a local-only workflow. Model licenses are separate.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Explore a graphical workflow for downloading and running compatible local language models.
Platforms
Windows · macOS · Linux
Guidance & sources
Check the model, runtime compatibility and your device memory before a large download.
Edition and requirements can change; check the official source before installing.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Build node-based image generation, image processing and supported video workflows.
Platforms
Windows · macOS · Linux
Guidance & sources
Capabilities depend on installed models, custom nodes and workflows; ComfyUI alone is not a video generator. Local core is free; optional cloud/partner nodes may cost money. Stable Diffusion model licenses differ.
Edition and requirements can change; check the official source before installing.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Build local image-generation and canvas workflows with supported models.
Platforms
Windows · macOS · Linux
Guidance & sources
Model licenses, hardware and supported workflows differ. Stable Diffusion is a model family, not a single universally licensed app.
Edition and requirements can change; check the official source before installing.
Choose by medium: GIMP for raster editing, Krita for painting, Inkscape for vectors, Blender for 3D and ShareX for Windows capture. Keep editable source files and export a separate delivery copy. Advanced creative work takes practice; begin with one small project.
Chat with supported local language models or optional remote providers.
Platforms
Windows · macOS · Linux
Guidance & sources
Download a compatible model and check RAM/GPU and its license. Local mode can work offline after downloads; enabling remote providers changes data processing and may cost money.
Edition and requirements can change; check the official source before installing.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Run supported language models locally and explore local document chat.
Platforms
Windows · macOS · Linux
Guidance & sources
Download a compatible model and check RAM/GPU and its license. Local mode can work offline after downloads; enabling remote providers changes data processing and may cost money.
Edition and requirements can change; check the official source before installing.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Self-host a web interface for local or remote model providers.
Platforms
Self-hosted server · Web
Guidance & sources
Current source license imposes a branding condition; do not classify it as unrestricted open source. Model/provider licenses and remote processing are separate; self-hosting needs administration.
Edition and requirements can change; check the official source before installing.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.
Transcribe supported audio locally with multilingual speech-recognition models.
Platforms
Windows · macOS · Linux
Guidance & sources
Download a compatible model and check RAM/GPU and its license. Local mode can work offline after downloads; enabling remote providers changes data processing and may cost money.
Edition and requirements can change; check the official source before installing.
Local AI runs model inference on your own device. After downloading software and models, supported local workflows can work offline; cloud features and extensions may still send data. RAM is system memory, VRAM is GPU memory. CPUs can run some models more slowly; a compatible GPU can accelerate supported workloads. Larger models require more memory. Quantization uses fewer bits to reduce size, trading some accuracy for easier running. Start small, check model licenses and measure on your actual machine.