Imference Desktop generate on your own GPU, or in the cloud. One app.
Run seven model families locally — SDXL, FLUX, Z-Image, Qwen-Image and more — private and free on your hardware. Or generate images and video through imference.com with credits or x402. No subscription, no lock-in.
Mode
Local
Your GPU, your images
Mode
Cloud
No GPU required
Models
7 local families
+ full cloud catalog, incl. video
Prompt
cinematic portrait, soft rim light, 85mm lens…
Two ways to generate
Switch between local and cloud at any time — same interface, same gallery.
Local — your GPU
Private, unlimited, free to run
One-click engine install, then generate across seven model families — SDXL, SD 1.5, Z-Image, FLUX, Chroma, Qwen-Image and Anima — directly on your machine. Images never leave your computer.
- NVIDIA (CUDA) & AMD (ROCm) on Windows · Apple Silicon on macOS
- Pick a model, hit Generate — weights download automatically, pre-tuned
- Or bring your own .safetensors checkpoint (Civitai & co.)
- Isolated environment — nothing touches your system Python
Cloud — imference.com
No GPU required, full model catalog
Generate on imference.com servers straight from the app. Pay with API credits or x402 (USDC on Base — no account needed).
- Works on any machine — laptops, no GPU, low VRAM
- Full hosted catalog: realism, anime — and video generation
- Same prompts, params and gallery as local mode
Built for actually making images
Not a wrapper around a text box — real controls, real history.
Per-model parameters
Format, steps, CFG, seed, quality tags, negative prompt — with sensible defaults pulled from the model catalog.
Image-to-image
Start from an existing image and steer it with a prompt — locally or in the cloud.
Metadata-rich gallery
Every image and video is saved with its prompt, model and params. Filter, search, and reopen in a fullscreen viewer.
Models that install themselves
Pick from the curated catalog and generate — the weights download automatically, already tuned per model.
Your own checkpoints
Add any .safetensors file — Civitai fine-tunes and merges included. Used in place, never copied or uploaded.
Queue & keep working
Stack generations, stop or cancel runs, even switch models mid-run — the app finishes the current image first.
Heads-up on first launch
The app isn't code-signed yet, so your OS will show a warning the first time you run it. That's expected — signing and in-app auto-update are on the way.
- Windows: SmartScreen → More info → Run anyway.
-
macOS:
right-click the app → Open (or
xattr -cr "/Applications/Imference Desktop.app").
Every release ships a checksums.txt so you can verify your download.
Requirements
System
Windows 10/11 (x64) or macOS 12+
Local generation
NVIDIA GPU (CUDA) or AMD Radeon RX 7000/9000 (ROCm) on Windows, Apple Silicon on macOS
~6–7 GB of disk per model, downloaded on demand
Cloud generation
An imference.com API key, or a funded x402 wallet
Start generating in minutes
Download, launch, pick local or cloud — that's it. Free and open on GitHub.