The Image Studio provides a local environment to generate high-resolution images. You can generate images through Stable Diffusion WebUI Forge or ComfyUI.
The studio supports FLUX.1, SDXL, and Stable Diffusion 1.5 base models.
Select the model family that matches your GPU memory and artistic goals:
| Model Family | Native Resolution | Minimum VRAM | Recommended Use Case |
|---|---|---|---|
| FLUX.1 (dev / schnell) | 1024 × 1024 | 12 GB | Best text rendering, complex prompts, and realistic anatomy. |
| SDXL (Stable Diffusion XL) | 1024 × 1024 | 8 GB | Photorealism, cinematic lighting, and custom community styles. |
| Stable Diffusion 1.5 | 512 × 512 | 4 GB | High-speed rendering and lightweight game assets. |
Note
You can download base models and checkpoints directly through the Model Management tab from Hugging Face or CivitAI.
Follow these steps to generate an image using the Fluid Studio Canvas and Creative Prompt Dock:
flowchart LR
E["1. Check Engine Card"] --> M["2. Select 🎨 Image"]
M --> P["3. Choose Preset"]
P --> D["4. Prompt in Dock"]
D --> G["5. Click Generate ↵"]
- Check the SD Forge card in the Telemetry Header or Ribbon.
- Ensure the status dot is green (Online).
- If stopped, click ▷ Start on the card to launch the engine on port
7860.
- Open the Studio workspace from the left Activity Rail.
- In the top modality selector bar, click 🎨 Image.
- The Center Stage canvas activates the interactive Image Viewport with pan/zoom preview controls.
Select a preset from the Studio Preset Bar or click a starter prompt chip directly within the canvas:
- Standard Square (1024 × 1024): Default setting for general art and character portraits.
- Landscape Wallpaper (1344 × 768): Wide aspect ratio for environments and desktop backgrounds.
- Portrait Photo (768 × 1152): Vertical aspect ratio for full-body human figures and posters.
- Classic SD 1.5 (512 × 512): Lightweight resolution for legacy models and rapid drafting.
- Prompt Textarea: Describe your subject, lighting, and composition in the autosizing prompt dock.
- Parameters Flyout (
⚙️ Settings): Click the settings button to slide open fine-tuning sliders:- Steps: 20 to 30 steps for SDXL, 4 to 8 for FLUX schnell.
- CFG Scale: 5.0 to 7.0 for SDXL, 1.0 for FLUX.
- Denoise: Adjust denoising strength (default: 0.75).
- Seed: Numeric seed (
-1for randomized variations). - Aspect Ratio: Click
1:1,16:9,9:16, or4:3pills.
- Reference Attachment (
📎): Optionally attach an image for image-to-image or style conditioning.
- Click the prominent Generate ↵ action button in the dock.
- Forge WebUI processes the prompt via GPU diffusion sampling.
- Inspect the completed image in the interactive Center Stage canvas.
Important
The application automatically saves rendered images to outputs/images/ with embedded prompt and parameter metadata.
LoRAs are compact model files (typically 50 MB to 200 MB). A LoRA modifies an existing checkpoint into a specific character, clothing style, or visual aesthetic.
- Open the Stable Diffusion (CivitAI) tab.
- Set the Type dropdown filter to LoRA.
- Sort results by Highest Rated or Most Downloaded.
- Type your desired aesthetic into the search bar (for example,
Pixel Art XLorDetail Tweaker). - Click ⬇ Download to Forge. The system stores the file in your models directory.
| Style Category | Search Term | Recommended Base Model | Effect |
|---|---|---|---|
| Pixel Art | Pixel Art XL |
SDXL | Generates authentic 16-bit sprites and backgrounds. |
| Retro 3D | PS1 Graphics |
SDXL / SD 1.5 | Recreates jagged low-polygon PlayStation 1 aesthetics. |
| Anime & Cartoon | Cel Shaded |
SDXL / Pony V6 | Applies clean outlines and flat anime color fills. |
| Micro-Detail | Detail Tweaker XL |
SDXL / FLUX | Sharpens skin pores, fabric weaves, and rim lighting. |
Add the trigger tag directly into your positive prompt box:
<lora:pixel_art_xl:0.8> a cyberpunk street market at night, 16-bit retro game asset
- Specify the filename of the LoRA between the colons.
- Add a weight value at the end (typically
0.6to1.0). - Lower the weight if the style distorts facial anatomy or colors.
When using ComfyUI as your studio backend:
- Add a Load LoRA node to your ComfyUI canvas.
- Connect the
MODELoutput from your Checkpoint Loader into themodelinput of the LoRA node. - Connect the
CLIPoutput into theclipinput of the LoRA node. - Select your downloaded LoRA filename inside the node dropdown menu.
- Set
strength_modelandstrength_clipto0.8. - Connect the modified outputs into your positive prompt and KSampler nodes.
High-resolution diffusion models require dedicated GPU VRAM. Review these memory guidelines:
- Automated LLM Offload: The VRAM Orchestrator automatically unloads Ollama models before diffusion starts.
- FLUX Models: FLUX.1-dev requires at least 12 GB VRAM in FP8 or NF4 quantization.
- Resolution Limits: Do not set resolutions above 1024 × 1024 on GPUs with 8 GB VRAM. Use high-resolution fix or latent upscalers instead.
Warning
Generating at resolutions above 1536 × 1536 without tiling causes Out-Of-Memory errors on 16 GB GPUs. Use the standard presets to maintain safe VRAM bounds.
- Studio Overview — Learn about modular feature packs and presets.
- Video Generation Guide — Transform static images into animated video clips.
- 3D Mesh Reconstruction Guide — Convert generated 2D images into 3D meshes.
- SD Forge Engine Guide — Configure your local Forge installation.
- CivitAI & Model Hub Guide — Manage checkpoints and embeddings.