Model reference · open weights

stable-diffusion-3.5-medium

Image city96 Text→image 1 build Its own licence terms 2k dl/mo

stable-diffusion-3.5-medium is an open-weight image model from city96. stable-diffusion-3.5-medium-gguf (GGUF) weighs 1.8 GB; the smallest configuration that runs it is RTX 3060 12 GB.

  • stable-diffusion-3.5-medium is a text-to-image model created by city96 as a direct GGUF conversion of the original stabilityai release.
  • It is a quantized version intended for use with the ComfyUI-GGUF custom node and supports English.
  • The model is published under an "other" license, meaning the original license terms and restrictions from the source model still apply.

Summary of the city96/stable-diffusion-3.5-medium-gguf model card, 2026-10-01

What it is

Released bycity96
Released2024-10-30
VRAM1.8 GB for the weights

What it runs on

Memory and cards for stable-diffusion-3.5-medium-gguf (GGUF)

1.8 GBweights, file size
5.0 GBworking memory, one 1024² image
537 MBruntime overhead
CardOne imageMemory
RTX 3060 12 GBfits11.6 GB
RTX 4060 Ti 16 GBfits15.4 GB
RTX 3090 24 GBfits23.4 GB
RTX 4090 24 GBfits23.4 GB
RTX 5090 32 GBfits31.0 GB
L40S 48 GBfits44.0 GB
A100 80 GBfits78.2 GB
H100 80 GBfits78.1 GB
RTX PRO 6000 Blackwell 96 GBfits93.8 GB
DGX Spark (GB10) 128 GB unifiedfits107 GB
H200 141 GBfits138 GB
B200 180 GBfits176 GB

From the model card

What city96 says about stable-diffusion-3.5-medium

Read the model card

This is a direct GGUF conversion of stabilityai/stable-diffusion-3.5-medium

As this is a quantized model not a finetune, all the same restrictions/original license terms still apply.

The model files can be used with the ComfyUI-GGUF custom node.

Place model files in ComfyUI/models/unet - see the GitHub readme for further install instructions.

Please refer to this chart for a basic overview of quantization types.

Quoted from the model card on Hugging Face. The full card is behind the Hugging Face link above.

How it works

How image models work

Text promptwhat to makeText encoderunderstands itDiffusion stepsdenoise to pixelsImagePNG / JPEGA diffusion model starts from noise and denoises it, guided by your prompt, into a finished image.

Running it yourself

Run it on a rented GPU

Rent a machine by the hour. ComfyUI is installed on it. Open ComfyUI through the tunnel and load the workflow from the model's card on Hugging Face. Its model loader takes .safetensors files: replace it with Unet Loader (GGUF) and choose this model's file there.

# on your rented machine (the ssh line is on its page in the console)
# the GGUF loader nodes, once per machine
git clone https://github.com/city96/ComfyUI-GGUF /opt/ComfyUI/custom_nodes/ComfyUI-GGUF
pip install -r /opt/ComfyUI/custom_nodes/ComfyUI-GGUF/requirements.txt

# get REPO FILE FOLDER: one file into /workspace/models/FOLDER, where ComfyUI loads it from
get() { hf download "$1" "$2" --local-dir /workspace/hf-files && mkdir -p "/workspace/models/$3" && mv "/workspace/hf-files/$2" "/workspace/models/$3/$4"; }

# the model (1.7 GB)
get city96/stable-diffusion-3.5-medium-gguf sd3.5_medium-Q4_K_M.gguf unet

start-comfyui
Renting a GPU: connect, tunnels, ComfyUI
# on your computer, in a second terminal: ComfyUI in your browser at http://localhost:8188
# HOST and PORT are your machine's, from its page in the console
ssh -L 8188:localhost:8188 dev@HOST -p PORT
© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms