Krea 2 VRAM requirements in ComfyUI, file by file

Krea 2 VRAM in ComfyUI: every file at every precision, which one loads on 12 to 32 GB cards, and what Krea says it needs.

Updated · estimates are labelled as estimates
On this page

Krea 2 is Krea's open image model. ComfyUI's own workflow for Turbo downloads 18.6 GB and needs 13.4 GB on the card at its largest stage, the diffusion stage. The smallest card listed here that loads that set with room to spare is the RTX 5060 Ti.

What Krea 2 is

Krea's open image model from June 2026: a 12-billion-parameter transformer in two versions. Turbo samples in eight steps in Krea's own example and is what ComfyUI's template loads; Raw is the undistilled model, the one trainers point LoRAs at.

Turbo: 12B · Raw: 12B Text encoder: Qwen3-VL 4B Krea 2 Community License June 2026

Which Krea 2 file for your card

For each memory size, the highest-precision diffusion file that loads with room to spare (a file that only just fits is used when nothing else does), and what is left for working memory while it samples. "Streams" means even the smallest file is larger than the card: ComfyUI still runs it by streaming weights from system RAM, more slowly. FP4 files are offered only to the Blackwell cards that run them natively.

Turbo

CardUsableFile that loadsLargest stageLeft for the workVerdict
RTX 4070 12 GB smallest: FP8 scaled 13.4 GB – Streams
RTX 5060 Ti 16 GB FP8 scaled 13.4 GB 2.6 GB Loads fully
RTX 4090 24 GB FP8 scaled 13.4 GB 10.6 GB Loads fully
RTX 5090 32 GB BF16 26.5 GB 5.5 GB Loads fully
RTX 6000 Ada 48 GB BF16 26.5 GB 21.5 GB Loads fully
H100 SXM 80 GB BF16 26.5 GB 53.5 GB Loads fully
RTX PRO 6000 Blackwell 96 GB BF16 26.5 GB 69.5 GB Loads fully
DGX Spark 126 GB BF16 26.5 GB 99.5 GB Loads fully

ComfyUI's template for Turbo: FP8 scaled diffusion model, FP8 scaled text encoder. Stages: text encoder 5.2 GB, diffusion 13.4 GB. Download 18.6 GB.

Raw

CardUsableFile that loadsLargest stageLeft for the workVerdict
RTX 4070 12 GB smallest: FP8 scaled 13.4 GB – Streams
RTX 5060 Ti 16 GB FP8 scaled 13.4 GB 2.6 GB Loads fully
RTX 4090 24 GB FP8 scaled 13.4 GB 10.6 GB Loads fully
RTX 5090 32 GB BF16 26.5 GB 5.5 GB Loads fully
RTX 6000 Ada 48 GB BF16 26.5 GB 21.5 GB Loads fully
H100 SXM 80 GB BF16 26.5 GB 53.5 GB Loads fully
RTX PRO 6000 Blackwell 96 GB BF16 26.5 GB 69.5 GB Loads fully
DGX Spark 126 GB BF16 26.5 GB 99.5 GB Loads fully

ComfyUI's template for Raw: BF16 diffusion model, FP8 scaled text encoder. Stages: text encoder 5.2 GB, diffusion 26.5 GB. Download 31.8 GB.

Check another card, or pick a file by hand, in the ComfyUI VRAM checker.

Every Krea 2 file ComfyUI loads

Sizes are the exact byte counts Hugging Face reports. Files marked "template" are the ones ComfyUI's official workflow downloads.

PartPrecisionSizeVersionFile
Diffusion model BF16 26.28 GB Turbo krea2_turbo_bf16.safetensors
Diffusion model FP8 scaled · template 13.14 GB Turbo krea2_turbo_fp8_scaled.safetensors
Diffusion model INT8 ConvRot 13.49 GB Turbo krea2_turbo_int8_convrot.safetensors
Diffusion model MXFP8 13.53 GB Turbo krea2_turbo_mxfp8.safetensors
Diffusion model NVFP4 7.67 GB Turbo krea2_turbo_nvfp4.safetensors
Diffusion model BF16 26.28 GB Raw krea2_raw_bf16.safetensors
Diffusion model FP8 scaled 13.14 GB Raw krea2_raw_fp8_scaled.safetensors
Diffusion model INT8 ConvRot 13.49 GB Raw krea2_raw_int8_convrot.safetensors
Text encoder BF16 8.88 GB all qwen3vl_4b_bf16.safetensors
Text encoder FP8 scaled · template 5.24 GB all qwen3vl_4b_fp8_scaled.safetensors
VAE BF16 · template 0.25 GB all qwen_image_vae.safetensors

What Krea says it needs

Krea states no VRAM figure on either model card or in its GitHub README. The closest it comes: "Architecture: Diffusion Transformer with 12 billion parameters" (source).

Those figures and this page measure different things. A vendor's figure covers its own script at its default resolution, working memory included. The tables here count the weights each stage holds, which is exact, and leave the working memory as the space that remains, because it depends on your resolution and attention kernel and nobody publishes it per model. If generation runs out of memory with a file that loads, lower the resolution before stepping down a precision. ComfyUI's README says it "can run even the biggest open source models on as low as 4GB vram + 8GB ram" by streaming weights (source), so "streams" means slower, not impossible.

LoRA training for Krea 2

Only configurations a trainer's own documentation gives a VRAM figure for. Settings change the figure a great deal, so read the setting column before trusting the number.

TrainerVRAMSettingSource
diffusion-pipe · Raw24 GBrank 32 at 512 pixels, model held in FP8docs

What Krea 2 is good at, and what to watch

Good at

  • Images from 1K up to 2K.
  • Style work: Comfy-Org repackages a set of Krea's own LoRAs alongside the model.
  • FP8 and INT8 files at about half the size of the BF16 original.

Watch out for

  • Commercial use only while your company's revenue over the last twelve months is under US$1 million.
  • The NVFP4 and MXFP8 files are made for Blackwell cards (RTX 50, RTX PRO 6000); on anything older, use FP8 or INT8.
  • Krea publishes no VRAM figure.

Questions

How much VRAM does Krea 2 need?

In ComfyUI's own workflow for Turbo, the largest stage is the diffusion stage at 13.4 GB, and the whole workflow downloads 18.6 GB. Working memory for generation comes on top and grows with resolution. Krea publishes no VRAM figure.

Can I run Krea 2 on 8 GB of VRAM?

Not fully. The smallest file here, FP8 scaled at 13.1 GB, needs 13.4 GB at its largest stage, more than 8 GB holds. ComfyUI will still run it by streaming weights from system RAM, which works but is slower.

Can I run Krea 2 on 12 GB of VRAM?

Not fully. The smallest file here, FP8 scaled at 13.1 GB, needs 13.4 GB at its largest stage, more than 12 GB holds. ComfyUI will still run it by streaming weights from system RAM, which works but is slower.

Can I run Krea 2 on 16 GB of VRAM?

Yes, with the FP8 scaled file: its largest stage, the diffusion stage, is 13.4 GB, which loads on 16 GB with 2.6 GB left for working memory.

Which Krea 2 file should I download?

The highest precision that loads fully on your card. On a 24 GB card that is the FP8 scaled file (13.1 GB). ComfyUI's template downloads FP8 scaled with the FP8 scaled text encoder. FP4 files (NVFP4) are made for Blackwell cards (RTX 50, RTX PRO 6000, DGX Spark); on older cards, use FP8 or INT8.

  • Files and sizes: Comfy-Org/Krea-2 on Hugging Face, read from the Hugging Face API. Template files: ComfyUI's Krea 2 guide.
  • Model and licence: krea/Krea-2-Turbo and the Krea 2 Community License.
  • Card memory: manufacturers' specifications, as on each GPU page. The checker covers NVIDIA cards; ComfyUI also runs on AMD and Apple silicon, where file formats and speed differ and we have not sized them.
  • Fit rule: a stage loads fully under 85% of usable memory and is tight up to 95%, the same margins the LLM pages use.