Krea 2 is Krea's open image model. ComfyUI's own workflow for Turbo downloads 18.6 GB and needs 13.4 GB on the card at its largest stage, the diffusion stage. The smallest card listed here that loads that set with room to spare is the RTX 5060 Ti.
What Krea 2 is
Krea's open image model from June 2026: a 12-billion-parameter transformer in two versions. Turbo samples in eight steps in Krea's own example and is what ComfyUI's template loads; Raw is the undistilled model, the one trainers point LoRAs at.
Turbo: 12B · Raw: 12B Text encoder: Qwen3-VL 4B Krea 2 Community License June 2026
Which Krea 2 file for your card
For each memory size, the highest-precision diffusion file that loads with room to spare (a file that only just fits is used when nothing else does), and what is left for working memory while it samples. "Streams" means even the smallest file is larger than the card: ComfyUI still runs it by streaming weights from system RAM, more slowly. FP4 files are offered only to the Blackwell cards that run them natively.
Turbo
| Card | Usable | File that loads | Largest stage | Left for the work | Verdict |
|---|---|---|---|---|---|
| RTX 4070 | 12 GB | smallest: FP8 scaled | 13.4 GB | – | Streams |
| RTX 5060 Ti | 16 GB | FP8 scaled | 13.4 GB | 2.6 GB | Loads fully |
| RTX 4090 | 24 GB | FP8 scaled | 13.4 GB | 10.6 GB | Loads fully |
| RTX 5090 | 32 GB | BF16 | 26.5 GB | 5.5 GB | Loads fully |
| RTX 6000 Ada | 48 GB | BF16 | 26.5 GB | 21.5 GB | Loads fully |
| H100 SXM | 80 GB | BF16 | 26.5 GB | 53.5 GB | Loads fully |
| RTX PRO 6000 Blackwell | 96 GB | BF16 | 26.5 GB | 69.5 GB | Loads fully |
| DGX Spark | 126 GB | BF16 | 26.5 GB | 99.5 GB | Loads fully |
ComfyUI's template for Turbo: FP8 scaled diffusion model, FP8 scaled text encoder. Stages: text encoder 5.2 GB, diffusion 13.4 GB. Download 18.6 GB.
Raw
| Card | Usable | File that loads | Largest stage | Left for the work | Verdict |
|---|---|---|---|---|---|
| RTX 4070 | 12 GB | smallest: FP8 scaled | 13.4 GB | – | Streams |
| RTX 5060 Ti | 16 GB | FP8 scaled | 13.4 GB | 2.6 GB | Loads fully |
| RTX 4090 | 24 GB | FP8 scaled | 13.4 GB | 10.6 GB | Loads fully |
| RTX 5090 | 32 GB | BF16 | 26.5 GB | 5.5 GB | Loads fully |
| RTX 6000 Ada | 48 GB | BF16 | 26.5 GB | 21.5 GB | Loads fully |
| H100 SXM | 80 GB | BF16 | 26.5 GB | 53.5 GB | Loads fully |
| RTX PRO 6000 Blackwell | 96 GB | BF16 | 26.5 GB | 69.5 GB | Loads fully |
| DGX Spark | 126 GB | BF16 | 26.5 GB | 99.5 GB | Loads fully |
ComfyUI's template for Raw: BF16 diffusion model, FP8 scaled text encoder. Stages: text encoder 5.2 GB, diffusion 26.5 GB. Download 31.8 GB.
Check another card, or pick a file by hand, in the ComfyUI VRAM checker.
Every Krea 2 file ComfyUI loads
Sizes are the exact byte counts Hugging Face reports. Files marked "template" are the ones ComfyUI's official workflow downloads.
| Part | Precision | Size | Version | File |
|---|---|---|---|---|
| Diffusion model | BF16 | 26.28 GB | Turbo | krea2_turbo_bf16.safetensors |
| Diffusion model | FP8 scaled · template | 13.14 GB | Turbo | krea2_turbo_fp8_scaled.safetensors |
| Diffusion model | INT8 ConvRot | 13.49 GB | Turbo | krea2_turbo_int8_convrot.safetensors |
| Diffusion model | MXFP8 | 13.53 GB | Turbo | krea2_turbo_mxfp8.safetensors |
| Diffusion model | NVFP4 | 7.67 GB | Turbo | krea2_turbo_nvfp4.safetensors |
| Diffusion model | BF16 | 26.28 GB | Raw | krea2_raw_bf16.safetensors |
| Diffusion model | FP8 scaled | 13.14 GB | Raw | krea2_raw_fp8_scaled.safetensors |
| Diffusion model | INT8 ConvRot | 13.49 GB | Raw | krea2_raw_int8_convrot.safetensors |
| Text encoder | BF16 | 8.88 GB | all | qwen3vl_4b_bf16.safetensors |
| Text encoder | FP8 scaled · template | 5.24 GB | all | qwen3vl_4b_fp8_scaled.safetensors |
| VAE | BF16 · template | 0.25 GB | all | qwen_image_vae.safetensors |
What Krea says it needs
Krea states no VRAM figure on either model card or in its GitHub README. The closest it comes: "Architecture: Diffusion Transformer with 12 billion parameters" (source).
Those figures and this page measure different things. A vendor's figure covers its own script at its default resolution, working memory included. The tables here count the weights each stage holds, which is exact, and leave the working memory as the space that remains, because it depends on your resolution and attention kernel and nobody publishes it per model. If generation runs out of memory with a file that loads, lower the resolution before stepping down a precision. ComfyUI's README says it "can run even the biggest open source models on as low as 4GB vram + 8GB ram" by streaming weights (source), so "streams" means slower, not impossible.
LoRA training for Krea 2
Only configurations a trainer's own documentation gives a VRAM figure for. Settings change the figure a great deal, so read the setting column before trusting the number.
| Trainer | VRAM | Setting | Source |
|---|---|---|---|
| diffusion-pipe · Raw | 24 GB | rank 32 at 512 pixels, model held in FP8 | docs |
What Krea 2 is good at, and what to watch
Good at
- Images from 1K up to 2K.
- Style work: Comfy-Org repackages a set of Krea's own LoRAs alongside the model.
- FP8 and INT8 files at about half the size of the BF16 original.
Watch out for
- Commercial use only while your company's revenue over the last twelve months is under US$1 million.
- The NVFP4 and MXFP8 files are made for Blackwell cards (RTX 50, RTX PRO 6000); on anything older, use FP8 or INT8.
- Krea publishes no VRAM figure.
Questions
How much VRAM does Krea 2 need?
In ComfyUI's own workflow for Turbo, the largest stage is the diffusion stage at 13.4 GB, and the whole workflow downloads 18.6 GB. Working memory for generation comes on top and grows with resolution. Krea publishes no VRAM figure.
Can I run Krea 2 on 8 GB of VRAM?
Not fully. The smallest file here, FP8 scaled at 13.1 GB, needs 13.4 GB at its largest stage, more than 8 GB holds. ComfyUI will still run it by streaming weights from system RAM, which works but is slower.
Can I run Krea 2 on 12 GB of VRAM?
Not fully. The smallest file here, FP8 scaled at 13.1 GB, needs 13.4 GB at its largest stage, more than 12 GB holds. ComfyUI will still run it by streaming weights from system RAM, which works but is slower.
Can I run Krea 2 on 16 GB of VRAM?
Yes, with the FP8 scaled file: its largest stage, the diffusion stage, is 13.4 GB, which loads on 16 GB with 2.6 GB left for working memory.
Which Krea 2 file should I download?
The highest precision that loads fully on your card. On a 24 GB card that is the FP8 scaled file (13.1 GB). ComfyUI's template downloads FP8 scaled with the FP8 scaled text encoder. FP4 files (NVFP4) are made for Blackwell cards (RTX 50, RTX PRO 6000, DGX Spark); on older cards, use FP8 or INT8.
- Files and sizes: Comfy-Org/Krea-2 on Hugging Face, read from the Hugging Face API. Template files: ComfyUI's Krea 2 guide.
- Model and licence: krea/Krea-2-Turbo and the Krea 2 Community License.
- Card memory: manufacturers' specifications, as on each GPU page. The checker covers NVIDIA cards; ComfyUI also runs on AMD and Apple silicon, where file formats and speed differ and we have not sized them.
- Fit rule: a stage loads fully under 85% of usable memory and is tight up to 95%, the same margins the LLM pages use.