_1782732888602_1783579734627.webp?width=1400&height=450&quality=80&resize=cover)
COMMUNITY PAGE
Run Krea2 now on Floyo
Home / Model / Krea 2 on Floyo
AI IMAGE GENERATION
Run Krea 2 on Floyo
The aesthetic-first image model. 12.9B parameter DiT trained from scratch on real images only (zero synthetic data). Style references, moodboard generation, and wide aesthetic diversity by design. Two variants: Raw (LoRA fine-tuning) and Turbo (2K images in 2 seconds).
Run Krea AI's Krea 2 through ComfyUI in your browser. No API key, no installs, no local GPU.
|
Parameters 12.9B (dense DiT) |
Turbo Speed ~2 seconds (2K, 8 steps) |
|
Training Data Real images only (zero synthetic) |
Arena Rank #1 independent lab (Artificial Analysis) |
No installation. Runs in browser. Updated July 2026.
image generation
krea 2
krea 2 turbo
krea ai
lora styles
open source
text to image
Turn a text prompt into a finished image with Krea 2 Turbo, Krea AI's fast open-source image model. Type what you want to see, hit run, and get a result in seconds.
Krea 2 for Text to Image
Turn a text prompt into a finished image with Krea 2 Turbo, Krea AI's fast open-source image model. Type what you want to see, hit run, and get a result in seconds.
Image generation
Krea2
text to image
turbo
Write a prompt and Krea 2 Turbo generates a high-resolution image at 1920x1280 in 8 steps, using a fast turbo-distilled model with a Qwen 3 VL text encoder.
Krea 2 Turbo · Text to Image For Anime
Write a prompt and Krea 2 Turbo generates a high-resolution image at 1920x1280 in 8 steps, using a fast turbo-distilled model with a Qwen 3 VL text encoder.
What you get?
What You Get
Krea 2 is Krea AI's first foundation image model, built from scratch with a 12.9 billion parameter dense Diffusion Transformer. Trained on billions of real images with zero synthetic or AI-generated data in the pretraining set. Designed for aesthetic diversity and creative exploration: the same prompt produces a wide range of distinct interpretations, not four variations of the same concept. Two open-weight variants on HuggingFace: Krea 2 Raw (undistilled, built for LoRA fine-tuning) and Krea 2 Turbo (8-step distilled, 2K images in about 2 seconds). Style reference and moodboard support. #1 among independent labs on Artificial Analysis. Available as ComfyUI nodes on Floyo.
KREA 2 WORKFLOWS ON FLOYO












What is Krea 2?
Krea 2 is Krea AI's first foundation image model, announced May 12, 2026 and open-sourced June 22-24, 2026. It is a 12.9 billion parameter dense Diffusion Transformer trained from scratch on billions of real images. No fine-tune. No distillation of someone else's model. No synthetic or AI-generated data in the pretraining set. Ranked #1 among independent labs and #6 overall on the Artificial Analysis text-to-image leaderboard.
The design philosophy is different from most image models. Krea 2 prioritizes aesthetic diversity and creative exploration over deterministic prompt adherence. Give it the prompt "horse rearing" and you get four distinct visual interpretations: different angles, styles, moods, and compositions. Most competing models give you four slightly varied versions of the same default interpretation. Krea built this intentionally. Great art pushes boundaries, and the model reflects that belief.
Style references are the primary control mechanism. Upload a single reference image or curate a moodboard (a collection of images representing a visual direction), and Krea 2 generates output that matches the aesthetic without copying the reference. This gives creative teams a natural workflow: show the model what you want rather than describing it in text. The moodboard system handles open-ended aesthetics that incorporate many styles, themes, or motifs in a single direction.
Two open-weight checkpoints shipped with the release. Krea 2 Raw is the undistilled mid-training checkpoint designed for LoRA fine-tuning and research. It gives you a malleable latent space with genuine room to push the model toward custom styles, characters, and subjects. Krea 2 Turbo is the distilled inference engine: 8 steps, 2K images in about 2 seconds on consumer hardware. Use Raw when you need maximum quality or training flexibility. Use Turbo when speed matters more than the last degree of refinement.
On Floyo, both variants run through native ComfyUI nodes on H100 NVL GPUs. Two workflows cover standard generation and Turbo generation. No model download, no 12.9B weight management, no ComfyUI version requirements to handle.
What are Krea 2's technical specifications?
Krea 2 is a 12.9B parameter dense single-stream DiT with 28 transformer blocks at width 6144, grouped-query attention with gated sigmoid attention, SwiGLU MLPs at 4x expansion, and 3D axial RoPE. Text encoder is Qwen3-VL-4B-Instruct with multi-layer feature aggregation. Two variants: Raw (undistilled, for fine-tuning) and Turbo (8-step distilled, ~2 seconds at 2K). Trained on real images only with no synthetic data.
| Spec | Details |
|---|---|
| Developer | Krea AI |
| Architecture | Dense single-stream DiT (28 blocks, width 6144, 3D axial RoPE) |
| Parameters | 12.9 billion (dense, no MoE) |
| Attention | Grouped-query attention with gated sigmoid attention |
| MLP | SwiGLU at 4x expansion |
| Text Encoder | Qwen3-VL-4B-Instruct (multi-layer feature aggregation) |
| Training Data | Billions of real images (zero synthetic/AI-generated in pretraining) |
| Training Pipeline | Pretraining, midtraining, SFT, preference optimization, reinforcement learning |
| Prompt Expander | Trained with GRPO (Group Relative Policy Optimization) |
| Max Resolution | Native 2K |
| KREA 2 RAW | |
| Purpose | Undistilled mid-training checkpoint for LoRA fine-tuning and research |
| Inference Steps | Full (higher step count for maximum quality) |
| HuggingFace | krea/Krea-2-Raw |
| KREA 2 TURBO | |
| Purpose | Distilled inference engine for fast generation |
| Inference Steps | 8 steps (~2 seconds at 2K on consumer GPU) |
| HuggingFace | krea/Krea-2-Turbo |
| PLATFORM | |
| Artificial Analysis | #1 independent lab, #6 overall (as of June 2026) |
| ComfyUI | ComfyUI 0.26.0+ via standard diffusion model nodes |
| License | Krea 2 Community License (free commercial for individuals/small teams; enterprise inquire separately) |
| ComfyUI Access | Native support on Floyo (2 workflows) |
| Release Dates | May 12, 2026 (announcement) / May 18 (GA) / June 22-24 (open weights + technical report) |
What can you create with Krea 2?
Krea 2 covers photographic realism, product and advertising imagery, architectural visualization, design illustration, automotive photography, gaming asset styles, retail product shots, editorial content, concept art, fashion lookbooks, and moodboard-driven creative direction. The model handles all of these from text prompts without switching between separate specialized models. Style references and moodboards steer the aesthetic.
| Capability | What It Does | Use Case |
|---|---|---|
| Style Reference | Upload a single reference image and Krea 2 generates output that matches the aesthetic without copying the content. Show the model what you want rather than describing it in text. | Brand consistency, campaign direction, editorial style |
| Moodboard Generation | Curate a collection of images representing a visual direction. The model synthesizes themes, motifs, and styles across the full moodboard into new output. Works for open-ended aesthetics. | Creative direction, pitch decks, exploratory ideation |
| Aesthetic Diversity | The same prompt produces a wide range of distinct visual interpretations. Different angles, styles, moods, and compositions. Intentional by design, not a lack of consistency. | Brainstorming, creative exploration, concept variation |
| Turbo Generation | Krea 2 Turbo generates 2K images in about 2 seconds (8 inference steps). Fast enough to stay in creative flow without waiting between iterations. | Rapid iteration, live brainstorming, high-volume production |
| LoRA Fine-Tuning (Raw) | Krea 2 Raw is an undistilled mid-training checkpoint with malleable latent space. Fine-tune for custom styles, characters, and subjects with genuine room to move the model. | Custom styles, character IP, brand-specific training |
| Pipeline Integration | Chain with video models in ComfyUI. Generate with Krea 2, animate with Wan 2.7 or Vidu Q3, add voiceover with Fish Audio S2. Or convert to 3D with TRELLIS 2 or Meshy v6. | Multi-model production pipelines |
How does Krea 2 compare to other image models?
Krea 2 leads on aesthetic diversity and creative exploration among open-weight models. Ideogram V4 leads on text rendering (0.97 OCR) and layout control. GPT Image 2 leads on overall Elo and instruction fidelity. FLUX.2 dev leads on raw parameter count (32B). ERNIE Image leads on prompt enhancement at smaller scale. Krea 2's edge: the widest aesthetic range from a single prompt, zero synthetic training data, and style reference/moodboard control as a first-class feature.
| Model | Parameters | Training Data | Style Control | Turbo Speed |
|---|---|---|---|---|
| Krea 2 | 12.9B | Real only (zero synthetic) | Reference + moodboard | ~2s (8 steps, 2K) |
| Ideogram V4 | 9.3B | Structured JSON captions | Bounding-box + palette | Turbo preset available |
| FLUX.2 dev | 32B | Not disclosed | Via Kontext | FLUX Schnell |
| ERNIE Image | 8B | Not disclosed | Prompt enhancer | Turbo (8 steps) |
Source: Krea 2 Technical Report (June 23, 2026), Artificial Analysis text-to-image leaderboard, HuggingFace model cards, Medium analysis, BuildFastWithAI review, and ExplainX technical breakdown as of July 2026.
What are Krea 2's key features?
Krea 2's feature set is designed around a core philosophy: creative exploration over deterministic output. The model prioritizes breadth and visual surprise alongside quality and control. This is an intentional design decision, not a limitation. Every feature reflects this: diversity in output, style reference as the primary control, moodboards for open-ended direction, and a training set built from real images.
Aesthetic Diversity by Design
Most image models converge on a default interpretation of any given prompt. Krea 2 was built to diverge. Give it "horse rearing" and you get four distinct compositions: different styles, camera angles, lighting moods, and artistic treatments. This is not randomness. The training pipeline (pretraining, midtraining, SFT, preference optimization, reinforcement learning) was tuned to balance breadth with quality at every stage.
Style Reference System
Upload a single reference image and the model generates output that captures the aesthetic (color palette, texture, mood, composition style) without copying the content. This is the primary creative control mechanism. It replaces long style descriptions in text prompts with visual examples. The model deeply understands aesthetics because that was the explicit training objective.
Moodboard Generation
Not everything can be conveyed with a single reference. Curate a collection of images and the model synthesizes themes, motifs, and styles across the full set into new output. A moodboard combining "Cyber Zine" textures, "Minimalist Ink" linework, and "Thermal Airbrush" color grading produces images that blend all three directions. This maps directly to how creative teams already work.
Zero Synthetic Training Data
The pretraining set contains no AI-generated or synthetic images. This is unusual. Most frontier models include synthetic data to fill gaps or boost specific capabilities. Krea chose not to because synthetic data can flatten aesthetic range and create homogeneous output. The result is output that community members have called "the most uncensored nonlobotomized open source image model" available.
Two Checkpoints: Raw + Turbo
Krea 2 Raw is the undistilled mid-training checkpoint. It has a malleable latent space that accepts LoRA fine-tuning with genuine room to push the model toward new domains. Krea 2 Turbo is distilled to 8 inference steps and generates 2K images in about 2 seconds. Use Raw for maximum quality and training flexibility. Use Turbo when you need speed for rapid iteration or bulk production.
GRPO-Trained Prompt Expander
The built-in prompt expander was trained using Group Relative Policy Optimization, the same technique used for reasoning models. It enriches short prompts with visual detail before generation. This is why a three-word prompt like "samurai drawing blade" can produce a fully realized cinematic composition with dramatic lighting, wide-angle perspective, and atmospheric depth.
How does Krea 2 work?
Krea 2 is a flow-matching model with a dense 12.9B parameter single-stream Diffusion Transformer. The backbone uses 28 transformer blocks at width 6144, grouped-query attention with gated sigmoid attention, SwiGLU MLPs at 4x expansion, and 3D axial RoPE for positional encoding. The text encoder is Qwen3-VL-4B-Instruct with a multi-layer feature aggregation method that pulls semantic information from multiple intermediate layers.
The training pipeline has five stages: pretraining on curated real images, midtraining on higher-quality subsets, supervised fine-tuning (SFT) for alignment, preference optimization (PO) for human aesthetic preferences, and reinforcement learning (RL) for creative diversity. The multi-stage design is what gives the model its distinctive balance of quality and variety. Each stage refines a different aspect of output.
The Turbo variant uses distillation to compress the full inference pipeline from many steps to 8. The distillation preserves most of the quality while cutting generation time to about 2 seconds for a 2K image on consumer hardware. On Floyo's H100 NVL GPUs, both variants run faster than on consumer cards.
On Floyo, Krea 2 runs through native ComfyUI nodes. Two workflows are available: standard generation (Krea 2 for Text to Image) and fast generation (Krea 2 Turbo Text to Image). Both accept text prompts. Style reference and moodboard features are available through the Krea platform; on Floyo, the text-to-image pipeline handles generation directly.
Fair warning: Krea 2 prioritizes aesthetic diversity over deterministic consistency. If you need four nearly identical outputs from one prompt (product photography with fixed angles), a model like Nano Banana or GPT Image 2 may give more predictable results. Krea 2 is strongest when you want creative surprise: brainstorming, concept exploration, editorial imagery, and moodboard-driven work. The open weights use the Krea 2 Community License (free commercial for individuals and small teams; enterprise use requires a separate license from opensource@krea.ai).
Frequently Asked Questions
Common questions about running Krea 2 on Floyo.
You can start with Floyo's free pricing plan. To continue using the service beyond the free tier, upgrade your Floyo pricing plan. Krea 2 is open-weight, so there is no additional API cost beyond your Floyo plan.
Open Floyo in your browser, search "Krea 2" in the template library, and pick the standard or Turbo workflow. Click Run, write your prompt, and generate. Floyo handles the H100 GPU, ComfyUI environment, and 12.9B model weights. No local install, no Python setup.
Krea AI. The team includes Sangwu Lee, Erwann Millon, Le Zhuo, Matthew Newton, and others. Krea 2 was announced May 12, 2026, reached general availability May 18, and the open weights and 58-page technical report were published June 22-24, 2026. Krea 2 is Krea's first model trained entirely from scratch. Previously Krea operated as a SaaS aggregator of third-party models.
Raw is the undistilled mid-training checkpoint. Use it for maximum quality output and as a base for LoRA fine-tuning. Turbo is distilled to 8 inference steps and generates 2K images in about 2 seconds. Use it for rapid iteration and bulk production. Both share the same 12.9B architecture. Floyo has separate workflows for each.
Different strengths. Krea 2 leads on aesthetic diversity: the same prompt produces a wide range of distinct interpretations. Ideogram V4 leads on text rendering (0.97 OCR), layout control (bounding boxes), and deterministic design output. Use Krea 2 for creative exploration and moodboard-driven work. Use Ideogram V4 for typography, poster design, and brand assets. Both are available on Floyo.
Yes. Floyo runs ComfyUI, which lets you chain multiple models. Generate with Krea 2, animate with Wan 2.7 or Vidu Q3, add voiceover with Fish Audio S2 or ElevenLabs, convert to 3D with TRELLIS 2. All in one pipeline, all in your browser.
The Krea 2 Community License allows free commercial use for individuals and small teams. Enterprise use requires a separate license (contact opensource@krea.ai). On Floyo, your usage is governed by Floyo's terms alongside the model license.
No. The pretraining set contains zero synthetic or AI-generated images. Krea made this an explicit design choice because synthetic data can flatten aesthetic range and create homogeneous output. The model was trained on billions of curated real images to preserve the full breadth of visual diversity.
Try Krea 2 on Floyo
12.9B aesthetic-first image model trained from scratch on real images. Style references, moodboard generation, and 2K Turbo in 2 seconds. Run it in your browser.
Try Krea 2 Now → Browse All ModelsRelated Reading
AI Ad Creatives for Social and Web
Character and Concept Design on Floyo
Last updated: July 2026. Specs from Krea 2 Technical Report (June 23, 2026), Krea AI official blog, HuggingFace model cards (krea/Krea-2-Raw, krea/Krea-2-Turbo), GitHub (krea-ai/krea-2), Artificial Analysis text-to-image leaderboard, Medium analysis, BuildFastWithAI review, and ExplainX technical breakdown.
Run Krea 2 through ComfyUI on Floyo. 12.9B parameter aesthetic-first image model trained from scratch on real images only. Two variants: Raw for LoRA fine-tuning, Turbo for 2K images in 2 seconds. #1 among independent labs on Artificial Analysis. Free to try.

