Floyo
Floyo
Workflows
API
Pricing
Floyo
Floyo
Workflows
API
Pricing
Run FLUX.2 Max on Floyo hero

COMMUNITY PAGE

Run FLUX.2 Max on Floyo

Home / Model / FLUX.2 Max on Floyo

AI IMAGE GENERATION

Run FLUX.2 Max on Floyo

The most capable model in the FLUX.2 family. 32B parameter architecture with Mistral-3 24B vision-language backbone + Rectified Flow Transformer. Web-grounded generation with real-time context, character consistency across scenes, multi-reference editing (up to 10 images), retexturing, and sub-10 second inference.

Run Black Forest Labs' FLUX.2 Max through ComfyUI in your browser. No API key, no installs, no local GPU.

Parameters

~32B (Mistral-3 VLM + Flow)

Grounded Generation

Real-time web context

Resolution

Up to 4 megapixel

Speed

Sub-10 seconds per image

No installation. Runs in browser. Updated July 2026.

FLUX.2 Max · Text to Image

flux 2 max

photorealistic

text rendering

text to image

Write a prompt and FLUX.2 Max by Black Forest Labs generates the highest-quality image in the FLUX.2 family, with photorealistic detail, accurate text rendering, and strong prompt adherence.

FLUX.2 Max · Text to Image

Write a prompt and FLUX.2 Max by Black Forest Labs generates the highest-quality image in the FLUX.2 family, with photorealistic detail, accurate text rendering, and strong prompt adherence.

FLUX.2 Max Edit · Image to Image

flux 2 max

image editing

image to image

multi-image

photorealistic

Upload an image to edit and up to eight reference images, describe the change, and FLUX.2 Max applies it with the highest editing consistency in the FLUX.2 family.

FLUX.2 Max Edit · Image to Image

Upload an image to edit and up to eight reference images, describe the change, and FLUX.2 Max applies it with the highest editing consistency in the FLUX.2 family.

What you get? 

FLUX.2 Max is the flagship image generation and editing model from Black Forest Labs, the company founded by the original creators of Stable Diffusion. A ~32 billion parameter architecture coupling the Mistral-3 24B vision-language model with a Rectified Flow Transformer. The most capable variant in the FLUX.2 family (above Pro, Flex, and Klein). Web-grounded generation with real-time context, character consistency across images, multi-reference editing with up to 10 images, retexturing, spatial reasoning, text rendering, hex color control, and cinematic photorealism. Up to 4 megapixel output in sub-10 seconds. 32K token context window. Available as a ComfyUI API node on Floyo.

What is FLUX.2 Max?

FLUX.2 Max is the top-tier model in Black Forest Labs' FLUX.2 image generation family, released December 2025. It couples the Mistral-3 24B vision-language model with a Rectified Flow Transformer for a total architecture of approximately 32 billion parameters. It is the most capable FLUX variant for professional-grade image generation and editing, sitting above FLUX.2 Pro, Flex, and Klein.

The standout capability is web-grounded generation. FLUX.2 Max integrates real-time web context before generating images. Ask for a trending product, a current event, or the latest fashion and the model looks it up. This is the same concept as GPT Image 2's world knowledge and Seedream 5.0 Pro's web search, but built into the FLUX architecture. No manual reference sourcing needed for current subjects.

Character consistency is a first-class feature. Create a character once and maintain their facial features, proportions, expressions, and visual identity across images, scenes, styles, and complex edits. Upload up to 10 reference images simultaneously to lock product appearance, character identity, or brand style across an entire campaign. This multi-reference system is where FLUX.2 Max pulls ahead of single-reference competitors.

The FLUX.2 family is built by the original architects of Stable Diffusion. Robin Rombach, Andreas Blattmann, and team left Stability AI to found Black Forest Labs. The jump from FLUX.1 to FLUX.2 is not incremental. The architecture went from 12B parameter diffusion to 32B parameter flow matching with a full vision-language backbone. This is why FLUX.2 Max understands physics, spatial relationships, and complex multi-clause prompts at a level that FLUX.1 could not.

On Floyo, FLUX.2 Max runs through ComfyUI API nodes on H100 NVL GPUs. Write a prompt, optionally upload reference images, and generate. No BFL API key, no dashboard setup, no local hardware requirements.

What are FLUX.2 Max's technical specifications?

FLUX.2 Max couples the Mistral-3 24B vision-language model with a Rectified Flow Transformer for approximately 32B total parameters. 32K token context window for detailed multi-part prompts. Up to 4 megapixel output. Sub-10 second generation. Web-grounded generation with real-time context. Multi-reference editing with up to 10 images. Character consistency across scenes. Retexturing with geometry preservation. LM Arena score 1168.

Spec Details
DeveloperBlack Forest Labs (BFL), Freiburg, Germany
ArchitectureMistral-3 24B VLM + Rectified Flow Transformer (~32B total)
Context Window32K tokens (~46,864 tokens reported)
Max ResolutionUp to 4 megapixel
SpeedSub-10 seconds per image
Web GroundingReal-time web context integration (current events, products, trends)
Multi-ReferenceUp to 8-10 reference images simultaneously
Character ConsistencyIdentity preservation across images, scenes, styles, and edits
RetexturingSurface and material redesign with geometry, shape, and lighting preserved
Text RenderingTypography, UI mockups, signage, logos (precise at small sizes)
Color ControlHex code color specification with no approximation
LM ArenaElo 1168
FamilyFLUX.2 Max > Pro > Flex > Klein
Founded ByOriginal creators of Stable Diffusion (Robin Rombach, Andreas Blattmann)
ComfyUI AccessAPI-based node on Floyo
Release DateNovember 25, 2025 (FLUX.2 family) / December 2025 (Max)

What can you create with FLUX.2 Max?

FLUX.2 Max covers product photography, e-commerce imagery, character design, brand identity, campaign assets, cinematic visuals, logo design, retexturing, filmmaking previsualization, UI mockups, and web-grounded generation of current products and trends. The multi-reference system and character consistency make it suited for campaign work where visual identity must hold across dozens of images.

Capability What It Does Use Case
Web-Grounded GenerationGenerate images with real-time web context. The model looks up current products, events, trends, and styles before rendering. No manual reference sourcing.Trending product shots, current event visuals, seasonal campaigns
Character ConsistencyCreate a character once and preserve their facial features, proportions, expressions, and identity across images, scenes, and styles. Survives complex edits.Brand ambassadors, recurring characters, campaign series
Multi-Reference EditingUpload up to 10 reference images to lock product, character, or brand appearance. The model synthesizes identity from multiple angles and contexts.Product catalogs, brand guidelines, visual identity systems
RetexturingRedesign surfaces and materials while preserving shape, geometry, and lighting. Wood to marble, matte to chrome, fabric to leather. Precise and controlled.Material exploration, product variants, interior design
Typography and UIPrecise text rendering at small sizes. Complex typography layouts, UI mockups, and logo designs. Hex color specification with no approximation.Poster design, packaging mockups, app UI concepts
Pipeline IntegrationChain with video models in ComfyUI. Generate with FLUX.2 Max, animate with Wan 2.7 or Vidu Q3, add voiceover with ElevenLabs. Or convert to 3D with TRELLIS 2.Multi-model production pipelines

How does FLUX.2 Max compare to other image models?

FLUX.2 Max leads on multi-reference editing (up to 10 images), retexturing, and web-grounded generation among commercial API models. GPT Image 2 leads on overall Elo and batch consistency. Ideogram V4 leads on text rendering (0.97 OCR) and bounding-box layout. Seedream 5.0 Pro leads on grounded editing and layer separation. FLUX.2 Max's edge: the largest context window (32K tokens), strongest character consistency across edits, and the highest-capacity architecture in the FLUX family.

Model Parameters Web Grounding Multi-Ref Retexturing
FLUX.2 Max ~32B Yes (real-time) Up to 10 images Yes (geometry-preserving)
GPT Image 2 GPT-5.4 backbone Knowledge cutoff + search 8 batch output Multi-turn editing
Ideogram V4 9.3B No Image-to-image No
Seedream 5.0 Pro Not disclosed Yes (web search) 2-10 references Grounded editing

Source: Black Forest Labs official documentation (bfl.ai), LM Arena leaderboard, VentureBeat launch coverage, AI Wiki FLUX.2 entry, WaveSpeed Blog, innFactory analysis, and MindStudio model card as of July 2026.

What are FLUX.2 Max's key features?

FLUX.2 Max's feature set is built on the most capable architecture Black Forest Labs has ever shipped. The Mistral-3 24B vision-language backbone gives it world knowledge and spatial reasoning. The Rectified Flow Transformer gives it photorealistic rendering. Combined, they produce an image model that understands what things look like, where they should go, and how they interact with light and physics.

Web-Grounded Generation

FLUX.2 Max integrates real-time web context into the generation process. Ask for a product that launched last week and the model looks it up. Request an image tied to a current event and the model pulls in relevant visual information. This eliminates the "stale training data" problem where models render outdated versions of products, logos, and public figures. No manual reference uploading needed for current subjects.

Character Consistency Across Scenes

Create a character and maintain their facial features, proportions, expressions, clothing, and visual identity across completely different images. Change the scene, change the style, change the lighting. The character stays recognizable. This works across complex edits, multiple references, and changing environments. For campaign work where a brand character or product ambassador appears across 50+ images, this is the feature that matters most.

Multi-Reference System (Up to 10 Images)

Upload up to 10 reference images simultaneously. The model synthesizes identity from multiple angles, contexts, and conditions. This is not single-image style transfer. It is multi-view identity extraction. A product photographed from 5 angles, a character in 3 outfits, and 2 environmental references can all feed into a single generation. The model understands what each reference contributes and composites them coherently.

Retexturing

Redesign surfaces and materials with precision. Swap leather to denim, wood to marble, matte to chrome. FLUX.2 Max preserves the shape, geometry, lighting, and spatial context of the original object while replacing only the surface material. This is a production feature: product teams can explore material variants without re-rendering or re-photographing.

32K Token Context Window

The largest context window among image generation models. Write detailed, multi-paragraph prompts with scene descriptions, character specifications, lighting directions, composition instructions, and style references. The model processes the full prompt without truncation. This is the Mistral-3 VLM backbone at work: it reads prompts the way a language model reads a document, with full contextual understanding across the entire input.

Photorealistic Physics and Lighting

Fabric textures with visible weave. Architectural materials with correct reflection properties. Skin with subsurface scattering. Lighting that follows physical rules: shadows fall correctly, reflections match the environment, and specular highlights respond to material roughness. The gap between FLUX.2 Max output and real photography is narrow enough that the model is used for marketplace product images that look indistinguishable from studio shoots.

How does FLUX.2 Max work?

FLUX.2 Max couples two architectures: a Mistral-3 24B vision-language model for understanding and a Rectified Flow Transformer for rendering. The VLM processes your prompt (and any reference images) with full contextual understanding, including world knowledge, physics, and spatial reasoning. The flow transformer converts that understanding into a photorealistic image through a learned transport path from noise to signal.

The Mistral-3 backbone is what separates FLUX.2 from models using CLIP or T5 text encoders. A 24B VLM reads your prompt the way a language model reads an essay: understanding relationships between clauses, interpreting ambiguity, and maintaining coherence across long, complex instructions. The 32K token context window means no truncation even for highly detailed production prompts.

Web-grounded generation works by querying current web context before the generation step. The model fetches relevant information about subjects, products, events, or styles that exist in the real world and integrates that context into the visual output. This happens automatically when the model detects that current real-world knowledge would improve the output.

On Floyo, FLUX.2 Max runs through ComfyUI API nodes on H100 NVL GPUs. Your prompt and optional reference images are sent to inference servers, and the generated image returns to your ComfyUI canvas. You can chain it with video models (Wan 2.7, Vidu Q3, Seedance 2.0 Mini), 3D generators (TRELLIS 2, Meshy v6), and audio models (ElevenLabs, Fish Audio S2) in the same workflow.

Fair warning: FLUX.2 Max is API-based, not open source. The open-weight variant is FLUX.2 Dev (32B, non-commercial license for self-hosting). Max is the premium closed tier with the highest quality and all features (web grounding, full multi-reference). API pricing applies through your Floyo API Wallet. The FLUX.2 family requires ISO 27001 and SOC 2 Type II compliance for enterprise deployments, which BFL has obtained.

Frequently Asked Questions

Common questions about running FLUX.2 Max on Floyo.

Is FLUX.2 Max free to use on Floyo?

You can start with Floyo's free pricing plan. Floyo gives $0.25 in free API credits on signup. To continue using the service beyond the free tier, upgrade your Floyo pricing plan. FLUX.2 Max runs as an API node, so generation costs come from your API Wallet (separate from your plan's GPU time).

How do I run FLUX.2 Max without installing anything?

Open Floyo in your browser, search "FLUX" in the template library, and pick a FLUX.2 Max workflow. Click Run, write your prompt, optionally upload reference images, and generate. Floyo handles the ComfyUI environment and API connection. No BFL dashboard, no API key management, no local install.

Who made FLUX.2 Max?

Black Forest Labs (BFL), a German AI company founded by the original creators of Stable Diffusion, including Robin Rombach and Andreas Blattmann. The FLUX.2 family launched November 25, 2025. FLUX.2 Max is the flagship top-tier variant. BFL holds ISO 27001 and SOC 2 Type II certifications for enterprise deployment.

What is the difference between FLUX.2 Max, Pro, Flex, and Klein?

Max is the top tier: highest quality, web grounding, strongest editing consistency, and all features. Pro is production-grade at a lower price. Flex specializes in typography and fine detail preservation. Klein is optimized for speed and rapid prototyping. Use Max for final commercial assets. Use Klein for fast iteration during concepting.

How does FLUX.2 Max compare to GPT Image 2?

FLUX.2 Max leads on multi-reference editing (10 images vs batch output), retexturing, and explicit web-grounded generation. GPT Image 2 leads on overall Elo, batch consistency (8 images per prompt), and conversational multi-turn editing. Both have world knowledge and physics understanding. FLUX.2 Max uses a larger context window (32K tokens). Choose based on whether you need reference-based editing (FLUX) or conversational iteration (GPT Image 2).

Can I combine FLUX.2 Max with other AI models in one workflow?

Yes. Floyo runs ComfyUI, which lets you chain multiple models. Generate with FLUX.2 Max, animate with Wan 2.7 or Vidu Q3, add voiceover with ElevenLabs or Fish Audio S2, convert to 3D with TRELLIS 2 or Meshy v6. All in one pipeline, all in your browser.

What is web-grounded generation?

The model searches the web in real time before generating. If you ask for a current product, trending style, or recent event, it fetches relevant visual and factual context and integrates it into the output. This means generated images reflect current reality, not stale training data. No manual reference uploading needed for real-world subjects.

Can I use FLUX.2 Max output commercially?

Yes. FLUX.2 Max is a commercial API product from Black Forest Labs. Generated images can be used for products, marketing, client work, and any commercial context under BFL's terms of service. On Floyo, your usage is governed by Floyo's terms alongside the BFL license.

Try FLUX.2 Max on Floyo

Black Forest Labs' flagship image model. 32B parameters, web-grounded generation, 10-image multi-reference, character consistency, retexturing, and 4MP output. Run it in your browser.

Try FLUX.2 Max Now → Browse All Models

Related Reading

AI Ad Creatives for Social and Web

Character and Concept Design on Floyo

Top AI Models on Floyo

Last updated: July 2026. Specs from Black Forest Labs official documentation (bfl.ai), FLUX.2 launch blog (November 25, 2025), LM Arena leaderboard, VentureBeat launch coverage, AI Wiki FLUX.2 entry, WaveSpeed Blog analysis, innFactory model breakdown, MindStudio model card, and Digital Applied production guide.

Table of Contents
OVERVIEW

Run FLUX.2 Max online through ComfyUI on Floyo. Black Forest Labs' flagship ~32B parameter image model with web-grounded generation, 10-image multi-reference editing, character consistency, retexturing, 4MP photorealism, and 32K token context window. Built by the creators of Stable Diffusion. Sub-10 second generation. No install, no GPU, browser-based. Free to try.