
COMMUNITY PAGE
Run FLUX.2 Max on Floyo
Home / Model / FLUX.2 Max on Floyo
AI IMAGE GENERATION
Run FLUX.2 Max on Floyo
The most capable model in the FLUX.2 family. 32B parameter architecture with Mistral-3 24B vision-language backbone + Rectified Flow Transformer. Web-grounded generation with real-time context, character consistency across scenes, multi-reference editing (up to 10 images), retexturing, and sub-10 second inference.
Run Black Forest Labs' FLUX.2 Max through ComfyUI in your browser. No API key, no installs, no local GPU.
|
Parameters ~32B (Mistral-3 VLM + Flow) |
Grounded Generation Real-time web context |
|
Resolution Up to 4 megapixel |
Speed Sub-10 seconds per image |
No installation. Runs in browser. Updated July 2026.
flux 2 max
photorealistic
text rendering
text to image
Write a prompt and FLUX.2 Max by Black Forest Labs generates the highest-quality image in the FLUX.2 family, with photorealistic detail, accurate text rendering, and strong prompt adherence.
FLUX.2 Max · Text to Image
Write a prompt and FLUX.2 Max by Black Forest Labs generates the highest-quality image in the FLUX.2 family, with photorealistic detail, accurate text rendering, and strong prompt adherence.
flux 2 max
image editing
image to image
multi-image
photorealistic
Upload an image to edit and up to eight reference images, describe the change, and FLUX.2 Max applies it with the highest editing consistency in the FLUX.2 family.
FLUX.2 Max Edit · Image to Image
Upload an image to edit and up to eight reference images, describe the change, and FLUX.2 Max applies it with the highest editing consistency in the FLUX.2 family.
What you get?
FLUX.2 Max is the flagship image generation and editing model from Black Forest Labs, the company founded by the original creators of Stable Diffusion. A ~32 billion parameter architecture coupling the Mistral-3 24B vision-language model with a Rectified Flow Transformer. The most capable variant in the FLUX.2 family (above Pro, Flex, and Klein). Web-grounded generation with real-time context, character consistency across images, multi-reference editing with up to 10 images, retexturing, spatial reasoning, text rendering, hex color control, and cinematic photorealism. Up to 4 megapixel output in sub-10 seconds. 32K token context window. Available as a ComfyUI API node on Floyo.
_1783065533774_1784225434247.webp?width=1400&height=620&quality=80&resize=cover)
_1783065533774_1784225434247.webp?width=1400&height=620&quality=80&resize=cover)
_1783065533774_1784225440855.webp?width=1400&height=620&quality=80&resize=cover)
_1783065533774_1784225440855.webp?width=1400&height=620&quality=80&resize=cover)
_1783067283308_1784225446513.webp?width=1400&height=620&quality=80&resize=cover)
_1783067283308_1784225446513.webp?width=1400&height=620&quality=80&resize=cover)


_1783065533774_1784225434247.webp?width=104&height=104&quality=80&resize=cover)
_1783065533774_1784225440855.webp?width=104&height=104&quality=80&resize=cover)
_1783067283308_1784225446513.webp?width=104&height=104&quality=80&resize=cover)

What is FLUX.2 Max?
FLUX.2 Max is the top-tier model in Black Forest Labs' FLUX.2 image generation family, released December 2025. It couples the Mistral-3 24B vision-language model with a Rectified Flow Transformer for a total architecture of approximately 32 billion parameters. It is the most capable FLUX variant for professional-grade image generation and editing, sitting above FLUX.2 Pro, Flex, and Klein.
The standout capability is web-grounded generation. FLUX.2 Max integrates real-time web context before generating images. Ask for a trending product, a current event, or the latest fashion and the model looks it up. This is the same concept as GPT Image 2's world knowledge and Seedream 5.0 Pro's web search, but built into the FLUX architecture. No manual reference sourcing needed for current subjects.
Character consistency is a first-class feature. Create a character once and maintain their facial features, proportions, expressions, and visual identity across images, scenes, styles, and complex edits. Upload up to 10 reference images simultaneously to lock product appearance, character identity, or brand style across an entire campaign. This multi-reference system is where FLUX.2 Max pulls ahead of single-reference competitors.
The FLUX.2 family is built by the original architects of Stable Diffusion. Robin Rombach, Andreas Blattmann, and team left Stability AI to found Black Forest Labs. The jump from FLUX.1 to FLUX.2 is not incremental. The architecture went from 12B parameter diffusion to 32B parameter flow matching with a full vision-language backbone. This is why FLUX.2 Max understands physics, spatial relationships, and complex multi-clause prompts at a level that FLUX.1 could not.
On Floyo, FLUX.2 Max runs through ComfyUI API nodes on H100 NVL GPUs. Write a prompt, optionally upload reference images, and generate. No BFL API key, no dashboard setup, no local hardware requirements.
What are FLUX.2 Max's technical specifications?
FLUX.2 Max couples the Mistral-3 24B vision-language model with a Rectified Flow Transformer for approximately 32B total parameters. 32K token context window for detailed multi-part prompts. Up to 4 megapixel output. Sub-10 second generation. Web-grounded generation with real-time context. Multi-reference editing with up to 10 images. Character consistency across scenes. Retexturing with geometry preservation. LM Arena score 1168.
| Spec | Details |
|---|---|
| Developer | Black Forest Labs (BFL), Freiburg, Germany |
| Architecture | Mistral-3 24B VLM + Rectified Flow Transformer (~32B total) |
| Context Window | 32K tokens (~46,864 tokens reported) |
| Max Resolution | Up to 4 megapixel |
| Speed | Sub-10 seconds per image |
| Web Grounding | Real-time web context integration (current events, products, trends) |
| Multi-Reference | Up to 8-10 reference images simultaneously |
| Character Consistency | Identity preservation across images, scenes, styles, and edits |
| Retexturing | Surface and material redesign with geometry, shape, and lighting preserved |
| Text Rendering | Typography, UI mockups, signage, logos (precise at small sizes) |
| Color Control | Hex code color specification with no approximation |
| LM Arena | Elo 1168 |
| Family | FLUX.2 Max > Pro > Flex > Klein |
| Founded By | Original creators of Stable Diffusion (Robin Rombach, Andreas Blattmann) |
| ComfyUI Access | API-based node on Floyo |
| Release Date | November 25, 2025 (FLUX.2 family) / December 2025 (Max) |
What can you create with FLUX.2 Max?
FLUX.2 Max covers product photography, e-commerce imagery, character design, brand identity, campaign assets, cinematic visuals, logo design, retexturing, filmmaking previsualization, UI mockups, and web-grounded generation of current products and trends. The multi-reference system and character consistency make it suited for campaign work where visual identity must hold across dozens of images.
| Capability | What It Does | Use Case |
|---|---|---|
| Web-Grounded Generation | Generate images with real-time web context. The model looks up current products, events, trends, and styles before rendering. No manual reference sourcing. | Trending product shots, current event visuals, seasonal campaigns |
| Character Consistency | Create a character once and preserve their facial features, proportions, expressions, and identity across images, scenes, and styles. Survives complex edits. | Brand ambassadors, recurring characters, campaign series |
| Multi-Reference Editing | Upload up to 10 reference images to lock product, character, or brand appearance. The model synthesizes identity from multiple angles and contexts. | Product catalogs, brand guidelines, visual identity systems |
| Retexturing | Redesign surfaces and materials while preserving shape, geometry, and lighting. Wood to marble, matte to chrome, fabric to leather. Precise and controlled. | Material exploration, product variants, interior design |
| Typography and UI | Precise text rendering at small sizes. Complex typography layouts, UI mockups, and logo designs. Hex color specification with no approximation. | Poster design, packaging mockups, app UI concepts |
| Pipeline Integration | Chain with video models in ComfyUI. Generate with FLUX.2 Max, animate with Wan 2.7 or Vidu Q3, add voiceover with ElevenLabs. Or convert to 3D with TRELLIS 2. | Multi-model production pipelines |
How does FLUX.2 Max compare to other image models?
FLUX.2 Max leads on multi-reference editing (up to 10 images), retexturing, and web-grounded generation among commercial API models. GPT Image 2 leads on overall Elo and batch consistency. Ideogram V4 leads on text rendering (0.97 OCR) and bounding-box layout. Seedream 5.0 Pro leads on grounded editing and layer separation. FLUX.2 Max's edge: the largest context window (32K tokens), strongest character consistency across edits, and the highest-capacity architecture in the FLUX family.
| Model | Parameters | Web Grounding | Multi-Ref | Retexturing |
|---|---|---|---|---|
| FLUX.2 Max | ~32B | Yes (real-time) | Up to 10 images | Yes (geometry-preserving) |
| GPT Image 2 | GPT-5.4 backbone | Knowledge cutoff + search | 8 batch output | Multi-turn editing |
| Ideogram V4 | 9.3B | No | Image-to-image | No |
| Seedream 5.0 Pro | Not disclosed | Yes (web search) | 2-10 references | Grounded editing |
Source: Black Forest Labs official documentation (bfl.ai), LM Arena leaderboard, VentureBeat launch coverage, AI Wiki FLUX.2 entry, WaveSpeed Blog, innFactory analysis, and MindStudio model card as of July 2026.
What are FLUX.2 Max's key features?
FLUX.2 Max's feature set is built on the most capable architecture Black Forest Labs has ever shipped. The Mistral-3 24B vision-language backbone gives it world knowledge and spatial reasoning. The Rectified Flow Transformer gives it photorealistic rendering. Combined, they produce an image model that understands what things look like, where they should go, and how they interact with light and physics.
Web-Grounded Generation
FLUX.2 Max integrates real-time web context into the generation process. Ask for a product that launched last week and the model looks it up. Request an image tied to a current event and the model pulls in relevant visual information. This eliminates the "stale training data" problem where models render outdated versions of products, logos, and public figures. No manual reference uploading needed for current subjects.
Character Consistency Across Scenes
Create a character and maintain their facial features, proportions, expressions, clothing, and visual identity across completely different images. Change the scene, change the style, change the lighting. The character stays recognizable. This works across complex edits, multiple references, and changing environments. For campaign work where a brand character or product ambassador appears across 50+ images, this is the feature that matters most.
Multi-Reference System (Up to 10 Images)
Upload up to 10 reference images simultaneously. The model synthesizes identity from multiple angles, contexts, and conditions. This is not single-image style transfer. It is multi-view identity extraction. A product photographed from 5 angles, a character in 3 outfits, and 2 environmental references can all feed into a single generation. The model understands what each reference contributes and composites them coherently.
Retexturing
Redesign surfaces and materials with precision. Swap leather to denim, wood to marble, matte to chrome. FLUX.2 Max preserves the shape, geometry, lighting, and spatial context of the original object while replacing only the surface material. This is a production feature: product teams can explore material variants without re-rendering or re-photographing.
32K Token Context Window
The largest context window among image generation models. Write detailed, multi-paragraph prompts with scene descriptions, character specifications, lighting directions, composition instructions, and style references. The model processes the full prompt without truncation. This is the Mistral-3 VLM backbone at work: it reads prompts the way a language model reads a document, with full contextual understanding across the entire input.
Photorealistic Physics and Lighting
Fabric textures with visible weave. Architectural materials with correct reflection properties. Skin with subsurface scattering. Lighting that follows physical rules: shadows fall correctly, reflections match the environment, and specular highlights respond to material roughness. The gap between FLUX.2 Max output and real photography is narrow enough that the model is used for marketplace product images that look indistinguishable from studio shoots.
How does FLUX.2 Max work?
FLUX.2 Max couples two architectures: a Mistral-3 24B vision-language model for understanding and a Rectified Flow Transformer for rendering. The VLM processes your prompt (and any reference images) with full contextual understanding, including world knowledge, physics, and spatial reasoning. The flow transformer converts that understanding into a photorealistic image through a learned transport path from noise to signal.
The Mistral-3 backbone is what separates FLUX.2 from models using CLIP or T5 text encoders. A 24B VLM reads your prompt the way a language model reads an essay: understanding relationships between clauses, interpreting ambiguity, and maintaining coherence across long, complex instructions. The 32K token context window means no truncation even for highly detailed production prompts.
Web-grounded generation works by querying current web context before the generation step. The model fetches relevant information about subjects, products, events, or styles that exist in the real world and integrates that context into the visual output. This happens automatically when the model detects that current real-world knowledge would improve the output.
On Floyo, FLUX.2 Max runs through ComfyUI API nodes on H100 NVL GPUs. Your prompt and optional reference images are sent to inference servers, and the generated image returns to your ComfyUI canvas. You can chain it with video models (Wan 2.7, Vidu Q3, Seedance 2.0 Mini), 3D generators (TRELLIS 2, Meshy v6), and audio models (ElevenLabs, Fish Audio S2) in the same workflow.
Fair warning: FLUX.2 Max is API-based, not open source. The open-weight variant is FLUX.2 Dev (32B, non-commercial license for self-hosting). Max is the premium closed tier with the highest quality and all features (web grounding, full multi-reference). API pricing applies through your Floyo API Wallet. The FLUX.2 family requires ISO 27001 and SOC 2 Type II compliance for enterprise deployments, which BFL has obtained.
Frequently Asked Questions
Common questions about running FLUX.2 Max on Floyo.
You can start with Floyo's free pricing plan. Floyo gives $0.25 in free API credits on signup. To continue using the service beyond the free tier, upgrade your Floyo pricing plan. FLUX.2 Max runs as an API node, so generation costs come from your API Wallet (separate from your plan's GPU time).
Open Floyo in your browser, search "FLUX" in the template library, and pick a FLUX.2 Max workflow. Click Run, write your prompt, optionally upload reference images, and generate. Floyo handles the ComfyUI environment and API connection. No BFL dashboard, no API key management, no local install.
Black Forest Labs (BFL), a German AI company founded by the original creators of Stable Diffusion, including Robin Rombach and Andreas Blattmann. The FLUX.2 family launched November 25, 2025. FLUX.2 Max is the flagship top-tier variant. BFL holds ISO 27001 and SOC 2 Type II certifications for enterprise deployment.
Max is the top tier: highest quality, web grounding, strongest editing consistency, and all features. Pro is production-grade at a lower price. Flex specializes in typography and fine detail preservation. Klein is optimized for speed and rapid prototyping. Use Max for final commercial assets. Use Klein for fast iteration during concepting.
FLUX.2 Max leads on multi-reference editing (10 images vs batch output), retexturing, and explicit web-grounded generation. GPT Image 2 leads on overall Elo, batch consistency (8 images per prompt), and conversational multi-turn editing. Both have world knowledge and physics understanding. FLUX.2 Max uses a larger context window (32K tokens). Choose based on whether you need reference-based editing (FLUX) or conversational iteration (GPT Image 2).
Yes. Floyo runs ComfyUI, which lets you chain multiple models. Generate with FLUX.2 Max, animate with Wan 2.7 or Vidu Q3, add voiceover with ElevenLabs or Fish Audio S2, convert to 3D with TRELLIS 2 or Meshy v6. All in one pipeline, all in your browser.
The model searches the web in real time before generating. If you ask for a current product, trending style, or recent event, it fetches relevant visual and factual context and integrates it into the output. This means generated images reflect current reality, not stale training data. No manual reference uploading needed for real-world subjects.
Yes. FLUX.2 Max is a commercial API product from Black Forest Labs. Generated images can be used for products, marketing, client work, and any commercial context under BFL's terms of service. On Floyo, your usage is governed by Floyo's terms alongside the BFL license.
Try FLUX.2 Max on Floyo
Black Forest Labs' flagship image model. 32B parameters, web-grounded generation, 10-image multi-reference, character consistency, retexturing, and 4MP output. Run it in your browser.
Try FLUX.2 Max Now → Browse All ModelsRelated Reading
AI Ad Creatives for Social and Web
Character and Concept Design on Floyo
Last updated: July 2026. Specs from Black Forest Labs official documentation (bfl.ai), FLUX.2 launch blog (November 25, 2025), LM Arena leaderboard, VentureBeat launch coverage, AI Wiki FLUX.2 entry, WaveSpeed Blog analysis, innFactory model breakdown, MindStudio model card, and Digital Applied production guide.
Run FLUX.2 Max online through ComfyUI on Floyo. Black Forest Labs' flagship ~32B parameter image model with web-grounded generation, 10-image multi-reference editing, character consistency, retexturing, 4MP photorealism, and 32K token context window. Built by the creators of Stable Diffusion. Sub-10 second generation. No install, no GPU, browser-based. Free to try.
_1783065533774.png?width=400&height=300&quality=80&resize=cover)
_1783067283308.png?width=400&height=300&quality=80&resize=cover)