Skip to main content

Flux Schnell vs Dev vs Pro vs Max (2026): Which to Use

By Johnathan Kwok · VibeDex Research

TL;DR

FLUX.2 Max is the strongest current Flux-family model in the public Sonnet leaderboard. It posts a VibeDex Score (our 0–5 blind-benchmark score) of 3.965 versus FLUX.2 Pro at 3.900 — a 0.065 gap that sits within the benchmark's noise band, so treat the two as a near-tie. Flux Dev (3.504) and Flux Schnell (3.350) trail clearly. FLUX.2 Pro remains the practical value pick because it costs half as much as Max.

prompt-0109

One prompt, four Flux tiers — click any image to zoom

High fashion editorial photograph of a model emerging from a swimming pool at twilight, water cascading off a metallic gold lamé gown... (full...

Flux Schnell ($0.001) - High fashion editorial photograph of a model emerging from a swimming pool at twilight, water cascading off a metallic gold lamé gown... (full 50-prompt benchmark prompt)

Flux Schnell ($0.001)

Flux Dev ($0.003) - High fashion editorial photograph of a model emerging from a swimming pool at twilight, water cascading off a metallic gold lamé gown... (full 50-prompt benchmark prompt)

Flux Dev ($0.003)

FLUX.2 Pro ($0.035) - High fashion editorial photograph of a model emerging from a swimming pool at twilight, water cascading off a metallic gold lamé gown... (full 50-prompt benchmark prompt)

FLUX.2 Pro ($0.035)

FLUX.2 Max ($0.070) - High fashion editorial photograph of a model emerging from a swimming pool at twilight, water cascading off a metallic gold lamé gown... (full 50-prompt benchmark prompt)

FLUX.2 Max ($0.070)

Current Flux Ladder

This table includes Flux-family models that appear in the current public Sonnet image leaderboard. FLUX 1.1 Pro remains in legacy data, but is not mixed into this current view. Note the shape of the ladder: Max and Pro are separated by just 0.065 points — within the benchmark's noise band — while the drop from Pro to Dev (0.396) is the real quality step.

#ModelVibeDex ScoreCost/ImageTier
1FLUX.2 Max3.96$0.070Premium
2FLUX.2 Pro3.90$0.035Standard
3Flux Dev3.50$0.003Budget
4Flux Schnell3.35$0.001Budget

Same Prompt, Four Tiers

Every Flux tier was run on the same 50-prompt set in our blind benchmark. Below are two more prompts from that set rendered by all four models — a luxury-watch product shot where fine detail separates the tiers, and a neon night scene where even the budget tiers hold up. Click any image for the full-size render and judge the gap yourself.

prompt-0073

Ultra high resolution product photograph of a luxury watch, every microscopic detail of the dial visible, sapphire crystal catching light, brushed...

Flux Schnell ($0.001) - Ultra high resolution product photograph of a luxury watch, every microscopic detail of the dial visible, sapphire crystal catching light, brushed steel texture on case, leather strap grain visible, zero noise, magazine quality

Flux Schnell ($0.001)

Flux Dev ($0.003) - Ultra high resolution product photograph of a luxury watch, every microscopic detail of the dial visible, sapphire crystal catching light, brushed steel texture on case, leather strap grain visible, zero noise, magazine quality

Flux Dev ($0.003)

FLUX.2 Pro ($0.035) - Ultra high resolution product photograph of a luxury watch, every microscopic detail of the dial visible, sapphire crystal catching light, brushed steel texture on case, leather strap grain visible, zero noise, magazine quality

FLUX.2 Pro ($0.035)

FLUX.2 Max ($0.070) - Ultra high resolution product photograph of a luxury watch, every microscopic detail of the dial visible, sapphire crystal catching light, brushed steel texture on case, leather strap grain visible, zero noise, magazine quality

FLUX.2 Max ($0.070)

prompt-0063

Cyberpunk Tokyo alley at night, neon signs reflecting on wet pavement, cinematic color grading

Flux Schnell ($0.001) - Cyberpunk Tokyo alley at night, neon signs reflecting on wet pavement, cinematic color grading

Flux Schnell ($0.001)

Flux Dev ($0.003) - Cyberpunk Tokyo alley at night, neon signs reflecting on wet pavement, cinematic color grading

Flux Dev ($0.003)

FLUX.2 Pro ($0.035) - Cyberpunk Tokyo alley at night, neon signs reflecting on wet pavement, cinematic color grading

FLUX.2 Pro ($0.035)

FLUX.2 Max ($0.070) - Cyberpunk Tokyo alley at night, neon signs reflecting on wet pavement, cinematic color grading

FLUX.2 Max ($0.070)

Renders are the actual benchmark outputs from GCS. Per-prompt judge scores are not shown here; the citable numbers are the model-level VibeDex Scores in the ladder table above.

Which Flux Should You Use?

FLUX.2 Max

Strengths

  • +Highest current Sonnet score in the Flux family
  • +Best pick when Flux-family quality matters most

Limitations

  • Costs twice as much as FLUX.2 Pro
  • Still below the overall public leaderboard top cluster

FLUX.2 Pro

Strengths

  • +Best practical Flux value
  • +Half the listed cost of FLUX.2 Max

Limitations

  • Trails Max by 0.065 points (within the noise band)
  • No longer a top-five public leaderboard model

Flux Dev / Schnell

Strengths

  • +Low listed costs for experimentation
  • +Useful when Flux ecosystem compatibility matters

Limitations

  • Both trail the current FLUX.2 models clearly
  • Not strong enough for final-image quality compared with current leaders

Related Vibedex Benchmarks

Methodology: Rankings and scores in this article align to VibeDex's current Sonnet 4.6 blind benchmark: 50 prompts, 3 passes, and 150 judgments per model across visual fidelity, physics, subject integrity, and instruction adherence. See our full methodology

FAQ

Which Flux model is best in the current benchmark?

FLUX.2 Max leads the current Sonnet Flux-family comparison at 3.965.

Is FLUX.2 Pro still the value pick?

Yes for most paid Flux use cases. It costs half as much as FLUX.2 Max while trailing by 0.065 points.

Why is FLUX 1.1 Pro missing?

FLUX 1.1 Pro is not present in the current public Sonnet leaderboard, so this rewrite does not mix it into the current comparison.

See how every model stacks up

The Vibedex leaderboard ranks 18 image models on a 50-prompt blind benchmark, judged by Claude Sonnet 4.6 across visual fidelity, physics, subject integrity, and instruction adherence.

See the leaderboard