🔬 Tested by TokenRouter

Same prompt. Every model. Judge for yourself.

We run a fixed set of prompts against every model in the catalog and keep the outputs — so you can compare what they actually produce, not what the docs claim.

The prompt Create a clean infographic: a vertical bar chart titled 'HIFIBOTS USERS' with exactly three labeled bars: Jan = 10, Feb = 25, Mar = 40. Show each label and value as legible text.
flux-2-flex black-forest-labs ✓ passed 12s
flux-2-max black-forest-labs ✓ passed 23s
flux-2-pro black-forest-labs ✓ passed 11s
flux-kontext-max black-forest-labs ✓ passed 143s
flux-kontext-pro black-forest-labs ✗ failed 6s
FLUX.1 Kontext [max] black-forest-labs ✗ failed 9s
FLUX.1 Kontext [pro] black-forest-labs ✓ passed 6s
FLUX1.1 [pro] black-forest-labs ✗ failed 3s
FLUX.2 [dev] black-forest-labs ✓ passed 7s
FLUX.2 [flex] black-forest-labs ✓ passed 11s
FLUX.2 [max] black-forest-labs ✓ passed 17s
seedream-4-0-250828 bytedance ✓ passed 8s
gemini-3-pro-image google ✓ passed 17s $0.16224
gemini-3.1-flash-image google ✓ passed 6s $0.09078
gemini-3.1-flash-lite-image google ✓ passed 4s
nano-banana-pro-preview google ✓ passed 16s
OpenAI: GPT-5 Image openai ✓ passed 48s $0.20619
OpenAI: GPT-5 Image Mini openai ✓ passed 41s $0.04204
OpenAI: GPT-5.4 Image 2 openai ✓ passed 108s $0.22668
gpt-image-1 openai ✓ passed 28s
gpt-image-1-mini openai ✓ passed 33s
gpt-image-1.5 openai ✓ passed 37s
gpt-image-1.5-2025-12-16 openai ✓ passed 28s
gpt-image-2 openai ✓ passed 24s
gpt-image-2-all openai ✓ passed 26s
gpt-image-2-vip openai ✓ passed 46s
Qwen Image qwen ✓ passed 10s
Juggernaut Lightning Flux by RunDiffusion rundiffusion ✓ passed 4s
Juggernaut Pro Flux by RunDiffusion 1.0.0 rundiffusion ✗ failed 4s
stabilityai/stable-diffusion-3.5-medium stabilityai ✗ failed 7s $0.025
SD XL stabilityai ✗ failed 4s
Tongyi-MAI/Z-Image-Turbo tongyi-mai ✓ passed 4s $0.025

Outputs generated 2026-08-26. Costs are measured from the provider's own billing where reported. Nothing here is simulated.