How 20 AI image models render: “Exactly four ceramic mugs sit on a pale wooden shelf: two blue mugs on the left, one white mug in the center, and one red mug on the right. A green saucer is beneath the center mug, nothing else.”
Same prompt, sent verbatim, at each model's default settings and a fixed seed. Judge: every stated element present, correct count, correct placement. 184 community votes so far.
Fresh prompt from the daily stream · voting closed 2026-09-25
Checklist: 4 ceramic mugs · 2 blue mugs on the left · 1 white mug in the center · 1 red mug on the right · green saucer beneath the center mug · nothing else
20 text-to-image models rendered this prompt adherence prompt with the identical text at their default settings and a fixed seed. ImagineArt 1.5 Pro is the community favourite so far, winning 68% of its 19 head-to-head votes on this prompt. Grok Imagine Image Quality was the fastest at 5.0s.
What to look for: the right count of objects, the right spatial relationships, the stated colours and attributes, and nothing added.
- ImagineArt 1.5 Pro68%ImagineArt52.1s · 55 cr Try
- 5.0s · 80 cr Try
- Kling v3 Omni61%Kuaishou40.0s · 35 cr Try
- Nano Banana 2 Lite61%Google10.4s · 36 cr Try
- Krea 2 Large59%Krea30.1s · 80 cr Try
- Qwen Image 3.058%Alibaba11.3s · 32 cr Try
- 53%15.6s · 50 cr Try
- Seedream 5.0 Pro53%ByteDance32.4s · 38 cr Try
- Recraft v4 Pro47%Recraft18.5s · 300 cr Try
- Qwen Image 3.0 Pro44%Alibaba10.1s · 36 cr Try
- GPT Image 2.5 Sunburst42%OpenAI27.2s · 28 cr Try
- Recraft v442%Recraft9.1s · 45 cr Try
- ImagineArt 241%ImagineArt28.3s · 60 cr Try
- Muse Image 1.040%Meta8.5s · 12 cr Try
- GPT Image 2.5 Flare33%OpenAI20.0s · 28 cr Try
- Krea 2 Medium26%Krea17.5s · 40 cr Try
- Ideogram v425%Ideogram10.8s · 40 cr Try
- Midjourney Niji 725%Midjourney114.5s · 160 cr Try
- GPT Image 224%OpenAI54.1s · 60 cr Try
FAQ
- Which AI model renders this prompt best?
- ImagineArt 1.5 Pro leads this prompt in the BudgetPixel AI Image Arena, winning 68% of its 19 blind head-to-head votes. Vote on the arena page to add your judgement.
- Were the prompts adjusted per model?
- No. Every model received the exact same prompt, unmodified, at its default settings, with no negative prompt or style suffix. Some models rewrite prompts internally, and that is part of what is being compared.
- Can I run this prompt myself?
- Yes. Each tile’s “Try” link opens the BudgetPixel workshop with this prompt and that model preselected.