Category

The best AI for generating images

GPT Image 2 is first today, because it is the only model that wins both Arena boards: generating from text and editing. But if cost matters, Nano Banana 2 gets close for less, and if it has to run on your own hardware, exactly one row of this table is open to you.

  • six models, all ranked
  • numbers from vendor docs and leaderboards
  • an Iran column with a stated method

Last checked: This is a reference page. It is re-checked against the vendor sources and updated when a new version ships. What changed

Today ranking, built on what has actually been published

Both board scores are read from Arena; prices and specifications come from each vendor own documentation. An empty cell means nothing was published, not that the number is zero.

The best AI for generating images
Rank Tool Score Text to image 30 Image editing 20 Price per image 20 Reference inputs 10 Weight openness 10 Access from Iran 10
1 GPT Image 2 Current pick 73.5 1,381 Source 1,463 Source $0.053 per image Source not verified 0/3 No weights published. The model page lists three API endpoints and no download route of any kind blocked / no working route We read the supported-countries list for the OpenAI API on developers.openai.com today and Iran is not on it. openai.com itself returns 403 to this server, so that list is the only readable document, and it speaks about the API rather than the ChatGPT app. No network test was run from inside Iran.
2 Nano Banana 2 65.5 1,264 Source 1,385 Source $0.067 per image Source 10 reference inputs Source 0/3 No weights published, and Google offers no download route for any Gemini image model blocked / no working route The Gemini API available-regions page lists the countries where the API works, and Iran is not on that list. That is reading a page, not measuring a network.
3 Nano Banana Pro 46.4 1,246 Source 1,389 Source $0.134 per image Source 6 reference inputs Source 0/3 No weights published, and Google offers no download route for any Gemini image model blocked / no working route The Gemini API available-regions page lists the countries where the API works, and Iran is not on that list. That is reading a page, not measuring a network.
4 Seedream 5.0 Pro 34.4 1,258 Source 1,393 Source not verified not verified not verified not verified
5 FLUX.2 [max] 34.0 1,162 Source 1,262 Source $0.070 per image Source 8 reference inputs Source 0/3 No weights for this member. Black Forest Labs publishes klein weights and not max, and says so in its own comparison table not verified
6 FLUX.2 [klein] 4B 30.0 1,030 Source 1,188 Source $0.014 per image Source 4 reference inputs Source 3/3 Weights under Apache 2.0 on Hugging Face, plus an undistilled Base variant the vendor itself points at for fine-tuning and LoRA, with a published training guide not verified

An empty cell means we could not verify that number, not that the tool scored zero.

How this ranking is calculated

Every criterion below has a weight and a source. Change a weight and the whole table recomputes. There is no hand-placed position anywhere in this hub.

Criterion Weight Evidence
Text to image 30 Arena (Text to Image)Human preference on generating an image from a text prompt. It carries the most weight because it is what most people want from these models
Image editing 20 Arena (Image Edit)Human preference on editing an existing image. Counted separately because the two boards do not agree: several models sit higher on one and lower on the other
Price per image 20 vendor stated specificationThe cost of one image at the cheapest size the vendor publishes a number for. Lower is better. Every cell is tied to a specific size, and that condition is written on each model page
Reference inputs 10 vendor stated specificationHow many reference images one generation accepts. For a consistent character or product this number matters more than anything else
Weight openness 10 defined scale, with a written reason per assignmentTaken from the licence the vendor published, not from our opinion: from API only up to open weights under a free licence with a documented retraining path
Access from Iran 10 our access column, with its method statedToday this column separates nobody: all three candidates with a document are in the same situation. The body explains why

Why GPT Image 2 came first, with an empty cell

Fifty of the hundred points in this table belong to the two Arena boards, and this model wins both. On text-to-image it scores 1381.1 while the runner up sits at 1263.8, and on editing 1462.7 against 1392.6. Those are wide gaps and neither confidence interval comes near the next model.

The interesting part is that it won with an empty column. OpenAI does not publish a ceiling for reference images anywhere in its documentation; the sample code passes four images, but a sample is not a limit, so we left that cell empty and the model gave up ten points. Even ten points down, it finishes eight ahead of second place.

One condition is written into the leaderboard row itself and gets ignored almost everywhere: the row that came first is "gpt-image-2 (medium)", the medium quality tier. High quality at the same size costs $0.211, four times as much, and has no leaderboard row at all. The number in our price column is the $0.053 of the configuration that actually earned the rank.

The cheapest route to a good result

Nano Banana 2. Along with Nano Banana Pro it is one of only two candidates that fills all six columns, and it is the only one that combines full coverage with a high score. It wins the reference column outright with ten object images, and it undercuts the first-place model on price.

Then there is the finding that should change a purchase: Nano Banana Pro, Google expensive flagship, lands third here, below its own cheap model. On text-to-image it is genuinely behind and the confidence intervals confirm it, on editing the two are level, and it costs twice as much. If you are choosing between the two Google models, the cheaper one is almost always the answer.

The bottom row is the only one you can download

FLUX.2 klein 4B finishes sixth, and that should not be read as bad. It wins the price column outright at $0.014, roughly a quarter of what first place costs, and it is the only candidate that scores at all on weight openness. The other five all score zero there, because none of them has published weights.

So if your requirement is that the model runs on your own hardware, this table does not offer six options. It offers one. It runs on a graphics card with about 13GB of memory and its licence is Apache 2.0, so commercial work is allowed. Its quality sits well below the top and you have to accept that; in exchange it needs no account, no bank card, and no membership of anybody country list.

The Iran column separates nobody this month

This has to be said plainly because the table does not show it. Three candidates have data in this column: GPT Image 2 and both Google models. All three are in exactly the same situation, which is that Iran is absent from their supported-country lists and there is no official payment route. When every value in a column is equal, the engine awards all of them full marks. So this column creates no difference at all among the top three.

What it actually does is dock ten points from the other three, not because their situation is worse but because their vendors published no document to read. For FLUX we opened the full documentation index today: it contains no terms page and no country list, and the terms URL on the main site returns 404. ByteDance has published nothing comparable for Seedream. So in practice this column is measuring whether a vendor publishes a country list, which is not the same question as whether the thing works from Iran. Until more documents exist, that is what it is, and we are not hiding it.

Two criteria we deliberately did not build

First, resolution ceiling. Three vendors write three different things and the three do not add up. Google writes "up to 4K" for its models and nowhere prints pixel dimensions for any tier from 0.5K to 4K. OpenAI is the most precise: 3840 pixels on the long edge and a total between 655,360 and 8,294,400. Black Forest Labs says 4 megapixels. A label with no number, an exact number, and a third unit. You cannot build a column out of that, so we kept the values as facts on each model page and built no column.

Second, text inside the image. For Persian speakers this may be the single most important question, and it is exactly the one nobody publishes a number for. Google calls its text rendering advanced. OpenAI writes that its model can still struggle with text. Black Forest Labs calls flex specialised for typography. Those are three adjectives and not one measurement. A column built out of adjectives is not a column. The key exists in our data structure and stays empty until somebody prints a number.

Why Midjourney is not in the table at all

Because we have read nothing from it. midjourney.com returns 403 to our server: the home page, the updates page and its documentation alike. One difference from OpenAI is worth writing down: with openai.com at least robots.txt answers and does not forbid crawling, whereas midjourney.com/robots.txt is itself a 403, so even the rule we would be expected to follow cannot be read.

And unlike OpenAI, whose model holds first place on both boards, Midjourney appears on neither Arena image board: seventy seven models on text-to-image, fifty three on editing, and its name is on neither list. So we have no facts and no evidence, and we will not manufacture a row with no numbers in it. If you read somewhere that Midjourney is the best, ask where the number is.

One row that ranked on half a table

Seedream 5.0 Pro finishes fourth having filled only two of the six columns. It sits high on both boards, fifth on editing and eighth on text-to-image, and that alone was enough to edge past FLUX.2 max by four tenths of a point, while FLUX filled four columns. Read that carefully: Seedream position means it is good at the two things we measured, not that it is better than FLUX.

Those four empty columns are not our forgetfulness either. The Seedream 5.0 Pro page opens and is full of samples, but carries no model ID, no price, no reference ceiling and no date. The Seedream 5.0 page returns 200 with an empty body. This row coverage is exactly 50 out of 100, right on the boundary: one more empty cell and it would have dropped out of the table into the "not enough data" list beneath it.

When not to take our first pick

  • If you need Persian text inside the image, this table does not answer you: no vendor publishes a number on text quality and we have not tested these six on Persian ourselves.
  • If you want a monthly subscription with a graphical interface, this is the wrong page; this table measures models and its prices are per-image API prices.
  • If it has to run locally and offline, only the bottom row is any use to you and the rest are not options at all.
  • If your image needs several consistent characters, ignore the order of this table and go straight to the reference column.

Where it falls short

  • There is no resolution column, because three vendors publish three incompatible units.
  • There is no column for text quality inside the image, because nobody publishes a number.
  • The Iran column makes no difference at all among the three candidates that have a document.
  • Prices are tied to a specific size, and for FLUX they are per megapixel, so the table figure is a floor rather than the price of a large image.
  • Seedream 5.0 Pro ranks on 50 percent coverage and its position has to be read with that caveat.
  • Arena scores are tied to serving configurations: the first row is the medium quality tier and the Nano Banana 2 row carries a web-search suffix.
  • Midjourney is absent because no page of it opens for our server and it has no row on either board.
  • We have not run these six side by side on one prompt ourselves; every number here is published.

Our take

If you must pick one and budget is not the binding constraint, GPT Image 2 is the clearest choice in this table and it wins by a distance. If cost matters, take Nano Banana 2 and you will probably never notice the difference in daily work. And if your requirement is that the model belongs to you and goes nowhere, FLUX.2 klein 4B is the only option and you accept lower quality for it. Three different answers for three different constraints, and none of them is a best without conditions.

Questions people actually ask

What is the best AI image generator

GPT Image 2, because it is the only model that wins both Arena leaderboards. If price matters, Nano Banana 2 gets close for less money.

What is the cheapest AI image generator

FLUX.2 klein 4B, starting at $0.014 per image. Run it locally and it costs the vendor nothing at all, because its weights are published under Apache 2.0.

Which AI writes Persian text inside an image

We do not know, and no vendor publishes a number on it. Three of them describe their text rendering with adjectives like advanced, but none of them says anything about Persian. We have not tested it either, and until we do we will not write it.

Which image model can I run on my own computer

Only FLUX.2 klein 4B, whose weights are published under Apache 2.0 and which runs on a graphics card with about 13GB of memory. The other five publish no weights at all.

Can these be used from Iran

For GPT Image 2 and both Google models, Iran is not on the supported-country list and there is no official route. For FLUX and Seedream no such document has been published at all. We do not sell accounts and we do not suggest ways around it.

Why is Midjourney not in this table

Because its site returns 403 to our server, even its robots.txt, and it has no row on either Arena image board. We have no facts and no evidence, so we will not build a row with no numbers in it.

If the result has to be a publishable image rather than a raw output, that is the work we do in our graphic design service