Images

AI Photo Generator

Three image models behind one prompt box: a 1-credit model for volume, a 14-reference model for products, and a model that gets text inside the frame right.

Make your first imageBrowse the 5 generators

An AI photo generator turns a written prompt into an image. MakeViral puts three models in one form: Z-Image at 1 credit, Nano Banana 2 at 2 credits, and GPT Image 2.5 at 2 credits, all at 1K. You get 1 to 4 images per run, 9:16 first, and you pick the model per job.

What does this AI photo generator actually do?

You type a prompt, pick a model, pick a size, and get 1 to 4 images back. There is no canvas, no layers and no brush. The form is the whole interface.

Three things are worth knowing before your first run:

  • You choose the model, per image. The models are not tiers of one product. They fail in different ways, and the cheap one is the right answer more often than people expect.
  • 9:16 is offered first. Everything else on this site renders vertical, so the image sizes start where the videos end up. Other aspect ratios are in the same chip row.
  • Reference images are a model feature, not a site feature. Nano Banana 2 takes up to 14 of them, GPT Image 2.5 takes one, and Z-Image takes none.

Open the photo form and the model chips sit above the prompt box. If you already know which one you want, go straight in: Z-Image, Nano Banana 2 or GPT Image 2.5.

The images come out as stills. This surface never returns a video. If motion is what you need, that is the AI influencer generator or the video tools.

Which model should you use?

Start from the job, not from the model name. The decision table first, the specifications second.

The jobModelCredits at 1KWhy this one
Background plate for a slideshow or an adZ-Image1Nobody looks at the plate. Pay the least.
Bulk concepts, 20 ideas in one sittingZ-Image1Cheap enough to throw most away
Your product, kept recognizableNano Banana 22Up to 14 reference images hold the shape and the color
A face you need again next weekNano Banana 22References keep the same person across runs
A packshot with a readable labelGPT Image 2.52It is the strongest of the three at text inside the frame
A thumbnail with two words on itGPT Image 2.52Same reason: legible words, correctly spelled
A busy scene with many objectsNano Banana 22Long prompts, up to 4,000 characters, get followed
ModelCredits 1K / 2K / 4KReference imagesPrompt ceilingTypical wait
Z-Image1, 1K onlyNone, text prompt only900 characters5 to 15 seconds
Nano Banana 22 / 3 / 4Up to 144,000 characters10 to 30 seconds
GPT Image 2.52 / 2 / 41No stated limit20 to 60 seconds

Nano Banana 2 is Google's image model and GPT Image 2.5 is OpenAI's image model. Z-Image runs here as a model choice with no vendor claim attached.

Two things fall straight out of that table. GPT Image 2.5 costs the same at 2K as at 1K, so if you picked it, take the 2K. And Z-Image has no 2K or 4K at all, so anything you will crop into or print needs one of the other two.

How do you write a prompt that returns a usable product shot?

A weak product prompt names the product. A strong one names the photograph.

Write in this order, and the hit rate climbs immediately:

  1. Subject. What the thing is, in plain words. "A matte black 500ml water bottle."
  2. Surface and setting. "On a wet concrete ledge, morning, city street behind."
  3. Light. The single most useful clause in the prompt. "Hard side light from the left, deep shadow on the right."
  4. Lens language. "Shot on a 50mm lens, shallow depth of field, shot from slightly below."
  5. Framing for the crop. "Centered, space above the bottle for a headline."
  6. What to exclude. "No hands, no text, no logo."

Two habits that fix most bad results:

Describe the light, not the mood. "Cinematic" and "premium" are not instructions. "One softbox at 45 degrees, dark background, no fill" is.

Leave room for the overlay. If the image is going under a slideshow headline or above a caption, say where the empty space goes. Recropping a full frame costs another credit.

Prompt ceilings matter here. Z-Image stops at 900 characters, which is about six of those clauses, so keep it tight or move the shot to Nano Banana 2 with its 4,000-character ceiling. For the packshot version of this workflow, the AI product photo generator opens on the same form with product defaults.

What does a reference image actually change?

A prompt describes a category. A reference image pins an instance. That is the whole difference, and it decides which model you pick.

What references are good for:

  • Your actual product. The real bottle, the real label, the real color. A prompt cannot recover a shape it has never seen.
  • One person across many images. Upload the same face and the creator stays the same creator. That is what the AI influencer generator uses to keep a persona consistent.
  • A style you already own. Feed back an image you liked and ask for a different angle of it.
  • A set. Six references of the same object from different sides gives the model the geometry, not a guess at it.

What references do not fix: a bad prompt. The reference tells the model what the object is, and the prompt still has to say what the photograph is.

The practical limits are the model limits. Nano Banana 2 takes up to 14 reference images, which is enough for an object and a person and a set in one run: start at the Nano Banana 2 image generator. GPT Image 2.5 takes exactly one, so choose it when the reference is the label rather than the object. Z-Image takes none, which is the honest reason it costs 1 credit.

Why is the cheap model the right default?

Most images on a content site are not the subject. They are what the subject sits on.

A slideshow background, a gradient behind a hook, a texture under a chat screen, a scene you will crop to 20% of frame: none of those reward 4 credits. They reward volume, because the tenth attempt is usually better than the first, and at 1 credit you can afford a tenth attempt.

Spend the extra credit when one of these is true:

  • A human face is the subject, and it has to survive being looked at.
  • A real product has to stay itself, which needs references.
  • Words have to be legible inside the image, which is GPT Image 2.5.
  • You need 2K or 4K, which Z-Image does not offer.

A working split for a week of content: plates and concepts on Z-Image, hero shots and people on Nano Banana 2, anything with type on GPT Image 2.5.

Run the arithmetic once and the habit sticks. Forty background plates on Z-Image is 40 credits. The same forty on the 4K tier of Nano Banana 2 is 160 credits, which is a third of the $49.99 plan, spent on images nobody will look at directly.

What are these images for on this site?

This is not a general art tool bolted on. Each of the four uses feeds another surface.

UseBest modelWhere it goes next
Ad creative and product shotsNano Banana 2 for the product, GPT Image 2.5 for the labelCut into a conversation ad or a demo clip
Slideshow backgroundsZ-Image at 1 creditBehind the hook slide in the slideshow generator
AI creator portraitsNano Banana 2 at 2 credits, Z-Image at 1The persona portrait in the AI influencer generator
Thumbnails and coversGPT Image 2.5YouTube, TikTok covers, anywhere words sit in the frame

The creator portrait is the one that pays for itself. Two credits buys a face you then reuse across a whole campaign, and the same face in every ad is what makes a set of ads look like one brand instead of five stock photos.

For user-generated-content style shots, where the product looks like somebody's kitchen counter rather than a studio, use the UGC photo generator. The prompt rules invert there: ask for phone camera light, slight clutter and an off-center crop.

What does an image cost?

Priced per image, by model and resolution. Nothing is priced per prompt or per attempt, so a run of four images costs four times one image.

Model1K2K4K
Z-Image1 creditNot offeredNot offered
Nano Banana 22 credits3 credits4 credits
GPT Image 2.52 credits2 credits4 credits

In dollars, on the $49.99 plan for 500 credits, a credit is about $0.10:

  • A Z-Image plate: about 10 cents.
  • A Nano Banana 2 hero at 1K: about 20 cents. At 4K, about 40 cents.
  • A run of four GPT Image 2.5 images at 2K: 8 credits, about 80 cents.

Plans, read from pricing on September 17, 2026: $9.99 for a one-time 100-credit pack, $49.99 per month for 500 credits, $199.99 per month for 3,000 credits. On the larger plan a credit is about $0.067, so the same four-image run is about 54 cents.

New accounts start with zero credits, so the first image needs a plan. Cancel anytime, 14-day refund window on unused credits.

What the photo generator does not do

Short list, and it matters more than the feature list.

  • No editor. No layers, no masking, no inpainting, no brush and no object removal. You change the prompt and run it again.
  • No video. This surface returns stills. Motion lives on the AI influencer generator and the video tools, and nothing here exports 16:9 or long form.
  • No scheduling, no auto-posting, no connected accounts. You download the file and post it yourself.
  • No ad account integration. Images do not flow into Meta or TikTok Ads Manager. You upload them there.
  • No API and no batch import. One run at a time in the browser, 1 to 4 images per run.
  • No analytics. Nothing here tells you which image performed. That data stays in your ad platform.
  • No guarantee of a rights-cleared output. Do not prompt for a real person, a brand mark you do not own, or a character somebody else owns.

If an editor is what you actually need, get the generation here and do the retouching in the tool you already have. The alternatives comparisons list what the neighboring tools cover.

What you get

What is included

Three models, one prompt box

Switch between Z-Image, Nano Banana 2 and GPT Image 2.5 without leaving the form or rewriting the prompt.

Images from 1 credit

Z-Image renders at 1K for a single credit, which makes background plates and throwaway concepts cheap enough to iterate on.

Up to 14 reference images

Nano Banana 2 accepts up to 14 references, so your real product and your recurring creator stay consistent across runs.

Legible text inside the image

GPT Image 2.5 is the pick when words have to appear in the frame and be spelled correctly, like a label or a thumbnail.

9:16 first, up to 4K

Vertical sizes lead the chip row because that is where this content ends up, with 2K and 4K available on two of the models.

How it works

From idea to finished file

  1. Pick the model

    Z-Image for volume and backgrounds, Nano Banana 2 for products and people, GPT Image 2.5 when text has to be readable in the frame.

  2. Write the shot, not the object

    Name the subject, the setting, the light, the lens and the crop. Add reference images when the object has to stay itself.

  3. Run 1 to 4 images

    Choose the aspect ratio and the resolution, generate, and download the ones you want. You pay per image, not per attempt.

Who it is for

Who uses it

  • Ecommerce and DTC
  • Shopify stores
  • App and SaaS marketing
  • Ad creative
  • Slideshow backgrounds
  • Creator portraits
  • Thumbnails
  • Affiliate content
FAQ

AI Photo Generator questions

What is the best AI photo generator model here?
There is no single best one, which is why all three are in the form. Z-Image at 1 credit is right for backgrounds and volume. Nano Banana 2 at 2 credits handles real products and recurring faces with up to 14 reference images. GPT Image 2.5 at 2 credits is the one to use when text must be readable inside the image.
How much does one AI image cost?
Z-Image is 1 credit at 1K, which is its only size. Nano Banana 2 is 2 credits at 1K, 3 at 2K and 4 at 4K. GPT Image 2.5 is 2 credits at 1K and 2K, and 4 at 4K. On the $49.99 plan for 500 credits, a credit is about 10 cents.
What is Nano Banana 2?
Nano Banana 2 is Google's image model, offered here as one of three model choices. On this surface it accepts up to 14 reference images and prompts up to 4,000 characters, and it usually returns in 10 to 30 seconds. It is the default pick for product shots and for keeping one person consistent.
What is Z-Image and why is it 1 credit?
Z-Image is the cheapest model in the form. It renders at 1K only, takes a text prompt with no reference images, caps the prompt at 900 characters and usually returns in 5 to 15 seconds. Fewer inputs and one resolution is exactly why it costs a single credit per image.
Can I upload my own product photo as a reference?
Yes, on two of the three models. Nano Banana 2 takes up to 14 reference images, which is the right choice for a product that has to stay recognizable. GPT Image 2.5 takes one. Z-Image takes none and works from the text prompt alone.
Which model handles text inside the image?
GPT Image 2.5. It is the strongest of the three at rendering legible, correctly spelled words inside the frame, which matters for labels, packaging, posters and thumbnails. It also costs the same at 2K as at 1K, so take the larger size when you pick it.
How many images can I generate at once?
One to four per run. You pay per image, so a four-image run costs four times a single image on that model and resolution. Aspect chips start at 9:16 because most of this content ends up vertical, and other ratios are in the same row.
Can I edit an image after it is generated?
No. There is no editor, no layers, no masking and no inpainting on this surface. You adjust the prompt or the reference images and run it again. Retouching happens in whatever image editor you already use, after you download the file.

Make your first image

Three image models behind one prompt box: a 1-credit model for volume, a 14-reference model for products, and a model that gets text inside the frame right.

Cancel anytime. 14-day refund window on unused credits.