
Create images with AI Studio
Choose an AI image model, add references, control count, style, and ratio, then review image results on your canvas without wasting credits.
An image request usually starts with a deceptively small sentence: “Make this feel more premium.” The useful part is everything around it. Which product must stay recognizable? Is the result a rough direction or the final campaign visual? Does it need readable text? Is speed more important than preserving the reference exactly?
The short version
Choose Image, leave Model, Style, and Ratio on Auto, and keep Count at 1 for the first draft. Add references only when the result must preserve something specific, then describe the subject, composition, audience, and details that must not change. Review the first result before increasing quality or generating more versions.
Choose Image in AI Studio when the result should be a visual object you can place, compare, comment on, and share from the canvas. AI Studio can generate a fresh image from text or use selected and attached images as references, but the model you choose changes what it can preserve, how long it usually takes, and the relative cost.
Use another output when the job is not primarily visual:
| What you need | Better choice |
|---|---|
| A brief, proposal, summary, or long explanation | Doc |
| Structured facts that need rows and columns | Table |
| A browsable one-pager, product page, or interactive preview | Web page |
| A spoken draft or podcast-style result | Audio file |
| AI Studio to decide from the prompt and references | Auto |
For the broader creation flow, read Create content with AI Studio. This article stays focused on Image.
Start with the least expensive useful draft
- Open the canvas where the image belongs.
- Select Image in AI Studio.
- Leave Model, Style, and Ratio on Auto for the first attempt.
- Keep Count at 1 until the prompt and references produce the right direction.
- Add selected canvas objects or image attachments only when the result must preserve something specific.
- Describe the subject, composition, audience, visual treatment, and details that must not change.
- Submit, then review the image on the canvas before generating more versions.
This workflow keeps the first decision cheap. A weak prompt multiplied by Count 10 is still a weak prompt, only ten times.
The picker puts the important tradeoffs in one place:
Choose a model by the job, not the brand
Auto is the right default for most requests. It considers the prompt, references, expected time, and cost, then chooses a compatible route. If a selected canvas image or attached image is present, text-only models are unavailable and Auto stays on a reference-capable path.
Choose a specific model when you know which tradeoff matters. The Cost value below is the relative value shown in the picker.
| Model | Image references | Best use | Speed hint | Cost | Important limitation |
|---|---|---|---|---|---|
| Auto | Yes | Most first attempts; lets ALLO balance prompt, references, time, and cost. | Tailored pick | Auto | The chosen model can change with the request. |
| Nano Banana 🍌 | Reference/edit | Rough drafts that should preserve a small reference set. | ~15s | x1 | Lower detail; supports up to 3 references and tops out around 1024px. |
| Nano Banana 2 🍌 | Reference/edit | Polished social visuals, simple diagrams, multilingual text, and ordinary one- or two-reference work. | ~30s | x1.7 | More expensive than Nano Banana; use Pro for the hardest dense layouts. |
| Nano Banana Pro 🍌 | Reference/edit | Presentation-ready key visuals, dense infographics, complex storyboards, multilingual layouts, and many-reference continuity. | 40s+ | x3.4 | The most expensive Nano Banana choice; save it for detail that survives review. |
| GPT Image 2 · Low | Reference/edit | Low-detail GPT drafts and narrow exact-edit iterations. | 60s-2min | x0.2 | Lower detail than the default tier. |
| GPT Image 2 | Reference/edit | Precise edits, text and layout work, and changing one element while keeping the rest stable. | 60s-2min | x0.8 | Slower than the fast Google and Flux routes. |
| GPT Image 2 · High | Reference/edit | Final-quality GPT output when the highest tier is deliberate. | 60s-2min | x3.2 | Expensive; do not use it to discover the direction. |
| GPT Image mini · Low | Reference/edit | The cheapest OpenAI draft or quick edit. | 60s-2min | x0.1 | Lowest-detail GPT Image mini tier and no upscale support. |
| GPT Image mini | Reference/edit | Cleaner low-cost edits and layout drafts. | 60s-2min | x0.2 | Smaller output ceiling than GPT Image 2 and no upscale support. |
| GPT Image mini · High | Reference/edit | The most polished mini-model result when you still want the mini family. | 60s-2min | x0.9 | Costs close to stronger general choices while keeping the mini family’s output limits. |
| Flux 2 Pro | Text-only | Fast text-to-image generation for a fresh visual with no image to preserve. | ~15s | x0.75 | Selected or attached images are not supported; AI Studio currently keeps output within standard ratios and a bounded 1MP request. |
| Imagen 4 Fast | Text-only | The fastest simple text-to-image draft. | ~5s | x0.5 | No image editing or reference preservation; lower detail and 1024px output. |
| Imagen 4 | Text-only | Balanced photorealistic or illustrative generation from a prompt. | ~10s | x1 | No image editing or reference preservation. |
| Imagen 4 Ultra | Text-only | The highest Imagen 4 quality for a fresh text-only image. | ~20s | x1.5 | No image editing or reference preservation; use it after the direction is settled. |
The labels matter:
- Reference/edit models accept selected or attached images and can preserve or change them.
- Text-only models create from the written prompt. Flux 2 Pro and every Imagen 4 option become unavailable when an image reference is present.
- A text-only model is not a worse model. It is the wrong model only when the request depends on an existing image.
The Nano Banana 2 and Pro catalog can handle larger reference sets and higher-resolution work than the base Nano Banana route. GPT Image 2 is the stronger GPT choice for precise edits and larger output. GPT Image mini is for cheaper drafts, not for upscaling. Imagen 4 is strongest when you want a new photorealistic or illustrative image and do not need to preserve a source.
The picker’s x value is a relative cost estimate normalized to Nano Banana 🍌 at x1. It is not an exact AI credit quote. Actual workspace credits depend on provider usage, input and reference tokens, generated output, and Count. The accounting service converts that usage and rounds it to whole credits in the workspace ledger.
Set Count, Style, and Ratio deliberately
Count
Count accepts 1-10 images. One is the default because every additional result starts another image job and can consume more credits.
- Use 1 while testing the prompt, references, and model.
- Increase Count when the direction is already right and you need variants.
- Use a smaller batch with an expensive model, then keep the best result instead of repeating a large batch.
Each result gets its own placeholder. In a multi-image request, deleting one processing placeholder cancels that output slot. Deleting every image placeholder cancels the remaining task.
Style
| Style | Use it when |
|---|---|
| Auto | The prompt already describes the look, or you want ALLO to interpret the references. |
| Comic | The result should read as illustrated panels, characters, or a graphic narrative. |
| Infographic | The result needs visual hierarchy, labeled sections, or an explanatory composition. |
Style is direction, not a substitute for a concrete prompt. “Infographic” still needs the message, audience, hierarchy, and facts that belong in the visual.
Ratio
| Ratio | Good starting point |
|---|---|
| Auto | Let the prompt and references determine the shape. |
| 1:1 | Square concept tiles, profile visuals, and compact comparisons. |
| 4:3 | Visual briefs, diagrams, and conventional presentation content. |
| 3:4 | Portrait cards, posters, and document-like visuals. |
| 16:9 | Slides, banners, video covers, and wide presentation visuals. |
| 9:16 | Stories, Reels, Shorts, and other vertical social formats. |
All current picker models support these five explicit ratios. Auto chooses among them from the request. If a saved draft contains a choice the selected model cannot use, AI Studio normalizes the request to a supported route rather than sending an invalid combination.
Give the model references it can actually use
A good image prompt answers five questions:
- What is the subject?
- What should the composition show?
- Who will see it?
- What visual treatment should it use?
- What must stay unchanged?
For example:
Create a 16:9 launch visual for a client presentation. Keep the bottle shape and label from reference #1. Use the warm amber palette from reference #2. Place the product on the right with open space for a title on the left. Do not add text inside the image.
That prompt separates identity, style, layout, and constraints. “Make this premium” leaves all four for the model to guess.
Attachments are numbered before selected canvas objects, so references such as #1 and #2 stay clear. Use those numbers when several images play different roles. Keep only references that affect the result. A mood image, product image, old layout, and unrelated screenshot in one request create four competing instructions.
The safest selection rule is simple:
- Use Nano Banana 2 🍌 when references, multilingual text, or a moderately dense layout matter.
- Use Nano Banana Pro 🍌 when the reference set or layout is genuinely difficult and the result is headed for external review.
- Use GPT Image 2 for a precise edit or explicit text and layout instruction.
- Use GPT Image mini · Low or GPT Image mini to test an edit cheaply.
- Use Flux 2 Pro or Imagen 4 only when you want a fresh image from text.
- Use Auto when none of those needs dominates.
This walkthrough shows a reference image and canvas context becoming a finished image beside the source material:
Spend credits on decisions, not retries
Use this sequence for expensive work:
- Write the prompt with Count 1, Model Auto, Style Auto, and Ratio Auto.
- Generate a first composition.
- Fix the prompt before increasing quality. Add the missing subject, layout, reference number, or constraint.
- Use a low-cost tier to test a precise edit.
- Move to Nano Banana Pro, GPT Image 2 · High, or Imagen 4 Ultra only when the direction is stable.
- Increase Count only when you need useful alternatives, not reassurance.
Manual edits can be cheaper than asking the model to rediscover a nearly finished result. If one label, crop, or small canvas detail is wrong, fix it directly when possible.
Deleting a finished image does not refund credits. A successful result can consume credits even when you dislike it, remove it, or generate another version. For the workspace balance and ledger, read Understand AI credits.
Watch progress and cancel from the canvas
After submission, AI Studio places image placeholders on the canvas. The server updates each placeholder as its image finishes, so a batch can reveal usable results before every slot is done.
While a placeholder is processing, selection and deletion remain available. Editing, preview, download, duplicate, copy, lock, and thumbnail actions remain unavailable until the result finishes.
To cancel:
- Delete one processing image placeholder to cancel that slot.
- Delete every image placeholder from the request to cancel the whole image task.
- Deleting the canvas also stops generation still running there.
Do not resubmit because a slower model still shows progress. GPT Image 2 and GPT Image mini show a 60s-2min picker hint, and the relative speed hints are expectations rather than guarantees.
When image generation does not behave as expected
| Problem | What to check |
|---|---|
| A model is unavailable | Remove selected or attached image references if you intended to use Flux 2 Pro or Imagen 4. Otherwise choose Auto or a reference/edit model. |
| The result ignores a reference | Name the reference number and say what must be preserved. Remove unrelated references. |
| Text is wrong or unreadable | Use a model suited to text and layout, simplify the amount of text, or add final text as an editable canvas object. |
| The composition is right but the format is wrong | Set Ratio explicitly before regenerating. |
| Only some images finish | Keep the completed results. A multi-image batch progresses by slot, and a failed slot does not require deleting the successful ones. |
| The placeholder stays in progress | Wait for the model’s normal speed range, then refresh the canvas and check whether the placeholder still exists. Do not submit the same costly request twice. |
| The request is blocked by credits | Ask a workspace admin to add credits in Billing, then review Count and model cost before retrying. |
| The result is generic | Add audience, composition, references, visual treatment, and constraints before changing to a more expensive model. |
If you contact support, send the canvas link, approximate request time, selected Model, Count, Style, Ratio, a prompt summary, whether image references were selected or attached, and whether credits were blocked or consumed. Do not send private customer images or source files unless support specifically needs them and you are allowed to share them.
Related guides
- Use AI Studio in a canvas
- Create content with AI Studio
- Review AI output and share web pages
- Understand AI credits
- When an AI result is not right