Large Language ModelsGenerate imagesGenerate videos

Gemini API Free Tier Limits: Models, Rate Limits and Image Generation

The Gemini API free tier still works for text models, but image generation and Veo video now require billing. This article breaks down per-project rate limits, which models stay free, what each paid tier requires, and where to make AI images without a Google Cloud account.

Gemini API Free Tier Limits: Models, Rate Limits and Image Generation
Cristian Da Conceicao
Founder of Picasso IA

The Gemini API free tier still exists, but it no longer looks like the one people described a year ago. Text models can be called without a credit card, while every model that draws pictures or renders video now sits behind billing, and the daily quotas on the remaining models have shifted more than once. If you are budgeting a side project, a prototype or a classroom demo around these limits, the details decide whether your app works on launch day or throws errors by lunchtime.

This article lays out what the free tier includes, which models you can call without billing, how rate limits behave in a real project, and why image generation is the sticking point. You will also see what the paid tiers require and where to make AI images without a Google Cloud account.

💡 Quick answer: Text models such as Gemini 2.5 Flash have free access with tight per-project quotas. Image models like Nano Banana and Imagen 4, plus Veo 3.1 video, need a billed project.

What the Free Tier Includes

Every new Gemini API project starts on the free tier. The only requirement is an active project, so there is no billing account to link and no card to enter. That makes it the fastest way to test a prompt, wire up a prototype, or check how a model handles your own data.

Low-angle view down a data center aisle lined with black server racks and steady green status lights

Limits Apply Per Project

Google measures usage per project, not per API credential. Creating five credentials inside one project does not give you five quotas, because they all draw from the same pool. Splitting one app across several projects to stack allowances is fragile, and it is worth avoiding. Plan around a single project's allowance and design your app to live inside it.

RPM, TPM and RPD Explained

Four counters decide whether a request goes through:

  • RPM counts requests per minute.
  • TPM counts input tokens per minute.
  • RPD counts requests per day and resets at midnight Pacific time.
  • IPM counts images per minute and applies to image models.

Exceeding any single counter triggers a rate limit error, even when the other three look healthy. A chatbot that sends short messages usually trips RPM first. A document summarizer that sends long files trips TPM. A hobby app that runs all afternoon trips RPD.

Two hands typing on a laptop beside a notebook full of pencil tally marks and a hand-drawn grid

LimitWhat it countsWho hits it first
RPMRequests in a 60 second windowChat apps, bursty scripts
TPMInput tokens in a 60 second windowLong documents, big contexts
RPDRequests since the last resetBatch jobs, all day tools
IPMImages per minuteImage generation loops

Why the Numbers Keep Moving

Google's documentation no longer prints a per-model free tier table. The exact RPM, TPM and RPD for your project appear in the rate limit view inside Google AI Studio, and they can differ from project to project. In December 2025, developers reported steep cuts, with Gemini 2.5 Flash falling to roughly 20 requests per day in some reports. Third-party tables from early 2026 listed ranges of 5 to 15 RPM, around 250,000 TPM and 100 to 1,000 RPD. Treat those as reported figures, not promises.

💡 Tip: Read limits from your own AI Studio dashboard and from the error messages your app receives. Hardcoding a number from a blog post, including this one, is how apps break after a quiet quota change.

Which Models Work Without Billing

Not every model shares the same free status. At the time of writing, the pricing page draws a clean line between text models and everything that produces media.

Text Models You Can Call Free

The Flash family is the free tier's home. According to the pricing page, the Gemini 3.x Flash models and Gemini 2.5 Flash list input and output as "free of charge" in the Standard tier. The lighter Flash-Lite variants go further and stay free across Standard, Batch and Flex. Flash-class text-to-speech models are also free in the Standard and Priority tiers.

Gemini 3.5 Flash is the model most newcomers should start with, since it handles chat, code and image inputs at a speed that suits quotas this small. Batch and Flex do not come free on the larger Flash models, so a bulk job on those tiers means billing.

Models Locked Behind Billing

Everything that outputs pictures, music or video needs a paid project. Pro-class text models are the gray area, so confirm them in AI Studio before building on them.

Model familyFree tierNotes
Gemini Flash and Flash-Lite textYesTight per-project quotas
Gemini 3.1 Pro PreviewNot in third-party checksVerify in AI Studio
Nano Banana 2 (Gemini 3.1 Flash Image)NoBilling required
Nano Banana Pro (Gemini 3 Pro Image)NoBilling required
Imagen 4 familyNoBilled accounts only
Veo 3.1NoNot available on free
Lyria musicNoBilling required

Three colleagues around a light oak table, one pointing at a laptop while the others take notes

A team choosing a model should ask one question first: does the feature need media output? If yes, the free tier is already off the table, and the real decision is whether to pay Google directly or use another route.

Rate Limits in Real Projects

A limit on paper and a limit in production feel very different. On paper, 15 requests per minute sounds fine. In production, a page that fires four calls on every load will hit that ceiling with a handful of visitors.

A quick sizing exercise helps. Divide your daily request limit by the number of calls one user action makes. Suppose your project allows 100 requests per day and each action makes three calls: you can serve about 33 actions before the reset, which is a few dozen visitors at most. Run that math with the numbers shown in your own AI Studio dashboard before you promise anything to users.

Reading a 429 Error

When a quota trips, the API answers with HTTP 429. The message names the limit that failed, so read it before guessing.

What you seeLikely causeFix
Errors in short bursts, then recoveryRPMQueue requests and space them out
Errors only on long promptsTPMTrim context or split documents
Fine all morning, failing by eveningRPDCache results and cut call volume
Image model returns a zero quotaNo free tierLink billing or use another platform

Retry Logic That Works

Exponential backoff with a little randomness handles the per-minute limits well. This pattern fits in a dozen lines:

import random, time

def call_with_backoff(send, max_tries=5):
    for attempt in range(max_tries):
        try:
            return send()
        except RateLimitError:
            wait = (2 ** attempt) + random.random()
            time.sleep(wait)
    raise RuntimeError("Rate limited after all retries")

Backoff solves RPM. It cannot solve RPD, because that quota returns only at midnight Pacific time, no matter how politely you retry.

A round vintage wall clock with brass hands showing midnight in a dim room lit by a warm desk lamp

Four habits stretch a small free quota further:

  • Cache repeated answers so identical questions cost one request.
  • Trim prompts, since fewer input tokens protect your TPM.
  • Batch related questions into a single request when the task allows it.
  • Route simple jobs to a Flash-Lite model and save the bigger model for hard ones.

Image Generation Needs Billing

This is where the confusion starts, because older advice still fills the search results.

A photographer in a daylight studio studying a calibrated monitor showing a blurred mountain landscape

Nano Banana and Imagen Status

Early in 2026, several posts reported that Gemini 2.5 Flash Image allowed up to 500 free images per day at 1024 by 1024. That claim is outdated. The current pricing page lists no free tier for any image generation model:

  • Nano Banana 2, called Gemini 3.1 Flash Image in the API, has no free tier.
  • Nano Banana Pro, called Gemini 3 Pro Image in the API, has no free tier.
  • The Imagen 4 family works only with a linked billing account.

Then there is the ghost quota. In AI Studio, a free project can show a counter such as "0 / 25" next to Imagen, which looks like a daily allowance. A Google developer forum thread from February 2026 reports that requests still fail with a message saying the Imagen API is only accessible to billed users. Enabling billing does not add a free daily allowance either. The only free usage comes from the $300 introductory credit given to new Google Cloud users, valid for 90 days. After it expires, pay-as-you-go rates apply immediately.

A hand holding a plain dark metal card above a laptop keyboard on a tidy desk

💡 Watch out: A quota number on screen is not a promise of access. If a model sits in the "No" column of the pricing page, no amount of waiting will make it work on the free tier.

Veo and Video Limits

Video follows the same rule. Veo 3.1 is listed as "not available" on the free tier across its standard pricing variants, so a demo that renders clips needs a paid project from day one. If you want video clips without any API setup, PicassoIA hosts Veo 3.1, Veo 3.1 Fast and Veo 3.1 Lite for direct use.

Moving Up the Tiers

Paid tiers raise the ceiling and open up the media models. Moving up happens automatically once you meet the conditions below.

A whiteboard with a black marker sketch of three ascending steps connected by arrows

What Each Tier Requires

TierHow you qualifySpend cap
FreeActive project or free trialNot applicable
Tier 1Billing account linked$250
Tier 2$100+ spent and 3+ days since first payment$2,000
Tier 3$1,000+ spent and 30+ days since first payment$20,000 to $100,000+

Each tier also has a spend-based limit measured over a rolling 10 minute window: $10 for Tier 1, $50 for Tier 2 and $200 for Tier 3. A move from Free to Tier 1 takes effect instantly after billing setup. Later upgrades arrive within about 10 minutes of meeting the criteria.

Stay on the free tier if:

  • You are prototyping with text only.
  • Your traffic is a few hundred requests per day or less.
  • A failed call has no cost to your users.

Move to Tier 1 if:

  • You need Nano Banana, Imagen or Veo output.
  • Daily limits break your app before the day ends.
  • Real customers depend on the response.

Free Image Generation Without Billing

If your goal is pictures rather than text, there is a shorter path than a Google Cloud project. PicassoIA hosts Google's image and video models in the browser, so there is no Google Cloud billing account, no project setup and no quota dashboard to watch.

Gemini Models on PicassoIA

A young man in a sunlit café using a tablet on a stand beside a latte with foam art

The Google lineup that needs billing through the API is available directly on the platform:

NeedModel
Fast, flexible image creation and editingNano Banana 2
High resolution detailNano Banana Pro
Lighter, quicker draftsNano Banana 2 Lite
Photorealistic stillsImagen 4 and Imagen 4 Ultra
Speed over polishImagen 4 Fast
Chat, code and writingGemini 3.5 Flash and Gemini 3 Flash

Because these are separate models, you can send the same prompt to two of them and compare the results. Imagen 4 suits photorealistic stills, while Nano Banana 2 is the better pick when you want to edit or fuse existing images. Picking by task rather than by name keeps your results consistent.

How to Use Nano Banana 2

Nano Banana 2 is the model to try first, because it creates images from text and also edits or fuses existing ones.

  1. Open the Nano Banana 2 page on PicassoIA.
  2. Write a prompt that names the subject, the setting, the lighting and the lens.
  3. Add a reference image if you want to edit a photo or blend two images together.
  4. Pick the aspect ratio that fits your layout, such as 16:9 for a blog header.
  5. Generate, review the result, and adjust one detail of the prompt at a time.

Specific prompts win. A vague request like "a desk" gives a generic result, while "a developer's oak desk at golden hour, steam rising from a mug, 35mm lens, shallow depth of field" gives you something you can publish.

A top-down view of a creative workspace with a laptop, tablet, color swatches and a cup of tea

💡 Tip: Change only one thing per attempt. If the lighting is wrong, edit the lighting words and leave the rest alone. You will see exactly what each phrase does.

Quick Answers

Is the Gemini API Free?

Partly. Text models in the Flash and Flash-Lite families have a free tier with per-project limits. Pro-class text models, every image model, and Veo video need a billed project.

Can the Free Tier Make Images?

No. The current pricing page lists no free tier for Nano Banana, Gemini image models or Imagen. Posts claiming hundreds of free daily images describe an earlier setup. For image creation without a Google Cloud project, use a platform like PicassoIA with Nano Banana 2 or Imagen 4.

Do Limits Reset Daily?

Requests per day reset at midnight Pacific time. Per-minute limits recover as the rolling window passes, which is why backoff fixes bursts but not a spent daily quota.

Make Your Own Images Today

The free tier is a fine place to test text prompts, but the moment you need a picture, a billing account becomes the price of entry on Google's side. You can skip that step. Open Nano Banana 2 on PicassoIA, type the scene you have in mind, and see the result in seconds. Try Imagen 4 when you want a photorealistic still, or Veo 3.1 Fast when a still is not enough. Run the same prompt through two models, compare the results side by side, and keep the one that fits your project. Every model is listed at picassoia.com/en/all-models, so you can compare your options before you commit to one.

Share this article