Video Generation API: Free Options, Pricing and Cost
A video generation API bills by the second, and the gap between $0.05 and $0.60 per second adds up fast. This article lays out which options are free, what per-second tiers really cost, how retries inflate the bill, and a simple formula to budget your first month of clips.
A video generation API does not bill you by the month. It bills you by the second of footage, and that one detail decides whether your project costs a lunch or a laptop. Google lists Veo 3.1 Lite at $0.05 per second at 720p and Veo 3.1 at $0.40 per second, an eightfold gap for the same length of clip. This article shows which options are genuinely free, what the paid ones cost, and how to work out your own number before you write a single line of code.
💡 Prices move. The figures below come from Google's Gemini API price table and PicassoIA's API documentation, both checked in October 2026. Providers change rates often, so confirm the live page before you lock a budget.
What a Video API Actually Does
A video API is a web endpoint that accepts a text prompt, and sometimes a starting image, and returns a finished clip. The catch is time. Rendering a few seconds of footage takes far longer than a normal API call, so most providers use an asynchronous pattern instead of making you hold a connection open.
The Request, Poll, Fetch Loop
The loop has three steps:
Create: you send a POST with your prompt and settings. The provider answers at once with a prediction ID and a status of starting.
Poll: you check that ID on a timer until the status becomes succeeded, failed or canceled.
Fetch: you download the MP4 from the returned URL and store it somewhere you control.
PicassoIA's API follows this Replicate-style pattern. The example renders on the PicassoIA Video model page finished in roughly 30 to 80 seconds, so plan your polling interval around that range instead of hammering the endpoint.
What You Are Really Paying For
Four things move the final bill:
Seconds of output: the base unit for nearly every provider, so a 10 second clip costs twice a 5 second clip.
Resolution: 720p, 1080p and 4K usually sit at different rates.
Model tier: lite, fast and flagship versions of one family can differ by a factor of eight.
Retries: every take you throw away is a take you still paid for.
Free Options That Really Exist
Free is a loaded word in this market. Some options are free to use, some are free to download, and some are free only for the first afternoon. Here is the short version before the detail:
Google's Gemini API price table marks Veo 3.1 as not available on the free tier, so every second is billed from the first request. Other API providers sometimes grant a small credit at signup. Treat that credit as a trial: enough to test prompts and measure latency, not enough to run a product. Consumer apps that hand out daily credits usually do not extend them to API calls.
Open Weights You Host Yourself
Open models such as LTX Video and Wan 2.1 1.3b carry no per-second fee, but they are not free. You pay in GPU hours, setup time and maintenance. A simple comparison is the hourly GPU price divided by the clips rendered per hour. With an illustrative $1.00 per hour GPU that renders 12 clips an hour, each clip costs about $0.08 before idle time. That wins against a hosted API only while the GPU stays busy.
PicassoIA API: Free Predictions
PicassoIA documents its developer API at https://api.picassoia.com/v1 and states: "API predictions are currently free. They use no credits." Four models are exposed: PicassoIA Video, Seedance 2.5 Lite with synchronized audio, and two image models for generation and editing. Authentication is a Bearer token that starts with pia_sk_, and an account can hold up to two.
The limits matter as much as the price:
5 predictions queued or running at once per account, shared across all tokens and MCP connections
10 MB request body
5 MB per image sent as a data URL
4,000 characters per prompt
💡 Read the plan line. The documentation says accounts on the Infinite plan can use the API, and other accounts receive a 403 plan_required error. The pricing page, however, lists API access on Pro+, Elite and Infinite. Those two statements disagree today, so check which plan your account needs before you build on it. Free per prediction is not the same as free of any plan.
Real Per-Second Prices Compared
Google's table is the cleanest public example of how tiers scale, because one model family appears at three price points. These are per-second rates from the Gemini API pricing page, with clip costs worked out for comparison. Scale them to the length you actually render.
Inside one family the spread is large. Lite against Standard is 8x at 720p and 5x at 1080p. Other families publish their own rates, among them Seedance 2.5, Kling v3 Video, Wan 2.7 T2V and LTX 2.3 Fast. Rates differ by provider and change often, so compare them with the formula later in this article rather than trusting a screenshot.
A cheaper tier is not automatically a worse choice. Drafts, storyboards and social clips rarely need the flagship, which you can save for the final render.
Vendor Risk Is a Cost Too
OpenAI announced that the API for Sora 2 would shut down on September 24, 2026, and anyone who built a product on it had to move. A low rate means little if the endpoint disappears. Keep prompts and settings provider-neutral, and favor a platform that puts several models behind one interface.
A Clip Versus a Shoot Day
A 10 second clip at the lite tier costs $0.50. Even ten takes cost $5.00, a rounding error next to a camera rental or a crew day. The comparison stops being fair when you need a specific real person, product or location, because generation approximates and a camera captures. Use the API for concept work, b-roll, backgrounds and short social clips, and film what must be exact.
Costs That Hide Behind the Rate
The sticker price is the cheapest number you will ever see. Three other costs sit behind it.
Retries Eat the Budget
Video models do not hand you a perfect clip on the first try, so the sticker price understates the real one. Divide the price per clip by the share of takes you can actually use. At $0.50 per clip and a hit rate of one in four, each usable clip costs $2.00. Once a take works, lock its seed so you can reproduce it, since models such as PicassoIA Video accept a seed for repeatable output.
💡 Draft cheap, finish once. Test prompts on the lowest tier and resolution, then spend the higher rate only on the prompts that already work.
Resolution Multiplies Fast
Google's own table shows the effect: Veo 3.1 Fast costs $0.10 per second at 720p and $0.30 per second at 4K, triple the price for the same prompt. On PicassoIA Video you can pick 480p for the fastest render or 720p for sharper output. Draft at the low setting and keep the high setting for the final pass.
Storage and Engineering Time
Every finished MP4 needs storage and bandwidth, and every integration needs someone to write the polling loop, handle failures and watch for rate limits. For a small project that is a few evenings. For a team, it is a line item. Download results promptly and copy them to your own storage instead of relying on a provider URL staying alive.
Build a Budget in Five Minutes
You do not need a spreadsheet to price a project. You need one formula and an honest guess about retries.
The Formula
Monthly cost = usable clips × seconds per clip × price per second × takes per usable clip
Take 200 usable clips of 8 seconds each. That is 1,600 seconds of finished footage. At $0.10 per second with every take usable, the bill is $160. If you need three takes per usable clip, it is $480.
Three Sample Monthly Budgets
Same workload, three tiers, using the Veo rates listed earlier:
The gap between the cheapest and the most expensive row is 8x before retries and 8x after them, which is why the tier you pick matters more than any other lever. Whichever tier you choose, run a pilot of 20 clips, measure your real takes per usable clip, and replace my 3 with your own number before you commit.
How to Use PicassoIA Video
PicassoIA ships its own text and image to video model, PicassoIA Video, so it is a practical place to test the whole loop. The model page lists a fixed 5 second clip at 24 frames per second, rendered with synchronized audio.
Write the prompt in order. Describe the subject and starting pose, then the motion, then the camera move, then the light. The model schema says cinematic, chronological descriptions work best.
Add an input image if you want a specific first frame. The clip uses it as the opening shot and inherits its aspect ratio.
Choose a resolution. 480p renders fastest, 720p is sharper.
💡 Check the live schema. The model page and the API reference do not list identical resolution and duration ranges. Read the current schema for the model before you hard-code any value.
Call It From Code
The API wraps every setting in an input object. Create the prediction, then poll its ID:
curl -s -X POST https://api.picassoia.com/v1/models/picassoia/picassoia-video/predictions \
-H "Authorization: Bearer $PICASSOIA_TOKEN" \
-H "Content-Type: application/json" \
-d '{"input": {"prompt": "A fishing boat rocks on calm water at sunrise, mist lifts off the surface, slow dolly-in, warm low light", "resolution": "480p", "aspect_ratio": "16:9"}}'
curl -s https://api.picassoia.com/v1/predictions/PREDICTION_ID \
-H "Authorization: Bearer $PICASSOIA_TOKEN"
The first response returns 201 Created with an id, a starting status and eta.next_poll_in_seconds. When the status reads succeeded, the output field holds the video URL.
Limits That Cap Your Throughput
Money is not the only ceiling. PicassoIA allows 5 predictions queued or running at the same time per account, and that pool is shared across every token and MCP connection you own. If a render takes about a minute, five slots give you at most 300 clips per hour, no matter how large your budget is.
Plan around the limit:
Queue on your side: run a worker pool of five or fewer and feed it from a job list.
Respect the poll hint: wait the eta.next_poll_in_seconds the first response gives you between checks.
Resubmit failures: a failed status ends that prediction, so send a new one instead of polling again.
Keep payloads small: stay under 10 MB per request and 5 MB per image data URL.
Try It Yourself on PicassoIA
The cheapest way to find your own number is to make a few clips. Open PicassoIA Video and render a 480p draft of a shot you actually need. Then run the same idea through Seedance 2.5 Lite for synchronized audio and compare the two side by side.
When a draft works, count your takes, drop the numbers into the formula, and decide whether the free API route, a hosted per-second tier or your own GPU fits your volume. Picasso IA puts text to image, text to video, image editing and lipsync under one roof, so you can test all of it before you commit to a single provider.
Create your own images and videos with Picasso IA today. Start with one draft, measure what it costs, and scale only after the numbers hold.