Image to Video API: Free, Cheap and Pricing Compared
Image to video APIs charge anywhere from nothing to $2.00 for a single 5 second clip. This article lines up the free routes, the cheap per-second options and the hidden costs, with prices checked in October 2026 and a working PicassoIA API example.
A five second clip made from a single still can cost nothing, a quarter, or two dollars, depending on which image to video API you call and which tier you pick. That spread is the real story of pricing this month, and it is why a quick look at one provider's page rarely predicts what a project will cost. Below you get the free routes, the cheap routes, the expensive ones, the arithmetic for a 5 second clip, and the places where the price on the page is not the price on the invoice.
💡 Prices move monthly. Every number here was checked in October 2026 against a provider page or a named tracker. Use them as a snapshot, then confirm on the provider's own page before you commit volume.
What an Image to Video API Does
Still In, Clip Out
You send one image as the opening frame plus a short motion prompt. The API returns an MP4, usually 5 to 10 seconds long, at anywhere from 480p to 4K. Many models now generate synchronized audio in the same pass, and that changes the price (more on that below). The image fixes the subject, the colors and the framing, so the result is far more predictable than text to video, where the model invents everything from scratch.
The inputs look alike across providers: an image URL or upload, a prompt written as a chronological description ("the camera pushes in slowly while steam rises from the cup"), a resolution, a duration, and sometimes a seed for repeatable results.
Why Jobs Run Asynchronously
Rendering takes real time, so almost every video API works in two steps: create a job, then poll it until the status reads succeeded or failed. Example runs listed on PicassoIA's own model pages show the range: about 30 to 80 seconds for 5 second clips on Picasso IA Video, and roughly 100 to 190 seconds for 10 to 15 second clips on Seedance 2.5 Lite. Your code has to wait, retry and handle failure, and a job that fails is still your time and sometimes your money.
Free Image to Video APIs
"Free" comes in three flavors, and only one of them is free the way people hope.
Free Tiers That Exist
The PicassoIA API. The docs say "API predictions are currently free. They use no credits," with an Infinite plan required. Details are in the PicassoIA section below.
Hosted open models with a free tier. ModelsLab lists a free tier for LTX 2.3 Fast image to video, then pay per use after that.
Consumer app credits. One tracker listed Kling 3.0 with 66 daily credits and Runway with 125 monthly credits as of February. Those are app allowances, not API calls, so do not build a product on them.
💡 The catch: free tiers come with caps, queues and the word "currently" in their terms. They are perfect for testing prompts and risky as the only plan for a paid product.
Open Weights on Your Own GPU
Open models such as Wan 2.2 I2V, LTX Video and Stable Video Diffusion carry no per-clip fee. You pay in hardware instead. Reported needs are roughly 10 to 16 GB of VRAM for Stable Video Diffusion at 576x1024 and 24 GB or more for Wan 2.2 I2V A14B at 720p. Hugging Face also hosts demo Spaces, but expect queues.
Renting a GPU by the hour moves the bill back to a per-clip number, so measure the GPU seconds each clip uses before you call it free.
Cheap Per-Second Prices
Per-second billing is the easiest to compare: multiply the rate by the clip length and you have the cost.
The Google rows include audio. Seedance 2 is billed per token instead ($0.014 per 1,000 tokens on fal's page), so you cannot convert it to a per-second rate without knowing how many tokens a clip uses. Run one test clip and read the invoice.
Resolution moves the bill too. On Google's page, Veo 3.1 Fast costs $0.10 per second at 720p, $0.12 at 1080p and $0.30 at 4K. Same model, triple the price. Draft at 720p and upgrade only the clips you keep.
Same Model, Four Prices
Kling 3.0 shows up at about $0.029 per second on one tracker and at $0.14 for the Pro tier on fal. The gap comes from the tier (standard or pro), whether audio is on, the reseller's margin, and the date of the snapshot. Compare like with like: same resolution, same duration, same audio setting, same tier.
What Cheap Actually Buys
Low prices usually come from smaller variants, not from a discount on the flagship. A Lite or Fast build trades peak detail and the highest resolutions for speed and volume. Seedance 2.5 Lite, for example, is limited to 480p and 720p on purpose. For social clips, product loops and animated storyboards, that is plenty. For a 4K hero shot on a client's homepage, it is not.
Whatever tier you pick, test it on your own images before you buy volume. The same few things break clips across almost every model: hands in motion, readable text inside the frame, fast camera moves and crowds of small figures. A calm subject with one clear motion, such as steam, drifting clouds, a slow push-in or a head turn, is where cheap models look their best. Run five of your real photos through a cheap tier and a mid tier, then compare the results side by side. If you cannot tell them apart at the size you publish, the cheaper one wins.
The credit spread is real. Apiframe's August table lists Seedance 1.5 Pro at 720p for 4 seconds at 15 credits, and Luma Ray 2 at 720p for 5 seconds at 131 credits. That is about 3.75 credits per second against 26.2, roughly seven times more.
Subscriptions follow a simple break-even rule: monthly fee divided by cost per clip. One tracker put Runway plans between $12 and $76 a month. A $12 plan equals about 34 clips at $0.35 each. If you make fewer clips than that, pay per second.
A 100 Clip Budget
Here is one take for each of 100 clips at 5 seconds and 720p:
💡 Plan for retries. Few prompts land on the first try. If you average 3 takes per usable clip, multiply every number in that table by 3.
Which Option Fits You
Testing prompts and ideas: use a free route, either the PicassoIA API or a free hosted tier, and accept the caps.
Social clips at volume: a Lite or Fast tier at 720p, such as Veo 3.1 Lite or Kling 2.5 Turbo Pro, keeps 100 clips between $25 and $35 on one take each.
Client work at 1080p or higher:Veo 3.1 Fast or Kling 3.0 Pro, with a retry budget on top.
Privacy or offline work: open weights such as Wan 2.2 I2V on your own GPU, if you already own the hardware.
Occasional use: skip subscriptions, which only pay off past the break-even count above.
The PicassoIA API
Models and Limits
The API lives at https://api.picassoia.com/v1 and uses a Bearer token that starts with pia_sk_. Endpoints follow the familiar Replicate pattern: create a prediction, poll it, fetch the output. Four models are exposed:
The limits matter for any batch job: 5 predictions in progress at once per account (a sixth returns 429 concurrency_limit), a 10 MB request body, 5 MB per data URL image, prompts up to 4,000 characters, and 2 tokens per account.
What It Costs Today
The docs state that API predictions are currently free and use no credits, and that the price is "Free with the Infinite plan." Without that plan, creating a prediction returns 403 plan_required. One wrinkle: the pricing page lists API Access and MCP Connections as new features on its plans, while the docs require Infinite. Treat the docs as the rule and check the plan page before you build on it. And remember the word currently.
A Working Request
This request animates a photo with Seedance 2.5 Lite. The parameters go inside an input object:
curl -s -X POST https://api.picassoia.com/v1/models/picassoia/seedance-2.5-lite/predictions \
-H "Authorization: Bearer $PICASSOIA_TOKEN" \
-H "Content-Type: application/json" \
-d '{"input": {"image": "https://example.com/first-frame.jpg", "prompt": "The camera pushes in slowly as steam rises from the cup and morning light shifts across the table", "resolution": "720p", "duration": 5}}'
You get a 201 response with an id, a status of starting, a polling URL and an eta. Poll GET /v1/predictions/{id} until the status is succeeded, failed or canceled. The documented Python loop is short:
import os, time, requests
API = "https://api.picassoia.com/v1"
HEADERS = {"Authorization": f"Bearer {os.environ['PICASSOIA_TOKEN']}"}
def run(model, input):
r = requests.post(f"{API}/models/{model}/predictions", json={"input": input}, headers=HEADERS)
p = r.json()
if not r.ok:
raise RuntimeError(f"{p['code']}: {p['detail']}")
while p["status"] not in ("succeeded", "failed", "canceled"):
time.sleep((p.get("eta") or {}).get("next_poll_in_seconds", 2))
p = requests.get(p["urls"]["get"], headers=HEADERS).json()
if p["status"] != "succeeded":
raise RuntimeError(p["error"] or p["status"])
return p["output"]
How to Use Seedance 2.5 Lite
No code needed if you only want a few clips. Seedance 2.5 Lite runs from the model page, and its description says Wonder members can generate without limits.
Pick the first frame. Upload the image you want to animate. The clip uses it as the opening frame and inherits its aspect ratio, so the aspect ratio setting is ignored.
Write the motion as a timeline. Describe what moves first, what the camera does, and how the light changes. Cinematic, chronological prompts work best.
Start at 480p. It renders fastest. Switch to 720p once the motion looks right.
Choose 5 or 10 seconds. Five seconds is enough for a product teaser or a social loop.
Decide on audio. Synchronized audio is on by default. Turn off save_audio for a silent clip.
Lock a seed when you like a result and want to reproduce it. Add a last frame image if you need the clip to end on a specific shot.
Three starting prompts for calm, low-risk motion:
Product shot: "Slow push-in on the bottle while soft window light moves across the glass, nothing else moves."
Portrait: "She turns her head toward the camera and smiles slightly, hair stirring in a light breeze, shallow depth of field."
Landscape: "Clouds drift over the ridge, mist moves through the valley, the camera rises slowly in golden hour light."
Need the first frame itself? Generate it with Picasso IA Image, fix details with Picasso IA Image Editor Pro, then animate it. Picasso IA Video is the simpler alternative: every clip is a fixed 5 seconds at 24 frames per second, at 480p or 720p.
Hidden Costs to Watch
Failed and Retried Clips
The price per second is the price of one attempt. Hands that melt, faces that drift and a camera move that ignores your prompt all mean another take. At 3 takes per usable clip, a $0.25 clip really costs $0.75. Ask each provider whether failed or canceled predictions are billed, and track takes per usable clip for your own content, because that ratio changes with every model.
Audio, Length and Concurrency
Audio can double the price. The same Apiframe table lists Seedance 1.5 Pro at 720p for 4 seconds at 15 credits, and 30 credits with audio.
Length scales the bill. A 10 second clip costs twice a 5 second clip on any per-second plan.
Concurrency costs time. With 5 jobs at once and about a minute per clip, 100 clips take roughly 20 minutes of waiting. Slower models stretch that into hours.
Make Your Own Clip Today
The cheapest way to find out what your clips will cost is to make one. Pick a photo you already own, such as a product shot, a landscape or a portrait. Open Seedance 2.5 Lite or Picasso IA Video on Picasso IA, add a one sentence motion prompt, and render a 480p draft. Then try three prompt variations and write down how many it took to get a clip you would post.
If you need a fresh first frame, create it with Picasso IA Image and animate it in the same session. When the draft works, move to 720p, and if your volume justifies it, call the API with the request above. Browse every video and image model at picassoia.com/en/all-models and run your own comparison before you commit to a paid provider.