Modellix cover: two-line title AI VIDEO COST CALCULATOR and THE RANGE BEHIND EVERY PER-SECOND PRICE over a four-strata glass metering column, MODELLIX wordmark

Almost every “AI video generation cost calculator” you can find does the same thing: take one per-second price, multiply it by a number of seconds, show a total. That arithmetic is correct. The input is usually wrong. On the 112 video models listed on Modellix as of September 15, 2026, 88 charge more than one rate, and the spread inside a single model can exceed 11× between its cheapest and most expensive tier. A calculator that collapses that into one number is not a calculator — it is a guess with a decimal point.

So this page does the arithmetic honestly, with three inputs instead of one: the rate for the tier you actually rendered, the seconds you keep, and the attempts it takes to get them. It is written for developers paying per second through an API — not for people shopping for a subscription tool, and not to crown a winner. We run Modellix, so we have a commercial interest in the numbers below; they all come from model pages we also sell, captured on the date shown, and where we could not verify something we say so instead of filling the gap.

How to calculate AI video generation cost

The calculation is one multiplication, but it has three inputs and only one of them is a price:

cost = rate for the tier you rendered ($/sec) × (seconds of output you keep × attempts per kept second)

Two of those inputs are yours. The rate is not. That split is the whole reason a cost estimate goes wrong:

  • The rate belongs to the model you picked and the tier you rendered at — resolution, or a mode flag, depending on the model. It is per second of generated video.
  • Seconds of output you keep is your finished runtime, which is the number you were probably already estimating from a script or a storyboard.
  • Attempts per kept second is the multiplier nobody puts in the calculator. A ten-second shot that takes four generations is 40 billed seconds, and your invoice will say 40, not 10.

Two details matter when you go looking for the rate. First, it is on the model page, under a Pricing heading, listed per tier — not in the header chip, which shows only the span. Second, “tier” is not always resolution, and not every model has more than one. Both are covered below.

Diagram of a metering assembly with three gate modules labelled RATE TIER, SECONDS and ATTEMPTS feeding a graduated collection vessel labelled BILLED SECONDS

Figure: the three inputs to an AI video generation cost estimate. The rate tier is set by the model; seconds and attempts are set by you. Captured 2026-09-15.

The per-second rate is a range, not a number

Of the 112 video models live in the Modellix catalog as of September 15, 2026, 24 carry a single rate and 88 carry a range — the model page prints them as $0.0430~$0.4670/sec because the price depends on which tier you select. Catalog-wide, rates run from $0.0050/sec (vidu/viduq2-pro-digital-human, a single-tier model) to $1.4400/sec (alibaba/happyhorse-1.0-video-edit at 1080P) — a 288× spread across models that are all called by the same kind of request.

Here is what six of those tier tables actually say, taken from each model page’s Pricing section on September 15, 2026:

Model Tier axis Tiers as published Top ÷ bottom
bytedance/seedance-2.0-t2v resolution 480p 0.0700 · 720p 0.1500 · 1080p 0.3700 · 4k 0.7800 11.14×
bytedance/seedance-2.0-v2v resolution 480p 0.0430 · 720p 0.0930 · 1080p 0.2280 · 4k 0.4670 10.86×
google/gemini-omni-1.1-flash-t2v resolution 360p 0.0338 · 720p 0.1000 · 1080p 0.1520 · 4k 0.3041 9.00×
minimax/hailuo-02-i2v resolution 512P 0.0153 · 768P 0.0504 · 1080P 0.0738 4.82×
alibaba/wan3.0-v2v resolution 480P 0.3000 · 720P 0.6000 · 1080P 1.2000 4.00×
kling/kling-avatar mode, not resolution std 0.0448 · pro 0.0896 2.00×

Two things in that table are worth more than the table itself.

The tier axis is not always resolution. kling/kling-avatar prices on a mode flag — std or pro. If your estimate assumes “1080p will cost more than 720p”, you will miss the models where the axis is something else entirely, and you will have no idea which tier you were even quoting.

More resolution is not always more money. google/veo-3.1-t2v prices 720p and 1080p at the same $0.3600/sec, and only the 4k tier moves to $0.5400. The intuitive rule — “higher resolution, higher rate” — holds for bytedance/seedance-2.0-t2v and alibaba/wan3.0-v2v and fails here, which is why the tier table has to be read per model rather than assumed.

Screenshot of the Pricing section on the Modellix model page for bytedance/seedance-2.0-t2v, showing Unit: $/sec and four resolution tiers

Figure: the Pricing table on bytedance/seedance-2.0-t2v, retrieved 2026-09-15 — four resolution tiers from 0.0700 to 0.7800 per second, the 11.14× span that the header price chip shows only as a range.

Two of the calculators ranking on this query handle the problem in opposite ways. One keeps vendor units separate “instead of inventing a conversion between image outputs, video credits and exported minutes” (aipricing.guru). The other pre-loads one rate per model and applies generic quality multipliers — HD as 1.5× and 4K as 2× the base rate (abacktools) — while stating that its data is “based on publicly available API and subscription rates as of 2025”, with a model list that still leads with Runway Gen-4, Kling 1.6, Pika 2.2 and Google Veo 2. A 1.5× step is a reasonable guess for some models and wrong by more than 7× for Seedance 2.0, whose real step from 480p to 4k is 11.1×.

The units really are that heterogeneous. Google bills its Gemini video models by output tokens and publishes the conversion itself: 5,792 tokens per second of 720p video, which it translates to “an effective price of approximately $0.10 per second” (Gemini API pricing). Runway meters the same category in credits, and one plan’s 625 credits buy 52 seconds of one model or 104 seconds of another (Runway pricing) — so a credit is not a unit of video, it is a unit of a specific model. None of that is a defect; it just means a per-second figure is only comparable to another per-second figure when both were produced the same way.

Worked example: one 5-second clip at four resolutions

Take a single 5-second clip on bytedance/seedance-2.0-t2v, the model whose tier table is in the screenshot above. Each line is the tier rate multiplied by 5, so you can check it against the table:

Tier Rate 5 seconds
480p $0.0700/sec 5 × $0.0700 = $0.3500
720p $0.1500/sec 5 × $0.1500 = $0.7500
1080p $0.3700/sec 5 × $0.3700 = $1.8500
4k $0.7800/sec 5 × $0.7800 = $3.9000

Same prompt, same model, same 5 seconds — 11.14× apart purely from the tier. If your estimate said “about 35 cents for a 5-second clip”, it was right for 480p and wrong for everything you would actually ship.

For contrast, the same 5 seconds on google/veo-3.1-t2v costs 5 × $0.3600 = $1.80 at 720p, 5 × $0.3600 = $1.80 at 1080p, and 5 × $0.5400 = $2.70 at 4k. Note that the 720p and 1080p figures are identical, because that model does not charge for the step — which is exactly the kind of thing a single “rate” column destroys.

Worked example: 60 seconds, and the two-minute question

A finished video is more than one clip, and every clip in it gets retried. Here is the arithmetic for a 60-second finished video, stated so you can substitute your own numbers: 12 shots × 5 seconds = 60 seconds of output you keep, × 3 attempts = 180 seconds of video actually generated. The bill is charged on 180, not 60.

Model Tier Arithmetic Cost of the 60-second cut
bytedance/seedance-2.0-t2v 480p 180 × $0.0700 $12.60
bytedance/seedance-2.0-t2v 720p 180 × $0.1500 $27.00
bytedance/seedance-2.0-t2v 1080p 180 × $0.3700 $66.60
bytedance/seedance-2.0-t2v 4k 180 × $0.7800 $140.40
google/veo-3.1-t2v 720p or 1080p 180 × $0.3600 $64.80
google/veo-3.1-t2v 4k 180 × $0.5400 $97.20
google/gemini-omni-1.1-flash-t2v 720p 180 × $0.1000 $18.00
google/gemini-omni-1.1-flash-t2v 1080p 180 × $0.1520 $27.36
kling/kling-avatar std 180 × $0.0448 $8.06
kling/kling-avatar pro 180 × $0.0896 $16.13
minimax/hailuo-02-i2v 512P 180 × $0.0153 $2.75
minimax/hailuo-02-i2v 1080P 180 × $0.0738 $13.28
alibaba/wan3.0-v2v 480P 180 × $0.3000 $54.00
alibaba/happyhorse-1.0-video-edit 1080P 180 × $1.4400 $259.20

The last row is the one you will not find on the calculators ranking on this query: they are built around text-to-video, and a 60-second cut is not where those costs land. Image-to-video (58 models) and video-to-video (28 models) together account for 86 of the 112 video models here, and they behave differently at the same resolution: minimax/hailuo-02-i2v at 512P is $2.75 for the same 60 seconds that alibaba/happyhorse-1.0-video-edit bills at $259.20.

The most-asked version of this question is about two minutes, so: a 2-minute finished video is 24 shots × 5 seconds × 3 attempts = 360 seconds generated. bytedance/seedance-2.0-t2v at 480p is 360 × $0.0700 = $25.20; google/gemini-omni-1.1-flash-t2v at 720p is 360 × $0.1000 = $36.00; google/veo-3.1-t2v at 720p is 360 × $0.3600 = $129.60. Same runtime, 5× between the cheapest and dearest of those three.

The attempt count deserves its own line, because it is the input you control and the one that scales everything. Holding the model and tier fixed at bytedance/seedance-2.0-t2v 720p and the finished video at 60 seconds:

Attempts per shot Seconds generated Arithmetic Cost
1 60 60 × $0.1500 $9.00
2 120 120 × $0.1500 $18.00
3 180 180 × $0.1500 $27.00
6 360 360 × $0.1500 $54.00

Three versus six attempts is a $27 swing on a single 60-second cut — on the same model and tier, more than the step from 480p to 720p, and it costs nothing but prompt discipline to move it the other way. Picking the attempts number is your judgement call; publishing it as an assumption is ours. The 12-shot, 5-second, 3-attempt figures above are one worked example, not a standard. Others in this space assume three to ten attempts per usable shot (LTX) or three to five for complex motion, and the fact that two published estimates disagree by that much is the reason the number belongs in your inputs rather than in a vendor’s table.

Estimate Your Own Workload

Log in to Modellix and run the same arithmetic on the models you actually plan to call.

Login

What the per-second rate does not include

Input assets. For image-to-video and video-to-video, the reference image or source video has to get somewhere the API can reach it. On Modellix that is the File API, a free, API-key-authenticated upload for image, video and audio assets that feeds directly into prediction requests — so a first frame is not a cost line in the estimate. It is not storage you should lean on, though: uploaded files are retained for about 7 days by default, long enough for a production run and not long enough to treat as an archive. If you are multiplying a per-second rate to get a monthly figure, the input path is not part of it.

Account-level discounts are not rate discounts. The Modellix changelog lists a 10% discount on the first top-up of each tier — that is money in your account balance, applied at checkout, and it is a different mechanism from any per-model discount published on a model page. If you are modelling cost per second, the top-up discount does not change the rate; it changes how far your balance goes. Do not multiply them together and do not treat one as evidence for the other.

Tasks that were not charged. The media request log records cost as 0 when nothing was billed, so the log — not an assumption — is where you check whether a failed or skipped generation cost you anything. Read the status and the cost on the same row.

Other vendors’ units. Do not convert a subscription plan’s credits into seconds and compare the result to a per-second API rate. Runway’s own plan page shows why: 625 credits is 52 seconds of Gen-4.5 or 104 seconds of Gen-4 Turbo, so the same credit buys two different amounts of video inside one vendor, let alone across two. The honest move is to keep the units apart, which is what one competitor’s calculator does as a matter of stated policy. One more honest caveat on catalog scope: OpenAI’s Sora 2 and its Videos API are not in the 112 models priced here, and OpenAI’s documentation states they are deprecated with a shutdown date of September 24, 2026 (OpenAI video generation guide). A calculator that still lists Sora as a line item is quoting a model you cannot build on for much longer.

Two linked skills carry the mechanics this page deliberately does not repeat: how a text-to-video request is submitted and polled, and the same for image-to-video, which is the path most product features actually take.

Compare video models under one key

The reason this calculator can cover all 112 video models in one place is not that we are generous with data — it is that they are all callable with one credential. Switching models means changing a string in the request, not opening an account. That matters for costing specifically, because the only defensible way to compare per-second rates is to send the same prompt to several models at the same tier and compare what came back, and that comparison is impossible if each model sits behind a separate contract.

If you would rather not hand-copy rates from model pages, the catalog is queryable: GET /api/v1/models returns the live model list with slugs, types and descriptions, and each model page carries its own Pricing table at a predictable path — https://www.modellix.ai/models/bytedance/seedance-2.0-t2v, and so on for every entry. The full model directory is the index; one API, many video models covers what actually changes between them at the request level.

I should be precise about what that buys you, and what it does not. It is not a claim that Modellix is the cheapest place to run any of these models. Some of the models above are cheaper elsewhere, some are not, and we have not run that comparison here. The claim is narrower: one key and one billing surface makes the measurement cheap, and measuring is the part most estimates skip.

Media API Reference

See the request fields, the upload step, and how to read a task result for any video model.

View Docs

Verify the estimate against the actual bill

Everything above is an estimate built from published rates. An estimate you cannot falsify is not worth much, so the last step is to check the arithmetic against what you were actually charged — and this is where the request log does the work. Modellix exposes media request logs at GET https://api.modellix.ai/api/v1/logs, paginated, with a time window of at most 30 days. Each entry carries the task_id, the model (provider and model_name), the status, the wall-clock duration, the input the task was submitted with, and — the field this whole page has been building toward — a cost.

Read the unit before you reconcile anything, because cost is not in dollars. The Get Logs reference documents it as the billed amount in sub-pennies, where 1 equals USD $0.0001, the same unit the console balance uses, to be divided by 10,000 for dollars, and 0 if the task was not charged. A log row reading cost: 12 is $0.0012, not twelve dollars, and the fastest way to convince yourself your estimate is wrong is to skip that division.

That one field turns the estimate into a closed loop: multiply the tier rate by the generated seconds, predict the number, then read the log row for the same task and compare. Where they disagree, the cause is almost always one of three inputs — a tier you mis-set, an attempt you forgot, or a duration the model clipped. All three are visible in the same log entry, because it records the input parameters each request was made with.

The same loop is worth running on a video-to-video pass, which is where estimates drift furthest from reality: upscaling an existing clip bills per second of output just like generation, but it sits in the video-to-video band where rates in this catalog range from $0.0120/sec to $1.4400/sec depending on the model and tier.

Frequently Asked Questions About AI Video Generation Cost

How much does it cost to make an AI video?
With a per-second API, the cost is the tier rate multiplied by the seconds of video generated — including retries. At rates measured September 15, 2026, a single 5-second clip runs from 5 × $0.0050 = $0.025 at the cheapest tier in the catalog to 5 × $1.4400 = $7.20 at the dearest. A 60-second finished video, assuming 12 shots of 5 seconds and 3 attempts each, works out to 180 billed seconds, which came to $12.60 at bytedance/seedance-2.0-t2v 480p and $259.20 at alibaba/happyhorse-1.0-video-edit 1080P on the same day.

How much does a 2-minute video cost to generate?
Applying the same shape — 24 shots × 5 seconds × 3 attempts — gives 360 seconds of generated video. bytedance/seedance-2.0-t2v at 480p is 360 × $0.0700 = $25.20; google/gemini-omni-1.1-flash-t2v at 720p is 360 × $0.1000 = $36.00; google/veo-3.1-t2v at 720p is 360 × $0.3600 = $129.60. Change the shot count, shot length or attempt count and every figure moves proportionally.

Why is one model listed as a price range like $0.0700~$0.7800/sec?
Because that model charges different rates for different tiers, and the header shows the span. The model page’s Pricing section lists each tier separately — for bytedance/seedance-2.0-t2v, 480p at 0.0700, 720p at 0.1500, 1080p at 0.3700 and 4k at 0.7800 per second, as published on September 15, 2026. Always estimate from the tier you will actually render.

Is the per-second rate always based on resolution?
No. Most video models here tier on resolution, but not all — kling/kling-avatar tiers on a mode flag, with std at $0.0448/sec and pro at $0.0896/sec. Read the Pricing table for the specific model rather than assuming the axis.

Which AI video model is the cheapest?
We do not answer that, and you should be sceptical of any single number that does. Rates vary by tier, tier axes differ, and 88 of the 112 video models here charge more than one rate, so “cheapest” is only meaningful once you fix the model, the resolution and the duration. Browsing the AI video generator collection by output type is a more useful first move than sorting by price: picking the lowest rate in a catalog is a five-minute exercise, while picking the right model for a shot is the actual work, and the tables above give you the inputs to do it.

Does an AI video generator app cost the same as the API?
No, and the two should not be converted into each other. Consumer apps meter in credits or monthly allowances where one credit buys different amounts of video depending on the model, while API access meters in seconds (or, at Google, in output tokens). This page covers the API side only.

Do failed generations get billed?
Check the log row rather than assuming either way: the media request log records cost as 0 when a task was not charged, next to the status for that same task, so a failure’s actual cost is answerable per request on your own account.


Provider rates, tier tables and model counts reflect Modellix catalog data captured on September 15, 2026 and change frequently — validate against each model page’s Pricing section before committing to a budget. Access image and video models, including the leading Chinese models, through a single API key at modellix.ai.