MODELLIX editorial cover reading VEO 3 FAST API over TIERS & PER-SECOND RATES, with three glowing glass gateways of increasing height on a dark amber technical base

“Veo 3 Fast” is a name that no longer maps to a live endpoint. Google shut veo-3.0-fast-generate-001 down on June 30, 2026, and the fast tier that answers today is veo-3.1-fast-generate-preview. That is not a footnote you can skip. As of September 13, 2026, most of the API-facing pages ranking for this term still sell the retired name: pollo.ai files “Veo 3 Fast” under its Legacy API index, Segmind documents POST /v2/veo-3-fast, and kie.ai prices “Veo 3 Fast” per 8-second clip — and none of the pages we checked mention the shutdown. A disclosure before we go on: Modellix publishes this blog and is itself an aggregator that routes Google’s video models, so the second half of this page has a commercial interest behind it. Every number below carries a source and a read date so you can check it without trusting us.

This page is the fast-tier counterpart to our Veo 3 series documentation — the tier’s live model codes, its spec ceiling, its per-second price at every resolution, the decision rule for choosing it over Standard or Lite, and the full call lifecycle on Modellix’s Google catalogue as of September 13, 2026. Two boundaries up front: Modellix does not carry the retired Veo 3.0 generation at all, so nothing here will revive an old integration, and if your workload already lives inside Google Cloud, going direct to the Gemini API is a legitimate choice we will describe rather than argue against.

Is the Veo 3 Fast API still available? The verified answer

Not under that model ID. Google’s deprecations table retires veo-3.0-fast-generate-001 — released September 9, 2025, shut down June 30, 2026 — and names veo-3.1-fast-generate-preview as the recommended replacement (Gemini API deprecations, read September 13, 2026). The same table is explicit about what shutdown means: once a model is shut down, “it is completely turned off, and the endpoint is no longer available.” There is no re-enable flag and no legacy tier. An even earlier id, veo-3.0-fast-generate-preview, went dark on November 12, 2025; the current fast id was released October 15, 2025 and carries no shutdown date announced as of September 13, 2026.

Three first-party surfaces, all read on September 13, 2026, agree:

Surface What it says today Read
Gemini API deprecations Veo table: veo-3.0-fast-generate-001 shut down June 30, 2026veo-3.1-fast-generate-preview. veo-3.1-fast-generate-preview: no shutdown date announced 2026-09-13
Gemini API model index The media-models table lists Veo 3.1 and Veo 3.1 Lite. Neither veo-3.0-fast-generate-001 nor veo-3.0-generate-001 appears 2026-09-13
Gemini API pricing The Veo section contains only the Veo 3.1 tiers (Standard / Fast / Lite). The Veo 3.0 rows have been removed 2026-09-13

One wrinkle is worth naming, because it is exactly how a dead model keeps looking alive: Google’s own Veo guide — page footer dated September 9, 2026 — still carries the legacy Veo 3 / Veo 3 Fast specification rows, and its feature table still labels that generation “Stable” — even though the same page’s model cards already read “Veo 3 (Deprecated)” and “Veo 3 Fast (Deprecated).” Read that against the deprecations table and the model index and the disagreement is one stale cell, not a different reality. The resolution rule is simple and unchanged: the deprecations table owns availability, the pricing page owns rates, and legacy spec tables are history. (The video generation overview now leads with Gemini Omni Flash and keeps Veo 3.1 for scene extension, last-frame control, and legacy pipelines.)

That internal disagreement is why every “Veo 3 Fast” listing deserves three checks before you build on it:

  1. Does it use a dated Google model code? Google’s codes are specific: veo-3.0-fast-generate-001, veo-3.1-fast-generate-preview. Strings like veo3-fast, veo-3-fast, or google/veo-3.1-fast are vendor aliases — the pattern Segmind’s POST /v2/veo-3-fast and pollo.ai’s Legacy API index both follow.
  2. Does it cross-reference the deprecations page? A page discussing Veo 3 Fast availability without the June 30, 2026 date was written before the fact and never updated.
  3. Do the stated specs match the model card? Veo 3 Fast was 720p and 1080p, 8 seconds — no 4K. Aspect ratio is the one place that same page contradicts itself: 16:9 only on Google’s feature table, while the parameter table on that page also lists 9:16 — one more reason to read the model card rather than a listing. A “Veo 3 Fast” listing advertising 4K is describing a different model.

What Veo 3 Fast actually is, and what the fast tier costs you in specs

Fast is a latency-and-cost tier inside the same generation family, not a separate architecture. On Google’s published tables, Veo 3.1 and Veo 3.1 Fast share the same rows for audio, frame rate, duration values, aspect ratios, and videos-per-request: natively generated audio that is always on, 24fps, durations of 4, 6, or 8 seconds, 16:9 or 9:16, and one output video per request (Veo API parameters and specifications, read September 13, 2026). The model card adds a text-input ceiling of 1,024 tokens.

Where the tiers actually diverge is resolution, and how resolution is bolted to duration:

Capability Veo 3.1 & Veo 3.1 Fast Veo 3.1 Lite Veo 3 & Veo 3 Fast (retired)
Resolution 720p; 1080p and 4K only at 8 seconds 720p; 1080p only at 8 seconds 720p & 1080p
Duration 8s only if 1080p, 4K, or reference images are used 8s only if 1080p or reference images 8 seconds only
Audio Always on (included in the per-second rate) Always on Always on

Two constraints bite in practice, and both are exact-match enums rather than free text:

  • personGeneration differs by mode. Text-to-video (and video extension) accepts allow_all only; image-to-video, interpolation, and reference-image jobs accept allow_adult only. Modellix’s request schema mirrors this exactly — its fast text-to-video route lists allow_all as the sole value, and its fast image-to-video route lists allow_adult (fast T2V request schema, fast I2V request schema, read September 13, 2026). Copying the wrong value from the other mode returns 400.
  • 1080p and 4K are not available at 4 or 6 seconds. If your shot list says “six seconds, 1080p,” the fast tier cannot serve it — you either go to eight seconds or drop to 720p.

One thing we will not put a number on: per-tier latency. Google’s documentation describes the fast tier qualitatively — built for “high quality and optimizing for speed and business use cases” (Gemini API video generation, read September 13, 2026) — and publishes no per-tier render seconds. Every “5× faster” or “10–15 seconds per clip” figure we found on ranking pages is unsourced, so treat any such claim, including any we might make later, as unverified.

Veo 3 Fast pricing per second: Google’s list rates vs the parameter-level grid

Veo is billed per second of output, with audio included by default, and there is no free tier. These are Google’s own list rates (Gemini API pricing, read September 13, 2026):

Google list rate (USD per output second, audio default) 720p 1080p 4K
Veo 3.1 Standard $0.40 $0.40 $0.60
Veo 3.1 Fast $0.10 $0.12 $0.30
Veo 3.1 Lite $0.05 $0.08 not supported
Free tier Not available Not available Not available

Google’s table gives you three tiers and three columns. What it does not give you is a per-model, per-parameter grid — which is the thing you need when your own pricing page has to quote a number. Here is the same family as Modellix publishes it, read from the live model pages on September 13, 2026:

Modellix route (USD per second) 720p 1080p 4K
google/veo-3.1-t2v $0.3600 $0.3600 $0.5400
google/veo-3.1-fast-t2v $0.0900 $0.1080 $0.2700
google/veo-3.1-fast-i2v $0.0900 $0.1080 $0.2700
google/veo-3.1-lite-t2v $0.0450 $0.0720 not supported
google/veo-3.1-lite-i2v $0.0450 $0.0720 not supported

Source: the live price dimension table on each route’s model page — fast T2V, fast I2V, standard T2V — plus the Lite route, each read September 13, 2026; the per-second unit definition is documented in Modellix pricing (to-video is billed by USD/sec, where sec is output duration). Rates move; these are dated snapshots, not a quote.

Three glass gateways of increasing height on a dark amber technical base, representing three price tiers with a second step inside each

The fast tier is not one price: the same 8-second task costs three different amounts depending on resolution, and the tiers stack on top of that.

Turn the grid into money and the ordering gets interesting. For an 8-second clip: $0.72 on fast at 720p, $0.864 at 1080p, $2.16 at 4K; $2.88 on Standard at 720p or 1080p; $0.36 on Lite at 720p. Standard’s 720p rate is exactly fast’s 720p rate. But fast at 4K ($0.2700/sec) is only about 25% cheaper than Standard at 720p ($0.3600/sec). The conclusion is uncomfortable for the usual advice: choosing the fast tier does not automatically save money — dropping resolution does. Fast solves latency; the resolution column solves the bill.

Two honest notes on the surrounding market. First, reseller markup is real and legitimate, but it is not the underlying rate: DeepInfra lists google/veo-3.1-fast at $0.1500 per second as of September 13, 2026 — above Google’s list rate for the same tier, which is what a wrapper looks like. Second, undated price claims are unfalsifiable: one ranking guide asserts $5–15 per video without linking a single Google page, and another cites a Vertex figure of $0.75 per second that does not match the tiers Google publishes today. If a page will not date its numbers, you cannot audit them.

A four-second and a six-second panel sealed beneath a low translucent ceiling while an eight-second panel rises through into the high-resolution tier

Duration gates resolution: 1080p and 4K are only reachable at eight seconds, which is why the tier choice and the duration choice cannot be made separately.

When the fast tier is the right pick — and when it is not

The rule is not “fast is for when you need speed.” It is this: pick the resolution you can ship, and the duration that resolution forces, and only then pick the tier.

If your constraint is… Use Why
Interactive iteration, previews, bulk background generation, 720p is acceptable Fast (720p) Lowest-rate tier that keeps the full 3.1 feature set; $0.0900/sec on Modellix as of 2026-09-13
Product-facing 1080p, clip is 8 seconds anyway Fast (1080p) $0.1080/sec — still roughly a third of Standard’s rate, and 1080p only exists at 8s
You need 4K at volume Fast (4K), or re-check Standard Fast 4K is $0.2700/sec vs Standard 720p $0.3600/sec — the gap is small enough that quality may win
You need 4 or 6 seconds at 1080p Neither — change duration or resolution 1080p is unavailable below 8 seconds on every current Veo tier
Lowest-cost 720p, quality is second priority Veo 3.1 Lite (720p) $0.0450/sec on Modellix as of 2026-09-13; no 4K, and no reference-image support in Google’s table
Reference-image or video-extension work Standard / Fast, duration fixed at 8 Reference and extension modes require duration: "8"

Two decision inputs that are not about price. Latency is the tier’s reason to exist, and it is also the one thing no first-party page quantifies — so if latency is your deciding factor, measure it on your own prompt before committing, rather than trusting a comparison table. Vendor surface matters more than people expect: if your infrastructure already sits inside Google Cloud, the deprecations page points at the GA models on the Gemini Enterprise Agent Platform as a replacement path, and staying on one cloud with one billing line is a real advantage. The aggregator case is the inverse — one key, one balance, the same request shape across Standard, Fast, and Lite, and parameter-level rates you can quote internally. That is a genuine difference and it is not an argument that Modellix is the cheapest way to reach Veo; the tier you pick changes the bill by roughly 4×, and no vendor choice can undo that. The six live routes are listed together on the Veo 3.1 series page, and the full tier-by-tier price breakdown sits in our Veo 3.1 price analysis rather than being repeated here.

Calling Veo 3 Fast on Modellix: submit, poll, retrieve

The media API is asynchronous: one POST to create a task, one GET to collect it. The paths no longer carry an /async suffix — that was removed on June 11, 2026, and older /async endpoints still work, so nothing breaks if your code base predates the change (product updates, read September 13, 2026). Base host for all media generation is https://api.modellix.ai.

1
2
3
4
5
6
7
8
9
10
11
12
13
# Submit a fast-tier text-to-video task (read 2026-09-13)
curl -sS -X POST "https://api.modellix.ai/api/v1/google/veo-3.1-fast-t2v" \
-H "Authorization: Bearer $MODELLIX_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"prompt": "Time-lapse of cherry blossoms falling in a peaceful Japanese garden",
"aspectRatio": "16:9",
"duration": "8",
"resolution": "1080p"
}'
# -> {"code":0,"message":"success","data":{"status":"pending",
# "task_id":"task-t2v002","model_id":"google/veo-3.1-fast-t2v",
# "get_result":{"method":"GET","url":"https://api.modellix.ai/api/v1/tasks/task-t2v002"}}}

The submit response hands you the polling URL, so there is no need to build it by hand. Poll it until data.status leaves pending, then read the asset URL out of data.result.resources[]; the same response carries billing.status and billing.amount for that one task (Get task result, read September 13, 2026):

1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
import os, time, requests

KEY = os.environ["MODELLIX_API_KEY"]
API = "https://api.modellix.ai/api/v1"
H = {"Authorization": f"Bearer {KEY}"}

r = requests.post(f"{API}/google/veo-3.1-fast-t2v", headers=H, json={
"prompt": "Time-lapse of cherry blossoms falling in a peaceful Japanese garden",
"aspectRatio": "16:9", "duration": "8", "resolution": "1080p",
}).json()
task_id = r["data"]["task_id"]

while True:
d = requests.get(f"{API}/tasks/{task_id}", headers=H).json()["data"]
if d["status"] in ("success", "failed", "canceled"):
break
time.sleep(10)

if d["status"] == "success":
print(d["result"]["resources"][0]["url"]) # signed asset URL
print("billed", d["billing"]["amount"], d["billing"]["status"])
A submitted task entering a glass chamber, branching into a polling loop and a webhook delivery chute, both feeding a single retrieved result panel

One submit, two ways to collect: poll the task, or let Modellix deliver it to your endpoint when it reaches a terminal state.

Veo 3.1 Fast Request Reference

See the exact request body, enums, and async task format for google/veo-3.1-fast-t2v and -i2v.

View Docs

Before you write anything else by hand, two discovery endpoints save an integration day. GET /api/v1/models returns every active model with slug, type, docs_url, and description — filter it for image-to-video and you have the current catalogue rather than a hardcoded list (List models, read September 13, 2026). And modellix-cli model get-schema google/veo-3.1-fast-t2v prints the request contract without an API key, which means your schema check and your build are not gated on the same permission (Get schema, read September 13, 2026). If you also need the rest of the family’s request surface — reference images, last-frame interpolation, video extension — our Veo 3.1 API guide covers the routes this page deliberately leaves out.

Parameters, inputs, and where your starting frame comes from

Text-to-video takes six fields, and only one is required.

Field Type Required Accepted values Notes
prompt string yes free text Minimum length 1
negativePrompt string no free text what to discourage
aspectRatio string no 16:9, 9:16 defaults to 16:9
duration string no 4, 6, 8 string, not integer
resolution string no 720p, 1080p, 4k 1080p and 4k require duration: "8"
personGeneration string no allow_all text-to-video only accepts allow_all

Source: fast T2V request schema, read September 13, 2026. The image-to-video route swaps prompt semantics and adds a starting frame: image is required, lastFrame is optional and forces duration: "8", personGeneration accepts allow_adult instead, and image mode and reference mode are mutually exclusive (fast I2V request schema, read September 13, 2026).

Image-to-video needs the frame to be reachable over HTTP, which is the first thing most integrations get wrong — a path to a local file in the image field will fail. Modellix’s File API exists for exactly this: POST /api/v1/media/files with a multipart/form-data body (field name file) returns a file_id and a URL you can pass straight into the image field (Upload media file, read September 13, 2026). Five constraints decide whether you can build on it:

  • Uploads are not billed and are not subject to balance admission.
  • Retention is about 7 days by default — this is the constraint that shapes architecture, not the price. Long-lived assets must live in your own storage; the File API is a hand-off, not an asset library.
  • Maximum file size is 16 MB, and the default cap is 10 files per team with 2 concurrent uploads.
  • Images, video, and audio are all accepted (image extensions include png, webp, jpg, heic), so the same endpoint serves the frame you are animating and any asset you feed a different model later.
  • The list and delete endpoints (GET/DELETE /api/v1/media/files) let you free quota deterministically instead of waiting for expiry.

Running it in production: webhooks, cost logs, and the errors you will hit

Polling is fine for a six-second 720p clip. It is the wrong shape for a queue of 4K tasks, and the median failure mode is not a crash — it is a worker parked in time.sleep while your concurrency budget sits idle. Modellix supports webhook delivery instead: add an X-Webhook-URL header to the prediction creation call and the result is pushed when the task reaches a terminal state (prediction.task.succeeded, prediction.task.failed, or prediction.task.canceled) (REST API webhooks, read September 13, 2026).

Four delivery rules decide how you write the receiver:

  • Your endpoint must be a public HTTPS address — no localhost, no private ranges, no credentials embedded in the URL.
  • Return a 2xx to acknowledge; 200 with a plain-text ok, or 204 with an empty body, is the documented pattern.
  • Delivery is retried on 429, 5xx, network timeouts, and temporary network errors, and is not retried on 3xx, 4xx other than 429, invalid URLs, or permanent connection errors. Retried deliveries carry an incrementing X-Modellix-Retry-Count.
  • Deduplicate on X-Modellix-Delivery-ID, not on the task id — that is the documented idempotency key, and it is the difference between a 4K clip landing once in your CDN and landing four times.

Then the bill. Media request logs are queryable per team over a window of at most 30 days: GET /api/v1/logs?start_time=&end_time= (UNIX seconds) returns entries carrying task_id, status, and cost, and request logs record input parameters as well, so a price change or a surprise 4K task is attributable to a specific request rather than to a monthly total (Get logs, read September 13, 2026). If you resell access to your own users, tag each call with the optional X-Mdlx-User-Id header (8–128 characters, ASCII letters, digits, -, and _) and the same log endpoint will filter by that id — note that the list response does not echo the field back, so filter by the query parameter. For the account side, the console supports email alerts when the team balance drops below a threshold you set (product updates, read September 13, 2026); a depleted balance surfaces as 402, not as a generation failure.

Errors follow one shape — code equals the HTTP status, message is "<Category>: <detail>":

Status Meaning Retry?
400 Missing parameter, invalid enum, malformed body No — fix the payload
401 Missing, invalid, or expired API key No
402 Insufficient balance or account in arrears No — top up
404 Unknown task id, model, or provider No
429 Rate or concurrency limit exceeded Yes — exponential backoff
500 / 503 Internal error / temporarily unavailable Yes — bounded retry

Source: REST API error handling on the same documentation page cited for webhooks, read September 13, 2026. The most common 400 on this tier is not a typo — it is a number sent where a string is required ("8", not 8) or an allow_all sent to the image-to-video route.

Run Veo 3.1 Fast on One Key

Log in to call the fast, standard, and Lite Veo 3.1 routes from one balance with per-request cost logs.

Login

Frequently Asked Questions

Is the Veo 3 Fast API free?
No. Google’s pricing page lists the free tier for every Veo 3.1 tier as “Not available” (read September 13, 2026), and the fast tier is billed per second of output. Modellix ended its signup credit on August 19, 2026; trial credit is requested by emailing support@modellix.ai. Any page offering free Veo 3 Fast minutes is offering a trial allowance or a different model.

Which model ID do I use today?
veo-3.1-fast-generate-preview on the Gemini API, or the -fast-i2v / -fast-t2v variants. On Modellix, the slug is google/veo-3.1-fast-t2v or google/veo-3.1-fast-i2v. The retired string veo-3.0-fast-generate-001 no longer resolves anywhere.

Should I use Veo 3.1 or Veo 3.1 Fast?
Decide on resolution and duration first. Fast is the right pick when 720p or 8-second 1080p is enough and latency or volume matters. Choose standard Veo 3.1 when you need the highest fidelity or when the same output has to hold up as a finished asset rather than a draft — and note that at 4K the fast tier is only about 25% cheaper than standard at 720p, so the tier alone is not the saving.

Does Veo 3 Fast generate audio?
Yes, and it is always on. Audio is generated with the video in the same pass, which is why Google’s per-second prices are quoted as “video with audio price (default)” — you are not paying a separate line for it, and you cannot switch it off on the current tiers.

Can I still call veo-3.0-fast-generate-001?
No. It was shut down on June 30, 2026, and Google states that a shut-down model’s endpoint is no longer available. If a listing still answers under that name, you are calling a vendor alias whose upstream target the vendor can change without telling you.

Why do so many pages still sell “Veo 3 Fast”?
Because a listing is cheap to leave up and expensive to audit. Of the API-facing pages ranking for this term on September 13, 2026, most — including Segmind’s and pollo.ai’s — still present the retired name, and the ones we read did not disclose the shutdown.

Do I need a new API key or account to move to the fast tier?
No. Standard, Fast, and Lite run on the same key, the same base host, and the same submit-poll lifecycle; the tier is a model slug, not a platform. That is also why the tier decision is reversible in production — change the slug, keep the client. Our Veo 3.1 API key guide walks through the credential path if you are coming from a different provider.

Can I see output before paying?
Yes, on Google’s side, through AI Studio’s Veo preview surfaces. On Modellix, the Playground runs the same media models in a browser with the same account and pricing, so a prompt can be validated there before it becomes an API call.

The decision in one paragraph

If the term that brought you here is the one Google retired, the actionable version of it is veo-3.1-fast-generate-preview — and the only decision that meaningfully changes your outcome is resolution and duration, because they are coupled and they set the price: $0.0900 per second at 720p, $0.1080 at 8-second 1080p, $0.2700 at 4K on Modellix as of September 13, 2026, against $0.3600 on the standard tier. Pick the resolution your product can ship, let it fix your duration, then pick the tier — and wire the webhook and the cost log before the first 4K job, not after it.

If you are still choosing between model families rather than between tiers of one family, how Kling and Veo 3 compare and our image-to-video API guide are the better next reads; this page assumed you had already settled on the Google family and only needed the fast tier.


Model availability and pricing reflect public information as of September 13, 2026 and change frequently — Google has retired two Veo generations since April 2025, and a tier name is not a contract. Validate against each provider’s live pricing before committing. Access image and video models through a single API key at modellix.ai. Modellix publishes this blog, operates the aggregator described above, and routes Google’s media models behind one API key.