Veo 3.1 Fast
Veo 3.1 Fast by Google AI: billed per media unit (second, image or minute of audio), 480 context. Call it through the odnoga LLM gateway; see the pricing page for the media rate.
Veo 3.1 Fast is served by Google AI and called through odnoga with the same OpenAI-compatible request shape as every other model in the catalog. Its 480-token context window sets how much input you can send in one request.
Specification and price
| Vendor | Google AI |
|---|---|
| Model ID | veo-3.1-fast-generate-preview |
| Context window | 480 tokens |
| Max output | 8K tokens |
| Pricing | Per media unit (second / image / minute) — see pricing |
| Capabilities | — |
Vendor cost is the list price per million tokens as recorded in the odnoga catalog; your price applies your plan margin with the same formula that bills every request — see pricing. Pricing
Call it through odnoga
const res = await fetch("https://api.odnoga.com/v1/chat/completions", {
method: "POST",
headers: {
Authorization: `Bearer ${process.env.ODNOGA_API_KEY}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
model: "veo-3.1-fast-generate-preview",
messages: [{ role: "user", content: "Hello" }],
}),
});
Same request shape for every vendor in the catalog — swap the model id and odnoga handles keys, routing, limits and cost accounting.
How to use Veo 3.1 Fast
Derived from the odnoga catalog record for this model.
Best for
- High-volume, latency-sensitive calls: classification, extraction, routing, short rewrites and summaries.
Not the right pick when
- Anything that needs to read an image — this model takes text only.
- Agent loops that must call your functions — tool calling is not available here.
- Pipelines that require guaranteed JSON — parse defensively or pick a model with enforced JSON.
- Large documents in one request — the window is 480 tokens, so you will need chunking.
- Anything where a wrong answer is costly without a human check — no model in the catalog removes that requirement.
Practical tips through odnoga
- 01Pin the model id in a managed prompt version, so a model swap is a version change you can compare and roll back, not an edit in application code.
- 02Compare it against 2–8 other models on the same frozen test cases in the evaluation laboratory before you make it the production default.
- 03For repeated identical deterministic calls, odnoga answer reuse returns the stored answer and bills no vendor tokens — turn it off for creative output.
- 04Set a fallback model on the route so a vendor incident degrades quality instead of returning an error, and a per-tenant budget so one caller cannot spend the month.
Veo 3.1 Fast compared
| Model | Context | Input / 1M | Output / 1M | Capabilities |
|---|---|---|---|---|
| Veo 3.1 Fast | 480 | — | — | — |
| Gemini 2.5 Computer Use | 131K | $1 | $5 | Vision, Tools, JSON mode, Streaming |
| Gemini 2.5 Flash Image (Nano Banana) | 33K | $0.3 | $2.50 | Vision |
Questions
- How much does Veo 3.1 Fast cost per 1M tokens?
- Google AI lists — / — per 1M input / output tokens in the odnoga catalog. Through odnoga you pay that vendor price plus your plan margin, and every request is recorded with both numbers.
- What is Veo 3.1 Fast best for?
- High-volume, latency-sensitive calls: classification, extraction, routing, short rewrites and summaries.
- Can I switch to Veo 3.1 Fast without changing my code?
- Yes. odnoga exposes one OpenAI-compatible endpoint, so switching means sending "veo-3.1-fast-generate-preview" as the model id — or changing it in the managed prompt version, with no application deploy.
- How large is the Veo 3.1 Fast context window?
- 480 tokens of input, with up to 8K tokens of output per response.
All models · Pricing · Docs · Compare models in the evaluation lab
Where this model sits
Other Google AI models
Gemini 2.5 Computer Use
$1 / $5 per 1M
Gemini 2.5 Flash Image (Nano Banana)
$0.3 / $2.50 per 1M
Gemini 2.5 Flash Native Audio (Live)
$0.5 / $2 per 1M
Gemini 2.5 Flash TTS
$0.5 / $10 per 1M
Gemini 2.5 Pro
$1.25 / $10 per 1M
Gemini 2.5 Pro TTS
$1 / $20 per 1M
Gemini 3 Flash Preview
$0.5 / $3 per 1M
Gemini 3 Pro Image (Nano Banana Pro)
$2 / $12 per 1M