We dropped our image preprocessing pipeline entirely. The Vidu Q1 API takes the prompt and ships the clip — nothing else needed.
The Vidu Q1 API generates 5-second 1080p video clips from a natural-language text prompt — no image input required. Each Vidu Q1 API call supports adjustable motion amplitude (auto / small / medium / large), optional background music, and a reproducible seed, all through a single OpenAI-compatible REST endpoint at 15% below the official rate.
SampleRun the model to replace this previewVidu Q1 API: Text-to-Video
Use the Vidu Q1 API to generate 5-second 1080p video clips from a text prompt alone — no image input required.
The Vidu Q1 API (video_generate.vidu_q1_t2v_novita) generates 5-second 1080p video clips from a natural-language text prompt — no image or reference input is required. Each Vidu Q1 API call supports adjustable motion amplitude (auto, small, medium, or large), optional background music, and a reproducible integer seed, all via a single OpenAI-compatible REST call on the Vidu Q1 API. Because there is no image to host, the text-to-video task is the fastest way to turn an idea into motion — write a sentence, send it, and receive a finished clip for storyboards, ad variants, and rapid concept work.
Pricing for the Vidu Q1 API text-to-video is $0.34 per generation — 15% below the official published rate. One RouterBase API key covers the Vidu Q1 API and every other model in the catalog; no separate Vidu account is needed. Because the Vidu Q1 API is async and OpenAI-compatible, you can fan out dozens of prompt variations in parallel and collect every clip from a signed CDN URL.
Six reasons teams ship on the Vidu Q1 API
From pure text input to per-call motion control — what makes the Vidu Q1 API text-to-video task stand out.
Pure text input
The Vidu Q1 API needs only a natural-language prompt — no image upload, no reference file, no preprocessing — so an idea becomes a clip in one request.
5s 1080p per call
Every Vidu Q1 API call returns a 5-second 1080p clip. The fixed duration and resolution remove encoding guesswork and keep cost predictable across a batch.
Motion amplitude control
Pass auto, small, medium, or large per Vidu Q1 API call. Small suits calm scenes, large drives sweeping action, and auto lets the Vidu Q1 API read the prompt and choose.
Optional background music
Add bgm=true and the Vidu Q1 API scores the clip with a fitting track; omit the flag for a silent output you can pair with your audio.
15% below official rate
$0.34 per generation on the Vidu Q1 API — 15% below the Vidu Q1 published price. Flat pricing regardless of motion or BGM, so spend stays predictable.
OpenAI-compatible REST API
One endpoint, one key — the Vidu Q1 API uses the same JSON shape and Bearer header as all 200+ RouterBase models, so switching to it is only a model ID change.

Get started with the Vidu Q1 API in 3 steps
From sign-up to your first text-to-video call in under 5 minutes.
Write your prompt
Describe the scene you want the Vidu Q1 API to generate — subject, setting, action, and mood. The more vivid the prompt, the more purposeful the resulting motion.
POST to the Vidu Q1 endpoint
Send your prompt, optional motion_amplitude, bgm flag, and seed in JSON. The Vidu Q1 API returns an async job; poll GET /v1/images/generations/{id} until status=success.
Download and embed the clip
The success response from the Vidu Q1 API contains a signed CDN URL under the results array. Drop it into a <video> tag or copy the file to your bucket — the 1080p MP4 is ready to use.
Pay only for what you use
RouterBase passes through partner-tier pricing. Compared against the model's official published API rate.
| Resolution | RouterBase | Official | Save |
|---|---|---|---|
| generation | $0.340 | −15% |
One run at the default settings (generation) costs $0.340.With $1 you can run this model approximately 2 times.
What teams build with the Vidu Q1 API
Real production loads running on the Vidu Q1 API and the RouterBase catalog.
One sentence prompt, one call — a polished 5-second clip. Our content team can't believe it's this simple.
Pure text-to-video means zero image hosting overhead. It cut our pipeline from four steps to one.
We iterate on prompts and the Vidu Q1 API handles the rest. Motion amplitude auto keeps things dynamic without extra tuning.
I added the Vidu Q1 API in an afternoon. Text in, video out — it's exactly the interface I wanted.
RouterBase puts 200+ models behind one key. The Vidu Q1 API is our go-to for rapid concept videos.
At $0.34 per clip it fits inside our ad budget — and we generate dozens of variants per campaign.
Same OpenAI-compatible JSON, different model ID. The Vidu Q1 API was live in our system within an hour.
15% off the official rate adds up fast when you generate hundreds of clips per week for clients.
Frequently Asked Questions
Common questions about the Vidu Q1 API text-to-video task.
The Vidu Q1 API text-to-video task generates a 5-second 1080p clip from a natural-language prompt only — no image or reference input is required. It is the fastest entry point on the Vidu Q1 API for turning a written idea into motion.