PixVerse V5.5 Text-to-Video API

pixverse/v5.5/text-to-video
PixVerse V5.5 redefines AI video generation with native audio-visual synchronization and multi-shot camera control, all driven by natural language prompts.
Output
$13 / 1K tokens
Released
Dec 8, 2025

How to use PixVerse V5.5 Text-to-Video API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to pixverse/v5.5/text-to-video.
import requests, time

headers = {"Authorization": "Bearer " + AIMLAPI_KEY}
job = requests.post(
    "https://api.aimlapi.com/v2/video/generations",
    headers=headers,
    json={
      "model": "pixverse/v5.5/text-to-video",
      "prompt": "A serene timelapse of clouds over a mountain range"
    },
).json()
gid = job["id"]

while True:
    res = requests.get(f"https://api.aimlapi.com/v2/video/generations?generation_id={gid}", headers=headers).json()
    if res.get("status") in ("completed", "error"):
        break
    time.sleep(5)
print(res["video"]["url"])
const headers = {
  Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
  "Content-Type": "application/json",
};
const job = await (await fetch("https://api.aimlapi.com/v2/video/generations", {
  method: "POST",
  headers,
  body: JSON.stringify({
    "model": "pixverse/v5.5/text-to-video",
    "prompt": "A serene timelapse of clouds over a mountain range"
  }),
})).json();

let res;
do {
  await new Promise((r) => setTimeout(r, 5000));
  res = await (await fetch(`https://api.aimlapi.com/v2/video/generations?generation_id=${job.id}`, { headers })).json();
} while (!["completed", "error"].includes(res.status));
console.log(res.video.url);
# submit the generation — the response contains the job "id"
curl -X POST https://api.aimlapi.com/v2/video/generations \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"pixverse/v5.5/text-to-video","prompt":"A serene timelapse of clouds over a mountain range"}'

# then poll until status is "completed" — video.url holds the result
curl "https://api.aimlapi.com/v2/video/generations?generation_id={id}" -H "Authorization: Bearer $AIMLAPI_KEY"

OpenAI-compatible — swap the base URL and it works with your existing SDK.

PixVerse V5.5 Text-to-Video API Pricing

TypePrice
Output
$13 / 1K tokens

PixVerse V5.5 Text-to-Video Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Text-to-Video Arena
1201 (Elo)
Human preference Elo from blind pairwise comparisons of generated videos, measured independently by Artificial AnalysisSourceJuly 16, 2026

PixVerse V5.5 Text-to-Video vs other models

ModelInputOutputContextBest for
$13 / 1K tokens
Video generation
$0.26 / sec (variable)
Video generation
$0.13 / sec (variable)
Video generation
$1.95 / 1M tokens
$22.75 / 1M tokens
Video generation
$0.09243–$1.014 / sec (by resolution)
Cinematic video + native audio

Frequently asked questions

PixVerse V5.5 Text-to-Video takes text as input and returns video.

PixVerse V5.5 Text-to-Video became available on December 8, 2025.

PixVerse V5.5 Text-to-Video is priced at $13 / 1K tokens.

PixVerse V5.5 Text-to-Video is billed per generation — a fixed charge per output rather than by prompt length.

PixVerse V5.5 Text-to-Video was built by PixVerse.

Send a request to v2/video/generations with pixverse/v5.5/text-to-video as the model id.

Yes. PixVerse V5.5 Text-to-Video is served through AI/ML API, so the same key and endpoint format used for other models applies.

Yes, it generates audio-visual synchronization alongside the video output.

Yes, it includes multi-shot camera control for generated videos.

Start building with PixVerse V5.5 Text-to-Video

Get API Key
1000+ models, one API.