Multimodal video generation
The API is presented as a platform for advanced multimodal models, with documented video workflows based on text, images, keyframes, and reference inputs.
Vidu API is a developer platform for integrating AI video generation into applications. It provides API workflows for text-, image-, keyframe-, and reference-based video creation, along with access to other Vidu model capabilities.
Vidu API is a developer-focused platform for integrating Vidu’s multimodal video-generation models into applications. It supports documented workflows for creating video from text, images, keyframes, and references, with additional platform options for image-to-video, start/end-to-video, and real-time video.
The service includes API documentation, a quickstart, an API Playground, model and pricing information, and credit-based account billing. It is intended for developers and enterprises building or scaling video-driven products rather than only using a standalone video editor.
The API is presented as a platform for advanced multimodal models, with documented video workflows based on text, images, keyframes, and reference inputs.
The text-to-video documentation describes JSON request data and token-based authorization for submitting video-generation requests.
Platform navigation lists image-to-video, start/end-to-video, reference-to-video, and real-time video alongside text-to-video.
The platform organizes capabilities across video generation models including Vidu Q3, Q2, and Q1, with separate categories for image, audio, avatar, and editing models.
Accounts use rechargeable credits, while plans specify validity periods and concurrent-job allowances. Tasks above a plan’s concurrency limit enter a queue.
A quickstart, API documentation, API Playground, pricing console, and Help Center provide resources for implementation and evaluation.
Developers can use the API documentation and quickstart to connect Vidu generation workflows to a video-driven product or application.
Applications that start with prompts, images, keyframes, or reference material can select the corresponding generation workflow instead of using a single text-only process.
Teams can review the model catalog, use the API Playground, and compare credit consumption by model, resolution, task type, and duration.
Creators, studios, and professional content teams can use higher-volume credit packages with stated concurrency allowances for ongoing generation workloads.
Products that require more immediate or interactive output can investigate the platform’s listed real-time video capability, subject to the current API documentation and account limits.
The platform is described for enterprises and developers building and scaling video-driven products with Vidu’s multimodal models.
The platform lists text, images, keyframes, and reference inputs, as well as image-to-video, start/end-to-video, and real-time video workflows. Specific input requirements depend on the selected API operation.
The text-to-video documentation shows an Authorization header using a token and JSON as the data exchange format. Consult the current API documentation for the exact request and credential setup.
Usage is based on rechargeable credits. The pricing page states a standard rate of $0.005 per credit, while actual consumption varies by capability, model, resolution, task type, and duration. Credits have plan-specific validity periods.
The pricing page states that tasks exceeding a plan’s fixed concurrent-job allowance automatically enter a queue. The allowance varies by recharge plan.
segmind.com
Segmind automates image and video workflows with APIs, Pixelflow templates, model access, and plans for individuals, teams, and enterprises.
thinkdiffusion.com
Cloud workspace for open-source AI image and video tools
producer.ai
Google Flow Music is a generative AI platform for creating, remixing, and sharing songs, with music video generation, audio remixing, and tool building in a web app with free and paid plans.
onorca.dev
Orca is an Agent Development Environment for shipping with coding agents, running parallel CLI agents in isolated worktrees with desktop and mobile companion workflows.
rizzle.com
Rizzle turns articles into editorial-grade video for publishers and enterprises, with AI assistance, human review, and multi-channel distribution and monetization.
easysite.ai
EZsite AI is an AI website builder that turns a URL into a full-stack React or Vue.js application, with hosting, custom domains, code export, and backend features for deployable team projects.