One Video API, Every Input Type
An AI video API usually means text-to-video. The ModelsLab video API is wider than that: a prompt, a still image, an existing clip or an audio track can each be the input, and the same key covers all of them. Text-to-video and image-to-video generate new footage; video-to-video restyles footage you have; lip sync matches a speaker’s mouth to a new audio track; motion control drives a subject with a reference motion.
The model is a parameter, not a separate integration. Open-weight models such as Wan and LTX run on ModelsLab GPUs and are covered by your plan. Third-party models — Kling, Veo, Seedance, Minimax, Runway, Vidu — sit behind the same endpoint and are billed per call from wallet balance, with the price shown on each model page before you send the request.