AI APIs for Developers
Filter by use case
AI Model APIs
191 modelsMinmax
MiniMax H3 Start/ End Frame To Video
Give H3 a start frame and an end frame — it generates the transition in native 2K with sound. Deterministic in and out points for edit-ready clips.
Minmax
MiniMax H3 Image To Video
Feed one image and a prompt, get a native 2K clip with sound. Preserves subject identity, lighting, and on-image text across motion.
Minmax
MiniMax-H3 Reference To Video
Pass reference images, video, or audio and describe how they relate. H3 handles subject, style, and motion transfer in one unified call.
Minmax
MiniMax H3 Text to Video (Hailuo-03)
MiniMax H3 is a frontier AI video model that turns text prompts into stunning 2K videos in just seconds. Generate 5–15 second cinematic clips in seven aspect ratios, perfect for social media, marketing, and professional content.
ltx
LTX 2.3 Reframe Video-to-Video
Reframe any video to a new aspect ratio — AI outpainting fills the missing areas so 16:9 footage becomes clean 9:16, 1:1, or 4:5 without cropping.
ltx
LTX 2.3 Fast Image To Video
LTX-2.3 Fast Image-to-Video is a powerful AI model that turns a single still image into a dynamic video clip using a text prompt to guide motion, camera moves, and atmosphere
ltx
LTX 2.3 Fast Text To Video
LTX-2.3 Fast Text-to-Video is an advanced AI model that converts text descriptions into high-quality short videos. It can generate cinematic visuals with synchronized audio, such as sound effects and ambience.
Tencent
Vidu Q3 Turbo Image to Video
The fastest way to bring images to life. Vidu Q3 Turbo animates photos into 1–16 second videos with synced audio and stable subjects — at a fraction of Pro pricing.
Vidu
Vidu Q3 Turbo Text to Video
Q3 quality at nearly half the price. Vidu Q3 Turbo turns text prompts into 1–16 second videos with native audio and smart scene cuts — built for speed and scale
Tencent
Vidu Q3 Pro Image to Video
Bring static images to life with Vidu Q3 Pro on ModelsLab — 4–16 second videos with synced audio, smart scene cuts, and stable subject consistency in 720P,1080p,2K,4K.
Tencent
Vidu Q3 Pro Text to Video
Vidu's flagship Q3 Pro model is live on ModelsLab. Turn text prompts into 1–16 second cinematic videos with native audio, smart scene cuts, and 1080p,2K & 4K output.
Bytedance
Seedream 5.0 Pro Text To Image
Generate stunning AI images from text prompts with Seedream 5.0 Pro. Create photorealistic visuals, illustrations, product renders, marketing assets, and concept art through a fast, developer-friendly API.
Bytedance
Seedream 5.0 Pro Image To Image
Next-generation image creation and editing model delivering ultra-fast 4K resolution outputs, multi-image reference support, natural language editing, and versatile style transfer for creative workflows.
Gemini Omni Video Edit
Gemini Omni Video Edit is live on ModelsLab. Upload any video clip, describe what you want changed in plain text, and the model applies the edit while preserving the parts you want to keep — no timeline, no keyframes.
Gemini Omni Text To Video
Google's Gemini Omni is live on ModelsLab. Describe any scene in plain text and get a cinematic video clip with synchronized audio — powered by Gemini's world knowledge and physics understanding.
Gemini Omni Image to Video
Gemini Omni Image to Video is live on ModelsLab. Upload any image as a first frame or reference, add a text prompt, and get a cinematic video clip with synchronized audio — no editing software needed.















