alibaba/wan-3.0/text-to-video
alibaba
text-to-video
Alibaba WAN 3.0 Text-to-Video generates videos from text prompts with flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls. Ready-to-use REST inference API, best performance, no cold starts.
black-forest-labs/flux-3/video-extend
black-forest-labs
video-extend
Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene.
bytedance/seedream-v5.0-pro/edit
bytedance
image-to-image
Seedream 5.0 Pro API Preview Edit is ByteDance's advanced image editing model for single-image and multi-reference image generation.
openai/gpt-image-2/text-to-image
openai
text-to-image
OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts.
alibaba/happyhorse-1.1/video-extend
alibaba
video-extend
Alibaba Happy Horse 1.1 (Video Extend) extends existing videos with seamless AI-generated continuation, supporting 720p/1080p output. Natural, motion-consistent extension.
bytedance/seedance-2.5/video-edit
bytedance
video-to-video
Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output.