Kwaivgi

Kwaivgi (developed by Kuaishou) is a industry-leading provider of commercial AI video generation models. The platform hosts the complete Kling AI model family, spanning from cost-effective Kling 1.6/2.1 tiers up to the flagship Kling 3.0 and 4K-capable Omni models. Engineered for real-world physical simulation and fluid spatiotemporal motion, Kwaivgi supports text-to-video, image-to-video, video editing, local replacements, Motion Control transfer, and precise Lip-sync alignment.

Kwaivgi

All Models

Kling V3.0

4 variants available

Kling Video O1

Kling Video O1

Kling Video O1 (Omni One) is Kuaishou's industry-first unified multimodal video model that merges video generation and editing into a single engine. Integrating text-to-video, image-to-video, element referencing, localized inpainting, and video restyling, it supports referencing up to 7 subjects simultaneously to lock character and prop consistency. Creators can perform conversational video editing via natural language prompts—ideal for advertising, VFX, and end-to-end film production.

OfficialText-to-VideoImage-to-VideoVideo-to-Video
$0.084/SEC
Kling v3.0 Image-to-Video

Kling v3.0 Image-to-Video

Kling v3.0 Image-to-Video is Kuaishou's next-generation multimodal AI video model. Built on the unified Omni architecture, it takes static images or subject references to generate up to 15-second cinematic videos in up to 4K resolution. It features native audio-visual synchronization, multilingual lip-sync, and enhanced subject consistency to prevent visual drift, along with intelligent multi-shot control—ideal for commercial ads, film VFX, and narrative short videos.

OfficialImage-to-Video
$0.084/SEC
Kling v3.0 Text-to-Video

Kling v3.0 Text-to-Video

Kling v3.0 Text-to-Video is Kuaishou's next-generation AI video model. Powered by the native Omni architecture, it accurately parses complex prompt text to generate up to 4K cinematic-grade videos up to 15 seconds long. It natively supports integrated audio-video generation (ambient audio, music, and multilingual lip-sync) alongside exceptional visual realism, multi-shot coherence, and complex physical simulation—ideal for commercial advertising, film VFX, and content creation.

OfficialText-to-Video
$0.084/SEC

Kwaivgi Models API Pricing Details

ModelPricing (USD)Our Pricing (USD)Discount
Kling Video O1$0.084/SECStart from$0.084/SEC
Kling v3.0 Image-to-Video$0.084/SECStart from$0.084/SEC
Kling v3.0 Text-to-Video$0.084/SECStart from$0.084/SEC
Kling 3.0 Omni$0.084/SECStart from$0.084/SEC

Explore models from other providers