Wan 2.7 Video API is Alibaba Tongyi Lab's AI video suite, covering four generation modes — Text-to-Video, Image-to-Video, Reference-to-Video, and Video Edit. It supports full-modality input (text, image, video, audio) and outputs 720P–1080P video. Access all four modes through a single API on Kie.ai.

Model Type:
Pricing: Pricing: 16 credits per second for 720p (~$0.08) and 24 credits per second for 1080p (~$0.12). High-tier top-ups (+10% bonus) bring effective pricing down to ~$0.072 per second for 720p and ~$0.108 per second for 1080p.
Input

The positive prompt for the generation

0/5000

The negative prompt for the generation

0/500

Click to upload or drag and drop

Supported formats: MPEG, WAV, X-WAV, AAC, MP4, OGG Maximum file size: 50MB

Audio URL for the generation

The resolution of output

The ratio of output

Total duration of output

Wether to enable the prompt optimizer

Wether to enable the watermark

The random seed to use for the generation.(>=0,<=2147483647)

A configurable parameter. Defaults to true in the Playground.

There is no guarantee that everything can be filtered out; if you are not satisfied with the results, you will need to make your own arrangements.
Output
output typevideo

README

Complete guide to using Wan 2.7 Video

Wan 2.7 Video API – T2V, I2V, R2V & Video Edit

Original image

Product videos no longer require a production crew. Use the Wan 2.7 T2V API to generate polished ad creatives and demo clips directly from text prompts — or feed in image references via I2V to lock in brand visuals across every output. Multi-angle grid input lets teams iterate on creative directions fast, cutting turnaround from days to minutes.

Short-form platforms demand volume and consistency. With Wan 2.7 AI video generation on Kie.ai, content teams can produce lifestyle clips, hook videos, and scene transitions at scale — then use the VideoEdit API to swap backgrounds, shift tone, or localize content for different markets via a single natural-language instruction. One API call replaces hours of post-production.

Translate scripts into visual storyboards before a single frame is shot. Wan 2.7 API supports multi-shot narrative sequencing and camera movement control, letting directors and writers preview scenes with real motion — not static sketches. R2V mode keeps character appearance consistent across shots, making it practical for animatics and pitch decks alike.

Repurpose archival or low-quality footage without reshooting. The Wan 2.7 VideoEdit API handles black-and-white colorization, degraded footage restoration, and full style transfers — all triggered by plain-language instructions. Studios and archivists can modernize entire video libraries through batch API calls on Kie.ai, with the original motion structure preserved throughout.

4.9/ 5
36,885 ratings
Tap a star to rate

Wan 2.7 Video API is a suite of AI video generation models released by Alibaba's Tongyi Lab in April 2026. It includes four sub-models — T2V, I2V, R2V, and VideoEdit — covering the full creative pipeline from text-to-video generation to instruction-based editing. On Kie.ai, all four modes are accessible via a unified API endpoint.

Wan 2.7 API supports four modes: Text-to-Video (T2V) for prompt-driven generation, Image-to-Video (I2V) for animating static images, Reference-to-Video (R2V) for multi-reference character and voice replication, and VideoEdit for natural-language instruction-based editing of existing footage. Each mode targets a distinct stage of the video production workflow.

Wan 2.7 AI video API accepts text prompts, single images, 9-grid multi-image sets, existing video clips (2–10 seconds for VideoEdit), and audio files for voice reference in R2V mode. This full-modality input support — text, image, video, and audio — makes it one of the most flexible video generation APIs available on Kie.ai.

Wan 2.7 VideoEdit takes an existing 2–10 second video and edits it based on a natural-language instruction. Supported operations include local element swaps, background changes, style transfer, and black-and-white colorization — all while preserving the original motion structure. No re-generation is required.

T2V generates video from a text prompt alone — no image input needed. I2V takes one image or a 9-grid multi-image set as input, with optional first and last frame control so the model auto-infers motion in between. I2V is better suited when visual consistency and subject identity matter more than creative freedom.

Wan 2.7 introduces several capabilities not available in Wan 2.6: first and last frame control, 9-grid multi-image input, up to five simultaneous references in R2V, and the full VideoEdit instruction mode. Overall video quality, temporal consistency, and style range have also been significantly improved.

Wan 2.7 api can be accessed through Kie AI with an API key. Developers can test wan 2.7 video api in playground, review documentation, and integrate wan 2.7 ai video into applications using standard API requests.

The key capability to test is multimodal referencing for more consistent visual guidance from references; try MiniMax H3 (Hailuo-03) on Kie.ai for that workflow.