VideoAnonymized

Wan 2.7 Enhanced

Alibaba's open-weights video model that generates 1080p, audio-enabled clips up to 15 seconds in text-to-video and image-to-video modes.

Maker
Alibaba
Modality
Video + audio
Max duration
15 seconds
Max resolution
1080p

Overview

What is Wan 2.7 Enhanced

Wan 2.7 Enhanced is Alibaba's open-weights video generation model, available on Venice in text-to-video and image-to-video variants. It produces 1080p clips up to 15 seconds long with native audio, supports multiple aspect ratios, and runs under an anonymized privacy tier with zero prompt retention.

Using it anonymously on Venice

On Venice, Wan 2.7 Enhanced runs under an anonymized privacy tier with zero prompt retention — your prompts are not stored, profiled, or used for training. Because the model carries open weights and is tagged uncensored, you get permissionless sovereignty over creative workflows without Big-Tech surveillance. You simply pay per clip, scaling from $0.72 for a 720p·5s render.

AnonymizedNo prompt trainingTEE · hardware enclaveEnd-to-end encrypted

Assessment

Strengths and limitations

Strengths
  • Fully open weights and uncensored: suitable for self-hosting, fine-tuning, and creative sovereignty outside closed platforms.
  • Generates 1080p clips up to 15 seconds with native audio and three aspect ratios: 16:9, 9:16, and 1:1.
  • Dual modality via two Venice variants: text-to-video (wan-2-7-enhanced-text-to-video) and image-to-video (wan-2-7-enhanced-image-to-video).
  • Cinematic motion quality with believable camera grammar and strong texture detail for landscapes, products, and atmospheric scenes.
  • Runs privately on Venice with anonymized, zero-retention inference — no prompt storage or profiling.
Limitations
  • Hard 1080p / 15-second ceiling; no 4K or extended-duration mode for longer storytelling.
  • No tool-use, vision, reasoning, or web-search integrations — it is a pure generative video model.
  • Crowd scenes can show identity drift, and complex physics remain a step behind leading closed rivals.
  • Audio is generated natively but may not match dedicated audio pipelines; post-production sound mixing is still recommended for professional work.
  • Open-weights quality depends on inference optimization; the Venice-hosted version is tuned, but self-hosted setups vary.

Samples

Sample outputs

Generated on Venice with our standard prompt suite — the same prompts we run through every model of this type, so you can judge it like-for-like.

Cinematic landscape

Slow aerial drone shot gliding over a misty mountain valley at golden hour, sunlight piercing clouds onto a winding river, ancient pine forests on either side, ultra-smooth motion, professional color grading, atmospheric haze, 4K

Seamless loop

Calm ocean waves rolling onto a black-sand beach at sunrise, golden light on wet sand, foam dissolving into the shore, a single silhouetted figure at the waterline, smooth continuous forward push, serene cinematic atmosphere

Urban cinematic

A slow tracking shot through a rain-soaked Tokyo street at night, neon reflecting in puddles, steam rising from a food cart, a person with a translucent umbrella, shallow depth of field, teal-and-orange grade, smooth steady camera

Compare every video model on these prompts

Capabilities

What it supports

  • Text to video
  • Image to video
  • Reference to video
  • Native audio generation

Variants

Wan 2.7 Enhanced model variants

Wan 2.7 Enhanced runs on Venice as 2 variants of the same underlying model. Pick by what you're starting from: a written prompt, a still image, reference images, or an existing clip. Each variant is its own model id on the API; the generation quality is the same across the family.

VariantWhat it isClip lengthsResolutionsAspect ratiosAudioModel ID
Text to VideoflagshipGenerate a clip from a written prompt5s, 10s, 15s1080p, 720p16:9, 9:16, 1:1wan-2-7-enhanced-text-to-video
Image to VideoAnimate a still image into motion5s, 10s, 15s1080p, 720pwan-2-7-enhanced-image-to-video

Capability data comes straight from the Venice model API and refreshes with every catalog ingest. The specs and pricing on this page are captured from the flagship variant; pass the model id of the variant you want to the API.

Wan 2.7 Enhanced Text to Video

Generate a clip from a written prompt. Supports clips of 5s, 10s, 15s, 1080p, 720p output, 16:9, 9:16, 1:1 aspect ratios, with native audio.

wan-2-7-enhanced-text-to-video

Wan 2.7 Enhanced Image to Video

Animate a still image into motion. Supports clips of 5s, 10s, 15s, 1080p, 720p output, with native audio.

wan-2-7-enhanced-image-to-video

Specifications

Datasheet

Maker
Alibaba
Released
April 2026
Modality
Text-to-video, image-to-video
Max resolution
1080p
Open weights
Yes
Resolutions
1080p, 720p
Clip lengths
5s, 10s, 15s
Mode
text-to-video
Aspect ratios
16:9, 9:16, 1:1
Audio
Yes
Privacy on Venice
Anonymized — prompts not stored
Available on Venice since
Jun 2026

API

Call it from your code

Venice exposes this model through the REST API. Queue a generation with the model id.

curl https://api.venice.ai/api/v1/video/queue \
  -H "Authorization: Bearer $VENICE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "wan-2-7-enhanced-text-to-video",
    "prompt": "Aerial drone shot over a misty mountain valley at golden hour"
  }'

# Use the returned queue_id with https://api.venice.ai/api/v1/video/retrieve.
# Call /video/complete after downloading if needed.

Pricing

What it costs on Venice

Pay per clip on Venice — price scales with resolution and duration (5s–15s), from $0.72.

1080p · 5s
$1.10
Per clip
720p · 5s
$0.72
Per clip

New Venice accounts include a free daily allowance and 500 welcome credits — no credit card required.

Alternatives

How it compares

ModelBest forMax resolutionMax durationOpen weightsPrice (Venice)
Wan 2.7 EnhancedThe open-weights, audio-enabled generation suite with zero prompt retention.1080p15sYesfrom $0.68
Wan 2.7The base open-weights Wan model. Enhanced adds native audio and Venice-specific optimizations.Yesfrom $0.55
Kling O3 ProKuaishou's closed flagship — strong cinematic storytelling but no self-hosting or open fine-tuning.Nofrom $0.46
Vidu Q3Budget-friendly closed rival with smooth motion; lacks open-weight flexibility.Nofrom $0.27

The open-weights, audio-enabled generation suite with zero prompt retention.

Use cases

What it is good for

  1. 01Short-form social content (TikTok/Reels) in native 9:16 with audio.
  2. 02Product and brand promos where rapid 5–15 second clips are sufficient.
  3. 03Animating still photos into motion for editorial or e-commerce.
  4. 04Prototyping cinematic B-roll and camera moves before committing to a full shoot.
  5. 05Uncensored creative and artistic video generation without platform guardrails.

Prompting

Getting better results

Tag the desired aspect ratio explicitly in the prompt, e.g., 'cinematic 16:9' or 'vertical 9:16'.

Use cinematographic language — specify lens, light, and camera move (drone shot, dolly-in, golden hour) for smoother results.

Iterate with 5-second clips at 720p to test composition cheaply, then scale to 1080p and 10s–15s.

For image-to-video, start with a high-resolution still and describe the motion you want rather than re-describing the scene.

Version history

Wan 2.1
2025

Earlier open-weights release praised for motion quality.

Wan 2.7
2026-04

Base model family adding image-to-video, editing, and audio conditioning.

Wan 2.7 Enhanced
2026-06

Venice-optimized variant with native audio and anonymized inference.

FAQ

Frequently asked questions

Wan 2.7 Enhanced is Alibaba's open-weights AI video model hosted on Venice. It generates 1080p clips up to 15 seconds long from text or images, with native audio, multiple aspect ratios, and fully anonymized inference.

Pricing is per clip and scales with resolution and duration. A 720p·5s clip starts at $0.72, while a 1080p·5s clip is $1.10. Longer 10s and 15s renders cost more accordingly. There is no subscription.

Yes. Wan 2.7 Enhanced is an open-weights model, meaning the weights are available for self-hosting and fine-tuning outside Venice. On Venice you pay per clip rather than buying a subscription.

You can try Wan 2.7 Enhanced on Venice using trial credits; after that, you pay per clip starting at $0.72 with no subscription required.

Yes. The family includes the sibling model wan-2-7-enhanced-image-to-video, which animates a still image into a 5–15 second motion clip at 1080p or 720p with audio.

Wan 2.7 Enhanced is the Venice-optimized variant with confirmed native audio, anonymized inference, and per-clip pricing. The base Wan 2.7 is also open-weights but may lack the same audio stack and privacy wrapper on Venice. Choose Enhanced for production audio and private, pay-as-you-go access.

It supports 1080p and 720p resolutions, with clip lengths of 5, 10, and 15 seconds. You can generate in 16:9, 9:16, or 1:1 aspect ratios.

Yes. On Venice it runs under an anonymized privacy tier — your prompts are not stored, profiled, or used for training. There is zero prompt retention.

Yes. Both the text-to-video and image-to-video variants output clips with native audio generation enabled.

Use Wan 2.7 Enhanced anonymously

Venice does not store your prompts. Chat history stays in your browser.