Wan 2.7 Enhanced
Alibaba's open-weights video model that generates 1080p, audio-enabled clips up to 15 seconds in text-to-video and image-to-video modes.
Generate videoGet API keyWhat is Wan 2.7 Enhanced?
Wan 2.7 Enhanced is Alibaba's open-weights video generation model, available on Venice in text-to-video and image-to-video variants. It produces 1080p clips up to 15 seconds long with native audio, supports multiple aspect ratios, and runs under an anonymized privacy tier with zero prompt retention.
Use Wan 2.7 Enhanced privately on Venice
On Venice, Wan 2.7 Enhanced runs under an anonymized privacy tier with zero prompt retention — your prompts are not stored, profiled, or used for training. Because the model carries open weights and is tagged uncensored, you get permissionless sovereignty over creative workflows without Big-Tech surveillance. You simply pay per clip, scaling from $0.72 for a 720p·5s render.
What can Wan 2.7 Enhanced do?
- •Fully open weights and uncensored — suitable for self-hosting, fine-tuning, and creative sovereignty outside closed platforms.
- •Generates 1080p clips up to 15 seconds with native audio and three aspect ratios: 16:9, 9:16, and 1:1.
- •Dual modality via two Venice variants — text-to-video (wan-2-7-enhanced-text-to-video) and image-to-video (wan-2-7-enhanced-image-to-video).
- •Cinematic motion quality with believable camera grammar and strong texture detail for landscapes, products, and atmospheric scenes.
- •Runs privately on Venice with anonymized, zero-retention inference — no prompt storage or profiling.
- •Hard 1080p / 15-second ceiling; no 4K or extended-duration mode for longer storytelling.
- •No tool-use, vision, reasoning, or web-search integrations — it is a pure generative video model.
- •Crowd scenes can show identity drift, and complex physics remain a step behind leading closed rivals.
- •Audio is generated natively but may not match dedicated audio pipelines; post-production sound mixing is still recommended for professional work.
- •Open-weights quality depends on inference optimization; the Venice-hosted version is tuned, but self-hosted setups vary.
Sample outputs
Generated on Venice with our standard prompt suite — the same prompts we run through every model of this type, so you can judge it like-for-like.
Slow aerial drone shot gliding over a misty mountain valley at golden hour, sunlight piercing clouds onto a winding river, ancient pine forests on either side, ultra-smooth motion, professional color grading, atmospheric haze, 4K
Calm ocean waves rolling onto a black-sand beach at sunrise, golden light on wet sand, foam dissolving into the shore, a single silhouetted figure at the waterline, smooth continuous forward push, serene cinematic atmosphere
A slow tracking shot through a rain-soaked Tokyo street at night, neon reflecting in puddles, steam rising from a food cart, a person with a translucent umbrella, shallow depth of field, teal-and-orange grade, smooth steady camera
Wan 2.7 Enhanced model variants
Wan 2.7 Enhanced runs on Venice as 2 variants of the same underlying model — pick by what you're starting from: a written prompt, a still image, reference images, or an existing clip. Each variant is its own model id on the API; the generation quality is the same across the family.
| Variant | What it is | Clip lengths | Resolutions | Aspect ratios | Audio | Model ID |
|---|---|---|---|---|---|---|
| Text to Videoflagship | Generate a clip from a written prompt | 5s, 10s, 15s | 1080p, 720p | 16:9, 9:16, 1:1 | wan-2-7-enhanced-text-to-video | |
| Image to Video | Animate a still image into motion | 5s, 10s, 15s | 1080p, 720p | — | wan-2-7-enhanced-image-to-video |
Capability data comes straight from the Venice model API and refreshes with every catalog ingest. The specs and pricing on this page are captured from the flagship variant; pass the model id of the variant you want to the API.
Wan 2.7 Enhanced Text to Video
Generate a clip from a written prompt. Supports clips of 5s, 10s, 15s, 1080p, 720p output, 16:9, 9:16, 1:1 aspect ratios, with native audio.
wan-2-7-enhanced-text-to-videoWan 2.7 Enhanced Image to Video
Animate a still image into motion. Supports clips of 5s, 10s, 15s, 1080p, 720p output, with native audio.
wan-2-7-enhanced-image-to-videoHow to use Wan 2.7 Enhanced via API
Venice exposes this model through the REST API. Queue a generation with wan-2-7-enhanced-text-to-video.
curl https://api.venice.ai/api/v1/video/queue \
-H "Authorization: Bearer $VENICE_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "wan-2-7-enhanced-text-to-video",
"prompt": "Aerial drone shot over a misty mountain valley at golden hour"
}'
# Use the returned queue_id with https://api.venice.ai/api/v1/video/retrieve.
# Call /video/complete after downloading if needed.Specifications
Pricing
Pay per clip on Venice — price scales with resolution and duration (5s–15s), from $0.72.
New Venice accounts include a free daily allowance and 500 welcome credits — no credit card required.
Wan 2.7 Enhanced vs alternatives
| Model | Max resolution | Max duration | Open weights | Price (Venice) |
|---|---|---|---|---|
| Wan 2.7 Enhanced | 1080p | 15s | Yes | from $0.68 |
| Wan 2.7 | — | — | Yes | from $0.55 |
| Kling O3 Pro | — | — | No | from $0.46 |
| Vidu Q3 | — | — | No | from $0.27 |
The open-weights, audio-enabled generation suite with zero prompt retention.
What is Wan 2.7 Enhanced good for?
- •Short-form social content (TikTok/Reels) in native 9:16 with audio.
- •Product and brand promos where rapid 5–15 second clips are sufficient.
- •Animating still photos into motion for editorial or e-commerce.
- •Prototyping cinematic B-roll and camera moves before committing to a full shoot.
- •Uncensored creative and artistic video generation without platform guardrails.
Prompting tips
- •Tag the desired aspect ratio explicitly in the prompt, e.g., 'cinematic 16:9' or 'vertical 9:16'.
- •Use cinematographic language — specify lens, light, and camera move (drone shot, dolly-in, golden hour) for smoother results.
- •Iterate with 5-second clips at 720p to test composition cheaply, then scale to 1080p and 10s–15s.
- •For image-to-video, start with a high-resolution still and describe the motion you want rather than re-describing the scene.
Version history
Earlier open-weights release praised for motion quality.
Base model family adding image-to-video, editing, and audio conditioning.
CurrentVenice-optimized variant with native audio and anonymized inference.
Frequently asked questions
Wan 2.7 Enhanced is Alibaba's open-weights AI video model hosted on Venice. It generates 1080p clips up to 15 seconds long from text or images, with native audio, multiple aspect ratios, and fully anonymized inference.
Pricing is per clip and scales with resolution and duration. A 720p·5s clip starts at $0.72, while a 1080p·5s clip is $1.10. Longer 10s and 15s renders cost more accordingly. There is no subscription.
Yes. Wan 2.7 Enhanced is an open-weights model, meaning the weights are available for self-hosting and fine-tuning outside Venice. On Venice you pay per clip rather than buying a subscription.
You can try Wan 2.7 Enhanced on Venice using trial credits; after that, you pay per clip starting at $0.72 with no subscription required.
Yes. The family includes the sibling model wan-2-7-enhanced-image-to-video, which animates a still image into a 5–15 second motion clip at 1080p or 720p with audio.
Wan 2.7 Enhanced is the Venice-optimized variant with confirmed native audio, anonymized inference, and per-clip pricing. The base Wan 2.7 is also open-weights but may lack the same audio stack and privacy wrapper on Venice. Choose Enhanced for production audio and private, pay-as-you-go access.
It supports 1080p and 720p resolutions, with clip lengths of 5, 10, and 15 seconds. You can generate in 16:9, 9:16, or 1:1 aspect ratios.
Yes. On Venice it runs under an anonymized privacy tier — your prompts are not stored, profiled, or used for training. There is zero prompt retention.
Yes. Both the text-to-video and image-to-video variants output clips with native audio generation enabled.
Related models
Run Wan 2.7 Enhanced privately.
No prompt logging. No data used for training. Free to start — no credit card.
