Website profiles · Technology insights · Alternatives

wan3.net Paid content Multilingual

Categories: Artificial Intelligence

Create 2–30 second videos with WAN3 from text, images, or reference media. Choose 720p or 1080p output, with optional generated audio.

Visit website

Updated: 2026-09-23 05:42 Language: English (default) Access: Normal

Profile views 3 Outbound visits 0
WAN3 AI Video Generator Full homepage screenshot
Editorial Review

Website Review

What is WAN3 AI Video Generator?

WAN3 AI Video Generator is a browser-based creative workspace at WAN3 AI Video Generator for turning text, images or reference media into short AI videos and still images. You type a prompt (up to 5,000 characters), pick a workflow and settings such as aspect ratio, clip length and resolution, then spend credits to generate the result. It is an independent studio built around the Wan family of models rather than the model owner's own site, so treat it as a front end and workspace, not the source of the underlying model.

What you can do there

  • Text to video — describe a scene and generate a clip.
  • Image to video — animate a still you supply.
  • Reference to video — use existing footage or media to guide motion or style.
  • Video editing and motion transfer — rework or re-drive existing clips.
  • Image generation — create stills for campaigns, products or characters.

The page frames output as roughly 2–30 second videos at 720p or 1080p, with optional generated audio, and shows preset "featured ideas" such as product reveals, fashion street style and cinematic scenes for inspiration.

Who it suits

User Fit Trade-off
Marketer testing product or campaign visuals Fast prompt-to-clip loop, image-to-video for existing product shots Credit spend per attempt; commercial licence not included on lower tiers
Social creator publishing weekly Standard plan credits cover regular short-form output Credits refresh monthly and do not roll over
Agency or client-work studio Pro tier targets campaigns and advanced model access Highest monthly cost; still credit-metered

The credit model matters most

Generation is metered in credits, and the plan tiers differ mainly in monthly credits and which video models you can reach — the entry tier is limited to earlier video models, while higher tiers unlock Wan 3.0. Credits expire monthly rather than accumulating, so estimate your real output before choosing: a handful of five-second clips per month needs far less than weekly publishing or client revisions.

Next step: open the studio, run one short text-to-video test at the lowest settings to gauge output quality and credit burn, then compare that against your actual monthly clip count before subscribing. If you also want to see how the model family is presented by its originator, check Wan for reference.

How do WAN3 credits work, and which subscription plan is best for my video volume?

Credits are the currency you spend each time you generate on WAN3. You pick a workflow (text to video, image to video, reference to video, motion transfer, editing or image generation), set duration and resolution, and the interface shows an estimated credit cost before you commit. The page's own plan descriptions give you a rough conversion: about 7 five-second Wan 2.7 videos at 720p per 1,200 credits, and about 14 five-second Wan 3.0 videos at 720p per 2,500 credits.

Two rules matter more than the headline numbers. First, subscription credits refresh monthly and do not roll over, so unused credits are lost rather than saved for a heavy month. Second, higher tiers lower the effective cost per credit (roughly $24.92 per 1,000 on Basic versus $19.96 per 1,000 on Standard, based on the listed monthly prices). Annual subscriptions and one-time credit packs also exist, which helps if your workload is spiky rather than steady.

Note the tier gap that affects model access, not just volume: Basic lists Wan 2.7 and earlier video models, while Standard and above list Wan 3.0. If your project depends on the newest video model, the choice is effectively Standard or higher.

Matching plan to volume

Your situation Sensible starting point Why
Testing prompts, occasional short clips Basic Cheapest entry; enough credits for a handful of 720p videos
Publishing new video weekly Standard More credits, lower per-credit cost, Wan 3.0 access
Client work, campaigns, longer or higher-res output Pro Largest credit pool; the listed example is about 35 five-second Wan 3.0 videos at 720p

Treat those video counts as planning estimates, not guarantees: cost scales with duration and resolution, so 1080p or 10–30 second outputs will consume credits faster than the five-second, 720p examples. Also check the commercial license terms for your tier before using output for paid work, since the Basic and Standard descriptions shown do not include it.

A practical next step: list your realistic monthly output (how many clips, typical length, 720p or 1080p, which model), multiply by the per-generation estimates shown in the interface, and compare that total against each tier. If you are within about 20% of a tier's credits, take the next tier up rather than buying top-up packs mid-month. WAN3 is an independent workspace and does not own the Wan model, so confirm current model availability and terms on WAN3 before committing to an annual plan.

Is WAN3 the official Alibaba Wan website, and can I use its videos commercially?

No, WAN3 is not Alibaba's official Wan website. It describes itself as an independent image and video workspace, and its own FAQ says it is not the official Alibaba Wan site. It also states that it does not own the Wan model — it's a third-party front end that gives you access to Wan 3.0, Wan 2.7 and earlier video models, plus image-generation models.

That distinction matters for two reasons: you're relying on a third party for access, uptime and billing, and the commercial rights you get depend on that third party's terms, not just on Alibaba's model licence.

Commercial use

Based on the plan information shown on the site, commercial use is not included on the lower tiers. The Basic and Standard monthly plans both list "Commercial license not included," while the Pro plan is positioned for client work, campaigns and advanced model access. So if you need to publish, monetise or hand footage to a client, the practical answer is that you should not assume the cheaper plans cover you — check the Pro plan terms and any separate licence wording before you build a paid deliverable on top of it.

Also note that credits refresh monthly and do not roll over on subscriptions, so unused credits are lost rather than banked.

What to check before you commit

  • The exact licence text for the plan you're buying, and whether it covers client work, ads, resale or redistribution.
  • Whether commercial rights attach to the output itself or only to your account tier.
  • What happens to rights if you downgrade or stop paying.
  • Whether the underlying model provider imposes additional restrictions that flow through to you.

Practical next step

If your use is personal, experimental or internal, a lower tier is a reasonable place to test prompts, reference-to-video and image-to-video workflows. If your use is client-facing, price the Pro tier into the job and get the licence terms in writing before you deliver anything.

For comparison, the official Wan model materials and terms are published by Alibaba, and other hosted access options exist — for example Replicate and fal.ai host Wan-family models with their own licensing and commercial-use terms, which are worth reading alongside WAN3's.

How do I write effective Wan 3.0 video prompts for text-to-video generation?

Write prompts the way you'd brief a camera operator: subject, action, setting, camera behaviour, lighting and mood, in that order. Wan 3.0 text-to-video rewards concrete, physical description over abstract adjectives, and the workspace on WAN3 AI Video Generator gives you a 5,000-character prompt field plus visible controls for aspect ratio, duration (5s in the default view) and resolution, so the prompt only has to carry what the settings don't.

A workable prompt skeleton

  • Subject: who or what, with one or two distinguishing details ("a weathered fisherman in a yellow oilskin").
  • Action: one clear motion per shot ("hauls a net hand over hand"), not a sequence of events.
  • Setting: place, time of day, weather ("harbour wall at dawn, low mist").
  • Camera: shot size and movement ("medium shot, slow push in", "handheld, slight drift").
  • Light and mood: direction and quality of light, colour feel ("cold blue pre-dawn light, muted palette").

Example: "Medium shot of a weathered fisherman in a yellow oilskin hauling a net hand over hand on a harbour wall at dawn, low mist, slow push in, cold blue pre-dawn light, muted palette, 16:9."

Practical habits that matter more than prompt length

  1. Change one variable at a time. If the camera move is wrong, fix the camera clause before rewriting the whole prompt.
  2. Front-load the subject. Early words tend to dominate; burying the subject behind style language invites drift.
  3. Name one camera instruction only. "Slow push in" plus "orbit" plus "crane up" usually produces mush.
  4. Use motion verbs the model can render: walking, turning, pouring, drifting, unfolding. Avoid internal states ("she realises", "he regrets").
  5. Keep clips short when testing. A 5-second generation costs a fraction of a longer one and tells you whether the concept works before you spend on duration and resolution.
  6. Reuse what works. Save winning prompts as templates and swap only the subject and setting.

Common failure modes and the fix

Symptom Likely cause Prompt fix
Subject changes appearance mid-clip Too many competing descriptors Cut to two or three fixed traits, repeat them consistently
Camera feels static or chaotic No move, or several conflicting moves State exactly one move and its speed
Scene reads as generic stock Abstract mood words only Replace "beautiful, amazing" with physical detail and named light
Action never completes Multi-beat narrative in one clip Reduce to a single continuous action

When to switch workflow instead

Text-to-video is the wrong tool when the exact look of a person, product or location matters. WAN3 also lists image-to-video and reference-to-video among its toolkit options, and those let you anchor appearance with a still or reference clip while the prompt handles motion, camera and light. A practical rule: if you'd be annoyed by the subject changing, start from an image; if you're exploring a mood or camera idea, start from text.

For a concrete next step, write your first prompt in the five-part order above, generate at 5 seconds and the lowest resolution that still reads clearly, then revise only the clause that failed. Keep a short log of prompt, settings and result — after five or six iterations you'll have a personal template far more useful than any generic prompt list.

What is the difference between image-to-video and reference-to-video in WAN3?

Image-to-video starts from a single still and animates it. Reference-to-video uses one or more reference images to guide the look of a scene you describe, rather than animating the reference itself.

How they differ in practice

Image to video Reference to video
What you provide One image as the first frame Reference image(s) plus a prompt
What the model does Extends that frame into motion Carries the reference's style, subject or look into a new scene
Best for Making a photo, product shot or illustration move Keeping a character, product or visual style consistent across shots
Main trade-off Motion is constrained by what is already in the frame More freedom in composition, but less frame-level control

Choosing between them

Use image-to-video when the first frame matters most — a product photo that should rotate, a portrait that should blink and turn, a still landscape you want to bring to life. The result usually stays close to your source image.

Use reference-to-video when you care about consistency more than a specific frame. A reader scenario: an illustrator with a fixed character design can supply that design as a reference and prompt new scenes — the character stays recognisable while the setting changes. The same applies to brand work, where a product's shape and colour need to survive multiple shots.

Both appear in the site's toolkit alongside text-to-video, video editing and motion transfer, so you can combine approaches — for example, generate a scene with a reference, then animate a chosen frame with image-to-video.

Next step: open the create panel, pick Reference to video, and run one short 5-second test at 720p before committing credits to a longer clip. Compare it against the same prompt run through Image to video using your strongest still, and judge which gives you more usable control for your project.

For the model's own workflow options, see WAN3 AI Video Generator.

How can I create a 2–30 second video with generated audio at 720p or 1080p?

Use the site's Text to video, Image to video, or Reference to video workflow, pick the Wan 3.0 model, set your duration between 2 and 30 seconds, and select 720p or 1080p. If generated audio is offered for your chosen model and duration, enable it in the same settings panel before you create.

The page shows a single creation screen with a prompt field (0/5000 characters), a workflow selector, a model selector, and output settings such as aspect ratio, duration, resolution, and quantity. Credit cost is estimated next to the settings, so you can check the cost before committing.

H3 Practical steps

  1. Choose a starting point: text, an uploaded image, or reference media.
  2. Select Wan 3.0 as the model.
  3. Set duration (2–30s), resolution (720p or 1080p), and aspect ratio.
  4. Turn on audio generation if available for that combination.
  5. Write a prompt describing scene, subject, camera movement, lighting, and any sound you want.
  6. Review the estimated credit cost, then create.

H3 What to watch for

  • Longer durations and 1080p usually cost more credits than short 720p clips.
  • Audio support can vary by model, so confirm it is available for Wan 3.0 at your chosen length.
  • The site is an independent workspace, not the official Alibaba Wan site, and it does not own the Wan model.

For a first test, make a 5-second 720p clip with audio to confirm the workflow and cost, then scale up to your target length and resolution.

Related questions

More questions →
How to Create Your First Wan 3.0 Video: Text, Image, or Video Input

Wan 3.0 is a browser-based AI video generator that turns text, a photo, or an existing clip into a video of up to 1080p and 30 seconds. To make your first clip, pick the starting mode that matches the material you already have, write a prompt that names the subject, style, lighting, mood, and composition, then set resolution, duration, and audio before you generate. Expect a credit cost somewhere in the 98–5880 range, since the price scales with length and resolution.

Pick your starting mode

The generator offers four entry points. Choose based on what you have on hand, not on which sounds most advanced.

Mode Use it when What you provide
Text to video You have an idea but no footage or image A written prompt
Image to video You have a still you want to animate A photo plus a prompt for motion
Video to video You have a clip you want restyled or transformed A reference video plus direction
Video editing You want to adjust an existing clip Your video plus edit instructions

If you are starting from nothing, text to video is the fastest path to a finished result. If you already have a strong image, image to video usually gives you more control over the opening frame, because the model starts from your composition rather than inventing one.

Write a prompt that actually guides the model

The prompt field accepts up to 2000 characters, and the site's own tip is to be detailed and specific. A useful prompt covers five things:

  • Subject — who or what is on screen
  • Style — realistic, animated, cinematic, and so on
  • Lighting — soft daylight, neon, backlit, overcast
  • Mood — calm, tense, playful
  • Composition — wide shot, close-up, centered subject

For example, instead of "a dog running," try "a golden retriever sprinting across a wet beach at sunrise, low camera angle, warm backlight, energetic mood, shallow depth of field."

If your idea is thin, use the Expand option next to the prompt field. It fills out a short idea into a fuller description, which is useful when you know the vibe but not the wording.

Set resolution, duration, and audio

Before generating, you choose:

  • Resolution — 480p, 720p, or 1080p
  • Duration — anywhere from 2 to 30 seconds
  • Audio — whether to generate sound with the video
  • Thinking mode — an option available in the generator panel

Higher resolution and longer clips both push the credit cost up. The full range runs from 98 to 5880 credits, so a short 480p test costs far less than a 30-second 1080p render. The generate button shows the credit price for your current settings (the example on the page reads 455 credits), so you can check the number before committing.

A practical approach: render a short, low-resolution version first to confirm the motion and framing, then re-run at higher settings once you like the direction.

Check your credits before you generate

Because cost scales with length and resolution, the same prompt can be cheap or expensive depending on settings. The generator displays the credit total on the generate button, so read it before you click. If you are experimenting with prompt wording, keep duration short and resolution low until the content is right — then spend the credits on a final high-quality pass.

The site lists a pricing page, so credit costs and any current offers are worth checking there rather than assuming a fixed rate.

Review and iterate

After the clip renders, judge it against what you asked for. If the motion is wrong, adjust the action wording in the prompt. If the framing is off, change the composition terms or switch to image to video and supply a starting frame that already has the composition you want. If the style is wrong, name the style more explicitly.

Switching input mode is often faster than rewriting a prompt repeatedly. A still image locks in the look; a reference video locks in the motion. Use whichever constraint your problem actually is.

What to know about the service

This is an independent, browser-based AI creation service. Its footer states that references to Wan 3.0 describe the creative product on the site and do not imply ownership, partnership, or endorsement by a model provider. All gallery videos and images are AI-generated synthetic content and do not depict real people or events unless explicitly stated.

How to Turn an Image into a Video with Wan 3.0

Wan 3.0 on wan3.dev can animate a still image into a clip of up to 1080p and 30 seconds. Switch the generator to Image to video, upload your photo, then write a prompt that describes the motion and camera move rather than re-describing the picture. Expect a single run to cost hundreds of credits — the site's own example shows a 455-credit generate button — so decide your duration and resolution before you click.

Before you start

  • Source image: a clear, well-lit still. The model animates what it sees, so a blurry or cluttered photo gives it little to work with.
  • Prompt: up to 2000 characters. The site's own tip is to describe subject, style, lighting, mood, and composition — for image-to-video, add motion and camera direction on top of that.
  • Credits: check your balance before generating. The interface shows "Available Credits" next to the generate button, and the example run is priced at 455 credits.
  • Settings you'll choose: resolution (480p / 720p / 1080p), aspect ratio, duration (2–30s), and whether to generate audio.

Step-by-step: image to video

1. Open the generator and pick Image to video

The site lists four starting points: Text to video, Image to video, Video to video, and Video editing. Choose Image to video so the generator treats your upload as the first frame rather than as a loose reference.

2. Upload your image

Select the photo you want to animate. This becomes the visual anchor — the subject, framing, and lighting of the output should stay recognisable to it.

3. Write a motion prompt

This is where most first attempts go wrong. Don't describe what's already in the photo; describe what should change.

Weak prompt: "A woman standing in a field."

Useful prompt: "The woman turns her head slowly toward the camera, hair moving in a light breeze; slow push-in, warm late-afternoon light, calm mood."

Cover four things:

  • Subject motion — what moves, and how fast
  • Camera move — push in, pull out, pan, orbit, or hold still
  • Lighting and mood — keeps the clip consistent with the source image
  • Style — only if you want the output to shift away from the photo's look

4. Set resolution, ratio, and duration

Duration runs from 2 to 30 seconds and resolution from 480p to 1080p. Longer and higher-resolution clips are the expensive end of that range, so a first test at a short duration and lower resolution is the cheaper way to check whether your prompt produces the motion you want.

5. Decide on audio

The generator has a Generate audio toggle. If you don't need sound, leaving it off keeps the run simpler; if you do, the site advertises audio generation from text descriptions.

6. Check the credit cost, then generate

The generate button displays its price before you commit — the on-page example is 455 credits. Credit cost scales with the settings you chose, so if the number is higher than you expected, shorten the clip or drop the resolution and check again.

7. Review and adjust

Watch the result against your source image. Two common failure modes:

What you see What to change
Motion is wrong or the subject does something unexpected Rewrite the motion sentence to name one clear action instead of several
The clip drifts away from the original image Remove style words from the prompt and lean on lighting/mood language that matches the photo
Too little happens Add an explicit camera move, or lengthen the duration slightly

What it costs and what you get

  • Output: up to 1080p, up to 30 seconds, from text, image, or video input.
  • Model options: the generator lists Wan 3.0 with a 2–30s range at 480p/720p/1080p and a credit band of 98–5880 credits depending on settings.
  • Pricing page: wan3.dev/pricing is linked from the main navigation and currently shows a 50% OFF label. Check it for current credit packs rather than assuming a price.

Practical notes

  • Test cheap, finish expensive. Run a short, low-resolution version to validate the prompt, then re-run at your target settings once the motion looks right.
  • One action per clip. A single clear movement reads better than a list of instructions, and it's easier to debug when the output is wrong.
  • The site is independent. Its footer states that references to Wan 3.0 describe the creative product on the website and don't imply ownership, partnership, or endorsement by a model provider. Treat it as a browser-based creation service, not the model's official home.
  • Gallery as reference. The workflow showcase on the homepage demonstrates subject action, camera direction, image-led motion, and visual transformation — useful for calibrating how much motion a prompt should ask for.
How to Write Wan 3.0 Video Prompts That Get the Clip You Want

Wan 3.0's prompt box accepts up to 2,000 characters and its own tip tells you to describe the subject, style, lighting, mood, and composition. That is the whole method: name what is on screen, say what it does, say how the camera behaves, then set the light and the look. Prompts written this way work for all three starting points — text, a photo, or a reference clip — and you can adjust one element at a time when a result misses instead of rewriting from scratch.

The prompt pattern

Build every prompt in this order. Later elements matter less if the earlier ones are vague, so fix the top of the list first.

  1. Subject — who or what is on screen, with enough detail to be specific ("a woman in a yellow raincoat," not "a person").
  2. Action — the single main thing that happens during the clip.
  3. Camera — how the shot moves or holds (static, slow push in, tracking alongside).
  4. Lighting — time of day and light quality (overcast morning, warm lamp light, hard midday sun).
  5. Mood and style — the feeling and the visual treatment (documentary, soft and quiet, high contrast).
  6. Composition — framing and where the subject sits in frame (wide shot, centered, subject on the left third).

A working example:

A woman in a yellow raincoat walks toward the camera along a wet pier, slow push in, overcast morning light, calm and slightly melancholic, wide shot with the subject centered, muted color palette.

Compare that to "a woman walking on a pier, cinematic, 4K, beautiful." The second prompt gives the model nothing to act on — no direction of movement, no camera instruction, no light. Vague quality adjectives are the most common reason a clip comes back looking like a generic stock shot.

Adapting the pattern to each input type

The same six elements apply, but what you should spend characters on changes with how you start.

Start What the model already has What your prompt should carry most
Text to video Nothing All six elements — this is the only case where you describe the subject from zero
Image to video Subject, framing, and look from the photo Action, camera direction, and how much the scene should change
Video to video Existing motion and structure What to keep, what to transform, and the target style

For image-to-video, don't re-describe the photo. If the image already shows a person at a desk, write the motion: "she turns her head toward the window, camera slowly drifts right, late afternoon light." Describing the subject again wastes characters and can push the model away from the source image.

For video-to-video, be explicit about the transformation. "Keep the camera movement and timing; change the setting to a snowy street at night, cool blue tones" tells the model what is fixed and what is not.

How much to write

The box holds 2,000 characters, but length is not the goal — coverage is. A prompt that names all six elements usually lands between 40 and 90 words. Below that, you are probably missing camera or lighting. Far above it, you risk contradicting yourself, which is worse than saying less.

Two habits that keep prompts tight:

  • One main action per clip. If you want a character to walk in, sit down, and open a laptop, that is three beats competing for a clip that may run as short as a few seconds. Pick the beat you actually need.
  • Cut adjectives that don't change pixels. "Stunning," "amazing," and "masterpiece" don't tell the model what to render. "Backlit," "shallow depth of field," and "handheld" do.

Matching prompt detail to your settings

Your prompt and your settings have to agree, or the result will feel wrong even when the model did what you asked.

  • Duration. A prompt describing a slow, multi-stage action needs room to play out. On a short clip, describe a single moment instead of a sequence.
  • Resolution. Fine detail you ask for — text on a sign, small objects, intricate patterns — is more likely to survive at higher resolution. At lower settings, favor larger shapes and simpler compositions.
  • Audio. If you enable audio, say what should be heard. Ambient sound, dialogue, or music are different requests, and silence is a valid one to state.
  • Aspect ratio. Write the composition to match the frame. A "wide establishing shot" fights a vertical frame; "centered subject with headroom" suits it.

When the result is wrong, change one thing

Rewriting the whole prompt after a bad clip makes it impossible to know what fixed it. Instead, diagnose by category:

  • Wrong subject or wrong action → the top of your prompt is too vague. Add one concrete detail.
  • Right subject, wrong feeling → lighting and mood are underspecified. Name the light source and time of day.
  • Static or awkward motion → you described a scene but not a camera. Add an explicit camera instruction.
  • Composition is off → state the framing and where the subject sits in frame.
  • It ignored part of the prompt → you likely have two competing actions. Cut to one.

Change one element, regenerate, and compare. Two or three passes usually get you to a prompt worth saving and reusing with a different subject.

A reusable template

Fill this in and you have a complete prompt:

[Subject with one distinguishing detail] [does one action], [camera instruction], [lighting and time of day], [mood], [framing and subject placement], [style or color treatment].

Keep your best prompts in a note. Because the pattern is modular, you can swap the subject and keep the camera, lighting, and style — which is the fastest way to get a consistent look across several clips.

How Wan 3.0 Video Credits Work: What Determines the Cost of a Clip

Wan 3.0 charges credits per generation, and the price of a clip is quoted on the Generate button before you confirm. On the site's generator page, the model selection line lists a range of 98–5880 credits for Wan 3.0, alongside settings of 2–30 seconds and 480p/720p/1080p. Your live "Available Credits" balance sits next to that quoted cost, so you can compare the two before spending anything. The practical rule: the longer the clip and the higher the resolution, the closer you land to the top of that range.

Where the credit cost comes from

Credits are the in-app currency for this independent browser-based service. You don't buy a fixed number of videos — you spend credits per generation, and the amount depends on the settings you choose at that moment.

The generator page shows three things side by side:

Element What it tells you
Available Credits Your current balance
Generate (455 credits) The cost of the clip with your current settings
Model Selection line The full settings range: 2–30s, 480p/720p/1080p, 98–5880 credits

The 455-credit figure in the page evidence is the quoted cost for one specific configuration — not a fixed price for every clip. Change the duration or resolution and that number moves.

What pushes a clip toward 5880 credits

Two settings dominate the cost:

  • Duration (2–30 seconds). A 30-second clip is roughly fifteen times the length of a 2-second one, and the top of the credit range reflects that.
  • Resolution (480p/720p/1080p). Higher resolution means more frames to render at greater detail, so 1080p sits well above 480p.

Optional features can also change the total:

  • Generate audio — the text-to-video panel includes an audio toggle, and the site describes generating "videos with audio."
  • Thinking mode — listed as an available option on the generator panel.

Because these are optional, the safest habit is to read the credit figure on the Generate button after you've set duration, resolution, audio, and thinking mode — not before.

A workflow that keeps credit spend predictable

  1. Write your prompt first. The prompt field allows up to 2000 characters, and the site advises describing subject, style, lighting, mood, and composition. A vague prompt that produces the wrong clip costs you credits twice.
  2. Set duration and resolution before anything else. These two drive most of the cost.
  3. Decide on audio and thinking mode. Toggle them on or off and watch the quoted figure update.
  4. Read the Generate button. That number is your actual cost for this render.
  5. Compare against Available Credits. If the balance is short, the Pricing page is the top-up route.
  6. Test cheap, then commit. Draft an idea at a short duration and lower resolution, confirm the motion and framing work, then re-run at 1080p or a longer length.

Step 6 is the single biggest lever on total spend. A 1080p, 30-second render with audio is the most expensive configuration the page describes; the same idea validated first at low settings costs a fraction of that.

Common questions about the credit range

Why does the site list a range instead of one price? Because cost is a function of your settings, not a flat fee. 98 credits represents the cheapest combination the page supports; 5880 represents the most expensive.

Is the 455-credit figure the standard price? No. It's the quote shown for one configuration on the page at the time it was captured. Your own quote depends on your settings.

Do failed or abandoned generations cost credits? The available page evidence doesn't state a refund or failure policy, so treat the quoted cost as what you commit when you press Generate.

Does the balance carry over or expire? Not addressed in the available material.

Is there a free tier? The page shows an Available Credits balance and a Pricing page with a 50% OFF signal, but the evidence doesn't state whether any credits are granted free or whether login is required to generate. Don't assume either.

What to check before you generate

  • The Generate button figure — this is the only number that matters for your next clip.
  • Duration and resolution — the two settings that move cost the most.
  • Audio and thinking mode toggles — optional add-ons that can raise the total.
  • Available Credits — if it's below the quoted cost, top up via Pricing first.

If you're new to the tool, start with the shortest duration and lowest resolution that still tests your idea. Once the motion, camera direction, and subject look right, re-run at the quality you actually need. That sequence keeps you near the bottom of the 98–5880 range while you iterate, and reserves the expensive end for clips you already know work.

How to Use Reference to Video in Wan 3.0

Reference to video in Wan 3.0 means starting from a clip instead of text or a still image: you upload a reference video, then describe what the new clip should inherit (subject action, camera movement, style) and what it should change (scene, wardrobe, lighting). On wan3.dev this sits under the Video to video start option, alongside Text to video, Image to video, and Video editing. It fits you best when you already have footage whose motion or look you want to reuse but whose content you want to replace. If you only have a written idea or a single photo, text-to-video or image-to-video is the cheaper, simpler path.

What the reference actually controls

Wan 3.0's generator lists three ways to start — Text / Image / Video — and the site describes the workflow showcase as covering "subject action, camera direction, image-led motion, and visual transformation." In practice, a reference clip is your strongest lever on:

  • Motion — how the subject moves and how fast the action reads
  • Camera language — pans, pushes, handheld feel, framing changes
  • Style and grade — color, contrast, texture, overall look
  • Transformation direction — what the clip turns into over its duration

What it does not reliably lock down is identity and exact detail. A clean reference gives the model a clear motion signal; a busy or low-quality one gives it noise, and the output drifts.

Preparing the reference clip

The site's stated output ceiling is up to 1080p and up to 30 seconds, so treat that as the target range for what you feed in.

Reference quality Effect on the result
Single clear subject, simple background Motion and camera transfer most faithfully
Multiple subjects or fast cuts Model averages the motion; output looks muddled
Heavy motion blur or low resolution Weak motion signal, softer output
Long clip with several distinct shots Only part of the motion is reflected; trim to one shot

Practical prep: trim to one continuous shot, keep the subject large enough in frame to read, and avoid clips where the camera and the subject both move hard at once. If your reference is longer than the clip you want, cut it to roughly the length you plan to generate.

Writing the prompt: separate "keep" from "change"

The prompt field on the generator accepts up to 2000 characters, with a tip to describe subject, style, lighting, mood, and composition. For reference-to-video, split your prompt into two explicit halves:

  1. Inherit — name the motion and camera you want carried over: "keep the slow push-in and the walking pace of the reference."
  2. Replace — name everything else: new setting, wardrobe, time of day, color palette.

Example prompt structure:

Keep the reference's forward tracking shot and the subject's steady walking rhythm. Change the setting to a rain-soaked night market, swap the outfit to a red jacket, shift the grade to cool blue with warm practical lights, keep the framing centered.

Vague prompts like "make it like the reference but different" give the model nothing to hold onto; the output tends to copy the reference too literally or ignore it entirely.

Settings and cost before you generate

The generator exposes model selection (Wan 3.0), resolution ratio, duration, a generate-audio toggle, and a thinking mode. The listed range is 2–30 seconds at 480p / 720p / 1080p, costing 98–5880 credits depending on those choices — so both length and resolution move the price, and the low end of that range is a short low-resolution clip while the high end is a long 1080p one. The interface shows a live credit figure next to the Generate button (the page displays an example of 455 credits for one configuration), so set duration and resolution first and read the number before committing.

For a first reference-to-video attempt, generate short and at a lower resolution to confirm the motion transfers the way you expect, then re-run at 1080p once the prompt is right. That keeps failed attempts cheap. Note that pricing and any discounts are shown on the site's own Pricing page — check there rather than assuming a rate.

When the output drifts from your reference

Most failures fall into three buckets:

  • Motion ignored — the clip looks static or generic. Your reference likely had weak or ambiguous motion. Swap in a clip with one obvious, sustained movement.
  • Reference copied too literally — you get the same scene with minor changes. Your "replace" instructions were too soft. State the new scene, wardrobe, and lighting as explicit changes.
  • Muddy result — usually a busy reference or too many simultaneous changes. Simplify to one inherited motion plus one or two replacements, then build up.

Change one variable per retry: fix the reference first, then the prompt, then the settings. Changing all three at once tells you nothing about which one caused the improvement.

A quick decision check

Use reference-to-video when you have a clip whose movement or look is the point and you want new content in that mold. Use image-to-video when a single frame defines the shot and you only need it to move. Use text-to-video when you have no source material and want the model to invent the motion. And if your goal is to rework an existing clip rather than generate a new one from it, the Video editing option is the closer match.

Website Overview

Page metadata, canonical configuration and social previews work together to provide more consistent search and sharing presentation. An active inbound-mail setup with incomplete authentication may leave the domain more open to impersonation. Provider hosting alone does not close that gap.

Domain and Registration

The domain was registered less than a year ago and has limited historical evidence to assess. Transfer-protection status is present, helping reduce the risk of unauthorized domain transfers. The registrar is Cloudflare, Inc., a widely used domain service provider. The domain uses the common .net extension, which is not an independent safety signal.

DNS and Email

The observed email authentication setup is incomplete: DMARC is missing. Nameservers are provided by Cloudflare, indicating managed DNS hosting. MX records point to the Cloudflare Email Routing email service. No CNAME was found; the observed records resolve directly to addresses. TXT records include verification markers for Google. Such markers may also remain after a service stops being used.

TLS and Certificates

The certificate uses an RSA 2048-bit public key, offering broad client compatibility. The server supplied a complete certificate chain. No organization name is present in the certificate; the available fields are consistent with domain validation. The certificate was issued by Let's Encrypt, commonly associated with automated certificate services. The certificate's total validity is about 89 days, consistent with a short renewal cycle.

HTTP and Browser Security

CORS permits any origin to read this response. This is common for public resources; sensitive responses need narrower handling. No X-Powered-By header was found, reducing one common source of backend fingerprinting information. All six checked browser-security headers are present. Their effectiveness still depends on the policy values and application behavior. No obvious internal addresses or debug information were found in the headers. The Server header contains the custom value Vercel.

Technology Stack Analysis

The public page identifies Next.js, Google Analytics, Vercel without precise versions, leaving fewer clues for version-specific scanning.

Search and Social Sharing

Twitter Card metadata is configured. JSON-LD includes Organization data, helping describe the organization as an entity. The page declares 13 language or regional alternatives using hreflang. The title has 58 characters, within a common display range. A meta description is present, with 134 characters.

Hosting and Email

DNSCloudflare
HostingVercel
EmailCloudflare Email Routing
Location United States flagWalnut, California, United States 76.76.21.21

User reviews (0)

  • No reviews yet.

Pages, Search and Sharing

Meta descriptionCreate 2–30 second videos with WAN3 from text, images, or reference media. Choose 720p or 1080p output, with optional generated audio.
Canonical URLhttps://wan3.net
LanguageEnglish (default) · Multilingual
Twitter Cardsummary_large_image
adsbot-google 1 allowed · 24 disallowed
  • Allow/
  • Disallow/api/
  • Disallow/admin/
  • Disallow/zh/api/
  • Disallow/zh/admin/
  • Disallow/ja/api/
  • Disallow/ja/admin/
  • Disallow/ko/api/
  • Disallow/ko/admin/
  • Disallow/es/api/
  • Disallow/es/admin/
  • Disallow/de/api/
  • Disallow/de/admin/
  • Disallow/fr/api/
  • Disallow/fr/admin/
  • Disallow/pt/api/
  • Disallow/pt/admin/
  • Disallow/it/api/
  • Disallow/it/admin/
  • Disallow/ru/api/
  • Disallow/ru/admin/
  • Disallow/ar/api/
  • Disallow/ar/admin/
  • Disallow/th/api/
  • Disallow/th/admin/
adsbot-google-mobile 1 allowed · 24 disallowed
  • Allow/
  • Disallow/api/
  • Disallow/admin/
  • Disallow/zh/api/
  • Disallow/zh/admin/
  • Disallow/ja/api/
  • Disallow/ja/admin/
  • Disallow/ko/api/
  • Disallow/ko/admin/
  • Disallow/es/api/
  • Disallow/es/admin/
  • Disallow/de/api/
  • Disallow/de/admin/
  • Disallow/fr/api/
  • Disallow/fr/admin/
  • Disallow/pt/api/
  • Disallow/pt/admin/
  • Disallow/it/api/
  • Disallow/it/admin/
  • Disallow/ru/api/
  • Disallow/ru/admin/
  • Disallow/ar/api/
  • Disallow/ar/admin/
  • Disallow/th/api/
  • Disallow/th/admin/
All bots 2 allowed · 24 disallowed
  • Allow/
  • Allow/_next/
  • Disallow/api/
  • Disallow/admin/
  • Disallow/zh/api/
  • Disallow/zh/admin/
  • Disallow/ja/api/
  • Disallow/ja/admin/
  • Disallow/ko/api/
  • Disallow/ko/admin/
  • Disallow/es/api/
  • Disallow/es/admin/
  • Disallow/de/api/
  • Disallow/de/admin/
  • Disallow/fr/api/
  • Disallow/fr/admin/
  • Disallow/pt/api/
  • Disallow/pt/admin/
  • Disallow/it/api/
  • Disallow/it/admin/
  • Disallow/ru/api/
  • Disallow/ru/admin/
  • Disallow/ar/api/
  • Disallow/ar/admin/
  • Disallow/th/api/
  • Disallow/th/admin/

Registration details RDAP / WHOIS

RegistrarCloudflare, Inc.
Registered2026-08-08
Expires2027-08-08
Domain statusclient transfer prohibited
Nameserversnucum.ns.cloudflare.com、wesley.ns.cloudflare.com
DNSSECunsigned

DNS records

TypeNameValueTTLPriority
Awan3.net76.76.21.21300
MXwan3.netroute2.mx.cloudflare.net30027
MXwan3.netroute1.mx.cloudflare.net30051
MXwan3.netroute3.mx.cloudflare.net30080
NSwan3.netnucum.ns.cloudflare.com86400
NSwan3.netwesley.ns.cloudflare.com86400
TXTwan3.netgoogle-site-verification=LrfSrnaD2D-1soAYt8YICp3QdgqP0bIFXPOpngTyn8Q300
TXTwan3.netv=spf1 include:_spf.mx.cloudflare.net ~all300

TLS and certificates

AssessmentNormal configuration
Supported protocolsTLSv1.2、TLSv1.3
Negotiated protocolTLSv1.3
Certificate subjectwan3.net
IssuerLet's Encrypt
Valid until2026-11-10T09:59 · Remaining when checked: 48 days
Verification detailsCertificate trust: Passed · Hostname match: Passed

HTTP response headers

HeaderValue
content-typetext/html; charset=utf-8
cache-controlpublic, max-age=0, must-revalidate
serverVercel
strict-transport-securitymax-age=63072000; includeSubDomains; preload
content-security-policydefault-src 'self'; script-src 'self' 'unsafe-inline' https://www.googletagmanager.com https://www.google-analytics.com https://www.googleadservices.com https://googleads.g.doubleclick.net https://bat.bing.com https://bat.bing.net https://*.clarity.ms https://c.bing.com https://sa.douni.one https://challenges.cloudflare.com https://accounts.google.com/gsi/client; worker-src 'self' blob:; style-src 'self' 'unsafe-inline' https://accounts.google.com/gsi/style; img-src 'self' data: blob: https:; media-src 'self' data: blob: https:; font-src 'self' data:; connect-src 'self' https://www.google-analytics.com https://*.google-analytics.com https://www.googletagmanager.com https://google.com https://www.google.com https://www.googleadservices.com https://pagead2.googlesyndication.com https://*.doubleclick.net https://bat.bing.com https://bat.bing.net https://*.clarity.ms https://c.bing.com https://sa.douni.one https://*.ingest.sentry.io https://*.ingest.us.sentry.io https://*.vercel-insights.com https://challenges.cloudflare.com https://*.r2.cloudflarestorage.com https://accounts.google.com/gsi/; frame-src 'self' https://challenges.cloudflare.com https://accounts.google.com/gsi/; frame-ancestors 'none'; object-src 'none'; manifest-src 'self'; base-uri 'self'; form-action 'self'; upgrade-insecure-requests
x-frame-optionsSAMEORIGIN
x-content-type-optionsnosniff
referrer-policystrict-origin-when-cross-origin
permissions-policycamera=(), microphone=(), geolocation=()
access-control-allow-origin*
set-cookieRedacted

Identified technologies

Next.jsGoogle AnalyticsVercel

Recent Updates

  • Screenshots