Reference to Video (R2V)

Reference to Video.
Same Character.
Every Clip.

R2V turns a reference image or clip into a starring role. Feed the model your character, product, or voice and it keeps them consistent across every shot — no subscription, pay per generation with credits.

Consistent identityAcross every shot
Per-reference voiceLip-synced dialogue
No subscriptionCredit-based pricing
3 R2V enginesOne credit balance
Reference to Video (R2V)
Smooth motion
Rich details
Cinematic lighting
Built for creators, storytellers, and teams who want production-ready results in minutes.

Model Options

Choose the workflow that fits your scene

Start with a prompt, add references when needed, and pick the generation path that matches the output.

MiniMax H3 R2V

Reference-to-video with per-reference voice and lip-synced dialogue. Best when your character has to speak: upload a face reference plus optional voice audio and the model performs the line.

Per-reference voice slots Talking-head dialogue Multi-character scenes 720P-1080P output
Generate with MiniMax H3 R2V

Wan 3.0 R2V

Wan 3.0's reference mode keeps subjects consistent while handling stronger camera moves and action beats. A good fit when the shot needs energy, not just a talking face.

Reference image conditioning Stronger motion handling First-frame control Up to 1080P
Generate with Wan 3.0 R2V

Why use it on Moosky AI

Powerful tools. Seamless experience.

Character consistency

The same face, outfit, and product across every clip in your project — not a new person each render.

Reference-driven audio

Attach voice references and get lip-synced speech that matches the person on screen.

Multi-character scenes

Give each character its own reference slot and keep the cast stable in shared shots.

Three R2V engines

Switch between MiniMax, Wan 3.0, and Seedance R2V from the same credit balance.

No subscription

Pay per generation. R2V is the fastest way to burn a subscription you barely use.

Commercial use

Use reference-driven clips in ads, shorts, and client work.

How It Works

Three simple steps to finished output

1

Upload references

2

Describe the scene

3

Generate consistent clips

Credit-Based Generation

Simple, honest pricing

No subscriptions. No hidden fees. Pay only for the credits you use.

Pay as you goOnly pay when you generate.
No monthly feesNo subscriptions. No commitments.
Credits never expireUse your credits whenever you want.
Commercial useUse outputs for personal or commercial projects.
View credits & pricing

Frequently Asked Questions

What is reference to video (R2V) in AI?

R2V is a generation mode where the model takes one or more reference images (or clips) and conditions the video on them, keeping a character, object, or style consistent from shot to shot. Instead of describing your hero in every prompt, you show the model what they look like once and it carries that identity through the scene.

How is R2V different from image to video?

Image-to-video animates one still frame — the picture itself becomes the first shot. R2V goes further: the references define identity, not composition, so you can place the same character in new scenes, angles, and outfits while they stay recognizably themselves.

Which R2V model should I choose?

Pick MiniMax H3 R2V when your characters speak — it handles per-reference voice and lip-synced dialogue best. Pick Wan 3.0 R2V when the shot is about motion and camera work. Seedance 2.0 R2V sits in between with strong cinematic framing. All three run on the same Moosky credits.

Can R2V keep the same character across multiple clips?

Yes — that is the point. Reuse the same reference image as the seed for every clip in the project and the model preserves the face and styling. For dialogue scenes, add a voice reference so the character sounds consistent too.

How much does R2V cost on Moosky?

R2V generations are billed in credits like everything else on Moosky (100 credits per $1), and the exact cost of each clip is shown in the composer before you run. There is no subscription and credits never expire — you pay for the clips you actually generate.

Disclosure

Moosky AI is an independent platform provider and reseller of AI model access. Product names, model names, and company names shown on this page belong to their respective owners. Moosky AI is not officially affiliated with, sponsored by, or endorsed by those providers unless expressly stated. Page content, examples, and previews may include AI-assisted and synthetic content.