Your image is ready. You know how the scene should develop after that frame. But you face another list of models: which one accepts this material, where is the right mode, and why does one name have several versions? We changed the order of that decision and moved the essentials to the top of the Video node.

Before: start with a name

Choose a familiar model, open its settings, then find out whether it can do the job.

Now: start with an action

Name the task, see the available models for it, then configure the next run.

“What do you want to do?” is a better first question

The selector has six roles: six answers to “what do you want to do?” There are more ways to use them: one role can define just the starting frame or both the opening and ending of a video. Pick a task below to jump to its walkthrough.

Start with the task. All six roles in the actual Vini Studio selector.
Start with the task. All six roles in the actual Vini Studio selector.

The interface screenshots were captured in the local studio using a sample project. These are working menus and a real node, not a drawn mockup.

Role: the task. Model: what performs it. Settings: how the next run works. “Video from a starting frame” selects image-to-video: the image defines the starting point rather than limiting the clip to animating that picture. And video-to-video alone does not tell you whether you want an edit, an extension, or motion transfer. That is why the selector names tasks, not just file types.
Video from a starting frameThe image sets the opening; the prompt defines what happens next
Create from referencesUse characters, style, or audio as guidance
Create from a descriptionA new clip without source material
Edit a clipKeep the basis, change the content
Extend a clipCreate the frames that follow the source
Transfer motionMotion from video, appearance from an image

Compatible models first. Explanations for the rest

After you choose a role, models are ordered by support for that operation and the type of material connected. Matches come first. Other models available to your account stay visible, with a reason why they do not fit this choice, such as an unsupported operation or an input that this role does not use.

For Extend a clip, the Seedance versions are compatible. Kling and other models stay in the list with an “Operation not supported” explanation.
For Extend a clip, the Seedance versions are compatible. Kling and other models stay in the list with an “Operation not supported” explanation.
This is not a best-to-worst ranking. Compatibility does not guarantee a beautiful result. Before running, the studio separately checks input types, counts, and durations.

Fewer lines on the node. More useful information at a glance

The role sits beside your node’s name. The model and current settings sit immediately below it. You no longer need to reopen a large panel just to remember which run you were setting up.

A detail of the actual Video node: Grok Video 1.5 with 8 seconds and 480p selected. These are the next run’s settings, not the model’s limits.
A detail of the actual Video node: Grok Video 1.5 with 8 seconds and 480p selected. These are the next run’s settings, not the model’s limits.
Role icon
Click to choose a different action.
Model name
Opens the models for the current role.
Timer and other icons
Open a specific setting: duration, resolution, variant, or audio, where supported by the model.
The header describes the next run. Changing the settings does not turn a finished clip into a new result. Its preview stays intact; the new parameters apply when you generate again.

The selected duration is not the model’s limit

The node screenshot has eight seconds selected. For video from a starting frame, Grok Video 1.5 supports up to 15 seconds, while Seedance 2.5 supports up to 30. Click Compare and select up to three models to see the difference. The view shows available information about maximum duration, resolution, and inputs for the chosen role and variant.

A variant is not a universal quality score. Seedance 2.0, for example, offers Draft, Fast, and Full; FLUX 3 offers Draft and Full. FLUX 3 Draft has a fixed 720p output. Changing the variant can change the available settings; matching variant names do not make different models equivalent.

Seconds mean different things too: for generation, the length of the new clip; for editing, possibly “same as source video”; for extension, the length of the added segment, not the whole film. Resolution sets frame size, while a variant selects a version within a model. Fixed parameters do not get a separate selection button in the compact header.

Kling, Seedance 2.5, and Grok Video 1.5 compared for video from a starting frame. The icons show limits, not the node’s current settings.
Kling, Seedance 2.5, and Grok Video 1.5 compared for video from a starting frame. The icons show limits, not the node’s current settings.
On the node: what I chose. In comparison: what I can choose.

Every role, explained through a task

Each walkthrough tells you what to connect, which role to choose, and what result to request. The images and clips come from our earlier projects, not a new model benchmark. The prompts are separate examples of intent, not promises to reproduce the shown results.

01Your image defines the first frame of the video

What to connect
A starting image + what happens next
What you get
A clip that starts from your frame

Choose “Video from a starting frame”: this is image-to-video. Connect the image to the first-frame input, then choose a model and duration. The image sets the initial scene; the prompt describes what happens next: a character’s actions, a developing scene, a transformation, or camera movement. Simply animating a picture is one possible use, not the definition of the role.

The tomato and cucumber characters from Garden City. This image can define the video’s first frame; it is not a generated result of the prompt below.
The tomato and cucumber characters from Garden City. This image can define the video’s first frame; it is not a generated result of the prompt below.
Example prompt

The camera slowly pushes in. The cucumber tilts his head slightly toward the tomato; she shifts her gaze toward him. Preserve the composition and soft lighting. No dialogue.

A first frame does not always rule out additional references. Kling, for example, accepts character Elements in this same role: the opening image sets the scene’s start, while Elements help describe the characters. This is a capability of that model within the role, not a rule for every image-to-video model.

Not one frame, but an opening and an ending

In the same Video from a starting frame role, choose a model that supports a last frame, such as Seedance 2.5. The first image sets the opening; the second sets where the scene should arrive. Connect them to the first-frame and last-frame inputs, then describe the transition. These are endpoints, not simply two character references.

Example: first and last frame

Need to control the middle of the clip too?

Choose FLUX 3 under Video from a starting frame. Connect intermediate images to keyframe inputs and set their times in seconds. For example: a helicopter at the opening, the transformation starting at three seconds, and a robot at the end. Up to 10 frames are supported, including the first and last. These are timed visual targets, not a collection of character appearances.

A real screenshot from our keyframes case: four images guide the stages of a transformation. Captured before the compact header update; it illustrates inputs and timing, not the new node design.
A real screenshot from our keyframes case: four images guide the stages of a transformation. Captured before the compact header update; it illustrates inputs and timing, not the new node design.
How to place video keyframes in time

02You have a character and an object and need a new scene

What to connect
Character, object, motion, or audio references + a prompt
What you get
A new scene guided by the chosen references

Choose “Create from references”: this is reference-to-video. These images do not have to become the opening frame: they guide who or what should appear in a new scene. Connect the materials you need and give each a clear job in the prompt.

Generated characters and an artifact from our Seedance 2.5 case. These are visual references, not a sequence of frames for the next video.
Generated characters and an artifact from our Seedance 2.5 case. These are visual references, not a sequence of frames for the next video.
Example prompt

Use the selected character’s reference for their appearance, and the Key image for the artifact’s shape. The character notices the Key on a table; the camera racks focus from the face to the object. Do not add the other characters to the shot.

Not every model accepts everything together. Images, video, and audio can only be combined in an operation that supports them. An audio input does not by itself mean that audio alone is enough to start a run.

Image + video + audio, in one run

Seedance 2.5 accepts all three together under Create from references: an image guides the character or object, a clip guides motion or atmosphere, and audio provides a sound reference. This is one combined operation, not three generations in sequence. Input video and audio share a 30-second total budget; that is separate from output duration. Explain each file’s purpose in the prompt; the new shot does not have to reproduce a frame from the reference clip.

Example prompt

Use the image for the Key’s shape, the video for the slow camera push-in and drifting smoke, and the audio for rhythm and atmosphere. Create a new shot of the Key floating above a wrecked starship. Do not copy characters from the reference clip. No speech.

Only have audio: can you still create a video?

Yes: choose Seedance 2.5 and Create from references, connect the audio, and describe the visuals. This operation does not require an image or a video. Seedance 2.0 is different: you must add an image or video alongside the audio. Here, audio-to-video is a use of the references role, not a separate button available for every model.

Example prompt

Create visuals for the audio reference: glowing waves travel across a dark lake, their movement following the music’s rhythm. One continuous shot, without characters or speech.

This is not the output-audio switch. An audio reference is an input; “model generates”, “source audio”, and “no audio” control the output soundtrack. An audio reference by itself does not promise precise character lip-sync.

03Only have an idea? Create from scratch

What to connect
A text description only
What you get
A new clip without source images or video

Choose “Create from a description”: this is text-to-video. No image, video, or audio input is needed; the prompt defines the scene. Use it for a new shot that does not need to begin with a particular image. If matching an existing character matters, use references instead.

Example prompt

Wide shot of an abandoned greenhouse at dawn. Enormous leaves grow between glass panels. The camera moves slowly along the aisle; dust floats in the shafts of sunlight. No people or speech.

04You like the clip and want to change one detail

What to connect
An existing clip + a description of changes
What you get
An edited version of that shot

Choose “Edit a clip” and connect the source video. Start the prompt with the change, then specify what should stay. You do not need to choose new-video generation just because you recognize the model’s name.

An earlier Seedance 2.5 edit: source on the left, edited faceplate logo on the right. The background sign stayed unchanged. The new interface helps you choose this kind of task; it does not promise the same precision from every model.
Example prompt

Replace the logo on the character’s faceplate with Vini Studio. Preserve the character, car, camera movement, and confetti. Do not change the sign in the background.

Editing requests changes within an existing shot. For a new scene merely guided by your clip, choose references. For events after its ending, choose extension.

Find the materials and details in Seedance 2.5: references and video editing.

05The clip ends too soon: shoot what happens next

What to connect
An existing clip + what should happen next
What you get
A continuation after the source clip ends

Choose “Extend a clip”, connect the source video, and choose a model that supports this operation, such as Seedance 2.5. Describe what happens after the clip’s final moment. Duration describes the new continuation: five seconds means five more seconds of action, not shortening the source to five.

Example prompt

Continue events after the end of the source video. Preserve the camera’s direction, lighting, and rain. The character gives a small wave, then the camera settles on a medium shot. Do not restart the scene. No speech.

This is neither editing nor slow motion: you request new frames after the existing ones.

06Keep the movements, use a different character

What to connect
A character image + a clip with the desired movements
What you get
Appearance and setting from the image, actions from the clip

Choose Kling and “Transfer motion”. Connect the desired character image as the first frame and a movement clip as the source video. The image supplies the appearance and setting; the clip supplies the actions. For example, a character from your illustration performs the dance from the motion clip.

Example prompt

Use the image for the character and environment. Transfer the character’s actions from the motion clip while preserving the image character’s appearance and clothing.

This is not “replace the character inside the old video”: the source scene need not remain, because the image supplies the visual basis. Duration follows the motion clip; you can preserve its audio or switch sound off.

Try it on your canvas

  1. Add Video using “+”, or drag a wire from source material into empty canvas space and choose Video.
  2. Choose a role, then a model. Open comparison here without leaving your canvas.
  3. Review duration, variant, resolution, and audio. Available controls depend on the chosen operation.
  4. Create the node or apply the selection to an existing one. Review its inputs and cost estimate before running.

When creating from a wire, you still choose the role. The selector then uses the material type to order models; it does not automatically guess the task.

The update is available to all studio users. The models have not become identical, and we have not started choosing for you. Their differences are now presented in terms of what you want to do.

Build your own hero in Vini Studio

A cloud AI video generator: from an idea in words to a finished clip with sound, 9:16 or 16:9. The project is in closed beta — I am looking for authors and partners.

Get started in the studio →Sign in with Telegram or Google. No account needed in advance — it is created for you.