Multi-Ref Video: up to four characters and a setting in one clip

Multi-Ref Video makes one clip from references: up to four subjects (characters, animals, objects) and one setting, connected to the generator, then a scene you describe in your own words, calling each reference by its name. The same characters come back from one clip to the next, together, in the same place.

Multi-Ref Video with Lea and Lucas connected, a setting, and a scene that mentions @Lea and @Lucas

Start with a character sheet, not a LoRA

This is the step that decides whether a character looks like itself. Multi-Ref Video does not use the LoRA. A character's likeness comes from its character sheet: five reference views of the same character (Bust, front; Bust ¾ left; Bust ¾ right; Full body, front; Full body, back or profile).

A character without a character sheet is rendered from a single image, and the likeness is clearly weaker. It is the most common surprise: a character with a well-trained LoRA but no sheet comes out as a loose version of itself.

The app flags it. In Multi-Ref Video, such a character shows No character sheet, and a click opens its profile. If you click Generate with one connected, a Limited likeness window offers Create the character sheet or Generate anyway.

The Limited likeness window, shown when a connected character has no character sheet

To create one, open My characters, pick the character, then its Character sheets section, and click New sheet. Drop your own images, generate the views from the character's characteristics (with its image LoRA when it has no photo), or complete the missing views from the ones you have. Make one sheet per outfit or attitude: when you connect the character, Multi-Ref Video asks which sheet to use. Everything is explained in Character sheets.

Creating a character needs a licence. In the demo, Lea and Lucas come with complete character sheets, so you can try right away.

Objects, animals and settings

Characters are not the only references. In the Library, two pages hold the others: Objects & animals and Places & sets. Each item gathers up to five images, a title and a description, and connects to Multi-Ref Video the same way as a character. A clip takes one setting at most.

Library, Objects & animals: the window to create an item, with a title, a description and up to five images

Build the scene

  1. In Multi-Ref Video, use Add a subject (a character, an animal or an object, up to four) and Add a setting (one). Each reference is connected to the generator.

    The Add a subject window

  2. Describe the scene and mention each reference with @ and its name, for example: "@Lea hands a letter to @Lucas on a station platform, evening light".

  3. If you want help, the AI button rewrites your idea into a text ready for the video engine (shot, action, camera, light, sound, lines) and keeps the @references. It works online at no credit cost, or offline on your own computer once the local assistant is installed (2.8 GB, at the first click or from Settings > Components).

    A scene idea typed in the text field, before clicking AI

    The same scene rewritten by the AI assistant

Each reference has a Strength from 0 to 2: 0 ignores it, 1 is normal, above 1 it takes more room in the clip. The seed is automatic by default, a new one per render: fix it to redo the same shot while you change one thing at a time (see seed).

A reference card up close, with its Strength slider and its Character sheet label

Write the scene

The AI button applies these rules for you. If you write the scene yourself:

  • Call each reference by @ and its name. The app links it to its images at render time.
  • Say how many people are in the frame and where each one stands ("Lea on the left, Lucas on the right, nobody else in the foreground"). That is what keeps a character from appearing twice.
  • One action after another, with visible gestures: the model shows what can be seen, not what is felt.
  • A short line of dialogue in quotes, tied to whoever says it: about two to three words per second of clip. Longer, and the speech skips words or turns into a voice-over.
  • Describe what is there, never what must not happen: the engine ignores negations.
  • A character that is not human (a creature, a dragon): say what it is in the scene, and describe it as such on its profile ("wings, claws, a tail", not "a dragon arm outfit"), otherwise the engine turns it into a person in costume.
  • The setting: an image in the video's format (portrait, landscape or square); otherwise it is cropped to the center.

Mode, format and price

  • Mode: Quality or Ultra. Fast is not available with several references (see Fast, Quality or Ultra).
  • Format: portrait, landscape or square, in 480p (Quality only), 720p or 1080p.
  • Duration: automatic, or 5, 10 or 20 seconds, with or without sound.

Multi-Ref Video settings: mode, format, duration, resolution, sound, seed and price

In Quality, a Multi-Ref clip costs the same as a regular clip of the same format (see cloud video). Ultra has its own fixed price per clip in Multi-Ref Video, shown in the app before you render, and runs on Cloud GPU only.

In the demo, Multi-Ref Video can be tried with the free video, in 720p and up to 5 seconds.

On your computer or online

On Windows, Multi-Ref Video can also render on your own graphics card, in Quality mode. It needs a recent NVIDIA card with at least 8 GB and a component downloaded once (2.2 GB, from the tab or from Settings > Components). A local render takes about twice as long as a clip without references. On Mac, local video is not available yet: switch to Cloud GPU.

Online, a render waits up to 30 minutes for a free machine, then has 30 minutes to compute. The progress bar follows the real render time and shows when it is waiting for a machine or finishing the clip. If a render is abandoned, it is cancelled and nothing is charged.

Pick a clip up again

The finished clip preview, with Edit the prompt, Regenerate and Show file

In the Gallery (Videos tab), Reuse on a Multi-Ref clip reopens Multi-Ref Video with its references connected again, along with its text, format and seed. Change one thing, generate again.

For a single character across images and ordinary clips, see character consistency.

Windows · Mac · Cloud or local

Generate privately. Tonight.

Download tendre.AI, activate once, and run it forever, fully offline, on your own GPU.

OS Windows 10/11GPU NVIDIA 8GB+VRAM 12GB recommendedDisk 20GB

Mac: Apple silicon (M1 or later), macOS 13 or later. Intel Macs are not supported.

Local image generation on Mac needs an M4 with 24 GB of memory, or an M2 or M3 Pro, Max or Ultra with 24 GB. Below that, everything runs online with credits. Retouching and video run online on Mac for now.

Everything about tendre.AI for Mac

No NVIDIA GPU? Start 100% in the cloud: 10 free images every day, then 1 credit per image. An NVIDIA card (8GB+) unlocks unlimited local generation.