Skip to main content
HeyGen Video 1.0 is a general-purpose video model from HeyGen that generates the picture, the spoken line, and the ambience in one pass. You describe the shot and the dialogue in the prompt, and the clip comes back with synchronized speech, so a talking-head or product video needs no shoot, no presenter, and no separate voice-over step. In ComfyUI the model runs through two partner nodes. HeyGen Video 1.0 Reference to Video handles text to video on its own and reference to video once images, videos, or audio are connected; HeyGen Video 1.0 Image to Video starts from a first frame. Both are cloud API nodes: they need a Comfy account with credits and download no local model. See Partner Nodes pricing for the per-second rates. Every run returns a 5 to 15 second clip at 480p or 768p, with the dialogue, room tone, and any music already in the video.
To use the Partner Nodes, you need to ensure that you are logged in properly and using a permitted network environment. Please refer to the Partner Nodes Overview section of the documentation to understand the specific requirements for using the Partner Nodes.
Make sure your ComfyUI is updated.Workflows in this guide can be found in the Workflow Templates. If you can’t find them in the template, your ComfyUI may be outdated.If nodes are missing when loading a workflow, possible reasons:
  1. You are not using the latest ComfyUI version (Nightly version)
  2. Some nodes failed to import at startup

Available workflows

Text to Video

Generate a clip from a prompt alone. The prompt carries the action, the spoken line, and the sound, so one node run returns a finished talking clip.

Run on Comfy Cloud

Open in Comfy Cloud

Download Workflow

Download JSON or search “HeyGen Video 1.0: Text to Video” in Template Library
The text-to-video template runs the Reference to Video node with nothing connected, which is the text-to-video mode of that node.

Image to Video

Animate one image. The image is the first frame of the clip, and the output keeps its aspect ratio, so crop the image when you need a different shape. The prompt drives what happens and what is said.

Run on Comfy Cloud

Open in Comfy Cloud

Download Workflow

Download JSON or search “HeyGen Video 1.0: Image to Video” in Template Library
Input material

model_red_curls_ring.png

Load into the LoadImage node that feeds the HeyGen Video 1.0 Image to Video node

Reference to Video

Keep people, products, and places consistent across the shot by connecting them as references: up to 9 images, 3 videos, and 3 audio clips, 12 in total. The prompt then addresses them as @Image1, @Video1, and @Audio1, numbered per type in input order.

Run on Comfy Cloud

Open in Comfy Cloud

Download Workflow

Download JSON or search “HeyGen Video 1.0: Reference to Video” in Template Library
Input materials

model_black_blazer_red_earrings.png

Load into the first reference image slot · @Image1

cognac_leather_handbag.png

Load into the second reference image slot · @Image2

Node controls

The Image to Video node takes the image, the prompt, the duration, the resolution, and the seed. It has no aspect ratio control, because the first frame sets the shape of the clip.

Prompting tips

  • Write the spoken line into the prompt. Put the dialogue in quotes as part of the description of the shot, so the model speaks it and matches the lip movement. Both templates ship long example prompts that show the shape: the shot and the subject first, the line in quotes, then the sound and the camera language.
  • Describe the sound with the picture. Room tone, a piece of music, or the noise an action makes is written into the prompt rather than set with a control. The sample prompts end with the ambience, for example quiet room tone with a faint kettle whistle, or a soft glass clink as an object is set down.
  • Keep a product or presenter consistent. Connect the product or the person as a reference image and name it in the prompt, then describe how it must stay the same, such as keeping the exact color, proportions, and hardware of a bag from @Image2.
  • Shape the shot with references. With aspect_ratio set to auto, the framing follows the first reference image, so the reference you connect first decides the shape of the video.
  • Expect variation between runs. The seed does not lock the result; the same settings can still produce a different take.

Get started

  1. Update ComfyUI to the latest version
  2. Go to Template Library, search for HeyGen Video 1.0