> ## Documentation Index
> Fetch the complete documentation index at: https://dripart-chore-sync-comfy-api-v2-spec.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# HeyGen Video 1.0 in ComfyUI: Text, Image, and Reference to Video

> Generate 5 to 15 second clips with synced dialogue from a prompt, a first frame, or up to 12 reference images, videos, and audio clips in ComfyUI

**HeyGen Video 1.0** is a general-purpose video model from HeyGen that generates the picture, the spoken line, and the ambience in one pass. You describe the shot and the dialogue in the prompt, and the clip comes back with synchronized speech, so a talking-head or product video needs no shoot, no presenter, and no separate voice-over step.

In ComfyUI the model runs through two partner nodes. **HeyGen Video 1.0 Reference to Video** handles text to video on its own and reference to video once images, videos, or audio are connected; **HeyGen Video 1.0 Image to Video** starts from a first frame. Both are cloud API nodes: they need a Comfy account with credits and download no local model. See [Partner Nodes pricing](/tutorials/partner-nodes/pricing) for the per-second rates.

Every run returns a 5 to 15 second clip at 480p or 768p, with the dialogue, room tone, and any music already in the video.

<Tip>
  To use the Partner Nodes, you need to ensure that you are logged in properly and using a permitted network environment. Please refer to the [Partner Nodes Overview](/tutorials/partner-nodes/overview) section of the documentation to understand the specific requirements for using the Partner Nodes.
</Tip>

<Tip>
  <Tabs>
    <Tab title="Local users">
      Make sure your ComfyUI is updated.

      * [Download ComfyUI](https://www.comfy.org/download)
      * [Update Guide](/installation/update_comfyui)

      Workflows in this guide can be found in the [Workflow Templates](/interface/features/template).
      If you can't find them in the template, your ComfyUI may be outdated.

      If nodes are missing when loading a workflow, possible reasons:

      1. You are not using the latest ComfyUI version (Nightly version)
      2. Some nodes failed to import at startup
    </Tab>

    <Tab title="Cloud users">
      * [Cloud](https://cloud.comfy.org) will update after ComfyUI stable release.

      So, if you find any core node missing in this document, it might be because the new core nodes have not yet been released in the latest stable version. Please wait for the next stable release.
    </Tab>
  </Tabs>
</Tip>

## Available workflows

### Text to Video

Generate a clip from a prompt alone. The prompt carries the action, the spoken line, and the sound, so one node run returns a finished talking clip.

<video controls className="w-full aspect-video" src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/output/api_heygen_video_1_t2v.mp4" />

<CardGroup cols={2}>
  <Card title="Run on Comfy Cloud" icon="cloud" href="https://cloud.comfy.org/?template=api_heygen_video_1_t2v&utm_source=docs&utm_medium=referral&utm_campaign=heygen-video-1">
    Open in Comfy Cloud
  </Card>

  <Card title="Download Workflow" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_heygen_video_1_t2v.json">
    Download JSON or search "HeyGen Video 1.0: Text to Video" in Template Library
  </Card>
</CardGroup>

The text-to-video template runs the Reference to Video node with nothing connected, which is the text-to-video mode of that node.

### Image to Video

Animate one image. The image is the first frame of the clip, and the output keeps its aspect ratio, so crop the image when you need a different shape. The prompt drives what happens and what is said.

<video controls className="w-full aspect-video" src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/output/api_heygen_video_1_i2v.mp4" />

<CardGroup cols={2}>
  <Card title="Run on Comfy Cloud" icon="cloud" href="https://cloud.comfy.org/?template=api_heygen_video_1_i2v&utm_source=docs&utm_medium=referral&utm_campaign=heygen-video-1">
    Open in Comfy Cloud
  </Card>

  <Card title="Download Workflow" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_heygen_video_1_i2v.json">
    Download JSON or search "HeyGen Video 1.0: Image to Video" in Template Library
  </Card>
</CardGroup>

**Input material**

<Card title="model_red_curls_ring.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/model_red_curls_ring.png">
  Load into the `LoadImage` node that feeds the HeyGen Video 1.0 Image to Video node
</Card>

### Reference to Video

Keep people, products, and places consistent across the shot by connecting them as references: up to 9 images, 3 videos, and 3 audio clips, 12 in total. The prompt then addresses them as `@Image1`, `@Video1`, and `@Audio1`, numbered per type in input order.

<video controls className="w-full aspect-video" src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/output/api_heygen_video_1_r2v.mp4" />

<CardGroup cols={2}>
  <Card title="Run on Comfy Cloud" icon="cloud" href="https://cloud.comfy.org/?template=api_heygen_video_1_r2v&utm_source=docs&utm_medium=referral&utm_campaign=heygen-video-1">
    Open in Comfy Cloud
  </Card>

  <Card title="Download Workflow" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_heygen_video_1_r2v.json">
    Download JSON or search "HeyGen Video 1.0: Reference to Video" in Template Library
  </Card>
</CardGroup>

**Input materials**

<CardGroup cols={2}>
  <Card title="model_black_blazer_red_earrings.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/model_black_blazer_red_earrings.png">
    Load into the first reference image slot · `@Image1`
  </Card>

  <Card title="cognac_leather_handbag.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/cognac_leather_handbag.png">
    Load into the second reference image slot · `@Image2`
  </Card>
</CardGroup>

## Node controls

| Control | Values | Applies to | Effect |
| - | - | - | - |
| `prompt` | text, up to 32000 characters | both nodes | The shot, the spoken line, and the sound. Reference media is addressed as `@Image1`, `@Video1`, `@Audio1` |
| `duration` | 5 to 15 seconds | both nodes | Slider, default 5 |
| `resolution` | `480p`, `768p` | both nodes | Output resolution |
| `aspect_ratio` | `auto`, `16:9`, `9:16`, `1:1`, `4:3`, `3:4`, `21:9` | Reference to Video | `auto` is 16:9 with nothing connected. With references it follows the first reference image, or the first reference video when no images are connected |
| `seed` | integer | both nodes | Results can still vary between runs with the same seed |
| `reference_images` | up to 9 | Reference to Video | One image per slot. A batched input on a single slot is rejected |
| `reference_videos` | up to 3 | Reference to Video | One video per slot |
| `reference_audios` | up to 3 | Reference to Video | Needs at least one reference image or video connected. A voice reference wants a few seconds of clean speech |

The Image to Video node takes the image, the prompt, the duration, the resolution, and the seed. It has no aspect ratio control, because the first frame sets the shape of the clip.

## Prompting tips

* **Write the spoken line into the prompt**. Put the dialogue in quotes as part of the description of the shot, so the model speaks it and matches the lip movement. Both templates ship long example prompts that show the shape: the shot and the subject first, the line in quotes, then the sound and the camera language.
* **Describe the sound with the picture**. Room tone, a piece of music, or the noise an action makes is written into the prompt rather than set with a control. The sample prompts end with the ambience, for example quiet room tone with a faint kettle whistle, or a soft glass clink as an object is set down.
* **Keep a product or presenter consistent**. Connect the product or the person as a reference image and name it in the prompt, then describe how it must stay the same, such as keeping the exact color, proportions, and hardware of a bag from `@Image2`.
* **Shape the shot with references**. With `aspect_ratio` set to `auto`, the framing follows the first reference image, so the reference you connect first decides the shape of the video.
* **Expect variation between runs**. The seed does not lock the result; the same settings can still produce a different take.

## Get started

1. Update ComfyUI to the latest version
2. Go to Template Library, search for `HeyGen Video 1.0`


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.