model parameter:
grok-imagine-video: the previous generation model, supports optional image inputgrok-imagine-video-1.5: the latest model, always requires an input image and supports 1080p output
What Grok Imagine Video 1.5 is good at
- Image-to-video generation: produces high-quality video from a single input image
- Native audio: sound effects, ambience, and dialogue are synthesized in the same pass, with no separate audio pipeline needed
- Realistic motion: motion stays synchronized with the generated audio
- Up to 1080p output: the 1.5 model supports 1080p resolution
- Flexible duration: video length from 1 to 15 seconds
Use it in ComfyUI
Grok Imagine Video 1.5 workflow
Run the image-to-video workflow in ComfyUI, locally or on Comfy Cloud