model parameter:
grok-imagine-video: the previous generation model, supports optional image inputgrok-imagine-video-1.5: the latest model, always requires an input image and supports 1080p output
What Grok Imagine Video 1.5 is good at
- Image-to-video generation: produces high-quality video from a single input image
- Native audio: sound effects, ambience, and dialogue are synthesized in the same pass, with no separate audio pipeline needed
- Realistic motion: motion stays synchronized with the generated audio
- Up to 1080p output: the 1.5 model supports 1080p resolution
- Flexible duration: video length from 1 to 15 seconds
Grok Imagine Video 1.5: Image to Video
Run on Comfy Cloud
Open in Comfy Cloud
Download Workflow
Download JSON or search “Grok Imagine Video 1.5” in Template Library
Download Sample Input Image
Get the example input image for this workflow.
Workflow Overview
This workflow uses three nodes:- LoadImage: provides the starting image frame
- GrokVideoNode: the core node configured with the
grok-imagine-video-1.5model - SaveVideo: saves the generated video with native audio
Steps to Run
- Upload a starting image: use the LoadImage node to load your reference image
- Enter your prompt: describe the motion, atmosphere, and scene dynamics in the GrokVideoNode node
- Select model: ensure
grok-imagine-video-1.5is selected - Set resolution: choose output resolution (
720precommended) - Set duration: choose the video length in seconds
- Set seed: control whether the node re-runs; outputs are not reproducible regardless of seed
- Click Queue: press
Ctrl+Enterto generate
Output
The generated video includes native audio synchronized with the motion, saved automatically via the SaveVideo node.Tips
- Use high-quality input images for best results
- The prompt works best when it describes both the visual scene and the motion dynamics
- For different results, try varying the seed value