Can I use this model without an image?
Yes. This model supports text-to-video; simply provide a prompt to describe the content you want to generate. It is recommended to specify the subject, scene, action, and camera movement, organizing the information around a single shot first. When calling it, explicitly specify grok-imagine-video:reverse to avoid using other default models.
Is a prompt required for image-to-video?
When providing image_url, prompt can be left blank. However, if you want the subject to move in a specific direction, or want the camera to push in or rotate, adding an action description makes your intent easier to express. The image provides the visual starting point, while the text explains how you want the scene to change; they serve different roles.
Can this model generate a 30-second video?
This model supports durations of 1–15 seconds, with a default of 6 seconds, and is not suitable for directly submitting a 30-second task. For longer single-shot videos, consider grok-imagine-video-1.5-fast:reverse, which supports durations of 6–30 seconds; for multiple shots, you can also generate them separately and edit them together.
What should I do after a generation request returns task_id?
task_id is a task identifier, not a video URL. When using async, you can check progress through POST /grok/tasks; when callback_url is set, you can wait for the completion notification. After confirming that the result status is succeeded, read video_url; pending indicates that processing is still in progress.
Is :reverse an independent native video model?
:reverse is a suffix for the public call ID, used to distinguish endpoints, and should not be treated as an independent vendor model name. When using this endpoint, retain the complete ID. It shares the basic text-to-video and image-to-video creation methods with the :official endpoint, but this does not mean that all optional parameters are exactly the same.