Can I use this model without an image?
Yes. This model supports text-to-video; simply provide a prompt to describe what you want to generate. It is recommended to specify the subject, scene, action, and camera movement, and organize the information around a single shot first. When calling it, explicitly enter grok-imagine-video:reverse to avoid using another default model.
Is a prompt required for image-to-video?
When providing image_url, prompt can be left blank. However, if you want the subject to move in a specific direction, or want the camera to push in or rotate, adding an action description makes it easier to express your intent. The image provides the visual starting point, while the text explains how you want the scene to change; they serve different purposes.
Can this model generate a 30-second video?
This model supports durations of 1–15 seconds, with a default of 6 seconds, and is not suitable for directly submitting a 30-second task. For longer single-shot videos, consider grok-imagine-video-1.5-fast:reverse, which supports durations of 6–30 seconds; for multiple shots, you can also generate them separately and edit them together.
What should I do after a generation request returns task_id?
task_id is a task identifier, not a video URL. When using async, you can check progress through POST /grok/tasks; when callback_url is set, you can wait for the completion notification. After confirming that the result status is succeeded, read video_url; pending means it is still being processed.
Is :reverse an independent native video model?
:reverse is a suffix for the public call ID used to distinguish entry points, and should not be treated as an independent vendor model name. When using this entry point, retain the full ID. It shares the basic text-to-video and image-to-video creation methods with the :official entry point, but this does not mean that all available parameters are exactly the same.