Can I use grok-imagine-video without an image?
You can start creating text-to-video content from a text prompt. Simply describe the subject, environment, and actions you want to appear; there is no need to prepare a static image first. If you already have a clear visual design, you can instead use image input and focus the prompt on motion and camera changes.
What should I write in the prompt for image-to-video?
Prioritize describing what happens next in the image, such as a person turning around, an object moving, or the camera pushing in, rather than merely repeating what is already in the image. After submitting an image through image_url, action prompts can help convey creative intent; the result should still be checked for subject appearance and visual details.
What is the difference between it and grok-imagine-video-1.5:official?
They are different invocation IDs. grok-imagine-video is intended for creative workflows starting from text and images, while grok-imagine-video-1.5:official explicitly uses image-to-video and requires image_url to be submitted. When migrating, choose again based on the input material rather than simply replacing the name.
Can I directly select fun, normal, or spicy?
Do not submit these names as dedicated switches for this model. When creating, naturally describe the desired atmosphere, actions, and presentation in the prompt, such as relaxed, restrained, or dramatic; style descriptions express generation intent and are not deterministic effect presets.
How do I obtain the video after asynchronous submission?
Save the task_id returned by the submission, track the status through task queries, or set callback_url to receive completion notifications. When the state in the result is succeeded and video_url exists, then display or retrieve the video; pending means it is still being processed, while failed should enter the error-handling flow.