Can minimax-i2v accept text only?
You should provide a first-frame reference image link and create with a text description. minimax-i2v is an image-to-video model and is not suitable for use in a text-only generation workflow. If you do not have a reference image and want to generate visuals directly from text, you can choose minimax-t2v.
How should reference images be submitted? Can I use Base64?
Submit the image link through first_image_url; Base64 encoding is not supported. First prepare the image as a URL accessible to the service, then make the request; local paths or encoded image strings cannot directly replace this link.
What should the prompt focus on describing?
It is recommended to focus on the desired actions, environmental changes, and camera intent, rather than repeatedly listing every detail in the image. Start by trying one main dynamic subject, then adjust the description based on the results; prompts guide generation and are not equivalent to precise frame-by-frame control instructions.
Is it the same model as minimax-i2v-director?
No. Both are image-to-video options, but minimax-i2v-director is a separate director-mode model with a stronger focus on creative control. When using minimax-i2v, work around the first-frame image and prompt, and do not treat director-mode control settings as default features.
How do I get the video after submitting asynchronously?
After setting async=true, first save the returned task_id, then query the task to obtain its status and result; you can also set callback_url to receive a completion callback. Once generation is complete, retrieve the video from video_url in the result; do not regard receiving a task ID as meaning the video is already complete.