Can minimax-i2v accept text input only?
You should provide a first-frame reference image link and create based on a text description. minimax-i2v is an image-to-video model and is not suitable for use in a text-only generation workflow. If you do not have a reference image and want to generate visuals directly from text, you can choose minimax-t2v.
How should reference images be submitted? Can Base64 be used?
Submit the image link through first_image_url; Base64 encoding is not supported. First prepare the image as a URL that the service can access, then send the request; local paths or encoded image strings cannot directly replace this link.
What should prompts focus on describing?
It is recommended to focus on the desired actions, environmental changes, and camera intent, rather than repeatedly listing every detail in the image. Start by trying one main dynamic subject, then adjust the description based on the results; prompts guide generation and are not equivalent to precise frame-by-frame control instructions.
Is it the same model as minimax-i2v-director?
No. Both are image-to-video options, but minimax-i2v-director is a separate director-mode model with a stronger focus on creative control. When using minimax-i2v, work around the first-frame image and prompt, and do not treat director-mode control settings as default features.
How do I obtain the video after asynchronous submission?
After setting async=true, first save the returned task_id, then query the task to obtain its status and results; you can also set callback_url to receive a completion callback. Once generation is complete, read video_url in the results to obtain the video; do not treat receiving a task ID as meaning the video is already complete.