Do I need to prepare an image to call happyhorse-1.0-t2v?
No. For text-to-video, use action=generate and provide a prompt describing the subject, scene, action, and style. If you want generation to strictly begin from an image, choose the corresponding i2v model instead of adding a first-frame requirement to this text-generation model.
How can I ensure that I am calling 1.0 rather than 1.1?
Explicitly specify model=happyhorse-1.0-t2v in the request, together with action=generate. The default text-to-video option is 1.1, so if you need to consistently use the 1.0 workflow, save the complete model configuration rather than saving only the prompt or reusing examples that omit the model.
Which durations and aspect ratios are supported?
The platform's text-to-video entry supports 3–15 seconds, with a default of 5 seconds; available aspect ratios are 16:9, 9:16, 1:1, 4:3, and 3:4. It is recommended to determine the aspect ratio based on the final display placement before writing composition descriptions; content exceeding the single-generation range can be split into multiple generations and edited afterward.
Must I keep the connection open while generating?
You can submit an asynchronous task using async=true, save the task_id, and query it through the task API; you can also provide callback_url to receive the result when the task is complete. Before retrieving the video, check the task status and distinguish between pending, succeeded, and error before arranging downloads and subsequent processing.
Is it suitable for generating complete short films with dialogue?
It is suitable for generating dynamic visual clips from text, but dialogue, music, or lip-sync should not be regarded as guaranteed delivery capabilities of this model. When producing short films with sound, you can first generate and review the visuals, then add voice-over, music, and sound editing so that audio and visuals each meet the project requirements.