Will the reference image directly become the first frame of the video?
This model uses reference images to guide characters and visual elements, and should not be used as a fixed first-frame workflow. If you need an image with a specific composition as the opening frame, choose happyhorse-1.1-i2v. r2v is better suited for preserving the reference subject while rearranging the environment, actions, and camera through text.
How do I make prompts correspond to multiple reference images?
Put 1–9 image URLs in image_urls, and reference them in order in the prompt using character1, character2, and so on. In addition to specifying the corresponding assets, clearly describe the subject's actions, scene, and the purpose of each asset. Avoid merely listing names without explaining their relationships.
Which key items need to be explicitly specified when calling it?
Set action to reference_to_video, model to happyhorse-1.1-r2v, and submit prompt and image_urls. Do not rely on the default text-to-video action. Then choose resolution, ratio, and duration according to your needs to form a complete reference-image generation request.
Can it generate audio or preserve the original video's sound?
This model has native audio capabilities, but reference-image generation does not mean preserving the original video's sound, nor does it mean specifying voice-over or lip sync. The origin usage for preserving original audio belongs to video editing tasks; if audio has specific delivery requirements, check the generated result and arrange any necessary audio post-production.
How do I retrieve asynchronously generated videos?
After submitting with async, save the task_id and query the task through /happyhorse/tasks; you can also provide callback_url to receive completion notifications. The result includes the task status and video_url. Statuses may be pending, succeeded, or error. Download the video for inspection after successful completion.