What must be provided when calling 1.0-r2v?
Submit a request to /happyhorse/videos, explicitly set model to happyhorse-1.0-r2v and action to reference_to_video, and provide a prompt and 1–9 image_urls. Do not rely on default actions or default models, otherwise they may not match the reference-image generation task.
Can it be used with only one reference image?
Yes, the number of reference images can start from 1. One image is suitable for providing a subject or visual direction, while multiple images are suitable for supplementing different reference relationships. If you require this image to directly become the first frame, choose the i2v variant; R2V focuses on using reference information to create new clips.
How can prompts accurately reference multiple images?
In the order arranged in image_urls, use names such as character1 and character2 to refer to the corresponding images, and clearly specify each purpose. For example, state that the subject references the first image and visual elements reference the second image, then describe the action and camera. After adjusting the image order, you should also update the prompt accordingly.
Can dialogue be generated directly and the original video audio preserved?
Do not design tasks assuming that dialogue, lip-sync, or original audio preservation are confirmed features of 1.0-r2v. Its stated purpose is to generate video from image and text references; preserving existing video audio is an operation of video-edit. When specific audio content is required, voice-over and mixing can be completed in post-production.
How do I get the video after asynchronous submission?
After submitting with async, retain the task_id and query it through /happyhorse/tasks; you can also provide callback_url to receive the completion result. Task statuses include pending, succeeded, and error. After success, obtain video_url and use the returned duration and resolution to check whether the final video meets the requirements of this task.