Is omni-flash a standard Gemini conversational model?
No. omni-flash here is designed for video creation: submit text and optional assets through POST /gemini/videos to receive video results. It is not a Gemini Flash model that returns standard text answers through a conversational interface; during development, organize inputs, status queries, and result storage around video tasks.
Can I generate a video with just one image?
You can submit an image link through image_urls while also providing the required prompt. Clearly describe how the subject should move, how the camera should move, and which visual characteristics you want to preserve. Images are used to guide generation, so do not write only “make it move,” or it will be difficult to convey the shot effect you actually need.
What assets are needed to edit an existing video?
You need to submit a video link in video_urls, provide at least one reference image in image_urls, and use prompt to describe the editing goal. You can specify requirements for style, scene, or visual elements, and should also explain which layout needs to be preserved. Submitting only a video without an image does not meet the input requirements for this workflow.
Can I choose portrait orientation and 1080p output?
Yes. aspect_ratio supports 16:9 and 9:16, while resolution supports 720p and 1080p; the defaults are 16:9 and 720p respectively. Determine the delivery aspect ratio before creation, especially the subject position and camera movement direction, to avoid cropping after generation that shifts the composition away from the original intent.
How do I obtain a video after asynchronous generation?
After setting async to true, save the returned task_id and submit a query request to /gemini/tasks using that value as the id; you can also set callback_url to receive completion notifications. Once the task succeeds, read video_url from the result and download it for storage. Continue waiting while it is pending, and check the error message if it fails.