Google releases Gemini Omni multimodal generation model
谷歌发布Gemini Omni多模态生成模型
Google released Gemini Omni, a multimodal generation model that combines image, video, and text inputs to generate videos grounded in Gemini’s real-world knowledge; the post does not disclose model size, pricing, or availability.
Why it matters: HKR-H/K/R all pass: Google’s Gemini Omni adds combined image, video, and text inputs for video generation. Missing parameters, pricing, and rollout timing keep it at the low end of the must-write band.