Skip to content
AI HOT (Curated Pool)

Google releases Gemini Omni multimodal generation model

谷歌发布Gemini Omni多模态生成模型

Google released Gemini Omni, a multimodal generation model that combines image, video, and text inputs to generate videos grounded in Gemini’s real-world knowledge; the post does not disclose model size, pricing, or availability.

Why it matters: HKR-H/K/R all pass: Google’s Gemini Omni adds combined image, video, and text inputs for video generation. Missing parameters, pricing, and rollout timing keep it at the low end of the must-write band.

Read the original ↗Export Markdown