Google's multimodal model family — massive context windows, grounding with Search, and deep integration with Google Cloud.
Five passes over the same idea, each from a different angle. Do them in order, or jump to whichever you need.
Gemini (Google DeepMind) is a natively multimodal model supporting text, image, video, and audio. Gemini 1.5 Pro offers 1M+ token context. Key features include grounding with Google Search, code execution, and integration with Vertex AI. Gemini Flash offers lower-cost, faster inference for latency-sensitive use cases.