Google Gemini is a family of natively multimodal large language models developed by Google DeepMind, capable of reasoning across text, images, audio, video, and code. Released in tiers (such as Ultra, Pro, Flash, and Nano) it spans data-centre to on-device deployment and powers Google’s assistant and developer APIs. It is a leading frontier model used for chat assistants, agents, and integrated productivity tools.

Content

  • Gemini models are trained from the ground up to process multiple modalities, enabling tasks that mix vision, audio, and language with long context windows. Tiered variants trade capability for latency and cost, from on-device Nano to high-capability Ultra/Pro, served through the Gemini API and Vertex AI. The models support tool use and function calling, making them a backbone for agentic applications and search integration.