Browse models from Google
Models
Google: Gemma 3 12B
- Model ID: google/gemma-3-12b-it
- Tokens: 491.80M
- Description: Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities, including structured outputs and function calling. Gemma 3 12B is the second largest in the family of Gemma 3 models after Gemma 3 27B.
Google: Gemini 2.5 Flash
- Model ID: google/gemini-2.5-flash
- Tokens: 4.26B
- Description: Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater accuracy and nuanced context handling.
Google: Gemini 2.5 Pro
- Model ID: google/gemini-2.5-pro
- Tokens: 3.25B
- Description: Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy and nuanced context handling. Gemini 2.5 Pro achieves top-tier performance on multiple benchmarks, including first-place positioning on the LMArena leaderboard.
Google: Gemini 2.5 Flash Lite
- Model ID: google/gemini-2.5-flash-lite
- Tokens: 24.72B
- Description: Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models.
Google: Gemini 2.5 Flash Image (Nano Banana)
- Model ID: google/gemini-2.5-flash-image
- Tokens: 233.13M
- Description: Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is a state-of-the-art image generation model with contextual understanding. It is capable of image generation, edits, and multi-turn conversations.
Google: Gemini 3.1 Flash Lite Preview
- Model ID: google/gemini-3.1-flash-lite-preview
- Tokens: 15.07B
- Description: Gemini 3.1 Flash-Lite is Google's most cost-efficient Gemini model, optimized for low latency use cases for high-volume, cost-sensitive LLM traffic.
Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)
- Model ID: google/gemini-3.1-flash-image-preview
- Tokens: 743.23M
- Description: Designed for speed and efficiency, the Gemini 3.1 Flash Image generation model is effective for quick, interactive responses and high throughput.
Google: Gemini 3.1 Pro Preview
- Model ID: google/gemini-3.1-pro-preview
- Tokens: 25.93B
- Description: Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models.
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)
- Model ID: google/gemini-3-pro-image-preview
- Tokens: 1.68B
- Description: Nano Banana Pro is an AI image generation model that is a significant upgrade from its predecessor, promising to move beyond simple pattern matching to a more reasoning-driven system.
Google: Gemini 3 Flash Preview
- Model ID: google/gemini-3-flash-preview
- Tokens: 79.40B
- Description: Gemini 3 Flash Preview is a low-latency model in the Gemini 3 family, optimized for fast, high-throughput inference.
Google: Veo 3.1 Lite
- Model ID: google/veo-3.1-lite-generate-001
- Tokens: -
- Description: Veo 3.1 Lite is Google DeepMind's most cost-efficient AI video generation model, designed for professional-grade video capabilities.
Google: Veo 3.1 Fast
- Model ID: google/veo-3.1-fast-generate-001
- Tokens: 32.00K
- Description: Veo 3.1 Fast is a speed-optimized variant of Google DeepMind's flagship video generation model.
Google: Veo 3.1
- Model ID: google/veo-3.1-generate-001
- Tokens: 352.04K
- Description: Veo 3.1 is Google's state-of-the-art model for generating high-fidelity videos featuring stunning realism and natively generated audio.