Skip to main content
Documentation
close
Get Started
Get Started with Google Cloud
Product List
Cloud Customer Care
Featured Products
Agent Platform
Apigee API Management
BigQuery
Compute Engine
Cloud CDN
Cloud Run
Cloud Storage
Cloud SQL
Gemini Enterprise
Google Kubernetes Engine
Looker
Cross-product Tools
Access and resources management
Costs and usage management
Infrastructure as code
SDK, languages, frameworks, and tools
Technology Areas
AI and ML
Application development
Application hosting
Compute
Data analytics and pipelines
Databases
Distributed, hybrid, and multicloud
Industry solutions
Migration
Networking
Observability and monitoring
Security
Storage
/
Console
English
Deutsch
Español – América Latina
Français
Indonesia
Italiano
Português – Brasil
עברית
中文 – 简体
中文 – 繁體
日本語
한국어
Sign in
Gemini Enterprise Agent Platform
Start free
Overview
Studio
Agents
Models
Notebooks
Pricing
Agent Platform
Generative AI
Engineering Blog
Documentation
More
Overview
Studio
Agents
Models
Notebooks
Pricing
More
Engineering Blog
Console
Overview
Beginner's guide
Get started
Get started with Agent Platform
Develop Gemini API code with the Gen AI SDK
Connect to the Knowledge MCP server
Get an API key
Configure application default credentials
Migrate from Google AI Studio to Agent Platform
Get started with Gemini 3
Developer guides for Gemini models
Gemini 3.8 Flash
Gemini 3.7 Flash
Gemini 3.6 Flash
Gemini 3.5 Flash
Google GenAI libraries
Generative AI cookbook
Access Gemini models using OpenAI libraries
Express mode
Overview
Console tutorial
API tutorial
Select models
Model Garden
Overview of Model Garden
Use models in Model Garden
Test model capabilities
Google Models
All Google models
Gemini
Migrate to the latest Gemini models
Pro
3.1 Pro
3 Pro Image
2.5 Pro
Flash
3.8 Flash
3.7 Flash
3.6 Flash
3.5 Flash
3.1 Flash Image
3 Flash
2.5 Flash
2.5 Flash Image
2.5 Flash Live API
Flash-Lite
3.5 Flash-Lite
3.1 Flash-Lite Image
3.1 Flash-Lite
2.5 Flash-Lite
Omni
Omni 1.1 Flash
Omni Flash
Cyber
3.8 Flash Cyber
Transcribe
Gemini 3.5 Transcribe
Translate
Gemini 3.5 Live Translate
Embedding
Gemini Embedding 2
Robotics
Gemini Robotics ER 2
Overview
Spatial reasoning
Agentic capabilities
Task orchestration
Video understanding
Veo
Veo 3
Veo 3.1
Lyria
Lyria 2
Lyria 3
Virtual Try-On
Model versions
Partner Models
Partner models overview
Claude
Overview
Request predictions
Quotas for Anthropic Claude models
Batch predictions
Structured outputs
Prompt caching
Count tokens
Web search
Safety classifiers
Cyber Verification Program
Model details
Claude Fable 5.1
Claude Opus 5
Claude Sonnet 5
Claude Fable 5
Claude Opus 4.8
Claude Opus 4.7
Claude Sonnet 4.6
Claude Opus 4.6
Claude Opus 4.5
Claude Sonnet 4.5
Claude Opus 4.1
Claude Haiku 4.5
Claude Opus 4
Claude Sonnet 4
Grok
Overview
Responses API
Function calling
Structured output
Reasoning
Model details
Grok 4.1 Fast
Grok 4.20
Grok 4.3
Grok 4.6
Mistral AI
Overview
Model details
Mistral Medium 3
Mistral OCR (25.05)
Mistral Small 3.1 (25.03)
Codestral 2
Deploy partner models from Model Garden
Partner model deprecations
Open Models
Overview
AlphaFold 3
AlphaGenome
DeepSeek
Overview
DeepSeek-V3.2
DeepSeek-V3.1
DeepSeek-R1-0528
DeepSeek-OCR
Embedding (e5)
Multilingual E5 Small
Multilingual E5 Large
Google Gemma
Model-as-a-Service (MaaS)
Gemma-4-26B-A4B-IT MaaS
Use Gemma
Tutorial: Deploy and inference Gemma (GPU)
Tutorial: Deploy and inference Gemma (TPU)
Kimi
Overview
Kimi K2 Thinking
Llama
Overview
Request predictions
Model details
Llama 4 Maverick
Llama 4 Scout
Llama 3.3
MiniMax
Overview
MiniMax M2
OpenAI
Overview
OpenAI gpt-oss-120b
OpenAI gpt-oss-20b
Qwen
Overview
Qwen 3 Next Instruct 80B
Qwen 3 Next Thinking 80B
Qwen 3 Coder
Qwen 3 235B
ZAI.org
Overview
GLM 5.2
GLM 5
GLM 4.7
Managed open models (MaaS)
Overview
Use open models via Model as a Service (MaaS)
Grant access to open models
API
Call MaaS APIs for open models
Function calling
Thinking
Structured output
Batch prediction
Open model deprecations
Self-deployed open models
Overview
Deploy open models
Deploy open models from Model Garden
Deploy open models with prebuilt containers
Deploy open models with a custom vLLM container
Deploy models with custom weights
Use Hugging Face Models
Tutorials
Optimize model performance with advanced features in Model Garden
Hex-LLM
Comprehensive guide to vLLM for Text and Multimodal LLM Serving (GPU)
vLLM TPU
xDiT
Deploy Llamma 3 models with SpotVM and Reservations
Build
Prompt design
Introduction to prompting
Prompting strategies
Overview
Give clear and specific instructions
Use system instructions
Include few-shot examples
Add contextual information
Structure prompts
Compare prompts
Instruct the model to explain its reasoning
Break down complex tasks
Experiment with parameter values
Prompt iteration strategies
Task-specific prompt guidance
Design multimodal prompts
Design chat prompts
Capabilities
Safety
Overview
Responsible AI
System instructions for safety
Configure content filters
Gemini for safety filtering and content moderation
Abuse monitoring
Process blocked responses
Content Credentials
AI Content Detection API
Text and code generation
Text generation
System instructions
Structured output
Content generation parameters
Image generation
Generate images with Gemini
Generate images from video with Gemini
Edit images with Gemini
Gemini image generation best practices