Model APIs
Interact with Gemini and other generative AI models for text and multimodal generation, real-time streaming, embeddings, customization, and evaluation.
Interactions API
Use a unified, stateful, and streaming interface to interact with models and agents.
Live API
Stream low-latency, bidirectional voice and video interactions with Gemini models.
Embeddings APIs
Generate text and multimodal embeddings for semantic search, classification, and retrieval.
Tuning API
Customize model behavior using supervised fine-tuning.
Gen AI evaluation API
Evaluate generative AI models and agents programmatically using rubrics and metrics.
RAG APIs
Ground model responses and connect your agents to external data sources using Retrieval-Augmented Generation (RAG).
RAG API v1
Manage RAG corpora, import files, and retrieve relevant context with RAG Engine.
Retrieval and generation output of RAG
Understand the structure of retrieval queries and grounded generation responses.
Platform APIs
Use SDKs, command-line tools, and service-level APIs to manage Agent Platform resources, infrastructure, and developer environments.
Client libraries
Build with the Google Gen AI SDK and Agent Platform SDK in Python, Go, Java, Node.js, and C#.
REST API reference
Explore all HTTP REST resources and methods for Agent Platform services.
RPC API reference
View gRPC service definitions and protocol buffer messages for Agent Platform.
gcloud CLI reference
Manage Agent Platform resources from the command line.
Agent Platform in express mode
Prototype and test generative AI capabilities using streamlined REST endpoints and an API key.
MCP reference
Connect Model Context Protocol (MCP) clients and agents to Agent Platform tools.
Agent Retrieval
Build, manage, and query high-scale vector similarity search indexes.
Agent Gateway
Configure and govern network traffic between agents, MCP servers, and external tools.
Agent Platform Workbench
Manage notebook instances and development environments using REST, RPC, CLI, and client libraries.