Skip to main content
Cloud / Google Cloud / Products / Deployments (formerly Agent Engine) - AI Agent Runtime

Deployments (formerly Agent Engine) - AI Agent Runtime

Deployments/Agent Runtime (formerly Vertex AI Agent Engine) is Google Cloud's managed runtime for autonomous AI agents in production.

AI/ML
Pricing Model Pay-per-use (compute time + invocations)
Availability Global, EU regions
Data Sovereignty EU regions available
Reliability SLA as published by the provider SLA

Deployments, also referred to as Agent Runtime, is Google’s managed runtime environment for operating autonomous AI agents in production. The service evolved from Vertex AI Agent Engine, introduced in 2024, and is now part of the Gemini Enterprise Agent Platform (formerly Vertex AI). It targets development teams that want to run agent frameworks such as the Agent Development Kit (ADK), LangChain, LangGraph, or LlamaIndex at scale without managing their own infrastructure.

What is Deployments (Agent Runtime)?

Deployments handles all the infrastructure surrounding the operation of LLM agents: automatic scaling, session and memory management via a Memory Bank, logging/tracing, and integration into the Google Cloud ecosystem. Developers deploy their existing agent frameworks directly onto the runtime without needing to manage containers or Kubernetes clusters themselves. The service supports tool use — the ability of agents to call external functions and APIs — and manages the associated state across multiple conversation turns.

An important distinction from the related Agent Studio (formerly Vertex AI Agent Builder): Agent Studio focuses on low-code creation of RAG applications, search systems, and chatbots via a graphical interface. Deployments/Agent Runtime, on the other hand, is the programmatic runtime environment for code-first agents that execute complex workflows, orchestrate multiple tools, and integrate into existing systems.

The service offers resource controls (CPU, memory, concurrency limits), custom service accounts and agent identities, and monitoring via traces and logs. Integration with other Google Cloud services lets agents access enterprise data directly. Security features such as VPC Service Controls and IAM-based access control make the service production-ready for enterprise applications.

Integration with innFactory

As a certified Google Cloud partner, innFactory supports you in designing and operating AI agent architectures on Deployments/Agent Runtime — from choosing the right agent framework and integrating tools to production-ready deployment on the Gemini Enterprise Agent Platform.

Contact us for a consultation on Deployments/Agent Runtime and autonomous AI systems.

Typical Use Cases

Production operation of LLM agents
Multi-agent systems
Tool use and function calling for AI
Orchestration of agent workflows

Note: All product information on this page has been compiled with care, but is provided without guarantee and may be outdated or incomplete. Cloud services evolve rapidly — features, pricing, SLAs, and availability change frequently. Authoritative and up-to-date information can only be found on the official product page of Google Cloud (official documentation). This page does not represent an offer by Google Cloud.

Google Cloud Partner

innFactory is a certified Google Cloud Partner. We provide expert consulting, implementation, and managed services.

Google Cloud Partner

Similar Products from Other Clouds

Other cloud providers offer comparable services in this category. As a multi-cloud partner, we help you choose the right solution.

AWS

Amazon Augmented AI (A2I) - Human Review for ML

Amazon Augmented AI (A2I) enables human review of ML predictions. Human-in-the-loop workflows for AI quality assurance.

Pricing Pay-per-use: price per human review task
SLA SLA as published by the provider (A2I is part of Amazon SageMaker AI)
Compare →
AWS

Amazon Bedrock AgentCore - AI Agent Runtime

Amazon Bedrock AgentCore: serverless runtime and services to securely run, scale, govern and observe production AI …

Pricing Pay-per-use (consumption-based, …
SLA N/A
Compare →
AWS

Amazon Bedrock Agents (Classic): Status and Alternative

Amazon Bedrock Agents is now Bedrock Agents Classic and in maintenance mode. AWS recommends Bedrock AgentCore for new …

Pricing Pay-per-use (model tokens and connected …
SLA SLA as published by the provider
Compare →
AWS

Amazon Bedrock Data Automation - Structure Data

Amazon Bedrock Data Automation turns documents, images, audio, and video into structured outputs via API: for IDP, media …

Pricing Pay-per-use (per page / per image / per …
SLA N/A
Compare →
AWS

Amazon Bedrock Guardrails - Safety for Generative AI

Amazon Bedrock Guardrails filters harmful content, protects PII, and checks responses for factual accuracy, …

Pricing Pay-per-use (billed per evaluated text …
SLA SLA as published by the provider
Compare →
AWS

Amazon Bedrock Knowledge Bases: Managed RAG

Amazon Bedrock Knowledge Bases: a fully managed RAG service for precise, verifiable AI answers grounded in your …

Pricing Pay-per-use (embeddings, vector storage, …
SLA 99.9%
Compare →

80 comparable products found across other clouds.

Ready to start with Deployments (formerly Agent Engine) - AI Agent Runtime?

Our certified Google Cloud experts help you with architecture, integration, and optimization.

Schedule Consultation