OpenAI compatible API · Attested · Public status

Google Vertex AI

Explore Google Vertex AI models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Google Vertex AIgoogle-vertex

No logs (prepaid)

All providers

ProviderGoogle Vertex AI
Provider websitehttps://cloud.google.com/vertex-ai
Models9 public models
Prepaid routes9
BYOK routes0
Zero data retentionprepaid only
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteTrustedRouter's managed Vertex AI account is covered by contractual Zero Data Retention. This guarantee applies only to TrustedRouter-funded prepaid routes. TrustedRouter does not invoke Google Search or Maps grounding or Gemini Live session resumption on these routes. Google AI Studio is classified separately.
Policy source

Measured performance

56 samples

Continuously sampled across Google Vertex AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT1384 ms
Effective throughput58 tok/s n=5
Uptime100.00%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
google/gemini-3.6-flash 1150 ms 1150 ms 47 tok/s n=1 100.00% 10
google/gemini-2.5-flash-lite 1304 ms 1304 ms 100.00% 5
google/gemini-3.5-flash-lite 1327 ms 1327 ms 93 tok/s n=1 100.00% 8
google/gemini-2.5-flash 1384 ms 1384 ms 100.00% 5
google/gemini-3.1-flash-lite 1516 ms 1516 ms 100.00% 6
google/gemini-3-flash-preview 1591 ms 1591 ms 100.00% 5
google/gemini-3.5-flash 3014 ms 3013 ms 58 tok/s n=2 100.00% 3
google/gemini-3.1-pro-preview 3346 ms 3346 ms 82 tok/s n=1 100.00% 7
google/gemini-2.5-pro 4474 ms 4474 ms 100.00% 7

Google Vertex AI performance history · Full provider & model leaderboard.

Provider models

Models served by Google Vertex AI.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
google/gemini-2.5-flash
Google: Gemini 2.5 Flash
1,048,576 1 $0.315/1M $2.625/1M prepaid
google/gemini-2.5-flash-lite
Google: Gemini 2.5 Flash Lite
1,048,576 1 $0.105/1M $0.42/1M prepaid
google/gemini-2.5-pro
Google: Gemini 2.5 Pro
IQ 103#58 1,048,576 1 $1.3125/1M $10.5/1M prepaid
google/gemini-3-flash-preview
Google: Gemini 3 Flash Preview
IQ 116#25 1,048,576 1 $0.525/1M $3.15/1M prepaid
google/gemini-3.1-flash-lite
Google: Gemini 3.1 Flash Lite
IQ 101#65 1,048,576 1 $0.2625/1M $1.575/1M prepaid
google/gemini-3.1-pro-preview
Google: Gemini 3.1 Pro Preview
IQ 127#9 1,048,576 1 $2.1/1M $12.6/1M prepaid
google/gemini-3.5-flash
Google: Gemini 3.5 Flash
IQ 124#13 1,048,576 1 $1.575/1M $9.45/1M prepaid
google/gemini-3.5-flash-lite
Google: Gemini 3.5 Flash Lite
1,048,576 1 $0.315/1M $2.625/1M prepaid
google/gemini-3.6-flash
Google: Gemini 3.6 Flash
1,048,576 1 $1.575/1M $7.875/1M prepaid
Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.