GCP

GCP

GCP

Storage Intelligence Advisor GA: Cloud Storage zero-setup metrics, Flexible CUDs for G2/G4 GPUs, App Topology billing change

Storage Intelligence Advisor for Cloud Storage is GA with zero-setup metrics and anomalies. GCP widened GPU CUDs and moved App Topology to usage billing.

Sep 20, 2026·3mgoogle-cloudcloud-storage
GCP

Cloud Run Instances (Preview): long-lived, individually addressable serverless containers

GCP adds Cloud Run instances for long-lived, addressable containers; Gemini's API expands agentic video capabilities; GKE unveils agent-focused tooling.

Sep 19, 2026·3mcloud-rungemini-api
GCP

Gemini Enterprise Agent Platform preview — console RL fine-tuning and deferred execution tier for agent workloads

Gemini Enterprise Agent Platform preview adds console RL fine-tuning and a deferred execution tier, forcing platform teams to treat model lifecycle, scheduling, and cost ops.

Sep 18, 2026·3mgoogle-cloudgemini-enterprise
GCP

GCP Cloud Run: monthly spend caps and persistent instances for predictable serverless pricing

Cloud Run supports enforceable monthly spend caps and a persistent instance pricing option, now making serverless agent workloads more cost-predictable.

Sep 16, 2026·3mcloud-rungemini-enterprise
GCP

Google Gemini API key migration: move from unrestricted/standard API keys to service-account-backed auth keys

Google Gemini API is deprecating unrestricted and standard API keys in favor of service-account-backed auth keys; migrate by mid-June and September 2026.

Sep 15, 2026·3mgoogle-cloudgemini-api
GCP

GKE defaults, gateway.dev hostname change for model-routing gateways, and App Topology billing

Cloud: model-routing gateways now use GATEWAY_ID-PROJECT_NUMBER.REGION.gateway.dev. App Topology API billing changes and GKE defaults updated—plan accordingly.

Sep 14, 2026·3mgkecloud-run
GCP

Gemini Pro preview in Vertex AI, Flash‑Lite rollout, and Cloud Run worker pools GA

Google previewed a Gemini Pro in Vertex AI and rolled Flash‑Lite into Vertex AI and the Gemini API. Cloud Run worker pools GA brings always‑on inference options.

Sep 13, 2026·3mgeminivertex-ai
GCP

GKE rapid channel 1.36.4-gke.1082000: default for new clusters; older rapid and alpha builds removed

New GKE rapid-channel clusters default to 1.36.4-gke.1082000; older rapid/alpha builds were removed. Pin cluster versions, audit CNI and topology billing ASAP.

Sep 11, 2026·3mgkekubernetes
GCP

Google Cloud: 50% Provisioned Throughput Credit for Gemini Flash Through 2026; GKE 1.37 in Rapid

Google Cloud gives a 50% Provisioned Throughput credit for Gemini Flash through Dec 31, 2026, and adds GKE 1.37 to the Rapid channel, changing TCO now.

Sep 10, 2026·3mgoogle-cloudgemini-api
GCP

Cloud Run Instances (Preview): long-lived, addressable workloads and per-instance pricing

Cloud Run Instances (Preview) expose long‑lived, addressable workloads with per‑instance billing (1 vCPU + 1 GiB steady cost). Rethink always‑on agents.

Sep 9, 2026·3mcloud-rungke
GCP

Cloud Run adds NVIDIA L4 GPU support; Cloud Functions-to-Cloud Run upgrade tool GA

Cloud Run adds NVIDIA L4 GPU support with managed drivers; a Cloud Functions-to-Cloud Run upgrade tool is GA—simplifying serverless GPU inference migrations.

Sep 8, 2026·3mgcpcloud-run
GCP

GKE 1.36: Dataplane V2 Emits CNI cniVersion 1.1.0 — Upgrade Risk for CNI Plugins

Dataplane V2 in GKE 1.36 emits CNI configs with cniVersion 1.1.0. Plugins lacking 1.1.0 semantics can fail to set up pod networking during upgrades — validate.

Sep 6, 2026·3mgkekubernetes
76 more