Associate Cloud Engineer Certification Path
A developer is documenting their journey toward Google's Associate Cloud Engineer Certification, highlighting the certification path provided by Google Skills and MentorMe Collective. The path include…
A developer is documenting their journey toward Google's Associate Cloud Engineer Certification, highlighting the certification path provided by Google Skills and MentorMe Collective. The path include…
Google Kubernetes Engine (GKE) can reduce the cost per AI agent by over 30% and increase agent density by more than 40% per vCPU by using GKE Agent Sandbox with gVisor instead of microVMs, according t…
Google is overhauling its global data center architecture to support autonomous AI agents, which require persistent, multi-step processing. The company's facilities now process over 3 quadrillion toke…
Anthropic released Claude Opus 5 on July 24, making it the default model on Claude Max and the strongest model on Claude Pro, with pricing at $5 per million input tokens and $25 per million output tok…
Ray AI libraries (Serve, Data, Train) now support Google TPU slices through a topology field that reserves a whole ICI-connected slice, preventing multi-host deployment hangs. Ray Serve serves LLMs vi…
Google Cloud has published a security blueprint for AI workloads on Google Kubernetes Engine, outlining a three-layer approach covering infrastructure, model integrity, and application security. The d…
Google Cloud has released a blueprint for securing AI workloads on Google Kubernetes Engine (GKE), consolidating controls across multiple services to address infrastructure, model supply chain, and ap…
Google Cloud demonstrated elastic training on TPUs, where a worker failure during multi-node LLM training was recovered in under two minutes without restarting the job. Using the JAX AI stack (MaxText…
Google Cloud introduced a multi-node KV cache offloading solution using GKE and Managed Lustre for large language model inference, achieving over 50% TCO savings and nearly 60% reduction in GPU-hour r…
Google Cloud and Anyscale announced optimizations for Ray Serve LLM on Google Kubernetes Engine (GKE) that deliver up to 5x higher throughput and 8x lower latency for large language model inference. T…
Ray Serve LLM, in partnership with Google Kubernetes Engine, announced major performance improvements achieving up to 4.4x higher throughput on prefill-heavy workloads and 24x higher on decode-heavy w…
Anthropic's Model Context Protocol (MCP) enables standardized context integration for LLMs. Google has published a guide for deploying a remote MCP server on Google Kubernetes Engine (GKE) in 30 minut…
Yahoo partnered with Google Cloud to build Seller Agent, an agentic digital media buying platform using Google Data Cloud graph technologies. The platform reduces multi-week manual campaign processes …
WALT Labs joined Anthropic's Claude Partner Network as a Registered Partner after completing the learning path and earning the Claude Certified Associate certification, enabling the company to help or…
Google Cloud's GKE Inference Gateway delivers up to 92.8% shorter wait times and 62.6% lower inter-token latency compared to the next leading managed Kubernetes service, according to an independent be…
A developer detailed how to build interrupt-resilient AI workloads on Google Kubernetes Engine (GKE) by handling Spot VM evictions. The approach involves catching the SIGTERM signal sent by Kubernetes…
Google engineers deployed a Gemma 3 large language model across two GKE clusters in separate regions, using four TPU v6e chips per cluster, to test multi-cluster failover with the Inference Gateway. T…
General availability of GKE Agent Sandbox, an open-source, Kubernetes-based execution environment designed for secure and scalable AI agent workloads, featuring capabilities like pod snapshots for idl…
At the Google I/O conference, NVIDIA and Google Cloud announced new resources for their joint developer community of over 100,000 members, including learning paths for JAX on NVIDIA GPUs and an NVIDIA…