cd/entity/GKE· home entities GKE
grep -l @gke /news/*.json | wc -l → 22

GKE

mentions 22 type Organization page 1/2 feed RSS

// recent coverage 22 mentions

13:00
2026-08-05
cloud.google.com
ai-infrastructure

Unlocking the future of shared storage: Filestore on Colossus

Google Cloud announced that Filestore, its first-party NFS file service, now incorporates a cloud-native backend storage layer built directly on Colossus, Google's foundational distributed storage sys…

16:00
2026-07-31
cloud.google.com
ai-infrastructure

What’s new in AI infrastructure and orchestration this month

Google Cloud announced the general availability of Managed Lustre, a high-performance storage solution powered by DDN's EXAScaler, offering throughput from 125 MB/s to 1000 MB/s per TiB and scaling up…

09:53
2026-07-23
cast.ai
artificial-intelligence

Multi-Cloud and Cross-Region GPU Capacity for Kubernetes AI

GPU utilization across production Kubernetes clusters averages just 5%, yet half of GPU cloud providers surveyed by SemiAnalysis report being completely sold out of H100 and H200 capacity, according t…

16:25
2026-07-20
developers.googleblog.com
ai-infrastructure

Run Ray on TPU, Part 1: The foundations

Ray 2.55 makes Google Cloud TPUs a first-class accelerator, with official pre-built images and support across core libraries. The distributed-computing framework now treats TPU slices as schedulable r…

00:00
2026-07-16
modelplane.ai
ai-infrastructure

Modelplane v0.2: more clouds, and traffic you can direct

Modelplane released v0.2 of its open-source inference orchestration platform, adding support for Nebius and Azure AKS cloud providers, weighted request routing for canarying and traffic splitting, and…

16:31
2026-06-30
cast.ai
artificial-intelligence

Karpenter vs Cluster Autoscaler: Which to Use in 2026

Karpenter, an open-source node provisioner, directly provisions cloud instances without node groups, while Cluster Autoscaler scales pre-defined node groups. Karpenter is faster, cheaper, and more fle…

12:40
2026-06-30
dev.to
ai-agents

Cutting Idle Agent Costs by 90% with Agent Substrate

Agent Substrate, a new platform for running AI agents, reduces idle agent costs by 90% compared to traditional Kubernetes Pod deployments. By using an actor model with checkpoint/restore and gVisor, i…

01:53
2026-06-18
letsdatascience.com
large-language-models

Google releases OpenRL for LLM fine-tuning

Google released OpenRL, an open-source API for fine-tuning large language models on Kubernetes clusters, aiming to decouple infrastructure from AI research and improve GPU utilization by running multi…

09:00
2026-06-17
pydantic.dev
developer-tools

The pod that did not survive Tuesday's deploy

Pydantic Logfire launched a new Kubernetes cluster view that surfaces pod restart loops, memory limits, and rollout failures from a single page. The feature uses standard OpenTelemetry receivers to ag…

07:02
2026-06-17
dev.to
artificial-intelligence

Prompt Engineering Patterns for SRE Playbooks and Postmortems

An engineer proposes treating prompt engineering for site reliability engineering (SRE) workflows as reusable infrastructure, introducing three production-tested patterns: structured context injection…

22:21
2026-06-02
letsdatascience.com
ai-infrastructure

Google Cloud Demonstrates Multi-Cluster TPU Inference Setup

Google Cloud demonstrated a multi-cluster inference setup deploying an LLM across two regional GKE clusters using TPU v6e accelerators and GKE managed DRANET for networking. The experiment, documented…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics