cd/entity/Cast AI· home› entities› Cast AI
grep -l @cast ai /news/*.json | wc -l → 58

Cast AI

mentions 58 type Organization page 1/3 feed RSS

// recent coverage 58 mentions

08:35
2026-10-07
cloudnativenow.com
ai-infrastructure

AKS Is Growing – Along with Its Kubernetes Utilization Problem

Average GPU utilization across Kubernetes clusters measured in Cast AI's 2026 State of Kubernetes Optimization Report was just 5%, and only 2% on Azure Kubernetes Service (AKS), meaning 98 cents of ev…

15:10
2026-09-21
cast.ai
ai-agents

What Is AI SRE, and Where Does Cost Automation Fit?

Gartner formalized AI SRE as an analyst category in January 2026 and projects that 85% of enterprises will use AI SRE tooling by 2029, up from less than 5% today, according to its Market Guide for AI …

12:14
2026-09-14
cast.ai
ai-infrastructure

AKS Cost Optimization: A Guide to Reducing Spend in 2026

AKS clusters average just 8% CPU utilization, 20% memory utilization, and 2% GPU utilization, the lowest of any major cloud provider, according to Cast AI's 2026 State of Kubernetes Optimization Repor…

12:22
2026-09-11
cast.ai
ai-infrastructure

GKE Cost Optimization: A Guide to Cutting Waste in 2026

GKE clusters average just 8% CPU and 6% GPU utilization, while EKS clusters average 5% GPU and AKS clusters average 2% GPU, according to the Cast AI 2026 State of Kubernetes Optimization Report, indic…

12:13
2026-09-11
cast.ai
ai-infrastructure

EKS Cost Optimization: The Engineer’s Guide for 2026

Average CPU utilization across Amazon Elastic Kubernetes Service (EKS) clusters has fallen to 8% in 2026 from 10% in 2025, with 69% of provisioned CPU now completely unused, up from 40% year-over-year…

20:12
2026-08-12
byteiota.com
artificial-intelligence

Kubernetes 1.37: DRA Slices One GPU Into Many Pods

Kubernetes 1.37, releasing August 26, brings pod-level resources to stable, DRA device taints to GA, and advances GPU partitioning, addressing the 5% average GPU utilization measured by Cast AI across…

13:30
2026-08-10
cast.ai
artificial-intelligence

LLM Inference Cost Optimization: Run AI Inference for Less

Cast AI benchmark testing shows that continuous batching at batch size 8 reduces Llama 3.1 70B inference cost on a single H100 from approximately $0.60-$0.80 per million tokens to $0.15-$0.25 per mill…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics