{"slug": "only-7-of-organizations-using-kubernetes-for-ai-do-so-daily", "title": "Only 7% of Organizations Using Kubernetes for AI Do So Daily", "summary": "Only 7% of organizations running AI workloads on Kubernetes deploy them daily, while 66% already deploy AI on Kubernetes and 47% do so occasionally, according to figures Apple Principal Engineer and CNCF Technical Oversight Committee member Katie Gamanji presented at Infobip Shift 2026. Gamanji said Dynamic Resource Allocation reached general availability in Kubernetes 1.34, adding a native way to request GPU and TPU capacity for AI workloads, and that the 7% deploying daily \"most likely implemented an automated retraining pipeline that treats the model as a dynamic component rather than a static element.\" The CNCF's AI Conformance Working Group is defining what platforms must support to run AI workloads reliably, with portability as the goal.", "body_md": "# Only 7% of Organizations Using Kubernetes for AI Do So Daily\n\nYesterday at [Infobip Shift 2026](https://shiftmag.dev/tag/infobip-shift-2026/), I listened to Apple Principal Engineer and CNCF Technical Oversight Committee member Katie Gamanji talk about the **state of cloud native and its slow turn towards AI**.\n\nShe said cloud native spent years becoming stable and unexciting, and AI teams now need that stability more than new tools. The big question is whether the people around that infrastructure can move as fast as the platform itself.\n\nFor teams already using Kubernetes in production, the next change is very concrete, Katie said:\n\nDynamic Resource Allocation, or DRA, reached general availability in Kubernetes 1.34, adding a native way to request GPU and TPU capacity for AI workloads.\n\n## Kubernetes is now officially boring\n\nKatie started with a number that captures how far Kubernetes has come: “98% of organizations in the CNCF’s annual report said they had adopted cloud native technologies.” After a decade of development, Kubernetes has become mainstream:\n\nKubernetes is now described as being boring, and I think this is a wonderful achievement.\n\nThe problem appears when AI workloads enter that mature infrastructure. According to the figures Katie presented, **66% of organizations are already deploying AI workloads on Kubernetes**. But only 7% are doing it daily, while 47% deploy them occasionally. Of the organizations using Kubernetes for AI, 23% have fully adopted the Kubernetes stack and 43% have partially adopted it.\n\nThat 67 to 7 gap points to one likeliest cause. The talk did not name the cause directly, but the pattern Katie described points to operational immaturity as the likeliest one:\n\nTeams fine-tune an existing model rather than build one, deploy it once to prove it works, and lack the automated retraining pipelines that would treat the model as a dynamic component.\n\nThe 7% deploying daily, she noted, “most likely implemented an automated retraining pipeline that treats the model as a dynamic component rather than a static element.” **Running an AI workload is not as simple as putting another application into a container**. Teams need model artifacts, rollback strategies, and training pipelines. Kubernetes is being positioned as the shared platform, though whether teams adopt it that way is another question.\n\n## DRA and AI Conformance are making Kubernetes the AI platform\n\nKatie compared this to DevOps: Kubernetes once brought developers and operations together, and now it could do the same for infrastructure teams and data scientists. But the analogy has limits – DevOps took years to work, and it is still not fully settled everywhere.\n\n**Data scientists**, for their part, **rarely want to operate Kubernetes**; they want a model served. Still, a platform that can run both traditional workloads and training or inference jobs reduces the need for separate AI infrastructure. \n\nKubernetes is also adapting for this: DRA in 1.34 gives teams a more reliable way to allocate accelerators, which matters most for teams deploying daily.\n\nThe [AI Conformance Working Group](https://www.cncf.io/announcements/2025/11/11/cncf-launches-certified-kubernetes-ai-conformance-program-to-standardize-ai-workloads-on-kubernetes/) defines what a platform must support to run AI workloads reliably, with portability as the goal as Katie said: \n\nIf you have two conformant platforms, it’s going to be easy to lift and shift one product to the next platform.\n\nOther initiatives target batch workloads, inference performance, and AI integration, and an Agent Sandbox effort is designing stateful, isolated runtimes for AI agents. The Serving Working Group completed its milestones and archived itself in February, continuing as a SIG on inference performance.\n\nThe list is still changing quickly. Katie expects much of this landscape to look different within six months to a year.\n\n## The cloud native playbook AI can borrow\n\nKatie divided the emerging open source AI ecosystem into three areas: **training, inference, and agents**. Training turns data into a model. Inference serves it. Agents connect it to the outside world.\n\nFor platform leads, maturity matters. Training is led by the [PyTorch Foundation](https://pytorch.org/foundation/), inference has many options, and agents are still young under the [Agentic AI Foundation](https://aaif.io/).\n\n**Cloud native experience can help AI tools mature faster**. Security, observability, and identity were already solved in cloud native, and the CNCF project pipeline shows that kind of progress at scale.e.\n\nIt also prunes: 28 archived projects. For anyone choosing a stack, the archive list is a practical filter, and Katie argued archival is a healthy sign, letting maintainers redirect energy toward projects that earn adoption.\n\n## Teams that engage now will shape the patterns everyone else follows\n\nFor Katie, the next phase of cloud native is about whether the people building and maintaining the ecosystem can keep pace. She pointed out that contributing also means production feedback, feature requests, documentation, and white papers all shape projects.\n\nBut in the end Katie’s talk left the hard questions unanswered, but the data shows why:\n\n98% of organizations trust Kubernetes with their infrastructure, two thirds have tried running AI on it, and almost nobody operates it continuously.\n\nThe specific bottlenecks, daily operational maturity, DRA adoption, and the conformance baseline, are still being defined, which means teams deploying AI on Kubernetes today are writing the patterns everyone else will copy:\n\nAll of the working groups’ meeting notes, invites, and repositories are public on GitHub. The teams that engage now, before the patterns harden, will not have to retrofit someone else’s choices in two years.", "url": "https://wpnews.pro/news/only-7-of-organizations-using-kubernetes-for-ai-do-so-daily", "canonical_source": "https://shiftmag.dev/only-7-of-organizations-using-kubernetes-for-ai-do-so-daily-12086/", "published_at": "2026-09-15 11:35:26+00:00", "updated_at": "2026-09-15 11:41:43.130573+00:00", "lang": "en", "topics": ["ai-infrastructure", "ai-policy", "mlops", "ai-agents"], "entities": ["Kubernetes", "Katie Gamanji", "Apple", "CNCF", "Infobip Shift 2026", "Dynamic Resource Allocation", "AI Conformance Working Group", "Agent Sandbox"], "alternates": {"html": "https://wpnews.pro/news/only-7-of-organizations-using-kubernetes-for-ai-do-so-daily", "markdown": "https://wpnews.pro/news/only-7-of-organizations-using-kubernetes-for-ai-do-so-daily.md", "text": "https://wpnews.pro/news/only-7-of-organizations-using-kubernetes-for-ai-do-so-daily.txt", "jsonld": "https://wpnews.pro/news/only-7-of-organizations-using-kubernetes-for-ai-do-so-daily.jsonld"}}