Your Agent Is Starving
GitHub has moved to per-token billing, ending the era of flat-rate pricing that absorbed AI agent inefficiency. A team of five developers using structured agent tooling spends $3,000 per month on AI costs while writing z…
AI Infrastructure news and analysis on Web Pulse: 32775 curated articles tracking the latest AI Infrastructure developments, tools, and research, updated continuously from vetted sources.
GitHub has moved to per-token billing, ending the era of flat-rate pricing that absorbed AI agent inefficiency. A team of five developers using structured agent tooling spends $3,000 per month on AI costs while writing z…
Anthropic's Claude Code service suffered a 90-minute outage on Tuesday, reducing its uptime to just one nine over the past 60 days. Meanwhile, GitHub CTO Vlad Fedorov published a second blog post in six weeks addressing …
Stripe announced 288 new products and features at its annual Sessions conference on Tuesday, including an expanded Agentic Commerce Suite with partnerships with Meta and Google that enables native checkout inside Faceboo…
Together AI has launched DeepSeek-V4 Pro, a 1.6T-parameter Mixture-of-Experts model with a 512K-token context window, priced at $2.10 per million input tokens and $4.40 per million output tokens. The model supports three…
PageIndex, an open-source project by Vectify AI, ranked #14 in GitHub Star Growth and #38 Overall on the Open Source Growth Index (OSSCAR) Q1 2026, published by Supabase and Commit VC. Vectify AI was also recognized in t…
Google has developed new networking technologies for its TPU 8 series AI accelerators, including the Boardfly inter-chip interconnect and the Virgo datacenter-scale Ethernet fabric, to support both inference and training…
NVIDIA announced that manufacturers are shifting from traditional design-build-test cycles to simulation-first workflows using OpenUSD and NVIDIA Omniverse. ABB Robotics achieved 99% simulation-to-real accuracy in its Ro…
Intel and AMD are jointly developing Advanced Performance Extensions (APX), a major expansion of the x86 instruction set that doubles the number of general-purpose registers from 16 to 32 and adds conditional instruction…
Microsoft and OpenAI have amended their seven-year partnership, ending exclusivity clauses that allowed each to work only with each other on AI development and cloud services. Under the new terms, OpenAI can now run its …
Together AI has made NVIDIA's Nemotron 3 Nano Omni model available on its platform, giving developers immediate access to a single open model that reasons across video, images, audio, and language. The 30-billion paramet…
Intel's data center CPU sales surged as AI inference demand shifted system designs from 8-to-1 GPU-to-CPU ratios to 4-to-1 or even 2-to-1 configurations, increasing CPU content per AI system. Intel CFO Dave Zinsner revea…
AWS and Anthropic deepened their product collaboration this week, with Anthropic now training its most advanced foundation models on AWS Trainium and Graviton infrastructure and launching Claude Cowork within Amazon Bedr…
AMD and Intel, through the x86 Ecosystem Advisory Group, are standardizing new matrix math instructions called ACE (AI Computation Extensions) to accelerate AI workloads on x86 processors. ACE introduces outer product op…
Meta's FAIR team documented a series of training failures in 2021 for their OPT-175B model, including repeated loss explosions and learning issues that required extensive hyperparameter tuning and architecture swaps. In …
Anthropic capped a user's Claude Max plan at 15 daily routines after a scheduled batch job failed at 2:36 AM, revealing the platform's hidden constraints on heavy usage. The user calculated a $137 API run for five cities…