The Copper Chokepoint
A structural copper deficit is emerging as demand from AI, electrification, and defense outpaces supply, with S&P Global projecting a 10-million-ton gap by 2040. The scramble for new supply is driving Western investment …
AI Infrastructure news and analysis on Web Pulse: 31608 curated articles tracking the latest AI Infrastructure developments, tools, and research, updated continuously from vetted sources.
A structural copper deficit is emerging as demand from AI, electrification, and defense outpaces supply, with S&P Global projecting a 10-million-ton gap by 2040. The scramble for new supply is driving Western investment …
A one-person AI business faces median gross margins of 52% versus 70-80% for mature SaaS, with 6.1% monthly churn on low-priced products and inference costs falling 50x per year. Over 54% of indie AI products make zero r…
A new open-source agent specification called AgentAz aims to reduce alert noise by ranking alerts based on actionability and recommending suppression rules, while never suppressing alerts linked to real incidents or crit…
A new open-source AI incident response agent, AgentAz, has been released under Apache-2.0, designed to autonomously handle low-risk remediation steps while escalating high-risk actions to human engineers. The agent opera…
A new AI-powered SOC alert triage agent, governed by the open-source AgentAz specification, enriches, correlates, scores, and recommends dispositions for raw security alerts. It reduces alert fatigue by deduplicating inc…
DeepSeek raised $7.4 billion in Series A funding led by Tencent, with CATL as a major investor, signaling a shift in Chinese AI funding toward non-ecosystem players as Alibaba and ByteDance sat out. The record-breaking r…
FERNme, a new user-owned memory layer for AI agents, updates memories with zero LLM calls using a Hebbian co-occurrence rule, keeping token costs flat and enabling users to see, edit, and own their data. The system achie…
Cloudflare launched Temporary Accounts for AI agents on June 19, 2026, allowing autonomous agents to deploy Workers via a single CLI flag without OAuth. Unclaimed accounts self-delete after 60 minutes, marking the first …
Cisco Foundation AI open-sourced FAPO, a Claude Code-driven system that autonomously optimizes multi-step LLM pipelines by attributing failures at the step level and proposing variants. In evaluations, FAPO outperformed …
Google DeepMind released Gemma 4 12B, an encoder-free multimodal model that processes visual and audio inputs directly through a single decoder-only transformer, enabling local agentic workflows on standard 16GB laptops.…
A developer argues that the evaluation and observability layers should be part of an agent's harness, not external tools. Using the formula Agent = Model × Harness, they explain that closed-loop agents require built-in l…
Armorer Labs has developed a pattern for agentic browser work that separates the agent planning loop from a control plane managing tool permissions, policy, human approvals, and run receipts. The approach emphasizes stru…
A developer detailed the architectural differences between NVIDIA's Ampere and Hopper GPU architectures, focusing on tensor core and memory bandwidth improvements. The Hopper architecture introduces Tensor Core 2.0 with …
A developer successfully ran a 35-billion-parameter Mixture-of-Experts model on a 2017 AMD RX 580 8GB GPU using Vulkan, bypassing CUDA and ROCm. The project, called Polaris Revival, achieved 17-18 tokens per second for L…
A new scheduling technique called TurboPrefill reduces waiting time for Vision Language Models by nearly half, from 9.0 to 4.6 seconds, without changing model weights or architecture. The optimization, validated on Qwen2…
MicroPhase Technology announced the AntSDR T510 AI, a single-board platform combining an AMD Zynq UltraScale+ RFSoC with an NVIDIA Jetson module for RF capture and GPU-accelerated AI inference. The board, listed as comin…
ARK Invest CEO Cathie Wood purchased 223,690 shares of Snowflake worth $52.45 million and $22 million in Tesla on June 18, signaling a capital rotation toward AI-adjacent growth stocks. The firm also trimmed its Roku pos…
Alloy, a new PyTorch backend and inference engine for Apple Silicon, has been released as a technical preview. The open-source project compiles Python GPU kernels to Metal and supports LLM serving with a drop-in compatib…
A developer spent two weeks optimizing a homelab with four RTX 3090s (96GB VRAM) for local LLM inference, achieving improvements like 40% throughput gain and 4x VRAM savings, but ultimately found that paid APIs were more…
A developer documented a field repair for a Codex Desktop update on Windows that broke browser extension connectivity and Computer Use. The fix involved patching the node_repl shim to handle missing sandbox metadata and …