NVIDIA Rosa Eyes TSMC 2nm/A16 Backside Power
NVIDIA's next-generation Rosa CPU may use TSMC's A16 process with back-side power delivery for a 2028 data-center platform, according to TrendForce and Wccftech reports. The technology could improve p…
NVIDIA's next-generation Rosa CPU may use TSMC's A16 process with back-side power delivery for a 2028 data-center platform, according to TrendForce and Wccftech reports. The technology could improve p…
A Fortran developer proposes adding a RESIDENT locality specifier to DO CONCURRENT to prevent implicit offload copying in systems with separate memory address spaces, allowing selective host-only para…
Ollama, an AI platform that simplifies running open-source models locally, has reached 8.9 million developers and is used by 85% of the Fortune 500. The company raised $65 million in Series B funding …
NVIDIA's closed-source cuda-checkpoint tool, which freezes and restores CUDA processes, suffers from slow PCIe transfers. Engineers reverse-engineered the tool to understand why checkpoint transfers f…
SpaceXAI released Grok 4.5, a general-purpose AI model trained alongside Cursor for coding, agentic tasks, and knowledge work, priced at $2 per million input tokens. The model claims superior token ef…
Xsight Labs is developing networking gear for AI and SpaceX satellites, with a focus on full switch programmability, on-path cores in DPUs, and open-sourcing the ISA. The company's 12.8T power-optimiz…
NVIDIA and Hugging Face are collaborating to bring the NVIDIA Isaac GR00T 1.7 model and Isaac Teleop framework to LeRobot, Hugging Face's open-source robotics library, aiming to provide developers wit…
A developer built SkillSpector Report, an open-source Python CLI that converts NVIDIA SkillSpector security scan output into PDF reports for easier sharing and review. The tool addresses the need for …
Samsung has begun mass production of the PM1763, its first PCIe Gen6 enterprise SSD, offering sequential reads up to 28,400 MB/s and writes up to 21,000 MB/s. The drive, built with 9th-generation V-NA…
Chinese food-delivery company Meituan trained a 1.6-trillion-parameter language model, LongCat-2.0, using approximately 50,000 domestic AI chips without any NVIDIA GPUs. The model activates only 48 bi…
NVIDIA released the Nemotron Post-Training v3 Prompt Atlas, an interactive embedding atlas for exploring post-training data used in AI agents. The company also highlighted its open data releases of ov…
Google Cloud has been named a Leader in the inaugural 2026 Gartner Magic Quadrant for AI Infrastructure, positioned highest for Ability to Execute and furthest for Completeness of Vision. The recognit…
DeepInfra opened a 1.7 MW Toronto data center on July 8, its ninth site and first outside the United States, hosting over 1,000 NVIDIA Blackwell B300 GPUs for low-latency AI inference. The expansion s…
NVIDIA's Nemotron 3 Ultra, tuned with LangChain's Deep Agents harness, achieved benchmark-leading performance at 10x lower inference cost than top closed models, enabling enterprises to build speciali…
NVIDIA and LangChain released a tutorial on creating a LangChain Deep Agents harness profile for NVIDIA Nemotron 3 Ultra to match proprietary frontier model intelligence. The profile optimizes agent p…
Scality is enhancing its object storage to support AI workloads by integrating with NVIDIA's NIXL library, enabling S3 over RDMA to bypass CPU bottlenecks and move data directly between flash and GPU …
Samsung Electronics has begun mass production of the PM1763, its first PCIe 6.0 enterprise SSD designed for AI and HPC servers. The drive delivers sequential read speeds up to 28,400 MB/s and write sp…
NVIDIA's NIM platform now offers free API access to Z.ai's GLM-5.2, a 753-billion-parameter open-source model with a 1,000,000-token context window, removing major infrastructure barriers for long-hor…
On April 14, 2026, NVIDIA launched NVIDIA Ising, a family of open-source AI models designed to solve quantum computing's scaling problems in calibration and error correction. The models, released unde…
DDN launched Infinia 2.4 at the RAISE Summit in Paris, adding multi-tenancy, governance, identity management, and POSIX support to its AI data platform for production AI factories. The update targets …