DeepSeek V4 Flash now runs from a backpack
A community tester ran DeepSeek-V4-Flash-0731 on a Bosgame M5 mini PC with an RTX PRO 6000 Max-Q eGPU, achieving 44 to 60 tokens per second depending on quantisation, with no hyperscaler involved. The…
A community tester ran DeepSeek-V4-Flash-0731 on a Bosgame M5 mini PC with an RTX PRO 6000 Max-Q eGPU, achieving 44 to 60 tokens per second depending on quantisation, with no hyperscaler involved. The…
Taiwan Semiconductor Manufacturing Co. reported Q2 2026 revenue of $40.2 billion, a 36% year-over-year increase, and net profit of NT$706.56 billion ($21.99 billion), up 77.4%, while raising its 2026 …
The European Commission has launched a call for tenders to establish up to 7 AI Gigafactories across the EU, backed by up to €10 billion in combined EU and national funding, with the goal of attractin…
A security researcher has published a detailed attempt to extract Google DeepMind's SynthID image watermark detector and explore adversarial machine learning attacks against it, revealing that the inv…
University of Toronto researchers unveiled GPUBreach at Black Hat USA 2026, a Rowhammer-based attack on NVIDIA GDDR6 GPU memory that escalates to a root shell on the host system in under 20 seconds, b…
DeepSeek founder Liang Wenfeng, in a leaked 3-hour-44-minute investor meeting transcript, revealed the lab's 'Costco strategy' for achieving AGI: pricing its API so GPU hardware pays for itself in ten…
A cost analysis across 24 providers and 378 GPU rental rates found that the price of one million output tokens ranges from $0.09 on a single AMD MI355X to $290.12 on eight NVIDIA H100s, with the gap d…
MSI's Vector 16 HX AI gaming laptop, featuring Intel's Core Ultra 9 275HX processor and NVIDIA's GeForce RTX 5080 Laptop GPU with 16GB of dedicated memory, is available at a discounted price on Amazon…
PCIe Gen6 servers will begin shipping in the second half of 2026, doubling per-lane transfer rates over Gen5 and enabling 800Gbps network bandwidth from a single x16 slot, but Gen5 storage will remain…
Fireworks AI, a Series D company valued at $17.5 billion, is hiring an AI Product Engineer for its Fireworks Nexus platform, which routes AI workloads from expensive proprietary models to open models …
DeepSeek's V4 Flash 0731, a 284-billion-parameter open-weight mixture-of-experts model released under the MIT license on July 31, runs at 107 tokens per second on DeepSeek's API and 267 on the fastest…
Alphabet, Meta Platforms, Microsoft, and Amazon are collectively investing nearly $2.4 trillion in AI infrastructure, with Morgan Stanley projecting a $1.5 trillion financing gap that private credit m…
NVIDIA has published a technical post analyzing how group size, head dimension, and sequence length affect dense attention performance in long-context inference, offering a co-design checklist to help…
OpenAI published a technical deep-dive revealing that GPT-5.6 Sol's low ARC-AGI-3 score was due to the evaluation harness discarding reasoning and truncating context, not model weakness. Enabling reta…
Amazon finalized its $50 billion investment in OpenAI on July 31, releasing the remaining $35 billion tranche after OpenAI met undisclosed performance milestones, as confirmed by an SEC filing. The de…
KUKA AG has deployed its Automation Management Platform (AMP) at KUKA Toledo Production Operations in Ohio, a facility that produces over 300 vehicle bodies per day and has generated more than 2 milli…
NVIDIA Corporation CEO Jensen Huang's 2024 claim that its Blackwell GPUs would reduce large language model inference operating costs and energy by up to 25x has been pared back to 10x, a 150% overstat…
A team of four developers won third place at the Dell x NVIDIA AI Hackathon by building Squidward, a 100% air-gapped, self-improving IT firewall that runs locally on a Dell GB10 with a Qwen3.6-27B mod…
Nuro, a Mountain View-based autonomous driving company founded in 2016, is hiring a Senior/Staff Software Engineer for AI Agent Infrastructure to build a platform that enables AI agents to operate aut…
NVIDIA released Video Codec SDK 13.1, adding AV1 Hierarchical Reference Mode with up to 31 B-frames, per-macroblock decode statistics for H.264 and HEVC, frame-accurate seek, and application-allocated…