The New Apple Silicon Lineup #
Apple isn't playing around with the professional tier either. Here is the breakdown of the new hardware:
Mac mini (M6 & M5 Pro): Starts at 6,999 CNY for M6 and 12,999 CNY for M5 Pro. The M6 features a 12-core CPU, 12-core GPU, and a dual 16-core Neural Engine. Apple claims a 4x boost in AI performance compared to the M4.Mac Studio (M5 Max & M5 Ultra): Starts at 19,999 CNY for M5 Max and 46,999 CNY for M5 Ultra. The M5 Ultra is a beast, utilizing a quad-die architecture with up to 36 CPU cores, 80 GPU cores, and a massive 1.2TB/s memory bandwidth.
If you are transitioning from an older Intel Mac or even an early M1, the leap in unified memory bandwidth and AI throughput is going to be massive for running large local models.
OpenAI's Jalapeño chip might actually beat NVIDIA #
While Apple is upgrading the edge, OpenAI is making moves on the data center side that should make NVIDIA nervous. The first performance benchmarks for OpenAI's in-house inference chip, Jalapeño, are out, and the numbers are staggering. According to tests conducted with SemiAnalysis on the InferenceX benchmark, Jalapeño is outperforming the GB200 and GB300 in several key areas:
Throughput: On GPT-OSS 120B, it hit 1459 tokens/s, which is 2.7x faster than the GB200.Latency: OnDeepSeekR1 670B, end-to-end latency was just 1.65s, compared to 5.99s for the GB300.Efficiency: The throughput per watt on Kimi K2.5 1T saw a 1.5x improvement.
What’s even more impressive is how they built it. They used their own models, GPT-Astra and Codex, to accelerate the circuit design and verification process. This is a perfect real-world example of an AI workflow accelerating AI hardware development. They managed to reduce the area of SIMD units by 8% and matrix engines by 10% just by using AI-generated kernels.
ByteDance launches "Doubao Work" for seamless Agent workflows #
On the software side, ByteDance has officially released "Doubao Work." This is a significant step for LLM agents in a corporate environment. The standout feature is the ability for the Agent to operate within a virtual desktop and, more importantly, read the context from Feishu (Lark).
This solves one of the biggest pain points in prompt engineering for enterprise: context fragmentation. Instead of manually copying and pasting data from chats into a prompt, the agent can actually "see" the workspace context to perform tasks. It's a massive leap toward true autonomous AI agents in the workplace.
Other quick hits in the tech ecosystem #
Autonomous Driving Laws: New draft legislation suggests that for fully autonomous driving, the manufacturer (not the driver) will be held liable for traffic violations. This is a huge shift for the industry.OpenAI Subscription Update: Plus users will see a return of the 5-hour usage window forChatGPTWork and Codex starting August 26th to help manage compute loads.Memory Market Shifts: SK Hynix is shutting down its official flagship store on Taobao, though distributors claim existing warranties will remain handled by individual sellers.Robotics Progress: A subsidiary of Agibot (智元) just dominated a humanoid robot competition, taking home 7 gold medals.
It feels like we are hitting a convergence point where specialized AI hardware (like Jalapeño) and context-aware software agents (like Doubao Work) are finally starting to match the massive scale of the models themselves.
Nvidia Jetson Orin is being used in combat drones in Ukraine 5h ago
Apple's M6 and M5 Ultra chips are going to redefine local AI 22h ago
LLMs might finally let us backseat drive autonomous vehicles 1d ago
Apple's new AI integration might actually compromise your 3d ago
Why China's AI deployment looks so different from the West 3d ago
Apple's next AirPods might pack cameras — here's why that 6d ago
Next Tesla Denies Shanghai Data Center Shutdown →