cd/sources/martinalderson-auto-discovered· home sources Martinalderson (auto-discovered)
cat /sources/martinalderson-auto-discovered.feed | wc -l → 9

Martinalderson (auto-discovered)

articles 9 domain martinalderson.com → feed RSS
00:00
2026-08-16
martinalderson.com
artificial-intelligence

How I think about reducing AI costs

AI inference costs are becoming a major problem for many companies, with AI spend per employee per month rising sharply, according to the Ramp AI Index via a16z. To reduce costs, companies should audi…

00:00
2026-08-10
martinalderson.com
artificial-intelligence

Watch out for cache read costs

Cache read costs now dominate agentic AI workloads, accounting for up to 81.6% of total inference spend in a 100-turn session, according to an analysis by an unnamed author. The analysis shows that fo…

12:11
2026-07-16
martinalderson.com
artificial-intelligence

Winners and losers in the coming AI margin collapse

The AI market is bifurcating into expensive frontier models and cheap 'good enough' models, with Grok 4.5 priced at $6/MTok output signaling a margin collapse, according to analyst Ben Thompson. Semic…

00:00
2026-07-12
martinalderson.com
artificial-intelligence

Winners and losers in the coming AI margin collapse (part 2)

The AI market is bifurcating into expensive frontier models and cheap 'good enough' models, with Grok 4.5's aggressive pricing signaling a margin collapse. Hardware suppliers like semiconductor compan…

20:15
2026-07-06
martinalderson.com
large-language-models

GLM 5.2 and the coming AI margin collapse

GLM 5.2, an open-weights model from Z.ai, has emerged as a genuine competitor to frontier models like Opus and GPT, but its slow inference speed and lack of vision and web search capabilities limit it…

00:00
2026-07-06
martinalderson.com
artificial-intelligence

GLM 5.2 and the coming AI margin collapse (part 1)

GLM 5.2, an open-weights model from Z.ai, has emerged as a genuine competitor to frontier models like Opus and GPT, but its slow inference speed and lack of vision and web search capabilities limit it…

00:00
2026-06-15
martinalderson.com
large-language-models

A brief history of KV cache compression developments

The memory needed to store one token of context in large language models has fallen by roughly 100x since 2017, driven by advances in KV cache compression techniques such as Multi-Query Attention (MQA…

00:00
2026-06-08
martinalderson.com
artificial-intelligence

xAI is looking more like a datacentre REIT than a frontier lab

XAI has formed partnerships with Anthropic and Google to provide massive datacenter capacity, with fees reaching $1.25 billion per month for 300 MW of capacity. The deals, which include cancellation c…