cd /news/artificial-intelligence/this-ai-entrepreneur-is-developing-a… · home topics artificial-intelligence article
[ARTICLE · art-123172] src=technologyreview.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

This AI entrepreneur is developing agents that can plan ahead for the unexpected

Danijar Hafner, a 31-year-old AI researcher and former Google DeepMind scientist, is developing a stealth startup in San Francisco that uses model-based reinforcement learning to create AI agents capable of planning ahead and navigating unfamiliar environments, with the goal of enabling robots to operate in human spaces. Hafner's technique, which trains agents within world models that simulate physical reality, has produced systems like Dreamer 3, the first to solve the Minecraft Diamond challenge, and Dreamer 4, which learned from offline video data. His former manager Timothy Lillicrap says Hafner 'easily sits in the top half of 1%' of Google researchers.

read4 min views1 publishedSep 8, 2026
This AI entrepreneur is developing agents that can plan ahead for the unexpected
Image: MIT Tech Review AI

Danijar Hafner has a startup in stealth and a long track record of teaching AI agents about our world.

Danijar Hafner’s office in San Francisco’s SoMa district sits mostly empty. His brand-new startup is still in stealth mode and doesn’t even have its name on the door. On the day I visit, there’s only one other person there, and little in the way of furniture. But what it lacks in decor, it makes up for in robots. Humanoids of various shapes and sizes hang like marionettes from racks that run down the center of the wide-open space.

While Hafner, 31, won’t say too much about his new venture just yet, he describes it as a continuation of his longtime work to enable AI to navigate environments it has not encountered in training. The humanoids, which he imports from China, are the next evolution of this work—and its physical embodiment. Their ability to react in previously untested scenarios will be key to getting robots into human spaces. Because if you want to send a robot into a person’s home, for example, it needs to be able to handle a floor plan and furniture it’s never seen before. To achieve this, Hafner relies on something called model-based reinforcement learning. He develops world models—AI models designed to emulate physical reality—and trains agents within them. The agent essentially treats the model as a real-world simulation and learns how to act there. It then uses those experiences to make predictions (to dream or imagine, Hafner might say) about future outcomes. That allows agents—or the robots they’re embedded in—to navigate unfamiliar situations IRL.

“I get to interact with a lot of really smart people in research at Google, and he easily sits in the top half of 1%.”

Unlike other efforts, Hafner’s technique enables agents and the robots they control to execute massively complicated tasks without the real-world trial-and-error training that’s traditionally been used in robotics.

Hafner grew up in a rural town in northeastern Germany, where his parents were both classical musicians. He learned programming from a neighbor, and in high school he began taking online courses about AI, which quickly developed into a passion. “I was always fascinated with how thinking works,” he says. AI offered him a way to emulate it on a computer.

In 2015, as a second-year undergraduate studying engineering at Hasso Plattner Institute in Potsdam, he won a role as a student researcher at Google Brain. From there, he went on to a dozen internships and other positions at the company, including stints with Google Brain and Google DeepMind (the two have since merged under DeepMind) in the UK, Canada, and the US. He worked with industry legends including Geoffrey Hinton, who is often referred to as one of the godfathers of AI, and Ashish Vaswani, coauthor of the groundbreaking research paper “Attention Is All You Need,” which described the transformer technology used by today’s large language models.

One of Hafner’s former managers and coauthors at Google, Timothy Lillicrap, describes him as a standout among standouts. “I get to interact with a lot of really smart people in research at Google, and he easily sits in the top half of 1%,” Lillicrap says. “In many cases he would build, single-handedly, things it would take entire teams of engineers to build.”

Over the years, Hafner has honed and proved his approach by pitting agents trained within his world models against popular video games. His first breakthrough was PlaNet, a model that allowed agents to execute actions by planning ahead. His Dreamer 2 was the first agent to hit human-level performance playing Atari 2600 games using a world model. Dreamer 3 was the first one to solve the Minecraft Diamond challenge—successfully mining in-game gems on its own. And Dreamer 4 went a step beyond that by learning to mine diamonds from an offline data set of recorded game-play videos, without ever interacting with the game directly.

More recently, he’s begun to migrate his agents out of the virtual world and into physical reality. His DayDreamer project used the Dreamer algorithm to let robots operate themselves in novel environments and react to new experiences (such as being pushed over) without any specific training.

Today, Hafner is working on his new startup, which he left Google DeepMind to form in the fall of 2025. Though he’s coy about his next steps, it’s clear he’s dreaming big: “I was interested in solving a problem,” he hints, “that would change the world.”

Deep Dive

Artificial intelligence

A fundamental flaw leaves LLMs strikingly vulnerable to attack

It makes it easy to trick them into doing things they shouldn’t, such as telling you how to sabotage an aircraft’s navigation system.

Anthropic found a hidden space where Claude puzzles over concepts

A new technique has let the company probe deeper than ever into the weird workings of an LLM.

AI is more likely than humans to form biases when hiring

AI doesn’t just learn stereotypes from its training. It can cook up new ones, too.

Here’s why AI agents lie and cheat to reach their goals

The misbehavior is called reward hacking. This is what you need to know.

Stay connected

Get the latest updates from #

MIT Technology Review

Discover special offers, top stories, upcoming events, and more.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @danijar hafner 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/this-ai-entrepreneur…] indexed:0 read:4min 2026-09-08 ·