How to Create Your Own Personal AI Benchmark
Every head of evaluations built a personal AI benchmark after Wharton professor Ethan Mollick argued that widely cited tests like MMLU-Pro measure the wrong things, asking models for the approximate c…
Every head of evaluations built a personal AI benchmark after Wharton professor Ethan Mollick argued that widely cited tests like MMLU-Pro measure the wrong things, asking models for the approximate c…
A Wharton Blueprint for AI Agent Adoption study of 3,000 people finds that only 9% of the 59% of enterprises using agentic AI have converted it into autonomous workflows, with middle managers seeing r…
ChatTJB, a satirical San Francisco platform created by former Google project manager Tucker Bryant, replaces AI with human volunteers and has processed over 100,000 queries since April, peaking at 5,0…
A study of Perplexity's Comet browser found that 55% of agentic actions were personal, 30% professional, and 16% educational, based on hundreds of millions of anonymized interactions from July to Octo…
OpenAI's first telemetry-based study, 'How Organizations Use AI: Evidence from ChatGPT,' analyzing 17 million messages from over 1,500 organizations, finds enterprise AI adoption is broad but shallow,…
PwC's US senior partner and CEO Paul Griggs said the firm hired fewer consultants and accountants this year than two years ago, a trend he expects to continue as it hires more data scientists, enginee…
A developer reflects on how learning to code without AI in 2013 built structural understanding through desirable difficulty, contrasting it with the experience of the first generation to grow up with …
Researchers from Wharton, Carnegie Mellon University and Harvard Business School introduced BusinessCaseBench, a benchmark showing frontier AI models produce strong answers across 18 open-ended busine…
OpenAI CEO Sam Altman said he deleted TikTok after becoming addicted to the app while studying it for the development of Sora, OpenAI's AI-generated video social network that was scrapped in March. Al…
New research shows that people with access to AI-generated answers are less willing to admit "I don't know," even when those answers are wrong, while intellectual humility protects against believing A…
Anthropic's head of economics Peter McCrory published an X essay arguing AI has caused no material rise in US unemployment, directly contradicting CEO Dario Amodei's warnings of a white-collar bloodba…
New research published last week reveals that people using AI advice experience a one-third drop in accuracy while their confidence more than doubles, according to a study by researchers at École Norm…
Google Cloud published the Open Knowledge Format (OKF) on June 12, 2026, a specification for representing organizational knowledge as markdown files with YAML frontmatter. The format is designed to be…
Lucy Gill-Simmen, associate dean at Royal Holloway, University of London, warns that reliance on AI in education could stunt critical thinking and lead to 'cognitive surrender' and 'epistemic atrophy.…
Professor Lucy Gill-Simmen of Royal Holloway, University of London warns that AI's greatest danger is not its inaccuracies but its potential to make people stop thinking critically, leading to "episte…
Wharton professor Eric Bradlow said companies still don't know how to incorporate AI in a holistic way, citing organizational change and the need for humans in the loop as the biggest bottleneck. He a…
AI is not making people intellectually lazy but rather enabling pre-existing habits of cognitive offloading, according to a growing body of research. Studies from Wharton, Carnegie Mellon, and Anthrop…
Ethan Mollick, a Wharton professor and AI researcher, gave Anthropic's Claude Fable 5 access to Unity and an MCP server with a single prompt to create a first-person shooter, resulting in a playable W…
Syndio CEO Maria Colacurcio describes how AI agents she built to run her company made her three times more productive, but her daughter Sofia Frei, a Dartmouth student, warns of the technology's unreg…
A 30-month study of 26,811 Chinese secondary school students found that using generative AI boosted homework scores by 18% and reduced completion time by 30%, but led to a 20% drop in closed-book exam…