Hamel Husain explains why AI evals fail before the evaluation begins
Hamel Husain, a prominent AI engineer, argues that most AI evaluations fail because teams treat weak outputs as model failures when the real problem is product design, not the model itself. In the sec…