When Do LLMs Replace Fine-Tuned NLU? A Decision Framework for Intent Detection in Production Conversational Systems
A new arXiv study (2608.20371v1) comparing zero-shot Claude Haiku against fine-tuned NLU models for intent detection finds that fine-tuned RoBERTa beats Claude by 11.8 points on ATIS (95.9 vs. 84.1, p…