Cost-Aware Best-LLM Identification using Dueling Feedback
A new arXiv paper (2609.30360v1) formulates a variant of the multi-armed bandit problem for identifying the best large language model from a collection with heterogeneous querying costs, using dueling…