Putting Task Expertise into RL Achieves Performance on Text-to-SQL
A fine-tuned model called ReViSQL-K2.6, developed by Tinker using reinforcement learning with verifiable rewards, exceeds the human benchmark of 92.96% on the BIRD text-to-SQL task, reaching 92.96% when selecting from 16…