05:09
2026-07-28
sourcefeed.dev
artificial-intelligence
A $500 RL Fine-Tune Beat the Frontier. Sort Of.
Fermisense, a consultancy, spent roughly $500 of GPU time fine-tuning a Qwen3.5-9B open model with reinforcement learning and claims it beat GPT-5.5, Gemini 3.1 Pro, Claude Opus 4.8, and Claude Fable β¦