{"slug": "the-pit-crew", "title": "The Pit Crew", "summary": "Mark Pesce, a technology writer, reported that he used AI agents to optimize his local Qwen3.8-27B model, achieving a 20% speed improvement after GPT-5.6 Sol in Codex spent four hours tuning it. He also used the anonymous Ox Alpha model to draft configuration changes for the Hermes harness, which were applied by GPT-5.6 Terra, demonstrating a workflow where frontier models act as a 'pit crew' for home AI systems.", "body_md": "# The Pit Crew\n\nI've been having too much fun with the new home watershed AI models. But the cutting edge sometimes draws blood, and on Monday I had to disentangle a model from a task that it simply could not complete. Caught in a loop, it repeated the same mistake, over and over again.\n\nI'd heard of such things happening. I'd never seen it.\n\nI pulled back, raised the quality settings, let it go back to work. No change.\n\nSometimes the knife you think is 'sharp enough' turns out to be rather dull.\n\nSo, back into the drawer for the old dependable: [Qwen3.8-27B](https://thewatershed.markpesce.com/ai-comes-home/). (Eleven days old, which seems like a decade this month.)\n\nI'd put it aside for two reasons: it's very slow, and because it's very slow it can get \"caught\" in the Hermes harness I use it within.\n\nHermes is fantastic, but it has some settings that make it less than useful for long-horizon tasks. By this I mean tasks that run on the *nightshift* - from when I go to bed to when I wake up again. If a harness can't hold itself together for an eight-hour task, it's only getting in the way of the agent.\n\nThe speed issue is a classic optimisation problem, and the kind that's perfect to hand to a very smart AI agent - like GPT-5.6 Sol inside Codex. I told Sol to make my model fast and accurate and reliable, favouring accuracy and reliability over speed. It spent nearly four hours running a series of tests, taking benchmarks, tweaking settings, and testing again. At the end of all of that I have Qwen running 20% faster than it was before. Not insignificant - though I am fighting back envy of folks with GPUs big enough to run the model at 30 or even 130 tps. Luxury!\n\nThen onto Hermes, designed for agentic work of a more interactive variety than found on the unattended nightshift. I loaded up [Ox Alpha](https://wccftech.com/a-mysterious-ai-lab-is-offering-100-trillion-free-tokens-day-for-its-ox-alpha-model-as-evidence-points-to-zhipus-unreleased-glm/?ref=thewatershed.markpesce.com) in Hermes - an anonymous free model that's stunned users with its power, and raised some big questions about its origins - and asked it to draft a set of proposed changes to Hermes' configuration to accommodate my slow model chugging its way through the nightshift. I handed off those recommendations to Codex and GPT-5.6 Terra, because Hermes agents aren't allowed to manipulate their own configuration files.\n\nFor safety's sake, minds may *draft* changes to their own harnesses; they may never *apply* them. Ox proposed, Terra disposed, and neither could do the other's job. That policy is most of AI safety, practiced at the kitchen table without ceremony.\n\nSol's four hours were metered - rented frontier cognition, billed by the token, buying something *permanent*: a 20% improvement capitalised into a machine I own, compounding every night it runs, at no further cost. You don't rent intelligence to do the work anymore - you rent it, briefly, to make your *owned* intelligence better at doing the work.\n\nWith all of that up and running, I put a question to the nightshift - one I had posed earlier that day in the closing lines of \"[The Emerging Cognition Surplus](https://thewatershed.markpesce.com/the-emerging-cognition-surplus/)\": *What do you propose should be the next task?*\n\nIt had a good long think, examined its own work, found that work wanting, and proposed, as a first step, going back to make it whole and complete. \"*Should I set that up as the next task?*\" it asked.\n\nThe whole roster in one evening: Qwen working, Sol tuning, Ox drafting, Terra applying, and me signing off on the work. Four quick minds around one slow car. The frontier models now serve as *pit crew for the home stack*: rent the tuner, own the tuned.\n\nAnd as in racing, so here: nobody remembers the tire-changers, and no car finishes without them. But the chequered flag belongs to the slow car that *runs all night*.", "url": "https://wpnews.pro/news/the-pit-crew", "canonical_source": "https://thewatershed.markpesce.com/the-pit-crew/", "published_at": "2026-08-25 22:30:43+00:00", "updated_at": "2026-08-25 22:43:26.630776+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-agents", "ai-products", "ai-tools"], "entities": ["Mark Pesce", "Qwen3.8-27B", "Hermes", "GPT-5.6 Sol", "Codex", "Ox Alpha", "GPT-5.6 Terra"], "alternates": {"html": "https://wpnews.pro/news/the-pit-crew", "markdown": "https://wpnews.pro/news/the-pit-crew.md", "text": "https://wpnews.pro/news/the-pit-crew.txt", "jsonld": "https://wpnews.pro/news/the-pit-crew.jsonld"}}