10:10
2026-08-04
castform.com
artificial-intelligence
I RL-finetuned an LLM to unslop my writing
A developer trained a 4-billion-parameter language model using reinforcement learning to rewrite AI-generated text into more human-sounding prose, using AI detectors as reward functions. The project, โฆ