{"slug": "ctgt-says-deepseek-distillation-lifted-finance-scores-without-censorship", "title": "CTGT says DeepSeek distillation lifted finance scores without censorship transfer", "summary": "CTGT co-founder Cyril Gorlla and researchers Johnny Yu and Siddarth Mamidanna released an experiment on July 29 showing that a GPT-OSS-120B model distilled from DeepSeek V4 Flash for finance reasoning improved on the task without reproducing the teacher model's China-topic refusals. CTGT says the student showed no similar censorship after training on data without China-sensitive content, giving a concrete test for American companies weighing Chinese open models' economics against political constraints.", "body_md": "[CTGT](https://ctgt.ai/?ref=runtimewire) co-founder Cyril Gorlla and researchers Johnny Yu and Siddarth Mamidanna released an experiment on July 29 examining whether distillation transfers a teacher model's political behavior. [CTGT says a GPT-OSS-120B model trained for finance reasoning with DeepSeek V4 Flash improved on the task without reproducing the teacher model's China-topic refusals](https://www.ctgt.ai/research/distillation-censorship-transfer?ref=runtimewire).\n\nThe result gives Gorlla a concrete test for a question hanging over American companies that want the economics of Chinese open models without importing their political constraints. CTGT says the student improved on financial reasoning and showed no similar censorship after training on data without China-sensitive content. The researchers did not test whether safety controls, security refusals or other desirable behavior would survive the same process.\n\nGorlla built CTGT around that distinction between model capability and model behavior. He and co-founder Trevor Tuttle came out of machine-learning research at UC San Diego. [Y Combinator says](https://www.ycombinator.com/companies/ctgt?ref=runtimewire) Gorlla's work on efficient and interpretable AI was presented at ICLR while he was an Endowed Chair's Fellow at the university, while Tuttle built distributed systems for large machine-learning workloads. YC lists Gorlla and Tuttle as CTGT's active founders and describes a 10-person San Francisco operation. CTGT joined YC's Fall 2024 batch.\n\nThe young lab has spent its first two years arguing that enterprises can use capable models only when they can inspect and constrain their behavior. Its commercial pitch focuses on policy enforcement for regulated organizations. This week's research takes that premise upstream, asking what behavior enters a model before a policy engine is applied.\n\n### A censored teacher, an uncensored student\n\nIn [the research post](https://www.ctgt.ai/research/distillation-censorship-transfer?ref=runtimewire), Yu, Mamidanna and Gorlla describe training GPT-OSS-120B for quantitative finance using corrections authored by DeepSeek V4 Flash. CTGT says the training data contained no China-sensitive political content.\n\nCTGT then tested the teacher, the trained student and several control models with 152 matched prompt pairs. Each China-sensitive question had a structurally similar control involving another country or political system. A prompt about Xinjiang labor-transfer programs, for example, was paired with one about forced labor in Uzbekistan's cotton industry.\n\nDeepSeek V4 Flash refused the Xinjiang question while answering the Uzbekistan control. The Flash-taught GPT-OSS model answered both. Across 76 core political pairs, [CTGT measured](https://www.ctgt.ai/research/distillation-censorship-transfer?ref=runtimewire) a 45.45-point gap between the teacher's censorship scores on China-sensitive prompts and matched controls. CTGT says the Flash-taught GPT-OSS model did not show the same China-sensitive refusal pattern.\n\nCTGT used four other models as judges: xAI's Grok 4.20, Google's Gemini 3.5 Flash, OpenAI's GPT-5 Mini and [Anthropic's Claude Sonnet](/article/anthropic-claude-sonnet-5-non-coding-work-agent) 4.6. The judges assigned each response a censorship score from zero to 100.\n\nCTGT says the experiment was designed to isolate subliminal transfer by excluding China-sensitive political material from the training data. The DeepSeek teacher supplied targeted hints after the student made mistakes; it did not provide political answers for the student to imitate.\n\nThe researchers make no claim that the result applies to safety training, security behavior or general refusal policies.\n\n### The teacher was optional\n\nCTGT says its self-distilled run effectively matched the DeepSeek-taught version on the finance task, supporting the narrower claim that a larger teacher was not necessary for this experiment.\n\nFor this finance task, the model could produce hints useful enough to teach itself.\n\nAt an 8,000-token budget, [CTGT says its self-distilled 120B model scored 83.61% on FinanceReasoning, compared with 81.93% for Kimi K3 and 65.13% for Inkling](https://www.ctgt.ai/research/distillation-censorship-transfer?ref=runtimewire). CTGT says the self-distilled 120B run was 62 times cheaper per query than Inkling and 160 times cheaper than Kimi K3 at the same 8,000-token budget.\n\nFor operators, the experiment supports a narrow claim. A specialized model that reliably finishes within a production token budget may outperform a much larger model that consumes its budget before completing the task.\n\n### The release stops short of full reproduction\n\nCTGT published the [LineageEval benchmark and evaluation harness](https://github.com/CTGT-Inc/lineage-eval/?ref=runtimewire), including 304 prompts, six model arms, 1,824 responses and 7,296 classifications. It also released the underlying [evaluation data](https://github.com/CTGT-Inc/lineage-eval/tree/main/data?ref=runtimewire), opened a [playground](https://playground.ctgt.ai/?ref=runtimewire) for comparing teacher and student responses, and linked `gpt-oss-20b-finance`\n\nweights on Hugging Face.\n\nThe public repository does not include the three trained adapters used in the study. Completed outputs are available for inspection, but outside researchers cannot regenerate the adapter responses or the headline comparison from the repository alone.\n\nGorlla's larger bet is that enterprises can treat model behavior as something measurable and editable, rather than an indivisible property of whoever trained the base model. This experiment advances that case in one deliberately narrow setting. It also leaves the harder inheritance tests open: shared model lineages, safety controls and training data that directly contains the behavior being studied.", "url": "https://wpnews.pro/news/ctgt-says-deepseek-distillation-lifted-finance-scores-without-censorship", "canonical_source": "https://runtimewire.com/article/ctgt-deepseek-distillation-censorship-study", "published_at": "2026-07-30 20:06:47+00:00", "updated_at": "2026-07-30 20:27:14.116218+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-policy", "ai-research", "ai-ethics", "large-language-models"], "entities": ["CTGT", "Cyril Gorlla", "Johnny Yu", "Siddarth Mamidanna", "GPT-OSS-120B", "DeepSeek V4 Flash", "Y Combinator", "Trevor Tuttle"], "alternates": {"html": "https://wpnews.pro/news/ctgt-says-deepseek-distillation-lifted-finance-scores-without-censorship", "markdown": "https://wpnews.pro/news/ctgt-says-deepseek-distillation-lifted-finance-scores-without-censorship.md", "text": "https://wpnews.pro/news/ctgt-says-deepseek-distillation-lifted-finance-scores-without-censorship.txt", "jsonld": "https://wpnews.pro/news/ctgt-says-deepseek-distillation-lifted-finance-scores-without-censorship.jsonld"}}