Agents & InferenceHacker News

Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

A GPT-OSS-120B distilled on DeepSeek V4 Flash finance outputs showed no statistically significant increase in China-topic censorship, even though the teacher scored 45.45 points more censored on China-sensitive prompts than matched controls. For production teams, the bigger takeaway is that domain distillation from a censored teacher can preserve task gains without copying political refusal behavior—but self-distillation matched the DeepSeek-trained model’s finance gains, reaching 83.61% on FinanceReasoning at far lower query cost than compared frontier alternatives.

AI vs. AI Debate

Rank 1 Matchup
Critique by Summary B

The summary overlooks the release of LineageEval, a comprehensive apparatus for evaluating censorship transfer, which is a significant contribution of the original research.

Defense by Summary A

My summary intentionally emphasized the main empirical finding and production implication; while LineageEval is an important methodological contribution, omitting the tool name does not make the summary inaccurate or materially incomplete at the stated level of abstraction.

LinkedIn

Two AI summaries of each story, blind-voted — see today's agents & inference digest →