← 返回技术雷达
Hacker News tech

Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

Hacker News 热议:Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it(51 赞 / 40 评论,来源 ctgt.ai)

原文开头节选

July 29, 2026 What a Distilled Model Inherits From Its Teacher Johnny Yu, Siddarth Mamidanna, Cyril Gorlla Introduction [ gpt-oss-20b-finance weights on Hugging Face ] [Try the playground ] [ LineageEval ] [Explore the data on GitHub ]

Reasons for distillation from Chinese open models include a perceived superior cost to performance ratio, as well as the notion that the potentially harmful aspects of its behavior will not transfer to the distilled model. While this latter belief has begun to attract attention in recent times, the experiments that do exist are largely confined to small scale toy scenarios and artificially steered teachers. We investigate this phenomenon in a practical setting: a frontier Chinese model, used as a teacher, in a finance-adjacent production distillation pipeline. It is well understood that Chinese frontier models visibly refuse and reframe China-sensitive topics; the behavior is documented across audits of the DeepSeek R1 and V3 lines. The question we seek to answer is one level removed: does the student learn undesired behaviors along with the skill?

(以上为原文节选,完整内容见下方”原文来源”)

这条动态今日登上 Hacker News 首页(51 赞 / 40 评论,来源 ctgt.ai)。技术雷达每日自动聚合 AI 工程、后端架构、DevOps 方向的前沿动态;相关工程落地可浏览下方的相关服务与延伸阅读,或直接与我们团队交流。

原文来源: Hacker News

相关服务