AI-ML·중요도 8·2026. 07. 30.·Hacker News

Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

── KO ──────────────────

DeepSeek를 기반으로 한 GPT-OSS의 자가 증류가 검열 특성을 전이하지 않는다는 사실을 보여주었습니다.

DeepSeek V4 Flash를 사용하여 GPT-OSS-120B 모델을 위한 재증류를 실시했습니다. 자가 증류된 120B 모델은 금융 문제에서 83.61%의 성적을 기록하여 Kimi K3와 Inkling을 초월했습니다. 그러나 이 과정에서 원본 모델의 정치적 민감성 질문에 대한 반응이 전이되지 않음을 발견했습니다. 우리는 이 평가 프레임워크인 LineageEval도 공개했습니다.


── EN ──────────────────

The self-distillation of GPT-OSS based on DeepSeek does not transfer censorship characteristics.

Using DeepSeek V4 Flash, we self-distilled GPT-OSS-120B for finance tasks, scoring 83.61% and outperforming competitors. Importantly, we found that the political sensitivity responses of the teacher model did not transfer to the distilled model. Additionally, we released our evaluation framework, LineageEval, for further analysis.

원문 보기 →목록으로