引用本概念的论文(1) EasyOPD:面向大语言模型的易用式在策略蒸馏框架 EasyOPD: An Easy-to-use On-Policy Distillation Framework for Large Language Models arXiv 2607.11012
共现相关概念 在引用本概念的论文中,与下列概念同时出现的次数(降序)。 Cross-Tokenizer Distillation method · 共现 1 Distributed Reinforcement Learning methodology · 共现 1 Knowledge Distillation method · 共现 1 On-policy Distillation (OPD) method · 共现 1 Self-Distillation method · 共现 1