知识蒸馏 pytorch官网源码分析

pythonSuperman

已于 2024-02-29 21:08:57 修改

阅读量515

点赞数 4

于 2024-02-29 19:07:49 首次发布

本文链接：https://blog.csdn.net/llf000000/article/details/136377486

版权

深度学习同时被 2 个专栏收录

55 篇文章 3 订阅

订阅专栏

迁移学习

4 篇文章 0 订阅

订阅专栏

参考连接：

Knowledge Distillation Tutorial — PyTorch Tutorials 2.2.1+cu121 documentation

方法一：

知识蒸馏的损失函数只接受两个相同维度的输入，所以我们需要采取措施使他们在进入损失函数之前是相同维度的。我们将使用平均池化层对在教师模型卷积运算后的logits进行池化，使得logits维度和学生保持一样。

方法二：

原来的教师模型：

只有forward函数内有调整

class DeepNN(nn.Module):

    def forward(self, x):
        x = self.features(x)
        x = torch.flatten(x, 1)
        x = self.classifier(x)
        return x

加入了池化后的教师模型：

class ModifiedDeepNNCosine(nn.Module):

    def forward(self, x):
        x = self.features(x)
        flattened_conv_output = torch.flatten(x, 1)
        x = self.classifier(flattened_conv_output)
        flattened_conv_output_after_pooling = torch.nn.functional.avg_pool1d(flattened_conv_output, 2)
        return x, flattened_conv_output_after_pooling

原来的学生模型

class LightNN(nn.Module):

    def forward(self, x):
        x = self.features(x)
        x = torch.flatten(x, 1)
        x = self.classifier(x)
        return x

更新后的学生模型：

class ModifiedLightNNCosine(nn.Module):

    def forward(self, x):
        x = self.features(x)
        flattened_conv_output = torch.flatten(x, 1)
        x = self.classifier(flattened_conv_output)
        return x, flattened_conv_output

方法三：

Teacher accuracy: 75.60%
Student accuracy without teacher: 70.41%
Student accuracy with CE + KD: 70.19%
Student accuracy with CE + CosineLoss: 70.87%
Student accuracy with CE + RegressorMSE: 71.40%

pythonSuperman

关注

4
点赞
踩
10

收藏

觉得还不错? 一键收藏
0
评论
知识蒸馏 pytorch官网源码分析

知识蒸馏的损失函数只接受两个相同维度的输入，所以我们需要采取措施使他们在进入损失函数之前是相同维度的。我们将使用平均池化层对在教师模型卷积运算后的logits进行池化，使得logits维度和学生保持一样。只有forward函数内有调整。
复制链接

扫一扫