关于pytorch的保存和加载模型

最新推荐文章于 2025-02-25 10:46:56 发布

窥探

最新推荐文章于 2025-02-25 10:46:56 发布

阅读量181

点赞数 1

分类专栏：机器学习个人笔记

本文链接：https://blog.youkuaiyun.com/qq_41360787/article/details/104249642

版权

机器学习个人笔记专栏收录该内容

6 篇文章

订阅专栏

关于pytorch的保存和加载模型常用的两种方法

参考自pytorch官方文档：https://pytorch.org/tutorials/beginner/saving_loading_models.html

什么是state_dict：

举例如下

# Define model
class TheModelClass(nn.Module):
    def __init__(self):
        super(TheModelClass, self).__init__()
        self.conv1 = nn.Conv2d(3, 6, 5)
        self.pool = nn.MaxPool2d(2, 2)
        self.conv2 = nn.Conv2d(6, 16, 5)
        self.fc1 = nn.Linear(16 * 5 * 5, 120)
        self.fc2 = nn.Linear(120, 84)
        self.fc3 = nn.Linear(84, 10)

    def forward(self, x):
        x = self.pool(F.relu(self.conv1(x)))
        x = self.pool(F.relu(self.conv2(x)))
        x = x.view(-1, 16 * 5 * 5)
        x = F.relu(self.fc1(x))
        x = F.relu(self.fc2(x))
        x = self.fc3(x)
        return x
        # Initialize model
model = TheModelClass()

# Initialize optimizer
optimizer = optim.SGD(model.parameters(), lr=0.001, momentum=0.9)

# Print model's state_dict
print("Model's state_dict:")
for param_tensor in model.state_dict():
    print(param_tensor, "\t", model.state_dict()[param_tensor].size())

# Print optimizer's state_dict
print("Optimizer's state_dict:")
for var_name in optimizer.state_dict():
    print(var_name, "\t", optimizer.state_dict()[var_name])

output

Model's state_dict:
conv1.weight     torch.Size([6, 3, 5, 5])
conv1.bias   torch.Size([6])
conv2.weight     torch.Size([16, 6, 5, 5])
conv2.bias   torch.Size([16])
fc1.weight   torch.Size([120, 400])
fc1.bias     torch.Size([120])
fc2.weight   torch.Size([84, 120])
fc2.bias     torch.Size([84])
fc3.weight   torch.Size([10, 84])
fc3.bias     torch.Size([10])

Optimizer's state_dict:
state    {}
param_groups     [{'lr': 0.001, 'momentum': 0.9, 'dampening': 0, 'weight_decay': 0, 'nesterov': False, 'params': [4675713712, 4675713784, 4675714000, 4675714072, 4675714216, 4675714288, 4675714432, 4675714504, 4675714648, 4675714720]}]

保存和加载模型在pytorch中分为两种方式

方法一

save：

torch.save(model.state_dict(), PATH)

load:

model = TheModelClass(*args, **kwargs)
model.load_state_dict(torch.load(PATH))
model.eval()

这种方法也是比较推荐的一种方式，存储东西少，速度快

方法二：

save：

torch.save(model, PATH)

load：

# Model class must be defined somewhere
model = torch.load(PATH)
model.eval()

以上两种是较为常见的方法，其他方法可以在文章顶部的官方链接中查看

关于pytorch的 保存和加载模型

关于pytorch的 保存和加载模型常用的两种方法

什么是state_dict：

output

方法一

save：

load:

方法二：

save：

load：

关于pytorch的保存和加载模型

关于pytorch的保存和加载模型常用的两种方法