数据集读入

暖i

已于 2022-03-17 22:29:28 修改

阅读量260

点赞数

分类专栏： pytorch 文章标签： pytorch 深度学习 python

于 2022-03-17 22:28:39 首次发布

本文链接：https://blog.csdn.net/weixin_44270812/article/details/123562565

版权

pytorch 专栏收录该内容

2 篇文章 0 订阅

订阅专栏

Dataset类代码实战

在配置好环境后

from torch.utils.data import Dataset
from PIL import Image
import os

class MyData(Dataset):

    def __init__(self,root_dir,label_dir):
        self.root_dir = root_dir
        self.label_dir = label_dir
        self.path = os.path.join(self.root_dir,self.label_dir)//获取ants目录下的全部文件，这里是列表集合
        self.img_path = os.listdir(self.path)

    def __getitem__(self, idx):
        img_name = self.img_path[idx]
        img_item_path = os.path.join(self.root_dir,self.label_dir,img_name)
        img = Image.open(img_item_path)
        label = self.label_dir
        return img,label

    def __len__(self):
        return len(self.img_path)

root_dir = "dataset/train"
ants_label_dir = "ants"
bees_label_dir = "bees"
ants_dataset = MyData(root_dir,ants_label_dir)
bees_dataset = MyData(root_dir,bees_label_dir)

train_dataset = ants_dataset + bees_dataset//继承Dataset类，重载 '+' 直接把数据集相加起来

写txt文件，一般把image 和 label标签分开存储，这里简易的手写一个，评论区写的

root_dir = 'dataset/train'
target_dir = 'ants_image'
img_path = os.listdir(os.path.join(root_dir, target_dir))
label = target_dir.split('_')[0]
out_dir = 'ants_label'
for i in img_path:
    file_name = i.split('.jpg')[0]
    with open(os.path.join(root_dir, out_dir,"{}.txt".format(file_name)),'w') as f:
        f.write(label)

主要文件的读写，绝对路径相对路径，地址转义等问题还是不清晰，挖个坑

暖i

关注

0
点赞
踩
0

收藏

觉得还不错? 一键收藏
0
评论
数据集读入

Dataset类代码实战在配置好环境后from torch.utils.data import Datasetfrom PIL import Imageimport osclass MyData(Dataset): def __init__(self,root_dir,label_dir): self.root_dir = root_dir self.label_dir = label_dir self.path = os.path.jo
复制链接

扫一扫

专栏目录