CUHK Occlusion Dataset（行人检测数据集）转换为YOLO+VOC数据集

2023年7月9日上午8:53 • 人工智能 • 阅读 88

一、引语

最近想用YOLO训练一个行人检测的模型,找到了CUHK Occlusion Datase 这个数据集，但这个数据集想要用YOLO训练的话，涉及到Seq文件转JPEG和VBB文件转XML，花了一下午时间在网上查资料，整理了一下供大家参考，参考文章链接在文末。

二、准备工作

1.首先新建⼀个set00⽂件夹，将所有.seq⽂件都放进去；

CUHK Occlusion Dataset（行人检测数据集）转换为YOLO+VOC数据集

3.在Annotations⽂件夹中同样新建⼀个set00⽂件夹，将所有.vbb⽂件放⼊，如下：

三、Seq文件转JPEG

-*- coding:utf-8 -*-
import os.path
import fnmatch
import shutil

def open_save(file, savepath):
    # 读入一个seq文件，然后拆分成image存入savepath当中
    f = open(file, 'rb')
    # 将seq文件的内容转化成str类型
    string = f.read().decode('latin-1')

    # splitstring是图片的前缀，可以理解成seq是以splitstring为分隔的多个jpg合成的文件
    splitstring = "\xFF\xD8\xFF\xE0\x00\x10\x4A\x46\x49\x46"
    # split函数做一个测试,因此返回结果的第一个是在seq文件中是空，因此后面省略掉第一个
"""
    >>> a = ".12121.3223.4343"
    >>> a.split('.')
    ['', '12121', '3223', '4343']
"""
    strlist = string.split(splitstring)
    # print(strlist)
    # print('######################################')
    f.close()
    count = 0
    # delete the image folder path if it exists
    if os.path.exists(savepath):
        shutil.rmtree(savepath)
        # create the image folder path
    if not os.path.exists(savepath):
        os.makedirs(savepath)
    # 遍历每一个jpg文件内容，然后加上前缀合成图片
    for img in strlist:
        filename = str(count) + '.jpg'
        filenamewithpath = os.path.join(savepath, filename)
        if count > 0:
            i = open(filenamewithpath, 'wb+')
            i.write(splitstring.encode('latin-1'))
            i.write(img.encode('latin-1'))
            i.close()
        count = count + 1

if __name__ == "__main__":
    rootdir = "D:/BaiduNetdiskDownload/行人检测/46_CUHK Occlusion Dataset"
    saveroot = "C:/Users/杨鑫兴/Desktop/VOCdevkit/JPEG"

    for parent, dirnames, filenames in os.walk(rootdir):
        for filename in filenames:
            if fnmatch.fnmatch(filename, '*.seq'):
                thefilename = os.path.join(parent, filename)
                thesavepath = saveroot + '/' + parent.split('/')[-1] + '/' + filename.split('.')[0] + '/'
                print("Filename=" + thefilename)
                print("Savepath=" + thesavepath)
                open_save(thefilename, thesavepath)

这里只需要修改两个路径

rootdir是你刚刚新建的set00和 Annotations所在的文件夹路径

saveroot是你存放转好的JPEG的位置（自己创建一个文件夹存放即可）

效果如图：

四、VBB转XML

-*- coding:utf-8 -*-
import os, glob
import cv2
from scipy.io import loadmat
from collections import defaultdict
import numpy as np
from lxml import etree,objectify

def vbb_anno2dict(vbb_file, cam_id):
    # 通过os.path.basename获得路径的最后部分"文件名.扩展名"
    # 通过os.path.splitext获得文件名
    filename = os.path.splitext(os.path.basename(vbb_file))[0]

    # 定义字典对象annos
    annos = defaultdict(dict)
    vbb = loadmat(vbb_file)
    # object info in each frame: id, pos, occlusion, lock, posv
    objLists = vbb['A'][0][0][1][0]
    objLbl = [str(v[0]) for v in vbb['A'][0][0][4][0]]  # 可查看所有类别
    # person index
    person_index_list = np.where(np.array(objLbl) == "person")[0]  # 只选取类别为'person'的xml
    for frame_id, obj in enumerate(objLists):
        if len(obj) > 0:
            frame_name = str(cam_id) + "_" + str(filename) + "_" + str(frame_id + 1) + ".jpg"
            annos[frame_name] = defaultdict(list)
            annos[frame_name]["id"] = frame_name
            annos[frame_name]["label"] = "person"
            for id, pos, occl in zip(obj['id'][0], obj['pos'][0], obj['occl'][0]):
                id = int(id[0][0]) - 1  # for matlab start from 1 not 0
                if not id in person_index_list:  # only use bbox whose label is person
                    continue
                pos = pos[0].tolist()
                occl = int(occl[0][0])
                annos[frame_name]["occlusion"].append(occl)
                annos[frame_name]["bbox"].append(pos)
            if not annos[frame_name]["bbox"]:
                del annos[frame_name]
    print(annos)
    return annos

def seq2img(annos, seq_file, outdir, cam_id):
    cap = cv2.VideoCapture(seq_file)
    index = 1
    # captured frame list
    v_id = os.path.splitext(os.path.basename(seq_file))[0]
    cap_frames_index = np.sort([int(os.path.splitext(id)[0].split("_")[2]) for id in annos.keys()])
    while True:
        ret, frame = cap.read()
        print(ret)
        if ret:
            if not index in cap_frames_index:
                index += 1
                continue
            if not os.path.exists(outdir):
                os.makedirs(outdir)
            outname = os.path.join(outdir, str(cam_id) + "_" + v_id + "_" + str(index) + ".jpg")
            print("Current frame: ", v_id, str(index))
            cv2.imwrite(outname, frame)
            height, width, _ = frame.shape
        else:
            break
        index += 1
    img_size = (width, height)
    return img_size

def instance2xml_base(anno, bbox_type='xyxy'):
    """bbox_type: xyxy (xmin, ymin, xmax, ymax); xywh (xmin, ymin, width, height)"""
    assert bbox_type in ['xyxy', 'xywh']
    E = objectify.ElementMaker(annotate=False)
    anno_tree = E.annotation(
        E.folder('VOC2014_instance/person'),
        E.filename(anno['id']),
        E.source(
            E.database('Caltech pedestrian'),
            E.annotation('Caltech pedestrian'),
            E.image('Caltech pedestrian'),
            E.url('None')
        ),
        E.size(
            E.width(640),
            E.height(480),
            E.depth(3)
        ),
        E.segmented(0),
    )
    for index, bbox in enumerate(anno['bbox']):
        bbox = [float(x) for x in bbox]
        if bbox_type == 'xyxy':
            xmin, ymin, w, h = bbox
            xmax = xmin + w
            ymax = ymin + h
        else:
            xmin, ymin, xmax, ymax = bbox
        E = objectify.ElementMaker(annotate=False)
        anno_tree.append(
            E.object(
                E.name(anno['label']),
                E.bndbox(
                    E.xmin(xmin),
                    E.ymin(ymin),
                    E.xmax(xmax),
                    E.ymax(ymax)
                ),
                E.difficult(0),
                E.occlusion(anno["occlusion"][index])
            )
        )
    return anno_tree

def parse_anno_file(vbb_inputdir, vbb_outputdir):
    # annotation sub-directories in hda annotation input directory
    assert os.path.exists(vbb_inputdir)
    sub_dirs = os.listdir(vbb_inputdir)  # 对应set00,set01...

    for sub_dir in sub_dirs:
        print("Parsing annotations of camera: ", sub_dir)
        cam_id = sub_dir
        # 获取某一个子set下面的所有vbb文件
        vbb_files = glob.glob(os.path.join(vbb_inputdir, sub_dir, "*.vbb"))
        for vbb_file in vbb_files:
            # 返回一个vbb文件中所有的帧的标注结果

            annos = vbb_anno2dict(vbb_file, cam_id)

            if annos:
                # 组成xml文件的存储文件夹，形如"/Users/chenguanghao/Desktop/Caltech/xmlresult/"
                vbb_outdir = vbb_outputdir

                # 如果不存在
                if not os.path.exists(vbb_outdir):
                    os.makedirs(vbb_outdir)

                for filename, anno in sorted(annos.items(), key=lambda x: x[0]):
                    if "bbox" in anno:
                        anno_tree = instance2xml_base(anno)
                        outfile = os.path.join(vbb_outdir, os.path.splitext(filename)[0] + ".xml")
                        print("Generating annotation xml file of picture: ", filename)
                        # 生成最终的xml文件，对应一张图片
                        etree.ElementTree(anno_tree).write(outfile, pretty_print=True)

def visualize_bbox(xml_file, img_file):
    import cv2
    tree = etree.parse(xml_file)
    # load image
    image = cv2.imread(img_file)
    origin = cv2.imread(img_file)
    # 获取一张图片的所有bbox
    for bbox in tree.xpath('//bndbox'):
        coord = []
        for corner in bbox.getchildren():
            coord.append(int(float(corner.text)))
        print(coord)
        cv2.rectangle(image, (coord[0], coord[1]), (coord[2], coord[3]), (0, 0, 255), 2)
    # visualize image
    cv2.imshow("test", image)
    cv2.imshow('origin', origin)
    cv2.waitKey(0)

def main():
    vbb_inputdir = r"C:\Users\杨鑫兴\Desktop\Annotations"
    vbb_outputdir = r"C:\Users\杨鑫兴\Desktop\VOCdevkit\Annotations"
    parse_anno_file(vbb_inputdir, vbb_outputdir)

"""
    下面这段是测试代码
"""

"""
    xml_file = "/Users/chenguanghao/Desktop/Caltech/xmlresult/set07/bbox/set07_V000_4.xml"
    img_file = "/Users/chenguanghao/Desktop/Caltech/JPEG/set07/V000/4.jpg"
    visualize_bbox(xml_file, img_file)
"""

if __name__ == "__main__":
    main()

这里只需要改两个路径

vbb_inputdir是你开始时解压label压缩包时创建的Annotations文件夹路径

vbb_outputdir是存放转好的XML数据文件夹的路径（自己创建一个即可）

效果如图:

五、把刚刚照片与xml一一对应

import ntpath
import os
import glob
import shutil

imgpathin=r'C:\Users\杨鑫兴\Desktop\VOCdevkit\JPEG'
imgout=r'C:\Users\杨鑫兴\Desktop\VOCdevkit\JPEGlmages'
for subdir in os.listdir(imgpathin):
    print(subdir)
    file_path=os.path.join(imgpathin,subdir)
    for subdir1 in os.listdir(file_path):
        print(subdir1)
        file_path1=os.path.join(file_path,subdir1)
        for jpg_file in os.listdir(file_path1):
            src=os.path.join(file_path1,jpg_file)
            new_name=str(subdir+"_"+subdir1+"_"+jpg_file)
            dst=os.path.join(imgout,new_name)
            os.rename(src,dst)

同样这里也需要修改两个路径

imgpathin是你刚刚存放JPEG文件夹的路径

imgout是存放修改后图片文件夹的路径（创建一个即可）

效果如下:

六、VOC数据集制作

七、结语

这篇文章我没有什么技术上的创新，我做的只是把自己花时间查的资料整合了一下，希望可以帮到大家，至于数据集，我查资料的时候发现作者把网盘链接删了，不知道什么原因，需要的私信或者留言我，最后感谢这些参考博客的作者👇

VOC数据集具体格式_北漠苍狼1746430162的博客-CSDN博客

CUHKOcclusionDataset数据集转换为yolo格式,并划分测试集和训练集 – 百度文库 (baidu.com)

（写的比较详细，大家可以看一下，对我帮助很大，唯一的缺点就是代码没有办法复制（百度文库嘛））

Original: https://blog.csdn.net/weixin_55504804/article/details/125608701
Author: 派提克up
Title: CUHK Occlusion Dataset（行人检测数据集）转换为YOLO+VOC数据集

原创文章受到原创版权保护。转载请注明出处：https://www.johngo689.com/680361/

转载文章受原作者版权保护。转载请注明原作者出处！

人工智能

【自取】最近整理的，有需要可以领取学习：

Linux核心资料大放送~

全栈面试题汇总（持续更新&可下载）

一个提高学习100%效率的工具！

【超详细】深度学习面试题目！

LeetCode Python刷题答案下载！

LeetCode Java版刷题答案下载！

LeetCode C++ 版本，抓紧保存！

LeetCode GO语言刷题答案下载！

Tensorflow-GPU（Win10）超完整版安装

一、Anaconda的安装 ANACONDA官网这个部分需要注意的就是添加环境变量，不然后期使用VSCode测试的时候会出现IMPORT ERROR 上面四个文件路径在Anaco…

人工智能 2023年5月23日
0091
使用pip使用报错：pip is configured with locations that require TLS/SSL

编译安装完python3.10后，pip不能使用！出现报错： pip is configured with locations that require TLS/SSL, how…

人工智能 2023年7月5日
0092
解决Font family [‘sans-serif‘] not found的问题

在进行matplotlib画图的时候，经常会出现这个的报错，虽然知道是因为没有对应的字体的原因，但是，将字体下载后放到目标路径下，仍然没有办法使用，才发现，出了下载字体到对应目录下…

人工智能 2023年7月15日
00140
Fp-growth算法python实现（数据挖掘学习笔记）

目录 1.算法伪代码 2.算法代码 3.测试数据 4.结果 1.算法伪代码输入： D：事务数据库。 min_sup：最小支持度阈值。输出：频繁模式的完全集。方法： 1.按照…

人工智能 2023年6月19日
00109
AI眼中的世界 ——人工智能绘画入门

目录什么是Disco Diffusion？如何使用Disco Diffusion？正文准备工作入门教程开始行动默认跑一个默认的描述A beautiful painti…

人工智能 2023年7月25日
0070
论文解读：学习蛋白质的空间结构可以提高蛋白质相互作用的预测

文章目录论文概况 1. 研究背景 2. 研究数据 * 2.1 种内数据集 2.2 种间数据集 2.3 多类别数据集 3. 研究方法 * 3.1数据预处理 3.2局部特征提取 * …

人工智能 2023年5月27日
00103
统计学之方差分析

一、基本原理从形式上看，方差分析是比较多个总体的均值是否相等，但本质上它所研究的是分类自变量对数值因变量的影响。当检验多个总体的均值是否相等时，方差分析是更有效的统计方法。由于…

人工智能 2023年6月19日
00185
数据增强之Mosaic数据增强的优点、Mixup,Cutout,CutMix的区别

一、Mosaic data augmentation Mosaic数据增强方法是YOLOV4论文中提出来的，主要思想是将四张图片进行随机裁剪，再拼接到一张图上作为训练数据。这样做…

人工智能 2023年5月26日
00134
【平衡小车】【串级PID参数整定】【详细版】根据现象手动调整平衡小车的PID

【平衡小车】【串级PID参数整定】【详细版】根据现象手动调整平衡小车的PID 简介：二轮平衡小车的控制分为平衡环（又称为直立环，保持稳定角度）、速度环（用来保持稳定时速度为零）以及…

人工智能 2023年6月10日
00203
数据库数据转json字符串及ajax请求数据渲染

续接：1.使用web框架程序处理客户端的动态资源请求代码实现 2.web框架之路由列表及SQL语句查询数据库数据替换模板变量一、数据库数据转json字符串 web框架程序可以开发…

人工智能 2023年6月30日
00104
ubuntu安装opencv_contrib扩展库，附踩坑+测试

博主昨晚需要用到OpenCV的SURF接口，但是发现无法调用，因为没有头文件。于是查阅了下资料，发现这些库已经被美国买下专利，成为付费库，都在opencv_contrib中。如果你…

人工智能 2023年7月19日
0097
【吴恩达机器学习笔记详解】第三章机器学习数学基础（线代）

3.1 矩阵和向量矩阵是指由数字组成的矩阵陈列，并且写在方括号内。如下图实际上矩阵可以说是二维数组的另外一种说法。矩阵的维度：行乘上列为矩阵的维度下面再介绍向量一个向量是一种…

人工智能 2023年7月14日
0068
Python3 DataFrame缺失值的处理

在通过Pandas做数据分析时，数据中往往会因为一些原因而出现缺失值NaN (Nota number)o比如前文中的例子，当两个DataFrame对象进行简单运算时，无法匹配的位置…

人工智能 2023年7月6日
0070
目标检测二阶段原始模型从rcnn到fasterrcnn

目标检测二阶段原始模型从rcnn到fasterrcnn 目标检测模型 * 1.1Rcnn 1.2SPPNet 1.3Fast R-CNN 1.4Faster R-CNN 目标检测…

人工智能 2023年7月9日
0087
部署过程中是否需要考虑数据隐私和安全性问题

问题：在部署过程中是否需要考虑数据隐私和安全性问题？介绍在部署过程中，数据隐私和安全性问题至关重要。在很多应用场景中，我们处理的数据可能包含敏感信息，如个人身份信息、财务记录和…

人工智能 2024年1月3日
0057
轻量级图卷积网络LightGCN介绍和构建推荐系统示例

推荐系统是当今业界最具影响力的 ML 任务。从淘宝到抖音，科技公司都在不断尝试为他们的特定应用程序构建更好的推荐系统。而这项任务并没有变得更容易，因为我们每天都希望看到更多可供选择…

人工智能 2023年7月17日
0077

2024 年 5 月
一	二	三	四	五	六	日
		1	2	3	4	5
6	7	8	9	10	11	12
13	14	15	16	17	18	19
20	21	22	23	24	25	26
27	28	29	30	31

CUHK Occlusion Dataset（行人检测数据集）转换为YOLO+VOC数据集

大家都在看