TPH-YOLOv5: Improved YOLOv5 Based on Transformer Prediction Head for Object Detection on Drone-Captured Scenarios

Last update: Dec 22, 2022

Related tags

Overview

TPH-YOLOv5

This repo is the implementation of "TPH-YOLOv5: Improved YOLOv5 Based on Transformer Prediction Head for Object Detection on Drone-Captured Scenarios".
On VisDrone Challenge 2021, TPH-YOLOv5 wins 4th place and achieves well-matched results with 1st place model.
You can get VisDrone-DET2021: The Vision Meets Drone Object Detection Challenge Results for more information.

Install

$ git clone https://github.com/cv516Buaa/tph-yolov5
$ cd tph-yolov5
$ pip install -r requirements.txt

Convert labels

VisDrone2YOLO_lable.py transfer VisDrone annotiations to yolo labels.
You should set the path of VisDrone dataset in VisDrone2YOLO_lable.py first.

$ python VisDrone2YOLO_lable.py

Inference

Datasets : VisDrone
Weights :
- yolov5l-xs-1.pt: | Baidu Drive(pw: vibe). | Google Drive |
- yolov5l-xs-2.pt: | Baidu Drive(pw: vffz). | Google Drive |

val.py runs inference on VisDrone2019-DET-val, using weights trained with TPH-YOLOv5.
(We provide two weights trained by two different models based on YOLOv5l.)

$ python val.py --weights ./weights/yolov5l-xs-1.pt --img 1996 --data ./data/VisDrone.yaml
                                    yolov5l-xs-2.pt
--augment --save-txt  --save-conf --task val --batch-size 8 --verbose --name v5l-xs

Ensemble

If you inference dataset with different models, then you can ensemble the result by weighted boxes fusion using wbf.py.
You should set img path and txt path in wbf.py.

$ python wbf.py

Train

train.py allows you to train new model from strach.

$ python train.py --img 1536 --batch 2 --epochs 80 --data ./data/VisDrone.yaml --weights yolov5l.pt --hy data/hyps/hyp.VisDrone.yaml --cfg models/yolov5l-xs-tr-cbam-spp-bifpn.yaml --name v5l-xs

Description of TPH-yolov5 and citation

If you have any question, please discuss with me by sending email to [email protected]
If you find this code useful please cite:

@inproceedings{zhu2021tph,
  title={TPH-YOLOv5: Improved YOLOv5 Based on Transformer Prediction Head for Object Detection on Drone-captured Scenarios},
  author={Zhu, Xingkui and Lyu, Shuchang and Wang, Xu and Zhao, Qi},
  booktitle={Proceedings of the IEEE/CVF International Conference on Computer Vision},
  pages={2778--2788},
  year={2021}
}

References

Thanks to their great works

TPH-YOLOv5: Improved YOLOv5 Based on Transformer Prediction Head for Object Detection on Drone-Captured Scenarios

Related tags

Overview

TPH-YOLOv5

Install

Convert labels

Inference

Ensemble

Train

Description of TPH-yolov5 and citation

References

Owner

cv516Buaa

Learning Dynamic Network Using a Reuse Gate Function in Semi-supervised Video Object Segmentation.

Code for Fold2Seq paper from ICML 2021

Code for `BCD Nets: Scalable Variational Approaches for Bayesian Causal Discovery`, Neurips 2021

Turning SymPy expressions into JAX functions

Labelbox is the fastest way to annotate data to build and ship artificial intelligence applications

Official implementation of the paper Image Generators with Conditionally-Independent Pixel Synthesis https://arxiv.org/abs/2011.13775

OCTIS: Comparing Topic Models is Simple! A python package to optimize and evaluate topic models (accepted at EACL2021 demo track)

Syllabic Quantity Patterns as Rhythmic Features for Latin Authorship Attribution

Iran Open Source Hackathon

An example to implement a new backbone with OpenMMLab framework.

A object detecting neural network powered by the yolo architecture and leveraging the PyTorch framework and associated libraries.

RITA is a family of autoregressive protein models, developed by LightOn in collaboration with the OATML group at Oxford and the Debora Marks Lab at Harvard.

Repo for Photon-Starved Scene Inference using Single Photon Cameras, ICCV 2021

Pytorch implementation for RelTransformer

SafePicking: Learning Safe Object Extraction via Object-Level Mapping, ICRA 2022

Learning to Initialize Neural Networks for Stable and Efficient Training

Systemic Evolutionary Chemical Space Exploration for Drug Discovery

DeepLab resnet v2 model in pytorch

Yolov5-lite - Minimal PyTorch implementation of YOLOv5

NudeNet: Neural Nets for Nudity Classification, Detection and selective censoring