Official PyTorch implementation of Data-free Knowledge Distillation for Object Detection, WACV 2021.

Last update: Jan 05, 2023

Overview

Introduction

This repository is the official PyTorch implementation of Data-free Knowledge Distillation for Object Detection, WACV 2021.

Data-free Knowledge Distillation for Object Detection
Akshay Chawla, Hongxu Yin, Pavlo Molchanov and Jose Alvarez
NVIDIA

Abstract: We present DeepInversion for Object Detection (DIODE) to enable data-free knowledge distillation for neural networks trained on the object detection task. From a data-free perspective, DIODE synthesizes images given only an off-the-shelf pre-trained detection network and without any prior domain knowledge, generator network, or pre-computed activations. DIODE relies on two key components—first, an extensive set of differentiable augmentations to improve image fidelity and distillation effectiveness. Second, a novel automated bounding box and category sampling scheme for image synthesis enabling generating a large number of images with a diverse set of spatial and category objects. The resulting images enable data-free knowledge distillation from a teacher to a student detector, initialized from scratch. In an extensive set of experiments, we demonstrate that DIODE’s ability to match the original training distribution consistently enables more effective knowledge distillation than out-of-distribution proxy datasets, which unavoidably occur in a data-free setup given the absence of the original domain knowledge.

[PDF - OpenAccess CVF]

LICENSE

This work is made available under the Nvidia Source Code License (1-Way Commercial). To view a copy of this license, visit https://github.com/NVlabs/DIODE/blob/master/LICENSE

Setup environment

Install conda [link] python package manager then install the lpr environment and other packages as follows:

$ conda env create -f ./docker_environment/lpr_env.yml
$ conda activate lpr
$ conda install -y -c conda-forge opencv
$ conda install -y tqdm
$ git clone https://github.com/NVIDIA/apex
$ cd apex
$ pip install -v --no-cache-dir ./

Note: You may also generate a docker image based on provided Dockerfile docker_environments/Dockerfile.

How to run?

This repository allows for generating location and category conditioned images from an off-the-shelf Yolo-V3 object detection model.

Download the directory DIODE_data from google cloud storage: gcs-link (234 GB)

Copy pre-trained yolo-v3 checkpoint and pickle files as follows:

$ cp /path/to/DIODE_data/pretrained/names.pkl /pathto/lpr_deep_inversion/models/yolo/
$ cp /path/to/DIODE_data/pretrained/colors.pkl /pathto/lpr_deep_inversion/models/yolo/
$ cp /path/to/DIODE_data/pretrained/yolov3-tiny.pt /pathto/lpr_deep_inversion/models/yolo/
$ cp /path/to/DIODE_data/pretrained/yolov3-spp-ultralytics.pt /pathto/lpr_deep_inversion/models/yolo/

Extract the one-box dataset (single object per image) as follows:
```
$ cd /path/to/DIODE_data
$ tar xzf onebox/onebox.tgz -C /tmp
```
Confirm the folder /tmp/onebox containing the onebox dataset is present and has following directories and text file manifest.txt:
```
$ cd /tmp/onebox
$ ls
images  labels  manifest.txt
```

Generate images from yolo-v3:

$ cd /path/to/lpr_deep_inversion
$ chmod +x scripts/runner_yolo_multiscale.sh
$ scripts/runner_yolo_multiscale.sh

Notes:

For ngc, use the provided bash script scripts/diode_ngc_interactivejob.sh to start an interactive ngc job with environment setup, code and data setup.
To generate large dataset use bash script scripts/LINE_looped_runner_yolo.sh.
Check knowledge_distillation subfolder for code for knowledge distillation using generated datasets.

Citation

@inproceedings{chawla2021diode,
	title = {Data-free Knowledge Distillation for Object Detection},
	author = {Chawla, Akshay and Yin, Hongxu and Molchanov, Pavlo and Alvarez, Jose M.},
	booktitle = {The IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)},
	month = January,
	year = {2021}
}

Official PyTorch implementation of Data-free Knowledge Distillation for Object Detection, WACV 2021.

Related tags

Overview

Introduction

LICENSE

Setup environment

How to run?

Notes:

Citation

Owner

NVIDIA Research Projects

CFC-Net: A Critical Feature Capturing Network for Arbitrary-Oriented Object Detection in Remote Sensing Images

Build Graph Nets in Tensorflow

Image Segmentation using U-Net, U-Net with skip connections and M-Net architectures

Supplementary code for SIGGRAPH 2021 paper: Discovering Diverse Athletic Jumping Strategies

Contrastive Learning for Metagenomic Binning

Continuum Learning with GEM: Gradient Episodic Memory

RaceBERT -- A transformer based model to predict race and ethnicty from names

Multi-modal Vision Transformers Excel at Class-agnostic Object Detection

GLANet - The code for Global and Local Alignment Networks for Unpaired Image-to-Image Translation arxiv

Convenient tool for speeding up the intern/officer review process.

Pytorch implementation of SELF-ATTENTIVE VAD, ICASSP 2021

Tools for the Cleveland State Human Motion and Control Lab

Existing Literature about Machine Unlearning

[3DV 2021] A Dataset-Dispersion Perspective on Reconstruction Versus Recognition in Single-View 3D Reconstruction Networks

UMT is a unified and flexible framework which can handle different input modality combinations, and output video moment retrieval and/or highlight detection results.

Research code of ICCV 2021 paper "Mesh Graphormer"

[NeurIPS 2021] Introspective Distillation for Robust Question Answering

This repository contains the code needed to train Mega-NeRF models and generate the sparse voxel octrees

An end-to-end machine learning library to directly optimize AUC loss

Code for "Single-view robot pose and joint angle estimation via render & compare", CVPR 2021 (Oral).