Learning Open-World Object Proposals without Learning to Classify

Last update: Dec 22, 2022

Overview

Learning Open-World Object Proposals without Learning to Classify

Pytorch implementation for "Learning Open-World Object Proposals without Learning to Classify" (arXiv 2021)

Dahun Kim, Tsung-Yi Lin, Anelia Angelova, In So Kweon, and Weicheng Kuo.

@article{kim2021oln,
  title={Learning Open-World Object Proposals without Learning to Classify},
  author={Kim, Dahun and Lin, Tsung-Yi and Angelova, Anelia and Kweon, In So and Kuo, Weicheng},
  journal={arXiv preprint arXiv:2108.06753},
  year={2021}
}

Introduction

Humans can recognize novel objects in this image despite having never seen them before. “Is it possible to learn open-world (novel) object proposals?” In this paper we propose Object Localization Network (OLN) that learns localization cues instead of foreground vs background classification. Only trained on COCO, OLN is able to propose many novel objects (top) missed by Mask R-CNN (bottom) on an out-of-sample frame in an ego-centric video.

Cross-category generalization on COCO

We train OLN on COCO VOC categories, and test on non-VOC categories. Note our [email protected] evaluation does not count those proposals on the 'seen' classes into the budget (k), to avoid evaluating recall on see-class objects.

Method	AUC	[email protected]	[email protected]	[email protected]	[email protected]	[email protected]	Download
OLN-Box	24.8	18.0	26.4	33.4	39.0	45.0	model

Disclaimer

This repo is tested under Python 3.7, PyTorch 1.7.0, Cuda 11.0, and mmcv==1.2.5.

Installation

This repo is built based on mmdetection.

You can use following commands to create conda env with related dependencies.

conda create -n oln python=3.7 -y
conda activate oln
conda install pytorch=1.7.0 torchvision cudatoolkit=11.0 -c pytorch -y
pip install mmcv-full
pip install -r requirements.txt
pip install -v -e .

Please also refer to get_started.md for more details of installation.

Prepare datasets

COCO dataset is available from official websites. It is recommended to download and extract the dataset somewhere outside the project directory and symlink the dataset root to $OLN/data as below.

object_localization_network
├── mmdet
├── tools
├── configs
├── data
│   ├── coco
│   │   ├── annotations
│   │   ├── train2017
│   │   ├── val2017
│   │   ├── test2017

Testing

Our trained models are available for download here. Place it under trained_weights/latest.pth and run the following commands to test OLN on COCO dataset.

# Multi-GPU distributed testing
bash tools/dist_test_bbox.sh configs/oln_box/oln_box.py \
trained_weights/latest.pth ${NUM_GPUS}
# OR
python tools/test.py configs/oln_box/oln_box.py work_dirs/oln_box/latest.pth --eval bbox

Training

# Multi-GPU distributed training
bash tools/dist_train.sh configs/oln_box/oln_box.py ${NUM_GPUS}

Contact

If you have any questions regarding the repo, please contact Dahun Kim ([email protected]) or create an issue.

Learning Open-World Object Proposals without Learning to Classify

Related tags

Overview

Learning Open-World Object Proposals without Learning to Classify

Pytorch implementation for "Learning Open-World Object Proposals without Learning to Classify" (arXiv 2021)

Introduction

Cross-category generalization on COCO

Disclaimer

Installation

Prepare datasets

Testing

Training

Contact

Owner

Dahun Kim

Vehicle direction identification consists of three module detection , tracking and direction recognization.

Computational Methods Course at UdeA. Forked and size reduced from:

Gesture-Volume-Control - This Python program can adjust the system's volume by using hand gestures

Deep Learning Tutorial for Kaggle Ultrasound Nerve Segmentation competition, using Keras

Code for CVPR 2021 paper TransNAS-Bench-101: Improving Transferrability and Generalizability of Cross-Task Neural Architecture Search.

Source code for paper "Deep Diffusion Models for Robust Channel Estimation", TBA.

DAT4 - General Assembly's Data Science course in Washington, DC

Code for TIP 2017 paper --- Illumination Decomposition for Photograph with Multiple Light Sources.

Propagate Yourself: Exploring Pixel-Level Consistency for Unsupervised Visual Representation Learning, CVPR 2021

Phylogeny Partners

Companion repo of the UCC 2021 paper "Predictive Auto-scaling with OpenStack Monasca"

Human Pose Detection on EdgeTPU

a spacial-temporal pattern detection system for home automation

《Fst Lerning of Temporl Action Proposl vi Dense Boundry Genertor》(AAAI 2020)

Semantic Segmentation for Aerial Imagery using Convolutional Neural Network

Applying CLIP to Point Cloud Recognition.

A synthetic texture-invariant dataset for object detection of UAVs

Si Adek Keras is software VR dangerous object detection.

Automatically Build Multiple ML Models with a Single Line of Code. Created by Ram Seshadri. Collaborators Welcome. Permission Granted upon Request.

Deep Learning Package based on TensorFlow