Official implementation of "Generating 3D Molecules for Target Protein Binding"

Last update: Dec 07, 2022

Related tags

Overview

Generating 3D Molecules for Target Protein Binding

This is the official implementation of the GraphBP method proposed in the following paper.

Meng Liu, Youzhi Luo, Kanji Uchino, Koji Maruhashi, and Shuiwang Ji. "Generating 3D Molecules for Target Protein Binding".

Requirements

We include key dependencies below. The versions we used are in the parentheses. Our detailed environmental setup is available in environment.yml.

PyTorch (1.9.0)
PyTorch Geometric (1.7.2)
rdkit-pypi (2021.9.3)
biopython (1.79)
openbabel (3.3.1)

Preparing Data

Download and extract the CrossDocked2020 dataset:

wget https://bits.csb.pitt.edu/files/crossdock2020/CrossDocked2020_v1.1.tgz -P data/crossdock2020/
tar -C data/crossdock2020/ -xzf data/crossdock2020/CrossDocked2020_v1.1.tgz
wget https://bits.csb.pitt.edu/files/it2_tt_0_lowrmsd_mols_train0_fixed.types -P data/crossdock2020/
wget https://bits.csb.pitt.edu/files/it2_tt_0_lowrmsd_mols_test0_fixed.types -P data/crossdock2020/

Note: (1) The unzipping process could take a lot of time. Unzipping on SSD is much faster!!! (2) Several samples in the training set cannot be processed by our code. Hence, we recommend replacing the it2_tt_0_lowrmsd_mols_train0_fixed.types file with a new one, where these samples are deleted. The new one is available here.

Split data files:

python scripts/split_sdf.py data/crossdock2020/it2_tt_0_lowrmsd_mols_train0_fixed.types data/crossdock2020
python scripts/split_sdf.py data/crossdock2020/it2_tt_0_lowrmsd_mols_test0_fixed.types data/crossdock2020

Run

Train GraphBP from scratch:

CUDA_VISIBLE_DEVICES=${you_gpu_id} python main.py

Note: GraphBP can be trained on a 48GB GPU with batchsize=16. Our trained model is avaliable here.

Generate atoms in the 3D space with the trained model:

CUDA_VISIBLE_DEVICES=${you_gpu_id} python main_gen.py

Postprocess and then save the generated molecules:

CUDA_VISIBLE_DEVICES=${you_gpu_id} python main_eval.py

Reference

@article{liu2022graphbp,
      title={Generating 3D Molecules for Target Protein Binding},
      author={Meng Liu and Youzhi Luo and Kanji Uchino and Koji Maruhashi and Shuiwang Ji},
      journal={arXiv preprint arXiv:2204.09410},
      year={2022},
}

Official implementation of "Generating 3D Molecules for Target Protein Binding"

Related tags

Overview

Generating 3D Molecules for Target Protein Binding

Requirements

Preparing Data

Run

Reference

Owner

DIVE Lab, Texas A&M University

NeuTex: Neural Texture Mapping for Volumetric Neural Rendering

CKD - Collaborative Knowledge Distillation for Heterogeneous Information Network Embedding

Stochastic Tensor Optimization for Robot Motion - A GPU Robot Motion Toolkit

constructing maps of intellectual influence from publication data

EfficientNetv2 TensorRT int8

SASM - simple crossplatform IDE for NASM, MASM, GAS and FASM assembly languages

A collection of Reinforcement Learning algorithms from Sutton and Barto's book and other research papers implemented in Python.

MVP Benchmark for Multi-View Partial Point Cloud Completion and Registration

Code to reproduce experiments in the paper "Explainability Requires Interactivity".

Easy Parallel Library (EPL) is a general and efficient deep learning framework for distributed model training.

Constraint-based geometry sketcher for blender

Generative Art Using Neural Visual Grammars and Dual Encoders

LSTM built using Keras Python package to predict time series steps and sequences. Includes sin wave and stock market data

This python-based package offers a way of creating a parametric OpenMC plasma source from plasma parameters.

MINIROCKET: A Very Fast (Almost) Deterministic Transform for Time Series Classification

imbalanced-DL: Deep Imbalanced Learning in Python

Distributionally robust neural networks for group shifts

NeROIC: Neural Object Capture and Rendering from Online Image Collections

Official PyTorch implementation for Generic Attention-model Explainability for Interpreting Bi-Modal and Encoder-Decoder Transformers, a novel method to visualize any Transformer-based network. Including examples for DETR, VQA.

Quantization library for PyTorch. Support low-precision and mixed-precision quantization, with hardware implementation through TVM.