nlpcommon

nlpcommon, Python Text Tool. Python3开发。

Guide

Feature
Install
Usage
Dataset
Contact
Cite
Reference

Feature

nlpcommon is a python Open Source Toolkit for text classification. The goal is to implement text analysis algorithm, so as to achieve the use in the production environment.

nlpcommon has the characteristics of clear algorithm, high performance and customizable corpus.

Functions：

Classifier

Cluster

MiniBatchKmeans

While providing rich functions, nlpcommon internal modules adhere to low coupling, model adherence to inert loading, dictionary publication, and easy to use.

Install

Requirements and Installation

pip3 install nlpcommon

git clone https://github.com/shibing624/nlpcommon.git
cd nlpcommon
python3 setup.py install

Usage

data

Stopwrods

examples/base_demo.py:

import sys

sys.path.append('..')
from nlpcommon import stopwords

if __name__ == '__main__':
    print(len(stopwords), stopwords)

output:

2438 {'．', '大家', '孰知', '至于', './', '知道', '二话没说', '一何', '从宽', 'especially' ... }

Contact

Issue(建议)：
邮件我：xuming: [email protected]
微信我：加我微信号：xuming624, 进Python-NLP交流群，备注：姓名-公司名-NLP

Cite

如果你在研究中使用了nlpcommon，请按如下格式引用：

@software{nlpcommon,
  author = {Xu Ming},
  title = {nlpcommon: A Tool for Text NLP},
  year = {2021},
  url = {https://github.com/shibing624/nlpcommon},
}

License

授权协议为 The Apache License 2.0，可免费用做商业用途。请在产品说明中附加nlpcommon的链接和授权协议。

Contribute

项目代码还很粗糙，如果大家对代码有所改进，欢迎提交回本项目，在提交之前，注意以下两点：

在tests添加相应的单元测试
使用python setup.py test来运行所有单元测试，确保所有单测都是通过的

之后即可提交PR。

Reference

pytextclassifier

nlpcommon is a python Open Source Toolkit for text classification.

Related tags

Overview

nlpcommon

Feature

Classifier

Cluster

Install

Usage

data

Stopwrods

Contact

Cite

License

Contribute

Reference

Owner

xuming

Beautiful visualizations of how language differs among document types.

SASE : Self-Adaptive noise distribution network for Speech Enhancement with heterogeneous data of Cross-Silo Federated learning

GPT-3 command line interaction

(ACL-IJCNLP 2021) Convolutions and Self-Attention: Re-interpreting Relative Positions in Pre-trained Language Models.

A Flask Sentiment Analysis API, with visual implementation

iSTFTNet : Fast and Lightweight Mel-spectrogram Vocoder Incorporating Inverse Short-time Fourier Transform

Code for EmBERT, a transformer model for embodied, language-guided visual task completion.

DAGAN - Dual Attention GANs for Semantic Image Synthesis

Reproducing the Linear Multihead Attention introduced in Linformer paper (Linformer: Self-Attention with Linear Complexity)

KR-FinBert And KR-FinBert-SC

Simple, Pythonic, text processing--Sentiment analysis, part-of-speech tagging, noun phrase extraction, translation, and more.

Honor's thesis project analyzing whether the GPT-2 model can more effectively generate free-verse or structured poetry.

Sequence Modeling with Structured State Spaces

This repository contains all the source code that is needed for the project : An Efficient Pipeline For Bloom’s Taxonomy Using Natural Language Processing and Deep Learning

This repo is to provide a list of literature regarding Deep Learning on Graphs for NLP

Py65 65816 - Add support for the 65C816 to py65

A simple Streamlit App to classify swahili news into different categories.

This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.

Semi-automated vocabulary generation from semantic vector models

PyTorch Implementation of "Non-Autoregressive Neural Machine Translation"