nlpcommon

nlpcommon, Python Text Tool. Python3开发。

Guide

Feature
Install
Usage
Dataset
Contact
Cite
Reference

Feature

nlpcommon is a python Open Source Toolkit for text classification. The goal is to implement text analysis algorithm, so as to achieve the use in the production environment.

nlpcommon has the characteristics of clear algorithm, high performance and customizable corpus.

Functions：

Classifier

Cluster

MiniBatchKmeans

While providing rich functions, nlpcommon internal modules adhere to low coupling, model adherence to inert loading, dictionary publication, and easy to use.

Install

Requirements and Installation

pip3 install nlpcommon

git clone https://github.com/shibing624/nlpcommon.git
cd nlpcommon
python3 setup.py install

Usage

data

Stopwrods

examples/base_demo.py:

import sys

sys.path.append('..')
from nlpcommon import stopwords

if __name__ == '__main__':
    print(len(stopwords), stopwords)

output:

2438 {'．', '大家', '孰知', '至于', './', '知道', '二话没说', '一何', '从宽', 'especially' ... }

Contact

Issue(建议)：
邮件我：xuming: [email protected]
微信我：加我微信号：xuming624, 进Python-NLP交流群，备注：姓名-公司名-NLP

Cite

如果你在研究中使用了nlpcommon，请按如下格式引用：

@software{nlpcommon,
  author = {Xu Ming},
  title = {nlpcommon: A Tool for Text NLP},
  year = {2021},
  url = {https://github.com/shibing624/nlpcommon},
}

License

授权协议为 The Apache License 2.0，可免费用做商业用途。请在产品说明中附加nlpcommon的链接和授权协议。

Contribute

项目代码还很粗糙，如果大家对代码有所改进，欢迎提交回本项目，在提交之前，注意以下两点：

在tests添加相应的单元测试
使用python setup.py test来运行所有单元测试，确保所有单测都是通过的

之后即可提交PR。

Reference

pytextclassifier

nlpcommon is a python Open Source Toolkit for text classification.

Related tags

Overview

nlpcommon

Feature

Classifier

Cluster

Install

Usage

data

Stopwrods

Contact

Cite

License

Contribute

Reference

Owner

xuming

Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding

Coreference resolution for English, German and Polish, optimised for limited training data and easily extensible for further languages

Shared code for training sentence embeddings with Flax / JAX

💫 Industrial-strength Natural Language Processing (NLP) in Python

A model library for exploring state-of-the-art deep learning topologies and techniques for optimizing Natural Language Processing neural networks

تولید اسم های رندوم فینگیلیش

Search-Engine - 📖 AI based search engine

Recognition of 38 speech commands in russian. Based on Yandex Cup 2021 ML Challenge: ASR

Repo for Enhanced Seq2Seq Autoencoder via Contrastive Learning for Abstractive Text Summarization

CorNet Correlation Networks for Extreme Multi-label Text Classification

Smart discord chatbot integrated with Dialogflow to manage different classrooms and assist in teaching!

Dense Passage Retriever - is a set of tools and models for open domain Q&A task.

⚖️ A Statutory Article Retrieval Dataset in French.

Telegram bot to auto post messages of one channel in another channel as soon as it is posted, without the forwarded tag.

Geometry-Consistent Neural Shape Representation with Implicit Displacement Fields

Explore different way to mix speech model(wav2vec2, hubert) and nlp model(BART,T5,GPT) together

📔️ Generate a text-based journal from a template file.

Yet another Python binding for fastText

Natural language processing summarizer using 3 state of the art Transformer models: BERT, GPT2, and T5

NeuralQA: A Usable Library for Question Answering on Large Datasets with BERT