A Benchmark of Text Classification in PyTorch
We are trying to build a Benchmark for Text Classification including
Many Text Classification DataSet, including Sentiment/Topic Classfication, popular language(e.g. English and Chinese). Meanwhile, a basic word embedding is provided.
Implment many popular and state-of-art Models, especially in deep neural network.
We have done some dataset and models
- IMDB
- SST
- Trec
- BasicCNN
- KimCNN
- MultiLayerCNN
- InceptionCNN
- LSTM
- FastText
You should have install these librarys
python3 torch torchtext
Dataset will be automatically configured in current path, or download manually your data in Dataset, step-by step.
including
Glove embeding Sentiment classfication dataset IMDB
Run in default setting
python main.pyCNN
python main.py --model cnnLSTM
python main.py --model lstmWelcome your issues and contribution!!!