Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

Last update: Dec 13, 2022

Related tags

Text Data & NLP Diaformer

Overview

Diaformer

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation (AAAI 2022)

Diaformer is an efficient model for automatic diagnosis via symptoms sequence generation. It takes the sequence of symptoms as input, and predicts the inquiry symptoms in the way of sequence generation.

Figure 1: Illustration of symptom attention framework.

Requirements

Our experiments are conducted on Python 3.8 and Pytorch == 1.8.0. The main requirements are:

transformers==2.1.1
torch
numpy
tqdm
sklearn
keras
boto3

In the root directory, run following command to install the required libraries.

pip install -r requirement.txt

Usage

Download data

Download the datasets, then decompress them and put them in the corrsponding documents in \data. For example, put the data of Synthetic Dataset under data/synthetic_dataset.

The dataset can be downloaded as following links:
Build data

Switch to the corresponding directory of the dataset and just run preprocess.py to preprocess data and generate a vocabulary of symptoms.

Train and test

Train and test models by the follow commands.

Diaformer

# Train and test on Diaformer
# Run on MuZhi dataset
python Diaformer.py --dataset_path data/muzhi_dataset --batch_size 16 --lr 5e-5 --min_probability 0.009 --max_turn 20 --start_test 10 

# Run on Dxy dataset
python Diaformer.py --dataset_path data/dxy_dataset --batch_size 16 --lr 5e-5 --min_probability 0.012 --max_turn 20 --start_test 10 

# Run on Synthetic dataset
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Diaformer_GPT2

# Train and test on GPT2 variant of Diaformer
python GPT2_variant.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Diaformer_UniLM

# Train and test on UniLM variant of Diaformer
python UniLM_variant.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Ablation study

# run ablation study
# w/o Sequence Shuffle
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_sequence_shuffle

# w/o Synchronous Learning
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_synchronous_learning

# w/o Repeated Sequence
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_repeated_sequence

Generative inference

# save the model
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --model_output_path models
# use the trained model to output the results
python predict.py --dataset_path data/synthetic_dataset --min_probability 0.01 --max_turn 20 --pretrained_model models/ --result_output_path results.json

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

Related tags

Overview

Diaformer

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation (AAAI 2022)

Diaformer is an efficient model for automatic diagnosis via symptoms sequence generation. It takes the sequence of symptoms as input, and predicts the inquiry symptoms in the way of sequence generation.

Requirements

Usage

Owner

Junying Chen

Fake Shakespearean Text Generator

This python module is an easy-to-use port of the text normalization used in the paper "Not low-resource anymore: Aligner ensembling, batch filtering, and new datasets for Bengali-English machine translation". It is intended to be used for normalizing / cleaning Bengali and English text.

Sentello is python script that simulates the anti-evasion and anti-analysis techniques used by malware.

PyTorch implementation of NATSpeech: A Non-Autoregressive Text-to-Speech Framework

BiQE: Code and dataset for the BiQE paper

Under the hood working of transformers, fine-tuning GPT-3 models, DeBERTa, vision models, and the start of Metaverse, using a variety of NLP platforms: Hugging Face, OpenAI API, Trax, and AllenNLP

The source code of "Language Models are Few-shot Multilingual Learners" (MRL @ EMNLP 2021)

This project converts your human voice input to its text transcript and to an automated voice too.

Interactive Jupyter Notebook Environment for using the GPT-3 Instruct API

Command Line Text-To-Speech using Google TTS

Shared, streaming Python dict

Twitter-Sentiment-Analysis - Twitter sentiment analysis for india's top online retailers(2019 to 2022)

Sequence-to-sequence framework with a focus on Neural Machine Translation based on Apache MXNet

Simple, hackable offline speech to text - using the VOSK-API.

Collection of scripts to pinpoint obfuscated code

Topic Modelling for Humans

Translation for Trilium Notes. Trilium Notes 中文版.

This is the code for the EMNLP 2021 paper AEDA: An Easier Data Augmentation Technique for Text Classification

Input english text, then translate it between languages n times using the Deep Translator Python Library.