Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

Last update: Dec 13, 2022

Related tags

Text Data & NLP Diaformer

Overview

Diaformer

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation (AAAI 2022)

Diaformer is an efficient model for automatic diagnosis via symptoms sequence generation. It takes the sequence of symptoms as input, and predicts the inquiry symptoms in the way of sequence generation.

Figure 1: Illustration of symptom attention framework.

Requirements

Our experiments are conducted on Python 3.8 and Pytorch == 1.8.0. The main requirements are:

transformers==2.1.1
torch
numpy
tqdm
sklearn
keras
boto3

In the root directory, run following command to install the required libraries.

pip install -r requirement.txt

Usage

Download data

Download the datasets, then decompress them and put them in the corrsponding documents in \data. For example, put the data of Synthetic Dataset under data/synthetic_dataset.

The dataset can be downloaded as following links:
Build data

Switch to the corresponding directory of the dataset and just run preprocess.py to preprocess data and generate a vocabulary of symptoms.

Train and test

Train and test models by the follow commands.

Diaformer

# Train and test on Diaformer
# Run on MuZhi dataset
python Diaformer.py --dataset_path data/muzhi_dataset --batch_size 16 --lr 5e-5 --min_probability 0.009 --max_turn 20 --start_test 10 

# Run on Dxy dataset
python Diaformer.py --dataset_path data/dxy_dataset --batch_size 16 --lr 5e-5 --min_probability 0.012 --max_turn 20 --start_test 10 

# Run on Synthetic dataset
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Diaformer_GPT2

# Train and test on GPT2 variant of Diaformer
python GPT2_variant.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Diaformer_UniLM

# Train and test on UniLM variant of Diaformer
python UniLM_variant.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Ablation study

# run ablation study
# w/o Sequence Shuffle
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_sequence_shuffle

# w/o Synchronous Learning
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_synchronous_learning

# w/o Repeated Sequence
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_repeated_sequence

Generative inference

# save the model
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --model_output_path models
# use the trained model to output the results
python predict.py --dataset_path data/synthetic_dataset --min_probability 0.01 --max_turn 20 --pretrained_model models/ --result_output_path results.json

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

Related tags

Overview

Diaformer

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation (AAAI 2022)

Diaformer is an efficient model for automatic diagnosis via symptoms sequence generation. It takes the sequence of symptoms as input, and predicts the inquiry symptoms in the way of sequence generation.

Requirements

Usage

Owner

Junying Chen

RuCLIP-SB (Russian Contrastive Language–Image Pretraining SWIN-BERT) is a multimodal model for obtaining images and text similarities and rearranging captions and pictures. Unlike other versions of the model we use BERT for text encoder and SWIN transformer for image encoder.

Code repository of the paper Neural circuit policies enabling auditable autonomy published in Nature Machine Intelligence

Code for evaluating Japanese pretrained models provided by NTT Ltd.

NeuralQA: A Usable Library for Question Answering on Large Datasets with BERT

ChessCoach is a neural network-based chess engine capable of natural-language commentary.

Various Algorithms for Short Text Mining

NumPy String-Indexed is a NumPy extension that allows arrays to be indexed using descriptive string labels

This is a really simple text-to-speech app made with python and tkinter.

customer care chatbot made with Rasa Open Source.

Google and Stanford University released a new pre-trained model called ELECTRA

LV-BERT: Exploiting Layer Variety for BERT (Findings of ACL 2021)

PyTorch implementation of NATSpeech: A Non-Autoregressive Text-to-Speech Framework

leaking paid token generator that was a shit lmao for 100$ haha

Baseline code for Korean open domain question answering(ODQA)

Searching keywords in PDF file folders

spaCy plugin for Transformers , Udify, ELmo, etc.

Blue Brain text mining toolbox for semantic search and structured information extraction

Words_And_Phrases - Just a repo for useful words and phrases that might come handy in some scenarios. Feel free to add yours

RoNER is a Named Entity Recognition model based on a pre-trained BERT transformer model trained on RONECv2

Client library to download and publish models and other files on the huggingface.co hub