Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

Last update: Dec 13, 2022

Related tags

Text Data & NLP Diaformer

Overview

Diaformer

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation (AAAI 2022)

Diaformer is an efficient model for automatic diagnosis via symptoms sequence generation. It takes the sequence of symptoms as input, and predicts the inquiry symptoms in the way of sequence generation.

Figure 1: Illustration of symptom attention framework.

Requirements

Our experiments are conducted on Python 3.8 and Pytorch == 1.8.0. The main requirements are:

transformers==2.1.1
torch
numpy
tqdm
sklearn
keras
boto3

In the root directory, run following command to install the required libraries.

pip install -r requirement.txt

Usage

Download data

Download the datasets, then decompress them and put them in the corrsponding documents in \data. For example, put the data of Synthetic Dataset under data/synthetic_dataset.

The dataset can be downloaded as following links:
Build data

Switch to the corresponding directory of the dataset and just run preprocess.py to preprocess data and generate a vocabulary of symptoms.

Train and test

Train and test models by the follow commands.

Diaformer

# Train and test on Diaformer
# Run on MuZhi dataset
python Diaformer.py --dataset_path data/muzhi_dataset --batch_size 16 --lr 5e-5 --min_probability 0.009 --max_turn 20 --start_test 10 

# Run on Dxy dataset
python Diaformer.py --dataset_path data/dxy_dataset --batch_size 16 --lr 5e-5 --min_probability 0.012 --max_turn 20 --start_test 10 

# Run on Synthetic dataset
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Diaformer_GPT2

# Train and test on GPT2 variant of Diaformer
python GPT2_variant.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Diaformer_UniLM

# Train and test on UniLM variant of Diaformer
python UniLM_variant.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10

Ablation study

# run ablation study
# w/o Sequence Shuffle
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_sequence_shuffle

# w/o Synchronous Learning
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_synchronous_learning

# w/o Repeated Sequence
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --no_repeated_sequence

Generative inference

# save the model
python Diaformer.py --dataset_path data/synthetic_dataset --batch_size 16 --lr 5e-5 --min_probability 0.01 --max_turn 20 --start_test 10 --model_output_path models
# use the trained model to output the results
python predict.py --dataset_path data/synthetic_dataset --min_probability 0.01 --max_turn 20 --pretrained_model models/ --result_output_path results.json

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

Related tags

Overview

Diaformer

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation (AAAI 2022)

Diaformer is an efficient model for automatic diagnosis via symptoms sequence generation. It takes the sequence of symptoms as input, and predicts the inquiry symptoms in the way of sequence generation.

Requirements

Usage

Owner

Junying Chen

🚀 RocketQA, dense retrieval for information retrieval and question answering, including both Chinese and English state-of-the-art models.

Client library to download and publish models and other files on the huggingface.co hub

KakaoBrain KoGPT (Korean Generative Pre-trained Transformer)

Pre-Training with Whole Word Masking for Chinese BERT

This is a modification of the OpenAI-CLIP repository of moein-shariatnia

Generate product descriptions, blogs, ads and more using GPT architecture with a single request to TextCortex API a.k.a Hemingwai

A design of MIDI language for music generation task, specifically for Natural Language Processing (NLP) models.

GrammarTagger — A Neural Multilingual Grammar Profiler for Language Learning

Sentiment Analysis Project using Count Vectorizer and TF-IDF Vectorizer

NLP command-line assistant powered by OpenAI

My Implementation for the paper EDA: Easy Data Augmentation Techniques for Boosting Performance on Text Classification Tasks using Tensorflow

Code for EMNLP 2021 main conference paper "Text AutoAugment: Learning Compositional Augmentation Policy for Text Classification"

StarGAN - Official PyTorch Implementation

Contains descriptions and code of the mini-projects developed in various programming languages

Diaformer: Automatic Diagnosis via Symptoms Sequence Generation

iBOT: Image BERT Pre-Training with Online Tokenizer

🦆 Contextually-keyed word vectors

Fuzzy String Matching in Python

A paper list of pre-trained language models (PLMs).

KoBERTopic은 BERTopic을 한국어 데이터에 적용할 수 있도록 토크나이저와 BERT를 수정한 코드입니다.