Code for the paper: Sequence-to-Sequence Learning with Latent Neural Grammars

Last update: Dec 23, 2022

Related tags

Overview

Sequence-to-Sequence Learning with Latent Neural Grammars

Code for the paper:
Sequence-to-Sequence Learning with Latent Neural Grammars
Yoon Kim
arXiv Preprint

Dependencies

The code was tested in python 3.7 and pytorch 1.5. We also use a slightly modified version of the Torch-Struct library, which is included in the repo and can be installed via:

cd pytorch-struct
python setup.py install

Data

For convenience we include the datasets used in the paper in the data/ folder. Please cite the original papers when using the data (i.e. Lake and Baroni 2018 for SCAN/MT, and Lyu et al. 2021 for StylePTB).

Training

SCAN

To train the model on (for example) the length split:

python train_scan.py --train_file data/SCAN/tasks_train_length.txt --save_path scan-length.pt

For prediction and evaluation:

python predict_scan.py --data_file data/SCAN/tasks_test_length.txt --model_path scan-length.pt

Style Transfer

To train on (for example) the active-to-passive task:

python train_styleptb.py --train_file data/StylePTB/ATP/train.tsv --dev_file data/StylePTB/ATP/valid.tsv --save_path styleptb-atp.pt

To predict:

python predict_styleptb.py --data_file data/StylePTB/ATP/test.tsv --model_path styleptb-atp.pt 
--out_file styleptb-atp-pred.txt

We use the nlg-eval package to calculate the various metrics.

Machine Translation

To train on MT:

python train_mt.py --train_file_src data/MT/train.en --train_file_tgt data/MT/train.fr 
--dev_file_src data/MT/dev.en --dev_file_tgt data/MT/dev.fr --save_path mt.pt

To predict on the daxy test set:

python predict_mt.py --data_file data/MT/test-daxy.en --model_path mt.pt --out_file mt-pred-daxy.txt

For the regular test set:

python predict_mt.py --data_file data/MT/test.en --model_path mt.pt --out_file mt-pred.txt

We use the multi-bleu script to calculate BLEU.

Training Stability

We observed training to be unstable and the approach required several runs across different seeds to perform well. For reference we have posted logs of some example runs in the logs/ folder.

License

MIT

Code for the paper: Sequence-to-Sequence Learning with Latent Neural Grammars

Related tags

Overview

Sequence-to-Sequence Learning with Latent Neural Grammars

Dependencies

Data

Training

SCAN

Style Transfer

Machine Translation

Training Stability

License

Owner

Yoon Kim

Recognition of 38 speech commands in russian. Based on Yandex Cup 2021 ML Challenge: ASR

中文生成式预训练模型

A Streamlit web app that generates Rick and Morty stories using GPT2.

本插件是pcrjjc插件的重置版，可以独立于后端api运行

A practical and feature-rich paraphrasing framework to augment human intents in text form to build robust NLU models for conversational engines. Created by Prithiviraj Damodaran. Open to pull requests and other forms of collaboration.

Py65 65816 - Add support for the 65C816 to py65

Full Spectrum Bioinformatics - a free online text designed to introduce key topics in Bioinformatics using the Python

TextFlint is a multilingual robustness evaluation platform for natural language processing tasks,

Galois is an auto code completer for code editors (or any text editor) based on OpenAI GPT-2.

The Classical Language Toolkit

Translate U is capable of translating the text present in an image from one language to the other.

Paddlespeech Streaming ASR GUI

Material for GW4SHM workshop, 16/03/2022.

Journey is a NLP-Powered Developer assistant

A tool helps build a talk preview image by combining the given background image and talk event description

:mag: Transformers at scale for question answering & neural search. Using NLP via a modular Retriever-Reader-Pipeline. Supporting DPR, Elasticsearch, HuggingFace's Modelhub...

Random Directed Acyclic Graph Generator

Meta learning algorithms to train cross-lingual NLI (multi-task) models

SentAugment is a data augmentation technique for semi-supervised learning in NLP.

A unified tokenization tool for Images, Chinese and English.