• Stars
    star
    153
  • Rank 241,885 (Top 5 %)
  • Language
    Python
  • Created over 4 years ago
  • Updated over 2 years ago

Reviews

There are no reviews yet. Be the first to send feedback to the community and the maintainers!

Repository Details

Code for ACL2020 paper: Few-shot Slot Tagging with Collapsed Dependency Transfer and Label-enhanced Task-adaptive Projection Network

Few-shot Slot Tagging

This is the code of the ACL 2020 paper: Few-shot Slot Tagging with Collapsed Dependency Transfer and Label-enhanced Task-adaptive Projection Network.

Notice: Better implementation availiable now!

  • A new and powerfull platform is now availiable for general few-shot learning problems!!

  • It fully support current experiments with better interface and flexibility~ (E.g. supoort newer huggingface/transformers)

Try it at: https://github.com/AtmaHou/MetaDialog

Get Started

Requirement

python >= 3.6
pytorch >= 0.4.1
pytorch_pretrained_bert >= 0.6.1
allennlp >= 0.8.2
pytorch-nlp

Step1: Prepare BERT embedding:

  • Download the pytorch bert model, or convert tensorflow param by yourself as follow:
export BERT_BASE_DIR=/users4/ythou/Projects/Resources/bert-base-uncased/uncased_L-12_H-768_A-12/

pytorch_pretrained_bert convert_tf_checkpoint_to_pytorch
  $BERT_BASE_DIR/bert_model.ckpt
  $BERT_BASE_DIR/bert_config.json
  $BERT_BASE_DIR/pytorch_model.bin
  • Set BERT path in the file ./scripts/run_L-Tapnet+CDT.sh to your setting:
bert_base_uncased=/your_dir/uncased_L-12_H-768_A-12/
bert_base_uncased_vocab=/your_dir/uncased_L-12_H-768_A-12/vocab.txt

Step2: Prepare data

Tips: The numbers in file name denote cross-evaluation id, you can run a complete experiment by only using data of id=1.

  • Set test, train, dev data file path in ./scripts/run_L-Tapnet+CDT.sh to your setting.

For simplicity, your only need to set the root path for data as follow:

base_data_dir=/your_dir/ACL2020data/

Step3: Train and test the main model

  • Build a folder to collect running log
mkdir result
  • Execute cross-evaluation script with two params: -[gpu id] -[dataset name]
Example for 1-shot Snips:
source ./scripts/run_L-Tapnet+CDT.sh 0 snips
Example for 1-shot NER:
source ./scripts/run_L-Tapnet+CDT.sh 0 ner

To run 5-shots experiments, use ./scripts/run_L-Tapnet+CDT_5.sh

Model for Other Setting

We also provide scripts of four model settings as follows:

  • Tap-Net
  • Tap-Net + CDT
  • L-WPZ + CDT
  • L-Tap-Net + CDT

You can find their corresponding scripts in ./scripts/ with the same usage as above.

Project Architecture

Root

  • the project contains three main parts:
    • models: the neural network architectures
    • scripts: running scripts for cross evaluation
    • utils: auxiliary or tool function files
    • main.py: the entry file of the whole project

models

  • Main Model
    • Sequence Labeler (few_shot_seq_labeler.py): a framework that integrates modules below to perform sequence labeling.
  • Modules
    • Embedder Module (context_embedder_base.py): modules that provide embeddings.
    • Emission Module (emission_scorer_base.py): modules that compute emission scores.
    • Transition Module (transition_scorer.py): modules that compute transition scores.
    • Similarity Module (similarity_scorer_base.py): modules that compute similarities for metric learning based emission scorer.
    • Output Module (seq_labeler.py, conditional_random_field.py): output layer with normal mlp or crf.
    • Scale Module (scale_controller.py): a toolkit for re-scale and normalize logits.

utils

  • utils contains assistance modules for:
    • data processing (data_helper.py, preprocessor.py),
    • constructing model architecture (model_helper.py),
    • controlling training process (trainer.py),
    • controlling testing process (tester.py),
    • controllable parameters definition (opt.py),
    • device definition (device_helper)
    • config (config.py).

Updates - New branch: fix_TapNet_svd_issue

Thanks Wangpeiyi9979 for pointing out the problem of TapNet implementation (issue), which is caused by port differences of cupy.linalg.svd and svd() in pytorch.

The corrected codes are included in new branch named fix_TapNet_svd_issue, because we found correction of TapNet will slightly degrade performance (still the best).

License for code and data

Apache License 2.0

More Repositories

1

Task-Oriented-Dialogue-Research-Progress-Survey

A datasets and methods survey about task-oriented dialogue, including recent datasets and SOTA leaderboards.
1,239
star
2

MetaDialog

Platform for few-shot natural language processing: Text Classification, Sequene Labeling.
Python
219
star
3

FewShotMultiLabel

Code for AAAI2021 paper: Few-Shot Learning for Multi-label Intent Detection.
Python
103
star
4

Seq2SeqDataAugmentationForLU

This repo is code for the COLING 2018 paper: Sequence-to-sequence Data Augmentation for Dialogue Language Understanding.
Python
77
star
5

PromptSlotTagging

Code for ACL22 findings paper: Inverse is Better! Fast and Accurate Prompt for Slot Tagging
Python
26
star
6

Pascal-Simple-Compiler

哈工大编译原理实验 编译器
C++
16
star
7

Bi-LSTM_PosTagger

An easy-to-use sequence labeling project(get SoA on ATIS data) with pytorch
Python
15
star
8

FewShotJoint

Python
12
star
9

UserSimulator

Code for CCL 2019 BEST Poster Paper: A Corpus-free State2Seq User Simulator for Task-oriented Dialogue
OpenEdge ABL
7
star
10

Sequence-to-Sequence-User-Simulator

Python
6
star
11

atma

Light NLP Tool: atma-0.4.1, commonly-used & tested NLP tools: sentence level bleu, tokenizer, proxy crawler included
Python
6
star
12

HIT-secondary-trading-platform-

哈工大二手交易平台
HTML
6
star
13

NNlearning

This repo contains code solution for Stanford CS224N natural language class, and some project relative code.
Python
2
star
14

simple_BP

A simple 3-3-4 three layer neural network for HIT PR homework2
Python
2
star
15

BigPrimeNumber

MSRA homework, implement a big integer packet and generate the k th prime.
C++
1
star
16

Parzen-Window-And-Top-K

Machine Learning non-parameter method: Parzen window and Top-k
Python
1
star
17

git_homework

A good example of Django development ,which is HIT Software Engineering's homework
1
star
18

scene2

for homework
Python
1
star
19

CruelWorld

Game & Machine-Learning attempt?
Python
1
star
20

scene1

for homework scene
HTML
1
star