The Submission for SIMMC 2.0 Challenge 2021

Last update: Jul 26, 2022

Related tags

Deep Learning simmc2.0

Overview

The Submission for SIMMC 2.0 Challenge 2021

challenge website

Requirements

python 3.8.8
pytorch 1.8.1
transformers 4.8.2
apex for multi-gpu
nltk

Preprocessing

Download Data

Download the data provided by the challenge organizer and put it in the data folder.
Unzip data files

Image saving

Preprocess the image files in advance. The preprocessed result has the image name as the key and visual as the value.

python3 image_preprocessor.py
python3 image_preprocessor_final.py

The result(.pickle) is saved in res folder.

Step 1 (ITM)

First, the model is post-trained by image-to-text matching. Here, image is each object and text is the visual metadata of the object. Code is provided in the ITM folder.

Step 2 (BTM)

Second, pretraining is performed to use background reprsentation of image in subtasks. Similar to ITM, it is trained to match image and text, and the image is the background of the dialog and the text is the entire context of the dialog. Code is provided in the BTM folder.

Step 3

This is the learning process for each subtask. You can train the model in each folder (sub1, sub2_1, sub2_2, sub2_3, sub2_4, sub4).

Model

All models can be downloaded from the following link

model.pt is a model for evaluating devtest, and the result is saved in the dstc10-simmc-entry folder. model_final.pt is a model for evaluating teststd, and the result is saved in the dstc10-simmc-final-entry folder. However, the training of the model was not completed within the challenge period, so we inferred to model.pt for the teststd data in subtask2.

Evlauation

Using the evaluation script suggested by the challenge organizer

The SIMMC organizers introduce the scripts:

(line-by-line evaluation) $ python -m gpt2_dst.scripts.evaluate \ --input_path_target={PATH_TO_GROUNDTRUTH_TARGET} \ --input_path_predicted={PATH_TO_MODEL_PREDICTIONS} \ --output_path_report={PATH_TO_REPORT} (Or, dialog level evaluation) $ python -m utils.evaluate_dst \ --input_path_target={PATH_TO_GROUNDTRUTH_TARGET} \ --input_path_predicted={PATH_TO_MODEL_PREDICTIONS} \ --output_path_report={PATH_TO_REPORT} $ python tools/response_evaluation.py \ --data_json_path={PATH_TO_GOLD_RESPONSES} \ --model_response_path={PATH_TO_MODEL_RESPONSES} \ --single_round_evaluation $ python tools/retrieval_evaluation.py \ --retrieval_json_path={PATH_TO_GROUNDTRUTH_RETRIEVAL} \ --model_score_path={PATH_TO_MODEL_CANDIDATE_SCORES} \ --single_round_evaluation ">


     
      
$ python tools/disambiguator_evaluation.py \
	--pred_file="{PATH_TO_PRED_FILE}" \
	--test_file="{PATH_TO_TEST_FILE}" \


      
       
(line-by-line evaluation)
$ python -m gpt2_dst.scripts.evaluate \
  --input_path_target={PATH_TO_GROUNDTRUTH_TARGET} \
  --input_path_predicted={PATH_TO_MODEL_PREDICTIONS} \
  --output_path_report={PATH_TO_REPORT}

(Or, dialog level evaluation)
$ python -m utils.evaluate_dst \
    --input_path_target={PATH_TO_GROUNDTRUTH_TARGET} \
    --input_path_predicted={PATH_TO_MODEL_PREDICTIONS} \
    --output_path_report={PATH_TO_REPORT}
    

       
        
$ python tools/response_evaluation.py \
    --data_json_path={PATH_TO_GOLD_RESPONSES} \
    --model_response_path={PATH_TO_MODEL_RESPONSES} \
    --single_round_evaluation


        
         
$ python tools/retrieval_evaluation.py \
    --retrieval_json_path={PATH_TO_GROUNDTRUTH_RETRIEVAL} \
    --model_score_path={PATH_TO_MODEL_CANDIDATE_SCORES} \
    --single_round_evaluation

DevTest Results

Subtask #1: Multimodal Disambiguation

Test Method	Accuracy
GPT2 from CO(Challenge Organizer)	73.9
Ours	92.28

Subtask #2: Multimodal Coreference Resolution

Test Method	Object F1
GPT2 from CO	0.366
Ours-1 (sub2_1)	0.595
Ours-2 (sub2_2)	0.604
Ours-3 (sub2_3)	0.607
Ours-4 (sub2_4)	0.608

Subtask #3: Multimodal Dialog State Tracking

No Training/Testing

Subtask #4: Multimodal Dialog Response Generation

Generation

Baseline	BLEU
GPT2 from CO	0.192
MTN-SIMMC2 from CO	0.217
Ours	0.285

Retrieval

No Training/Testing

The G|oogl|e challenge for Quantum Coalition Hackathon 2021

Qchack 2021 Google Challenge This is a challenge for the brave 2021 qchack.io participants. Instructions Hello, intrepid qchacker, welcome to the G|o

18 May 4, 2022

This repo contains the official code of our work SAM-SLR which won the CVPR 2021 Challenge on Large Scale Signer Independent Isolated Sign Language Recognition.

Skeleton Aware Multi-modal Sign Language Recognition By Songyao Jiang, Bin Sun, Lichen Wang, Yue Bai, Kunpeng Li and Yun Fu. Smile Lab @ Northeastern

128 Dec 8, 2022

The Submission for SIMMC 2.0 Challenge 2021

Related tags

Overview

The Submission for SIMMC 2.0 Challenge 2021

Requirements

Preprocessing

Step 1 (ITM)

Step 2 (BTM)

Step 3

Model

Evlauation

DevTest Results

You might also like...

The G|oogl|e challenge for Quantum Coalition Hackathon 2021

This repo contains the official code of our work SAM-SLR which won the CVPR 2021 Challenge on Large Scale Signer Independent Isolated Sign Language Recognition.

The 1st place solution of track2 (Vehicle Re-Identification) in the NVIDIA AI City Challenge at CVPR 2021 Workshop.

Waymo motion prediction challenge 2021: 3rd place solution

UIUCTF 2021 Public Challenge Repository

Music Source Separation; Train & Eval & Inference piplines and pretrained models we used for 2021 ISMIR MDX Challenge.

EMNLP'2021: Simple Entity-centric Questions Challenge Dense Retrievers

4th place solution for the SIGIR 2021 challenge.

Meli Data Challenge 2021 - First Place Solution

Owner

Submission to Twitter's algorithmic bias bounty challenge

ManiSkill-Learn is a framework for training agents on SAPIEN Open-Source Manipulation Skill Challenge (ManiSkill Challenge), a large-scale learning-from-demonstrations benchmark for object manipulation.

Source code for Zalo AI 2021 submission

(under submission) Bayesian Integration of a Generative Prior for Image Restoration

A Comprehensive Analysis of Weakly-Supervised Semantic Segmentation in Different Image Domains (IJCV submission)

A supplementary code for Editable Neural Networks, an ICLR 2020 submission.

This is the code for our KILT leaderboard submission to the T-REx and zsRE tasks. It includes code for training a DPR model then continuing training with RAG.

This is the pytorch implementation for the paper: Learning Accurate Performance Predictors for Ultrafast Automated Model Compression, which is in submission to TPAMI

Top #1 Submission code for the first https://alphamev.ai MEV competition with best AUC (0.9893) and MSE (0.0982).

CVPR 2021 Challenge on Super-Resolution Space

The Submission for SIMMC 2.0 Challenge 2021

Related tags

Overview

The Submission for SIMMC 2.0 Challenge 2021

Requirements

Preprocessing

Step 1 (ITM)

Step 2 (BTM)

Step 3

Model

Evlauation

DevTest Results

You might also like...

The G|oogl|e challenge for Quantum Coalition Hackathon 2021

This repo contains the official code of our work SAM-SLR which won the CVPR 2021 Challenge on Large Scale Signer Independent Isolated Sign Language Recognition.

The 1st place solution of track2 (Vehicle Re-Identification) in the NVIDIA AI City Challenge at CVPR 2021 Workshop.

Waymo motion prediction challenge 2021: 3rd place solution

UIUCTF 2021 Public Challenge Repository

Music Source Separation; Train & Eval & Inference piplines and pretrained models we used for 2021 ISMIR MDX Challenge.

EMNLP'2021: Simple Entity-centric Questions Challenge Dense Retrievers

4th place solution for the SIGIR 2021 challenge.

Meli Data Challenge 2021 - First Place Solution

Owner

Submission to Twitter's algorithmic bias bounty challenge

ManiSkill-Learn is a framework for training agents on SAPIEN Open-Source Manipulation Skill Challenge (ManiSkill Challenge), a large-scale learning-from-demonstrations benchmark for object manipulation.

Source code for Zalo AI 2021 submission

(under submission) Bayesian Integration of a Generative Prior for Image Restoration

A Comprehensive Analysis of Weakly-Supervised Semantic Segmentation in Different Image Domains (IJCV submission)

A supplementary code for Editable Neural Networks, an ICLR 2020 submission.

This is the code for our KILT leaderboard submission to the T-REx and zsRE tasks. It includes code for training a DPR model then continuing training with RAG.

This is the pytorch implementation for the paper: *Learning Accurate Performance Predictors for Ultrafast Automated Model Compression*, which is in submission to TPAMI

Top #1 Submission code for the first https://alphamev.ai MEV competition with best AUC (0.9893) and MSE (0.0982).

CVPR 2021 Challenge on Super-Resolution Space

This is the pytorch implementation for the paper: Learning Accurate Performance Predictors for Ultrafast Automated Model Compression, which is in submission to TPAMI