The 2nd place solution of 2021 google landmark retrieval on kaggle.

Last update: Dec 13, 2022

Related tags

Deep Learning Google_Landmark_Retrieval_2021_2nd_Place_Solution

Overview

Google_Landmark_Retrieval_2021_2nd_Place_Solution

The 2nd place solution of 2021 google landmark retrieval on kaggle.

Environment

We use cuda 11.1/python 3.7/torch 1.9.1/torchvision 0.8.1 for training and testing.

Download imagenet pretrained model ResNeXt101ibn and SEResNet101ibn from IBN-Net. ResNest101 and ResNeSt269 can be found in ResNest.

Prepare data

Download GLDv2 full version from the official site.
Run python tools/generate_gld_list.py. This will generate clean, c2x, trainfull and all data for different stage of training.
Validation annotation comes from all 1129 images in GLDv2. We expand the competition index set to index_expand. Each query could find all its GTs in the expanded index set and the validation could be more accurate.

Train

We use 8 GPU (32GB/16GB) for training. The evaluation metric in landmark retrieval is different from person re-identification. Due to the validation scale, we skip the validation stage during training and just use the model from last epoch for evaluation.

Fast Train Script

To make quick experiments, we provide scripts for R50_256 trained for clean subset. This setting trains very fast and is helpful for debug.

python -m torch.distributed.run --standalone --nnodes=1 --nproc_per_node=8 --master_port 55555 --max_restarts 0 train.py --config_file configs/GLDv2/R50_256.yml

Whole Train Pipeline

The whole training pipeline for SER101ibn backbone is listed below. Other backbones and input size can be modified accordingly.

python -m torch.distributed.run --standalone --nnodes=1 --nproc_per_node=8 --master_port 55555 --max_restarts 0 train.py --config_file configs/GLDv2/SER101ibn_384.yml
python -m torch.distributed.run --standalone --nnodes=1 --nproc_per_node=8 --master_port 55555 --max_restarts 0 train.py --config_file configs/GLDv2/SER101ibn_384_finetune.yml
python -m torch.distributed.run --standalone --nnodes=1 --nproc_per_node=8 --master_port 55555 --max_restarts 0 train.py --config_file configs/GLDv2/SER101ibn_512_finetune.yml
python -m torch.distributed.run --standalone --nnodes=1 --nproc_per_node=8 --master_port 55555 --max_restarts 0 train.py --config_file configs/GLDv2/SER101ibn_512_all.yml

Inference(notebooks)

With four models trained, cd to submission/code/ and modify settings in landmark_retrieval.py properly.
Then run eval_retrieval.sh to get submission file and evaluate on validation set offline.

General Settings

REID_EXTRACT_FLAG: Skip feature extraction when using offline code.
FEAT_DIR: Save cached features.
IMAGE_DIR: competition image dir. We make a soft link for competition data at submission/input/landmark-retrieval-2021/
RAW_IMAGE_DIR: origin GLDv2 dir
MODEL_DIR: the latest models for submission
META_DIR: saves meta files for rerank purpose
LOCAL_MATCHING and KR_FLAG disabled for our submission.

Fast Inference Script

Use R50_256 model trained from clean subset correspongding to the fast train script. Set CATEGORY_RERANK and REF_SET_EXTRACT to False. You will get about mAP=32.84% for the validation set.

Whole Inference Pipeline

Copy cache_all_list.pkl, cache_index_train_list.pkl and cache_full_list.pkl from cache to submission/input/meta-data-final
Set REF_SET_EXTRACT to True to extract features for all images of GLDv2. This will save about 4.9 million 512 dim features for each model in submission/input/meta-data-final.
Set REF_SET_EXTRACT to False and CATEGORY_RERANK to before_merge. This will load the precomputed features and run the proposed Landmark-Country aware rerank.
The notebooks on kaggle is exactly the same file as in base_landmark.py and landmark_retrieval.py. We also upload the same notebooks as in kaggle in kaggle.ipynb.

Kaggle and ICCV workshops

The challenge is held on kaggle and the leaderboard can be found here. We rank 2nd(2/263) in this challenge.
Kaggle Discussion post link here
ICCV workshop slides coming soon.

Thanks

The code is motivated by AICITY2021_Track2_DMT, 2020_1st_recognition_solution, 2020_2nd_recognition_solution, 2020_1st_retrieval_solution.

Citation

If you find our work useful in your research, please consider citing:

@inproceedings{zhang2021landmark,
 title={2nd Place Solution to Google Landmark Retrieval 2021},
 author={Zhang, Yuqi and Xu, Xianzhe and Chen, Weihua and Wang, Yaohua and Zhang, Fangyi},
 year={2021}
}

You might also like...

🏆 The 1st Place Submission to AICity Challenge 2021 Natural Language-Based Vehicle Retrieval Track (Alibaba-UTS submission)

AI City 2021: Connecting Language and Vision for Natural Language-Based Vehicle Retrieval 🏆 The 1st Place Submission to AICity Challenge 2021 Natural

82 Dec 29, 2022

The 1st place solution of track2 (Vehicle Re-Identification) in the NVIDIA AI City Challenge at CVPR 2021 Workshop.

AICITY2021_Track2_DMT The 1st place solution of track2 (Vehicle Re-Identification) in the NVIDIA AI City Challenge at CVPR 2021 Workshop. Introduction

91 Dec 21, 2022

Waymo motion prediction challenge 2021: 3rd place solution

Waymo motion prediction challenge 2021: 3rd place solution 📜 Technical report 🗨️ Presentation 🎉 Announcement 🛆Motion Prediction Channel Website 🛆

158 Jan 8, 2023

BirdCLEF 2021 - Birdcall Identification 4th place solution

BirdCLEF 2021 - Birdcall Identification 4th place solution My solution detail kaggle discussion Inference Notebook (best submission) Environment Use K

42 Jan 2, 2023

4th place solution for the SIGIR 2021 challenge.

SIGIR-2021 (Tinkoff.AI) How to start Download train and test data: https://sigir-ecom.github.io/data-task.html Place it under sigir-2021/data/. Run py

4 Jul 1, 2022

Meli Data Challenge 2021 - First Place Solution

My solution for the Meli Data Challenge 2021

23 Mar 9, 2022

The sixth place winning solution (6/220) in 2021 Gaofen Challenge.

SwinTransformer + OBBDet The sixth place winning solution (6/220) in the track of Fine-grained Object Recognition in High-Resolution Optical Images, 2

46 Dec 2, 2022

Codebase for the solution that won first place and was awarded the most human-like agent in the 2021 NeurIPS Competition MineRL BASALT Challenge.

KAIROS MineRL BASALT Codebase for the solution that won first place and was awarded the most human-like agent in the 2021 NeurIPS Competition MineRL B

37 Oct 30, 2022

1st place solution in CCF BDCI 2021 ULSEG challenge

1st place solution in CCF BDCI 2021 ULSEG challenge This is the source code of the 1st place solution for ultrasound image angioma segmentation task (

30 Nov 22, 2022

Comments

关于模型forward的疑惑

在make_model.py的第253行: return cls_score, global_feat # global feature for triplet loss 这里给triplet loss的global_feat是经过了bn之后的特征，请问作者这里用了bn之后的特征而不是之前的特征的原因。我理解的是triplet更适用之前的欧式空间的特征进行度量、而bn后的特征近似球面、适用于arc而非triplet。个人理解不一定准确、因此请教下作者，望解答，谢谢！

opened by Wzj02200059 2
landmark retireval 的问题【Questions from email】

您好, 我在github 上阅读了您的论文关于GLD-v2数据集的图像检索。有个问题想请教一下关于评估。我所下载的数据集GLD-v2中大部分类别都是小于10张的。所以请问map@100这个评估方法你们是在所有的训练数据集中使用的吗？还是只是对于图像超过100张的类别计算的?

很期待您的回复 Tianyi Hu

opened by WesleyZhang1991 1

The 2nd place solution of 2021 google landmark retrieval on kaggle.

Related tags

Overview

Google_Landmark_Retrieval_2021_2nd_Place_Solution

Environment

Prepare data

Train

Fast Train Script

Whole Train Pipeline

Inference(notebooks)

General Settings

Fast Inference Script

Whole Inference Pipeline

Kaggle and ICCV workshops

Thanks

Citation

You might also like...

🏆 The 1st Place Submission to AICity Challenge 2021 Natural Language-Based Vehicle Retrieval Track (Alibaba-UTS submission)

The 1st place solution of track2 (Vehicle Re-Identification) in the NVIDIA AI City Challenge at CVPR 2021 Workshop.

Waymo motion prediction challenge 2021: 3rd place solution

BirdCLEF 2021 - Birdcall Identification 4th place solution

4th place solution for the SIGIR 2021 challenge.

Meli Data Challenge 2021 - First Place Solution

The sixth place winning solution (6/220) in 2021 Gaofen Challenge.

Codebase for the solution that won first place and was awarded the most human-like agent in the 2021 NeurIPS Competition MineRL BASALT Challenge.

1st place solution in CCF BDCI 2021 ULSEG challenge

Comments

关于模型forward的疑惑

landmark retireval 的问题【Questions from email】

Owner

This is the solution for 2nd rank in Kaggle competition: Feedback Prize - Evaluating Student Writing.

10th place solution for Google Smartphone Decimeter Challenge at kaggle.

Google Landmark Recogntion and Retrieval 2021 Solutions

Simple Linear 2nd ODE Solver GUI - A 2nd constant coefficient linear ODE solver with simple GUI using euler's method

My 1st place solution at Kaggle Hotel-ID 2021

Kaggle Lyft Motion Prediction for Autonomous Vehicles 4th place solution

7th place solution of Human Protein Atlas - Single Cell Classification on Kaggle

Kaggle | 9th place (part of) solution for the Bristol-Myers Squibb – Molecular Translation challenge

Kaggle | 9th place single model solution for TGS Salt Identification Challenge

2nd solution of ICDAR 2021 Competition on Scientific Literature Parsing, Task B.