Codebase for BMVC 2021 paper "Text Based Person Search with Limited Data"

Last update: Nov 24, 2022

Related tags

Deep Learning TextReID

Overview

Text Based Person Search with Limited Data

This is the codebase for our BMVC 2021 paper.

Please bear with me refactoring this codebase after CVPR deadline 😅

Abstract

Text-based person search (TBPS) aims at retrieving a target person from an image gallery with a descriptive text query. Solving such a fine-grained cross-modal retrieval task is challenging, which is further hampered by the lack of large-scale datasets. In this paper, we present a framework with two novel components to handle the problems brought by limited data. Firstly, to fully utilize the existing small-scale benchmarking datasets for more discriminative feature learning, we introduce a cross-modal momentum contrastive learning framework to enrich the training data for a given mini-batch. Secondly, we propose to transfer knowledge learned from existing coarse-grained large-scale datasets containing image-text pairs from drastically different problem domains to compensate for the lack of TBPS training data. A transfer learning method is designed so that useful information can be transferred despite the large domain gap. Armed with these components, our method achieves new state of the art on the CUHK-PEDES dataset with significant improvements over the prior art in terms of Rank-1 and mAP.

Comments

Research prepared to obtain a diploma degree in computer and Automation Engineering.

Hello!

My research focuses on Person search using Visual-Textual Attributes. Having said that, I would like to use your model to assist me in my project, but I have some issues when I finish train and test the model. My problem is trying to write code to run the model to get the same response as the photo. so Can you help me please!

opened by ram7772 6
Cannot find test_query and train_query folders
Hi @BrandonHanx

In the ReadMe file, it is mentioned to setup the datasets dir as follows:

└── cuhkpedes ├── annotations │ ├── test.json │ ├── train.json │ └── val.json ├── clip_vocab_vit.npy └── imgs ├── cam_a ├── cam_b ├── CUHK01 ├── CUHK03 ├── Market ├── test_query └── train_query

After downloading the cuhkpedes data set, we get only the imgs folder, containing cam_a, cam_b and CUHK01 folders. there is no test_query and train_query folders. Also, these folders are not in the repository. Could you provide more information regarding on these folders, more exactly, what kind of information they contain and how they must be set up?

Also, there are few more folders that are not part of the cuhkpedes, such as CUHK03 and Market. Do we need these data sets to reproduce the results?

Best regards, liviust
opened by liviust 5
some problem in training and testing

Hello

I have some problem. first: I don't find test_query and train_query file when I get images from [Dr. Shuang Li] second: I have this problem for testing and training.

opened by ram7772 4
Problem about the clip_vocab_vit.npy

Hi :) I have a question about the pre-processing document clip_vocab_vit.npy. My understanding is that it contains the tensor of the CLIP-Text-Encoder output corresponding to each word (total 9408). My question is, the output dimension of CLIP-TEXT-ENCODER is 1024, but the tensor dimension of each word in clip_vocab_vit.npy is 512. Is there some other operation in it? Thanks

opened by Frost-Yang-99 2
There is only caption_all.json in the dataset CUHK-PEDES, what are the train.json and test.json in the dataset part
Describe the bug A clear and concise description of what the bug is.

To Reproduce Steps to reproduce the behavior:

Go to '...'

Click on '....'

Scroll down to '....'

See error

Expected behavior A clear and concise description of what you expected to happen.

Screenshots If applicable, add screenshots to help explain your problem.

Desktop (please complete the following information):

OS: [e.g. iOS]

Browser [e.g. chrome, safari]

Version [e.g. 22]

Smartphone (please complete the following information):

Device: [e.g. iPhone6]

OS: [e.g. iOS8.1]

Browser [e.g. stock browser, safari]

Version [e.g. 22]

Additional context Add any other context about the problem here.
opened by SwimKY 1

Releases(v0.1.1)

v0.1.1(Dec 10, 2021)

Full Changelog: https://github.com/BrandonHanx/TextReID/compare/v0.1.0...v0.1.1
Source code(tar.gz)
Source code(zip)

Owner

Xiao Han

Ph.D. student @ UoSurrey CVSSP, B.Eng. @ ZJU ISEE

GitHub Repository

Implementation of DocFormer: End-to-End Transformer for Document Understanding, a multi-modal transformer based architecture for the task of Visual Document Understanding (VDU)

DocFormer - PyTorch Implementation of DocFormer: End-to-End Transformer for Document Understanding, a multi-modal transformer based architecture for t

171 Jan 06, 2023

Codebase for BMVC 2021 paper "Text Based Person Search with Limited Data"

Related tags

Overview

Text Based Person Search with Limited Data

Abstract

Comments

Research prepared to obtain a diploma degree in computer and Automation Engineering.

Cannot find test_query and train_query folders

some problem in training and testing

Problem about the clip_vocab_vit.npy

There is only caption_all.json in the dataset CUHK-PEDES, what are the train.json and test.json in the dataset part

Releases(v0.1.1)

v0.1.1(Dec 10, 2021)

Owner

Xiao Han

Implementation of DocFormer: End-to-End Transformer for Document Understanding, a multi-modal transformer based architecture for the task of Visual Document Understanding (VDU)

Official PyTorch Implementation of Unsupervised Learning of Scene Flow Estimation Fusing with Local Rigidity

Rocket-recycling with Reinforcement Learning

Fast, flexible and easy to use probabilistic modelling in Python.

On the adaptation of recurrent neural networks for system identification

The pytorch implementation of the paper "text-guided neural image inpainting" at MM'2020

The implementation of "Bootstrapping Semantic Segmentation with Regional Contrast".

Notes taking website build with Docker + Django + React.

Regulatory Instruments for Fair Personalized Pricing.

Py4fi2nd - Jupyter Notebooks and code for Python for Finance (2nd ed., O'Reilly) by Yves Hilpisch.

Code for paper "Which Training Methods for GANs do actually Converge? (ICML 2018)"

[ArXiv 2021] One-Shot Generative Domain Adaptation

Code for our paper at ECCV 2020: Post-Training Piecewise Linear Quantization for Deep Neural Networks

Sound and Cost-effective Fuzzing of Stripped Binaries by Incremental and Stochastic Rewriting

This tool uses Deep Learning to help you draw and write with your hand and webcam.

Code release of paper Improving neural implicit surfaces geometry with patch warping

Elevation Mapping on GPU.

FaceAnon - Anonymize people in images and videos using yolov5-crowdhuman

Official implementation for paper: A Latent Transformer for Disentangled Face Editing in Images and Videos.

Code for approximate graph reduction techniques for cardinality-based DSFM, from paper