Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness

Last update: Nov 29, 2022

Related tags

Deep Learning 2021-SIGIR-plackett-luce

Overview

Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness

This repository contains the code used for the experiments in "Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness" published at SIGIR 2021 (preprint available).

Citation

If you use this code to produce results for your scientific publication, or if you share a copy or fork, please refer to our SIGIR 2021 paper:

@inproceedings{oosterhuis2021plrank,
  Author = {Oosterhuis, Harrie},
  Booktitle = {Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR`21)},
  Organization = {ACM},
  Title = {Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness},
  Year = {2021}
}

License

The contents of this repository are licensed under the MIT license. If you modify its contents in any way, please link back to this repository.

Usage

This code makes use of Python 3, the numpy and the tensorflow packages, make sure they are installed.

A file is required that explains the location and details of the LTR datasets available on the system, for the Yahoo! Webscope, MSLR-Web30k, and Istella datasets an example file is available. Copy the file:

cp example_datasets_info.txt local_dataset_info.txt

Open this copy and edit the paths to the folders where the train/test/vali files are placed.

Here are some command-line examples that illustrate how the results in the paper can be replicated. First create a folder to store the resulting models:

mkdir local_output

To optimize NDCG use run.py with the --loss flag to indicate the loss to use (PL_rank_1/PL_rank_2/lambdaloss/pairwise/policygradient/placementpolicygradient); --cutoff indicates the top-k that is being optimized, e.g. 5 for [email protected]; --num_samples the number of samples to use per gradient estimation (with dynamic for the dynamic strategy); --dataset indicates the dataset name, e.g. Webscope_C14_Set1. The following command optimizes [email protected] with PL-Rank-2 and the dynamic sampling strategy on the Yahoo! dataset:

python3 run.py local_output/yahoo_ndcg5_dynamic_plrank2.txt --num_samples dynamic --loss PL_rank_2 --cutoff 5 --dataset Webscope_C14_Set1

To optimize the disparity metric for exposure fairness use fairrun.py this has the additional flag --num_exposure_samples for the number of samples to use to estimate exposure (this must always be a greater number than --num_samples). The following command optimizes disparity with PL-Rank-2 and the dynamic sampling strategy on the Yahoo! dataset with 1000 samples for estimating exposure:

python3 fairrun.py local_output/yahoo_fairness_dynamic_plrank2.txt --num_samples dynamic --loss PL_rank_2 --cutoff 5 --num_exposure_samples 1000 --dataset Webscope_C14_Set1

Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness

Related tags

Overview

Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness

Citation

License

Usage

Owner

H.R. Oosterhuis

李云龙二次元风格化!打滚卖萌，使用了animeGANv2进行了视频的风格迁移

A collection of differentiable SVD methods and also the official implementation of the ICCV21 paper "Why Approximate Matrix Square Root Outperforms Accurate SVD in Global Covariance Pooling?"

Deep Q Learning with OpenAI Gym and Pokemon Showdown

Self-supervised Augmentation Consistency for Adapting Semantic Segmentation (CVPR 2021)

OBG-FCN - implementation of 'Object Boundary Guided Semantic Segmentation'

Paddle Graph Learning (PGL) is an efficient and flexible graph learning framework based on PaddlePaddle

The official implementation of You Only Compress Once: Towards Effective and Elastic BERT Compression via Exploit-Explore Stochastic Nature Gradient.

The 3rd place solution for competition

FL-WBC: Enhancing Robustness against Model Poisoning Attacks in Federated Learning from a Client Perspective

PyTorch implementation of the Crafting Better Contrastive Views for Siamese Representation Learning

Sample Code for "Pessimism Meets Invariance: Provably Efficient Offline Mean-Field Multi-Agent RL"

DeepLM: Large-scale Nonlinear Least Squares on Deep Learning Frameworks using Stochastic Domain Decomposition (CVPR 2021)

GRF: Learning a General Radiance Field for 3D Representation and Rendering

Data stream analytics: Implement online learning methods to address concept drift in data streams using the River library. Code for the paper entitled "PWPAE: An Ensemble Framework for Concept Drift Adaptation in IoT Data Streams" accepted in IEEE GlobeCom 2021.

Fully Convolutional DenseNets for semantic segmentation.

Official implementation for “Unsupervised Low-Light Image Enhancement via Histogram Equalization Prior”

Tutorials, assignments, and competitions for MIT Deep Learning related courses.

“英特尔创新大师杯”深度学习挑战赛赛道3：CCKS2021中文NLP地址相关性任务

Dyalog-apl-docset - Dyalog APL Dash Docset Generator

The PyTorch implementation of DiscoBox: Weakly Supervised Instance Segmentation and Semantic Correspondence from Box Supervision.

Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness

Related tags

Overview

Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and Fairness

Citation

License

Usage

Owner

H.R. Oosterhuis

李云龙二次元风格化!打滚卖萌，使用了animeGANv2进行了视频的风格迁移

A collection of differentiable SVD methods and also the official implementation of the ICCV21 paper "Why Approximate Matrix Square Root Outperforms Accurate SVD in Global Covariance Pooling?"

Deep Q Learning with OpenAI Gym and Pokemon Showdown

Self-supervised Augmentation Consistency for Adapting Semantic Segmentation (CVPR 2021)

OBG-FCN - implementation of 'Object Boundary Guided Semantic Segmentation'

Paddle Graph Learning (PGL) is an efficient and flexible graph learning framework based on PaddlePaddle

The official implementation of You Only Compress Once: Towards Effective and Elastic BERT Compression via Exploit-Explore Stochastic Nature Gradient.

The 3rd place solution for competition

FL-WBC: Enhancing Robustness against Model Poisoning Attacks in Federated Learning from a Client Perspective

PyTorch implementation of the Crafting Better Contrastive Views for Siamese Representation Learning

Sample Code for "Pessimism Meets Invariance: Provably Efficient Offline Mean-Field Multi-Agent RL"

DeepLM: Large-scale Nonlinear Least Squares on Deep Learning Frameworks using Stochastic Domain Decomposition (CVPR 2021)

GRF: Learning a General Radiance Field for 3D Representation and Rendering

Data stream analytics: Implement online learning methods to address concept drift in data streams using the River library. Code for the paper entitled "PWPAE: An Ensemble Framework for Concept Drift Adaptation in IoT Data Streams" accepted in IEEE GlobeCom 2021.

Fully Convolutional DenseNets for semantic segmentation.

Official implementation for “Unsupervised Low-Light Image Enhancement via Histogram Equalization Prior”

Tutorials, assignments, and competitions for MIT Deep Learning related courses.

“英特尔创新大师杯”深度学习挑战赛 赛道3：CCKS2021中文NLP地址相关性任务

Dyalog-apl-docset - Dyalog APL Dash Docset Generator

The PyTorch implementation of DiscoBox: Weakly Supervised Instance Segmentation and Semantic Correspondence from Box Supervision.

“英特尔创新大师杯”深度学习挑战赛赛道3：CCKS2021中文NLP地址相关性任务