Active search for high recall: a non-stationary extension of Thompson sampling

Published by NAVER LABS Europe at 27 February 2018

European Conference on Information Retrieval (ECIR), Grenoble, France, 26-28 March, 2018, Proceedings, pp. 722-728 (Part of LNCS, volume 10772)

Download

@book{book,
  author    = {Jean-Michel Renders}, 
  title     = {Active Search for High Recall: a Non-Stationary Extension of Thompson Sampling},
  publisher = {Springer, Cham},
  year      = 2018,
  volume    = 10772,
  address   = {The address},
  month     = March,
  isbn      = {978-3-319-76940-0}
}

Careers home

We consider the problem of Active Search, where a maximum of relevant objects – ideally all relevant objects – should be retrieved with the minimum effort or minimum time. Typically, there are two main challenges to face when tackling this problem: first, the class of relevant objects has often low prevalence and, secondly, this class can be multi-faceted or multi-modal: objects could be relevant for completely different reasons. To solve this problem and its associated issues, we propose an approach based on a non-stationary (aka restless) extension of Thompson Sampling, a well-known strategy for Multi-Armed Bandits problems. The collection is first soft-clustered into a finite set of components and a posterior distribution of getting a relevant object inside each cluster is updated after receiving the user feedback about the proposed instances. The “next instance” selection strategy is a mixed, two-level decision process, where both the soft clusters and their instances are considered. This method can be considered as an insurance, where the cost of the insurance is an extra exploration effort in the short run, for achieving a nearly “total” recall with less efforts in the long run.

INTERACTION

Equip robots to interact safely with humans, other robots and systems.

VISION

Perception to help robots understand and interact with the environment.

ACTION

Providing embodied agents with sequential decision-making capabilities to safely execute complex tasks in dynamic environments.

NAVER FRANCE Gender Equality 2026

All

Publications

Blog

News

Code & Data

Careers

People

Active search for high recall: a non-stationary extension of Thompson sampling

All

Publications

Blog

News

Code & Data

Careers

People

Cookie settings