Active Search for High Recall: A Non-stationary Extension of Thompson Sampling

Published 2017 in European Conference on Information Retrieval

ABSTRACT

We consider the problem of Active Search, where a maximum of relevant objects - ideally all relevant objects - should be retrieved with the minimum effort or minimum time. Typically, there are two main challenges to face when tackling this problem: first, the class of relevant objects has often low prevalence and, secondly, this class can be multi-faceted or multi-modal: objects could be relevant for completely different reasons. To solve this problem and its associated issues, we propose an approach based on a non-stationary (aka restless) extension of Thompson Sampling, a well-known strategy for Multi-Armed Bandits problems. The collection is first soft-clustered into a finite set of components and a posterior distribution of getting a relevant object inside each cluster is updated after receiving the user feedback about the proposed instances. The "next instance" selection strategy is a mixed, two-level decision process, where both the soft clusters and their instances are considered. This method can be considered as an insurance, where the cost of the insurance is an extra exploration effort in the short run, for achieving a nearly "total" recall with less efforts in the long run.

PUBLICATION RECORD

Publication year
2017
Venue
European Conference on Information Retrieval
Publication date
2017-12-27
Fields of study
Computer Science
Identifiers
DOI 10.1007/978-3-319-76941-7_68 arXiv 1712.09550
External record
Open on Semantic Scholar
Source metadata
Semantic Scholar

CITATION MAP

EXTRACTION MAP

CLAIMS

No claims are published for this paper.

CONCEPTS

No concepts are published for this paper.

REFERENCES

Scalability of Continuous Active Learning for Reliable High-Recall Text Classification
2016cited by this paper
Optimism in Active Learning
2015cited by this paper
Active Search and Bandits on Graphs using Sigma-Optimality
2015cited by this paper
WaterlooClarke: TREC 2015 Total Recall Track
2015cited by this paper
Multi-armed bandit problem with known trend
2015cited by this paper
Active Exploration in Networks: Using Probabilistic Relationships for Learning and Inference
2014cited by this paper
Contextual Bandit for Active Learning: Active Thompson Sampling
2014cited by this paper
Simple and Scalable Response Prediction for Display Advertising
2014cited by this paper
Introduction to special section on intelligent mobile knowledge discovery and management systems
2013cited by this paper
Bayesian Optimal Active Search and Surveying
2012cited by this paper

CITED BY

SeeSaw: Interactive Ad-hoc Search Over Image Databases
2022cites this paper
Active Search using Meta-Bandits
2020cites this paper
Cost Effective Active Search
2019cites this paper