paper-with-me

Papers

Monotonic Cardinality Estimation of Similarity Selection: A Deep Learning Approach

2020-02-15 · Yaoshu Wang, Chuan Xiao, Jianbin Qin, Xin Cao, Yifang Sun, Wei Wang, Makoto Onizuka

Due to the outstanding capability of capturing underlying data distributions, deep learning techniques have been recently utilized for a series of traditional database problems. In this paper, we investigate the possibilities of utilizing deep learning for cardinality estimation of similarity selection. Answering this problem accurately and efficiently is essential to many data management applications, especially for query optimization. Moreover, in some applications the estimated cardinality is supposed to be consistent and interpretable. Hence a monotonic estimation w.r.t. the query threshold is preferred. We propose a novel and generic method that can be applied to any data type and distance function. Our method consists of a feature extraction model and a regression model. The feature extraction model transforms original data and threshold to a Hamming space, in which a deep learning-based regression model is utilized to exploit the incremental property of cardinality w.r.t. the threshold for both accuracy and monotonicity. We develop a training strategy tailored to our model as well as techniques for fast estimation. We also discuss how to handle updates. We demonstrate the accuracy and the efficiency of our method through experiments, and show how it improves the performance of a query optimizer.

📄 PDF Abstract BibTeX arXiv:2002.06442

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningManagementregression

Similar Papers 제목 키워드 기반

Causal meets Submodular: Subset Selection with Directed Information

2016-12-01 · NeurIPS 2016 12 · Yuxun Zhou, Costas J. Spanos

We study causal subset selection with Directed Information as the measure of prediction causality. Two typical tasks, causal sensor placement and covariate selection, are correspondingly formulated into cardinality const…

Cardinality Estimation for High Dimensional Similarity Queries with Adaptive Bucket Probing

2026-04-06 · Zhonghan Chen, Qintian Guo, Ruiyuan Zhang, Xiaofang Zhou arxiv

In this work, we address the problem of cardinality estimation for similarity search in high-dimensional spaces. Our goal is to design a framework that is lightweight, easy to construct, and capable of providing accurate…

Asset pre-selection for a cardinality constrained index tracking portfolio with optional enhancement

2025-03-24 · N. Meade, C. A. Valle, J. E. Beasley

An index tracker is a passive investment reproducing the return and risk of a market index, an enhanced index tracker offers a return greater than the index. We consider the selection of a portfolio of given cardinality …

Sketched Sum-Product Networks for Joins

2025-06-16 · Brian Tsan, Abylay Amanbayev, Asoke Datta, Florin Rusu

Sketches have shown high accuracy in multi-way join cardinality estimation, a critical problem in cost-based query optimization. Accurately estimating the cardinality of a join operation -- analogous to its computational…

Investigation of cardinality classification for bacterial colony counting using explainable artificial intelligence

2026-04-21 · Minghua Zheng, Na Helian, Peter C. R. Lane, Yi Sun 외 arxiv

Automatic bacterial colony counting is a highly sought-after technology in modern biological laboratories because it eliminates manual counting effort. Previous work has observed that MicrobiaNet, currently the best-perf…

Density Estimation