paper-with-me

홈 › Papers

Where You Are Is Who You Are: User Identification by Matching Statistics

2015-12-09 · Farid M. Naini, Jayakrishnan Unnikrishnan, Patrick Thiran, Martin Vetterli

Most users of online services have unique behavioral or usage patterns. These behavioral patterns can be exploited to identify and track users by using only the observed patterns in the behavior. We study the task of identifying users from statistics of their behavioral patterns. Specifically, we focus on the setting in which we are given histograms of users' data collected during two different experiments. We assume that, in the first dataset, the users' identities are anonymized or hidden and that, in the second dataset, their identities are known. We study the task of identifying the users by matching the histograms of their data in the first dataset with the histograms from the second dataset. In recent works, the optimal algorithm for this user identification task is introduced. In this paper, we evaluate the effectiveness of this method on three different types of datasets and in multiple scenarios. Using datasets such as call data records, web browsing histories, and GPS trajectories, we show that a large fraction of users can be easily identified given only histograms of their data; hence these histograms can act as users' fingerprints. We also verify that simultaneous identification of users achieves better performance compared to one-by-one user identification. We show that using the optimal method for identification gives higher identification accuracy than heuristics-based approaches in practical scenarios. The accuracy obtained under this optimal method can thus be used to quantify the maximum level of user identification that is possible in such settings. We show that the key factors affecting the accuracy of the optimal identification algorithm are the duration of the data collection, the number of users in the anonymized dataset, and the resolution of the dataset. We analyze the effectiveness of k-anonymization in resisting user identification attacks on these datasets.

📄 PDF Abstract BibTeX arXiv:1512.02896

Code (0)

등록된 구현이 없습니다.

Tasks

User Identification

Similar Papers 제목 키워드 기반

Reverse Attitude Statistics Based Star Map Identification Method

2024-10-31 · Shunmei Dong, Qinglong Wang, Haiqing Wang, Qianqian Wang

The star tracker is generally affected by the atmospheric background light and the aerodynamic environment when working in near space, which results in missing stars or false stars. Moreover, high-speed maneuvering may c…

Bayesian OptimizationPosition

TAL: Two-stream Adaptive Learning for Generalizable Person Re-identification

2021-11-29 · Yichao Yan, Junjie Li, Shengcai Liao, Jie Qin 외

Domain generalizable person re-identification aims to apply a trained model to unseen domains. Prior works either combine the data in all the training domains to capture domain-invariant features, or adopt a mixture of e…

Domain GeneralizationGeneralizable Person Re-identificationMixture-of-ExpertsPerson Re-Identification+1

On the Non-asymptotic and Sharp Lower Tail Bounds of Random Variables

2018-10-21 · Anru R. Zhang, Yuchen Zhou

The non-asymptotic tail bounds of random variables play crucial roles in probability, statistics, and machine learning. Despite much success in developing upper bounds on tail probability in literature, the lower bounds …

Long Range Constraints for Neural Texture Synthesis Using Sliced Wasserstein Loss

2022-11-21 · Liping Yin, Albert Chua

In the past decade, exemplar-based texture synthesis algorithms have seen strong gains in performance by matching statistics of deep convolutional neural networks. However, these algorithms require regularization terms o…

TAGTexture Synthesis

Best Arm Identification for Contaminated Bandits

2018-02-26 · Jason Altschuler, Victor-Emmanuel Brunel, Alan Malek

This paper studies active learning in the context of robust statistics. Specifically, we propose a variant of the Best Arm Identification problem for \emph{contaminated bandits}, where each arm pull has probability $\var…

Active Learning