paper-with-me

Papers

A nonparametric framework for inferring orders of categorical data from category-real ordered pairs

2019-11-15 · Chainarong Amornbunchornvej, Navaporn Surasvadi, Anon Plangprasopchok, Suttipong Thajchayapong

Given a dataset of careers and incomes, how large a difference of income between any pair of careers would be? Given a dataset of travel time records, how long do we need to spend more when choosing a public transportation mode $A$ instead of $B$ to travel? In this paper, we propose a framework that is able to infer orders of categories as well as magnitudes of difference of real numbers between each pair of categories using Estimation statistics framework. Not only reporting whether an order of categories exists, but our framework also reports the magnitude of difference of each consecutive pairs of categories in the order. In large dataset, our framework is scalable well compared with the existing framework. The proposed framework has been applied to two real-world case studies: 1) ordering careers by incomes based on information of 350,000 households living in Khon Kaen province, Thailand, and 2) ordering sectors by closing prices based on 1060 companies' closing prices of NASDAQ stock markets between years 2000 and 2016. The results of careers ordering show income inequality among different careers. The stock market results illustrate dynamics of sector domination that can change over time. Our approach is able to be applied in any research area that has category-real ordered pairs. Our proposed "Dominant-Distribution Network" provides a novel approach to gain new insight of analyzing category orders. The software of this framework is available for researchers or practitioners within R package: EDOIF.

📄 PDF Abstract BibTeX arXiv:1911.06723

Code (1)

DarkEyes/EDOIF 공식 구현

Methods 이 논문이 사용한 방법론

Estimation Statistics Estimation statistics is a data analysis framework that uses a combination of effect sizes, confidence intervals, precision planning, and meta-analysis to plan experiments,…

Similar Papers 제목 키워드 기반

Efficient Inference of Gaussian Process Modulated Renewal Processes with Application to Medical Event Data

2014-02-19 · Thomas A. Lasko

The episodic, irregular and asynchronous nature of medical data render them difficult substrates for standard machine learning algorithms. We would like to abstract away this difficulty for the class of time-stamped cate…

Gaussian Processes

lgpr: An interpretable nonparametric method for inferring covariate effects from longitudinal data

2019-12-07 · Juho Timonen, Henrik Mannerström, Aki Vehtari, Harri Lähdesmäki

Longitudinal study designs are indispensable for studying disease progression. Inferring covariate effects from longitudinal data, however, requires interpretable methods that can model complicated covariance structures …

Gaussian Processes

Implicational Universals in Stochastic Constraint-Based Phonology

2018-10-01 · EMNLP 2018 10 · Giorgio Magri

This paper focuses on the most basic implicational universals in phonological theory, called T-orders after Anttila and Andrus (2006). It shows that the T-orders predicted by stochastic (and partial order) Optimality The…

Inferring Outcome Means of Exponential Family Distributions Estimated by Deep Neural Networks

2025-04-12 · Xuran Meng, Yi Li

While deep neural networks (DNNs) are widely used for prediction, inference on DNN-estimated subject-specific means for categorical or exponential family outcomes remains underexplored. We address this by proposing a DNN…

regression

Order is All You Need for Categorical Data Clustering

2024-11-19 · Yiqun Zhang, Mingjie Zhao, Hong Jia, Yang Lu 외

Categorical data composed of qualitative valued attributes are ubiquitous in machine learning tasks. Due to the lack of well-defined metric space, categorical data distributions are difficult to be intuitively understood…

AllAttributeCategorical data clusteringClustering