paper-with-me

홈 › Papers

Towards Statistically Significant Taxonomy Aware Co-location Pattern Detection

2024-06-29 · Subhankar Ghosh, Arun Sharma, Jayant Gupta, Shashi Shekhar

Given a collection of Boolean spatial feature types, their instances, a neighborhood relation (e.g., proximity), and a hierarchical taxonomy of the feature types, the goal is to find the subsets of feature types or their parents whose spatial interaction is statistically significant. This problem is for taxonomy-reliant applications such as ecology (e.g., finding new symbiotic relationships across the food chain), spatial pathology (e.g., immunotherapy for cancer), retail, etc. The problem is computationally challenging due to the exponential number of candidate co-location patterns generated by the taxonomy. Most approaches for co-location pattern detection overlook the hierarchical relationships among spatial features, and the statistical significance of the detected patterns is not always considered, leading to potential false discoveries. This paper introduces two methods for incorporating taxonomies and assessing the statistical significance of co-location patterns. The baseline approach iteratively checks the significance of co-locations between leaf nodes or their ancestors in the taxonomy. Using the Benjamini-Hochberg procedure, an advanced approach is proposed to control the false discovery rate. This approach effectively reduces the risk of false discoveries while maintaining the power to detect true co-location patterns. Experimental evaluation and case study results show the effectiveness of the approach.

📄 PDF Abstract BibTeX arXiv:2407.00317

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

+ ( 1 ) ⟷ 888 ⟷ ( 829 ) ⟷ 0881||How do I resolve a dispute on Expedia? How do I resolve a dispute on Expedia contact their support at + ( 1 ) ⟷ 888 ⟷ ( 829 ) ⟷ 0881 or + ( 1 ) ⟷ 805 ⟷ ( 330 ) ⟷ 4056. Provide booking details and explain the issue…

Similar Papers 제목 키워드 기반

Reducing False Discoveries in Statistically-Significant Regional-Colocation Mining: A Summary of Results

2024-07-01 · Subhankar Ghosh, Jayant Gupta, Arun Sharma, Shuai An 외

Given a set \emph{S} of spatial feature types, its feature instances, a study area, and a neighbor relationship, the goal is to find pairs $<$a region ($r_{g}$), a subset \emph{C} of \emph{S}$>$ such that \emph{C} is a s…

Sociology

Data Quality Antipatterns for Software Analytics

2024-08-22 · Aaditya Bhatia, Dayi Lin, Gopi Krishnan Rajbahadur, Bram Adams 외

Background: Data quality is vital in software analytics, particularly for machine learning (ML) applications like software defect prediction (SDP). Despite the widespread use of ML in software engineering, the effect of …

Creating a Fine Grained Entity Type Taxonomy Using LLMs

2024-02-19 · Michael Gunn, Dohyun Park, Nidhish Kamath

In this study, we investigate the potential of GPT-4 and its advanced iteration, GPT-4 Turbo, in autonomously developing a detailed entity type taxonomy. Our objective is to construct a comprehensive taxonomy, starting f…

Event Argument ExtractionRelation Extraction

Rare-Aware Autoencoding: Reconstructing Spatially Imbalanced Data

2026-04-02 · Alejandro Castañeda Garcia, Jan van Gemert, Daan Brinks, Nergis Tömen arxiv

Autoencoders can be challenged by spatially non-uniform sampling of image content. This is common in medical imaging, biology, and physics, where informative patterns occur rarely at specific image coordinates, as backgr…

Image Reconstruction

Converting Expert Deliberation into Financial Signals Through A Context-Aware NLP Pipeline

2026-08-19 · Vivek Batra, Kristin Chen, Sanjiv Das, Samuel Judge 외 arxiv

We introduce the CDSP (context-conditional deliberation signal pipeline), converting an investment committee's meeting transcripts into structured predictive features. CDSP segments the meeting transcripts into topical c…

Feature Engineering