paper-with-me

홈 › Papers

Measure of Strength of Evidence for Visually Observed Differences between Subpopulations

2021-01-02 · Xi Yang, Jan Hannig, Katherine A. Hoadley, Iain Carmichael, J. S. Marron

For measuring the strength of visually-observed subpopulation differences, the Population Difference Criterion is proposed to assess the statistical significance of visually observed subpopulation differences. It addresses the following challenges: in high-dimensional contexts, distributional models can be dubious; in high-signal contexts, conventional permutation tests give poor pairwise comparisons. We also make two other contributions: Based on a careful analysis we find that a balanced permutation approach is more powerful in high-signal contexts than conventional permutations. Another contribution is the quantification of uncertainty due to permutation variation via a bootstrap confidence interval. The practical usefulness of these ideas is illustrated in the comparison of subpopulations of modern cancer data.

📄 PDF Abstract BibTeX arXiv:2101.00362

Code (1)

mouseteeth/DiProPerm-test 공식 구현

Similar Papers 제목 키워드 기반

ResponseRank: Data-Efficient Reward Modeling through Preference Strength Learning

2025-12-31 · Timo Kaufmann, Yannick Metz, Daniel Keim, Eyke Hüllermeier arxiv

Binary choices, as often used for reinforcement learning from human feedback (RLHF), convey only the direction of a preference. A person may choose apples over oranges and bananas over grapes, but which preference is str…

Reinforcement Learning

How to Measure Evidence and Its Strength: Bayes Factors or Relative Belief Ratios?

2023-01-21 · Luai Al-Labadi, Ayman Alzaatreh, Michael Evans

Both the Bayes factor and the relative belief ratio satisfy the principle of evidence and so can be seen to be valid measures of statistical evidence. Certainly Bayes factors are regularly employed. The question then is:…

valid

HySim: An Efficient Hybrid Similarity Measure for Patch Matching in Image Inpainting

2024-03-21 · Saad Noufel, Nadir Maaroufi, Mehdi Najib, Mohamed Bakhouya

Inpainting, for filling missing image regions, is a crucial task in various applications, such as medical imaging and remote sensing. Trending data-driven approaches efficiency, for image inpainting, often requires exten…

Image InpaintingPatch MatchingTime SeriesTime Series Forecasting

Humans vs Vision-Language Models: A Unified Measure of Narrative Coherence

2026-03-26 · Nikolai Ilinykh, Hyewon Jang, Shalom Lappin, Asad Sayeed 외 arxiv

We study narrative coherence in visually grounded stories by comparing human-written narratives with those generated by vision-language models (VLMs) on the Visual Writing Prompts corpus. Using a set of metrics that capt…

Model inference for ranking from pairwise comparisons

2025-12-17 · Daniel Sánchez Catalina, George T. Cantwell arxiv

We consider the problem of ranking objects from noisy pairwise comparisons, for example, ranking tennis players from the outcomes of matches. We follow a standard approach to this problem and assume that each object has …