paper-with-me

홈 › Papers

Optimal Representative Sample Weighting

2020-05-18 · Shane Barratt, Guillermo Angeris, Stephen Boyd

We consider the problem of assigning weights to a set of samples or data records, with the goal of achieving a representative weighting, which happens when certain sample averages of the data are close to prescribed values. We frame the problem of finding representative sample weights as an optimization problem, which in many cases is convex and can be efficiently solved. Our formulation includes as a special case the selection of a fixed number of the samples, with equal weights, i.e., the problem of selecting a smaller representative subset of the samples. While this problem is combinatorial and not convex, heuristic methods based on convex optimization seem to perform very well. We describe rsw, an open-source implementation of the ideas described in this paper, and apply it to a skewed sample of the CDC BRFSS dataset.

📄 PDF Abstract BibTeX arXiv:2005.09065

Code (1)

cvxgrp/rsw 공식 구현

Similar Papers 제목 키워드 기반

Feature Weighting Improves Pool-Based Sequential Active Learning for Regression

2026-04-02 · Dongrui Wu arxiv

Pool-based sequential active learning for regression (ALR) optimally selects a small number of samples sequentially from a large pool of unlabeled samples to label, so that a more accurate regression model can be constru…

Active Learning

Constituency Optimisation Through Hamiltonian Representation Of Mandates (COTHROM): Algorithmic Redistricting of Irish Election Boundaries

2026-06-02 · Ruaidhrí Campion, Matthew Fenlon, Joshua Cooney Mercedal, Casey Farren-Colloty 외 arxiv

Electoral redistricting in Ireland's Proportional Representation Single Transferable Vote (PR-STV) system faces the challenge of selecting an optimally representative set of electoral boundaries from an enormous set of p…

Bias Correction of Learned Generative Models via Likelihood-free Importance Weighting

2019-03-27 · ICLR Workshop DeepGenStruct 2019 · Aditya Grover, Jiaming Song, Ashish Kapoor, Kenneth Tran 외

A learned generative model often gives biased statistics relative to the underlying data distribution. A standard technique to correct this bias is by importance weighting samples from the model by the likelihood ratio u…

Data Augmentation

Aggregation Weighting of Federated Learning via Generalization Bound Estimation

2023-11-10 · Mingwei Xu, Xiaofeng Cao, Ivor W. Tsang, James T. Kwok

Federated Learning (FL) typically aggregates client model parameters using a weighting approach determined by sample proportions. However, this naive weighting method may lead to unfairness and degradation in model perfo…

Federated LearningGeneralization Bounds

Rare Event Detection in Imbalanced Multi-Class Datasets Using an Optimal MIP-Based Ensemble Weighting Approach

2024-12-18 · Georgios Tertytchny, Georgios L. Stavrinides, Maria K. Michael

To address the challenges of imbalanced multi-class datasets typically used for rare event detection in critical cyber-physical systems, we propose an optimal, efficient, and adaptable mixed integer programming (MIP) ens…

Computational EfficiencyEvent Detection