A Biased Estimator for MinMax Sampling and Distributed Aggregation

Wolfrath, Joel; Chandra, Abhishek

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2404

Computer Science > Machine Learning

Title: A Biased Estimator for MinMax Sampling and Distributed Aggregation

Authors: Joel Wolfrath, Abhishek Chandra

(Submitted on 26 Apr 2024)

Abstract: MinMax sampling is a technique for downsampling a real-valued vector which minimizes the maximum variance over all vector components. This approach is useful for reducing the amount of data that must be sent over a constrained network link (e.g. in the wide-area). MinMax can provide unbiased estimates of the vector elements, along with unbiased estimates of aggregates when vectors are combined from multiple locations. In this work, we propose a biased MinMax estimation scheme, B-MinMax, which trades an increase in estimator bias for a reduction in variance. We prove that when no aggregation is performed, B-MinMax obtains a strictly lower MSE compared to the unbiased MinMax estimator. When aggregation is required, B-MinMax is preferable when sample sizes are small or the number of aggregated vectors is limited. Our experiments show that this approach can substantially reduce the MSE for MinMax sampling in many practical settings.

Subjects:	Machine Learning (cs.LG); Distributed, Parallel, and Cluster Computing (cs.DC); Applications (stat.AP)
Cite as:	arXiv:2404.17690 [cs.LG]
	(or arXiv:2404.17690v1 [cs.LG] for this version)

Submission history

From: Joel Wolfrath [view email]
[v1] Fri, 26 Apr 2024 20:39:08 GMT (205kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2404.17690

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: A Biased Estimator for MinMax Sampling and Distributed Aggregation

Submission history