Related Experiment Video
Updated: Aug 7, 2025

High Speed Droplet-based Delivery System for Passive Pumping in Microfluidic Devices
Published on: September 2, 2009
Data Flush
Xiaotong Shen1, Xuan Bi2, Rex Shen3
1School of Statistics, University of Minnesota, Minneapolis, MN 55455.
Abstract:
Data perturbation is a technique for generating synthetic data by adding "noise" to raw data, which has an array of applications in science and engineering, primarily in data security and privacy. One challenge for data perturbation is that it usually produces synthetic data resulting in information loss at the expense of privacy protection. The information loss, in turn, renders the accuracy loss for any statistical or machine learning method based on the synthetic data, weakening downstream analysis and deteriorating in machine learning. In this article, we introduce and advocate a fundamental principle of data perturbation, which requires the preservation of the distribution of raw data. To achieve this, we propose a new scheme, named data flush, which ascertains the validity of the downstream analysis and maintains the predictive accuracy of a learning task. It perturbs data nonlinearly while accommodating the requirement of strict privacy protection, for instance, differential privacy. We highlight multiple facets of data flush through examples.
Related Concept Videos
Discharge Summary Forms
Here's a detailed look at the key components and guidelines for preparing a discharge summary:
Displacement Current
Responses to Drought and Flooding
Cavity Drainage and Flashings in Masonry walls
Weep holes, strategically placed at the base of the cavity, are critical for draining accumulated water. These openings are created by leaving head...
Accelerating Fluids
The motion of the liquid within this infinitesimal cylinder is considered to obtain the pressure difference. Three vertical forces act on this liquid:
Heart Failure Drugs: Diuretics

