针对1:M微数据的背景知识攻击改进了天使化技术
Rabeeha Fazal1, Razaullah Khan2, Adeel Anjum3
1Department of Computer Science, COMSATS Institute of Information Technology, Islamabad, Pakistan.
PeerJ. Computer science
|June 22, 2023
概括
分享电子健康记录 (EHR) 带来隐私风险. 一个新的 (θ*,k) 实用算法增强了数据集的隐私和数据实用性,每个人 (1:M) 有多个记录.
科学领域:
- 医疗信息学 医疗信息学
- 数据 隐私 数据 隐私 数据
- 计算机科学 计算机科学
背景情况:
- 跨组织共享电子健康记录 (EHR) 为医疗治疗和研究提供了好处,但也引发了重大隐私问题.
- 传统的隐私模型通常假定每个人只有一条记录 (1:1数据集),这对于每个人多条记录 (1:M数据集) 的现实场景是不够的.
- 现有的隐私模型,如 θ-Sensitive k-Anonymity, (p,l) -angelization,和 (k,l) -diversity,表明1:M数据集的高功用损失和不充分的隐私.
研究的目的:
- 解决当前隐私模型在处理1M数据集方面的局限性.
- 提出一种新的算法,平衡1:M数据集的增强隐私与数据实用性.
- 评估拟议的算法的有效性与现有方法相比.
主要方法:
- 该研究分析了现有的隐私模型 (θ-敏感的k-匿名性, (p, l) 角化, (k, l) 多样性) 对1:M数据集的不适用性和高效益损失.
- 提出了一个新的算法, (θ*,k) -utility,以改善匿名1:M数据集的隐私和实用性保护.
- 使用现实世界数据集进行实验,以将拟议的方法与现有方法进行比较.
主要成果:
- 与现有模型相比,提出的 (θ*,k) 实用性算法在保护 1:M 数据集的隐私和数据实用性方面表现出卓越的性能.
- 当前的模型显示,当应用到1:M数据集时,显著的实用性损失和隐私漏洞.
- 实验结果验证了新算法在现实数据上的有效性.
结论:
- (θ*,k) 实用算法为1M EHR数据集的隐私保护数据共享提供了一个强大的解决方案.
- 这些发现凸显了传统隐私模型对于复杂,多记录数据集的不足.
- 这项研究有助于开发医疗保健中更安全,更实用的数据共享实践.
相关概念视频
Woodward–Hoffmann Selection Rules and Microscopic Reversibility
3.2K
Electrocyclic reactions, cycloadditions, and sigmatropic rearrangements are concerted pericyclic reactions that proceed via a cyclic transition state. These reactions are stereospecific and regioselective. The stereochemistry of the products depends on the symmetry characteristics of the interacting orbitals and the reaction conditions. Accordingly, pericyclic reactions are classified as either symmetry-allowed or symmetry-forbidden. Woodward and Hoffmann presented the selection criteria for...
3.2K
Masking and Demasking Agents
2.5K
EDTA titrations may necessitate masking and demasking agents to temporarily protect a particular metal ion in a mixture from the EDTA reaction. These agents facilitate the sequential analysis of the metal ions by forming stable complexes with some—but not all—metal ions during certain steps.
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on...
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on...
2.5K
Censoring Survival Data
151
Survival analysis is a statistical method used to analyze time-to-event data, often employed in fields such as medicine, engineering, and social sciences. One of the key challenges in survival analysis is dealing with incomplete data, a phenomenon known as "censoring." Censoring occurs when the event of interest (such as death, relapse, or system failure) has not occurred for some individuals by the end of the study period or is otherwise unobservable, and it might have many different...
151
Strategies for Assessing and Addressing Confounding
122
Confounding is a critical issue in epidemiological studies, often leading to misleading conclusions about associations between exposures and outcomes. It occurs when the relationship between the exposure and the outcome is mixed with the effects of other factors that influence the outcome. Given that, addressing confounding is of high importance for drawing accurate inferences in research.
Confounding can be addressed at both the design phase of a study and through analytical methods after data...
Confounding can be addressed at both the design phase of a study and through analytical methods after data...
122
Blinding
2.5K
Blinding is a commonly used method of not telling participants which treatment a subject is receiving. Blinding is a critical part of a randomized control trial or RCT. It reduces the bias that affects the results. In an RCT, blinding is used in the form of a placebo. A placebo effect occurs when untreated subjects falsely believe they have received the treatment and report improved symptoms. A placebo or a dummy treatment is administered to subjects to negate the bias caused by such an effect.
2.5K
Difference from Background: Limit of Detection
6.7K
The limit of detection (LOD) is the smallest amount of analyte that can be distinguished from the background noise. The LOD value corresponds to the concentration at which the analyte signal is three times larger than the standard deviation of the blank signal. Below this value, the analyte signal cannot be differentiated from the background noise. It is calculated by dividing the calibration slope by 3 times the standard deviation of the blank signals.
The LOD indicates the presence or absence...
The LOD indicates the presence or absence...
6.7K


