对比特币骨干节点属性的统计和集群分析
Dawei Xu1,2, Jiaqi Gao1, Liehuang Zhu1
1School of Cyberspace Security, Beijing Institute of Technology, Beijing, China.
PloS one
|November 8, 2023
概括
比特币的骨干节点,对于网络稳定至关重要,显示出令人惊的集中. 对超过127,000个节点的分析揭示了现实世界的组织联系,证明了比特币的存在.
科学领域:
- 计算机科学 计算机科学
- 网络分析 网络分析
- 加密货币安全性 加密货币安全性
背景情况:
- 比特币在一个分散的点对点 (P2P) 网络上运行.
- 比特币骨干节点显著影响网络稳定性和安全性.
- 了解节点行为对于评估网络完整性至关重要.
研究的目的:
- 分析比特币骨干节点的属性和分布.
- 调查比特币网络内的潜在集中.
- 为了使骨干节点变得匿名,并将它们与现实世界实体联系起来.
主要方法:
- 持续收集比特币节点一年的数据 (2021年7月至2022年6月).
- 过和分析了2,694个已识别的比特币骨干节点.
- 应用三个无监督机器学习算法用于节点集群.
- 集群节点的非匿名化,以识别部署者组织.
主要成果:
- 在比特币节点波动和洋网络节点波动之间发现了直接的相关性.
- 分析显示,比特币的骨干节点表现出集中化的特征.
- 集群识别了不同的骨干节点群体,其中103个节点成功地被匿名化.
- 获得了103个骨干节点部署者的组织信息.
结论:
- 这项研究为比特币骨干节点的集中提供了经验证据.
- 匿名化工作成功地将网络基础设施与现实世界组织联系起来.
- 这些发现强调了监测骨干节点行为对比特币整体健康的重要性.
相关概念视频
Cluster Sampling Method
11.9K
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
11.9K
Outliers and Influential Points
4.1K
An outlier is an observation of data that does not fit the rest of the data. It is sometimes called an extreme value. When you graph an outlier, it will appear not to fit the pattern of the graph. Some outliers are due to mistakes (for example, writing down 50 instead of 500), while others may indicate that something unusual is happening. Outliers are present far from the least squares line in the vertical direction. They have large "errors," where the "error" or residual is the...
4.1K
Statistical Analysis: Overview
6.6K
When we take repeated measurements on the same or replicated samples, we will observe inconsistencies in the magnitude. These inconsistencies are called errors. To categorize and characterize these results and their errors, the researcher can use statistical analysis to determine the quality of the measurements and/or suitability of the methods.
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
6.6K
Evolutionary Relationships through Genome Comparisons
5.8K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
5.8K
Central Tendency: Analysis
157
Measures of central tendency are tools used in biostatistics to identify the average or center of a dataset. They offer a single representative value for understanding and summarizing data distribution.
The mean is one such measure, calculated by totaling all values in a dataset and dividing by the number of values. For instance, the mean blood pressure reading (120, 130, 140, 150) would be 135. However, the mean can be affected by extreme values or outliers.
The median, another measure,...
The mean is one such measure, calculated by totaling all values in a dataset and dividing by the number of values. For instance, the mean blood pressure reading (120, 130, 140, 150) would be 135. However, the mean can be affected by extreme values or outliers.
The median, another measure,...
157
Quantifying and Rejecting Outliers: The Grubbs Test
1.6K
Sometimes, a data set can have a recorded numerical observation that greatly deviates from the rest of the data. Assuming that the data is normally distributed, a statistical method called the Grubbs test can be used to determine whether the observation is truly an outlier. To perform a two-tailed Grubbs test, first, calculate the absolute difference between the outlier and the mean. Then, calculate the ratio between this difference and the standard deviation of the sample. This...
1.6K


