关于私有云资源分配和使用的数据集
Paola Marques1, Mariana Mendes1, Thiago Emmanuel Pereira1
1Department of Computing and Systems, Federal University of Campina Grande, Campina Grande, Brazil.
Data in brief
|February 18, 2026
概括
研究人员发布了一个新的数据集,详细介绍了私人OpenStack云使用情况. 这个有价值的资源,有超过6400万条记录,有助于云计算研究和私有云分析.
科学领域:
- 计算机科学 计算机科学
- 云计算 云计算 云计算 云计算
- 数据科学数据科学数据科学
背景情况:
- 公共云提供商在商业上占主导地位,但由于特定的治理需求,私有云对学术和研究机构至关重要.
- 关于云资源使用的现有数据集主要集中在公共云上,在了解私有云环境方面存在差距.
- 这种稀缺性限制了对私有云运营模式和优化策略的研究.
研究的目的:
- 来自基于OpenStack的私有云的资源使用的全面数据集.
- 为研究私有云环境及其动态的研究人员提供有价值的资源.
- 促进在非商业云环境中对资源分配,利用和系统性能进行研究.
主要方法:
- 在近十二个月内 (从2024年5月23日到2025年5月16日) 从私有OpenStack云中收集了超过6400万条记录.
- 每五分钟定期查询OpenStack API和监控服务以收集数据.
- 匿名化敏感属性,仅保留系统生成的UUID,以保护隐私.
主要成果:
- 数据集包括基础设施细节,分配配额,用户与项目的关联,虚拟机规格和资源利用指标.
- 时间标记的条目允许对私有云系统动态进行时间分析.
- 该数据集为私有云的运营方面提供了详细的,时间标记的视图.
结论:
- 这一数据集弥合了私有,非商业云环境中公开可用的数据的差距.
- 它是探索云回归的学术机构和公司的宝贵资源.
- 这些发现支持对私有云管理,优化和安全的进一步研究.
相关概念视频
Energy Budgets
10.9K
Organisms must balance energy intake with the energy required for growth, maintenance and reproduction. These trade-offs result in a variety of survivorship and reproductive strategies, including semelparity and iteroparity. Semelparous species, like annual plants, have only one reproductive episode in their lifetimes and consequently have short lifespans. Iteroparous species, by contrast, have many reproductive events during their lifetimes but have relatively few offspring. These two...
10.9K
Cluster Sampling Method
14.9K
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
14.9K
Distributed Loads
1.0K
Distributed loads are a common type of load that engineers and scientists encounter in various practical situations. Distributed loads often refer to a type of load spread over a surface or a structure and can be modeled as continuous force per unit area.
For example, consider a bookshelf filled with books stacked vertically adjacent to each other. The weight of the books is evenly distributed over the length of the shelf. As a result, the pressure at different locations on the surface of the...
For example, consider a bookshelf filled with books stacked vertically adjacent to each other. The weight of the books is evenly distributed over the length of the shelf. As a result, the pressure at different locations on the surface of the...
1.0K
Maximum Size of Aggregate
587
The maximum size of aggregate is defined as the aperture of the sieve retaining 15 percent or more of the particles present in the aggregate sample. The aggregate's maximum size impacts the concrete's water requirement, workability, and strength. Larger aggregates reduce the surface area needing cement paste coverage, which can lower water needs, thereby allowing a decrease in the water-to-cement ratio when the desired workability and richness of the mix are to be maintained, which can...
587
Estimation of the Physical Quantities
8.1K
On many occasions, physicists, other scientists, and engineers need to make estimates of a particular quantity. These are sometimes referred to as guesstimates, order-of-magnitude approximations, back-of-the-envelope calculations, or Fermi calculations. The physicist Enrico Fermi was famous for his ability to estimate various kinds of data with surprising precision. Estimating does not mean guessing a number or a formula at random. Instead, estimation means using prior experience and sound...
8.1K
Probability Histograms
13.3K
A probability histogram is a visual representation of a probability distribution. Similar a typical histogram, the probability histogram consists of contiguous (adjoining) boxes. It has both a horizontal axis and a vertical axis. The horizontal axis is labeled with what the data represents. The vertical axis is labeled with probability. Each rectangular bar in the histogram is 1 unit wide, which suggests that the area under each bar equals the probability, P(x), where x is 1, 2, 3, and so on.
13.3K

