プライベートクラウドのリソース割り当てと利用に関するデータセット
Paola Marques1, Mariana Mendes1, Thiago Emmanuel Pereira1
1Department of Computing and Systems, Federal University of Campina Grande, Campina Grande, Brazil.
Data in brief
|February 18, 2026
まとめ
研究者は,プライベートなOpenStackクラウドの使用を詳細に記述する新しいデータセットをリリースしました. この貴重なリソースは,6400万件以上のレコードを備えており,クラウドコンピューティングの研究とプライベートクラウドの分析を支援しています.
科学分野:
- コンピュータサイエンス コンピュータサイエンス
- クラウドコンピューティング クラウドコンピューティング
- データサイエンス データサイエンス
背景:
- 公共クラウドプロバイダは商業的に優位だが,プライベートクラウドは特定のガバナンスニーズにより,学術機関や研究機関にとって不可欠である.
- クラウドリソース利用に関する既存のデータセットは,主にパブリッククラウドに焦点を当て,プライベートクラウド環境の理解にギャップを残しています.
- この希少性は,プライベートクラウドの運用パターンと最適化戦略の研究を制限しています.
研究 の 目的:
- プライベートなOpenStackベースのクラウドからのリソース使用に関する包括的なデータセットを提示します.
- プライベートクラウド環境とその動態を研究する研究者に貴重なリソースを提供すること.
- 非商用クラウド環境におけるリソースの割り当て,利用,およびシステムの性能に関する研究を促進する.
主な方法:
- 約12ヶ月間 (2024年5月23日から2025年5月16日まで) にわたって,プライベートなOpenStackクラウドから6400万件以上のレコードを収集しました.
- 定期的にOpenStack APIと監視サービスを5分ごとにクエリしてデータを収集しました.
- 機密性の高い属性を匿名化し,プライバシーのためにシステム生成のUUIDのみを保持します.
主要な成果:
- データセットには,インフラストラクチャの詳細,配分配分,ユーザーとプロジェクトの関連,仮想マシンの仕様,およびリソース利用メトリックが含まれています.
- タイムスタンプ付きエントリは,プライベートクラウドシステムのダイナミクスの時間分析を可能にします.
- このデータセットは,プライベートクラウドの運用面について,詳細でタイムスタンプ付きの見方を提供しています.
結論:
- このデータセットは,民間非商用クラウド環境からの公開可能なデータのギャップを埋めています.
- これは,クラウドの帰還を模索している学術機関や企業にとって貴重なリソースとなります.
- この発見は,プライベートクラウドの管理,最適化,およびセキュリティに関するさらなる研究を支援します.
関連する概念動画
Energy Budgets
10.9K
Organisms must balance energy intake with the energy required for growth, maintenance and reproduction. These trade-offs result in a variety of survivorship and reproductive strategies, including semelparity and iteroparity. Semelparous species, like annual plants, have only one reproductive episode in their lifetimes and consequently have short lifespans. Iteroparous species, by contrast, have many reproductive events during their lifetimes but have relatively few offspring. These two...
10.9K
Cluster Sampling Method
14.9K
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
14.9K
Distributed Loads
1.0K
Distributed loads are a common type of load that engineers and scientists encounter in various practical situations. Distributed loads often refer to a type of load spread over a surface or a structure and can be modeled as continuous force per unit area.
For example, consider a bookshelf filled with books stacked vertically adjacent to each other. The weight of the books is evenly distributed over the length of the shelf. As a result, the pressure at different locations on the surface of the...
For example, consider a bookshelf filled with books stacked vertically adjacent to each other. The weight of the books is evenly distributed over the length of the shelf. As a result, the pressure at different locations on the surface of the...
1.0K
Maximum Size of Aggregate
587
The maximum size of aggregate is defined as the aperture of the sieve retaining 15 percent or more of the particles present in the aggregate sample. The aggregate's maximum size impacts the concrete's water requirement, workability, and strength. Larger aggregates reduce the surface area needing cement paste coverage, which can lower water needs, thereby allowing a decrease in the water-to-cement ratio when the desired workability and richness of the mix are to be maintained, which can...
587
Estimation of the Physical Quantities
8.1K
On many occasions, physicists, other scientists, and engineers need to make estimates of a particular quantity. These are sometimes referred to as guesstimates, order-of-magnitude approximations, back-of-the-envelope calculations, or Fermi calculations. The physicist Enrico Fermi was famous for his ability to estimate various kinds of data with surprising precision. Estimating does not mean guessing a number or a formula at random. Instead, estimation means using prior experience and sound...
8.1K
Probability Histograms
13.3K
A probability histogram is a visual representation of a probability distribution. Similar a typical histogram, the probability histogram consists of contiguous (adjoining) boxes. It has both a horizontal axis and a vertical axis. The horizontal axis is labeled with what the data represents. The vertical axis is labeled with probability. Each rectangular bar in the histogram is 1 unit wide, which suggests that the area under each bar equals the probability, P(x), where x is 1, 2, 3, and so on.
13.3K

