中文

大规模设施fabric监控与网格环境中元数据服务的集成——GridMonitor

分布式、并行与集群计算 2007-05-23 v1 性能

摘要

网格计算由于协调使用大量异构、地理分布的资源进行高性能计算。有效监控这些计算资源至关重要,以便高效地在网格中使用。大量异构计算实体在网格中存在,使监控任务具有挑战性。在本工作中,我们描述了一个网格监控系统,称为 GridMonitor,用于捕获并提供来自大型计算设施的最重要信息。该网格监控系统由四个层次组成:本地监控、归档、发布和利用。该架构已应用于大规模 Linux 农场和网络基础设施上,可供包括调度服务和资源 broker 在内的许多高级网格服务使用。

关键词

引用

@article{arxiv.cs/0306073,
  title  = {GridMonitor: Integration of Large Scale Facility Fabric Monitoring with Meta Data Service in Grid Environment},
  author = {Rich Baker and Dantong Yu and Jason Smith and Anthony Chan and Kaushik De and Patrick McGuigan},
  journal= {arXiv preprint arXiv:cs/0306073},
  year   = {2007}
}

备注

Talk from the 2003 Computing in High Energy and Nuclear Physics (CHEP03), La Jolla, Ca, USA, March 2003, 8 pages, LaTeX, 1 eps figure, 4 ps figures, 1 style file, Monitoring, PSN MOET005