Metadata for Managing Grid Resources in Data Mining Applications

Metadata for Managing Grid Resources in Data Mining Applications
复制标题

用于管理数据挖掘应用程序中的网格资源的元数据

DOI:
10.1007/s10723-004-2809-x
复制
发表时间:
2004
影响因子:
5.5
通讯作者:
Paolo Trunfio
Paolo Trunfio
中科院分区:
计算机科学2区
文献类型:
--
作者:
C. Mastroianni;D. Talia;Paolo Trunfio

文献摘要

被引文献

相似文献

网格是一种在动态异构分布式环境中进行资源共享和协调使用的基础设施。有效地使用网格需要定义元数据来管理所涉及的资源的异构性,这些资源包括由不同组织提供的计算机、数据、网络设施和软件工具。当在网格上执行复杂应用程序(如数据密集型模拟和数据挖掘应用程序)时,元数据管理成为一个关键问题。本文讨论了基于网格的数据挖掘应用中异构资源管理的元数据模型。特别地,它讨论了如何在知识网格中表示和管理资源,知识网格是支持网格的分布式数据挖掘的框架。本文阐述了如何使用基于xml的元数据来描述数据挖掘工具、数据源、挖掘模型和执行计划,以及如何使用元数据来设计和执行网格上的分布式知识发现应用程序。
The Grid is an infrastructure for resource sharing and coordinated use of those resources in dynamic heterogeneous distributed environments. The effective use of a Grid requires the definition of metadata for managing the heterogeneity of involved resources that include computers, data, network facilities, and software tools provided by different organizations. Metadata management becomes a key issue when complex applications, such as data-intensive simulations and data mining applications, are executed on a Grid. This paper discusses metadata models for heterogeneous resource management in Grid-based data mining applications. In particular, it discusses how resources are represented and managed in theKnowledge Grid, a framework for Grid-enabled distributed data mining. The paper illustrates how XML-based metadata is used to describe data mining tools, data sources, mining models, and execution plans, and how metadata is used for the design and execution of distributed knowledge discovery applications on Grids.