【24h】

WorkflowHub: Community Framework for Enabling Scientific Workflow Research and Development

机译:WorkflowHub:社区框架,用于启用科学工作流程研究和开发

获取原文

摘要

Scientific workflows are a cornerstone of modern scientific computing. They are used to describe complex computational applications that require efficient and robust management of large volumes of data, which are typically stored/processed on heterogeneous, distributed resources. The workflow research and development community has employed a number of methods for the quantitative evaluation of existing and novel workflow algorithms and systems. In particular, a common approach is to simulate workflow executions. In previous work, we have presented a collection of tools that have been used for aiding research and development activities in the Pegasus project, and that have been adopted by others for conducting workflow research. Despite their popularity, there are several shortcomings that prevent easy adoption, maintenance, and consistency with the evolving structures and computational requirements of production workflows. In this work, we present WorkflowHub, a community framework that provides a collection of tools for analyzing workflow execution traces, producing realistic synthetic workflow traces, and simulating workflow executions. We demonstrate the realism of the generated synthetic traces by comparing simulated executions of these traces with actual workflow executions. We also contrast these results with those obtained when using the previously available collection of tools. We find that our framework not only can be used to generate representative synthetic workflow traces (i.e., with workflow structures and task characteristics distributions that resemble those in traces obtained from real-world workflow executions), but can also generate representative workflow traces at larger scales than that of available workflow traces.
机译:科学工作流是现代科学计算的基石。它们用于描述需要大量数据的复杂计算应用程序,这些数据通常在异构分布式资源上存储/处理。工作流程研究和开发社区采用了许多方法来定量评估现有和新颖的工作流程算法和系统。特别是,一种常见的方法是模拟工作流程执行。在以前的工作中,我们介绍了一系列已被用于解除Pegasus项目的研究和开发活动的工具,并通过其他工具进行工作流程研究。尽管他们受欢迎,但有几种缺点可以防止易于采用,维护和一致性,与生产工作流的不断发展和计算要求。在这项工作中,我们呈现WorkFlowHub,一个社区框架,提供了一个用于分析工作流执行跟踪的工具的集合,产生逼真的合成工作流程跟踪和模拟工作流程执行。通过将这些迹线的模拟执行与实际的工作流程执行进行比较,我们展示了生成的合成迹线的现实主义。我们还将这些结果与使用以前可用的工具收集时的结果对比。我们发现我们的框架不仅可以用于生成代表性的合成工作流程跟踪(即,使用类似于从真实世界工作流程执行中获取的迹线的工作流结构和任务特征分布),但也可以在更大的尺寸下生成代表工作流程图比可用的工作流程迹线。

著录项

相似文献

  • 外文文献
  • 中文文献
  • 专利
获取原文

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号