Article

Journal of Grid Computing

, Volume 9, Issue 4, pp 455-478

Adaptive Executions of Multi-Physics Coupled Applications on Batch Grids

  • Sivagama Sundari MurugavelAffiliated withSupercomputer Education and Research Centre, Indian Institute of Science Email author 
  • , Sathish S VadhiyarAffiliated withSupercomputer Education and Research Centre, Indian Institute of Science
  • , Ravi S NanjundiahAffiliated withCentre for Atmospheric & Oceanic Sciences, Indian Institute of Science

Rent the article at a discount

Rent now

* Final gross prices may vary according to local VAT.

Get Access

Abstract

Long running multi-physics coupled parallel applications have gained prominence in recent years. The high computational requirements and long durations of simulations of these applications necessitate the use of multiple systems of a Grid for execution. In this paper, we have built an adaptive middleware framework for execution of long running multi-physics coupled applications across multiple batch systems of a Grid. Our framework, apart from coordinating the executions of the component jobs of an application on different batch systems, also automatically resubmits the jobs multiple times to the batch queues to continue and sustain long running executions. As the set of active batch systems available for execution changes, our framework performs migration and rescheduling of components using a robust rescheduling decision algorithm. We have used our framework for improving the application throughput of a foremost long running multi-component application for climate modeling, the Community Climate System Model (CCSM). Our real multi-site experiments with CCSM indicate that Grid executions can lead to improved application throughput for climate models.

Keywords

Adaptive framework Batch systems Climate models Multi-component applications Rescheduling