A cache-based data intensive distributed computing architecture for "GRID" applications

Brian Tierney, Johnson, W, Jason Lee · CERN Document Server (European Organization for Nuclear Research) · 2000

Modern scientific computing involves organizing, moving, visualizing, and analyzing massive amounts of data from around the world, as well as employing large-scale computation.The distributed systems that solve largescale problems will always involve aggregating and scheduling many resources.Data must be located and staged, cache and network capacity must be available at the same time as computing capacity, etc.Every aspect of such a system is dynamic: locating and scheduling resources, adapting running application systems to availability and congestion in the middleware and infrastructure, responding to human interaction, etc.The technologies, the middleware services, and the architectures that are used to build useful highspeed, wide area distributed systems, constitute the field of data intensive computing, and are sometimes referred to as the "Data Grid".This paper explores the use of a network data cache in a Data Grid environment. 1.

Read the paper · More papers on PaperTik