usenix conference policies
DISC: A System for Distributed Data Intensive Scientific Computing
The increasing computation and data requirements of scientific applications have necessitated the use of distributed resources owned by collaborating parties. While existing distributed systems work well for computation that requires limited data movement, they fail in unexpected ways when the computation accesses, creates, and moves large amounts of data over wide-area networks. In this work, we analyzed the problems with existing systems and used the result of this analysis to design our own system. Realizing that it takes a long while for a new system to stabilize, we tried our best to reuse existing components. We added new components only when we could not get by with adding features to existing ones. We used our system to successfully process three terabytes of DPOSS image data in under a week by using idle CPUs in desktops and commodity clusters in the UW-Madison Computer Science Department and Starlight.
author = {George Kola and Tevfik Kosar and Jaime Frey and Miron Livny and Robert Brunner and Michael Remijan},
title = {{DISC}: A System for Distributed Data Intensive Scientific Computing },
booktitle = {First Workshop on Real, Large Distributed Systems (WORLDS 04)},
year = {2004},
address = {San Francisco, CA},
url = {https://www.usenix.org/conference/worlds-04/disc-system-distributed-data-intensive-scientific-computing},
publisher = {USENIX Association},
month = dec
}
connect with us