arrow
Return

Supporting scalable and distributed data subsetting and aggregation in large-scale seismic data analysis

delete2006-08-01
delete0
PRE
AI
X
Xilin Zhang *
B
Benjamin Rutt
C
Catalyuerek, U.
T
Tahsin Kurç
M
Mrinal K. Sen
DOI:10.1177/1094342006067471delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The ability to query and process very large, terabyte-scale datasets has become a key step in many scientific and engineering applications. In this paper, we describe the application of two middleware frameworks in an integrated fashion to provide a scalable and efficient system for execution of seismic data analysis on large datasets in a distributed environment. We investigate different strategies for efficient querying of large datasets and parallel implementations of a seismic image reconstruction algorithm. Our results on a state-of-the-art mass storage system coupled with a high-end compute cluster show that our implementation is scalable and can achieve about 2.9 Gigabytes per second data processing rate - about 70% of the maximum 4.2GB/s application-level raw I/O bandwidth of the storage platform.
Keywords:
seismic data analysis
data-driven applications
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

International Journal of High Performance Computing Applications cover
International Journal of High Performance Computing Applications
IF:
2.5
Papers:
1.1K
Citations:
1.3K

Organization

No organization information available