HR: 16:00h
AN: SF34A-01 [Abstracts]
TI: Distributed Technologies in a Data Pool
AU: * Keiser, K
EM: kkeiser@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899
United States
AU: Conover, H
EM: hconover@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899
United States
AU: Graves, S
EM: sgraves@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899
United States
AU: He, Y
EM: mhe@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899
United States
AU: Regner, K
EM: kregner@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899
United States
AU: Smith, M
EM: msmith@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899
United States
AB:
A Data Pool is an on-line repository providing interactive and programmatic access to data products through a variety of
services. The University of Alabama in Huntsville has developed and deployed such a Data Pool in conjunction with the
DISCOVER project, a collaboration with NASA and Remote Sensing Systems. DISCOVER provides long-term ocean and climate data
from a variety of passive microwave satellite instruments, including such products as sea-surface temperature and wind, air
temperature, atmospheric water vapor, cloud water and rain rate. The Data Pool provides multiple methods to access and
visualize these products, including conventional HTTP and FTP access, as well as data services that provide for enhanced
usability and interoperability, such as GridFTP, OPeNDAP, OpenGIS-compliant web mapping and coverage services, and custom
subsetting and packaging services.
This paper will focus on the distributed service technologies used in the Data Pool, which spans heterogeneous machines at
multiple locations. For example, in order to provide seamless access to data at multiple sites, the Data Pool provides
catalog services for all data products at the various data server locations. Under development is an automated metadata
generation tool that crawls the online data repositories regularly to dynamically update the Data Pool catalog with
information about newly generated data files. For efficient handling of data orders across distributed repositories, the
Data Pool also implements distributed data processing services on the file servers where the data resides. Ontologies are
planned to support automated service chaining for custom user requests. The UAH Data Pool is based on a configurable
technology framework that integrates distributed data services with a web interface and a set of centralized database
services for catalogs and order tracking. While this instantiation of the Data Pool was implemented to meet the needs of the
DISCOVER project, the framework was designed for reuse with other small- or large-scale data and information management
projects, and is especially suited for distributed collaborations.
UR: http://datapool.nsstc.nasa.gov
DE: 3354 Precipitation (1854)
DE: 3399 General or miscellaneous
DE: 1655 Water cycles (1836)
DE: 1699 General or miscellaneous
DE: 0815 Informal education
SC: Special Focus: Advances in Data Acquisition, Management, Analysis and Display [SF]
MN: 2004 AGU Fall Meeting