HR: 16:00h
AN: SF34A-01    [Abstracts]
TI: Distributed Technologies in a Data Pool
AU: * Keiser, K
EM: kkeiser@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899 United States
AU: Conover, H
EM: hconover@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899 United States
AU: Graves, S
EM: sgraves@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899 United States
AU: He, Y
EM: mhe@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899 United States
AU: Regner, K
EM: kregner@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899 United States
AU: Smith, M
EM: msmith@itsc.uah.edu
AF: Information Technology and Systems Center, University of Alabama in Huntsville, Huntsville, AL 35899 United States
AB: A Data Pool is an on-line repository providing interactive and programmatic access to data products through a variety of services. The University of Alabama in Huntsville has developed and deployed such a Data Pool in conjunction with the DISCOVER project, a collaboration with NASA and Remote Sensing Systems. DISCOVER provides long-term ocean and climate data from a variety of passive microwave satellite instruments, including such products as sea-surface temperature and wind, air temperature, atmospheric water vapor, cloud water and rain rate. The Data Pool provides multiple methods to access and visualize these products, including conventional HTTP and FTP access, as well as data services that provide for enhanced usability and interoperability, such as GridFTP, OPeNDAP, OpenGIS-compliant web mapping and coverage services, and custom subsetting and packaging services. This paper will focus on the distributed service technologies used in the Data Pool, which spans heterogeneous machines at multiple locations. For example, in order to provide seamless access to data at multiple sites, the Data Pool provides catalog services for all data products at the various data server locations. Under development is an automated metadata generation tool that crawls the online data repositories regularly to dynamically update the Data Pool catalog with information about newly generated data files. For efficient handling of data orders across distributed repositories, the Data Pool also implements distributed data processing services on the file servers where the data resides. Ontologies are planned to support automated service chaining for custom user requests. The UAH Data Pool is based on a configurable technology framework that integrates distributed data services with a web interface and a set of centralized database services for catalogs and order tracking. While this instantiation of the Data Pool was implemented to meet the needs of the DISCOVER project, the framework was designed for reuse with other small- or large-scale data and information management projects, and is especially suited for distributed collaborations.
UR: http://datapool.nsstc.nasa.gov
DE: 3354 Precipitation (1854)
DE: 3399 General or miscellaneous
DE: 1655 Water cycles (1836)
DE: 1699 General or miscellaneous
DE: 0815 Informal education
SC: Special Focus: Advances in Data Acquisition, Management, Analysis and Display [SF]
MN: 2004 AGU Fall Meeting