GCE-LTER ANNUAL REPORT of ACTIVITIES, Year Five (2004-2005) Award NSF OCE 99-82133
Total Page:16
File Type:pdf, Size:1020Kb
GCE-LTER ANNUAL REPORT OF ACTIVITIES, Year Five (2004-2005) Award NSF OCE 99-82133 Project Management Project Administration James Hollibaugh, Dept. of Marine Sciences, UGA; Steven C. Pennings, Univ. of Houston Administrative responsibilities have included organizing the GCE annual meeting, participating in LTER network coordinating committee meetings, and dealing with accounting, reporting and supplementary and continuation funding paperwork. Hollibaugh and Pennings split Coordinating Committee responsibilities with Pennings attending the spring meeting in Santa Barbara, CA and Hollibaugh attending the summer meeting in Fairbanks, AK. Hollibaugh also participated in a workshop (Coco Beach, FL) organized by LNO to develop a series of Grand Challenge questions to guide network synthesis in the next decade of LTER research. Science oversight responsibilities involved coordinating the efforts of scientists in different aspects of the study, and negotiating agreements with different agencies to supplement monitoring efforts. Our fourth year of Schoolyard LTER support provided an opportunity for teachers to participate in field research at the GCE-LTER site during the summer of 2004. Immediate supervision of the Schoolyard LTER is currently being provided by J.T. Hollibaugh, with the actual teacher training effort lead by Ms Trisha Hembree. Ms Hembree has been successful in augmenting the Schoolyard with funds from the Eisenhower Foundation. Web Site, Database, and Information Management Activities Wade Sheldon, Dept. of Marine Sciences, UGA Overview Information management activities this year were largely focused on strengthening existing support for LTER Network Information System initiatives, such as EML metadata, developing standards and software support for automatic synchronization of GCE metadata with the LTER Network Office and other metadata clearinghouses (NBII, KNB, SEEK Ecogrid), and supporting GCE investigators in data integration, analysis and synthesis projects. We also enhanced our IT infrastructure by acquiring a new database server and adding a network-attached RAID server to provide secure central storage and backup of GCE investigator data and to simplify data exchange and collaboration within subproject working groups. Near-real-time Climate and Hydrographic Data The fully automated climate data harvesting system, developed in November 2002 with supplemental NSF funding for ClimDB/HydroDB participants, continues to provide GCE participants with near-real-time data and plots from two climate stations (Marsh Landing on Sapelo Island and Hudson Creek/Meridian Landing) and one USGS streamflow gauge (Altamaha River at Doctortown). Real-time data are finalized monthly, re-sampled to daily values and submitted to the LTER climate and hydrological databases (ClimDB and HydroDB), along with data from the manually operated NWS climate station on Sapelo Island. Salinity, temperature and pressure data from moored hydrographic data loggers continue to be downloaded and processed on a bimonthly to quarterly basis. Provisional data and plots are provided to GCE investigators on the project web site within two days of acquisition, and finalized, quality-controlled data are added to the GCE Data Catalog at the end of each year to provide public access. The GCE project also continues to host the USGS Data Harvesting Service for HydroDB (see http://gce-lter.marsci.uga.edu/lter/research/tools/usgs_harvester.htm). Data from 52 USGS streamflow gauging stations are automatically harvested on a weekly basis for 10 LTER sites (AND, BES, FCE, GCE, KBS, KNZ, LUQ, NTL, PIE, SBC) and one USFS site. Recent provisional and finalized data are automatically acquired, standardized, quality-checked, formatted, and uploaded to HydroDB to provide the LTER community with the best available stream flow and precipitation data for synthetic research at no cost to individual site research programs. Software Development We have continued to improve on the GCE Data Toolbox for MATLAB software and offer a compiled version of this toolbox for public download on our web site (see http://gce-lter.marsci.uga.edu/lter/research/tools/data_toolbox.htm). Major new features were added this year to support data discovery and integration, including a comprehensive data indexing and search engine for performing metadata-based searches to identify data sets that meet a wide range of user-specified thematic, temporal, and geospatial criteria. All public data sets in the GCE Data Catalog and Data Portal web sites can be searched in addition to user-created or customized data sets stored in local directories; consequently, this tool serves both as an end-user data management tool and a remote data access client in addition to a search engine. Significant improvements were also made to data integration functions, allowing multiple related data sets stored in any number of directories to be merged into composite data sets in one step, with user-specified handling of metadata content, QA/QC flags and overlapping date ranges. Over 500 visitors downloaded the toolbox package this year (see download statistics below). Support for EML Metadata Support for EML 2.0, the new xml-based metadata standard adopted by the LTER Network in 2002, was significantly enhanced this year by adding EML support to the GCE taxonomic database. Standard and user-specified species lists can now be generated as stand-alone EML documents (in addition to HTML, text tables and CSV 2 files) to support exchange of taxonomic information across LTER. Two different EML implementations were developed (i.e. list form and tree form) and submitted for consideration as a network-wide standard. The GCE Information Manager (W. Sheldon) also headed a working group at the LTER Network Office (LNO) in May 2004 to establish best practices for LTER EML implementation. The best practices document and example schemas from this working group are currently under review by the LTER IM Committee, but all recommendations have already been implemented at GCE in the EML dynamically-generated for each data set in our public data catalog. We also provided logistical support to other LTER sites this year in implementing EML, including: Coweeta (CWT), Virginia Coast Reserve (VCR), Niwot Ridge (NWT) and Palmer Station (PAL). In addition, W. Sheldon collaborated with LNO and NCEAS on development and testing of a specification for automatically harvesting EML documents from LTER sites for inclusion in the KNB Metacat repository (and therefore the distributed Metacat and Ecogrid networks). This ‘Metacat Harvester’ is now operational and GCE EML documents are automatically added or updated in the LNO Metacat server on a weekly basis, and then synchronized to other Metacat servers across the country. The USGS National Biological Information Infrastructure project has also adopted this specification to harvest EML for inclusion in the NBII metadata clearinghouse, and GCE is the first LTER site to become established as an NBII Clearinghouse node. All GCE metadata can now be searched using the NCEAS Morpho application, the KNB Metacat web interface (http://knb.ecoinformatics.org), the NBII Mercury search engine (http://mercury.ornl.gov/nbii/), and the Ecogrid network being developed by the SEEK project. Corresponding data tables can also be automatically retrieved by these external systems using connection information in the metadata, with data access logged by the GCE database. In addition, the comprehensive EML implementation and support for automated data streaming developed at GCE has enabled NCEAS and the KNB and SEEK projects to prototype, test and demonstrate EML-based data analysis and workflow tools, such as Kepler (http://seek.ecoinformatics.org) using realistic ecological data sets, significantly aiding these projects. Data Catalog Additions Data submissions from both monitoring and directed study programs were similar to those in year 4, with 101 data sets added to the catalog or in the final stages of documentation for inclusion. GCE investigators have contributed a total of 280 documented data sets to date, 194 of which are currently accessible to the LTER and broader scientific community through our web-accessible public data catalog. This year we also became the first LTER site to develop support for automated metadata synchronization with the LNO/KNB Metacat repository, the NBII Metadata Clearinghouse, and the nascent SEEK Ecogrid Network (see EML Metadata section). These metadata-sharing collaborations dramatically increased the visibility and accessibility of GCE data and contributed to a 10-fold increase in data set downloads in Year 5 relative to Year 4. Detailed public data download statistics are provided below in the Web Site and Data Access Statistics section. Web Site Additions 3 The project calendar web application developed last year to promote communication within the project was significantly enhanced this year to improve display of upcoming events. All GCE participants can add events to the calendar, and event details are stored in a relational database to support dynamic calendar generation and searching. Support for title text searching was also added to the interactive filter panel on the public data catalog to improve data discovery, and an enhanced data search engine is currently under development in parallel with efforts at LNO to develop a standard data query interface for the LTER network. As mentioned above, additional options to display species lists in various EML metadata formats were added