Museum Data Exchange: Learning How to Share
Total Page:16
File Type:pdf, Size:1020Kb
Museum Data Exchange: Learning How to Share Final Report to The Andrew W. Mellon Foundation Günter Waibel Ralph LeVan Bruce Washburn OCLC Research A publication of OCLC Research Museum Data Exchange: Learning How to Share Museum Data Exchange: Learning How to Share Waibel, et. al., for OCLC Research © 2010 OCLC Online Computer Library Center, Inc. All rights reserved February 2010 OCLC Research Dublin, Ohio 43017 USA www.oclc.org ISBN: 1-55653-424-8 (978-1-55653-424-9) OCLC (WorldCat): 503338562 Please direct correspondence to: Günter Waibel Program Officer [email protected] Suggested citation: Waibel, Günter, Ralph LeVan and Bruce Washburn. 2010. Museum Data Exchange: Learning How to Share. Report produced by OCLC Research. Published online at: www.oclc.org/research/publications/library/2010/2010-02.pdf. www.oclc.org/research/publications/library/2010/2010-02.pdf February 2010 Waibel, et. al., for OCLC Research Page 2 Museum Data Exchange: Learning How to Share Contents Executive Summary .......................................................................................................................... 6 Introduction: Data Sharing in Fits and Starts .................................................................................... 8 Early Reception of CDWA Lite XML ................................................................................................... 10 Grant Overview ............................................................................................................................... 11 Phase 1: Creating Tools for Data Sharing ....................................................................................... 12 COBOAT and OAICATMuseum 1.0: Features and Functionality............................................ 14 Implementing and Refining the Suite of Tools ..................................................................... 16 Phase 2: Creating a Research aggregation ..................................................................................... 17 Legal agreements ............................................................................................................... 17 Harvesting Records ............................................................................................................. 17 Preparing for Data Analysis ................................................................................................. 19 Exposing the Research Aggregation to Participants ............................................................. 21 Phase 3: Analysis of the Research Aggregation .............................................................................. 22 Getting Familiar with the Data ............................................................................................. 23 Conformance to CDWA Lite, Part 1: Cardinality .................................................................... 26 Excursion: the Default COBOAT Mapping ........................................................................... 28 Conformance to CDWA Lite, Part 2: Controlled Vocabularies ............................................... 29 Economically Adding Value: Controlling More Terms ........................................................... 32 Connections: Data Values Used Across the Aggregation .................................................... 34 www.oclc.org/research/publications/library/2010/2010-02.pdf February 2010 Waibel, et. al., for OCLC Research Page 3 Museum Data Exchange: Learning How to Share Enhancement: Automated Creation of Semantic Metadata Using OpenCalais™ ................ 37 A Note about Record Identifiers .......................................................................................... 38 Patricia Harpring’s CCO Analysis ......................................................................................... 40 Third Party Data Analysis .................................................................................................... 41 Compelling Applications for Data Exchange Capacity...................................................................... 41 Conclusion: Policy Challenges Remain ........................................................................................... 44 Appendix A: Project Participants .................................................................................................... 46 Appendix B: Project Related URLs for Tools and Documents ........................................................... 47 Bibliography ................................................................................................................................... 48 Notes: ............................................................................................................................................ 50 Figures Figure 1. Draft system architecture for a CDWA Lite XML data extraction tool .................................. 13 Figure 2. Block diagram of COBOAT, its modules and configuration files ........................................ 15 Figure 3. Excerpt from a report detailing all data values for objectWorkType from a single contributor ...................................................................................................................... 20 Figure 4. Excerpt of a report detailing all of units of the information containing a data value across the research aggregation ...................................................................................... 21 Figure 5. Screenshot of the no-frills search interface to the MDE research aggregation .................. 22 Figure 6. Records contributed by MDE participants ........................................................................ 24 Figure 7. Use of CDWA Lite elements and attributes in the context of all possible units of information ..................................................................................................................... 24 Figure 8. Use of possible CDWA Lite elements and attributes across contributing institutions, take 1 .............................................................................................................................. 25 www.oclc.org/research/publications/library/2010/2010-02.pdf February 2010 Waibel, et. al., for OCLC Research Page 4 Museum Data Exchange: Learning How to Share Figure 9. Use of possible CDWA Lite elements and attributes across contributing institutions, take 2 .............................................................................................................................. 26 Figure 10. Any use of CDWA Lite required / highly recommended elements ................................... 26 Figure 11. Any use of CDWA Lite required / highly recommended elements ................................... 27 Figure 12. Use of CDWA Lite required / highly recommended elements by percentage ................... 28 Figure 13. Match rate of required / highly recommended elements to applicable controlled vocabularies .................................................................................................................. 30 Figure 14. Top 100 objectWorkTypes and their corresponding records for the Metropolitan Museum of Art .............................................................................................................. 32 Figure 15. Top 100 nameCreators and their corresponding records for the Harvard Art Museum......................................................................................................................... 33 Figure 16. Most widely shared values across the aggregation for nameCreator, nationalityCreator, roleCreator and objectWorkType ....................................................... 35 Figure 17. nationalityCreator: relating records, institutions, and unique values ............................. 36 Figure 18. objectWorkType: relating records, institutions, and unique values ................................ 36 Figure 19. Screenshot of a search result from the research aggregation ......................................... 39 Figure 20. objectWorkType spreadsheet (excerpt) for CCO analysis, including evaluation comments .................................................................................................................... 40 Figure 21. Overall scores from CCO evaluation—each bar represents a museum ............................ 41 www.oclc.org/research/publications/library/2010/2010-02.pdf February 2010 Waibel, et. al., for OCLC Research Page 5 Museum Data Exchange: Learning How to Share Executive Summary The Museum Data Exchange, funded by The Andrew W. Mellon Foundation, brought together a group of nine museums and OCLC Research to create tools for data sharing, build a research aggregation and analyze the aggregation. The project established infrastructure for standards-based metadata exchange for the museum community and modeled data sharing behavior among participating institutions. Tools The tools created by the project allow museums to share standards-based data using the Open Archives Initiative Protocol for Metadata Harvesting (OAI-PMH). • COBOAT allows museums to extract Categories for the Description of Works of Art (CDWA) Lite XML out of collections management systems. • OAICatMuseum 1.0 makes the data harvestable via OAI-PMH. COBOAT’s default configuration targets Gallery Systems’ TMS, but can be adjusted to work with other vendor-based or homegrown database systems. Both tools are a free download from: http://www.oclc.org/research/activities/museumdata/. Configuration files adapting COBOAT to different