1 / 14

Jane Greenberg janeg@email.unc & the Dryad Team

The DRYAD Repository ~~~~~~ INLS 720 visit to NESCent November 17, 2008. Jane Greenberg janeg@email.unc.edu & the Dryad Team. Motivation for Dryad. Small science repositories (SSR) Knowledge Network for Biocomplexity (KNB), Marine Metadata Initiative (MMI)

beau
Download Presentation

Jane Greenberg janeg@email.unc & the Dryad Team

An Image/Link below is provided (as is) to download presentation Download Policy: Content on the Website is provided to you AS IS for your information and personal use and may not be sold / licensed / shared on other websites without getting consent from its author. Content is provided to you AS IS for your information and personal use only. Download presentation by click this link. While downloading, if for some reason you are not able to download a presentation, the publisher may have deleted the file from their server. During download, if you can't get a presentation, the file might be deleted by the publisher.

E N D

Presentation Transcript


  1. The DRYAD Repository ~~~~~~ INLS 720 visit to NESCent November 17, 2008 Jane Greenberg janeg@email.unc.edu & the Dryad Team

  2. Motivation for Dryad • Small science repositories (SSR) • Knowledge Network for Biocomplexity (KNB), Marine Metadata Initiative (MMI) • Evolutionary biology • Publication process • Supplementary data (Evolution, American Naturalists) “Author,” “deposition date,” not“subject” “species,” ”geo. locator” • Data deposition (Genbank, TreeBase, Morphbank) • NESCent & SILS/Metadata Research Center ecology, paleontology, population genetics, physiology, systematics + genomics

  3. Dryad’s Goals • One-stop deposition and shopping for data objects supporting published research… • 108 data objects, 23 pubs. • American Naturalist, Evolution, • Support the acquisition, preservation, resource discovery, and reuse of heterogeneous digital datasets • Balance a need for low barriers, with higher-level … data synthesis Dryad Team NESCent • Todd Vision, Director of Informatics and Associate Professor, Biology, UNC • Hilmar Lapp, Assistant Director of Informatics • Ryan Scherle, Data Repository Architect UNC/SILS/MRC • Jane Greenberg, Associate Professor, SILS and MRC • Sarah Carrier, Research Assistant • Hollie White, Doctoral Fellow

  4. Dryad Repository model Titel (edit in slide master)

  5. Research and Development

  6. R & D: Accomplishments and Activities • Functional requirements and model • Workshops: Stakeholders (Dec. 06), SSR (May ‘07) • Repository analysis (Dube, et al. JCDL, 2007) - OAIS (Open Archival Information System), DSpace • Metadata architecture Level one application profile

  7. Functional requirements

  8. <DRIADE application profile> Bibliographic Citation Module • dcterms:bibliographicCitation/Citation information • DOI Data Object Module • dc:creator/Name* • dc:title/Data Set # • dc:identifier/Data Set Identifier • PREMIS:fixity/(hidden) • dc:relation/DOI of Published Article* • DDI:<depositr>/Depositor • DDI:<contact>/Contact Information • dc:rights/Rights Statement • dc:description/Description # • dc:subject/Keywords • dc:coverage / Locality Required* • dc:coverage/Date Range Required* • dc:software/Software* • dc:format/File Format • dc:format/File Size • dc:date/(Hidden) Required • dc:date/Date Modified* • Darwin Core: species/ Species, or Scientific* Key * = semi-automatic # = manual Everything else is automatic

  9. R & D: Accomplishments and Activities • Vocabulary analysis • NBII Thesaurus, LCSH, the Getty’s TGN • 600 keywords, Dryad partner journals • Facets: taxon, geographic name, time period, topic • W3C SKOS (Simple Knowledge Organisation Systems) • Instantiation study • Bibliographic relationships for life-cycle management (Coleman, 2002; Smiraglia, 1999, 2000, 2001, 2002, etc.; Tillett; FRBR, DCAM)

  10. Data object relationships B (=data set A annotated) A (=data set) A (=data set in Excel) A (=same data set in SAS) A (=same data set on paper) C (=data set A revised) A1 (=part 1 of a data set) A2 (=part 2 of a data set) A (=data set) A1 (=a subset of A)

  11. Instantiation Scenario: Sherry collects data on the survival and growth of the plant Borrichia frutescens (the bushy seaside tansy)… back at the lab she enters the exact same data into an excel spreadsheet and saves it on her hard drive. Question: What is the relationship between Sherry’s paper data sheet and her excel spreadsheet? Answer: Equivalent | Derivative | Whole-part | Sequential (circle one) Findings (20 participants) • In general, more seasoned scientists better grasp • Sequential data presented the most difficulty (less seasoned sci.) • Unanimous support: “very  extremely important”

  12. R & D: Accomplishments and Activities • Use-case study • Intensive interviews with evolutionary biologists about data sharing • Survey • International survey, launched via evoldir, ~ 400 respondents • PIM Exploratory study (Hollie White)

  13. HIVE model 04/10/2014 Titel (edit in slide master) 13

  14. =

More Related