scieee AI-readable full text Open interactive document viewer

How we organize Dataverse community efforts

Boyd, Ceilyn; Conzett, Philipp; Myers, James

Abstract

Lightening talk given at the Birdaro training program on October 9, 2025.

Full text

How we organize Dataverse community efforts Collaborative strategies for effective community participation The Global Dataverse Community Consortium Supporting Dataverse Repositories Around the World Lightning Talk at the Birdaro Training Program October 9, 2025 Presenting: Ceylin Boyd, Harvard University Philipp Conzett, UiT The Arctic University of Norway Contributed to the slides: Jim Myers, Global Dataverse Community Consortium OUTLINE ❖About the Dataverse Project and GDCC ❖Three types of community efforts ➢Community Meetings ➢Working Groups ➢Community Calls ❖Summary ❖Acknowledgements ABOUT THE PROJECT ABOUT THE DATAVERSE PROJECT Open-Source Repository Software Dataverse is open-source software that enables sharing, citation, and preservation of research data. Global and Multidisciplinary Support It supports data sharing across diverse disciplines and institutions worldwide. Enhancing Accessibility and Reproducibility The project aims to improve data accessibility and reproducibility in scholarly research. ABOUT THE GLOBAL DATAVERSE COMMUNITY CONSORTIUM (GDCC) Community Support and Coordination GDCC was established in 2018 to support the Dataverse community by coordinating and facilitating collaboration, community efforts, and resources worldwide. Governance and Strategic Direction GDCC provides governance and strategic guidance to keep the Dataverse Project aligned with user needs from a diverse community. Collaborative Venue GDCC provides a collaborative venue for institutions to leverage economies of scale in support of Dataverse repositories around the world, e.g., through the GDCC DataCite consortium membership that provides DOIs for Dataverse installations. THREE TYPES OF COMMUNITY EFFORTS DATAVERSE COMMUNITY MEETINGS Annual event since 2015, evolving from Harvard-hosted to globally distributed (Portugal, Mexico, North Carolina). Main venue for facilitating discussing past, ongoing, and future community efforts, and for fostering networking and collaboration. Organized by a rotating team: local hosts + other (more experienced) community members to support newcomers. Planning starts with a community survey to identify relevant topics (e.g., Metadata, Large Data Support, AI, Governance). Open call for contributions in diverse formats: talks, workshops, roundtables, demos, posters, WG sessions, etc. Challenges: engaging newcomers, travel limitations, and inclusivity in hybrid formats. WORKING GROUPS WGs form organically when community members identify a shared need or goal. Led by volunteer coordinators, ensuring alignment with broader community priorities. Open and inclusive membership, welcoming diverse contributors. Common characteristics: Regular meetings, progress documentation, and updates via mailing lists and community events. Challenges include limited capacity, shifting leadership, and extended timelines. Current WGs: ● Containerization ● Documentation ● Large Data Support ● pyDataverse