Showing posts with label data repositories. Show all posts
Showing posts with label data repositories. Show all posts

Sunday, 27 May 2018

Book - Exploring Research Data Management

By Andrew Cox and Eddy Verbaan 

See: http://www.facetpublishing.co.uk/title.php?id=302789#.WwpjbNVuaAw
 
Research Data Management (RDM) has become a professional topic of great importance internationally following changes in scholarship and government policies about the sharing of research data.
 
Exploring Research Data Management provides an accessible introduction and guide to RDM with engaging tasks for the reader to follow and develop their knowledge. 
 
Starting by exploring the world of research and the importance and complexity of data in the research process, the book considers how a multi-professional support service can be created then examines the decisions that need to be made in designing different types of research data service from local policy creation, training, through to creating a data repository. 
 
Coverage includes:
  • A discussion of the drivers and barriers to RDM
  • Institutional policy and making the case for Research Data Services
  • Practical data management
  • Data literacy and training researchers
  • Ethics and research data services
  • Case studies and practical advice from working in a Research Data Service.
Readership: This book will be useful reading for librarians and other support professionals who are interested in learning more about RDM and developing Research Data Services in their own institution. It will also be of value to students on librarianship, archives, and information management courses studying topics such as RDM, digital curation, data literacies and open science.

Thursday, 11 January 2018

CoreTrustSeal Certification Launched


CoreTrustSeal Certification Launched
Tokyo, Japan and The Hague, Netherlands – 11 September 2017

The ICSU World Data System (ICSU-WDS) and the Data Seal of Approval (DSA) are pleased to announce the launch of a new certification organization: CoreTrustSeal.

The CoreTrustSeal Board offers all interested data repositories a core-level certification based on the DSA–WDS Core Trustworthy Data Repositories Requirements catalogue and procedures. CoreTrustSeal Data Repository certification replaces the DSA certification and the WDS certification of Regular Members.

The CoreTrustSeal is a community-based nonprofit organization promoting sustainable and trustworthy data infrastructures. It is governed by a Standards and Certification Board consisting of members drawn from the Assembly of Reviewers (by election) and the wider repositories stakeholders (appointed).

‘We are driven by our commitments to offer professional certification tools and services to data repositories and to support our voluntary qualified reviewers to conduct audits under optimal conditions’ said Mustapha Mokrane, Chair of the ad hoc CoreTrustSeal Standards and Certification Board. CoreTrustSeal is developing a sustainable business model and as an initial step, will start charging a modest fee to cover administrative costs as of January 2018.

The CoreTrustSeal certification is envisioned as the initial level in a global framework for repository certification that also includes the extended and formal levels. Ultimately, the CoreTrustSeal will endeavour to provide core-level certification for other research entities such as data services and software.


Sunday, 10 December 2017


New report by OECD: "Business models for sustainable research data repositories"

http://www.oecd-ilibrary.org/science-and-technology/business-models-for-sustainable-research-data-repositories_302b12bb-en

Tuesday, 3 October 2017

A Comparative Review of Various Data Repositories

Usability Researcher Derek Murphy and Product Research Specialist Julian Gautier have put together a spreadsheet that compares Dataverse’s features, usage, and governance with other prominent online data repositories:

https://dataverse.org/blog/comparative-review-various-data-repositories

Tuesday, 28 February 2017

Curating Research Data - in two volumes

ACRL has posted the open access editions of the two-volume set titled Curating Research Data on their website at http://www.ala.org/acrl/publications/booksanddigitalresources/booksmonographs/catalog/publications! 
Edited by Lisa R. Johnston:

  • Volume 1 is a traditional edited volume with 12 chapters and explores the variety of reasons, motivations, and drivers for why data curation services are needed in the context of academic and disciplinary data repository efforts.  
  • Volume 2 is a how-to handbook of data curation techniques presented in 8 steps (from receive to reuse) and includes 30 case studies written by practitioners at institutional and disciplinary data repositories. 

COAR Research Data Management survey results

In September 2016, COAR (Confederation of Open Access Repositories) launched a Research Data Management (RDM) Interest Group to help the community expand their operations to support RDM. This group will provide a forum for discussion about managing research data, identify best practices and strategies, and support capacity building for RDM in the repository community.

RDM is wide ranging and there are already many organizations active in this area. The aim of COAR’s work is to promote and articulate the role of research institutions in providing access to research data and to help address the challenges of RDM for the repository community.

See the results of research data management survey, undertaken in December 2016

The COAR Research Data Management Interest Group has identified the following activities for 2016-2018:
  1. Webinars
  2. Develop an assessment framework to help institutions evaluate data repository systems
  3. Undertake a brief survey of RDM activities within the COAR membership
  4. Identify and link to key training material and resources
  5. Share collection policies for RDM in repositories
  6. Recommend metadata standards for research data sets

https://www.coar-repositories.org/activities/repository-content/rda-interest-group-the-long-tail-of-research-data/

Wednesday, 6 January 2016

Digital Curation Centre's new checklist: 'Where to keep research data'.

The Digital Curation Centre is pleased to announce 'Where to keep research data'. This new checklist is concerned with external third-party repositories that offer a managed service to the research community. It aims to assist research support staff whose task is to help researchers make informed choices about where to deposit data. 

The guidance describes different types of service, suggesting some pros and cons of each, and lists sources of information on where to find candidates of the various types.  It then introduces three levels of service that can help match a repository to the depositor’s requirements for their data collection.

See more at: http://www.dcc.ac.uk/resources/how-guides-checklists/where-keep-research-data
This checklist aims to assist research support staff in UK Higher Education Institutions whose task is to help researchers make informed choices about where to deposit data. - See more at: http://www.dcc.ac.uk/resources/how-guides-checklists/where-keep-research-data#sthash.cqR0LrnH.dpuf

Friday, 14 August 2015

Librarians as partners in research data service development at Griffith University

The purpose of the paper is to describe the evolution to date and future directions in research data policy, infrastructure, skills development and advisory services in an Australian university, with a focus on the role of librarians.

The authors have been involved in the development of research data services at Griffith, and the case study presents observations and reflections arising from their first-hand experiences.

Griffith University's organisational structure and 'whole-of-enterprise' approach has facilitated service development to support research data. Fostering strong national partnerships has also accelerated development of institutional capability. Policies and strategies are supported by pragmatic best practice guidelines aimed directly at researchers. Iterative software development and a commitment to well-supported enterprise infrastructure enable the provision of a range of data management solutions. Training programs, repository support and data planning services are still relatively immature. Griffith recognises that information services staff (including librarians) will need more opportunities to develop knowledge and skills to support these services as they evolve.


This case study provides examples of library-led and library-supported activities that could be used for comparative purposes by other libraries. At the same time, it provides a critical perspective by contrasting areas of good practice within the University with those of less satisfactory progress. While other institutions may have different constraints or opportunities, some of the major concepts within this paper may prove useful to advance the development of research data capability and capacity across the library profession.

Friday, 29 May 2015

Promoting Open Knowledge and Open Science: Current State of Repositories.

The report was produced on behalf of the COAR Aligning Repository Networks Committee, with significant input from many representatives of the repository community  - including SPARC.  It provides a high-level overview of the international repository landscape, as well as an interesting 
summary of the current repository environment around the world. It also explores potential new future directions that repositories might consider taking.

The report was also submitted  to the Global Research Council (GRC) and the Research Council’s UK as supplementary material for a GRC workshop that SPARC participated in on the future of scholarly communication held in April in London, and will be distributed to the member representatives of both organizations.

The full report is available here:

Heather Joseph
Executive Director, SPARC
21 Dupont Circle, Suite 800
Washington, DC 20036
+1 202 296 2296
heather@arl.org
http://sparc.arl.org

Tuesday, 7 April 2015

Scientific Data approves UK Data Service as recommended data depository

The UK Data Service is delighted to be the first UK-based social science repository to be listed as a recommended repository by Scientific Data.
Scientific Data, the open-access data journal of Nature Publishing Group, publishes the Data Descriptor article type and recommends that datasets accompanying manuscripts be deposited in established and trusted repositories, such as the UK Data Service ReShare. This ensures that these datasets are stably preserved for the longer-term, thoroughly peer-reviewed, and will be easily accessible to the research community after publication.
Using the UK Data Service’s ReShare repository, researchers can easily upload data collections in the social sciences, humanities and medical research, describe these collections and select the access conditions and licences that are best suited to their data. Researchers can then decide whether to publish these data either as fully open data or as safeguarded data that are made available under the UK Data Service’s End User Licence. Safeguarded data may be selected when anonymised data have been collected from human participants (via surveys or interviews), but where there is a risk of participant de-identification resulting from potential linkage to other data. The ReShare repository has inbuilt safeguards, such as requiring registration in order to access safeguarded datasets. This allows social science data to be shared and made available for research as openly as possible. The UK Data Service also reviews these data before their release, so researchers can be confident that they meet with the necessary ethical and legal requirements.
The UK Data Service fully supports the concept of scientific transparency and is increasingly working with journals to help support their policies. The collaboration with Scientific Data gives social science and humanities researchers the opportunity to increase the discoverability of their data via submission of a Data Descriptor, whilst maintaining UK Data Service’s safeguarding of sensitive data.
See: http://ukdataservice.ac.uk/news-and-events/newsitem/?id=4046 

Monday, 5 January 2015

Science magazine & Data in 2015

Data, eternal

Figure
IMAGE: STACEY PENTLAND PHOTOGRAPHY
During 2014, Science worked with members of the research community, other publishers, and representatives of funding agencies on many initiatives to increase transparency and promote reproducibility in the published research literature. Those efforts will continue in 2015. Connected to that progress, and an essential element to its success, an additional focus will be on making data more open, easier to access, more discoverable, and more thoroughly documented. My own commitment to these goals is deeply held, for I learned early in my career that interpretations come and go, but data are forever.
Figure
“…interpretations come and go, but data are forever.”
IMAGE: EVIRGEN/ISTOCKPHOTO.COM
During my qualifying exam to advance to Ph.D. candidacy, I drew a chalkboard cartoon of a then-new concept: that the weight of recently erupted oceanic volcanoes could elastically deform the surrounding seafloor, creating a deep depression and surrounding flexural arch. Afterward, H. W. Menard, the great marine geologist and a member of my exam committee, spread out a map of the Pacific. He pointed out places where there were older coral atolls (which marked former stands of sea level) that were either now uplifted or drowned in the vicinity of younger volcanoes. Using the distance from the young volcano to the atoll and the amount of uplift or depression, we were able to calibrate the long-term flexural strength of the Pacific seafloor under the weight of the volcanic loading. It was of no matter that Menard had published a paper years earlier using a subset of the uplifted atolls to argue for another hypothesis, which he now happily discarded in favor of the flexural warping one. It occurred to me that there was no database of “drowned and uplifted atolls” that one could access. Menard's prior publication provided only a biased sampling of all occurrences. Had he not been on my exam committee, that unique set of observations to constrain the flexural rigidity might never have presented itself.
Data, particularly those collected with public funding, should be used so that they do the most good. When the greatest number of creative and insightful minds can find, access, and understand the essential features that led to the collection of a data set, the data reach their highest potential. Although the situation has improved some four decades after my student days in terms of the number of public data repositories, requirements for making data available, and metadata standards, there is still a long way to go. So what can Science do to help in this regard, given that it covers many disciplines but is not deeply embedded in any one field?
There are many publicly and privately funded data repositories worldwide, not all of which are being used to their full potential. In 2015, we want to work with authors and readers to identify which of those repositories Scienceshould promote because they are well managed, have long-term support, and are responsive to community needs. For data that do not neatly fit into large-scale repositories, we will explore other available options. We also will evaluate different ways to tag data sets and integrate such tagging into our peer-review process. For example, one might associate a digital identifier for a data set with a figure in a paper. A reviewer could use such an identifier to find the particular data that are related to the figure. The hope is to work with repositories that allow bidirectional tagging so that it is easy for someone—a reviewer or reader—to identify the data used in a Science paper.
Along with improving the “discoverability” of data sets, Science hopes to inspire creative ways to visualize data sets to improve the communication of information and concepts and even facilitate the discoverability of new phenomena. What happens when you bring together those who collect large data sets with those who develop the tools to analyze and view them? Stay tuned!

Monday, 24 November 2014

Over 1,000 research data repositories indexed in re3data.org

in August 2012 re3data.org – the Registry of Research Data Repositories went online with 23 entries. Two years later the registry provides researchers, funding organisations, libraries and publishers with over 1,000 listed research data repositories from all over the world making it the largest and most comprehensive online catalog of research data repositories on the web. re3data.org provides detailed information about the research data repositories, and its distinctive icons help researchers easily identify relevant repositories for accessing and depositing data sets.

To more than 5,000 unique visitors per month re3data.org offers reliable orientation in the heterogeneous landscape of research data repositories. An average of 10 repositories are added to the registry every week. The latest indexed data infrastructure is the new CERN Open Data Portal: http://service.re3data.org/repository/r3d100011381

The project partners from the US and Germany are very happy about the steady growth and the success of the project.

re3data.org encourages funders, libraries and publishers to refer to re3data.org in their policies and guidelines as the primary source for identifying research data repositories.

Initial partners in re3data.org are the Berlin School of Library and Information Science at the Humboldt-Universität zu Berlin, the Library and Information Services department (LIS) of the GFZ German Research Centre for Geosciences, and the KIT Library at the Karlsruhe Institute of Technology (KIT).

This march re3data.org started a process of merging with Databib (http://databib.org), a similar initative at the University Purdue (West Lafayette, Indiana, USA). The aim of this cooperation is to serve the research community with a single, sustainable registry of research data repositories that incorporates the best features of both initiatives.

All records from Databib are now integrated in re3data.org. The process of merging will be finalised in the next months. By the end of 2015, re3data.org will become an imprint of DataCite and be included in its suite of services.

The work of re3data.org is funded by the German Research Foundation (DFG) in Germany and the Institute of Museum and Library Services (IMLS) in the United States.

URL of this announcement:
http://www.re3data.org/2014/11/over-1000-research-data-repositories-indexed-in-re3data-org

Further information:
Website: http://www.re3data.org
Twitter: https://twitter.com/re3data

--------------------------------------------------------------------------------
Heinz Pampel

Helmholtz Association
Helmholtz Open Science Coordination Office

re3data.org - Registry of Research Data Repositories

Helmholtz Centre Potsdam
GFZ German Research Centre for Geosciences
Library of the Albert Einstein Science Park

Friday, 7 November 2014

Software and Repositories


Join in the debate in the comments - ANDS would like to collate more about connecting all data products.
This email stream is from Research Data Management discussion list - sign up for great discussions and international data news


Date:    Tue, 4 Nov 2014 10:21:26 +0000
From:    Tint Hla Hla HTOO <thhhtoo@SMU.EDU.SG>
Subject: Research Data Curation

Hello
I'm a research data librarian from Singapore Management University. At our university, some researchers share their software, code, datasets, etc. on their personal websites and Github. Examples - here<http://libol.stevenhoi.org/> and here<http://olps.stevenhoi.org/>. I thought it might be a good idea to collect them and archive them in our institutional repository for long term access and availability. However, I'm also not sure if it is a good idea for the following reasons:

1) Limitations of the repository

- Our repository is on Digital Commons platform. Not OAIS compliant (lacking many preservation elements). No permanent data identifier (DOI/Handle, etc.).

2) Those data on their websites are already well organized and discoverable on google, etc. So, from the researchers' point of view, what value am I adding to their work by doing data curation?



I've been research data librarian for just about 6 months and not so sure about so many things. So any thought or comment is appreciated.



Thanks very much.

Tint

Ms Tint Hla Hla Htoo
Research Data Services Librarian | Li Ka Shing Library |
Singapore Management University | 70 Stamford Road, Singapore 178901 |
Tel: (65) 6808 7931 | Email: thhhtoo@smu.edu.sg<mailto:thhhtoo@smu.edu.sg> |

------------------------------

Date:    Tue, 4 Nov 2014 12:44:25 +0100
From:    "Jansen Erik (UB)" <erik.jansen@MAASTRICHTUNIVERSITY.NL>
Subject: Re: Research Data Curation

Hi Tint,

Dataverse Network (http://thedata.org/) might meet your requirements.
It's developed by IQSS Harvard University and being used by several Universities from allover the world.
http://thedata.org/book/dataverse-networks-around-world
Best,
Erik


[cid:image001.jpg@01CFF82D.10D84760]
Erik Jansen
dept. Systems
University Library
erik.jansen@maastrichtuniversity.nl <mailto:erik.jansen@maastrichtuniversity.nl>
www.maastrichtuniversity.nl <http://www.maastrichtuniversity.nl>

Grote Looiersstraat 17 - 6211 JH Maastricht - | Universiteitssingel 50 - 6229 ER Maastricht
P.O. Box 616, 6200 MD Maastricht, The Netherlands
T +31 43 38 82 615




Disclaimer<http://www.maastrichtuniversity.nl/web/show/id=6568779/langid=42>
[cid:image002.jpg@01CFF82D.10D84760]Please consider your environmental responsibility before printing this e-mail.




From: Research Data Management discussion list [mailto:RESEARCH-DATAMAN@JISCMAIL.AC.UK] On Behalf Of Tint Hla Hla HTOO
Sent: dinsdag 4 november 2014 11:21
To: RESEARCH-DATAMAN@JISCMAIL.AC.UK
Subject: Research Data Curation

Hello
I'm a research data librarian from Singapore Management University. At our university, some researchers share their software, code, datasets, etc. on their personal websites and Github. Examples - here<http://libol.stevenhoi.org/> and here<http://olps.stevenhoi.org/>. I thought it might be a good idea to collect them and archive them in our institutional repository for long term access and availability. However, I'm also not sure if it is a good idea for the following reasons:

1) Limitations of the repository

- Our repository is on Digital Commons platform. Not OAIS compliant (lacking many preservation elements). No permanent data identifier (DOI/Handle, etc.).

2) Those data on their websites are already well organized and discoverable on google, etc. So, from the researchers' point of view, what value am I adding to their work by doing data curation?



I've been research data librarian for just about 6 months and not so sure about so many things. So any thought or comment is appreciated.



Thanks very much.

Tint

Ms Tint Hla Hla Htoo
Research Data Services Librarian | Li Ka Shing Library |
Singapore Management University | 70 Stamford Road, Singapore 178901 |
Tel: (65) 6808 7931 | Email: thhhtoo@smu.edu.sg<mailto:thhhtoo@smu.edu.sg> |

------------------------------

Date:    Tue, 4 Nov 2014 14:05:20 +0000
From:    Rachel Proudfoot <R.E.Proudfoot@LEEDS.AC.UK>
Subject: Re: Research Data Curation

Hello Tint

Thanks for raising this. I'm sure you're not alone in considering how your local repository fits with other institutional and external systems; we are certainly discussing similar issues here in Leeds. We're envisaging a mixed economy where some data are held and curated in the local repository whereas other data are held in other internal/external services; but we'll want a central record or registry of the data/code etc. regardless of where it is sitting. I think we'll need to apply some basic 'trust and quality' criteria to any services we point to from our central registry. I would worry about digital material linked from a personal web site - it may be well organised and discoverable at the moment, but the danger is it could disappear overnight. We all know the web is littered with dead links.

In terms of creating DOIs for data in a repository, you could do that by signing up to DataCite. This would give some added value to your researchers. https://www.datacite.org/

I think your question is interesting because it prompts us to think about to what extent we want to develop local repositories for our researchers and to what extent we want to utilise services that exist outside the institution - like Erik's DataVerse suggestion.

For example, we have been discussing how to approach archiving software. The partnership between Github and Zenodo is attractive as it allows the depositor to apply a licence, create a DOI and archive the code (as I understand it - I'm sure someone will correct me if I'm off the mark) and Github seems widely used within the developer community.
We're considering whether, at the institution, we:

(i)                  Hold just a metadata record pointing to code in Github/Zenodo  or similar (using the DOI) or

(ii)                Also hold a copy of the code locally
We would be very interested to know what others services are doing.

Whatever your repository service looks like, it's likely to be a much better bet than relying on project/personal web sites.

Best wishes

Rachel


***
Rachel Proudfoot
Research Data Management Advisor
The University Library
University of Leeds
http://researchdata.leeds.ac.uk/
Tel: 0113 343 4554
Skype: rachel_proudfoot





From: Research Data Management discussion list [mailto:RESEARCH-DATAMAN@JISCMAIL.AC.UK] On Behalf Of Tint Hla Hla HTOO
Sent: 04 November 2014 10:21
To: RESEARCH-DATAMAN@JISCMAIL.AC.UK
Subject: Research Data Curation

Hello
I'm a research data librarian from Singapore Management University. At our university, some researchers share their software, code, datasets, etc. on their personal websites and Github. Examples - here<http://libol.stevenhoi.org/> and here<http://olps.stevenhoi.org/>. I thought it might be a good idea to collect them and archive them in our institutional repository for long term access and availability. However, I'm also not sure if it is a good idea for the following reasons:

1) Limitations of the repository

- Our repository is on Digital Commons platform. Not OAIS compliant (lacking many preservation elements). No permanent data identifier (DOI/Handle, etc.).

2) Those data on their websites are already well organized and discoverable on google, etc. So, from the researchers' point of view, what value am I adding to their work by doing data curation?



I've been research data librarian for just about 6 months and not so sure about so many things. So any thought or comment is appreciated.



Thanks very much.

Tint

Ms Tint Hla Hla Htoo
Research Data Services Librarian | Li Ka Shing Library |
Singapore Management University | 70 Stamford Road, Singapore 178901 |
Tel: (65) 6808 7931 | Email: thhhtoo@smu.edu.sg<mailto:thhhtoo@smu.edu.sg> |