DATA-X Pioneering Research Data Exhibition & Symposium

DATA-X has been a University of Edinburgh IS Innovation Fund project, also supported by the Data Lab and ASCUS. The project provided a dynamic platform for University of Edinburgh student researchers across all schools to come together and develop collaborate installations that explore data re-use and interdisciplinary boundaries. Research data are often invisible and complex to comprehend by the public and academic peers, with evolving technology and researcher-driving environments, DATA-X facilitate student researchers with the opportunity to visualize and communicate their research in a user-friendly format to audiences from within and outside the university.

After a series of successful and engaging DATA-X workshops, aimed to inform, shape and create ‘installations’ linked to digital data, the multidisciplinary teams (including students from the School of Architecture and Landscape Architecture, Edinburgh College of Art, Reid School of Music, the School of Engineering, The Centre for Synthetic and Systems Biology, the School of Chemistry, the Centre for Integrative Physiology and the Queen’s Medical Research Institute) continued to work on their installations throughout the summer in preparation for the DATA-X exhibition and Symposium.

DATA-X Exhibition: 

The DATA-X Exhibition ran from 26 November to 6 December 2016, in the Sculpture Court of the Edinburgh College of Art. A total of six physical installations were installed:

eTunes by Dr Siraj Sabihuddin

etunes1A collaborate project for novices to experience the process and creative input required in constructing a musical instrument from start to finish.

 

 

 

Feel the Heat by Nathalie Vladis and Julia Zaenker

feel-the-heatA data quilt, visualising world temperatures between 1961 to 1990. The installation included temperature data sets and interactive colouring maps for audience participation.

 

 

Inside the black box by Luis Fernando Montaño and Bohdan Mykhaylyk

black-boxAn installation simulating bacterial infections. The audience controls the bacterial infection by interactively administering treatment.

 

 

PUROS Sound Box by Dr. Sophia Banou, Dr. Christos Kakalis and Matt Giannotti

D:PDSound BoxSound Box 1_SB Model (1)An installation that ‘defines’ an ambient musical environment, that is conditioned by the movement of users on an interactive floor.

 

 

 

 

Sinterbot by Adela Rabell Montiell and Dr. Siraj Sabihuddin

sintering-process-300x179A hands on demonstration on the alternative use of an ordinary household microwave for sintering, in order to alter material by heat.

 

 

Surface of Significance by Lucas Godfrey and Matt Giannotti

SOS_PROMO1-300x240An audio-visual installation that reconceptualise geographic space. The installation explores the relationship between space, materiality and process.

 

 

 

The exhibition launch, on 26 November, also included three performance installations that serenaded the audience throughout the evening:

  • o ire by Prof. Nick Fells

A live audio performance during which the performance controller sculpt and shape sounds as the piece unfolds.

A composition based on wind data captured during Hurricane Matthew. Musicians captured the chaotic nature of the storm by moving around and inflecting sporadic sound intensity.

An excerpt of Oli Jan’s composition project ‘The Carnival of the Endangered Animals‘. The piece features sounds of endangered species on the IUCN Red List.

DATA-X Symposium

To accompany the exhibition, a DATA-X symposium was held on 1 December 2016 in the Main Lecture Theatre of the Edinburgh College of Art. PhD researchers presented their ‘installations’ and demonstrated the tools, processes and techniques behind the installation. This was an informal event and an open forum to facilitate discussion with an academic and non-academic audience. Guest speakers included Dr Jane Haley, Scientific Coordinator for Edinburgh Neuroscience and FUSION, and Dr James Howie, co-founder of ASCUS. Their talks entitled ‘FUSION –where art meets neuroscience’ and ‘ASCUS and the ASCUS Lab: catalysts for Artisience’, illustrated the efficacy of bridging the gap between the arts and sciences and how innovative, multidisciplinary projects can engage wider audiences and create novel public engagement initiatives.

The next and final phase of the project includes the creation of a DATA-X Exhibition Catalogue in which the students will publish their installations. Updates to follow soon.

Project Team

Data-X Project Manager: Stuart Macdonald (Associate Data Librarian at Edinburgh University Data Library)

Exhibition Coordinator: Dr. Rocio von Jungenfeld (Supported Research Data services at EDINA & Data Library)

Data-X PhD Interns:

Scully Beaver Lynch – PhD candidate in Architecture by Design, Edinburgh College of Art

Adela Rabell Montiel – PhD candidate in Cardiovascular Sciences, Edinburgh Medical School: Clinical Sciences

Cindy Nelson-Viljoen – PhD candidate in Archaeology, School of History, Classics and Archaeology

Dr. Siraj Sabihuddin – PhD in Electronic engineering, School of Engineering

Image credit: DATA-X blog. http://data-x.blogs.edina.ac.uk/

by Cindy Nelson-Viljoen
PhD Student Intern
EDINA and Data Library

Share

Highlights from the RDM Programme Progress Report: May to July 2016

The following key results were highlighted in the RDM Programme Progress Report:

  • There were 42 new users and 69 data management plans created with DMPOnline.
  • An additional 1.5PB has been procured for DataStore’s general capacity expansions.
  • The Roslin Institute has deposited 16 datasets into Data Vault.
  • DataShare upload release (2.1) went live on 23 May 2016.
  • There are now 334 dataset records in PURE, an increase of 124 records from the last reporting period (February to April 2016).
  • 54 datasets have been deposited into DataShare.
  • The University of Edinburgh was recommended as a preferred supplier on the Framework for the Research Data Management Shared Services for Jisc Services Ltd (JSL) for the following Lots:
  • Lot 2: Repository Interfaces
  • Lot 3: Data Exchange Interface
  • Lot 6: Research Data Preservation Tools Development
  • Lot 8: User Experience Enhancements
  • A total of 390 staff and postgraduates attended RDM courses and workshops during this quarter.
  • A total of 3,649 learners enrolled for the 5-week RDMS MOOC rolling course from March through July, 2016 and a total of 461 people completed the course in the same time frame.
  • There were 5,198 MANTRA sessions recorded from May to July with 58 to 60 percent identified as new users.
  • Set up an RDM Forum in collaboration with College of Arts, Humanities and Social Sciences (CAHSS) Research Officer and Research Outputs Co-ordinator. The first RDM forum is scheduled for Wednesday, 7 September 2016.

Data Management Planning highlights

We currently hold sample data management plans for grant applications submitted to the Arts and Humanities Research Council (AHRC) the Economic and Social Research Council (ESRC) and the Medical Research Council (MRC).

 Active Data Infrastructure highlights

DataStore

An additional 1.5PB has been procured for general capacity expansions. This capacity will primarily be deployed to the College of Medicine & Veterinary Medicine (CMVM) and the College of Science & Engineering (CSE).

MRC Institute of Genetics & Molecular Medicine (IGMM) has purchased an additional 1.2PB of capacity, and this is now deployed in their dedicated file system.

Data Stewardship highlights

DataShare

The large data sharing investigation was completed for DataShare and reported previously. Upload release (2.1) went live on 23 May 2016. Download release planned following ‘embargo release’ and ShareGeo spatial data migration.

Data Vault

There was a soft release of Data Vault in February 2016, with the Roslin Institute depositing 16 datasets during this quarter.

PURE

There are now 334 dataset records in PURE, an increase of 124 records from the last reporting period (February to April 2016).

Research Data Discovery Service (RDDS)

Two PhD interns are working on School engagement activities (dataset records into PURE / datasets into DataShare) for Divinity & Division of Infection and Pathway Medicine; contracts end 16 September 2016. One PhD intern retrospectively added DataShare metadata to PURE for data deposits prior to PURE Data Catalogue functionality; contract to end 16 September 2016. A fourth PhD intern (to work with School of Informatics) is awaiting for approval.

Data Management Support highlights

A total of 390 staff and postgraduates attended RDM courses and workshops during this quarter.

Other related research data management support activities to highlight

  • A talk was given ‘Understanding and overcoming challenges to sharing personal and sensitive data’ at the ReCon (Research Communication and Data Visualisation) Conference, 24th June 2016, The Edinburgh Centre for Carbon Innovation (ECCI).
  • ‘Working with sensitive data in research’ guide was written for research staff and students in social sciences.
  • Another guide is being written on ‘Sharing and retaining data’ for research staff and students in social sciences.
  • Set up an RDM Forum in collaboration with College of Arts, Humanities and Social Sciences (CAHSS) Research Officer and Research Outputs Co-ordinator. The first RDM forum is scheduled for Wednesday, 7 September 2016.

Other activities to highlight

The outcome of Jisc RDM Shared Services bid that was submitted in March 2016

The Procurement Panel has recommended University of Edinburgh as a preferred supplier on the Framework for the Research Data Management Shared Services for Jisc Services Ltd (JSL) for the following Lots:

  • Lot 2: Repository Interfaces
  • Lot 3: Data Exchange Interface
  • Lot 6: Research Data Preservation Tools Development
  • Lot 8: User Experience Enhancements

Unfortunately, the Procurement Panel has decided not to recommend University of Edinburgh for the following Lots:

  • Lot 1: Research Data Repository
  • Lot 4: Research Information and Administration Systems Integrations

National and International Engagement Activities

From May to June

Stuart Macdonald and Rocio von Jungenfeld ran three workshops for the IS Innovation Fund project, Data-X: Pioneering Research Data Exhibition, with PhD students from across the University. Introduction to Data-X: Pioneering Research Data Exhibition

In June

Stuart Macdonald presented peer-reviewed presentation to IASSIST conference, Bergen: Supporting the development of a national Research Data Discovery Service – a Pilot Project

Robin Rice presented a poster at Open Repositories 2016, Dublin: Data Curation Lifecycle Management at the University of Edinburgh

Pauline Ward presented a lightning talk at Open Repositories 2016, Dublin:  Growing Open Data: Making the sharing of XXL-sized research data files online a reality, using Edinburgh DataShare

Stuart Macdonald was an invited speaker at NFAIS (National Federation of Abstracting and Information Services) Fostering Open Science Virtual Seminar: NFAIS Fostering Open Science Virtual Seminar

In July

Robin Rice gave two presentations (invited and peer-reviewed) at LIBER 2016, Helsinki: University of Edinburgh RDM Training: MANTRA & beyond; Designing and delivering an international MOOC on Research Data Management and Sharing

Robin Rice filled in for Stuart Lewis as invited speaker for JISC-CNI 2016, London: Managing active research in the University of Edinburgh

This is the last quarterly report as the Research Data Management (RDM) Roadmap Project (August 2012 to July 2016) came to a close on 31 July 2016.

There will be discussions with the RDM Steering Group to decide how future reporting will be conducted. These reports will be released on the Research Data Blog as well.

Tony Mathys
Research Data Management Service Co-ordinator

 

Share

Research Data Management (RDM) Forum

RDM Forum is a newly created platform to bring together both researchers and research & IT support staff from across the University whose role involves helping academics in managing their research data. The aim of the Forum is to share good practice, exchange experiences as well as discuss current and future challenges related to data curation, preservation and publishing. We hope that the Forum will allow its participants to learn from one another and gain a new perspective on some common issues.

The Forum takes the form of meetings as well as e-mail updates (done through the RDM Forum mailing list) and an online platform (SharePoint website) for sharing useful resources, engaging with each other and keeping up-to-date with recent developments in RDM.

The first meeting took place on 7th December 2016. There were 24 in attendance and participants had the opportunity to introduce themselves, ask questions, and provide their expectations and suggestions for future RDM Forum meetings, which have been summarised below:

  • Overcoming challenges:
    • Supporting academic engagement
    • Going beyond funder requirements
    • Engagement beyond training
    • Avoiding last-minute arrangements
    • Addressing concerns about data sharing and reuse
  • Finding solutions that will work
    • Early training
    • Establishing workflows for standard processes
    • Developing an Information Governance structure for data
    • Sharing real-life scenarios
  • Forum structure
    • Forming several user groups focused on specific aspects of RDM
    • Organising meetings around specific themes
    • Updates from Research Data Service team
    • Forum as a platform for training
    • Forum to meet every two months at different locations

If you are interested in joining the Forum mailing list you can do so at: https://mlist.is.ed.ac.uk/lists/info/rdm-forum
RDM Forum SharePoint website (access by request) is available at:
https://uoe.sharepoint.com/sites/rdmforum

Cuna Ekmekcioglu
Senior Research Data Officer

Share

Highlights from the RDM Programme Progress Report: February to April 2016

The membership of the Research Data Service Virtual Team across four divisions of IS was confirmed and met for the first time (to replace the former action group meetings) on 11 February where it was agreed meetings would be held approximately every six weeks for information and decision-making.

In February, the DataShare metadata was mapped to the PURE metadata and staff in L&UC and Data Library trained each other for creating dataset records in Pure and reviewing submissions in DataShare. It was agreed that staff would create records in Pure for items deposited in DataShare until the company (Elsevier) provides a mechanism for automatically inputting records into Pure.

In March, Jisc announced that the University of Edinburgh was selected as a framework supplier for their new Research Data Management Shared Service.

A review of the existing ethics processes in each college is in progress with Jacqueline McMahon at the College of Arts, Humanities and Social Sciences (CAHSS) to create a University-wide ethics template. There is also engagement with the School ethics committees at the School of Health in Social Sciences (HiSS), Moray House School of Education (MHSE), Law and School of Social and Political Science (SPS) in CAHSS.

The Research Data Management and Sharing (RDMS) Coursera MOOC opened for enrolment on 1 March 2016. This was completed in partnership with the University of North Carolina-Chapel Hill CRADLE project. Research Data Management and Sharing (RDMS) MOOC stats from the Coursera Dashboard reveal that as of 23 May 2016, there have been 5,429 visitors and 1,526 active learners; 335 visitors have completed the course.

The large data sharing investigation was completed for DataShare and reported previously. (Two new releases in DataShare defined: upload and download). Upload release (2.1) to go live 23 May 2016.

PURE dataset functionality is now included in standard PURE and Research Data Management (RDM) training. There are now 210 dataset records in PURE.

Four PhD interns were hired in mid-March to act as College representatives for the IS Innovation Fund Pioneering Research Data Exhibition. They will be employed until mid-December 2016.

A total of 363 staff and postgraduates attended RDM courses and workshops during this quarter.

There were 30 new DMPonline users and 55 new plans created during this quarter.

There are now 210 dataset metadata records in PURE.

A total of 56 datasets were deposited in DataShare during this quarter.

The total number of DataStore users rose from 12,948 in the previous quarter to 13,239 in this quarter, an increase of 291 new users.

National and International Engagement Activities

In February

  • Stuart Lewis gave a DataVault presentation at the International Digital Curation Conference (IDCC) in Amsterdam.

In March

  • A University news item was released to mark the launch of the Research Data Management and Sharing (RDMS) MOOC on Coursera. http://www.ed.ac.uk/news/2016/dataskills-010316
  • Stuart MacDonald gave an RDM presentation to trainee physicians at the Royal College of Physicians Edinburgh Course: Critical appraisal and research for trainees, Edinburgh. http://www.slideshare.net/smacdon2/rdm-for-trainee-physicians
  • Three delegates from Göttingen University were hosted here. The delegates have shared interests in RDM and visited to gain more insight into RDM support and experiences here.
  • Robin Rice gave an invited talk about the RDMS MOOC and web-based Survey Documentation and Analysis (SDA) tool to Learning, Teaching and Web and elearning@Ed Showcase and Network monthly gathering.

In April

As part of my responsibilities to cover the one year interim of Kerry Miller’s maternity leave, I will be writing blogs for this page until Kerry returns next summer.

Prior to this post, I worked the past 12 years as the geospatial metadata co-ordinator at EDINA. My primary role was to promote and support research data management and sharing amongst UK researchers and students using spatial data and geographical information.

Tony Mathys
Research Data Management Service Co-ordinator


Share

Highlights from the RDM Programme Progress Report: November 2015 – January 2016

Data Seal of Approval have awarded DataShare Trusted Repository status their assessment of our service can be read at https://assessment.datasealofapproval.org/assessment_175/seal/html/. In addition a major new release of DataShare was completed in November, this makes the code open in Github as well as making general improvements to the look and feel of the website.

The ‘interim’ DataVault is now in final testing and will be rolled out on a request basis to those researchers who can demonstrate an urgent need to use the service now rather than waiting until the final version is ready later this year. The phase three funding for development of the DataVault has been received from Jisc, this runs from March to August, so the final version should be ready for launch sometime after this.  The project was presented at the International Digital Curation Conference in February 2016.

Over the three month period a total of 328 staff and PGR™s have attended a RDM course or workshop.

Work on the MANTRA MOOC is expected to be finalised in February and launched on 1st March, at the following URL: https://www.coursera.org/learn/data-management

Continue reading

Fostering open science in social science

FOSTER_logoOn 10th of June, the Data Library team ran two workshops in association with the EU Horizon 2020 project, FOSTER (Facilitate Open Science Training for European Research), and the Scottish Graduate School of Social Science.

The aim of the morning workshop, “Good practice in data management & data sharing with social research,� was to provide new entrants into the Scottish Graduate School of Social Science with a grounding in research data management using our online interactive training resource MANTRA, which covers good practice in data management and issues associated with data sharing.

The morning started with a brief presentation by Robin Rice on ‘open science’ and its meaning for the social sciences. Pauline Ward then demonstrated the importance of data management plans to ensure work is safeguarded and that data sharing is made possible. I introduced MANTRA briefly, and then Laine Ruus assigned different MANTRA units to participants and asked them to briefly go through the units and extract one or two key messages and report back to the rest of the group. After the coffee break we had another presentation on ethics, informed consent and the barriers for sharing, and we finished the morning session with a ‘Do’s and Dont’s exercise where we asked participants to write in post-it notes the things they remembered, the things they were taking with them from the workshop: green for things they should DO, and pink for those they should NOT. Here are some of the points the learners posted:

DO
– consider your usernames & passwords
– read the Data Protection Act
– check funder/institution regulations/policies
– obtain informed consent
– design a clear consent form
– give participants info about the research
– inform participants of how we will manage data
– confidentiality
– label your data with enough info to retrieve it in future
– develop a data management plan
– follow the certain policies when you re-use dataset[s] created by others
– have a clear data storage plan
– think about how & how long you will store your data
– store data in at least 3 places, in at least 2 separate locations
– backup!
– consider how/where you back up your data
– delete or archive old versions
– data preservation
– keep your data safe and secure with the help of facilities of fund bodies or university
– think about sharing
– consider sharing at all stages. Think about who will use my data next
– share data (responsibly)

DON’T
– unclear informed consent
– a sense of forcing participants to be part of research
– do not store sensitive information unless necessary
– don’t staple consent forms to de-identified data records/store them together
– take information security for granted
– assume all software will be able to handle your data
– don’t assume you will remember stuff. Document your data
– assume people understand
– disclose participants’ identity
– leave computer on
– share confidential data
– leave your laptop on the bus!
– leave your laptop on the train!
– leave your files on a train!
– don’t forget it is not just my data, it is public data
– forget to future proof

Robin Rice presenting at FOSTERing Open Science workshop

Our message was that open science will thrive when researchers:

  • organise and version their data files effectively,
  • provide comprehensive and sufficient documentation for others to understand and replicate results and thus cite the source properly
  • know how to store and transport your data safely and securely (ensuring backup and encryption)
  • understand legal and ethical requirements for managing data about human subjects
  • Recognise the importance of good research data management practice in your own context

The afternoon workshop on “Overcoming obstacles to sharing data about human subjects� built on one of the main themes introduced in the morning, with a large overlap of attendees. The ethical and regulatory issues in this area can appear daunting. However, data created from research with human subjects are valuable, and therefore are worth sharing for all the same reasons as other research data (impact, transparency, validation etc). So it was heartening to find ourselves working with a group of mostly new PhD students, keen to find ways to anonymise, aggregate, or otherwise transform their data appropriately to allow sharing.

Robin Rice introduced the Data Protection Act, as it relates to research with human subjects, and ethical considerations. Naturally, we directed our participants to MANTRA, which has detailed information on the ethical and practical issues, with specific modules on “Data protection, rights & access� and “Sharing, preservation & licensing�. Of course not all data are suitable for sharing, and there are risks to be considered.

In many cases, data can be anonymised effectively, to allow the data to be shared. Richard Welpton from the UK Data Archive shared practical information on anonymisation approaches and tools for ‘statistical disclosure control’, recommending sdcMicroGUI (a graphical interface for carrying out anonymisation techniques, which is an R package, but should require no knowledge of the R language).

DrNiamhMooreFinally Dr Niamh Moore from University of Edinburgh shared her experiences of sharing qualitative data. She spoke about the need to respect the wishes of subjects, her research gathering oral history, and the enthusiasm of many of her human subjects to be named in her research outputs, in a sense to own their own story, their own words.

Links:

Rocio von Jungenfeld & Pauline Ward
EDINA and Data Library

Share

EPSRC Expectations Awareness Survey

As many of you will already know EPSRC set out its research data management (RDM) expectations for institutions in receipt of EPSRC grant funding in April 2012, this included the development of an institutional ‘Roadmap’. EPSRC assessment of compliance with these expectations will begin on 1 May 2015 for research outputs published on or after that date.

In order to comply with EPSRC expectations and to implement the University’s RDM Policy, the University of Edinburgh has invested significantly in RDM services, infrastructure (incl. storage and security) and support as detailed in the University of Edinburgh’s RDM Roadmap.

In an effort to gauge the University of Edinburgh’s ‘readiness’ in relation to EPSRC’s RDM expectations, we are conducting a short survey of EPSRC grant holders.

The survey aims to find out more about researcher awareness of those expectations concerning the management and provision of access to EPSRC-funded research data as detailed in the EPSRC Policy Framework on Research Data.

We aim to conduct follow-up interviews with EPSRC grant holders who are willing to talk through these issues in a bit more detail to help shape the development of the RDM services at the University of Edinburgh.

We will endeavour to make available some of our findings shortly. In the meantime, if you want to use or refer to our survey we have posted a ‘demo’version below:
https://edinburgh.onlinesurveys.ac.uk/epsrc-expectations-awareness-demo

Should you decide to make use of our survey, let us know, as we can potentially share our data with each other to benchmark our progress.

(As an aside Oxford University have crafted a useful data decision tree for EPSRC-funded researchers at Oxford)

Regards
Stuart Macdonald
RDM Services Coordinator
stuart.macdonald@ed.ac.uk

Share

Open up! On the scientific and public benefits of data sharing

Research published a year ago in the journal Current Biology found that 80 percent of original scientific data obtained through publicly-funded research is lost within two decades of publication. The study, based on 516 random journal articles which purported to make associated data available, found the odds of finding the original data for these papers fell by 17 percent every year after publication, and concluded that “Policies mandating data archiving at publication are clearly needed� (http://dx.doi.org/10.1016/j.cub.2013.11.014).

In this post I’ll touch on three different initiatives aimed at strengthening policies requiring publicly funded data – whether produced by government or academics – to be made open. First, a report published last month by the Research Data Alliance Europe, “The Data Harvest: How sharing research data can yield knowledge, jobs and growth.�  Second, a report by an EU-funded research project called RECODE on “Policy Recommendations for Open Access to Research Data�, released last week at their conference in Athens.  Third, the upcoming publication of Scotland’s Open Data Strategy, pre-released to attendees of an Open Data and PSI Directive Awareness Raising Workshop Monday in Edinburgh.

Experienced so close together in time (having read the data harvest report on the plane back from Athens in between the two meetings), these discrete recommendations, policies and reports are making me just about believe that 2015 will lead not only to a new world of interactions in which much more research becomes a collaborative and integrative endeavour, playing out the idea of ‘Science 2.0’ or ‘Open Science’, and even that the long-promised ‘knowledge economy’ is actually coalescing, based on new products and services derived from the wealth of (open) data being created and made available.

‘The initial investment is scientific, but the ultimate return is economic and social’

John Wood, currently the Co-Chair of the global Research Data Alliance (RDA) as well as Chair of RDA-Europe, set out the case in his introduction to the Data Harvest report, and from the podium at the RECODE conference, that the new European commissioners and parliamentarians must first of all, not get in the way, and second, almost literally ‘plan the harvest’ for the economic benefits that the significant public investments in data, research and technical infrastructure are bringing.

CaptureThe report’s irrepressible argument goes, “Just as the World Wide Web, with all its associated technologies and communications standards, evolved from a scientific network to an economic powerhouse, so we believe the storing, sharing and re-use of scientific data on a massive scale will stimulate great new sources of wealth.â€� The analogy is certainly helped by the fact that the WWW was invented at a research institute (CERN), by a researcher, for researchers. The web – connecting 2 billion people, according to a McKinsey 2011 report, contributed more to GDP globally than energy or agriculture. The report doesn’t shy away from reminding us and the politicians it targets, that it is the USA rather than Europe that has grabbed the lion’s share of economic benefit– via Internet giants Google, Amazon, eBay, etc. – from the invention of the Web and that we would be foolish to let this happen again.

This may be a ruse to convince politicians to continue to pour investment into research and data infrastructure, but if so it is a compelling one. Still, the purpose of the RDA, with its 3,000 members from 96 countries is to further global scientific data sharing, not economies. The report documents what it considers to be a step-change in the nature of scientific endeavour, in discipline after discipline. The report – which is the successor to the 2010 report also chaired by Wood, “Riding the Wave: How Europe can gain from the rising tide of scientific data,” celebrates rather than fears the well-documented data deluge, stating,

“But when data volumes rise so high, something strange and marvellous happens: the nature of science changes.�

The report gives examples of successful European collaborative data projects, mainly but not exclusively in the sciences, such as the following:

  • Lifewatch – monitors Europe’s wetlands, providing a single point to collect information on migratory birds. Datasets created help to assess the impact of climate change and agricultural practices on biodiversity
  • Pharmacog – partnership of academic institutions and pharmaceutical companies to find promising compounds for Alzheimer’s research to avoid expensive late-stage failures of drugs in development.
  • Human Brain Project – multidisciplinary initiative to collect and store data in a standardised and systematic way to facilitate modelling.
  • Clarin – integrating archival information from across Europe to make it discoverable and usable through a single portal regardless of language.

The benefits of open data, the report claims, extends to three main groups:

  • to citizens, who will benefit indirectly from new products and services and also be empowered to participate in civic society and scientific endeavour (e.g. citizen science);
  • to entrepeneurs, who can innovate based on new information that no one organisation has the money or expertise to exploit alone;
  • to researchers, for whom the free exchange of data will open up new research and career opportunities, allow crossing of boundaries of disciplines, institutions, countries, and languages, and whose status in society will be enhanced.

‘Open by Default’

If the data harvest report lays out the argument for funding open data and open science, the RECODE policy recommendations focus on what the stakeholders can do to make it a reality. The project is fundamentally a research project which has been producing outputs such as disciplinary case studies in physics, health, bioengineering, environment and archaeology. The researchers have examined what they consider to be four grand challenges for data sharing.

  • Stakeholder values and ecosystems: the road towards open access is not perceived in the same way by those funding, creating, disseminating, curating and using data.
  • Legal and ethical concerns: unintended secondary uses, misappropriation and commercialization of research data, unequal distribution of scientific results and impacts on academic freedom.
  • Infrastructure and technology challenges: heterogeneity and interoperability; accessibility and discoverability; preservation and curation; quality and assessibility; security.
  • Institutional challenges: financial support, evaluating and maintaining the quality, value and trustworthiness of research data, training and awareness-raising on opportunities and limitations of open data.

Capture1RECODE gives overarching recommendations as well as stake-holder specific ones, a ‘practical guide for developing policies’ with checklist for the four major stakeholder groups: funders, data managers, research institutions and publishers.

‘Open Changes Everything’

The Scottish government event was a pre-release of the  open data strategy, which is awaiting final ministerial approval, though in its final draft, following public consultation. The speakers made it clear that Scotland wants to be a leader in this area and drive culture change to achieve it. The policy is driven in part by the G8 countries’ “Open Data Charterâ€� to act by the end of 2015 on a set of five basic principles – for instance, that public data should be open to all “by defaultâ€� rather than only in special cases, and supported by UK initiatives such as the government-funded Open Data Institute and the grassroots Open Knowledge Foundation.

Capture

Improved governance (or public services) and ‘unleashing’ innovation in the economy are the two main themes of both the G8 charter and the Scotland strategy. The fact was not lost on the bureaucrats devising the strategy that public sector organisations have as much to gain as the public and businesses from better availability of government data.

The thorny issue of personal data is not overlooked in the strategy, and a number of important strides have been taken in Scotland by government and (University of Edinburgh) academics recently on both understanding the public’s attitudes, and devising governance strategies for important uses of personal data such as linking patient records with other government records for research.

According to Jane Morgan from the Digital Public Services Division of the Scottish Government, the goal is for citizens to feel ownership of their own data, while opening up “trustworthy uses of data for public benefit.�

Tabitha Stringer, whose title might be properly translated as ‘policy wonk’ for open data, reiterated the three main reason for the government to embrace open data:

  • Transparency, accountability, supporting civic engagement
  • Designing and delivering public services (and increasingly digital services)
  • Basis for nnovation, supporting the economy via growth of products & services

‘Digital first’

The remainder of the day focused on the new EU Public Service Information directive and how it is being ‘transposed’ into UK legislation to be completed this year. In short, the Freedom of Information and other legislation is being built upon to require not just publication schemes but also asset lists with particular titles by government agencies. The effect of which, and the reason for the awareness raising workshop is that every government agency is to become a data publisher, and must learn how to manage their data not just for their own use but for public ‘re-users’. Also, for the first time academic libraries and other ‘cultural organisations’ are to be included in the rules, where there is a ‘public task’ in their mission.

‘Digital first’ refers to the charging rules in which only marginal costs (not full recovery) may be passed on, and where information is digital the marginal cost is expected to be zero, so that the vast majority of data will be made freely available.

keep-calm-and-open-data-11Robin Rice
EDINA and Data Library

 

 

Share

New release of Research Data MANTRA (Management Training) online course

The Research Data MANTRA course is an open, online training course that provides instruction in good practice in research data management. There are nine interactive learning units on key topics such as data management planning, organising and formatting data, using shared data and licensing your own data, as well as four data handling tutorials with open datasets for use in R, SPSS, NVivo and ArcGIS.

This fourth release of MANTRA has been revised and systematically updated with new content, videos, reading lists, and interactive quizzes. Three of the data handling tutorials have been rewritten and tested for newer software versions too.

New content in the online learning modules with the September, 2014 release:

  • New video footage from previous interviewees and introducing Richard Rodger, Professor of Economic and Social History and Stephen Lawrie, Professor of Psychiatry & Neuro-Imaging
  • Big Data now in Research Data Explained
  • Data citation and ‘reproducible research’ added to Documentation and Metadata
  • Safe password practice and more on encryption in Storage and Security
  • Refined information about the DPA and IPR in Data Protection, Rights and Access
  • Linked Open Data and CC 4.0 and CC0 now covered in Sharing, Preservation & Licensing

MANTRA home pageThis release will also be more stable and more accessible due to back-end enhancements. The flow of the learning units and usability of quizzes have been improved based on testing and feedback. We have simplified our feedback form and added a four-star rating button to the home page. A YouTube playlist for each unit is available on the Data Library channel.

MANTRA was originally created with funding from Jisc and is maintained by EDINA and Data Library, a division of Information Services, University of Edinburgh. It is an integral part of the University’s Research Data Management Programme and is designed to be modular and self-paced for maximum convenience; it is a non-assessed training course targeted at postgraduate research students and early career researchers.

Data management skills enable researchers to better organise, document, store and share data, making research more reproducible and preserving it for future use. Researchers in 144 countries used MANTRA last year, which is available without registration from the website. Postgraduate training organisations in the UK, Canada, and Australia have used the Creative Commons licensed material in the Jorum repository to create their own training. The website also hosts a ‘training kit’ for librarians wishing to increase their skills in supporting Research Data Management.

Visit MANTRA and consider recommending it to your colleagues and research students this term! http://datalib.edina.ac.uk/mantra/

Usage Statistics

According to Google Analytics, the following organisation’s websites were the top ten referrers to the MANTRA website for the academic year 2013-2014 (discounting Data Library, EDINA and Information Services):

  • Institute for Academic Development, University of Edinburgh
  • LIS Links (India)
  • Digital Curation Centre
  • eScience Portal for New England Libraries at University of Massachusetts Medical Library
  • Oxford University
  • University of Nebraska-Lincoln (USA)
  • Carleton University (Canada)
  • Glasgow University
  • Food and Agriculture Organization of the United Nations
  • Jisc

Social media sites Facebook, Twitter and Slideshare provided a large number of referrals; several more came from other UK institutions, and HEIs in Australia, the rest of Europe, and North America—University Library pages especially. Forty percent of sessions came  from a referring website.

Visitors to MANTRA over the year came from 144 countries. Google searches accounted for 4,000 sessions, 25% of the total. Nearly ten thousand visits were from new users (based on IP addresses) over the year from 22nd August, 2013 – 23rd August, 2014. Here is a link to a Google Analytics summary spreadsheet extracted from our account.

We expect to have more detailed usage statistics over the forthcoming year due to moving the learning units out of the authoring software (Xerte Online Toolkits) onto the main MANTRA website.

Postscript, 15 Sept: See my Storify story, “Research Data MANTRA Buzz” to find out who’s been talking about MANTRA on twitter!

Robin Rice
Data Librarian

 

 

Share

Dealing with Data Conference & RDM Service Launch – summary

University of Edinburgh Research Data Management Service LogoInformation Services (IS) held a half-day conference in the Main Library on the subject of ‘Dealing with Data’ to coincide with the launch of the University of Edinburgh’s Research Data Management support services on 26 August.

University researchers presented to over 120 delegates from across the disciplinary and support spectrum on many aspects of working with data, particularly research with novel methods of creating, using, storing, or sharing data. Subjects included Big Data for disease control, managing West Nilotic language sound files, sharing brain images, geospatial metadata services, visualising qualitative data via carpets!

Dealing with Data Conference

The RDM Programme team are currently collecting feedback and will report on this and the conference in more detail via this blog.

‘Dealing with Data Conference’ delegates then gathered in the Main Library foyer to hear brief talks by Professor Jeff Haywood, Professor Peter Clarke and Dr John Scally followed by the formal launch of the RDM Services by the University’s Principal, Sir Timothy O’Shea who underlined the successful collaboration between research and support service communities in establishing research support services worthy of a leading UK research-intensive university.

University of Edinburgh RDM Service launch by Sir Timothy O'Shea

A ‘storify’ story of tweets collected during the launch and the conference is available, with pictures and perspectives from various attendees.

The launch of the IS-led RDM Services is the culmination of work detailed in the RDM Roadmap which began in earnest in August 2012 following approval of the RDM Policy by the University Court in May 2011.

Details of available and planned RDM Services for University of Edinburgh researchers were reported on in the blogpost: RDM Roadmap: Completion of Phase 1

Conference presentations can be downloaded from Edinburgh Research Archive (ERA) at: https://www.era.lib.ed.ac.uk/handle/1842/9389

Stuart Macdonald
RDM Service Coordinator
stuart.macdonald@ed.ac.uk

Share