Byte by byte: developing our digital preservation capability
A brief timeline of the digital preservation journey at BGS and NGDC.
04/11/2021
In 2016, the initial thoughts I had to explore creating a digital continuity at BGS were starting to develop. I had finished my MSc thesis, ‘Exploring digital preservation requirements: A case study from the National Geoscience Data Centre (NGDC)’, which led to my peer-reviewed article being published in the Records Management Journal. I discovered that UKRI (then still RCUK) was a corporate member of the Digital Preservation Coalition (DPC), so I approached Juan Bicarregui, Chair of the DPC, to ask if I could join. I soon got into the habit of attending DPC events and meeting digital preservationists from around the world. Their jobs sounded fascinating so I took a postgraduate diploma in digital preservation at Aberystwyth University and learned more.
At BGS and NGDC, we hold lots of research data, both digital and analogue. I had run a small stakeholder survey as part of my thesis about the need to maintain the long-term accessibility and usability of our geoscience data. I started planning our work on a shoestring budget, talking to both research and data management staff, and it was clear that we needed a policy on how to deal with ‘aging’ digital data. Luckily, I had attended a workshop on how to write a digital preservation policy and, after researching other organisations’ policies, I wrote the first one for BGS.
I was also rummaging in the BGS legacy media store (which contains thousands of pieces of various storage media) learning how to use the National Archives’ file format identification tool DROID, and talking to colleagues about their floppy disks, Lotus spreadsheets, Bentley MicroStations, emulation setups and old LTO tapes. These discussions gave me the idea to set up a pop-up computer museum on the first World Digital Preservation Day (WDPD) in 2017.
It turned out to be very popular with BGS colleagues, some of whom had worked with the gadgets on display. Our data centre staff started to get involved in the outreach work and we were having more ad hoc discussions about adding preservation capability to our procedures.
During World Digital Preservation Day (WDPD) 2018, I delivered a taster preservation training session and a lunchtime talk about our preservation strategy, which was being developed at that time. We also published our first preferred file formats list and studied the PREMIS preservation metadata schema, with a view to developing a module to add to our Discovery Metadata.
In 2019, we ran a digital research data survey for BGS researchers to find out what was really happening at a grassroots level. The purpose was to inform our programme development and to strengthen links between research data management (RDM) and digital preservation. We then published an internal report describing our findings and started doing a gap analysis between the researchers’ needs and the RDM service provision. To showcase how we were combining our data management training course and preservation capability, I gave a talk at the International Digital Preservation Conference at the Eye Film Museum in Amsterdam.
Just before the lockdown in 2020, BGS data scientist Alex helped us by scanning terabytes of data on corporate shared drives to find out exactly what we’ve got. The updated US Library of Congress annual recommended formats statement, which now included GIS, geospatial and 3D data, was useful as we updated our preferred formats list. We were also exploring the technical side of creating checksums and running fixity checks in our ingestion workflow when things came to a halt in March 2020.
In 2021, we picked up the work with a blast as we finally created a dedicated digital preservation team! Every team member has been at BGS for quite a while (more than 15 years) so they are well versed in geoscience data, as well as our data management processes and workflows. The team members had top-up training through the National Archives courses, got access to the Digital Preservation Coalition website and training materials, and we had many lively discussions about enhancing our capability at team meetings. This gave everyone more confidence to integrate preservation thinking and activities within their existing workflows.
After the first six months of working together, I invited the team to provide feedback on our work so far.
Working in the data management area at BGS for over 20 years, I have seen many changes in how data is captured. It has now become clear that digital data preservation is a key issue for the future of the data we hold. When Jaana asked me to join the newly formed digital preservation team I was very keen to get involved. I have spent the first few months reading articles and blogs, and undertaking the TNA/DPC training to give me the skills to help develop and implement digital preservation strategies and workflows, in particular to look at the ingestion, access, use and reuse of digital information and take active steps to preserve it for the future.
Sally Stolworthy.
My work with digital preservation began at the start of my career 15 years ago, getting in on the ground floor with digital capture of analogue records, both for wider delivery to science and as disaster recovery. In doing so I embraced open, long-term reproduction standards so that no one (myself included) would have to repeat the capture exercise again. I then moved to managing marine data, involving gradually migrating our data holdings to non-proprietary formats where possible. The team was small and the work varied, so it was important to make sure that I could pick up my own work again in the future, as there’s little more embarrassing than not understanding your own work. This meant things like embedding naming conventions into files and folder structures and writing documentation that explains what is here, what was done and what is still to be done were important.
I expanded this experience out to the wider records collections and collaborated on guidance on implementation of scanning standards, ingestion of other organisations’ data and, perhaps most importantly, worked on getting data back out to scientists and public users. Helping users understand our data holdings means they can do innovative things with them and there is a reciprocity in that they then understand how to organise and document their work for others to benefit from.
Rob Cooke.
About the author

Jaana Pinnick
Data and Information Governance Manager
Relative topics
Latest blogs

Industry-leading data sharing partnership announced
02/11/2023
A data sharing partnership has been agreed between BGS and Ossian, allowing BGS to advance its knowledge of the rock and soil conditions under the seabed.

The art of boreholes: Essex artists visit the BGS to be inspired by our library of geological core
02/11/2023
Two UK-based artists visitors aim to turn art and earth science into a collaborative experience that facilitates discussion on land usage.

Good practice for sand mining
24/10/2023
Tom Bide and Clive Mitchell outline how BGS is working on geoscience-led solutions for the global issue of sand mining.

Rare hornet moth colony found at BGS Keyworth
03/10/2023
A colony of these rare clearwing moths has recently been discovered on site at the BGS headquarters in Keyworth.

Nurturing early career scientists: 20 years of undergraduate industrial placements at BGS
28/09/2023
Michael Watts, BGS Head of Inorganic Chemistry, and previous placement students reflect on their experiences working at BGS’s Inorganic Geochemistry Facility over the past 20 years.

BGS laboratory spotlight: isotopes as recorders of climate and environmental change
06/09/2023
How measuring oxygen and carbon isotopes in tiny fossils improves our understanding of past climate.

In photos: a volcanic field trip
31/08/2023
Volcanologist Samantha Engwell visited the Cascades in the United States to learn more about the 1980 Mount St Helens volcanic eruption.

Understanding Nottinghamshire’s groundwater microbial ecosystems
24/08/2023
PhD student Archita Bhattacharyya is undertaking a project focused on exploring the ecosystem of microorganisms in groundwater of England.

‘Core blimey!’ A PhD fieldwork trip to India
22/08/2023
PhD student Hamish Duncalf-Youngson recently visited Manipur, India, to assess the effects of aquaculture, environmental change and pollution at this internationally important site.

Midlands Innovation TALENT placement at BGS
15/08/2023
Jodie Brown revisits her time at BGS’s Stable Isotope Facility as part of the Midlands Innovation TALENT project, which aims to increase the status of technicians.

bluedot 2023: the importance of geological outreach
10/08/2023
Staff members from various disciplines across BGS worked over the weekend to engage festivalgoers with BGS’s work, specifically critical raw materials.

Boreholes aren’t boring!
31/07/2023
Work experience student Patrick visited BGS to learn more about being a professional rock lover.