This article describes the development of the digital infrastructure at a research data centre for audio-visual linguistic research data, the Hamburg Centre for Language Corpora (HZSK) at the University of Hamburg in Germany, over the past ten years. The typical resource hosted in the HZSK Repository, the core component of the infrastructure, is a collection of recordings with time-aligned transcripts and additional contextual data, a spoken language corpus. Since the centre has a thematic focus on multilingualism and linguistic diversity and provides its service to researchers within linguistics and other disciplines, the development of the infrastructure was driven by diverse usage scenarios and user needs on the one hand, and by the common technical requirements for certified service centres of the CLARIN infrastructure on the other. Beyond the technical details, the article also aims to be a contribution to the discussion on responsibilities and services within emerging digital research data infrastructures and the fundamental issues in sustainability of research software engineering, concluding that in order to truly cater to user needs across the research data lifecycle, we still need to bridge the gap between discipline-specific research methods in the process of digitalisation and generic digital research data management approaches.
Fields of Science and Technology classification (FOS)
05 social sciences, 0509 other social sciences, 050904 information & library sciences, 0602 languages and literature, 060201 languages & linguistics
free text keywords: Computer Science Applications, Media Technology, Communication, Business and International Management, Library and Information Sciences, research data, audio-visual data, linguistic data, domain-specific solutions, data quality, Certification, Service (systems architecture), Resource (project management), Spoken language, Process (engineering), Data quality, Contextual design, Multilingualism, Linguistics, Computer science, lcsh:Communication. Mass media, lcsh:P87-96, lcsh:Information resources (General), lcsh:ZA3040-5185