Eyra - Speech Data Acquisition System for Many Languages

dc.contributor.authorPetursson, Matthias
dc.contributor.authorKlüpfel, Simon
dc.contributor.authorGudnason, Jon
dc.contributor.departmentDepartment of Engineering
dc.date.accessioned2026-09-09T13:26:02Z
dc.date.available2026-09-09T13:26:02Z
dc.date.issued2016
dc.descriptionPublisher Copyright: © 2016 The Authors.en
dc.description.abstractSpeech data acquisition is particularly important for under-resourced languages. The data gathering is the most labour-intensive part of developing speech technologies such as automatic speech recognizers and synthesizers. It is therefore important to facilitate this process with as much automation and labour-cutting tools as possible. This paper describes a new open-source system called Eyra which enables distributed speech data collecting through a variety of devices. It addresses internet connectivity issues by allowing the data collectors to run the back-end server off a local laptop, thereby facilitating automatic quality control and less labour-intensive data uploading and compiling. It can also be used in a crowd-sourcing set-up where volunteers can donate voice samples through a desktop web-browser interface. An initial test shows that the system works well in an offline mode using smart-phones for data collection.en
dc.description.versionPeer revieweden
dc.format.extent8
dc.format.extent243365
dc.format.extent53-60
dc.identifier.citationPetursson, M, Klüpfel, S & Gudnason, J 2016, 'Eyra - Speech Data Acquisition System for Many Languages', Procedia Computer Science, vol. 81, pp. 53-60. https://doi.org/10.1016/j.procs.2016.04.029en
dc.identifier.doi10.1016/j.procs.2016.04.029
dc.identifier.issn1877-0509
dc.identifier.other250778207
dc.identifier.other71464031-4443-4b1d-8453-ddeb7d7bb40a
dc.identifier.other84976381580
dc.identifier.urihttps://hdl.handle.net/20.500.11815/8224
dc.language.isoen
dc.relation.ispartofseriesProcedia Computer Science; 81()en
dc.relation.urlhttps://www.scopus.com/pages/publications/84976381580en
dc.rightsinfo:eu-repo/semantics/openAccessen
dc.subjectautomatic speech recognitionen
dc.subjectinternationalizationen
dc.subjectspeech resource collectionen
dc.subjectunder-resourced languagesen
dc.subjectGeneral Computer Scienceen
dc.titleEyra - Speech Data Acquisition System for Many Languagesen
dc.type/dk/atira/pure/researchoutput/researchoutputtypes/contributiontojournal/conferencearticleen

Skrár

Original bundle

Niðurstöður 1 - 1 af 1
Nafn:
1-s2.0-S1877050916300436-main.pdf
Stærð:
237.66 KB
Snið:
Adobe Portable Document Format