Eyra - Speech Data Acquisition System for Many Languages
| dc.contributor.author | Petursson, Matthias | |
| dc.contributor.author | Klüpfel, Simon | |
| dc.contributor.author | Gudnason, Jon | |
| dc.contributor.department | Department of Engineering | |
| dc.date.accessioned | 2026-09-09T13:26:02Z | |
| dc.date.available | 2026-09-09T13:26:02Z | |
| dc.date.issued | 2016 | |
| dc.description | Publisher Copyright: © 2016 The Authors. | en |
| dc.description.abstract | Speech data acquisition is particularly important for under-resourced languages. The data gathering is the most labour-intensive part of developing speech technologies such as automatic speech recognizers and synthesizers. It is therefore important to facilitate this process with as much automation and labour-cutting tools as possible. This paper describes a new open-source system called Eyra which enables distributed speech data collecting through a variety of devices. It addresses internet connectivity issues by allowing the data collectors to run the back-end server off a local laptop, thereby facilitating automatic quality control and less labour-intensive data uploading and compiling. It can also be used in a crowd-sourcing set-up where volunteers can donate voice samples through a desktop web-browser interface. An initial test shows that the system works well in an offline mode using smart-phones for data collection. | en |
| dc.description.version | Peer reviewed | en |
| dc.format.extent | 8 | |
| dc.format.extent | 243365 | |
| dc.format.extent | 53-60 | |
| dc.identifier.citation | Petursson, M, Klüpfel, S & Gudnason, J 2016, 'Eyra - Speech Data Acquisition System for Many Languages', Procedia Computer Science, vol. 81, pp. 53-60. https://doi.org/10.1016/j.procs.2016.04.029 | en |
| dc.identifier.doi | 10.1016/j.procs.2016.04.029 | |
| dc.identifier.issn | 1877-0509 | |
| dc.identifier.other | 250778207 | |
| dc.identifier.other | 71464031-4443-4b1d-8453-ddeb7d7bb40a | |
| dc.identifier.other | 84976381580 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.11815/8224 | |
| dc.language.iso | en | |
| dc.relation.ispartofseries | Procedia Computer Science; 81() | en |
| dc.relation.url | https://www.scopus.com/pages/publications/84976381580 | en |
| dc.rights | info:eu-repo/semantics/openAccess | en |
| dc.subject | automatic speech recognition | en |
| dc.subject | internationalization | en |
| dc.subject | speech resource collection | en |
| dc.subject | under-resourced languages | en |
| dc.subject | General Computer Science | en |
| dc.title | Eyra - Speech Data Acquisition System for Many Languages | en |
| dc.type | /dk/atira/pure/researchoutput/researchoutputtypes/contributiontojournal/conferencearticle | en |
Skrár
Original bundle
1 - 1 af 1
- Nafn:
- 1-s2.0-S1877050916300436-main.pdf
- Stærð:
- 237.66 KB
- Snið:
- Adobe Portable Document Format