Publicação
Detection of voicing and place of articulation of fricatives with deep learning in a virtual speech and language therapy tutor
| dc.contributor.author | Anjos, Ivo | |
| dc.contributor.author | Eskenazi, Maxine | |
| dc.contributor.author | Marques, Nuno | |
| dc.contributor.author | Grilo, Margarida | |
| dc.contributor.author | Guimarães, Isabel | |
| dc.contributor.author | Magalhães, João | |
| dc.contributor.author | Cavaco, Sofia | |
| dc.contributor.institution | NOVALincs | |
| dc.contributor.institution | DI - Departamento de Informática | |
| dc.contributor.pbl | International Speech Communication Association | |
| dc.date.accessioned | 2026-01-16T17:14:54Z | |
| dc.date.available | 2026-01-16T17:14:54Z | |
| dc.date.issued | 2020 | |
| dc.description | ||
| dc.description.abstract | Children with fricative distortion errors have to learn how to correctly use the vocal folds, and which place of articulation to use in order to correctly produce the different fricatives. Here we propose a virtual tutor for fricatives distortion correction. This is a virtual tutor for speech and language therapy that helps children understand their fricative production errors and how to correctly use their speech organs. The virtual tutor uses log Mel filter banks and deep learning techniques with spectral-temporal convolutions of the data to classify the fricatives in children's speech by place of articulation and voicing. It achieves an accuracy of 90.40% for place of articulation and 90.93% for voicing with children's speech. Furthermore, this paper discusses a multidimensional advanced data analysis of the first layer convolutional kernel filters that validates the usefulness of performing the convolution on the log Mel filter bank. | en |
| dc.description.version | publishersversion | |
| dc.description.version | published | |
| dc.format.extent | 5 | |
| dc.format.extent | 674724 | |
| dc.identifier.doi | 10.21437/Interspeech.2020-2821 | |
| dc.identifier.issn | 2308-457X | |
| dc.identifier.other | PURE: 28751524 | |
| dc.identifier.other | PURE UUID: f6d5d230-f602-49e7-9371-6b0d0d2d707f | |
| dc.identifier.other | Scopus: 85093105936 | |
| dc.identifier.uri | http://hdl.handle.net/10362/199450 | |
| dc.identifier.url | https://www.scopus.com/pages/publications/85093105936 | |
| dc.language.iso | eng | |
| dc.peerreviewed | yes | |
| dc.relation | info:eu-repo/grantAgreement/FCT/5665-PICT/CMUP-ERI%2FTIC%2F0033%2F2014/PT | |
| dc.relation | info:eu-repo/grantAgreement/FCT/6817 - DCRRNI ID/UID%2FCEC%2F04516%2F2019/PT | |
| dc.subject | Convolutional neural networks | |
| dc.subject | Fricatives | |
| dc.subject | Speech and language therapy | |
| dc.subject | Language and Linguistics | |
| dc.subject | Human-Computer Interaction | |
| dc.subject | Signal Processing | |
| dc.subject | Software | |
| dc.subject | Modelling and Simulation | |
| dc.title | Detection of voicing and place of articulation of fricatives with deep learning in a virtual speech and language therapy tutor | en |
| dc.type | journal article | |
| degois.publication.firstPage | 3156 | |
| degois.publication.lastPage | 3160 | |
| degois.publication.title | Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH | |
| degois.publication.volume | 2020-October | |
| dspace.entity.type | Publication | |
| rcaap.rights | openAccess |
Ficheiros
Principais
1 - 1 de 1
