Home Catalogue search

eng

Refine your search:

Search in the Catalogues and Directories






	Sort by
Simple Search

Page: 1 2 3 4 5...292

Hits 1 – 20 of 5.836

1	A Bottleneck Auto-Encoder for F0 Transformations on Speech and Singing Voice
	Bous, Frederik; Roebel, Axel
	In: ISSN: 2078-2489 ; Information ; https://hal.archives-ouvertes.fr/hal-03599085 ; Information, MDPI, 2022, 13 (3), pp.102. ⟨10.3390/info13030102⟩ (2022)
	BASE
	Show details

2	Neural Vocoding for Singing and Speaking Voices with the Multi-Band Excited WaveNet
	Roebel, Axel; Bous, Frederik
	In: ISSN: 2078-2489 ; Information ; https://hal.archives-ouvertes.fr/hal-03599076 ; Information, MDPI, 2022, 13 (3), pp.103. ⟨10.3390/info13030103⟩ (2022)
	BASE
	Show details

3	Évaluation de la perception des sons de parole chez les populations pédiatriques : réflexion sur les épreuves existantes
	Meloni, Geneviève; Loevenbruck, Hélène; Vilain, Anne...
	In: ISSN: 0298-6477 ; EISSN: 2117-7155 ; Glossa ; https://hal.archives-ouvertes.fr/hal-03646757 ; Glossa, UNADREO - Union NAtionale pour le Développement de la Recherche en Orthophonie, 2022, 132, pp.1-27 ; https://www.glossa.fr/index.php/glossa/article/view/1043 (2022)
	BASE
	Show details

4	Learning and controlling the source-filter representation of speech with a variational autoencoder
	Sadok, Samir; Leglaive, Simon; Girin, Laurent...
	In: https://hal.archives-ouvertes.fr/hal-03650569 ; 2022 (2022)
	BASE
	Show details

5	Domestic Ubimus
	Keller, Damián; Simurra, Ivan; Messina, Marcello...
	In: EISSN: 2409-9708 ; EAI Endorsed Transactions on Creative Technologies ; https://hal-hprints.archives-ouvertes.fr/hprints-03602695 ; EAI Endorsed Transactions on Creative Technologies, EAI - European Alliance for Innovation, 2022, ⟨10.4108/eai.22-2-2022.173493⟩ (2022)
	BASE
	Show details

6	A comparative study of several parameterizations for speaker recognition ...
	Faundez-Zanuy, Marcos. - : arXiv, 2022
	BASE
	Show details

7	Speaker verification in mismatch training and testing conditions ...
	Faundez-Zanuy, Marcos; Slupinski, Adam. - : arXiv, 2022
	BASE
	Show details

8	Speech Segmentation Optimization using Segmented Bilingual Speech Corpus for End-to-end Speech Translation ...
	Fukuda, Ryo; Sudoh, Katsuhito; Nakamura, Satoshi. - : arXiv, 2022
	BASE
	Show details

9	A New Amharic Speech Emotion Dataset and Classification Benchmark ...
	Retta, Ephrem A.; Almekhlafi, Eiad; Sutcliffe, Richard. - : arXiv, 2022
	BASE
	Show details

10	The Norwegian Parliamentary Speech Corpus ...
	Solberg, Per Erik; Ortiz, Pablo. - : arXiv, 2022
	BASE
	Show details

11	Subspace-based Representation and Learning for Phonotactic Spoken Language Recognition ...
	Lee, Hung-Shin; Tsao, Yu; Jeng, Shyh-Kang; Wang, Hsin-Min. - : arXiv, 2022
	Abstract: Phonotactic constraints can be employed to distinguish languages by representing a speech utterance as a multinomial distribution or phone events. In the present study, we propose a new learning mechanism based on subspace-based representation, which can extract concealed phonotactic structures from utterances, for language verification and dialect/accent identification. The framework mainly involves two successive parts. The first part involves subspace construction. Specifically, it decodes each utterance into a sequence of vectors filled with phone-posteriors and transforms the vector sequence into a linear orthogonal subspace based on low-rank matrix factorization or dynamic linear modeling. The second part involves subspace learning based on kernel machines, such as support vector machines and the newly developed subspace-based neural networks (SNNs). The input layer of SNNs is specifically designed for the sample represented by subspaces. The topology ensures that the same output can be derived from ... : Published in IEEE/ACM Trans. Audio, Speech, Lang. Process., 2020, vol. 28, pp. 3065-3079 ...
	Keyword: Audio and Speech Processing eess.AS; Computation and Language cs.CL; FOS Computer and information sciences; FOS Electrical engineering, electronic engineering, information engineering; Machine Learning cs.LG; Sound cs.SD
	URL: https://dx.doi.org/10.48550/arxiv.2203.15576 https://arxiv.org/abs/2203.15576
	BASE
	Hide details

12	LPC Augment: An LPC-Based ASR Data Augmentation Algorithm for Low and Zero-Resource Children's Dialects ...
	Johnson, Alexander; Fan, Ruchao; Morris, Robin. - : arXiv, 2022
	BASE
	Show details

13	Automatic Dialect Density Estimation for African American English ...
	Johnson, Alexander; Everson, Kevin; Ravi, Vijay. - : arXiv, 2022
	BASE
	Show details

14	LINGUIST List Resources for Salish, Southern Puget Sound
	Damir Cavar, eLinguistics Foundation Board Member; Malgorzata E. Cavar, Director of Linguist List. - : The LINGUIST List (www.linguistlist.org), 2022
	BASE
	Show details

15	Variation in Spanish/s: Overview and New Perspectives
	Núñez-Méndez, Eva
	In: World Languages and Literatures Faculty Publications and Presentations (2022)
	BASE
	Show details

16	End-to-end contextual asr based on posterior distribution adaptation for hybrid ctc/attention system ...
	Zhang, Zhengyi; Zhou, Pan. - : arXiv, 2022
	BASE
	Show details

17	Towards Contextual Spelling Correction for Customization of End-to-end Speech Recognition Systems ...
	Wang, Xiaoqiang; Liu, Yanqing; Li, Jinyu. - : arXiv, 2022
	BASE
	Show details

18	SHAS: Approaching optimal Segmentation for End-to-End Speech Translation ...
	Tsiamas, Ioannis; Gállego, Gerard I.; Fonollosa, José A. R.. - : arXiv, 2022
	BASE
	Show details

19	Automatic Detection of Speech Sound Disorder in Child Speech Using Posterior-based Speaker Representations ...
	Ng, Si-Ioi; Ng, Cymie Wing-Yee; Wang, Jiarui. - : arXiv, 2022
	BASE
	Show details

20	Towards a Perceptual Model for Estimating the Quality of Visual Speech ...
	Aldeneh, Zakaria; Fedzechkina, Masha; Seto, Skyler. - : arXiv, 2022
	BASE
	Show details

Page: 1 2 3 4 5...292

© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern