Utilisation de la linguistique en reconnaissance de la parole : un état de l'art
Stéphane Huet, Pascale Sébillot, Guillaume Gravier
Abstract
To transcribe speech, automatic speech recognition systems use statistical methods, particularly hidden Markov model and N-gram models. Although these techniques perform well and lead to efficient systems, they approach their maximum possibilities. It seems thus necessary, in order to outperform current results, to use additional information, especially bound to language. However, introducing such knowledge must be realized taking into account specificities of spoken language (hesitations for example) and being robust to possible misrecognized words. This document presents a state of the art of these researches, evaluating the impact of the insertion of linguistic information on the quality of the transcription.
Create a lesson
Related papers
Do User-Authored Permission Policies Improve Protection Against AI Agent Overreach?
Ting Yan
Beyond Harassment: Exploring the Harm Experienced by People with Disabilities in Social Virtual Reality
Xinran Adeline Li, Kexin Zhang, Yuhang Zhao et al.
A Point-of-Prescription Safety-Check System for Adverse Drug Reactions in Rural Bangladeshi Hospitals: A Feasibility Study
Shahir Abdullah
Surrounded by Friends: Design and Evaluation of Immersive Layouts of Egocentric Network for Visual Analytics
Kentaro Takahira, Takanori Fujiwara, Wong Kam-Kwai et al.
Exploring Normativity in Stable Diffusion: Insights for XAI in the Arts
Michelle Dutoit, Baptiste Caramiaux
Dynamic Tree Colors: Adaptive Discriminable Hierarchies with Minimum Instability
Tobias Mertz, Steven Lamarr Reynolds, Jörn Kohlhammer