Integration of Phonotactic Features for Language Identification on Code-Switched Speech

نویسندگان

چکیده

In this paper, phoneme sequences are used as language information to perform code-switched identification (LID). With the one-pass recognition system, spoken sounds converted into phonetically arranged of sounds. The acoustic models robust enough handle multiple languages when emulating hidden Markov (HMMs). To determine similarity among our target languages, we reported two methods mapping. Statistical phoneme-based bigram (LM) integrated speech decoding eliminate possible phone mismatches. supervised support vector machine (SVM) is learn recognize phonetic mixed-language based on recognized sequences. As back-end decision taken by an SVM, likelihood scores segments with monolingual occurrence classify identity. corpus was tested Sepedi and English that often mixed. Our system evaluated measuring both ASR performance LID separately. systems have obtained a promising accuracy data-driven merging approach modelled using 16 Gaussian mixtures per state. respectively, proposed achieved acceptable accuracy.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

the effects of speech rate,prosodic features, and blurred speech on iranian efl learners listening comprehension

کلید واژه ها به زبان انگلیسی: effect of speech rate on listening comprehension, blurred speech,segmental and suprasegmental features,authentic speech,intelligibility, discrimination, omission, assimilation چکیده: سرعت مطالب شنیداری در کلام پیوسته بطور کلی همواره کابوسی بوده برای یادگیرنده های زبان دوم و بالاخص برای شنوندگان ایرانی. علی رغم عقل سلیم که کلام با سرعت کندتری فعالیتهای درک مطلب شن...

15 صفحه اول

Phonotactic Language Identification for Singing

In the past decades, many successful approaches for language identification have been published. However, almost none of these approaches were developed with singing in mind. Singing has a lot of characteristics that differ from speech, such as a wider variance of fundamental frequencies and phoneme durations, vibrato, pronunciation differences, and different semantic content. We present a new ...

متن کامل

A Neural Model for Language Identification in Code-Switched Tweets

Language identification systems suffer when working with short texts or in domains with unconventional spelling, such as Twitter or other social media. These challenges are explored in a shared task for Language Identification in Code-Switched Data (LICS 2016). We apply a hierarchical neural model to this task, learning character and contextualized word-level representations to make word-level ...

متن کامل

Selecting phonotactic features for language recognition

This paper studies feature selection in phonotactic language recognition. The phonotactic feature is presented by n-gram statistics derived from one or more phone recognizers in the form of high dimensional feature vectors. Two feature selection strategies are proposed to select the n-gram statistics for reducing the dimension of feature vectors, so that higher order n-gram features can be adop...

متن کامل

Features for factored language models for code-Switching speech

This paper presents investigations of features which can be used to predict Code-Switching speech. For this task, factored language models are applied and implemented into a state-of-the-art decoder. Different possible factors, such as words, part-of-speech tags, Brown word clusters, open class words and open class word clusters are explored. We find that Brown word clusters, part-of-speech tag...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

ژورنال

عنوان ژورنال: International journal on natural language computing

سال: 2022

ISSN: ['2278-1307', '2319-4111']

DOI: https://doi.org/10.5121/ijnlc.2022.11102