DOI 10.14209/its.2002.867
Guilherme de Oliveira Pinto, Filipe Leandro de F. Barbosa, Fernando Gil V. Resende Jr.
"This paper presents a Brazilian Portuguese TTS based on HMMs, which uses mel-cepstral coefficients as parameters of speech. We implemented an algorithm which performs a phoneme-based transcription to the Portuguese spoken in Rio de Janeiro. For a given text to be synthesized, after the phoneme transcription, static features are extracted from a sentence...
DOI 10.14209/its.2002.872
Rubem Dutra Ribeiro Fagundes, Juarez Sagebin Corrêa, Pierre Dumouchel
"The main goal of this work is to describe a new model for a large vocabulary continuous speech recognition system using a phonetic-phonological approach. This work proposes a statistical phonetic structure, applied at the phonetic-phonological level, to improve the speech recognition performance in systems with phonetic-phonological modeling. It is...
DOI 10.14209/its.2002.876
H. Sayoud, S. Ouamour, M. Boudraa
"Speaker indexing can broadly be divided into two problems: Locating the points of speaker change (Segmentation) and Identifying the speaker in each segment (Labeling). An important obstacle, in the speaker tracking, is the corruption of the speech signal during its recording or in a telephonic conversation. In this paper, we are interested in the...
DOI 10.14209/its.2002.882
Fernando S. Pacheco, Rui Seara
"This paper proposes a method to prosodic speech modification based on residual-excited linear predictive coding (RELP) applied to speech synthesis. In this way, pitch and time scale modifications are carried out requiring a very low computational complexity. Formal subjective test comparing the proposed approach with TD-PSOLA, under conditions of...
DOI 10.14209/its.2002.887
Carlos Alberto Ynoguti, Fábio Violaro
"The aim of this work is to enlighten some implementational issues on two widely used techniques for reduction in search space for large vocabulary speech recognition systems: tree organized lexicons and Viterbi Beam Search. We show how we can exploit some symmetries of the search space to use these techniques together and achieve a great reduction in...
DOI 10.14209/its.2002.892
Helder C. Bertan, Luís G. P. Meloni
"This paper presents a new method for Linear Spectrum Pair transcoding between G.729A and GSM AMR codecs. The transcoding in the bitstream domain gives better quality and lower complexity than the conventional method in the speech domain. Several simulation results are presented using Perceptual Quality Speech Measure showing a gain of about 0.45 in this...
DOI 10.14209/its.2002.897
Liselene de Abreu Borges, Miguel Arjona Ramírez, Rubem Dutra Ribeiro Fagundes
"This paper discusses speech recognition systems (SRS) using speaker adaptation techniques. The most recent speech recognition systems use Hidden Markov Models (HMM). For such systems, the eigenvoices speaker adaptation technique presents the best performance among other techniques usually suggested by researchers. This performance is due mainly to the...
DOI 10.14209/its.2002.902
Ivan R. S. Casella, Elvino S. Sousa, Paul Jean E. Jeszensky
"In this paper, we investigate the performance of a semi-blind spatial-temporal beamforming receiver for an asynchronous high data rate direct sequence wideband code division multiple access (DS-WCDMA) system in a microcellular environment. The presented receiver uses subspace channel identification to perform joint channel equalization, multipath energy...
DOI 10.14209/its.2002.908
Ivan R. S. Casella, Elvino S. Sousa, Paul Jean E. Jeszensky
"In this paper, a new training-based spatial-temporal receiver is proposed for an asynchronous wideband direct sequence code division multiple access (DS-CDMA) system. The presented receiver employs a hierarchical recursive least squares (HRLS) algorithm to reduce the computational complexity."
DOI 10.14209/its.2002.913
Gustavo Fraidenraich, Renato Baldini F., Celso de Almeida
"This paper presents simplified expressions for the mean bit error probability using random spreading sequences on AWGN and multipath Rayleigh fading channels using the multiuser MMSE and decorrelating detectors. It is assumed the multi processing gain schemes of multirate."
MultiuserCDMAmultirateMMSEdecorrelating