The Experts below are selected from a list of 15792 Experts worldwide ranked by ideXlab platform
Hiroya Fujisaki - One of the best experts on this subject based on the ideXlab platform.
-
prosodic comparison of declarative and interrogative utterances in standard colloquial bangla
2011 International Conference on Speech Database and Assessments (Oriental COCOSDA), 2011Co-Authors: Anal Haque Warsi, T K Basu, Keikichi Hirose, Hiroya FujisakiAbstract:This paper presents a comparative study of prosodic features of utterances of two types of Bangla sentences: the declarative type and the interrogative (‘yes-no’ question) type whose textual contents are identical except for the punctuation marks in Bangla. The study is based on the analysis of 44 utterances each of the declarative type and the interrogative type spoken by native speakers of Standard Colloquial Bangla (SCB). The results of F 0 contour analysis show that declarative utterances have a gradually falling F 0 contour with a terminal fall whereas interrogative utterances have a rising contour and a terminal rise. It is also observed that interrogative utterances have a positive swing of F 0 larger than that of declarative utterances, and the maximum of F 0 occurs within the prosodic word containing the interrogative information. It is shown that these differences can be represented in terms of differences in parameters of the F 0 contour Command-Response model. The validity of the analysis was confirmed by perceptual experiments using synthetic stimuli, demonstrating the feasibility of the Command-Response model in speech synthesis of Bangla.
-
INTERSPEECH - Analysis of Voice Fundamental Frequency Contours of Continuing and Terminating Phrases of Four Swiss German Dialects
2009Co-Authors: Adrian Leemann, Keikichi Hirose, Hiroya FujisakiAbstract:In the present study, the F0 contours of continuing and terminating prosodic phrases of 4 Swiss German dialects are analyzed by means of the Command-Response model. In every model parameter, the two prosodic phrase types show significant differences: continuing prosodic phrases indicate higher phrase Command magnitude and shorter durations. Locally, they demonstrate more distinct accent Command amplitudes as well as durations. In addition, continuing prosodic phrases have later rises relative to segment onset than terminating prosodic phrases. In the same context, fine phonetic differences between the dialects are highlighted.
-
Temporal organization of prosodic and segmental features in spoken Japanese
The Journal of the Acoustical Society of America, 2008Co-Authors: Hiroya Fujisaki, Sumio OhnoAbstract:It is apparent that prosodic and segmental features of speech must be temporally coordinated in order to produce a consistent and meaningful message. The precise mechanism for the coordination, however, has not been clear. The present study looks into this problem in the case of word accent in spoken Japanese. The speech material consisted of utterances of Japanese words that were identical in the word accent type, in the number of morae, in vowel constituents, but were different in consonantal constituents (including a 'null' consonant) at a certain intervocalic position. As for the prosodic features, the fundamental frequency contours were analyzed using the Command‐Response model, and the onset and the offset of the extracted accent Command were used as indices of prosodic timing. As for the segmental features, the formant frequency trajectories of the vowels were analyzed using another Command‐Response model, which allowed extraction of the onset of the articulatory Command for a vowel nucleus as an i...
-
Analysis of Tones in Cantonese Speech Based on the Command-Response Model
Phonetica, 2007Co-Authors: Keikichi Hirose, Hiroya FujisakiAbstract:As one of the major Chinese dialects, Cantonese has a tone system consisting of nine lexical tones and three additional changed tones, which is considerably more complex than that of Mandarin. The m
-
The Command‐Response model for F0 contours and its application to phonetics and phonology of tone
The Journal of the Acoustical Society of America, 2006Co-Authors: Hiroya FujisakiAbstract:The Command‐Response model for the F0 contours, first proposed for nontone languages such an Japanese and English, has been extended by the author and his co‐workers to cover F0 contour of tone languages and has been proved to be applicable to several dialects of Chinese including Mandarin, Cantonese, and Shanghainese, as well as other tone languages including Thai and Vietnamese. Namely, it allows one to approximate F0 contours of these languages/dialects with a very high accuracy using positive and negative tone Commands. This paper first presents some results of analysis of F0 contours of these languages/dialects to demonstrate that the model can capture the acoustic‐phonetic characteristics of their tones quite well. It then shows that coarse qualitative distinctions in the pattern of tone Commands serve to distinguish tones within a language/dialect, as well as to distinguish tone systems of different languages/dialects. Thus the model proves to be a useful tool, not only in the studies of phonetic f...
Sumio Ohno - One of the best experts on this subject based on the ideXlab platform.
-
Temporal organization of prosodic and segmental features in spoken Japanese
The Journal of the Acoustical Society of America, 2008Co-Authors: Hiroya Fujisaki, Sumio OhnoAbstract:It is apparent that prosodic and segmental features of speech must be temporally coordinated in order to produce a consistent and meaningful message. The precise mechanism for the coordination, however, has not been clear. The present study looks into this problem in the case of word accent in spoken Japanese. The speech material consisted of utterances of Japanese words that were identical in the word accent type, in the number of morae, in vowel constituents, but were different in consonantal constituents (including a 'null' consonant) at a certain intervocalic position. As for the prosodic features, the fundamental frequency contours were analyzed using the Command‐Response model, and the onset and the offset of the extracted accent Command were used as indices of prosodic timing. As for the segmental features, the formant frequency trajectories of the vowels were analyzed using another Command‐Response model, which allowed extraction of the onset of the articulatory Command for a vowel nucleus as an i...
-
analysis and synthesis of fundamental frequency contours of standard chinese using the Command Response model
Speech Communication, 2005Co-Authors: Hiroya Fujisaki, Changfu Wang, Sumio OhnoAbstract:While the tonal characteristics of Chinese syllables have been qualitatively described in traditional phonetics, quantitative analysis requires a mathematical model. This paper presents such a model for the fundamental frequency contours of Standard Chinese, based on an extension of a model that has already been proved to be applicable to non-tone languages including Japanese, English, and others. The model allows one to interpret a given fundamental frequency contour in terms of tone Commands and phrase Commands, and to analyze various tonal phenomena in quantitative terms. The paper then describes the results of analysis of fundamental frequency contours of a number of utterances, revealing systematic relationships between the timing of the tone Commands and the final of each syllable. The results are used to derive constraints for tone and phrase Command generation in speech synthesis. The validity of the rules is confirmed by evaluating the naturalness of prosody of synthetic speech. The validity of introducing these constraints in speech synthesis of Standard Chinese is confirmed by perceptual tests on naturalness of prosody as well as on intelligibility of tones, using speech synthesized with and without these constraints.
-
Analysis and synthesis of fundamental frequency contours of Standard Chinese using the Command–Response model
Speech Communication, 2005Co-Authors: Hiroya Fujisaki, Changfu Wang, Sumio OhnoAbstract:While the tonal characteristics of Chinese syllables have been qualitatively described in traditional phonetics, quantitative analysis requires a mathematical model. This paper presents such a model for the fundamental frequency contours of Standard Chinese, based on an extension of a model that has already been proved to be applicable to non-tone languages including Japanese, English, and others. The model allows one to interpret a given fundamental frequency contour in terms of tone Commands and phrase Commands, and to analyze various tonal phenomena in quantitative terms. The paper then describes the results of analysis of fundamental frequency contours of a number of utterances, revealing systematic relationships between the timing of the tone Commands and the final of each syllable. The results are used to derive constraints for tone and phrase Command generation in speech synthesis. The validity of the rules is confirmed by evaluating the naturalness of prosody of synthetic speech. The validity of introducing these constraints in speech synthesis of Standard Chinese is confirmed by perceptual tests on naturalness of prosody as well as on intelligibility of tones, using speech synthesized with and without these constraints.
-
Influences of various factors upon parameters of the Command–Response model for fundamental frequency contour generation
The Journal of the Acoustical Society of America, 1999Co-Authors: Sumio Ohno, Yoshikazu Hara, Hiroya FujisakiAbstract:The Command–Response model by Fujisaki and his co‐workers formulates the generation process of the fundamental frequency contour (henceforth F0 contour) in terms of a set of input Commands carrying linguistic and paralinguistic information and the mechanisms that respond to these Commands. The parameters of the mechanisms are time constant of the phrase control mechanism (α), that of the accent control mechanism (β), and the baseline frequency (Fb). The present study investigates the influences of speech rate, speaking style, and individual difference on these parameters as well as their variability. The speech material consisted of recordings of a short story by six native speakers of Japanese at three speech rates (slow, normal, and fast) and in two speaking styles (reading and conversational). The parameters were extracted by the method of analysis‐by‐synthesis, and the results were analyzed statistically. The analysis indicated that utterance‐to‐utterance variations are quite small in all three parameters for a given rate, style, and speaker. Among the three parameters, only α showed a small but systematic tendency to increase with the speech rate, while differences in speaking style did not affect these parameters. Individual differences were quite small in α and β, while Fb varied from speaker to speaker.
-
Application of the Command–Response model to the analysis, interpretation, and synthesis of fundamental frequency contours of speech of various languages
The Journal of the Acoustical Society of America, 1999Co-Authors: Hiroya Fujisaki, Sumio OhnoAbstract:A Command–Response model has been presented by Fujisaki and his co‐workers initially for the process of generation of the fundamental frequency contour (henceforth F0 contour) of the common Japanese. It consists of a set of input Commands carrying linguistic and paralinguistic information, and the mechanisms that respond to these Commands to generate both phrase and accent components, which, together with a baseline value, constitute the actual F0 contour. Subsequent works have shown that the model is applicable also to F0 contours of some other dialects of Japanese as well as of other languages including Chinese, English, German, Greek, Spanish, and Swedish. It has been argued, however, that the model may not be able to generate certain contour types that are not found in Japanese but are commonly used in other languages [e.g., P. Tayler, Speech Commun. 15, 183]. This paper shows how these contour types can actually be generated by the same mechanisms, with certain language‐specific timing, polarity and ...
Keikichi Hirose - One of the best experts on this subject based on the ideXlab platform.
-
prosodic comparison of declarative and interrogative utterances in standard colloquial bangla
2011 International Conference on Speech Database and Assessments (Oriental COCOSDA), 2011Co-Authors: Anal Haque Warsi, T K Basu, Keikichi Hirose, Hiroya FujisakiAbstract:This paper presents a comparative study of prosodic features of utterances of two types of Bangla sentences: the declarative type and the interrogative (‘yes-no’ question) type whose textual contents are identical except for the punctuation marks in Bangla. The study is based on the analysis of 44 utterances each of the declarative type and the interrogative type spoken by native speakers of Standard Colloquial Bangla (SCB). The results of F 0 contour analysis show that declarative utterances have a gradually falling F 0 contour with a terminal fall whereas interrogative utterances have a rising contour and a terminal rise. It is also observed that interrogative utterances have a positive swing of F 0 larger than that of declarative utterances, and the maximum of F 0 occurs within the prosodic word containing the interrogative information. It is shown that these differences can be represented in terms of differences in parameters of the F 0 contour Command-Response model. The validity of the analysis was confirmed by perceptual experiments using synthetic stimuli, demonstrating the feasibility of the Command-Response model in speech synthesis of Bangla.
-
INTERSPEECH - Analysis of Voice Fundamental Frequency Contours of Continuing and Terminating Phrases of Four Swiss German Dialects
2009Co-Authors: Adrian Leemann, Keikichi Hirose, Hiroya FujisakiAbstract:In the present study, the F0 contours of continuing and terminating prosodic phrases of 4 Swiss German dialects are analyzed by means of the Command-Response model. In every model parameter, the two prosodic phrase types show significant differences: continuing prosodic phrases indicate higher phrase Command magnitude and shorter durations. Locally, they demonstrate more distinct accent Command amplitudes as well as durations. In addition, continuing prosodic phrases have later rises relative to segment onset than terminating prosodic phrases. In the same context, fine phonetic differences between the dialects are highlighted.
-
Analysis of Tones in Cantonese Speech Based on the Command-Response Model
Phonetica, 2007Co-Authors: Keikichi Hirose, Hiroya FujisakiAbstract:As one of the major Chinese dialects, Cantonese has a tone system consisting of nine lexical tones and three additional changed tones, which is considerably more complex than that of Mandarin. The m
-
INTERSPEECH - F 0 models show Chinese speakers of Japanese insert intonational boundaries and drop pitch.
2007Co-Authors: Hiroko Hirano, Keikichi Hirose, Goh Kawai, Nobuaki MinematsuAbstract:We used a Command-Response additive F0 model to analyze F0 patterns of Japanese spoken by native speakers of Mandarin Chinese. Compared to native speakers of Japanese, we found that Chinese speakers exhibit the following characteristics: (a) higher pitch, (b) more phrases, (c) bunsetsu decomposition, and (d) utterance-final plunging. These characteristics physically manifest themselves as: (a) higher baseline F0, (b) more phrase Commands, (c) more accent Commands, and (d) negative Commands. These characteristics may be subjectively perceived as: (a) tinnier speech (possible L1 marker but does not degrade communication), (b) disjoint phrases (requires mental consolidation), (c) choppy prosodic words (requires reconstruction), and (d) abrupt utterance termination (possibly misconstrued as emphatic or rude). We believe these difficulties arose from tonal and syllable-timed interference, which can be overcome by prosodic control and planning. Index Terms: F0 contour, Command-Response model, L2 learning, Japanese, accent, phrase.
-
A General Approach for Automatic Extraction of Tone Commands in the Command-Response Model for Tone Languages
2006Co-Authors: Keikichi Hirose, Hiroya FujisakiAbstract:Although the Command-Response model for the process of F0 contour generation has been successfully applied to many languages, the inverse problem, viz., automatic derivation of the model parameters from an observed F0 contour, is more challenging, especially for tone languages which have tone Commands of both polarities. Since the polarities of tone Commands cannot be inferred directly from the F0 contour itself, the information on tone identity and timing need to be incorporated. The current study gives a general approach for the first-order estimation of tone Command parameters for tone languages, taking Mandarin and Cantonese as two examples. After a rule-based recognition of the tone Command patterns within each syllable, the timing and amplitude of tone Commands will be deduced. The experiments show that the method gives good results of analysis for both the two dialects.
Steffen Müller - One of the best experts on this subject based on the ideXlab platform.
-
ECC - Lateral trajectory tracking control for autonomous vehicles
2014 European Control Conference (ECC), 2014Co-Authors: Christian Rathgeber, Franz Winkler, Dirk Odenthal, Steffen MüllerAbstract:In this contribution a structure for high level lateral vehicle tracking control is presented. It is based on the two degrees of freedom structure that allows to separately define the Command Response and disturbance attenuation. The application of the disturbance observer guarantees robust compensation of the disturbances. An advantage of the presented structure is its robustness against variable vehicle parameters. Only a considerably reduced model is necessary which significantly simplifies the application process. Moreover the presented approach is characterized by its extensibility and its modularity.
Christian Rathgeber - One of the best experts on this subject based on the ideXlab platform.
-
ECC - Lateral trajectory tracking control for autonomous vehicles
2014 European Control Conference (ECC), 2014Co-Authors: Christian Rathgeber, Franz Winkler, Dirk Odenthal, Steffen MüllerAbstract:In this contribution a structure for high level lateral vehicle tracking control is presented. It is based on the two degrees of freedom structure that allows to separately define the Command Response and disturbance attenuation. The application of the disturbance observer guarantees robust compensation of the disturbances. An advantage of the presented structure is its robustness against variable vehicle parameters. Only a considerably reduced model is necessary which significantly simplifies the application process. Moreover the presented approach is characterized by its extensibility and its modularity.