JPS60257500A - Voice synthesization - Google Patents
Voice synthesizationInfo
- Publication number
- JPS60257500A JPS60257500A JP59113325A JP11332584A JPS60257500A JP S60257500 A JPS60257500 A JP S60257500A JP 59113325 A JP59113325 A JP 59113325A JP 11332584 A JP11332584 A JP 11332584A JP S60257500 A JPS60257500 A JP S60257500A
- Authority
- JP
- Japan
- Prior art keywords
- speech
- natural
- syllable
- timing point
- listening timing
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
Abstract
(57)【要約】本公報は電子出願前の出願データであるた
め要約のデータは記録されません。(57) [Summary] This bulletin contains application data before electronic filing, so abstract data is not recorded.
Description
【発明の詳細な説明】
産業上の利用分野
本発明は、予めファイリングしである音節単位の音声合
成パラメータを結合して、文章や文節等の音声を生成す
る音声合成方法に関するものである。DETAILED DESCRIPTION OF THE INVENTION Field of the Invention The present invention relates to a speech synthesis method for generating speech such as sentences and phrases by combining speech synthesis parameters in syllable units that have been filed in advance.
従来例の構成とその問題点
従来のこの種の音声合成力法には規制合成方法、あるい
は韻律のみを自然音声から抽出したものを用いる折衷方
法がある。゛規制合成では、文章中の文節継続時間長決
定規制に基づいて、予めファイリングしである単音節毎
の音声パラメータを接続し、これに音盛数等から定寸る
ピッチ情報を付与する等して合成パラメータを作成し、
これを用いて音声を合成する。ところが文章中での音節
の継続時間長や、ピッチ情報を規則によって付与するだ
め、合成音声の聞き易さ、自然性に難点があっ/こ。こ
のため、文章中での音節継続時間やピッチ情報を自然音
声から抽出したものを用いるなどした折衷方法が試みら
れている。しかしながら、いずれの方法でも、日本語特
有のリズムをコントロールするという考え方がとり入れ
られていないため、合成音声の聞き易さや自然性を改善
するには限界があった。Structure of conventional examples and their problems Conventional speech synthesis methods of this type include a restricted synthesis method and a compromise method that uses only prosody extracted from natural speech.゛Regulated synthesis involves connecting pre-filed audio parameters for each monosyllable based on the rules for determining the duration of a clause in a sentence, and adding pitch information determined from the number of tones, etc. Create synthesis parameters using
This is used to synthesize speech. However, there are problems with the ease of listening and naturalness of synthesized speech because it is not possible to assign pitch information and the duration of syllables in a sentence according to rules. For this reason, compromise methods have been attempted, such as using syllable duration and pitch information in sentences extracted from natural speech. However, none of these methods incorporates the idea of controlling the rhythm unique to Japanese, so there is a limit to the ability to improve the audibility and naturalness of synthesized speech.
発明の目的
本発明は、上記従来例の欠点を除去するものであり、音
節単位の音声合成パラメータを結合して自然なリズム感
を有する文節や文章等の音声を合成することを目的とし
ている。OBJECTS OF THE INVENTION The present invention eliminates the drawbacks of the above-mentioned conventional examples, and aims to synthesize speech such as phrases and sentences having a natural sense of rhythm by combining speech synthesis parameters in units of syllables.
発明の構成
本発明は、上記目的を達成するだめに、自然音声より抽
出した受聴タイミング点位置に1音節ファイルの受聴タ
イミング点を一致させるように音節結合を行うものであ
り、j″6節r1′l−位の音声合成パラメータを結合
して生成される文空や文節等の音声に自然なリズムを与
えることができるという効果を得るものである。Structure of the Invention In order to achieve the above object, the present invention performs syllable combination so that the listening timing point of a one-syllable file coincides with the listening timing point position extracted from natural speech. The effect is that a natural rhythm can be imparted to speech such as sentence spaces and phrases generated by combining the voice synthesis parameters of 'l-' order.
実施例の説明
以下に本発明の一実施例を説明する。第1図は[くもり
のちあめで ・−1という人気予報(自然音声)の各音
節毎の受聴タイミング点を抽出した例である。自然音声
の受聴クィミンダ点抽出手順は第2図に示す通りである
。この方法を用いるととKより自然音声の受聴タイミン
グ点は、容易に且つ正確に抽出できる。この手順を以下
に説明する。DESCRIPTION OF EMBODIMENTS An embodiment of the present invention will be described below. Figure 1 is an example of extracting the listening timing points for each syllable of the popular forecast (natural voice) [Kumori no Chiamede -1]. The procedure for extracting the natural speech listening comprehension point is as shown in FIG. Using this method, the listening timing point of natural speech can be easily and accurately extracted. This procedure will be explained below.
(1)受聴タイミング点を抽出すべき自然音声をA/J
)変換し、計算機の音声・ノ1.イル内に取り込む。(1) A/J the natural voice from which the listening timing point should be extracted.
) Convert the computer's audio/no 1. into the file.
(2)音声ファイル内のデータを:FR,み出し、CI
t、 T表示等を行って自然音声中からおおまかに音節
区間を切り出す。(2) Data in the audio file: FR, Extrusion, CI
t, T display, etc. to roughly cut out syllable sections from natural speech.
(3) 自然音声中で大まかに決めた音節区間を例えば
10m5の分析フレームに分割してLPO分析す(4)
あらかじめ、音節パラメータファイル内に格納しであ
る音節毎の音声パラメータの中がら受聴タイミング点乙
し−ムのパラメータを読み出し、分析フレーム毎にLP
C分析により得られた音声パラメータとの類似度を比較
する。(3) Divide roughly determined syllable intervals in natural speech into analysis frames of, for example, 10m5 and perform LPO analysis (4)
In advance, the parameters of the listening timing point are read out from among the speech parameters for each syllable stored in the syllable parameter file, and the LP is calculated for each analysis frame.
Compare the degree of similarity with the voice parameters obtained by C analysis.
(5) 類似度が最大となる分析フレームの中央ポイン
トを受聴タイミング点とする。(5) Set the central point of the analysis frame where the degree of similarity is maximum as the listening timing point.
第3図は、上記方法により決定した第3図の自然γ?声
の受聴タイミング点に、各音節毎の受聴タイミング点を
一致させて音声合成パラメータを結合して得た音声合成
波形と、その受聴タイミング(マで表示)である。第4
図は、従来の音節結合により得た音声合成波形と、それ
から抽出した受聴タイミング点(マで表示)である。Figure 3 shows the natural γ? of Figure 3 determined by the above method. These are a speech synthesis waveform obtained by combining speech synthesis parameters by matching the listening timing point of each syllable with the listening timing point of the voice, and its listening timing (indicated by a square). Fourth
The figure shows a speech synthesis waveform obtained by conventional syllable combination and listening timing points extracted from it (indicated by squares).
第1図と第4図の比較で、第4図の合成音声の受聴タイ
ミング点位置が自然音声の受聴タイミング点位置と著し
く異なっており、本発明による第3図の合成音の方が自
然なリズム感が得られることがわかる。A comparison between Figures 1 and 4 shows that the position of the listening timing point of the synthesized voice in Figure 4 is significantly different from that of the natural voice, and the synthesized voice of Figure 3 according to the present invention is more natural. You can see that you can get a sense of rhythm.
発明の効果
本発明は、上記のように自然音声より抽出しだ受聴タイ
ミング点位置に音節ファイルの受聴タイミング点を一致
させるように1″?節結合を行うようにしているので、
音節単位の1′6声合成バラメークを結合して生成され
る文章や文節等の音声に自然なリズムを与える利点を有
する。Effects of the Invention As described above, the present invention performs 1″? clause combination so that the listening timing point of the syllable file matches the listening timing point position extracted from natural speech.
It has the advantage of giving a natural rhythm to the sounds of sentences, phrases, etc. that are generated by combining 1'6-voice synthesis variations in syllable units.
第1図は自然音声の各節音毎の受聴タイミング点を示す
図、第2図は自然音声の受聴タイミング点抽出手順を示
す図、第3図は本発明の一実施例における音声合成方法
による合成音声波形を示す図、第4図は従来の方法によ
る合成音7h波形を示す図である。
代理人の氏名 弁理士 中 尾 敏 男 ほか1名第1
図
【川δ〕Fig. 1 is a diagram showing the listening timing points for each syllable of natural speech, Fig. 2 is a diagram showing the listening timing point extraction procedure of natural speech, and Fig. 3 is a diagram showing the listening timing point extraction procedure for natural speech. A diagram showing a synthesized speech waveform. FIG. 4 is a diagram showing a synthesized speech 7h waveform obtained by a conventional method. Name of agent: Patent attorney Toshio Nakao and 1 other person No. 1
Figure [River δ]
Claims (1)
イルの受聴タイミング点を一致させて音節結合を行うこ
とを特徴とする音声合成方法。A speech synthesis method characterized by performing syllable combination by matching a listening timing point of a syllable file to a listening timing point position extracted from natural speech.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP59113325A JP2548107B2 (en) | 1984-06-01 | 1984-06-01 | Speech synthesis method |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP59113325A JP2548107B2 (en) | 1984-06-01 | 1984-06-01 | Speech synthesis method |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| JPS60257500A true JPS60257500A (en) | 1985-12-19 |
| JP2548107B2 JP2548107B2 (en) | 1996-10-30 |
Family
ID=14609373
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP59113325A Expired - Lifetime JP2548107B2 (en) | 1984-06-01 | 1984-06-01 | Speech synthesis method |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JP2548107B2 (en) |
-
1984
- 1984-06-01 JP JP59113325A patent/JP2548107B2/en not_active Expired - Lifetime
Also Published As
| Publication number | Publication date |
|---|---|
| JP2548107B2 (en) | 1996-10-30 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| JPS62160495A (en) | Voice synthesization system | |
| JPS60257500A (en) | Voice synthesization | |
| JP3094622B2 (en) | Text-to-speech synthesizer | |
| JP2894447B2 (en) | Speech synthesizer using complex speech units | |
| JP2900454B2 (en) | Syllable data creation method for speech synthesizer | |
| JPH05224689A (en) | Speech synthesizer | |
| JP3241582B2 (en) | Prosody control device and method | |
| JP3367906B2 (en) | Speech synthesis method, speech synthesis device, recording medium recording speech synthesis program and speech segment record, method for creating the same, and recording medium recording speech segment record creation program | |
| JPS5880699A (en) | Voice synthesizing system | |
| JPS59155899A (en) | Voice synthesization system | |
| JPS6325700A (en) | Long note combination method | |
| JPS6021098A (en) | Synthesization of voice | |
| JPS60205596A (en) | Voice synthesizer | |
| JPS61173300A (en) | Voice synthesizer | |
| JPS6157997A (en) | Voice synthesization system | |
| JPS626299A (en) | Electronic singing apparatus | |
| JPH01118200A (en) | Voice synthesization system | |
| JPS63131191A (en) | Regular type voice synthesizer | |
| CN113178185A (en) | Singing synthesis method and system based on turning note processing method | |
| JPS60205597A (en) | Voice synthesizer | |
| JPH11327594A (en) | Speech synthesis dictionary creation system | |
| JPS59107391A (en) | Utterance training apparatus | |
| JPS6177897A (en) | Sentence-voice converter | |
| JPS5975296A (en) | Synthesization of voice | |
| JPS6147989A (en) | Voice time length data generator |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| EXPY | Cancellation because of completion of term |