JPH02210500A - Standard pattern registering system - Google Patents
Standard pattern registering systemInfo
- Publication number
- JPH02210500A JPH02210500A JP1031953A JP3195389A JPH02210500A JP H02210500 A JPH02210500 A JP H02210500A JP 1031953 A JP1031953 A JP 1031953A JP 3195389 A JP3195389 A JP 3195389A JP H02210500 A JPH02210500 A JP H02210500A
- Authority
- JP
- Japan
- Prior art keywords
- voice
- voices
- time
- registration
- similarity
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
Abstract
Description
【発明の詳細な説明】
挟哲分蟹−
本発明は、標準パターン登録方式、より詳細には、特定
話者用音声認識装置において、より正確な標準パターン
を登録する標準パターン登録方式従来、音声認識装置の
標準パターン登録方式においては登録のために発声され
た音声の区間を正確に検出したかどうかをチエツクする
有効な方法がないため、誤って区間検出した音声を標準
パターンとして登録してしまう場合があった。このよう
な問題に対して、音声を複数回発声させてM taを行
なう音声認識装置においては、標準パターン作成時、2
回目以降の発声時に1−回目の発声単語長と比較し、比
較結果がある範囲を超えた場合、標準パターンとして登
録しないようにし、標準パターンの精度を上げる方法が
提案されている。[Detailed Description of the Invention] The present invention relates to a standard pattern registration method, more specifically, a standard pattern registration method for registering a more accurate standard pattern in a speech recognition device for a specific speaker. In the standard pattern registration method of the recognition device, there is no effective way to check whether the section of the voice uttered for registration has been accurately detected, so the section of speech that is erroneously detected is registered as the standard pattern. There was a case. To solve this problem, in a speech recognition device that performs Mta by uttering a voice multiple times, when creating a standard pattern, 2
A method has been proposed in which the word length uttered for the first time and subsequent times is compared with the word length uttered for the first time, and if the comparison result exceeds a certain range, the word length is not registered as a standard pattern, thereby increasing the accuracy of the standard pattern.
しかしながら単語長の場合、同じ単語でも発声毎に大き
く変動し、従って範囲をきつくすると標準パターン登録
の際のリジェク1〜が多くなり、スムーズに登録が行な
えない可能性がでてくる。また逆に範囲を緩めてしまう
と今度は短いノイズがついたり、語頭や語尾が欠落して
もそのままイト録されてしまうという場合がでてくる。However, in the case of word length, even the same word varies greatly from utterance to utterance, and therefore, if the range is made too tight, there will be a large number of rejections during standard pattern registration, and there is a possibility that smooth registration will not be possible. On the other hand, if you loosen the range, short noises may be added, or even if the beginning or end of a word is missing, it may be recorded as it is.
上−−−煎
本発明は、上述のごとき実情に鑑みてなされたもので、
複数回発声させて標準パターンの登録を行なう音声認識
装置において、音声登録時に区間検出を誤った音声の登
録を防ぎ、正確な標準パターンを作成し、認識率を向上
させることを目的としてなされたものである。The present invention has been made in view of the above-mentioned circumstances.
In speech recognition devices that register standard patterns by uttering them multiple times, this technology was developed to prevent the registration of voices with incorrect segment detection during voice registration, create accurate standard patterns, and improve recognition rates. It is.
1M
本発明は、上記目的を達成するために、認識すべき音声
を予め登録しておき、音声が音声入力手段により入力さ
れた時、その入力音声を登録音声とパターン照合するこ
とにより認識を行なう音声認識装置の音声登録方式にお
いて、登録のために複数回発声された音声を各々入カバ
ソファに蓄えると同時に、隣合う発声をそれぞれパター
ン照合し、該照合結果がそれぞれ所定の類似度以上を示
し、かつバラツキが所定の範囲内であった場合のみ該複
数回発声を重ね合せて標準パターンとして登録を行なう
ことを特徴としたものである。以下、本発明の実施例に
基づいて説明する。1M In order to achieve the above object, the present invention registers the speech to be recognized in advance, and when the speech is inputted by the speech input means, performs recognition by pattern matching the input speech with the registered speech. In a voice registration method of a voice recognition device, voices uttered multiple times for registration are each stored in an input cover sofa, and at the same time, adjacent utterances are each pattern-matched, and the matching results each indicate a predetermined degree of similarity or higher; Moreover, only when the variation is within a predetermined range, the plurality of utterances are superimposed and registered as a standard pattern. Hereinafter, the present invention will be explained based on examples.
第1図は、本発明の一実施例を説明するための構成図、
第2図は、フロー図で、図中、1はマイクロフォン、2
は特徴量抽出部、3はパターン照合部、4は入カバソフ
ァ、5は標準パターン記憶部、6は結果出力部で、本発
明は、認識すべき音声を予め登録しておき、音声が音声
入力手段により入力された時、その入力音声を登録音声
とパターン照合することにより認識を行なう音声認識装
置の音声登録方式において、4γ録のために複数回(N
)回発声された音声を各々人力バッファに蓄えると同時
に、1回目と2回目、2回目と3回目、・・、N−1回
目とN回目というように隣合う発声をそれぞれパターン
照合し、該照合結果がそれぞれ所定の類似度以上を示し
、かつバラツキが所定の範囲内であった場合のみ該複数
回発声を重ね合せて標準パターンとして登録を行なうよ
うにしたものである。標準パターンの登録を行なう時、
複数回発声して登録するものであれば、いずれでも良い
が、ここでは仮に3回発声を行なって登録するものとし
て説明する。本発明において、標7((バターン登録の
際、まず音声入力手段(マイクロフォン)1より入力さ
れた音声は特徴量抽出部2で特徴量が抽出され、入力バ
ッファ4に保存される。FIG. 1 is a configuration diagram for explaining one embodiment of the present invention,
Figure 2 is a flow diagram, in which 1 is a microphone, 2
3 is a feature extraction unit, 3 is a pattern matching unit, 4 is an input cover sofa, 5 is a standard pattern storage unit, and 6 is a result output unit. In the present invention, the voice to be recognized is registered in advance, and the voice is input as a voice input. In the voice registration method of the voice recognition device, which performs recognition by pattern matching the input voice with the registered voice when the input voice is input by means of
) is stored in a manual buffer, and at the same time, patterns are matched between adjacent utterances such as the first and second, second and third, etc., N-1 and N-th times, and the corresponding Only when the matching results show a predetermined degree of similarity or higher and the variation is within a predetermined range, the plurality of utterances are superimposed and registered as a standard pattern. When registering a standard pattern,
Any method may be used as long as the voice is uttered multiple times and registered, but here, the description will be made assuming that the voice is uttered three times and then registered. In the present invention, when registering the mark 7 ((at the time of pattern registration), first, the feature quantity of the voice inputted from the voice input means (microphone) 1 is extracted by the feature quantity extracting section 2, and is stored in the input buffer 4.
1−回目の音声はそのまま人力バッファ4に保存され、
2回目が入力されると入力バッファ4に保存するととも
に、パターン照合部3において1回目と2回目の音声の
照合を行ない、所定の類似度以上を満たしているかをチ
エツクする。満たしていない場合には1回目又は2回目
のどちらかが区間検出を誤っていると判断されるので、
キャンセルし、最初の発声からやり直す。満たしている
場合には3回目の発声を促し、音声の取り込みを行なう
。3回目の音声が入力されると入力バッファ4に保存す
るとともに、パターン照合部3において2回目と3回目
の音声の照合を行ない、所定の類似度以上を濶だしてい
るかをチエツクする。満たしていない場合には3回目の
発声をキャンセルし、3回目の発声のみを再入力させ、
再度チエツクを行なう。満たしている場合には2回行な
った照合の各々の類似度のバラツキを調べ所定の範囲内
であるかどうかをチエツクし、範囲を超えているものに
ついては、やり直して登録を行なう。バラツキが所定の
範囲内であった場合には人力バッファに保存されている
3回の発声を重ね合わせて標準パターンとして標準パタ
ーン記憶部5にへ〉録する。The 1st audio is saved as is in the human buffer 4,
When the second input is made, it is stored in the input buffer 4, and the pattern matching section 3 compares the first and second voices to check whether they satisfy a predetermined degree of similarity or higher. If the conditions are not met, it will be determined that either the first or second time the section was detected incorrectly.
Cancel and start over from the first utterance. If the conditions are satisfied, the third utterance is prompted and the audio is captured. When the third voice is input, it is stored in the input buffer 4, and the second and third voices are compared in the pattern matching section 3 to check whether the similarity exceeds a predetermined degree. If the conditions are not met, cancel the third utterance and re-enter only the third utterance,
Check again. If the similarity is satisfied, the variation in similarity between the two comparisons is checked to see if it is within a predetermined range, and if it is outside the range, registration is performed again. If the variation is within a predetermined range, the three utterances stored in the manual buffer are superimposed and recorded in the standard pattern storage section 5 as a standard pattern.
助−米
以上の説明から明らかなように、本発明によると、複数
回発声させて標準パターンの登録を行なう音声認識装置
において、音声登録時に区間検出を誤った音声の登録を
防ぎ、正確な標準パターンを作成し、認識率を向上させ
ることがiif能どなる。As is clear from the above description, according to the present invention, in a speech recognition device that registers a standard pattern by uttering it multiple times, it is possible to prevent the registration of speech with incorrect section detection during speech registration, and to create an accurate standard pattern. Creating patterns and improving the recognition rate becomes an IIF function.
第1図は1本発明の一実施例を説明するための構成図、
第2図は、そのフロー図である。
1・・・マイクロフォン、2 特徴旦抽出部、3・パタ
ーン照合部、4・入力バッファ、5 ・標準パターン記
憶部、6 ・結果出力部。
特許出願人 株式会社 リコーFIG. 1 is a configuration diagram for explaining one embodiment of the present invention.
FIG. 2 is a flow diagram thereof. 1...Microphone, 2. Feature extraction section, 3. Pattern matching section, 4. Input buffer, 5. Standard pattern storage section, 6. Result output section. Patent applicant Ricoh Co., Ltd.
Claims (1)
力手段により入力された時、その入力音声を登録音声と
パターン照合することにより認識を行なう音声認識装置
の音声登録方式において、登録のために複数回発声され
た音声を各々入力バッファに蓄えると同時に、隣合う発
声をそれぞれパターン照合し、該照合結果がそれぞれ所
定の類似度以上を示し、かつバラツキが所定の範囲内で
あった場合のみ該複数回発声を重ね合せて標準パターン
として登録を行なうことを特徴とする標準パターン登録
方式。1. In the voice registration method of a voice recognition device, the voice to be recognized is registered in advance, and when the voice is input by a voice input means, recognition is performed by pattern matching the input voice with the registered voice. When each voice uttered multiple times is stored in an input buffer, each adjacent utterance is pattern-matched, and each of the matching results shows a predetermined degree of similarity or higher, and the variation is within a predetermined range. A standard pattern registration method characterized in that only the plurality of utterances are superimposed and registered as a standard pattern.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP1031953A JP2838848B2 (en) | 1989-02-10 | 1989-02-10 | Standard pattern registration method |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP1031953A JP2838848B2 (en) | 1989-02-10 | 1989-02-10 | Standard pattern registration method |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| JPH02210500A true JPH02210500A (en) | 1990-08-21 |
| JP2838848B2 JP2838848B2 (en) | 1998-12-16 |
Family
ID=12345323
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP1031953A Expired - Fee Related JP2838848B2 (en) | 1989-02-10 | 1989-02-10 | Standard pattern registration method |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JP2838848B2 (en) |
Cited By (6)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2007111197A1 (en) * | 2006-03-24 | 2007-10-04 | Pioneer Corporation | Speaker model registration device and method in speaker recognition system and computer program |
| WO2007111169A1 (en) * | 2006-03-24 | 2007-10-04 | Pioneer Corporation | Speaker model registration device, method, and computer program in speaker recognition system |
| JP2009211103A (en) * | 2003-03-25 | 2009-09-17 | Siemens Ag | Speaker-dependent voice recognition method and voice recognition system |
| WO2010086925A1 (en) * | 2009-01-30 | 2010-08-05 | 三菱電機株式会社 | Voice recognition device |
| JP2016076858A (en) * | 2014-10-07 | 2016-05-12 | パナソニックIpマネジメント株式会社 | Ringing tone registration system |
| JP2017535809A (en) * | 2014-10-22 | 2017-11-30 | クゥアルコム・インコーポレイテッドQualcomm Incorporated | Sound sample validation to generate a sound detection model |
-
1989
- 1989-02-10 JP JP1031953A patent/JP2838848B2/en not_active Expired - Fee Related
Cited By (10)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2009211103A (en) * | 2003-03-25 | 2009-09-17 | Siemens Ag | Speaker-dependent voice recognition method and voice recognition system |
| WO2007111197A1 (en) * | 2006-03-24 | 2007-10-04 | Pioneer Corporation | Speaker model registration device and method in speaker recognition system and computer program |
| WO2007111169A1 (en) * | 2006-03-24 | 2007-10-04 | Pioneer Corporation | Speaker model registration device, method, and computer program in speaker recognition system |
| JPWO2007111197A1 (en) * | 2006-03-24 | 2009-08-13 | パイオニア株式会社 | Speaker model registration apparatus and method in speaker recognition system, and computer program |
| JP4854732B2 (en) * | 2006-03-24 | 2012-01-18 | パイオニア株式会社 | Speaker model registration apparatus and method in speaker recognition system, and computer program |
| WO2010086925A1 (en) * | 2009-01-30 | 2010-08-05 | 三菱電機株式会社 | Voice recognition device |
| JP5172973B2 (en) * | 2009-01-30 | 2013-03-27 | 三菱電機株式会社 | Voice recognition device |
| US8977547B2 (en) | 2009-01-30 | 2015-03-10 | Mitsubishi Electric Corporation | Voice recognition system for registration of stable utterances |
| JP2016076858A (en) * | 2014-10-07 | 2016-05-12 | パナソニックIpマネジメント株式会社 | Ringing tone registration system |
| JP2017535809A (en) * | 2014-10-22 | 2017-11-30 | クゥアルコム・インコーポレイテッドQualcomm Incorporated | Sound sample validation to generate a sound detection model |
Also Published As
| Publication number | Publication date |
|---|---|
| JP2838848B2 (en) | 1998-12-16 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US5758021A (en) | Speech recognition combining dynamic programming and neural network techniques | |
| JP4408490B2 (en) | Method and apparatus for executing a database query | |
| JP2996019B2 (en) | Voice recognition device | |
| JP2838848B2 (en) | Standard pattern registration method | |
| JPH04318900A (en) | Multidirectional simultaneous sound collection type voice recognizing method | |
| JP2745562B2 (en) | Noise adaptive speech recognizer | |
| JPH0225517B2 (en) | ||
| JPS5852698A (en) | Voice recognition processing system | |
| JPS5934595A (en) | Voice recognition processing system | |
| JPS60107192A (en) | Pattern recognizing device | |
| JPH0484197A (en) | Continuous voice recognizer | |
| JPH0316038B2 (en) | ||
| JPS61260299A (en) | Voice recognition equipment | |
| JPS63306499A (en) | Reject system for unspecified speaker voice recognition equipment | |
| JPH02210499A (en) | Standard pattern registration method | |
| JPH04258999A (en) | Voice recognizing system | |
| JPS6193499A (en) | Audio pattern matching method | |
| JPS59124390A (en) | Candidate reduction voice recognition system | |
| JPS6312000A (en) | Voice recognition equipment | |
| JPS62245295A (en) | Specific speaker speech recognition device | |
| JPS61278896A (en) | Speaker collator | |
| JPH06100918B2 (en) | Voice recognizer | |
| JPH04134397A (en) | Voice recognizing device | |
| JPS6167898A (en) | Pattern creation method | |
| JPS61114298A (en) | Speaker collation system |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| FPAY | Renewal fee payment (event date is renewal date of database) |
Free format text: PAYMENT UNTIL: 20071016 Year of fee payment: 9 |
|
| FPAY | Renewal fee payment (event date is renewal date of database) |
Free format text: PAYMENT UNTIL: 20081016 Year of fee payment: 10 |
|
| LAPS | Cancellation because of no payment of annual fees |