JPH0310560Y2 - - Google Patents

Info

Publication number
JPH0310560Y2
JPH0310560Y2 JP1983154739U JP15473983U JPH0310560Y2 JP H0310560 Y2 JPH0310560 Y2 JP H0310560Y2 JP 1983154739 U JP1983154739 U JP 1983154739U JP 15473983 U JP15473983 U JP 15473983U JP H0310560 Y2 JPH0310560 Y2 JP H0310560Y2
Authority
JP
Japan
Prior art keywords
pattern
standard pattern
voice
input
section
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Expired
Application number
JP1983154739U
Other languages
Japanese (ja)
Other versions
JPS6063899U (en
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed filed Critical
Priority to JP1983154739U priority Critical patent/JPS6063899U/en
Publication of JPS6063899U publication Critical patent/JPS6063899U/en
Application granted granted Critical
Publication of JPH0310560Y2 publication Critical patent/JPH0310560Y2/ja
Granted legal-status Critical Current

Links

Description

【考案の詳細な説明】 [考案の技術分野] 本考案は音声認識装置に関する。[Detailed explanation of the idea] [Technical field of invention] The present invention relates to a speech recognition device.

[従来技術とその問題点] 従来、特定話者用の音声認識装置に於いては、
3つのモード即ち、登録モード、テストモード、
認識モードが設定できるようになつている。上記
テストモードは、登録したパタンが適当かどうか
をチエツクするモードであり、登録した同じ語を
再度入力し、パタンとの距離が閾値以下の時は、
登録パタンをそのままにしておき、閾値以上の時
は再入力した音声のパタンに書換える。この場
合、距離が閾値以下になるまで、音声入力とパタ
ンの書換えを繰返して行なう。上記のようにして
パタンの書換えが行なわれるが、上記従来の方法
では、登録単語が多ければ多いほど、登録時及び
テストモード時の手間がかかり、非常に煩わし
い。また、人間の声は経時変化があるため再登録
が必要であるが、その都度すべての単語を再登録
することは非常に面倒である。さらに、不明瞭な
登録を行なつたり、間違い易いパタンを登録する
と、誤認識を起し易い。
[Prior art and its problems] Conventionally, in speech recognition devices for specific speakers,
Three modes: registration mode, test mode,
Recognition mode can now be set. The above test mode is a mode to check whether the registered pattern is appropriate.If the same registered word is input again and the distance to the pattern is less than the threshold,
The registered pattern is left as it is, and when it exceeds the threshold value, it is rewritten to the re-input audio pattern. In this case, voice input and pattern rewriting are repeated until the distance becomes equal to or less than the threshold value. The pattern is rewritten as described above, but in the conventional method, the more words are registered, the more time and effort is required during registration and test mode, which is extremely troublesome. Furthermore, since human voices change over time, it is necessary to re-register them, but it is extremely troublesome to re-register all words each time. Furthermore, if the registration is unclear or if a pattern that is easily mistaken is registered, erroneous recognition is likely to occur.

[考案の目的] 本考案は上記の点に鑑みてなされたもので、登
録時のミスを減少させると共に、認識率を向上さ
せ、しかも、音声パタンの再登録の手間を軽減し
得る音声認識装置を提供することを目的とする。
[Purpose of the invention] The present invention has been made in view of the above points, and provides a speech recognition device that can reduce errors during registration, improve the recognition rate, and reduce the trouble of re-registering speech patterns. The purpose is to provide

[考案の要点] 本考案は音声認識装置において、距離が小さく
誤認識され易い単語の登録を禁止すると共に、経
時変化に伴う音声の変化に対応して標準パタンを
書換えるように構成したものである。
[Key points of the invention] This invention is a speech recognition device configured to prohibit the registration of words that are easily misrecognized due to their short distance, and to rewrite standard patterns in response to changes in speech over time. be.

[考案の実施例] 以下図面を参照して本考案の一実施例を説明す
る。第1図において11はマイクロフオンで、そ
の音声入力は音響処理部12へ送られる。この音
響処理部12は入力された音声信号をデジタル信
号に変換し、特徴抽出部13へ出力する。この特
徴抽出部13には始端終端検出部14が接続され
ており、入力信号のレベルから入力音声の始端終
端を検出して、その検出信号を制御部15へ出力
している。上記特徴抽出部13は入力音声の特徴
を抽出し、音声パタンとしてバツフア16へ一時
記憶する。このバツフア16に記憶された音声パ
タンはセレクタ17により選択され、標準パタン
メモリ18あるいは距離計算部19へ送られる。
上記セレクタ17は制御部15からの指令によつ
て動作する。また、上記標準パタンメモリ18及
び距離計算部19は、制御部15によつて制御さ
れる。上記標準パタンメモリ18には、認識すべ
き単語に対する標準パタンが書込まれており、制
御部15からの指令によつて記憶パタンが距離計
算部19へ順次読出される。この距離計算部19
は、入力音声のパタンと標準パタンとを比較して
その距離を計算し、最小値選択部20へ出力す
る。この最小値選択部20は信号部15によつて
制御され、距離計算部19の計算結果の中から最
小値を選択する。この最小値選択部20で選択さ
れた最小値は、制御部15へ送られると共に、セ
レクタ21を介して第1閾値判定回路22あるい
は第2閾値判定回路23へ送られる。上記セレク
タ21は制御部15からの指令によつて選択動作
する。上記第1閾値判定回路22及び第2閾値判
定回路23には、音声パタンと標準パタンとの最
小距離に対する閾値TH1,TH2が予め設定さ
れている。上記閾値TH1,TH2は、TH1<
TH2の関係にTH2を大きくとることにより、
入力音声の認識時に標準パタンの書換えを行わせ
る許容範囲を広げている。そして、上記第1閾値
判定回路22、第2閾値判定回路23の判定結果
は、制御部15へ送られる。この制御部15にキ
ー入力部24、表示部25が接続されており、最
小値選択部20で選択された結果に対する単語が
表示部25で表示される。上記キー入力部24
は、表示部25に表示された結果に対し、訂正操
作を行なつた場合、訂正信号が信号ラインaを介
して制御部15へ送られるようになつている。
[Embodiment of the invention] An embodiment of the invention will be described below with reference to the drawings. In FIG. 1, reference numeral 11 is a microphone, and its audio input is sent to a sound processing section 12. The acoustic processing section 12 converts the input audio signal into a digital signal and outputs it to the feature extraction section 13. A start/end detection section 14 is connected to the feature extraction section 13, which detects the start/end of the input audio from the level of the input signal and outputs the detection signal to the control section 15. The feature extractor 13 extracts the features of the input voice and temporarily stores them in the buffer 16 as a voice pattern. The audio pattern stored in the buffer 16 is selected by the selector 17 and sent to the standard pattern memory 18 or the distance calculation section 19.
The selector 17 operates according to a command from the control section 15. Further, the standard pattern memory 18 and the distance calculation section 19 are controlled by the control section 15. Standard patterns for words to be recognized are written in the standard pattern memory 18, and the stored patterns are sequentially read out to the distance calculation section 19 in response to instructions from the control section 15. This distance calculation section 19
compares the input voice pattern with the standard pattern, calculates the distance therebetween, and outputs the distance to the minimum value selection section 20. This minimum value selection section 20 is controlled by the signal section 15 and selects the minimum value from among the calculation results of the distance calculation section 19. The minimum value selected by the minimum value selection section 20 is sent to the control section 15 and also sent to the first threshold value determination circuit 22 or the second threshold value determination circuit 23 via the selector 21. The selector 21 performs a selection operation in response to a command from the control section 15. The first threshold value determination circuit 22 and the second threshold value determination circuit 23 are preset with threshold values TH1 and TH2 for the minimum distance between the audio pattern and the standard pattern. The above threshold values TH1 and TH2 are TH1<
By increasing TH2 in the TH2 relationship,
The allowable range for rewriting standard patterns when recognizing input speech has been expanded. Then, the determination results of the first threshold determination circuit 22 and the second threshold determination circuit 23 are sent to the control section 15. A key input section 24 and a display section 25 are connected to the control section 15, and the word corresponding to the result selected by the minimum value selection section 20 is displayed on the display section 25. The above key input section 24
When a correction operation is performed on the result displayed on the display section 25, a correction signal is sent to the control section 15 via the signal line a.

次に上記実施例の動作を第2図及び第3図のフ
ローチヤートを参照して説明する。登録モードに
おいて、マイクロフオン11より単語単位で音声
を入力すると、音響処理部12でデジタル信号に
変換され、特徴抽出部13へ送られる。これによ
り特徴抽出部13は入力音声の特徴を抽出し、バ
ツフア16へ出力する。一方、始端終端検出部1
4は、特徴抽出部13における信号入力レベルか
ら音声信号の始端を検出し、制御部15へ出力す
る。この制御部15は、上記始端検出信号によ
り、特徴抽出部13から出力される音声パタンの
バツフア16への読込を開始する。そして、始端
終端検出部14から終端検出信号が出力される
と、制御部15は、バツフア16に保持されてい
る音声パタンをセレクタ17を介して距離計算部
19へ転送し、第2図のステツプA1に示すよう
に標準パタンメモリ18に記憶されている各標準
パタンとの距離を計算する。この距離計算部19
の計算結果は最小値選択部20へ送られ、ステツ
プA2に示すように最小距離が選択される。この
最小値選択部20の選択結果は制御部15へ送ら
れ、それに対応する単語が表示部25において表
示される。また、上記最小値選択部20の選択結
果はセレクタ21を介して第1閾値判定回路22
へ送られ、ステツプA3に示すように最小距離が
閾値TH1より小さいか否か判定される。距離最
小値が閾値より小さい場合、つまり、入力音声が
標準パタンメモリ18の登録単語と類似している
場合には、ステツプA4に進み、表示部25にお
いて音声の再入力を指示する。また、上記ステツ
プA3で、最小距離が閾値TH1以上であると判
断された場合、つまり、他に類似している単語が
ないと判断された場合は、ステツプA5へ進んで
セレクタ17を標準パタンメモリ18側へ切替
え、入力音声のパタンを標準パタンとして標準パ
タンメモリ18に書込む。以上で1単語に対する
標準パタンの登録を終了する。以下、同様にして
他の単語に対するパタン登録が行なわれる。
Next, the operation of the above embodiment will be explained with reference to the flowcharts of FIGS. 2 and 3. In the registration mode, when speech is input word by word from the microphone 11, it is converted into a digital signal by the sound processing section 12 and sent to the feature extraction section 13. Thereby, the feature extraction unit 13 extracts the features of the input voice and outputs them to the buffer 16. On the other hand, the start end end detection section 1
4 detects the start end of the audio signal from the signal input level in the feature extraction section 13 and outputs it to the control section 15 . The control section 15 starts reading the audio pattern output from the feature extraction section 13 into the buffer 16 in response to the start edge detection signal. When the end detection signal is output from the start/end detection section 14, the control section 15 transfers the audio pattern held in the buffer 16 to the distance calculation section 19 via the selector 17, and performs the steps shown in FIG. As shown in A1, the distance from each standard pattern stored in the standard pattern memory 18 is calculated. This distance calculation section 19
The calculation result is sent to the minimum value selection section 20, and the minimum distance is selected as shown in step A2. The selection result of the minimum value selection section 20 is sent to the control section 15, and the corresponding word is displayed on the display section 25. Further, the selection result of the minimum value selection section 20 is sent to a first threshold value determination circuit 22 via a selector 21.
As shown in step A3, it is determined whether the minimum distance is smaller than a threshold value TH1. If the minimum distance value is smaller than the threshold, that is, if the input voice is similar to the registered word in the standard pattern memory 18, the process proceeds to step A4, and the display section 25 instructs to re-input the voice. If it is determined in step A3 that the minimum distance is greater than or equal to the threshold TH1, that is, if it is determined that there are no other similar words, the process proceeds to step A5 and the selector 17 is stored in the standard pattern memory. 18 side, and writes the pattern of the input voice into the standard pattern memory 18 as a standard pattern. This completes the registration of standard patterns for one word. Thereafter, pattern registration for other words is performed in the same manner.

次に上記パタン登録後、入力音声を認識する場
合の動作について説明する。認識モードにおい
て、1単語に対する音声を入力すると、第3図の
ステツプB1に示すように、上記の場合と同様に
距離計算部19で入力音声のパタンと標準パタン
との距離が計算される。次いでステツプB2に進
み、最小値選択部20において最小距離が求めら
れ、セレクタ21を介して第2閾値判定回路23
へ送られる。この第2閾値判定回路23ではステ
ツプB3に示すように上記最小距離が閾値TH2
以下か否かを判定し、最小距離が閾値TH2より
大きければ、ステツプB4に示すように認識結果
をリジエクトする。また、上記ステツプB3にお
いて最小距離が閾値TH2以下であると判定され
た場合は、ステツプB5へ進み、距離最小値に対
応する単語を認識結果として表示部25に表示す
る。この認識結果に対し、使用者が自分の意図し
た結果と相違した場合は、キー入力部24から訂
正指示を与える。キー入力部24から訂正信号が
入力された場合、制御部15は、ステツプB7に
示すように表示部25において再入力を指示す
る。また、上記ステツプB6において、訂正信号
が入力されなかつた場合は、上記バツフア16に
保持している入力音声のパタンを、標準パタンと
してセレクタ17を介して標準パタンメモリ18
に登録する。すなわち、認識モードにおいて、認
識結果が正しかつた場合は、入力音声のパタンを
標準パタンとして標準パタンメモリ18に再登録
するようにしている。この標準パタンメモリ18
へのパタン登録は、認識の都度行なわれる。
Next, a description will be given of the operation when recognizing the input voice after the above-mentioned pattern registration. In the recognition mode, when speech for one word is input, as shown in step B1 of FIG. 3, the distance calculating section 19 calculates the distance between the input speech pattern and the standard pattern, as in the above case. Next, the process proceeds to step B2, in which the minimum distance is determined in the minimum value selection section 20, and the minimum distance is determined via the selector 21 in the second threshold determination circuit 23.
sent to. In this second threshold determination circuit 23, as shown in step B3, the minimum distance is set to the threshold TH2.
It is determined whether the minimum distance is less than or equal to the threshold value TH2, and if the minimum distance is greater than the threshold value TH2, the recognition result is rejected as shown in step B4. If it is determined in step B3 that the minimum distance is less than or equal to the threshold TH2, the process proceeds to step B5, and the word corresponding to the minimum distance value is displayed on the display section 25 as a recognition result. If the recognition result is different from the user's intended result, the user issues a correction instruction from the key input unit 24. When a correction signal is input from the key input section 24, the control section 15 instructs re-input on the display section 25 as shown in step B7. Further, in step B6, if no correction signal is input, the pattern of the input voice held in the buffer 16 is transferred to the standard pattern memory 18 via the selector 17 as a standard pattern.
Register. That is, in the recognition mode, if the recognition result is correct, the pattern of the input voice is re-registered in the standard pattern memory 18 as a standard pattern. This standard pattern memory 18
Pattern registration is performed each time recognition is performed.

[考案の効果] 以上述べたように本考案によれば、登録モード
時、他の単語と距離が小さく誤認識され易い単語
については、標準パタンメモリ18への登録を行
なわないようにしたので、登録時のミスを減少し
て認識率を向上することができる。また、認識モ
ードにおいて、認識結果が正しい場合はその入力
音声のパタンを標準パタンメモリ18に再登録す
るようにし、かつ正しいと認識する範囲を広げる
ようにしたので、標準パタンの書換動作の頻度を
高めることができるようになり、音声の経時変化
による誤認識を防止できると共に、標準パタンの
再登録の手間を軽減することができる。
[Effects of the invention] As described above, according to the invention, in the registration mode, words that are easily misrecognized due to their small distance from other words are not registered in the standard pattern memory 18. It is possible to reduce errors during registration and improve recognition rate. In addition, in recognition mode, if the recognition result is correct, the pattern of the input voice is re-registered in the standard pattern memory 18, and the range of recognition as correct is expanded, so the frequency of rewriting the standard pattern is reduced. This makes it possible to prevent erroneous recognition due to changes in voice over time, and to reduce the effort required to re-register standard patterns.

【図面の簡単な説明】[Brief explanation of drawings]

図面は本考案の一実施例を示すもので、第1図
は回路構成を示すブロツク図、第2図は登録モー
ド時の動作を示すフローチヤート、第3図は認識
モード時の動作を示すフローチヤートである。 11……マイクロフオン、12……音響処理
部、13……特徴抽出部、14……始端終端検出
部、15……制御部、16……バツフア、17…
…セレクタ、18……標準パタンメモリ、19…
…距離計算部、20……最小値選択部、21……
セレクタ、22……第1閾値判定回路、23……
第2閾値判定回路、24……キー入力部、25…
…表示部。
The drawings show one embodiment of the present invention; FIG. 1 is a block diagram showing the circuit configuration, FIG. 2 is a flowchart showing the operation in registration mode, and FIG. 3 is a flowchart showing the operation in recognition mode. It's a chat. DESCRIPTION OF SYMBOLS 11...Microphone, 12...Acoustic processing section, 13...Feature extraction section, 14...Starting end end detection section, 15...Control section, 16...Buffer, 17...
...Selector, 18...Standard pattern memory, 19...
... Distance calculation section, 20 ... Minimum value selection section, 21 ...
Selector, 22...First threshold value determination circuit, 23...
Second threshold value determination circuit, 24...key input section, 25...
...Display section.

Claims (1)

【実用新案登録請求の範囲】 入力音声と標準パタンメモリに登録されている
各標準パタンとの距離を計算し、計算結果の中の
最小値を取出す距離計算手段と、 この距離計算手段により得られた最小値が第1
の閾値以上であるか否かを判断する第1の判断手
段と、 上記距離計算手段により得られた最小値が上記
第1の閾値より大きく設定された第2の閾値以上
であるか否かを判断する第2の判断手段と、 音声パタンの登録時入力音声に関する最小値が
上記第1の判断手段により第1の閾値以上と判断
された場合にこの入力音声の音声パタンを標準パ
タンとして新たに上記標準パタンメモリに登録
し、第1の閾値より小さいと判断された場合に音
声の再入力を指示する制御手段と、 入力音声の認識時入力音声に関する最小値が上
記第2の判断手段により第2の閾値以下と判断さ
れた場合に当該最小値に対応する標準パタンを認
識結果として選定し、当該標準パタンに関する単
語を出力する認識手段と、 この認識手段により出力された単語が適正であ
るときに選定された標準パタンを入力された音声
パタンに書換える標準パタン書換手段とを 具備したことを特徴とする音声認識装置。
[Scope of Claim for Utility Model Registration] Distance calculation means for calculating the distance between the input voice and each standard pattern registered in the standard pattern memory and taking the minimum value among the calculation results; The minimum value is the first
a first determination means for determining whether or not the minimum value obtained by the distance calculation means is equal to or greater than a second threshold set larger than the first threshold; a second determining means for determining the voice pattern; and a second determining means for determining the voice pattern of the input voice as a new standard pattern when the minimum value regarding the input voice at the time of registering the voice pattern is determined by the first determining means to be equal to or greater than the first threshold value. a control means for registering the standard pattern in the standard pattern memory and instructing re-input of the voice when it is determined that the voice is smaller than the first threshold; a recognition means that selects a standard pattern corresponding to the minimum value as a recognition result when it is determined that the minimum value is less than the threshold of 2, and outputs a word related to the standard pattern; and when the word output by this recognition means is appropriate; 1. A speech recognition device comprising: standard pattern rewriting means for rewriting a standard pattern selected by the user into an input speech pattern.
JP1983154739U 1983-10-05 1983-10-05 voice recognition device Granted JPS6063899U (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP1983154739U JPS6063899U (en) 1983-10-05 1983-10-05 voice recognition device

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP1983154739U JPS6063899U (en) 1983-10-05 1983-10-05 voice recognition device

Publications (2)

Publication Number Publication Date
JPS6063899U JPS6063899U (en) 1985-05-04
JPH0310560Y2 true JPH0310560Y2 (en) 1991-03-15

Family

ID=30341962

Family Applications (1)

Application Number Title Priority Date Filing Date
JP1983154739U Granted JPS6063899U (en) 1983-10-05 1983-10-05 voice recognition device

Country Status (1)

Country Link
JP (1) JPS6063899U (en)

Family Cites Families (2)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPS5681899A (en) * 1979-12-07 1981-07-04 Sanyo Electric Co Voice indentifier
JPS57141700A (en) * 1981-02-26 1982-09-02 Mitsubishi Electric Corp Voice recognizer

Also Published As

Publication number Publication date
JPS6063899U (en) 1985-05-04

Similar Documents

Publication Publication Date Title
JPH0962293A (en) Speech recognition dialogue device and speech recognition dialogue processing method
JPH0310560Y2 (en)
JP2002073061A (en) Speech recognition apparatus and method
US10818298B2 (en) Audio processing
JP2024119245A (en) Voice processing device, method and program
JP2002215184A (en) Voice recognition device and program
JPS584198A (en) Standard pattern registration system for voice recognition unit
JP2754960B2 (en) Voice recognition device
JPS645320B2 (en)
JP2024131933A (en) Audio processing device, method and program
JPH0236960B2 (en)
JPH07210186A (en) Voice registration device
JPH0619491A (en) Speech recognizing device
JPH09198079A (en) Voice recognition device
JPH0555880B2 (en)
JPS63305396A (en) Voice recognition equipment
JPH0634234B2 (en) Pattern recognizer
JPH0343639B2 (en)
JPH03155599A (en) Speech recognition device
JPS63281196A (en) Voice recognition equipment
JP2646539B2 (en) Standard pattern storage section management method
JPS637398B2 (en)
JPS60205600A (en) Voice recognition equipment
JPH01200293A (en) Sound recognizing system
JPS6011897A (en) Voice recognition equipment