JPH04155398A - Voice recognition device - Google Patents

Voice recognition device

Info

Publication number
JPH04155398A
JPH04155398A JP2280299A JP28029990A JPH04155398A JP H04155398 A JPH04155398 A JP H04155398A JP 2280299 A JP2280299 A JP 2280299A JP 28029990 A JP28029990 A JP 28029990A JP H04155398 A JPH04155398 A JP H04155398A
Authority
JP
Japan
Prior art keywords
template
unit
parameter
reject
similarity
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
JP2280299A
Other languages
Japanese (ja)
Inventor
Toshiki Kawamoto
河本 俊毅
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Ricoh Co Ltd
Original Assignee
Ricoh Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Ricoh Co Ltd filed Critical Ricoh Co Ltd
Priority to JP2280299A priority Critical patent/JPH04155398A/en
Publication of JPH04155398A publication Critical patent/JPH04155398A/en
Pending legal-status Critical Current

Links

Abstract

PURPOSE:To prevent a fluctuation of rejection rate by a word and a fluctuation of rejection rate by a template group, by deciding the rejection parameter to each temperature or to a template group to be the recognition object, and rejecting by using the parameter. CONSTITUTION:When a speaker pushes the key of a keyboard 9, a register signal is transmitted to a controller 8, and a switch circuit 3 is converted to a register 4 side. The voice of the speaker is converted to an electric signal and delivered to a voice feature extractor 2. The extracted feature amount is delivered to the register 4 to produce a template. The data of all the templates are stored in a template memory 6. After that, the register 4 delivers a finishing signal to a controller 8, the controller 8 makes a recognition member 5 carry out the matching of the templates each other, and delivers the degree of similarity of them to a rejection parameter deciding member 7. In the deciding member 7, the rejection parameter is determined as to every template.

Description

【発明の詳細な説明】 産業上の利用分野 本発明は、音声認識装置に関する。[Detailed description of the invention] Industrial applications The present invention relates to a speech recognition device.

従来の技術 従来、音声認識装置でリジェクトを行う場合、一定のリ
ジェクトパラメータでは発声する単語によってリジェク
ト率に異差が生じ、単語によってリジェクト率が非常に
高かったり、低がったりする。例えば、ある単語群では
りジェクト率が2%でも他の単語群では10%になるこ
ともある。
BACKGROUND ART Conventionally, when a speech recognition device performs a rejection, the rejection rate varies depending on the word to be uttered with a fixed rejection parameter, and the rejection rate may be very high or low depending on the word. For example, a rejection rate of 2% for one word group may be 10% for another word group.

また、このような現象は、新たに単語を追加登録した場
合などにも同様の現象が生じることがある。
Furthermore, a similar phenomenon may occur when a new word is additionally registered.

発明が解決しようとする課題 上述したようにリジェクトを行う場合、認識対象単語が
変わると、同じリジェクトパラメータを用いて認識結果
をリジェクトしていては、あるテンプレート群には適正
なりジェクトがかがるパラメータ値であっても、他のテ
ンプレート群で認識した時にはりジェクト率が高すぎた
り、低すぎたりすることがあり、このように従来の音声
認識装置ではテンプレート群によってリジェクト率が変
動するのを防ぐことができない。
Problems to be Solved by the Invention When performing rejection as described above, if the recognition target word changes, if the same rejection parameter is used to reject the recognition result, a certain group of templates may be rejected properly. Even if the parameter value is recognized, the rejection rate may be too high or too low when recognized using other template groups. cannot be prevented.

課題を解決するための手段 そこで、このような問題点を解決するために、請求項1
記載の発明では、入力音声の特徴を抽出する音声特徴抽
出部を設け、この音声特徴抽出部により抽出された入力
音声の特徴量を用いてテンプレートを作成する登録部を
設け、この登録部に登録された前記テンプレートを記憶
するテンプレート記憶部を設け、このテンプレート記憶
部に記憶されたテンプレート群を用いて各テンプレート
間のマツチングを行いその類似度を計算する認識部を設
け、登録時に前記認識部により求められた前記類似度の
大きさに応じてリジェクトパラメータをテンプレート毎
に決定するリジェクトパラメータ決定部を設けた。
Means for Solving the Problem Therefore, in order to solve such problems, claim 1
In the described invention, a voice feature extraction unit is provided for extracting the features of input voice, a registration unit is provided for creating a template using the feature amount of the input voice extracted by the voice feature extraction unit, and the template is registered in the registration unit. A template storage unit that stores the templates stored in the template is provided, and a recognition unit that performs matching between templates and calculates the degree of similarity using the template group stored in the template storage unit. A reject parameter determination unit is provided that determines a reject parameter for each template in accordance with the determined degree of similarity.

請求項2記載の発明では、入力音声の特徴を抽出する音
声特徴抽出部を設け、この音声特徴抽出部により抽出さ
れた入力音声の特徴量を用いてテンプレートを作成する
登録部を設け、この登録部に登録された前記テンプレー
トを記憶するテンプレート記憶部を設け、このテンプレ
ート記憶部に記憶されたテンプレート群を用いて各テン
プレート間のマツチングを行いその類似度を計算する認
識部を設け、登録時に前記認識部により求められた前記
類似度の大きさに応じてリジェクトパラメータをテンプ
レートの組合せ毎に決定するリジェクトパラメータ決定
部を設けた。
In the invention as set forth in claim 2, a voice feature extraction section is provided for extracting the features of the input voice, and a registration section is provided for creating a template using the feature amount of the input voice extracted by the voice feature extraction section. A template storage unit is provided to store the templates registered in the template storage unit, and a recognition unit is provided to match each template using the template group stored in the template storage unit and calculate the degree of similarity. A reject parameter determining unit is provided that determines a reject parameter for each combination of templates in accordance with the degree of similarity determined by the recognition unit.

請求項3記載の発明では、入力音声の特徴を抽出する音
声特徴抽出部を設け、この音声特徴抽出部により抽出さ
れた入力音声の特徴量を用いてテンプレートを作成する
登録部を設け、この登録部に登録された前記テンプレー
トを記憶するテンプレート記憶部を設け、このテンプレ
ート記憶部に記憶されたテンプレート群を用いて各テン
プレート間のマツチングを行いその類似度を計算する認
識部を設け、登録時に前記認識部により求められた前記
類似度の大きさに応じてリジェクトパラメータを決定し
以後前記テンプレートを追加する毎にそのリジェクトパ
ラメータの値を更新していくリジェクトパラメータ決定
部を設けた。
In the invention as set forth in claim 3, a voice feature extraction section is provided for extracting the features of the input voice, and a registration section is provided for creating a template using the feature amount of the input voice extracted by the voice feature extraction section. A template storage unit is provided to store the templates registered in the template storage unit, and a recognition unit is provided to match each template using the template group stored in the template storage unit and calculate the degree of similarity. A reject parameter determining unit is provided which determines a reject parameter according to the degree of similarity determined by the recognition unit and updates the value of the reject parameter each time the template is added thereafter.

請求項4記載の発明では、入力音声の特徴を抽出する音
声特徴抽出部を設け、この音声特徴抽出部により抽出さ
れた入力音声の特徴量を用いてテンプレートを作成する
登録部を設け、この登録部に登録された前記テンプレー
ト及びリジェクトパラメータを記憶するテンプレートリ
ジェクトパラメータ記憶部を設け、このテンプレートリ
ジェクトパラメータ記憶部に記憶された前記リジェクト
パラメータ及び前記テンプレートを読出すと共にテンプ
レート群を用いて各テンプレート間のマツチングを行い
その類似度を計算する認識部を設け、登録時に前記認識
部により求められた前記類似度の大きさに応じてリジェ
クトパラメータを決定し以後前記テンプレートを追加す
る毎に前記リジェクトパラメータの値を更新していくリ
ジェクトパラメータ決定部を設けた。
In the invention as set forth in claim 4, a voice feature extraction section is provided for extracting the features of the input voice, and a registration section is provided for creating a template using the feature amount of the input voice extracted by the voice feature extraction section. A template reject parameter storage unit is provided to store the template and reject parameters registered in the template reject parameter storage unit, and the reject parameter and the template stored in the template reject parameter storage unit are read out, and the template group is used to distinguish between the templates. A recognition unit that performs matching and calculates the degree of similarity is provided, and a reject parameter is determined according to the degree of similarity determined by the recognition unit at the time of registration, and the value of the reject parameter is determined every time the template is added thereafter. We have provided a reject parameter determination section that updates the parameters.

作用 請求項1記載の発明は、認識対象となるテンプレート毎
にリジェクトパラメータを決定し、このパラメータを用
いてリジェクトすることによって、単語によるリジェク
ト率の変動を防ぐことができる。
According to the invention described in claim 1, a rejection parameter is determined for each template to be recognized, and rejection is performed using this parameter, thereby making it possible to prevent fluctuations in the rejection rate depending on words.

請求項2記載の発明は、認識対象となるテンプレートの
組合わせ毎にリジェクトパラメータを決定し、このパラ
メータを用いてリジェクトすることによって、単語によ
るリジェクト率の変動を防ぐことができる。
According to the second aspect of the invention, a rejection parameter is determined for each combination of templates to be recognized, and rejection is performed using this parameter, thereby making it possible to prevent variations in the rejection rate due to words.

請求項3,4記載の発明は、認識対象となるテンプレー
ト群によってリジェクトパラメータを決定し、このパラ
メータを用いてリジェクトすることによって、テンプレ
ート群によるリジェクト率の変動を防ぐことができる。
According to the third and fourth aspects of the invention, a rejection parameter is determined based on a group of templates to be recognized, and rejection is performed using this parameter, thereby making it possible to prevent variations in the rejection rate due to the group of templates.

実施例 請求項1記載の発明の一実施例を第1図に基づいて説明
する。まず、その全体構成について述べる。音声を電気
的信号に変換するマイクロフォンlは、入力音声の特徴
を抽出する音声特徴抽出部2に接続されている。この音
声特徴抽出部2はスイッチ回路3に接続されている。こ
のスイッチ回路3は、登録部4及び認識部5と接続され
、切換えができるようになっている。前記登録部4は、
前記音声特徴抽出部2により抽出された入力音声の特徴
量を用いてテンプレートを作成することができる。この
登録部4には、登録された前記テンプレートを記憶する
テンプレート記憶部6が接続されている。
Embodiment An embodiment of the invention set forth in claim 1 will be described based on FIG. First, the overall structure will be described. A microphone 1 that converts voice into an electrical signal is connected to a voice feature extractor 2 that extracts features of input voice. This audio feature extraction section 2 is connected to a switch circuit 3. This switch circuit 3 is connected to a registration section 4 and a recognition section 5 so that switching can be performed. The registration unit 4
A template can be created using the feature amount of the input speech extracted by the speech feature extraction section 2. A template storage unit 6 that stores the registered template is connected to the registration unit 4.

また、このテンプレート記憶部6は前記認識部5と接続
されている。この認識部5は、前記テンプレート記憶部
6に記憶されたテンプレート群を用いて各テンプレート
間のマツチングを行いその類似度を計算する働きがある
。さらに、前記認識部5はリジェクトパラメータ決定部
7と接続されている。このリジェクトパラメータ決定部
7は、前記認識部5により求められた類似度の大きさに
応じてリジェクトパラメータをテンプレート毎に決定す
る働きがある。
Further, this template storage section 6 is connected to the recognition section 5. The recognition unit 5 has the function of matching templates using the template group stored in the template storage unit 6 and calculating the degree of similarity between the templates. Further, the recognition section 5 is connected to a reject parameter determination section 7. The reject parameter determining section 7 has the function of determining a reject parameter for each template according to the degree of similarity determined by the recognizing section 5.

さらに、この他の部分の構成として、上記各部と接続さ
れる制御部8が設けられている。この制御部8には、登
録するか認識するかを切換えるキーボード9と認識結果
を表示する認識表示部10とが接続されている。
Furthermore, as a configuration of other parts, a control section 8 is provided which is connected to each of the above-mentioned sections. Connected to this control section 8 are a keyboard 9 for switching between registration and recognition, and a recognition display section 10 for displaying recognition results.

このような構成において、まず、テンプレート登録時に
ついて述べる。話者がキーボード9の所定のキーを押す
ことによって、登録を行うことの信号が制御部8に伝え
られる。制御部8はその信号を受けると、スイッチ回路
3を登録部4側に切換える。次に、話者がマイクロフォ
ン1に向かって発声した音声が、電気信号に変換されて
音声特徴抽出部2に送られる。ここで、抽出された特徴
量は登録部4に送られて、これによりテンプレートが作
成される。この場合、話者がキーボード9によって設定
した単語数分のテンプレートが作成されると、全テンプ
レートのデータはテンプレート記憶部6に送られて記憶
される。その後、登録部4は制御部8に終了信号を送り
、制御部8はその信号を受は取ると、認識部5にテンプ
レート同士のマツチングを開始させる。この認識部5で
は、登録部4で作成されたテンプレート群を用いて各テ
ンプレート間のマツチングを行い、その類似度をリジェ
クトパラメータ決定部7に送る。これを全てのテンプレ
ートの組合わせに対して行い、その全ての類似度データ
をもとにしてリジェクトパラメータ決定部7ではリジェ
クトパラメータをテンプレート毎に決定する。
In such a configuration, first, the time of template registration will be described. When the speaker presses a predetermined key on the keyboard 9, a signal to perform registration is transmitted to the control unit 8. When the control section 8 receives the signal, it switches the switch circuit 3 to the registration section 4 side. Next, the voice uttered by the speaker into the microphone 1 is converted into an electrical signal and sent to the voice feature extraction section 2. Here, the extracted feature amount is sent to the registration unit 4, and a template is created thereby. In this case, once templates for the number of words set by the speaker using the keyboard 9 are created, the data of all the templates is sent to the template storage section 6 and stored therein. Thereafter, the registration section 4 sends an end signal to the control section 8, and when the control section 8 receives the signal, it causes the recognition section 5 to start matching the templates. The recognition unit 5 performs matching between templates using the template group created by the registration unit 4, and sends the degree of similarity to the rejection parameter determination unit 7. This is performed for all combinations of templates, and based on all of the similarity data, the reject parameter determining unit 7 determines reject parameters for each template.

例として、どれかのテンプレートとの類似度がX以上に
なったテンプレートAに対してはリジェクトパラメータ
はXにするように決定し、どれかのテンプレートとの類
似度がy以上、X以下になったテンプレートBに対して
はリジェクトパラメータはYにするように決定する。
For example, for template A whose similarity with any template is X or more, the reject parameter is determined to be X, and if the similarity with any template is y or more and X or less. For template B, the reject parameter is determined to be Y.

次に、認識時について述べる。話者はキーボード9の所
定のキーを押すことによって、認識を行うことの信号が
制御部8に伝えられる。制御部8はその信号を受は取る
と、スイッチ回路3を認識部5側に切換える。次に、話
者がマイクロフォンlに向かって発生した音声が、電気
信号に変換されて音声特徴抽出部2に送られる。ここで
抽出された特徴量が認識部5に送られて、テンプレート
記憶部6から送られたテンプレート群とマツチングが行
われる。これにより、1番目に類似度の大きいものを「
第−M3識候補J、2番目に大きいものを「第二認識候
補」として、それぞれそのテンプレートとの類似度の大
きさ、差、比等と、テンプレート登録時に決定された第
一認識候補のテンプレートに対するリジェクトパラメー
タとを用いてリジェクトするかどうかを判断し、リジェ
クトしない場合には「第一認識候補」を認識結果として
、また、リジェクトする場合にはリジェクトしたことを
伝える信号を認識表示部10に送る。
Next, the time of recognition will be described. When the speaker presses a predetermined key on the keyboard 9, a signal to perform recognition is transmitted to the control unit 8. When the control section 8 receives the signal, it switches the switch circuit 3 to the recognition section 5 side. Next, the voice generated by the speaker into the microphone l is converted into an electrical signal and sent to the voice feature extraction section 2. The feature amount extracted here is sent to the recognition unit 5 and matched with the template group sent from the template storage unit 6. As a result, the item with the highest degree of similarity is
-M3 recognition candidate J, with the second largest one as the "second recognition candidate", and the degree of similarity, difference, ratio, etc. with that template, and the template of the first recognition candidate determined at the time of template registration. It is determined whether or not to reject using the rejection parameters for the target, and if it is not rejected, the "first recognition candidate" is set as the recognition result, and if it is rejected, a signal indicating that it has been rejected is sent to the recognition display unit 10. send.

上述したように、認識対象となるテンプレート毎にリジ
ェクトパラメータを決定し、このパラメータを用いてリ
ジェクトすることによって、単語によるリジェクト率の
変動を防ぐことができる。
As described above, by determining a rejection parameter for each template to be recognized and rejecting using this parameter, it is possible to prevent fluctuations in the rejection rate depending on words.

次に、請求項2記載の発明の一実施例について説明する
。なお、前述した請求項1記載の発明の実施例で述べた
回路(第1図参照)と同一部分についての説明は省略し
、その同一部分については同一符号を用いる。
Next, an embodiment of the invention according to claim 2 will be described. Note that explanations of the same parts as those of the circuit (see FIG. 1) described in the embodiment of the invention according to claim 1 described above will be omitted, and the same parts will be denoted by the same reference numerals.

ここでは、リジェクトパラメータ決定部7の内部構成を
変えたものである。すなわち、リジェクトパラメータ決
定部7は、認識部5により求められた類似度の大きさに
応じて、リジェクトパラメータをテンプレートの組合わ
せ毎に決定するようにしたものである。
Here, the internal configuration of the reject parameter determining section 7 is changed. That is, the reject parameter determining unit 7 determines a reject parameter for each combination of templates according to the degree of similarity determined by the recognizing unit 5.

これにより、登録時には、リジェクトパラメータ決定部
7では、リジェクトパラメータをテンプレートの組合わ
せ毎に決定することになる。例えば、類似度がX以上に
なった組合わせのテンプレートA、Bに対してはリジェ
クトパラメータはXにするように決定し、類似度がy以
上、X以下になた組合わせのテンプレートA、Cに対し
てはリジェクトパラメータはYにするように決定する。
Thereby, at the time of registration, the reject parameter determination unit 7 determines a reject parameter for each combination of templates. For example, the reject parameter is determined to be X for templates A and B whose similarity is greater than or equal to X, and templates A and C whose similarity is greater than or equal to y and less than or equal to X. , the reject parameter is determined to be Y.

また、認識時には、「第一認識候補」と「第二認識候補
」とのテンプレートの組合わせで決定しているリジェク
トパラメータと、それぞれのテンプレートとの類似度の
大きさ、差、比等とを用いてリジェクトするかどうかを
判断し、リジェクトしない場合場合には「第一認識候補
」を認識結果として、また、リジェクトする場合にはリ
ジェクトしたことを伝える信号を認識表示部】Oに送る
ようにする。
In addition, during recognition, the rejection parameters determined by the combination of templates "first recognition candidate" and "second recognition candidate" and the degree of similarity, difference, ratio, etc. between each template are evaluated. If the recognition result is not rejected, the "first recognition candidate" is sent as the recognition result, and if the recognition result is rejected, a signal indicating the rejection is sent to the recognition display section ]O. do.

上述したように、認識対象となるテンプレートの組合わ
せ毎にリジェクトパラメータを決定し、このパラメータ
を用いてリジェクトすることによって、単語によるリジ
ェクト率の変動を防ぐことができる。
As described above, by determining a rejection parameter for each combination of templates to be recognized and rejecting using this parameter, it is possible to prevent fluctuations in the rejection rate depending on words.

次に、請求項3,4記載の発明の一実施例を第2図に基
づいて説明する。なお、前述した請求項l記載の発明の
実施例(第1図参照)と同一部分についての説明は省略
し、その同一部分については同一符号を用いる。
Next, an embodiment of the invention according to claims 3 and 4 will be described based on FIG. Note that the description of the same parts as in the embodiment of the invention described in claim 1 (see FIG. 1) described above will be omitted, and the same parts will be denoted by the same reference numerals.

ここでは、リジェクトパラメータ決定部7、テンプレー
トリジェクトパラメータ記憶部11、認識部5の内部構
成を変えたものである。すなわち、前記リジェクトパラ
メータ決定部7は、前記認識部5により求められた類似
度の大きさに応じてリジェクトパラメータを決定し、以
後テンプレートを追加する毎にリジェクトパラメータの
値を更新していく働きがある。また、前記テンプレート
リジェクトパラメータ記憶部11は、前記登録部4に登
録されたテンプレートを記憶するだけでなく、リジェク
トパラメータをも記憶する働きがある。
Here, the internal configurations of the reject parameter determination section 7, template reject parameter storage section 11, and recognition section 5 are changed. That is, the reject parameter determining unit 7 has the function of determining a reject parameter according to the degree of similarity determined by the recognizing unit 5, and updates the value of the reject parameter every time a template is added thereafter. be. Further, the template reject parameter storage section 11 has the function of not only storing templates registered in the registration section 4 but also storing reject parameters.

さらに、前記認識部5には、前記テンプレートリジェク
トパラメータ記憶部11に記憶された前記テンプレート
が読出される他に、前記リジェクトパラメータをも読出
す。
Further, in addition to reading out the template stored in the template rejection parameter storage unit 11, the recognition unit 5 also reads out the rejection parameters.

これにより、登録時には、リジェクトパラメータ決定部
7では、すべて類似度データ、及び、テンプレート数を
もとにしてリジェクトパラメータを決定する。例えば、
類似度がX以上の組合わせが全組合せのy%以上あれば
リジェクトパラメータはAにするように決定し、2%以
上であればリジェクトパラメータはBにするように決定
する。
As a result, at the time of registration, the reject parameter determining unit 7 determines reject parameters based entirely on the similarity data and the number of templates. for example,
If the number of combinations with a similarity of X or more is y% or more of all combinations, the reject parameter is determined to be A, and if it is 2% or more, the reject parameter is determined to be B.

このようにして決定されたリジェクトパラメータは、作
成されたテンプレートと共にテンプレートリジェクトパ
ラメータ記憶部11に送られ記憶される。
The reject parameters determined in this way are sent to the template reject parameter storage section 11 and stored together with the created template.

また、新たに違う単語群のテンプレートを作成する場合
や、新しい単語を追加する場合にも、上記と同様な動作
が行われ、リジェクトパラメータ決定部7で決定された
リジェクトパラメータは登録部4で作成された全テンプ
レートと共にテンプレートリジェクトパラメータ記憶部
11jこ送られて、上記テンプレート群が記憶されてい
る領域とは違う領域に記憶される。
Furthermore, when creating a new template for a different word group or when adding a new word, the same operation as above is performed, and the reject parameters determined by the reject parameter determining unit 7 are created by the registering unit 4. The template is sent to the template reject parameter storage unit 11j together with all the templates that have been created, and is stored in an area different from the area where the template group is stored.

上述したように、テンプレートを登録する時、全ての登
録が終わった時点で各テンプレート間の類似度を測定し
その類似度の大きさの分布やテンプレート数等をもとに
してリジェクトパラメータ値を決定する。そして、登録
したテンプレートをテンプレートリジェクトパラメータ
記憶部11に記憶させる時にリジェクトパラメータも同
時に記憶させるようにする。さらに、そのテンプレート
群をテンプレートリジェクトパラメータ記憶部11から
ロード(読出し)する時には、リジェクトパラメータも
同時にロードするようにする。このように動作させるこ
とによって、テンプレート群によってリジェクト率が変
動するのを防ぐことが可能となる。
As mentioned above, when registering templates, the degree of similarity between each template is measured when all registrations are completed, and the reject parameter value is determined based on the distribution of the degree of similarity, the number of templates, etc. do. Then, when the registered template is stored in the template reject parameter storage section 11, the reject parameters are also stored at the same time. Furthermore, when loading (reading) the template group from the template reject parameter storage section 11, the reject parameters are also loaded at the same time. By operating in this way, it is possible to prevent the rejection rate from varying depending on the template group.

発明の効果 請求項1記載の発明は、入力音声の特徴を抽出する音声
特徴抽出部を設け、この音声特徴抽出部により抽出され
た入力音声の特徴量を用いてテンプレートを作成する登
録部を設け、この登録部に登録された前記テンプレート
を記憶するテンプレート記憶部を設け、このテンプレー
ト記憶部に記憶されたテンプレート群を用いて各テンプ
レート間のマツチングを行いその類似度を計算する認識
部を設け、登録時に前記認識部により求められた前記類
似度の大きさに応じてリジェクトパラメータをテンプレ
ート毎に決定するりジエクトパラメータ決定部を設けた
ので、認識対象となるテンプレート毎にリジェクトパラ
メータを決定し、このパラメータを用いてリジェクトす
ることによって、単語によるリジェクト率の変動を防ぐ
ことができるものである。
Effects of the Invention The invention as set forth in claim 1 provides a voice feature extraction unit that extracts features of input voice, and a registration unit that creates a template using the feature amount of the input voice extracted by the voice feature extraction unit. , a template storage unit that stores the templates registered in the registration unit; a recognition unit that matches each template using the template group stored in the template storage unit and calculates the degree of similarity; Since a reject parameter determination unit is provided, a reject parameter is determined for each template according to the degree of similarity determined by the recognition unit at the time of registration, and a reject parameter is determined for each template to be recognized. By rejecting using this parameter, it is possible to prevent fluctuations in the rejection rate depending on words.

請求項2記載の発明は、入力音声の特徴を抽出する音声
特徴抽出部を設け、この音声特徴抽出部により抽出され
た入力音声の特徴量を用いてテンプレートを作成する登
録部を設け、この登録部に登録された前記テンプレート
を記憶するテンプレート記憶部を設け、このテンプレー
ト記憶部に記憶されたテンプレート群を用いて各テンプ
レート間のマツチングを行いその類似度を計算する認識
部を設け、登録時に前記認識部により求められた前記類
似度の大きさに応じてリジェクトパラメータをテンプレ
ートの組合せ毎に決定するりジエクトパラメータ決定部
を設けたので、認識対象となるテンプレートの組合わせ
毎にリジェクトパラメータを決定し、このパラメータを
用いてリジェクトすることによって、単語によるリジェ
クト率の変動を防ぐことができるものである。
The invention according to claim 2 provides a voice feature extraction section that extracts the features of the input voice, and a registration section that creates a template using the feature amount of the input voice extracted by the voice feature extraction section. A template storage unit is provided to store the templates registered in the template storage unit, and a recognition unit is provided to match each template using the template group stored in the template storage unit and calculate the degree of similarity. A reject parameter is determined for each combination of templates according to the degree of similarity determined by the recognition unit, and a reject parameter determination unit is provided, so that a reject parameter is determined for each combination of templates to be recognized. However, by rejecting using this parameter, it is possible to prevent fluctuations in the rejection rate depending on the word.

請求項3記載の発明は、入力音声の特徴を抽出する音声
特徴抽出部を設け、この音声特徴抽出部により抽出され
た入力音声の特徴量を用いてテンプレートを作成する登
録部を設け、この登録部に   ′登録された前記テン
プレートを記憶するテンプレート記憶部を設け、このテ
ンプレート記憶部に記憶されたテンプレート群を用いて
各テンプレート間のマツチングを行いその類似度を計算
する認識部を設け、登録時に前記認識部により求められ
た前記類似度の大きさに応じてリジェクトパラメータを
決定し以後前記テンプレートを追加する毎にそのリジェ
クトパラメータの値を更新していくリジェクトパラメー
タ決定部を設けたので、認識対象となるテンプレート群
によってリジェクトパラメータを決定し、このパラメー
タを用いてリジェクトすることによって、テンプレート
群によるリジェクト率の変動を防ぐことができるもので
ある。
The invention according to claim 3 provides a voice feature extraction section that extracts the features of the input voice, a registration section that creates a template using the feature amount of the input voice extracted by the voice feature extraction section, and The section is provided with a template storage section that stores the registered templates, a recognition section that matches each template using the template group stored in the template storage section and calculates the degree of similarity; A reject parameter determination unit is provided which determines a reject parameter according to the degree of similarity determined by the recognition unit and updates the value of the reject parameter each time the template is added. By determining a rejection parameter based on a group of templates and rejecting using this parameter, it is possible to prevent fluctuations in the rejection rate due to the group of templates.

請求項4記載の発明は、入力音声の特徴を抽出する音声
特徴抽出部を設け、この音声特徴抽出部により抽出され
た入力音声の特徴量を用いてテンプレートを作成する登
録部を設け、この登録部に登録された前記テンプレート
及びリジェクトパラメータを記憶するテンプレート記憶
部を設け、このテンプレート記憶部に記憶された前記リ
ジェクトパラメータ及び前記テンプレートを読出すと共
にテンプレート群を用いて各テンプレート間のマツチン
グを行いその類似度を計算する認識部を設け、登録時に
前記認識部により求められた前記類似度の大きさに応じ
てリジェクトパラメータを決定し以後前記テンプレート
を追加する毎に前記リジェクトパラメータの値を更新し
ていくリジェクトパラメータ決定部を設けたので、認識
対象となるテンプレート群によってリジェクトパラメー
タを決定し、このパラメータを用いてリジェクトするこ
とによって、テンプレート群によるリジェクト率の変動
を防ぐことができるものである。
The invention according to claim 4 provides a voice feature extraction section that extracts the features of the input voice, and a registration section that creates a template using the feature amount of the input voice extracted by the voice feature extraction section. A template storage section is provided for storing the templates and reject parameters registered in the template storage section, the reject parameters and the templates stored in the template storage section are read out, and the templates are matched using a group of templates. A recognition unit that calculates the degree of similarity is provided, a reject parameter is determined according to the degree of similarity determined by the recognition unit at the time of registration, and the value of the reject parameter is updated every time the template is added thereafter. Since several reject parameter determining units are provided, reject parameters are determined according to the group of templates to be recognized, and rejection is performed using these parameters, thereby making it possible to prevent fluctuations in the rejection rate due to the group of templates.

【図面の簡単な説明】[Brief explanation of drawings]

第1図は請求項1,2記載の発明の一実施例を示す回路
図、第2図は請求項3,4記載の発明の一実施例を示す
回路図である。 2・・・音声特徴抽出部、4・・・登録部、5・・・認
識部、6・・・テンプレート記憶部、7・・・リジェク
トパラメータ記憶部、11・・・テンプレートリジェク
トバラメータ記憶部 3,1 図
FIG. 1 is a circuit diagram showing an embodiment of the invention as claimed in claims 1 and 2, and FIG. 2 is a circuit diagram showing an embodiment of the invention as claimed in claims 3 and 4. 2... Audio feature extraction section, 4... Registration section, 5... Recognition section, 6... Template storage section, 7... Reject parameter storage section, 11... Template rejection parameter storage section 3 ,1 Figure

Claims (1)

【特許請求の範囲】 1、入力音声の特徴を抽出する音声特徴抽出部を設け、
この音声特徴抽出部により抽出された入力音声の特徴量
を用いてテンプレートを作成する登録部を設け、この登
録部に登録された前記テンプレートを記憶するテンプレ
ート記憶部を設け、このテンプレート記憶部に記憶され
たテンプレート群を用いて各テンプレート間のマッチン
グを行いその類似度を計算する認識部を設け、登録時に
前記認識部により求められた前記類似度の大きさに応じ
てリジェクトパラメータをテンプレート毎に決定するリ
ジェクトパラメータ決定部を設けたことを特徴とする音
声認識装置。 2、入力音声の特徴を抽出する音声特徴抽出部を設け、
この音声特徴抽出部により抽出された入力音声の特徴量
を用いてテンプレートを作成する登録部を設け、この登
録部に登録された前記テンプレートを記憶するテンプレ
ート記憶部を設け、このテンプレート記憶部に記憶され
たテンプレート群を用いて各テンプレート間のマッチン
グを行いその類似度を計算する認識部を設け、登録時に
前記認識部により求められた前記類似度の大きさに応じ
てリジェクトパラメータをテンプレートの組合せ毎に決
定するリジェクトパラメータ決定部を設けたことを特徴
とする音声認識装置。 3、入力音声の特徴を抽出する音声特徴抽出部を設け、
この音声特徴抽出部により抽出された入力音声の特徴量
を用いてテンプレートを作成する登録部を設け、この登
録部に登録された前記テンプレートを記憶するテンプレ
ート記憶部を設け、このテンプレート記憶部に記憶され
たテンプレート群を用いて各テンプレート間のマッチン
グを行いその類似度を計算する認識部を設け、登録時に
前記認識部により求められた前記類似度の大きさに応じ
てリジェクトパラメータを決定し以後前記テンプレート
を追加する毎にそのリジェクトパラメータの値を更新し
ていくリジェクトパラメータ決定部を設けたことを特徴
とする音声認識装置。 4、入力音声の特徴を抽出する音声特徴抽出部を設け、
この音声特徴抽出部により抽出された入力音声の特徴量
を用いてテンプレートを作成する登録部を設け、この登
録部に登録された前記テンプレート及びリジェクトパラ
メータを記憶するテンプレートリジェクトパラメータ記
憶部を設け、このテンプレートリジェクトパラメータ記
憶部に記憶された前記リジェクトパラメータ及び前記テ
ンプレートを読出すと共にテンプレート群を用いて各テ
ンプレート間のマッチングを行いその類似度を計算する
認識部を設け、登録時に前記認識部により求められた前
記類似度の大きさに応じてリジェクトパラメータを決定
し以後前記テンプレートを追加する毎に前記リジェクト
パラメータの値を更新していくリジェクトパラメータ決
定部を設けたことを特徴とする音声認識装置。
[Claims] 1. A voice feature extraction unit for extracting features of input voice is provided,
A registration unit is provided for creating a template using the features of the input voice extracted by the audio feature extraction unit, a template storage unit is provided for storing the template registered in the registration unit, and the template is stored in the template storage unit. A recognition unit is provided that performs matching between each template using the template group and calculates the degree of similarity, and a rejection parameter is determined for each template according to the degree of similarity determined by the recognition unit at the time of registration. What is claimed is: 1. A speech recognition device comprising: a rejection parameter determination section for determining a rejection parameter. 2. Provide a voice feature extraction unit that extracts the features of input voice,
A registration unit is provided for creating a template using the features of the input voice extracted by the audio feature extraction unit, a template storage unit is provided for storing the template registered in the registration unit, and the template is stored in the template storage unit. A recognition unit is provided that performs matching between each template using the template group and calculates the degree of similarity, and a rejection parameter is set for each combination of templates according to the degree of similarity determined by the recognition unit at the time of registration. 1. A speech recognition device comprising a reject parameter determination unit that determines a rejection parameter. 3. Provide a voice feature extraction unit that extracts the features of input voice,
A registration unit is provided for creating a template using the features of the input voice extracted by the audio feature extraction unit, a template storage unit is provided for storing the template registered in the registration unit, and the template is stored in the template storage unit. A recognition unit is provided that performs matching between templates using the template group and calculates the degree of similarity, and a rejection parameter is determined according to the degree of similarity determined by the recognition unit at the time of registration. A speech recognition device comprising a reject parameter determining unit that updates the value of a reject parameter each time a template is added. 4. Provide a voice feature extraction unit that extracts the features of input voice,
A registration unit is provided for creating a template using the features of the input voice extracted by the audio feature extraction unit, and a template reject parameter storage unit is provided for storing the template and reject parameters registered in the registration unit. A recognition unit is provided that reads out the reject parameters and the templates stored in the template rejection parameter storage unit, performs matching between templates using a template group, and calculates the degree of similarity. A speech recognition device comprising: a reject parameter determining unit that determines a reject parameter according to the magnitude of the similarity, and updates the value of the reject parameter each time the template is added thereafter.
JP2280299A 1990-10-18 1990-10-18 Voice recognition device Pending JPH04155398A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP2280299A JPH04155398A (en) 1990-10-18 1990-10-18 Voice recognition device

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP2280299A JPH04155398A (en) 1990-10-18 1990-10-18 Voice recognition device

Publications (1)

Publication Number Publication Date
JPH04155398A true JPH04155398A (en) 1992-05-28

Family

ID=17623052

Family Applications (1)

Application Number Title Priority Date Filing Date
JP2280299A Pending JPH04155398A (en) 1990-10-18 1990-10-18 Voice recognition device

Country Status (1)

Country Link
JP (1) JPH04155398A (en)

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH08297500A (en) * 1995-02-28 1996-11-12 Meidensha Corp Processing method for making erroneous recognition impossible to occur for discrete word speech recognition system

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH08297500A (en) * 1995-02-28 1996-11-12 Meidensha Corp Processing method for making erroneous recognition impossible to occur for discrete word speech recognition system

Similar Documents

Publication Publication Date Title
CN107591155B (en) Voice recognition method and device, terminal and computer readable storage medium
JPS6332394B2 (en)
KR100373989B1 (en) Method for identification of user using of cognition by syllable and its system
JPS6226038B2 (en)
JPS6361300A (en) Voice recognition system
JPH03274598A (en) Voice recognition device
JPS599080B2 (en) Voice recognition method
KR20010026288A (en) Voice recognition apparatus depended on speakers for being based on the phoneme like unit and the method
JPS58159598A (en) Monosyllabic voice recognition system
JPS6011897A (en) Voice recognition equipment
JPS5934597A (en) Voice recognition processor
JPS5968792A (en) Voice recognition equipment
JPH01209499A (en) Pattern matching method
JPS59124388A (en) Word voice recognition processing system
JPH08160986A (en) Voice recognition device
JPS6347797A (en) Word voice preselection system
JPH03155599A (en) Speech recognition device
JPS62123585A (en) Pattern recognizing device
JPS60147797A (en) Voice recognition equipment
JPS58159591A (en) Monosyllabic voice recognition system
JPS62103699A (en) voice recognition device
JPS63116199A (en) Voice dictionary storing system for voice input/output unit
JPS59125799A (en) Voice recognition equipment
JPH036698A (en) Cash register that tells received sum in voice
JPS59121398A (en) Voice recognition system