JPH02124599A - Continuously voiced speech recognition system - Google Patents
Continuously voiced speech recognition systemInfo
- Publication number
- JPH02124599A JPH02124599A JP63277771A JP27777188A JPH02124599A JP H02124599 A JPH02124599 A JP H02124599A JP 63277771 A JP63277771 A JP 63277771A JP 27777188 A JP27777188 A JP 27777188A JP H02124599 A JPH02124599 A JP H02124599A
- Authority
- JP
- Japan
- Prior art keywords
- voice
- recognition
- word
- speech
- corrected
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Pending
Links
Abstract
Description
【発明の詳細な説明】
[産業上の利用分野]
この発明は連続発声音声を認識する音声認識装置におい
て特に連続発声音声の誤認識時に部分訂正を行うための
連続発声音声認識方式に関するものである。[Detailed Description of the Invention] [Industrial Application Field] The present invention relates to a continuous utterance recognition method for performing partial correction when erroneously recognizing continuous utterance in a speech recognition device that recognizes continuous utterance. .
[従来の技術]
第4図は従来の連続発声音声認識方式を採用した連続発
声音声認識装置の構成ブロック図である。図において、
1は単語か登録された単語辞書、2は音声を入力するマ
イクロフォンで実現される音声入力手段、3は音声入力
手段2から入力された音声を単語辞書1に基づいて認識
する音声認識手段、4は音声認識手段3で認識された結
果を出力する認識結果出力手段である。認識結果出力手
段4はイヤホーン、スピーカ、CRT等で実現されるが
、ここではイヤホーン等を用い音声出力するものとする
。[Prior Art] FIG. 4 is a block diagram showing the configuration of a continuous speech recognition device that employs a conventional continuous speech recognition method. In the figure,
1 is a word dictionary in which words are registered; 2 is a voice input means realized by a microphone for inputting voice; 3 is a voice recognition means for recognizing the voice input from the voice input means 2 based on the word dictionary 1; 4 is a recognition result output means for outputting the result recognized by the voice recognition means 3. The recognition result output means 4 is realized by an earphone, a speaker, a CRT, etc., but here, it is assumed that the earphone or the like is used to output audio.
次に第5図に示すフローチャートを参照して従来方式に
ついて説明する。オペレータからの連続発声音声は、音
声入力手段2に入力され、電気信号に変換される(ステ
ップSL)。音声認識手段3は、その電気信号をデジタ
ル信号に変換し、単語辞書1から該当する単語を取り出
し、入力音声を認識する(ステップS2)。この認識処
理の後、認識結果をオペレータに確認させるため、音声
認識手段3は認識結果を認識結果出力手段4に与え、オ
ペレータに認識結果を連続的に音声で知らせる(ステッ
プS3)。即ち、認識結果出力手段4からは入力した連
続発声音声を認識した結果の音声が発声するが、その認
識結果が正しければ(ステップS4)処理は終了し、入
力連続発声音声と認識結果発声音声とが一部でも異なれ
ば、音声認識手段3が誤認識したことになる。したがっ
て、オペレータは上記のような誤認識があった時、再度
、連続発声音声を入力し直すという処理を行う。Next, the conventional method will be explained with reference to the flowchart shown in FIG. Continuously uttered voices from the operator are input to the voice input means 2 and converted into electrical signals (step SL). The speech recognition means 3 converts the electric signal into a digital signal, extracts the corresponding word from the word dictionary 1, and recognizes the input speech (step S2). After this recognition process, in order to have the operator confirm the recognition result, the voice recognition means 3 provides the recognition result to the recognition result output means 4, and continuously informs the operator of the recognition result by voice (step S3). That is, the recognition result output means 4 outputs the voice as a result of recognizing the input continuous voice, but if the recognition result is correct (step S4), the process ends and the input continuous voice and the recognition result voice are output. If there is even a partial difference, it means that the voice recognition means 3 has misrecognized the voice. Therefore, when the above-mentioned erroneous recognition occurs, the operator performs a process of inputting the continuous voice again.
[発明が解決しようとする課題]
従来の連続発声音声認識方式は上述したように処理する
ので、連続したN個の単語のうち、1つでも誤認識すれ
ば、認識処理を全部やり直し、誤認識がなくなるまで連
続発声音声の入力処理を繰り返す必要があり、認識処理
の効率を低下させるという問題点かあった。[Problem to be solved by the invention] Since the conventional continuous speech recognition method processes as described above, if even one of N consecutive words is misrecognized, the entire recognition process is redone and the misrecognition is eliminated. It is necessary to repeat the input process of continuous utterances until all the utterances are exhausted, which poses a problem of reducing the efficiency of the recognition process.
この発明は上記にような問題点を解消するためになされ
たもので、誤認識があった場合、誤認識の単語のみ認識
処理をやり直すことにより、訂正入力時の誤認識率を少
なくし、認識処理の効率を高めることができる連続発声
音声認識方式を提供することを目的とする。This invention was made to solve the above-mentioned problems, and when there is a misrecognition, by redoing the recognition process only for the misrecognized word, the misrecognition rate at the time of corrected input is reduced, and the recognition process is improved. The purpose of the present invention is to provide a continuous utterance speech recognition method that can improve processing efficiency.
[8題を解決するための手段]
この発明に係る連続発声音声認識方式は、音声入力手段
2から入力された音声のパワーを計算し、訂正する単語
と訂正しない単語との音声パワー差が一定値を越えると
訂正信号を出力する音声パワー計算手段5を備え、音声
認識時に誤認識された単語が含まれていた場合、その単
語を訂正するため、訂正する単語の音声パワーが訂正し
ない他の単語の音声パワーより大きく、かつ上記一定値
を越えるように音声入力手段2から再び音声入力し、音
声パワー計算手段5から訂正信号を出力させ、この訂正
信号を受けた音声認識手段3により単語を単語辞書1に
従って訂正し、認識結果出力手段4に正しい認識結果を
出力することを特徴とするものである。[Means for Solving 8 Problems] The continuous utterance speech recognition method according to the present invention calculates the power of the speech input from the speech input means 2, and calculates the speech power difference between the word to be corrected and the word not to be corrected to be constant. It is equipped with a voice power calculation means 5 that outputs a correction signal when the value exceeds the value, and when a word that is incorrectly recognized during speech recognition is included, in order to correct that word, the voice power of the word to be corrected is calculated by other words that are not corrected. The voice input means 2 inputs the voice again so that the voice power is higher than the voice power of the word and exceeds the above-mentioned fixed value, the voice power calculation means 5 outputs a correction signal, and the voice recognition means 3 which receives this correction signal recognizes the word. It is characterized in that it is corrected according to the word dictionary 1 and outputs the correct recognition result to the recognition result output means 4.
[作用]
この連続発声音声認識方式において、音声認識手段3が
ある単語を誤認識し、この誤認識の単語を含む認識結果
が認識結果出力手段4から出力されると、オペレータは
その単語を訂正するため次のような処理を行う。即ち、
訂正する単語の音声パワーが訂正しない他の単語の音声
パワーより大きく、かつ一定値を越えるように音声入力
手段2から再び音声を入力する。これにより、音声パワ
ー計算手段5からは訂正信号が出力され、音声認識手段
3は単語辞書1から訂正後の単語を取り出して認識する
。したがって、認識結果出力手段4からは訂正された単
語を含む正しい認識結果か出力される。[Operation] In this continuous utterance speech recognition method, when the speech recognition means 3 misrecognizes a certain word and the recognition result output means 4 outputs a recognition result including the misrecognized word, the operator corrects the word. To do this, perform the following processing. That is,
Speech is input again from the speech input means 2 so that the speech power of the word to be corrected is greater than the speech power of other words not to be corrected and exceeds a certain value. As a result, the speech power calculation means 5 outputs a correction signal, and the speech recognition means 3 extracts the corrected word from the word dictionary 1 and recognizes it. Therefore, the recognition result output means 4 outputs a correct recognition result including the corrected word.
[発明の実施例]
第1図はこの発明の一実施例に係る連続発声音声認識方
式を採用した連続発声音声認識装置の構成ブロック図で
ある。第1図において、第4図に示す構成要素に対応す
るものには同一の符号を付し、その説明を省略する。第
1図において、5は音声入力手段2から入力された音声
のパワーを計算し、音声パワー差(訂正する単語の音声
パワーと訂正しない単語の音声パワーとの差)が一定値
を越えると訂正信号を出力する音声パワー計算手段であ
る。[Embodiment of the Invention] FIG. 1 is a block diagram showing the configuration of a continuous utterance speech recognition apparatus employing a continuous utterance speech recognition method according to an embodiment of the present invention. In FIG. 1, components corresponding to those shown in FIG. 4 are designated by the same reference numerals, and their explanations will be omitted. In FIG. 1, 5 calculates the power of the voice input from the voice input means 2, and corrects it if the voice power difference (the difference between the voice power of the word to be corrected and the voice power of the word not to be corrected) exceeds a certain value. This is an audio power calculation means that outputs a signal.
次に第2図に示すフローチャー1・を参照してこの実施
例の動作について説明する。音声入力手段2はオペレー
タからの連続発声音声を入力して電気信号に変換する(
ステップNl>。音声認識手段3は、その電気信号をデ
ジタル信号に変換し、単語辞書1から該当する単語を取
り出し、入力音声を認識する(ステップN2)。一方、
音声パワ−計算手段5はステップN1で入力された連続
発声音声の各単語に対応する音声パワーを計算する(ス
テップN3)。音声認識手段3の認識結果を音声応答な
どでオペレータに確認させるため、認識結果出力手段4
は認識結果を連続的に出力する(ステップN4)。とこ
ろが、その認識結果は正しくなく(ステップN5)、そ
の認識結果に誤認識された単語が含まれていた時、オペ
レータは連続単語発声音声の再入力を行う(ステップN
6)。即ち、オペレータは再入力時に誤認識単語(訂正
する単語)のみを大きく(前記一定値を越えるように)
発声し、その他の単語(訂正しない単語)は最初の入力
時と同程度の大きさで発声する。この発声音声は、上記
と同様に音声認識手段3に入力され、認識処理される(
ステップN7)。また、この発声音声は音声パワー計算
手段5に入力され、音声パワーが計算される(ステップ
N8)。即ち、音声パワー計算手段5では入力一定値を
越えた単語を単語辞書1から取り出して認識し、最初に
誤認識した単語に代わって正しい単語を置き換える〈ス
テップN9>。その後、ステップN4へ戻り、認識結果
を連続的に出力し、オペレータは再度確認して認識結果
が正しければ(ステップN5)、この処理は終了する。Next, the operation of this embodiment will be explained with reference to flowchart 1 shown in FIG. The voice input means 2 inputs continuous voice from the operator and converts it into an electrical signal (
Step Nl>. The speech recognition means 3 converts the electric signal into a digital signal, extracts the corresponding word from the word dictionary 1, and recognizes the input speech (step N2). on the other hand,
The voice power calculation means 5 calculates the voice power corresponding to each word of the continuous utterance inputted in step N1 (step N3). In order to let the operator confirm the recognition result of the voice recognition means 3 by voice response etc., the recognition result output means 4 is used.
continuously outputs the recognition results (step N4). However, the recognition result is incorrect (step N5), and when the recognition result includes the incorrectly recognized word, the operator re-inputs the continuous word utterance audio (step N5).
6). In other words, when re-entering, the operator increases only the incorrectly recognized word (word to be corrected) (so as to exceed the above-mentioned certain value).
Other words (words that are not corrected) are uttered at the same loudness as when they were first input. This uttered voice is input to the voice recognition means 3 in the same manner as above, and is subjected to recognition processing (
Step N7). Further, this uttered voice is input to the voice power calculation means 5, and the voice power is calculated (step N8). That is, the voice power calculation means 5 extracts words whose input value exceeds a certain input value from the word dictionary 1, recognizes them, and replaces the initially erroneously recognized word with a correct word (step N9). Thereafter, the process returns to step N4, where the recognition results are continuously output, and the operator checks again, and if the recognition results are correct (step N5), this process ends.
第3図(1)〜(6)はこの実施例の動作を説明するた
めの信号波形図である。第3図(1)は最初の音声発声
信号2a、5a、la、3aを示し、第3図(2)は最
初の認識結果2a、5a、8a、3aを示2a、5a、
la、3aを示し、信号1aの音声パワーを他の信号2
a、5a、3aのそれよりも大きくしている。第3図(
4)は訂正時の仮の認識結果2a、5a、la、4aを
示す。この認識結果の1aは訂正されているが、4aは
誤認識されている。しかし、最初の認識結果で3aと正
しく認識され、音声認識手段3に設定されているので、
3aから4aへ変更されることはない。第3図(5)は
最初の入力時の音声パワー(破線)と訂正入力時の音声
パワー(実線)とを示す。第3図(6)は訂正時の認識
結果2a、5a、la、3aを示す。FIGS. 3(1) to 3(6) are signal waveform diagrams for explaining the operation of this embodiment. FIG. 3(1) shows the first voice utterance signals 2a, 5a, la, 3a, and FIG. 3(2) shows the first recognition results 2a, 5a, 8a, 3a 2a, 5a,
la, 3a, and the audio power of signal 1a is expressed as the other signal 2.
It is larger than that of a, 5a, and 3a. Figure 3 (
4) shows provisional recognition results 2a, 5a, la, and 4a at the time of correction. In this recognition result, 1a has been corrected, but 4a has been misrecognized. However, the first recognition result correctly recognized it as 3a and it was set to voice recognition means 3, so
There is no change from 3a to 4a. FIG. 3(5) shows the audio power at the time of initial input (broken line) and the audio power at the time of corrected input (solid line). FIG. 3(6) shows recognition results 2a, 5a, la, and 3a at the time of correction.
このように最初に誤認識された単語を含む認識結果であ
っても、訂正する単語の音声パワーのみ大きくして入力
すれば、正しい認識結果を得ることができる。Even if the recognition result includes a word that was initially misrecognized in this way, the correct recognition result can be obtained by increasing the voice power of the word to be corrected and inputting it.
[発明の効果]
以上のように本発明によれば、音声認識時に誤認識され
た単語が含まれていた場合、その単語を訂正するため、
訂正する単語の音声パワーが訂正しない他の単語の音声
パワーより大きく、かつ−定値を越えるように音声入力
手段から再び音声入力し、音声パワー計算手段から訂正
信号を出力させ、この訂正信号を受けた音声認識手段に
より単語を単語辞書に従って訂正し、認識結果出力手段
に正しい認識結果を出力するようにしたので、誤認識が
あった場合、誤認識の単語のみ認識処理が行われ、これ
により訂正入力時の誤認識率が少なくなり、部分訂正を
精度良く行うことができ、柔したがって認識処理効率の
向上を図れるという効果がイ岑られる。[Effects of the Invention] As described above, according to the present invention, when a word that is incorrectly recognized during speech recognition is included, in order to correct the word,
Input the voice again from the voice input means so that the voice power of the word to be corrected is larger than the voice power of other words to be corrected and exceeds a fixed value, output a correction signal from the voice power calculation means, and receive the correction signal. Since the words are corrected by the voice recognition means according to the word dictionary and the correct recognition result is output to the recognition result output means, if there is an erroneous recognition, only the erroneously recognized word is recognized and corrected. The effect is that the rate of recognition errors during input is reduced, partial correction can be performed with high accuracy, flexibility and recognition processing efficiency can be improved.
第1図はこの発明の一実施例に1系る連続発声音声認識
方式を採用した連続発声音声認識装置の構成ブロック図
、第2図はこの実施例の動作を説明するだめのフローチ
ャー1・、第3図(1)〜(6)はこの実施例の動1ヤ
を説明するための信号波形図、第4図は従来の連続発声
音声認識方式を採用した連続発声音声認識装置の構成ブ
ロック図、第5I2Iはこの従来例の動作を説明するた
めのフローチャー1・である。
1・・・・・・単語辞書、2・・・・・・音声入力手段
、3・・・・・音声認識手段、4・・・・・・認識結果
出力手段、5・・・・・・音声パワー計算手段。FIG. 1 is a block diagram of the configuration of a continuous speech recognition device that employs a continuous speech recognition system according to an embodiment of the present invention, and FIG. 2 is a flowchart 1 for explaining the operation of this embodiment. , Fig. 3 (1) to (6) are signal waveform diagrams for explaining the operation of this embodiment, and Fig. 4 is a structural block diagram of a continuous utterance speech recognition device employing a conventional continuous utterance speech recognition method. Figure 5I2I is a flowchart 1 for explaining the operation of this conventional example. 1... Word dictionary, 2... Voice input means, 3... Voice recognition means, 4... Recognition result output means, 5...... Voice power calculation means.
Claims (1)
力手段と、この音声入力手段から入力された音声を上記
単語辞書に基づいて認識する音声認識手段と、この音声
認識手段で認識された結果を出力する認識結果出力手段
とを備え、連続発声された音声を認識する連続発声音声
認識装置において、上記音声入力手段から入力された音
声のパワーを計算し、訂正する単語と訂正しない単語と
の音声パワー差が一定値を越えると訂正信号を出力する
音声パワー計算手段を設け、音声認識時に誤認識された
単語が含まれていた場合、その単語を訂正するため、訂
正する単語の音声パワーが訂正しない他の単語の音声パ
ワーより大きく、かつ上記一定値を越えるように上記音
声入力手段から再び音声入力し、上記音声パワー計算手
段から訂正信号を出力させ、この訂正信号を受けた上記
音声認識手段により単語を上記単語辞書に従って訂正し
、上記認識結果出力手段に正しい認識結果を出力するこ
とを特徴とする連続発声音声認識方式。A word dictionary in which words are registered, a voice input means for inputting voice, a voice recognition means for recognizing the voice input from the voice input means based on the word dictionary, and a result recognized by the voice recognition means. A continuous speech recognition device that recognizes continuously uttered speech includes a recognition result output means that outputs a recognition result output means, which calculates the power of the speech input from the speech input means and distinguishes between words to be corrected and words not to be corrected. A voice power calculation means that outputs a correction signal when the voice power difference exceeds a certain value is provided, and if a word that is incorrectly recognized during speech recognition is included, the voice power of the word to be corrected is adjusted to correct the word. The speech input means again inputs speech so that the speech power is greater than the speech power of other words to be corrected and exceeds the certain value, the speech power calculation means outputs a correction signal, and the speech recognition receives this correction signal. A continuous utterance speech recognition method, characterized in that the means corrects words according to the word dictionary and outputs correct recognition results to the recognition result output means.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP63277771A JPH02124599A (en) | 1988-11-02 | 1988-11-02 | Continuously voiced speech recognition system |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP63277771A JPH02124599A (en) | 1988-11-02 | 1988-11-02 | Continuously voiced speech recognition system |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| JPH02124599A true JPH02124599A (en) | 1990-05-11 |
Family
ID=17588098
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP63277771A Pending JPH02124599A (en) | 1988-11-02 | 1988-11-02 | Continuously voiced speech recognition system |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPH02124599A (en) |
-
1988
- 1988-11-02 JP JP63277771A patent/JPH02124599A/en active Pending
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| JPS63301998A (en) | Voice recognition responder | |
| JP3069531B2 (en) | Voice recognition method | |
| KR930011738B1 (en) | How to Correct End Point Error of Speech Recognition System | |
| JP2996019B2 (en) | Voice recognition device | |
| JP3437492B2 (en) | Voice recognition method and apparatus | |
| JP2001236091A (en) | Error correction method and apparatus for speech recognition result | |
| JP2585547B2 (en) | Method for correcting input voice in voice input / output device | |
| JP2712586B2 (en) | Pattern matching method for word speech recognition device | |
| JPH0754434B2 (en) | Voice recognizer | |
| JPS62255999A (en) | Word voice recognition equipment | |
| JPH02178699A (en) | Voice recognition device | |
| JPS59224900A (en) | Voice recognition method | |
| JPH10301595A (en) | Voice recognition and response device | |
| JPH05188998A (en) | Speech recognizing method | |
| JPS61238099A (en) | Word voice recognition equipment | |
| JPH0239199A (en) | Sound reference pattern registering system | |
| JPS61231629A (en) | Voice input device | |
| JPS59195739A (en) | Audio response unit | |
| JPH0449716B2 (en) | ||
| JP2999479B2 (en) | Dictionary update method for speech recognition device | |
| JPH01236000A (en) | Voice recognizing device | |
| JPS63279305A (en) | Voice controller | |
| JPH0158519B2 (en) | ||
| JPS59195299A (en) | Sepecific speaker's voice recognition equipment | |
| JPH02198499A (en) | Automatic update system for speech recognition device dictionary |