JPH11352995A - Voice recognition device - Google Patents
Voice recognition deviceInfo
- Publication number
- JPH11352995A JPH11352995A JP10158897A JP15889798A JPH11352995A JP H11352995 A JPH11352995 A JP H11352995A JP 10158897 A JP10158897 A JP 10158897A JP 15889798 A JP15889798 A JP 15889798A JP H11352995 A JPH11352995 A JP H11352995A
- Authority
- JP
- Japan
- Prior art keywords
- environment
- voice recognition
- voice
- recognition
- sound
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Abandoned
Links
Abstract
(57)【要約】
【課題】周囲の環境状態が音声認識に適切な環境か否か
を常に知らせることで音声認識率を高める。
【解決手段】話者の音声を認識して文字列情報を出力す
るとともに発声音と同時に外乱音を取込むと文字列スコ
ア値を算出して出力する音声認識手段1と、この音声認
識手段からの文字列情報を表示する表示装置2と、音声
認識手段からの文字列スコア値を受け取るとスコア値の
みを出力する外乱音環境検出手段3と、この外乱音環境
検出手段からのスコア値に基づいて環境が音声認識に適
切な環境か否かを判断し、判断結果を出力する外乱音環
境判断手段4と、この外乱音環境判断手段からの判断結
果に基づいて表示装置に外乱音レベルが音声認識可能な
範囲内にあるか否かを文字表示させる音声認識環境表示
制御手段5を備えている。
(57) [Summary] [Problem] To improve the speech recognition rate by always informing whether or not the surrounding environment state is appropriate for speech recognition. A speech recognition means for recognizing a speaker's voice, outputting character string information, and calculating and outputting a character string score value when a disturbance sound is taken in at the same time as a vocal sound. A display device 2 for displaying character string information, a disturbance sound environment detecting means 3 for outputting only a score value when a character string score value is received from the voice recognition means, and a score value from the disturbance sound environment detecting means. Means for determining whether or not the environment is appropriate for speech recognition, and outputting a result of the determination; and a display device for displaying the level of the disturbance sound on the display device based on the result of the determination from the means for determining the disturbance sound. There is provided a voice recognition environment display control means 5 for displaying a character indicating whether or not it is within a recognizable range.
Description
【0001】[0001]
【発明の属する技術分野】本発明は、入力される音声を
認識して文字列情報に変換し、表示する音声認識装置に
関する。BACKGROUND OF THE INVENTION 1. Field of the Invention The present invention relates to a speech recognition apparatus for recognizing input speech, converting the speech into character string information, and displaying the information.
【0002】[0002]
【従来の技術】従来の音声認識装置は、装置のセットア
ップ時に音声入力しない状態でのノイズレベルと音声入
力したときのノイズレベルを測定して設定し、その後
は、音声入力に対してこの設定したノイズレベルに基づ
いてノイズカット等の処理を行って音声認識を行うよう
になっていた。2. Description of the Related Art A conventional voice recognition apparatus measures and sets a noise level when no voice is input and a noise level when voice is input when setting up the apparatus, and thereafter sets the noise level for voice input. Conventionally, speech recognition is performed by performing processing such as noise cut based on the noise level.
【0003】[0003]
【発明が解決しようとする課題】しかしながら、このよ
うな音声認識装置を周囲の環境が時々刻々と変化するよ
うな場所で使用した場合、ノイズレベルが周囲の状況に
より変化し、装置のセットアップ時にのみノイズレベル
を測定して設定したのでは正確な音声認識ができなくな
り、認識率の低下を招くという問題があった。However, when such a voice recognition device is used in a place where the surrounding environment changes moment by moment, the noise level changes according to the surrounding conditions, and only when the device is set up. If the noise level is measured and set, accurate speech recognition cannot be performed, and there is a problem that the recognition rate is reduced.
【0004】各請求項記載の発明は、周囲の環境状態が
音声認識に適切な環境か否かを常にチェックして知らせ
ることができ、これにより話者は音声認識に適切な環境
のもとで音声入力ができ、認識率を高めることができる
音声認識装置を提供する。[0004] According to the invention described in each claim, it is possible to always check whether or not the surrounding environment state is an environment suitable for speech recognition, thereby informing a speaker under an environment suitable for speech recognition. Provided is a voice recognition device that can perform voice input and increase a recognition rate.
【0005】請求項4記載の発明は、さらに、話者に適
合する最適化情報に基づいて音声認識するので、認識率
をさらに高めることができる音声認識装置を提供する。
請求項5記載の発明は、さらに、音声の認識ができなか
ったときに音声に対応したエラー情報を表示でき、これ
により話者に対する発声時の注意を促すことができる音
声認識装置を提供する。[0005] The invention according to claim 4 further provides a speech recognition apparatus capable of further improving the recognition rate because speech recognition is performed based on optimization information suitable for a speaker.
The invention according to claim 5 further provides a voice recognition device that can display error information corresponding to voice when voice recognition cannot be performed, thereby prompting a speaker to pay attention when uttering.
【0006】[0006]
【課題を解決するための手段】請求項1記載の発明は、
入力する音声を認識して文字列情報を出力する音声認識
手段と、この音声認識手段からの文字列情報を表示する
表示装置と、音声認識に関する環境を検出する音声認識
環境検出手段と、この音声認識環境検出手段が検出した
環境情報により、今の環境が音声認識に適切な環境か不
適切な環境かを判断する音声認識環境判断手段と、この
音声認識環境判断手段が判断した結果を表示装置に表示
させる音声認識環境表示制御手段とを備えたものであ
る。According to the first aspect of the present invention,
Voice recognition means for recognizing input voice and outputting character string information, a display device for displaying character string information from the voice recognition means, voice recognition environment detection means for detecting an environment relating to voice recognition, Based on the environment information detected by the recognition environment detecting means, a voice recognition environment determining means for determining whether the current environment is suitable or inappropriate for voice recognition, and a display device for displaying a result determined by the voice recognition environment determining means. And a voice recognition environment display control means for displaying the information.
【0007】請求項2記載の発明は、請求項1記載の音
声認識装置において、音声認識環境検出手段として外乱
音のレベルを検出する外乱音環境検出手段を使用し、音
声認識環境判断手段として外乱音環境検出手段が検出し
た外乱音のレベルが音声認識可能な範囲内にあるか否か
を判断する外乱音環境判断手段を使用したものである。According to a second aspect of the present invention, in the voice recognition apparatus of the first aspect, a disturbance sound environment detecting means for detecting a level of disturbance sound is used as the sound recognition environment detecting means, and a disturbance sound detecting means is used as the sound recognition environment determining means. A disturbance sound environment judging means for judging whether or not the level of the disturbance sound detected by the sound environment detecting means is within a range in which speech can be recognized is used.
【0008】請求項3記載の発明は、請求項1記載の音
声認識装置において、音声認識環境検出手段として話者
の発生音量を検出する発生音量環境検出手段を使用し、
音声認識環境判断手段として発生音量環境検出手段が検
出した発生音量が音声認識に適切な音量であるか否かを
判断する音声認識環境判断手段を使用したものである。According to a third aspect of the present invention, in the voice recognition apparatus according to the first aspect, a generated sound volume environment detecting means for detecting a generated sound volume of a speaker is used as the voice recognition environment detecting means.
As the voice recognition environment determining means, a voice recognition environment determining means for determining whether or not the generated volume detected by the generated volume environment detecting means is appropriate for voice recognition is used.
【0009】請求項4記載の発明は、請求項1記載の音
声認識装置において、音声認識手段は、話者を認識する
ための複数の最適化情報を管理し、選択された音声入力
する話者に適合する最適化情報に基づいて音声認識する
ことにある。According to a fourth aspect of the present invention, in the voice recognition apparatus according to the first aspect, the voice recognition means manages a plurality of pieces of optimization information for recognizing the speaker, and the selected voice is input by the speaker. Speech recognition based on optimization information that conforms to
【0010】請求項5記載の発明は、請求項1記載の音
声認識装置において、音声認識手段は、入力する音声の
認識が不能のとき入力した音声に対応したエラー情報を
出力し、音声認識環境表示制御手段は、表示装置にエラ
ー情報を表示させることにある。According to a fifth aspect of the present invention, in the voice recognition device of the first aspect, the voice recognition means outputs error information corresponding to the input voice when the input voice cannot be recognized. The display control means is to display the error information on the display device.
【0011】[0011]
【発明の実施の形態】本発明の実施の形態を図面を参照
して説明する。なお、この実施の形態は音声認識装置を
物品販売の登録システムに適用した場合について述べ
る。Embodiments of the present invention will be described with reference to the drawings. This embodiment describes a case where the speech recognition apparatus is applied to an article sales registration system.
【0012】(第1の実施の形態)図1において、1は
入力する話者の音声を認識して文字列情報を出力する音
声認識手段、2はこの音声認識手段1からの文字列情報
を表示する表示装置である。すなわち、前記音声認識手
段1は、音声認識開始命令が入力された後、ある一定音
量レベル以上の音声が入力されたとき始めて音声認識処
理を開始し、音声入力が商品についてのものであれば図
3に示すような文字列情報を出力する。例えば、話者が
“商品A”と発声すると音声認識装置1はそれをマイク
ロホンから取込んで認識し、商品Aの文字列情報を表示
装置2に出力し、表示装置2は商品Aの文字列情報を表
示することになる。(First Embodiment) In FIG. 1, reference numeral 1 denotes a voice recognition means for recognizing an input speaker's voice and outputs character string information, and 2 denotes a character string information from the voice recognition means 1. It is a display device for displaying. That is, after the voice recognition start command is input, the voice recognition means 1 starts the voice recognition process only when a voice having a certain volume level or higher is input. The character string information as shown in FIG. For example, when the speaker utters “product A”, the speech recognition device 1 takes in the recognition from the microphone and recognizes it, outputs character string information of the product A to the display device 2, and displays the character string of the product A Information will be displayed.
【0013】また、前記音声認識手段1は発声音と同時
に外乱音も同時に取込むことになる。外乱音としては、
例えば、チャイムの音や周囲の人の会話の音等がある。
そして、前記音声認識手段1は、発声音と同時に外乱音
を取込むと、文字列スコア値を算出し、これを音声認識
環境手段である外乱音環境検出手段3に供給している。
なお、文字列スコア値とは、音声認識における確からし
さを%で示す値で、例えば、ある商品について音声入力
があったとき、図2に示すように、その音声入力に対す
る商品Aのスコア値が50%、商品Kのスコア値が40
%というように表わされ、スコア値が高いほどその商品
を示している確立が高いことを意味している。In addition, the voice recognition means 1 takes in a disturbance sound simultaneously with the utterance sound. As disturbance noise,
For example, there is a sound of a chime or a sound of a conversation of a surrounding person.
When the speech recognition unit 1 captures a disturbance sound at the same time as the utterance sound, the speech recognition unit 1 calculates a character string score value and supplies the character string score value to the disturbance sound environment detection unit 3 which is a speech recognition environment unit.
Note that the character string score value is a value indicating the likelihood of the speech recognition in%. For example, when there is a voice input for a certain product, as shown in FIG. 50%, product K score value is 40
%, The higher the score value, the higher the probability of indicating the product.
【0014】前記外乱音環境検出手段3は、音声認識手
段1から文字列スコア値を受け取ると、%の値であるス
コア値のみを切り取り、これを音声認識環境判断手段と
しての外乱音環境判断手段4に供給している。前記外乱
音環境判断手段4は、入力するスコア値から、今の環境
が音声認識に適切な環境か不適切な環境かを判断し、そ
の判断結果を音声認識環境表示制御手段5に供給してい
る。When the disturbance sound environment detection means 3 receives the character string score value from the speech recognition means 1, it cuts out only the score value which is a value of%, and uses it as a disturbance sound environment judgment means as a speech recognition environment judgment means. 4 The disturbance sound environment determination means 4 determines whether the current environment is appropriate or inappropriate for speech recognition from the input score value, and supplies the determination result to the speech recognition environment display control means 5. I have.
【0015】前記音声認識環境表示制御手段5は、前記
外乱音環境判断手段4からの判断結果に基づいて外乱音
レベルが音声認識可能な範囲内にあるか範囲外にあるか
をテキストとして前記表示装置2に出力し、文字表示さ
せるようになっている。The voice recognition environment display control means 5 displays the text as to whether the level of the disturbance sound is within the range in which the voice can be recognized or not, based on the judgment result from the disturbance sound environment judgment means 4 as text. The data is output to the device 2 and displayed as characters.
【0016】例えば、外乱音のレベルが高く音声認識で
きる環境の範囲外にあるときには、音声認識手段1が外
乱音を取込むと、この音量は一定レベル以上になってい
るので、音声認識手段1は音声認識処理を開始する。そ
して、図4の(a) に示すように、先ず、S1にて、音声
認識手段1から外乱音環境検出手段3に文字列スコア値
が送出される。このときの文字列スコア値は、商品C:
38%、商品K:40%、商品A:50%のように低い
値になっている。For example, if the level of the disturbance sound is outside the range of the environment in which speech recognition can be performed at a high level, and when the speech recognition unit 1 captures the disturbance sound, the volume is higher than a certain level. Starts voice recognition processing. Then, as shown in FIG. 4 (a), first, in S1, a character string score value is sent from the voice recognition means 1 to the disturbance sound environment detection means 3. The character string score value at this time is:
The values are as low as 38%, product K: 40%, and product A: 50%.
【0017】続いて、S2にて、外乱音環境検出手段3
は文字列スコア値からスコア値の部分のみを切り取る。
そして、切り取ったスコア値を外乱音環境判断手段4に
供給する。続いて、S3にて、外乱音環境判断手段4
は、入力されたスコア値から90%以上のスコア値が1
つもないことを判断し、フラグに「0」を設定する。な
お、一定時間経過後にはこのフラグを「1」に変更す
る。Subsequently, in S2, the disturbance sound environment detecting means 3
Cuts out only the score value part from the string score value.
Then, the cut score value is supplied to the disturbance sound environment determination means 4. Subsequently, in S3, the disturbance sound environment determining means 4
Indicates that a score value of 90% or more from the input score value is 1
It is determined that there is no connection, and “0” is set in the flag. Note that this flag is changed to “1” after a certain time has elapsed.
【0018】続いて、S4にて、音声認識環境表示制御
手段5はフラグが「0」に設定されていることを確認し
て、例えば「音声認識できない環境です」というテキス
トを表示装置2に送り、表示させる。これにより、話者
は今の環境が音声認識できない環境になっていることを
知ることができ、音声入力の作業を一時中断する。Subsequently, in S4, the voice recognition environment display control means 5 confirms that the flag is set to "0" and sends, for example, a text "environment in which voice cannot be recognized" to the display device 2. And display. As a result, the speaker can know that the current environment is an environment in which speech cannot be recognized, and temporarily suspends the work of voice input.
【0019】また、外乱音のレベルが低く音声認識でき
る環境の範囲内にあるときには、音声認識手段1が外乱
音を取込んでもこの音量は一定レベル以上になってはい
ないので、音声認識手段1は音声認識処理を開始するこ
とはない。従って、音声認識手段1から文字列や文字列
スコア値が出力されることはない。また、フラグは一旦
「0」に設定されても一定時間経過後には「1」になっ
ているので、音声認識環境表示制御手段5はフラグが
「1」に設定されていることを確認して、例えば「音声
認識できる環境です」というテキストを表示装置2に送
り、表示させる。これにより、話者は今の環境が音声認
識できる環境になっていることを知ることができる。Further, when the level of the disturbance sound is low and is within the range of the environment in which the voice can be recognized, the volume does not exceed a certain level even if the voice recognition means 1 takes in the disturbance sound. Does not start the voice recognition process. Therefore, no character string or character string score value is output from the voice recognition means 1. Further, even if the flag is once set to "0", the flag is set to "1" after a certain period of time, so the voice recognition environment display control means 5 confirms that the flag is set to "1". For example, a text “environment that can recognize voice” is sent to the display device 2 and displayed. This allows the speaker to know that the current environment is an environment in which speech can be recognized.
【0020】この環境下で話者が、例えば、“商品C”
と発声すると音量が一定レベル以上となって音声認識装
置1は音声認識処理を開始する。そして、図4の(b) に
示すように、先ず、S11にて、音声認識手段1から外
乱音環境検出手段3に文字列スコア値が送出される。こ
のときの文字列スコア値は、商品C:98%、商品K:
70%、商品A:60%のように高い値になる。In this environment, when the speaker, for example, "commodity C"
, The volume becomes equal to or higher than a certain level, and the voice recognition device 1 starts the voice recognition process. Then, as shown in FIG. 4 (b), first, in S11, a character string score value is sent from the speech recognition means 1 to the disturbance sound environment detection means 3. The character string score values at this time are: product C: 98%, product K:
High values such as 70% and product A: 60%.
【0021】続いて、S12にて、外乱音環境検出手段
3は文字列スコア値からスコア値の部分のみを切り取
る。そして、切り取ったスコア値を外乱音環境判断手段
4に供給する。続いて、S13にて、外乱音環境判断手
段4は、入力されたスコア値から90%以上のスコア値
が1つあることを判断し、フラグに「1」を設定する。Subsequently, in S12, the disturbance sound environment detecting means 3 cuts out only the score value part from the character string score value. Then, the cut score value is supplied to the disturbance sound environment determination means 4. Subsequently, in S13, the disturbance sound environment determination means 4 determines that there is one score value of 90% or more from the input score value, and sets "1" to the flag.
【0022】続いて、S14にて、音声認識環境表示制
御手段5はフラグが「1」に設定されていることを確認
して、「音声認識できる環境です」というテキストを表
示させる。また、音声認識手段1は、“商品C”の音声
入力を認識して「商品C」の文字列を表示装置2に供給
する。こうして、表示装置2は「商品C」の文字も表示
する。Subsequently, in S14, the voice recognition environment display control means 5 confirms that the flag is set to "1", and displays a text "environment in which voice can be recognized". Further, the voice recognition unit 1 recognizes the voice input of “product C” and supplies the character string of “product C” to the display device 2. In this way, the display device 2 also displays the character of “product C”.
【0023】このように、外乱音の環境を判断し、音声
認識ができない環境のときには、その旨を表示装置2に
表示して知らせているので、話者は常に音声認識が可能
な環境のもとで音声入力することが可能になり、認識率
を高めることができる。As described above, the environment of the disturbance sound is determined, and in the environment where the voice cannot be recognized, the fact is displayed on the display device 2 to notify the user that the environment cannot be recognized. Can input a voice, and the recognition rate can be increased.
【0024】なお、この実施の形態では、表示装置2に
対して、「音声認識できない環境です」、「音声認識で
きる環境です」というテキストを表示させるようにした
が必ずしもこれに限定するものではなく、例えば、「音
声認識できない環境です」のテキストの代わりに図5の
(a) に示すアイコン表示を行い、「音声認識できる環境
です」の代わりに図5の(b) に示すアイコン表示を行っ
てもよい。また、図6に示す赤と青の2色の光インジケ
ータ6を使用し、「音声認識できない環境です」のテキ
ストの代わりにこの光インジケータ6を赤点灯表示し、
「音声認識できる環境です」の代わりにこの光インジケ
ータ6を緑点灯表示してもよい。その他、音声を併用し
てもよい。In this embodiment, the display device 2 displays the texts "environment in which speech cannot be recognized" and "environment in which speech can be recognized". However, the present invention is not limited to this. For example, instead of the text "It is an environment where speech cannot be recognized",
The icon display shown in FIG. 5A may be performed, and the icon display shown in FIG. In addition, the red and blue light indicators 6 shown in FIG. 6 are used, and the light indicator 6 is displayed in red instead of the text "It is an environment where voice cannot be recognized".
The light indicator 6 may be displayed in green instead of "the environment in which voice can be recognized". In addition, voice may be used together.
【0025】(第2の実施の形態)図7において、11
は入力する話者の音声を認識して文字列情報を出力する
音声認識手段、12はこの音声認識手段11からの文字
列情報を表示する表示装置である。すなわち、前記音声
認識手段11は、音声認識開始命令が入力された後、あ
る一定音量レベル以上の音声が入力されたとき始めて音
声認識処理を開始し、音声入力が商品についてのもので
あれば前述した第1の実施の形態の音声入力手段1と同
様に図3に示すような文字列情報を出力する。例えば、
話者が“商品A”と発声すると音声認識装置11はそれ
をマイクロホンから取込んで認識し、商品Aの文字列情
報を表示装置12に出力し、表示装置12は商品Aの文
字列情報を表示することになる。(Second Embodiment) In FIG.
Is a voice recognition unit that recognizes the voice of the input speaker and outputs character string information, and 12 is a display device that displays the character string information from the voice recognition unit 11. That is, after the voice recognition start command is input, the voice recognition means 11 starts the voice recognition process only when a voice having a certain volume level or higher is input. The character string information as shown in FIG. 3 is output in the same manner as the voice input means 1 of the first embodiment. For example,
When the speaker utters “Product A”, the voice recognition device 11 takes in the product from the microphone and recognizes it, and outputs the character string information of the product A to the display device 12. Will be displayed.
【0026】13は音声認識環境検出手段としての発声
音量環境検出手段で、話者の発声音量を音圧として検出
し、dB値に変換した後、音声認識環境判断手段として
の発声音量環境判断手段14にそのdB値を供給するよ
うになっている。前記発声音量環境判断手段14は、受
け取ったdB値に基づいて環境を判断し、例えば、dB
値が60dB以上80dB以下のときには適切な発声音
量であると判断してフラグを「1」に設定し、また、d
B値が60dB未満か120dBを超えるときには不適
切な発声音量であると判断してフラグを「0」に設定す
るようになっている。そして、フラグを「0」に設定し
たときには、一定時間経過後にはフラグを「1」に変更
するようになっている。Reference numeral 13 denotes an utterance volume environment detecting means as a voice recognition environment detecting means, which detects the utterance volume of the speaker as sound pressure, converts the sound pressure into a dB value, and then makes a utterance volume environment determining means as a voice recognition environment determining means. 14 is supplied with the dB value. The utterance volume environment determination means 14 determines the environment based on the received dB value, and
When the value is not less than 60 dB and not more than 80 dB, it is determined that the sound volume is appropriate, and the flag is set to “1”.
When the B value is less than 60 dB or more than 120 dB, it is determined that the sound volume is inappropriate, and the flag is set to “0”. Then, when the flag is set to “0”, the flag is changed to “1” after a lapse of a predetermined time.
【0027】音声認識環境表示制御手段15は前記発声
音量環境判断手段14のフラグにより発声音量の環境を
判断し、フラグが「0」に設定されていることを確認し
て、例えば「発声音量が不適切です」というテキストを
表示装置2に送って表示させ、また、フラグが「1」に
設定されていることを確認して、例えば「発声音量が適
切です」というテキストを表示装置2に送って表示させ
るようになっている。このような構成においては、図8
に示すように、話者が、例えば、“商品A”と発声した
ときに、音量が大きく120dBに達すると、このとき
には発声音量環境判断手段14は、S21にて、フラグ
を「0」に設定し、音声認識環境表示制御手段15は、
S22にて、表示装置12に「発声音量が不適切です」
というテキストを表示させる。The voice recognition environment display control means 15 judges the environment of the sound volume based on the flag of the sound volume environment judgment means 14 and confirms that the flag is set to "0". The text "Inappropriate" is sent to the display device 2 for display, and after confirming that the flag is set to "1", for example, the text "Sound volume is appropriate" is sent to the display device 2. Is displayed. In such a configuration, FIG.
As shown in (2), when the speaker utters, for example, "product A" and the volume increases to 120 dB, the utterance volume environment determination means 14 sets the flag to "0" in S21. Then, the voice recognition environment display control means 15
At S22, the display device 12 displays "The utterance volume is inappropriate."
Is displayed.
【0028】このように、発声音量が大きすぎてマイク
ロホンに歪み等が生じて認識に支障を来すような場合に
は表示装置12に「発声音量が不適切です」という注意
を促すことができる。これにより、話者は発声音量に注
意して発声ができるようになる。As described above, in the case where the uttered sound volume is too large and the microphone is distorted or the like, which hinders recognition, it is possible to warn the display device 12 that "the uttered sound volume is inappropriate". . As a result, the speaker can speak while paying attention to the speech volume.
【0029】また、話者が、“商品A”と発声したとき
に、音量が小さく20dB程度にしか達しなかったとき
には、このときも発声音量環境判断手段14は、S21
にて、フラグを「0」に設定し、音声認識環境表示制御
手段15は、S22にて、表示装置12に「発声音量が
不適切です」というテキストを表示させる。When the speaker utters "Product A" and the volume is small and reaches only about 20 dB, the utterance volume environment judgment means 14 also returns to S21.
In step S22, the flag is set to "0", and the voice recognition environment display control means 15 causes the display device 12 to display the text "the utterance volume is inappropriate" in step S22.
【0030】このように、発声音量が小さすぎて認識に
支障を来すような場合にも表示装置12に「発声音量が
不適切です」という注意を促すことができる。これによ
り、話者は発声音量に注意して発声ができるようにな
る。As described above, even when the utterance volume is too low and the recognition is hindered, it is possible to warn the display device 12 that "the utterance volume is inappropriate". As a result, the speaker can speak while paying attention to the speech volume.
【0031】また、話者が、“商品A”と発声したとき
に、音量が70dB程度に達すると、このときは、音量
が60dB以上80dB以下の適切音量範囲に入ってい
るので、発声音量環境判断手段14は、S21にて、フ
ラグを「1」に設定し、音声認識環境表示制御手段15
は、S22にて、表示装置12に「発声音量が適切で
す」というテキストを表示させる。When the volume of the speaker reaches about 70 dB when the speaker utters "Product A", the volume is in the appropriate volume range of 60 dB to 80 dB. The determination means 14 sets the flag to "1" in S21, and the voice recognition environment display control means 15
Causes the display device 12 to display the text "Sound volume is appropriate" in S22.
【0032】このように、発声音量が適切な音量の範囲
に入っている場合には、話者に対して表示装置12に
「発声音量が適切です」という表示ができる。従って、
話者は、常に発声音量が適切になっているか否かを確認
しながら音声入力作業ができるので、発声音量を適切に
して音声入力を行うことが容易になり、認識率を高める
ことができる。なお、この実施の形態においても、表示
装置2にテキストを表示させる代わりにアイコン表示や
光インジケータによる表示、さらには、音声を併用して
もよい。As described above, when the uttered sound volume is within the appropriate sound volume range, the display device 12 can display "the uttered sound volume is appropriate" to the speaker. Therefore,
Since the speaker can always perform the voice input operation while checking whether or not the utterance volume is appropriate, it is easy to perform the voice input with the utterance volume appropriate, and the recognition rate can be increased. Also in this embodiment, instead of displaying text on the display device 2, an icon display, a display using a light indicator, or a voice may be used.
【0033】(第3の実施の形態)図9において、21
は入力する話者の音声を認識して文字列情報を出力する
音声認識手段、22はこの音声認識手段21からの文字
列情報を表示する表示装置である。すなわち、前記音声
認識手段21は、音声認識開始命令が入力された後、あ
る一定音量レベル以上の音声が入力されたとき始めて音
声認識処理を開始し、音声入力が商品についてのもので
あれば前述した第1の実施の形態の音声入力手段1と同
様に図3に示すような文字列情報を出力する。例えば、
話者が“商品A”と発声すると音声認識装置21はそれ
をマイクロホンから取込んで認識し、商品Aの文字列情
報を表示装置22に出力し、表示装置22は商品Aの文
字列情報を表示することになる。(Third Embodiment) In FIG.
Is a voice recognition unit that recognizes the voice of the input speaker and outputs character string information, and 22 is a display device that displays the character string information from the voice recognition unit 21. That is, the voice recognition means 21 starts the voice recognition process only when a voice having a certain volume level or more is input after the voice recognition start command is input, and if the voice input is for a commodity, The character string information as shown in FIG. 3 is output in the same manner as the voice input means 1 of the first embodiment. For example,
When the speaker utters "product A", the voice recognition device 21 takes in the product from the microphone and recognizes it, outputs character string information of the product A to the display device 22, and the display device 22 outputs the character string information of the product A. Will be displayed.
【0034】なお、ここでは省略しているが、この実施
の形態においても第1の実施の形態における外乱音環境
検出手段及び外乱音環境判断手段、あるいは第2の実施
の形態における発声音量環境検出手段及び発声音量環境
判断手段に相当する手段を備え、音声認識環境の良否や
発声音量環境の適切、不適切を表示装置22に表示でき
るようになっている。Although omitted here, also in this embodiment, the disturbing sound environment detecting means and the disturbing sound environment determining means in the first embodiment, or the sound volume environment detecting means in the second embodiment. Means and means corresponding to the sound volume environment determination means, so that the quality of the voice recognition environment and the appropriateness or inappropriateness of the sound volume environment can be displayed on the display device 22.
【0035】この実施の形態においては、前記音声認識
手段21は、さらに話者を認識するための複数の最適化
情報を管理し、選択された話者に適合する最適化情報に
基づいて音声認識できるようになっている。すなわち、
前記音声認識手段21に対して、外部から音声登録情報
等を入力し、内部メモリに、例えば、図10に示すよう
な、男性の声、女性の声、子供の声、大人の声、方言等
の複数の最適化情報を管理するようになっている。In this embodiment, the voice recognition means 21 further manages a plurality of optimization information for recognizing a speaker, and performs voice recognition based on the optimization information suitable for the selected speaker. I can do it. That is,
Voice registration information or the like is input from the outside to the voice recognition means 21 and, for example, a male voice, a female voice, a child voice, an adult voice, a dialect, or the like as shown in FIG. A plurality of pieces of optimization information are managed.
【0036】また、前記音声認識手段21は選択された
条件を最適化された話者に対する情報として音声認識環
境表示制御手段23に送出するようになっている。前記
音声認識環境表示制御手段23は受け取った最適化され
た話者に対する情報を前記表示装置22に表示するよう
になっている。なお、前記音声認識環境表示制御手段2
3はその他については前述した第1の実施の形態の音声
認識環境表示制御手段5、あるいは第2の実施の形態の
音声認識環境表示制御手段15と同様の制御機能を有す
るものである。The voice recognition means 21 sends the selected condition to the voice recognition environment display control means 23 as information for the optimized speaker. The voice recognition environment display control means 23 displays the received information on the optimized speaker on the display device 22. The voice recognition environment display control means 2
3 has the same control function as the voice recognition environment display control means 5 of the first embodiment or the voice recognition environment display control means 15 of the second embodiment.
【0037】このような構成においては、例えば、男性
の大人で方言Aを話す人が話者の場合に、その条件を選
択することにより、音声認識手段21は話者の発声入力
に対してより確実に認識できることになる。また、話者
が交替する場合に表示装置22に表示されている最適化
された話者に対する情報を確認することができ、もし、
自己に対する条件に適合していなければ条件を選択し直
すことができ、これにより、常に話者に適合した最適化
情報を設定することができ、音声認識率をさらに高める
ことができる。In such a configuration, for example, when a male adult who speaks the dialect A is a speaker, by selecting the condition, the voice recognition means 21 can respond more to the utterance input of the speaker. You can be surely recognized. In addition, when the speaker is changed, the information on the optimized speaker displayed on the display device 22 can be confirmed.
If the condition does not match the condition for the user, the condition can be selected again, whereby the optimization information suitable for the speaker can always be set, and the speech recognition rate can be further increased.
【0038】なお、ここでは、最適化情報として、男性
や女性など大まかな分類で設定したが、さらに、話者個
々に対応した最適化情報も設定できるようにすれば、個
々の話者の発声音を確実に認識することができ、例え
ば、商品の登録処理を行うPOS端末のようにある程度
決められたキャッシュが操作するような装置に適用した
場合にはきわめて有効となる。In this case, the optimization information is set according to rough classification such as male and female. However, if the optimization information corresponding to each speaker can be set, the utterance of each speaker can be set. This is extremely effective when applied to an apparatus in which a voice can be reliably recognized and, for example, a cache operated to a certain extent is operated, such as a POS terminal that performs product registration processing.
【0039】(第4の実施の形態)図11において、3
1は入力する話者の音声を認識して文字列情報を出力す
る音声認識手段、32はこの音声認識手段31からの文
字列情報を表示する表示装置である。すなわち、前記音
声認識手段31は、音声認識開始命令が入力された後、
ある一定音量レベル以上の音声が入力されたとき始めて
音声認識処理を開始し、音声入力が商品についてのもの
であれば前述した第1の実施の形態の音声入力手段1と
同様に図3に示すような文字列情報を出力する。例え
ば、話者が“商品A”と発声すると音声認識装置31は
それをマイクロホンから取込んで認識し、商品Aの文字
列情報を表示装置32に出力し、表示装置22は商品A
の文字列情報を表示することになる。(Fourth Embodiment) In FIG.
1 is a voice recognition means for recognizing the voice of the input speaker and outputting character string information, and 32 is a display device for displaying the character string information from the voice recognition means 31. That is, after the voice recognition start command is input, the voice recognition unit 31
The voice recognition process is started only when a voice having a certain volume level or higher is input, and if the voice input is for a product, it is shown in FIG. 3 similarly to the voice input unit 1 of the first embodiment described above. Output such string information. For example, when the speaker utters “product A”, the voice recognition device 31 takes in the recognition from the microphone, recognizes the product, outputs the character string information of the product A to the display device 32, and the display device 22
Will be displayed.
【0040】なお、ここでは省略しているが、この実施
の形態においても第1の実施の形態における外乱音環境
検出手段及び外乱音環境判断手段、あるいは第2の実施
の形態における発声音量環境検出手段及び発声音量環境
判断手段に相当する手段を備え、音声認識環境の良否や
発声音量環境の適切、不適切を表示装置32に表示でき
るようになっている。Although omitted here, also in this embodiment, the disturbing sound environment detecting means and the disturbing sound environment determining means in the first embodiment, or the sound volume environment detecting means in the second embodiment. Means and means corresponding to the sound volume environment determination means, so that the quality of the voice recognition environment and the appropriateness or inappropriateness of the sound volume environment can be displayed on the display device 32.
【0041】この実施の形態においては、前記音声認識
手段31は、さらに図12に示すような各種のエラー情
報を内部メモリに予め設定し、話者の発声音を取込んだ
とき、その発声音を認識する処理を行うが、このとき認
識ができずエラーとなったときには、前記内部メモリか
ら該当するエラー情報を読出して音声認識環境表示制御
手段33に送出するようになっている。In this embodiment, the voice recognizing means 31 further sets various kinds of error information as shown in FIG. In this case, if an error occurs because the recognition is not possible, the corresponding error information is read out from the internal memory and sent to the voice recognition environment display control means 33.
【0042】前記音声認識環境表示制御手段33は受け
取ったエラー情報を前記表示装置32に表示するように
なっている。なお、前記音声認識環境表示制御手段33
はその他については前述した第1の実施の形態の音声認
識環境表示制御手段5、あるいは第2の実施の形態の音
声認識環境表示制御手段15と同様の制御機能を有する
ものである。The voice recognition environment display control means 33 displays the received error information on the display device 32. The voice recognition environment display control means 33
Has the same control function as the voice recognition environment display control means 5 of the first embodiment or the voice recognition environment display control means 15 of the second embodiment.
【0043】このような構成においては、話者が発声し
て音声入力を行ったとき、認識エラーが発生すると音声
認識手段21はエラー内容を判断して該当するエラー情
報を内部メモリから読出して音声認識環境表示制御手段
33に送出する。例えば、話者が商品名の音声入力を行
ったときにエラーが発生すると、音声認識手段21は
「商品名がわかりません」というエラー情報を内部メモ
リから読出して音声認識環境表示制御手段33に送出す
る。In such a configuration, if a recognition error occurs when a speaker utters a voice and performs a voice input, the voice recognition means 21 determines the content of the error, reads out the corresponding error information from the internal memory, and outputs the voice information. It is sent to the recognition environment display control means 33. For example, if an error occurs when the speaker performs a voice input of the product name, the voice recognition unit 21 reads out the error information “I do not know the product name” from the internal memory and sends it to the voice recognition environment display control unit 33. Send out.
【0044】これにより、音声認識環境表示制御手段3
3は表示装置32を制御し、「商品名がわかりません」
というエラーメッセージを表示させる。こうして、話者
は商品名入力においてエラーが発生したことを把握で
き、改めてより明確な発音で商品名の発声を行うように
なる。このようにして話者に対してエラー内容を把握さ
せ、注意を促すことで認識率の向上に貢献できる。Thus, the voice recognition environment display control means 3
3 controls the display device 32, "I do not know the product name"
Is displayed. In this way, the speaker can recognize that an error has occurred in the input of the product name, and again speaks the product name with a clearer pronunciation. In this way, it is possible to make the speaker understand the contents of the error and call attention, thereby contributing to an improvement in the recognition rate.
【0045】[0045]
【発明の効果】以上詳述したように、各請求項記載の発
明によれば、周囲の環境状態が音声認識に適切な環境か
否かを常にチェックして知らせることができ、これによ
り話者は音声認識に適切な環境のもとで音声入力がで
き、認識率を高めることができる。As described in detail above, according to the invention described in each of the claims, it is possible to always check whether or not the surrounding environment state is suitable for speech recognition and to inform the user of the situation. Can input voice under an environment suitable for voice recognition, and can increase the recognition rate.
【0046】また、請求項4記載の発明によれば、さら
に、話者に適合する最適化情報に基づいて音声認識する
ので、認識率をさらに高めることができる。また、請求
項5記載の発明によれば、さらに、音声の認識ができな
かったときに音声に対応したエラー情報を表示でき、こ
れにより話者に対する発声時の注意を促すことができ
る。According to the fourth aspect of the present invention, the speech recognition is further performed based on the optimization information suitable for the speaker, so that the recognition rate can be further increased. Further, according to the invention described in claim 5, when the speech cannot be recognized, error information corresponding to the speech can be displayed, and thereby the speaker can be alerted when speaking.
【図1】本発明の第1の実施の形態を示すブロック図。FIG. 1 is a block diagram showing a first embodiment of the present invention.
【図2】同実施の形態における文字列スコア値の例を示
す図。FIG. 2 is a view showing an example of a character string score value in the embodiment.
【図3】同実施の形態における文字列の例を示す図。FIG. 3 is a diagram showing an example of a character string in the embodiment.
【図4】同実施の形態において音声認識ができない場合
と音声認識ができた場合の表示制御を示す流れ図。FIG. 4 is a flowchart showing display control in a case where speech recognition is not possible and a case where speech recognition is successful in the embodiment.
【図5】同実施の形態において音声認識ができない場合
と音声認識ができた場合の他の表示例を示す図。FIG. 5 is a view showing another display example in the case where speech recognition is not possible and the case where speech recognition is successful in the embodiment.
【図6】同実施の形態において音声認識ができない場合
と音声認識ができた場合の他の表示例を示す図。FIG. 6 is a diagram showing another display example in the case where speech recognition is not possible and the case where speech recognition is successful in the embodiment.
【図7】本発明の第2の実施の形態を示すブロック図。FIG. 7 is a block diagram showing a second embodiment of the present invention.
【図8】同実施の形態における適切音量の場合と不適切
音量の場合の表示制御を示す流れ図。FIG. 8 is a flowchart showing display control in the case of an appropriate volume and an inappropriate volume in the embodiment.
【図9】本発明の第3の実施の形態を示す要部ブロック
図。FIG. 9 is a main part block diagram showing a third embodiment of the present invention.
【図10】同実施の形態において音声認識手段が管理す
る最適化情報の例を示す図。FIG. 10 is a diagram showing an example of optimization information managed by a voice recognition unit in the embodiment.
【図11】本発明の第4の実施の形態を示す要部ブロッ
ク図。FIG. 11 is a main part block diagram showing a fourth embodiment of the present invention.
【図12】同実施の形態において音声認識手段が管理す
るエラー情報の例を示す図。FIG. 12 is a view showing an example of error information managed by the voice recognition means in the embodiment.
1,11,21,31…音声認識手段 2,12,22,32…表示装置 3…外乱音環境検出手段 4…外乱音環境判断手段 5,15,23,33…音声認識環境表示制御手段 13…発声音量環境検出手段 14…発声音量環境判断手段 1, 11, 21, 31 ... voice recognition means 2, 12, 22, 32 ... display device 3 ... disturbance sound environment detection means 4 ... disturbance sound environment determination means 5, 15, 23, 33 ... voice recognition environment display control means 13 ... Sound volume environment detection means 14 ... Sound volume environment judgment means
─────────────────────────────────────────────────────
────────────────────────────────────────────────── ───
【手続補正書】[Procedure amendment]
【提出日】平成10年8月26日[Submission date] August 26, 1998
【手続補正1】[Procedure amendment 1]
【補正対象書類名】明細書[Document name to be amended] Statement
【補正対象項目名】請求項3[Correction target item name] Claim 3
【補正方法】変更[Correction method] Change
【補正内容】[Correction contents]
【手続補正2】[Procedure amendment 2]
【補正対象書類名】明細書[Document name to be amended] Statement
【補正対象項目名】0008[Correction target item name] 0008
【補正方法】変更[Correction method] Change
【補正内容】[Correction contents]
【0008】請求項3記載の発明は、請求項1記載の音
声認識装置において、音声認識環境検出手段として話者
の発声音量を検出する発声音量環境検出手段を使用し、
音声認識環境判断手段として発声音量環境検出手段が検
出した発声音量が音声認識に適切な音量であるか否かを
判断する音声認識環境判断手段を使用したものである。[0008] According to a third aspect of the invention, the speech recognition system according to claim 1, using the outgoing voice volume environment detection means for detecting an outgoing voice volume of the speaker as a speech recognition environment detection means,
In which outgoing voice sound volume detected by the originating voice volume environment detection unit as a voice recognition environment determining means using speech recognition environment determination means for determining whether it is appropriate volume to speech recognition.
Claims (5)
力する音声認識手段と、この音声認識手段からの文字列
情報を表示する表示装置と、音声認識に関する環境を検
出する音声認識環境検出手段と、この音声認識環境検出
手段が検出した環境情報により、今の環境が音声認識に
適切な環境か不適切な環境かを判断する音声認識環境判
断手段と、この音声認識環境判断手段が判断した結果を
前記表示装置に表示させる音声認識環境表示制御手段と
を備えたことを特徴とする音声認識装置。1. A voice recognition means for recognizing an input voice and outputting character string information, a display device for displaying character string information from the voice recognition means, and a voice recognition environment detection for detecting an environment relating to voice recognition. Means and a voice recognition environment determining means for determining whether the current environment is appropriate or inappropriate for voice recognition based on the environment information detected by the voice recognition environment detecting means, And a voice recognition environment display control means for displaying the result on the display device.
を検出する外乱音環境検出手段であり、音声認識環境判
断手段が前記外乱音環境検出手段が検出した外乱音のレ
ベルが音声認識可能な範囲内にあるか否かを判断する外
乱音環境判断手段であることを特徴とする請求項1記載
の音声認識装置。2. The sound recognition environment detecting means is a disturbance sound environment detecting means for detecting a level of a disturbance sound, and the speech recognition environment determining means is capable of recognizing the level of the disturbance sound detected by the disturbance sound environment detecting means. 2. The speech recognition device according to claim 1, wherein the speech recognition device is a disturbance sound environment determination unit that determines whether the sound is within a range.
を検出する発生音量環境検出手段であり、音声認識環境
判断手段が前記発生音量環境検出手段が検出した発生音
量が音声認識に適切な音量であるか否かを判断する音声
認識環境判断手段であることを特徴とする請求項1記載
の音声認識装置。3. The voice recognition environment detecting means is a generated sound volume environment detecting means for detecting a generated sound volume of a speaker, and the voice recognition environment determining means is adapted so that the generated sound volume detected by the generated sound volume environment detecting means is suitable for voice recognition. 2. A speech recognition apparatus according to claim 1, wherein said speech recognition environment judgment means judges whether or not the sound volume is attained.
複数の最適化情報を管理し、選択された音声入力する話
者に適合する最適化情報に基づいて音声認識することを
特徴とする請求項1記載の音声認識装置。4. The speech recognition means manages a plurality of pieces of optimization information for recognizing a speaker, and performs speech recognition based on optimization information suitable for a selected speaker who inputs speech. The voice recognition device according to claim 1.
不能のとき入力した音声に対応したエラー情報を出力
し、音声認識環境表示制御手段は、表示装置にエラー情
報を表示させることを特徴とする請求項1記載の音声認
識装置。5. The voice recognition means outputs error information corresponding to the input voice when the input voice cannot be recognized, and the voice recognition environment display control means causes the display device to display the error information. The speech recognition device according to claim 1, wherein
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP10158897A JPH11352995A (en) | 1998-06-08 | 1998-06-08 | Voice recognition device |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP10158897A JPH11352995A (en) | 1998-06-08 | 1998-06-08 | Voice recognition device |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| JPH11352995A true JPH11352995A (en) | 1999-12-24 |
Family
ID=15681768
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP10158897A Abandoned JPH11352995A (en) | 1998-06-08 | 1998-06-08 | Voice recognition device |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPH11352995A (en) |
Cited By (9)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2003022092A (en) * | 2001-07-09 | 2003-01-24 | Fujitsu Ten Ltd | Dialog system |
| JP2004502985A (en) * | 2000-06-29 | 2004-01-29 | コーニンクレッカ フィリップス エレクトロニクス エヌ ヴィ | Recording device for recording voice information for subsequent offline voice recognition |
| WO2004070703A1 (en) * | 2003-02-03 | 2004-08-19 | Mitsubishi Denki Kabushiki Kaisha | Vehicle mounted controller |
| JP2006505003A (en) * | 2002-11-02 | 2006-02-09 | コーニンクレッカ フィリップス エレクトロニクス エヌ ヴィ | Operation method of speech recognition system |
| JP2006113439A (en) * | 2004-10-18 | 2006-04-27 | Ntt Data Corp | Voice automatic response device and program |
| KR100810275B1 (en) | 2006-08-03 | 2008-03-06 | 삼성전자주식회사 | Voice recognition device and method for vehicle |
| JP2015087649A (en) * | 2013-10-31 | 2015-05-07 | シャープ株式会社 | Utterance control device, method, utterance system, program, and utterance device |
| WO2016088410A1 (en) * | 2014-12-02 | 2016-06-09 | ソニー株式会社 | Information processing device, information processing method, and program |
| WO2017026239A1 (en) * | 2015-08-10 | 2017-02-16 | クラリオン株式会社 | Voice operating system, server device, in-vehicle equipment, and voice operating method |
-
1998
- 1998-06-08 JP JP10158897A patent/JPH11352995A/en not_active Abandoned
Cited By (15)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2004502985A (en) * | 2000-06-29 | 2004-01-29 | コーニンクレッカ フィリップス エレクトロニクス エヌ ヴィ | Recording device for recording voice information for subsequent offline voice recognition |
| JP4917729B2 (en) * | 2000-06-29 | 2012-04-18 | ニュアンス コミュニケーションズ オーストリア ゲーエムベーハー | Recording device for recording voice information for subsequent offline voice recognition |
| JP2003022092A (en) * | 2001-07-09 | 2003-01-24 | Fujitsu Ten Ltd | Dialog system |
| JP2006505003A (en) * | 2002-11-02 | 2006-02-09 | コーニンクレッカ フィリップス エレクトロニクス エヌ ヴィ | Operation method of speech recognition system |
| US7617108B2 (en) | 2003-02-03 | 2009-11-10 | Mitsubishi Denki Kabushiki Kaisha | Vehicle mounted control apparatus |
| WO2004070703A1 (en) * | 2003-02-03 | 2004-08-19 | Mitsubishi Denki Kabushiki Kaisha | Vehicle mounted controller |
| JP2006113439A (en) * | 2004-10-18 | 2006-04-27 | Ntt Data Corp | Voice automatic response device and program |
| KR100810275B1 (en) | 2006-08-03 | 2008-03-06 | 삼성전자주식회사 | Voice recognition device and method for vehicle |
| JP2015087649A (en) * | 2013-10-31 | 2015-05-07 | シャープ株式会社 | Utterance control device, method, utterance system, program, and utterance device |
| WO2016088410A1 (en) * | 2014-12-02 | 2016-06-09 | ソニー株式会社 | Information processing device, information processing method, and program |
| JPWO2016088410A1 (en) * | 2014-12-02 | 2017-09-14 | ソニー株式会社 | Information processing apparatus, information processing method, and program |
| US10642575B2 (en) | 2014-12-02 | 2020-05-05 | Sony Corporation | Information processing device and method of information processing for notification of user speech received at speech recognizable volume levels |
| WO2017026239A1 (en) * | 2015-08-10 | 2017-02-16 | クラリオン株式会社 | Voice operating system, server device, in-vehicle equipment, and voice operating method |
| JP2017037176A (en) * | 2015-08-10 | 2017-02-16 | クラリオン株式会社 | Voice operation system, server device, in-vehicle device, and voice operation method |
| US10540969B2 (en) | 2015-08-10 | 2020-01-21 | Clarion Co., Ltd. | Voice operating system, server device, on-vehicle device, and voice operating method |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US20220292259A1 (en) | Artificially intelligent order processing system | |
| JP4241376B2 (en) | Correction of text recognized by speech recognition through comparison of speech sequences in recognized text with speech transcription of manually entered correction words | |
| CN100524459C (en) | Method and system for speech recognition | |
| US20010003173A1 (en) | Method for increasing recognition rate in voice recognition system | |
| CN101589428A (en) | Vehicle Voice Recognition Device | |
| JP2005331882A (en) | Voice recognition device, method, and program | |
| JPH0784592A (en) | Voice recognizer | |
| US6629072B1 (en) | Method of an arrangement for speech recognition with speech velocity adaptation | |
| JP5077107B2 (en) | Vehicle drinking detection device and vehicle drinking detection method | |
| JPH11352995A (en) | Voice recognition device | |
| JP2005283647A (en) | Emotion recognition device | |
| WO2020153109A1 (en) | Presentation assistance device for calling attention to words that are forbidden to speak | |
| JP2008275987A (en) | Speech recognition device and conference system | |
| JPH0635497A (en) | Speech input device | |
| JP2003330491A (en) | Method, device, and program for voice recognition | |
| JP2010286627A (en) | Emotion estimation device and emotion estimation method | |
| JPH0527790A (en) | Voice input/output device | |
| JP2006113439A (en) | Voice automatic response device and program | |
| JP2008033198A (en) | Voice interaction system, voice interaction method, voice input device and program | |
| JP2017161581A (en) | Voice recognition device and voice recognition program | |
| JP2001005482A (en) | Voice recognition method and apparatus | |
| CN119404247A (en) | Emotion Detection in Interruption Analysis | |
| JP3285704B2 (en) | Speech recognition method and apparatus for spoken dialogue | |
| JP2664785B2 (en) | Voice recognition device | |
| JP3588929B2 (en) | Voice recognition device |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| A977 | Report on retrieval |
Free format text: JAPANESE INTERMEDIATE CODE: A971007 Effective date: 20040630 |
|
| A131 | Notification of reasons for refusal |
Free format text: JAPANESE INTERMEDIATE CODE: A131 Effective date: 20040706 |
|
| A762 | Written abandonment of application |
Free format text: JAPANESE INTERMEDIATE CODE: A762 Effective date: 20040903 |