JPH10312193A - Voice input device - Google Patents

Voice input device

Info

Publication number
JPH10312193A
JPH10312193A JP9135793A JP13579397A JPH10312193A JP H10312193 A JPH10312193 A JP H10312193A JP 9135793 A JP9135793 A JP 9135793A JP 13579397 A JP13579397 A JP 13579397A JP H10312193 A JPH10312193 A JP H10312193A
Authority
JP
Japan
Prior art keywords
word
dictionary
hierarchy
partial dictionary
partial
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Granted
Application number
JP9135793A
Other languages
Japanese (ja)
Other versions
JP3588975B2 (en
Inventor
Takeshi Ono
健 大野
Norimasa Kishi
則政 岸
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Nissan Motor Co Ltd
Original Assignee
Nissan Motor Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Nissan Motor Co Ltd filed Critical Nissan Motor Co Ltd
Priority to JP13579397A priority Critical patent/JP3588975B2/en
Publication of JPH10312193A publication Critical patent/JPH10312193A/en
Application granted granted Critical
Publication of JP3588975B2 publication Critical patent/JP3588975B2/en
Anticipated expiration legal-status Critical
Expired - Fee Related legal-status Critical Current

Links

Landscapes

  • Navigation (AREA)
  • Traffic Control Systems (AREA)

Abstract

(57)【要約】 【課題】 少ない入力回数で、大語彙音声入力できる音
声入力装置を提供する。 【解決手段】 マイク500とA/Dコンバータ501
は音声を取り込んでディジタルの音声信号に変えて信号
処理装置5に出力する。外部記憶装置503には階層化
した単語辞書を記憶させている。タッチパネル付きモニ
タ60は、認識結果を表示するとともに、タッチ入力で
部分辞書の変更が行なえる。信号処理装置5では、最初
に所定の最下位階層の部分辞書を入力し、音声信号と辞
書内の単語との一致度を演算し、一致度の最も高い単語
をその上位階層構造を含めモニタにおいて表示する。そ
の単語が認識したい単語でなければ、変更したい階層が
タッチ入力される。そのタッチ信号が検出され対応した
辞書を外部記憶装置503から読み込まれる。その後音
声により階層が決定され、その階層下の最下層辞書が読
み込まれる。目的語を新たに発話して辞書との照合によ
り音声認識される。
(57) [Summary] [PROBLEMS] To provide a voice input device capable of inputting a large vocabulary voice with a small number of input times. SOLUTION: A microphone 500 and an A / D converter 501 are provided.
Takes in a voice, converts it into a digital voice signal, and outputs it to the signal processing device 5. The external storage device 503 stores a hierarchical word dictionary. The monitor 60 with a touch panel displays the recognition result and can change the partial dictionary by touch input. The signal processing device 5 first inputs a predetermined partial dictionary at the lowest hierarchy, calculates the degree of coincidence between the voice signal and a word in the dictionary, and determines the word with the highest degree of coincidence on the monitor including its upper-layer structure. indicate. If the word is not the word to be recognized, the hierarchy to be changed is touch-input. The touch signal is detected, and the corresponding dictionary is read from the external storage device 503. After that, the hierarchy is determined by voice, and the lowest dictionary under the hierarchy is read. The target word is newly uttered, and voice recognition is performed by collation with the dictionary.

Description

【発明の詳細な説明】DETAILED DESCRIPTION OF THE INVENTION

【0001】[0001]

【発明の属する技術分野】この発明は、少ない入力回数
で大語彙認識を行なうことができる音声入力装置に関す
る。
BACKGROUND OF THE INVENTION 1. Field of the Invention The present invention relates to a voice input device capable of performing large vocabulary recognition with a small number of inputs.

【0002】[0002]

【従来の技術】音声を識別して情報を読み取る音声入力
装置は、手を介さずに指令やデータなどの入力ができる
ため、障害者用ワープロ、車載ナビゲーションといった
ような手操作が困難な場面に採用されつつある。典型的
な音声入力装置は、図13に示すように、音声を取り込
み、電気信号に変えて音声信号を出力するマイク500
と、アナログの音声信号をA/D変換してディジタル情
報に変換させるA/Dコンバータ501と、音声信号か
ら単語認識をするのに用いられる単語辞書を記憶した外
部記憶装置503と、CPUとメモリを持ち単語辞書を
メモリに読み込ませたうえでCPUが音声信号と辞書内
の単語との一致度判定により音声認識を行なう信号処理
装置502とを有する。スイッチ505は装置使用者に
操作され音声信号の取り込みタイミングを信号処理装置
502に与える。モニタ504は認識された単語を表示
する。
2. Description of the Related Art A voice input device for recognizing voice and reading information is capable of inputting commands and data without any hand, so that it is difficult to perform manual operations such as a word processor for persons with disabilities and in-vehicle navigation. It is being adopted. As shown in FIG. 13, a typical voice input device is a microphone 500 that captures voice and converts it into an electric signal and outputs a voice signal.
An A / D converter 501 for A / D converting an analog voice signal into digital information, an external storage device 503 storing a word dictionary used for word recognition from the voice signal, a CPU and a memory And a signal processing device 502 that reads a word dictionary into a memory and performs speech recognition by determining the degree of coincidence between a speech signal and a word in the dictionary. The switch 505 is operated by the user of the apparatus and gives the signal processing apparatus 502 the timing of capturing the audio signal. The monitor 504 displays the recognized word.

【0003】一般に音声による入力は、認識対象の語数
が多くなる程認識率が低下するため、大語彙認識の場合
階層化した辞書が使用される。入力の際、辞書の分類に
従って複数回音声入力を繰り返して階層を進み、認識対
象語数を絞り込んだ状態で認識し、誤認識率を低下させ
るようにしている。
[0003] In general, in the case of speech input, the recognition rate decreases as the number of words to be recognized increases, and a hierarchical dictionary is used for large vocabulary recognition. At the time of input, voice input is repeated a plurality of times in accordance with the dictionary classification, the hierarchy is advanced, recognition is performed in a state where the number of words to be recognized is narrowed down, and the false recognition rate is reduced.

【0004】上記従来の装置における音声認識の詳細に
ついて、図14のフローチャートに従って説明する。こ
こでは、JRの駅名である「桜木町」を音声で入力する
場合を示す。このような駅名入力は、自動車のナビゲー
ション装置の目的地設定などに用いられている。外部記
憶装置503には例えば図2のような階層化した単語辞
書を記憶させてある。
[0004] The details of speech recognition in the above-described conventional apparatus will be described with reference to the flowchart of FIG. Here, a case in which “Sakuragicho”, which is the JR station name, is input by voice is shown. Such a station name input is used for setting a destination of an automobile navigation device. The external storage device 503 stores, for example, a hierarchical word dictionary as shown in FIG.

【0005】まず、ステップ521において、音声入力
の初期状態で、信号処理装置502は階層化された単語
辞書の、階層の最上位の部分辞書を最初の認識対象とし
設定する。ここでは図2中の部分辞書aが設定される。
次に、装置使用者はスイッチ505を操作してマイク5
00に「施設」を発話する。これは「桜木町」が「施
設」という範疇に含まれることを使用者が認識している
ためである。スイッチ505が操作されたことは、ステ
ップ522において、検出される。これによって音声認
識処理が開始される。
First, in step 521, in an initial state of voice input, the signal processing device 502 sets a partial dictionary at the highest level of the hierarchical word dictionary as a first recognition target. Here, the partial dictionary a in FIG. 2 is set.
Next, the device user operates the switch 505 to operate the microphone 5.
At 00, say "facility". This is because the user has recognized that "Sakuragicho" is included in the category of "facilities". The operation of the switch 505 is detected in step 522. Thereby, the voice recognition processing is started.

【0006】ステップ523では、信号処理装置502
内のCPUが設定された階層の部分辞書を外部記憶装置
503からメモリに読み込ませる。初期では部分辞書a
が読み込まれる。信号処理装置502は、スイッチ50
5が操作されるまでの音声の平均パワーを演算してお
り、スイッチ505が操作された後、音声の平均パワー
に比べて音声信号の瞬間パワーが所定値以上大きくなっ
たとき、使用者が発話開始したと判断し、ステップ52
4において音声取り込みを開始する。
In step 523, the signal processing device 502
CPU reads the partial dictionary of the set hierarchy from the external storage device 503 into the memory. Initially a partial dictionary a
Is read. The signal processing device 502 includes the switch 50
The average power of the voice until the operation of the voice signal 5 is operated. When the instantaneous power of the voice signal becomes larger than the average power of the voice by a predetermined value or more after the switch 505 is operated, the user speaks. It is determined that it has started, and step 52
At 4, the voice capture is started.

【0007】ステップ525では、CPUが入力された
音声とメモリに読み込んだ部分辞書内の単語との一致度
を演算する。初期では「住所」、「施設」それぞれとの
一致度を演算する。
In step 525, the CPU calculates the degree of coincidence between the input speech and the words in the partial dictionary read into the memory. Initially, the degree of matching with each of "address" and "facility" is calculated.

【0008】そして、音声信号の瞬間パワーが所定値以
下になった時、使用者の発話が終了したと判断し、ステ
ップ526において、音声の取り込みを終了する。入力
音声と部分辞書内の単語との一致度の演算を終了する
と、ステップ527において、CPUが最も一致度高い
単語を選択する。ここでは「施設」の方が一致度が高い
ので選択される。
When the instantaneous power of the audio signal falls below a predetermined value, it is determined that the utterance of the user has ended, and in step 526, the capture of the audio ends. When the calculation of the matching degree between the input voice and the word in the partial dictionary is completed, in step 527, the CPU selects the word having the highest matching degree. Here, “facility” is selected because the degree of coincidence is higher.

【0009】ステップ528では、選択された一致度の
高い単語が認識の結果としてモニタ504において表示
される。ステップ529では、選択された単語が下の階
層の部分辞書名であるかどうかを判定し、辞書名である
場合ステップ523へ、そうでない場合ステップ530
へ進む。ここでは単語「施設」は図2の単語辞書に示さ
れるように下の階層の部分辞書eの辞書名となっている
ので、ステップ523に戻る。
In step 528, the selected word having a high degree of matching is displayed on the monitor 504 as a result of recognition. In step 529, it is determined whether or not the selected word is a partial dictionary name of a lower hierarchy. If the selected word is a dictionary name, the process proceeds to step 523; otherwise, the process proceeds to step 530.
Proceed to. Here, since the word “facility” is the dictionary name of the partial dictionary e in the lower hierarchy as shown in the word dictionary of FIG. 2, the process returns to step 523.

【0010】ステップ523では、CPUが外部記憶装
置503から新たに部分辞書eをメモリに読み込む。こ
のように、処理を繰り返すたび階層が進む。選択された
単語の下に部分辞書がなくなると、ステップ530で、
認識された単語は目的語となり出力される。ここでは
「桜木町」が出力される。
In step 523, the CPU newly reads the partial dictionary e from the external storage device 503 into the memory. As described above, the hierarchy advances each time the processing is repeated. When there are no partial dictionaries below the selected word, in step 530,
The recognized word is output as an object. Here, "Sakuragicho" is output.

【0011】[0011]

【発明が解決しようとする課題】上記のように従来の音
声入力装置においては、一つの単語を伝えるのに、その
単語の属する最下層までの部分辞書名を階層毎に入力し
なければならない。その階層は単語数が多いほど複雑に
なるので、大語彙の中からの認識になるほど使用者にと
って入力負担が大きいという問題があった。また使用者
が最も伝えたいのは目的語であるにも係わらず、目的語
の入力が最後で感覚的に違和感を抱かせ、使い勝手がよ
くないという問題があった。本発明は、上記従来の問題
点に鑑み、少ない入力回数で正確な音声入力を実現する
ことを目的としている。
As described above, in the conventional voice input device, in order to transmit one word, the partial dictionary name up to the lowest layer to which the word belongs must be input for each layer. Since the hierarchy becomes more complicated as the number of words increases, there is a problem that the recognition load from a large vocabulary imposes a greater input burden on the user. Further, there is a problem that although the user wants to convey the object most, the input of the object gives a sense of incongruity at the end and the usability is not good. The present invention has been made in view of the above-described conventional problems, and has as its object to realize an accurate voice input with a small number of input times.

【0012】[0012]

【課題を解決するための手段】発話音声を情報化して入
力する音声入力手段と、複数の単語を含む部分辞書から
なるとともに階層構造を持った単語辞書と、使用する部
分辞書を決定する部分辞書決定手段と前記音声入力手段
からの入力音声と決定された部分辞書内の単語との一致
度を演算する演算手段と、演算された単語のうち最も一
致度の高い単語を選択して出力する単語選択手段とを有
し、前記部分辞書決定手段は、初めに所定の最下層の部
分辞書を決定し、階層変更指令を受けた場合に階層変更
をして部分辞書を決定するものとした。
Means for Solving the Problems Speech input means for converting uttered speech into information and inputting the word, a word dictionary comprising a plurality of words and having a hierarchical structure, and a partial dictionary for determining a partial dictionary to be used Determining means, calculating means for calculating the degree of coincidence between the input voice from the voice input means and the determined word in the partial dictionary, and selecting and outputting the word having the highest degree of matching from the calculated words Selection means, wherein the partial dictionary determination means first determines a predetermined lowermost partial dictionary, and changes the hierarchy when receiving a hierarchy change instruction to determine the partial dictionary.

【0013】前記部分辞書決定手段は、上位階層への変
更指示を受けた場合、上位階層の部分辞書を決定し、入
力により単語選択手段が選択した単語が示す階層に変更
し、該階層下の所定の最下層部分辞書を決定することが
できる。前記部分辞書決定手段により初めに決定される
部分辞書は前回使用された最下層の部分辞書であること
が望ましい。また前回使用された最下層の部分辞書の代
わりに使用頻度の高い最下層の部分辞書を用いても可能
である。
The partial dictionary determining means, when receiving an instruction to change to a higher hierarchical level, determines a partial dictionary in the upper hierarchical level, changes to a hierarchical level indicated by the word selected by the word selecting means by input, and changes the hierarchical level to a level indicated by the word selected by the word selecting means. A predetermined bottom sub-dictionary can be determined. It is preferable that the partial dictionary first determined by the partial dictionary determining means is the lowest partial dictionary used last time. It is also possible to use a lower-level partial dictionary that is frequently used instead of the lowest-level partial dictionary used last time.

【0014】音声入力装置には表示手段を接続し、単語
選択手段が選択した単語及び上位の階層構造を表示する
のが望ましい。前記表示手段はタッチパネルを合わせ持
ち、使用者によってタッチされた階層に対応した階層変
更指令を前記部分辞書決定部に出力することも可能であ
る。また、タッチパネルに対するタッチ入力があったと
きに、変更可能な階層を表示手段上で表示するのが望ま
しい。
[0014] It is desirable that display means is connected to the voice input device to display the word selected by the word selection means and the higher hierarchical structure. The display unit may have a touch panel and output a hierarchy change command corresponding to the hierarchy touched by the user to the partial dictionary determination unit. Further, it is desirable that a changeable hierarchy is displayed on the display means when a touch input is made on the touch panel.

【0015】前記部分辞書決定手段は、目的語に続き入
力が行なわれたときに、上位階層への変更指示を受けた
とし、設定されている最下層部分辞書より上位の全ての
部分辞書を決定し、単語選択手段が選択した単語が示す
階層に階層変更を行なうことができる。
The partial dictionary determining means determines that a change instruction to a higher hierarchical level is received when an input is made following the object, and determines all partial dictionaries higher than the set lowest partial dictionary. Then, the hierarchy can be changed to the hierarchy indicated by the word selected by the word selecting means.

【0016】前記部分辞書決定手段は、目的語に続き入
力が行なわれたときに、設定されている最下層部分辞書
及びそれより上位の全ての部分辞書を決定し、単語選択
手段が選択した単語が上位部分辞書内の単語であれば、
上位階層への変更指示を受けたとし、前記単語が示す階
層に階層変更を行なうこともできる。
The sub-dictionary deciding means, when an input is made following the object, decides the lowest sub-dictionary set and all sub-dictionaries higher than the set sub-dictionary, and selects the word selected by the word selecting means. Is a word in the upper sub-dictionary,
Assuming that a change instruction to a higher hierarchy is received, the hierarchy can be changed to the hierarchy indicated by the word.

【0017】また、最初に入力された音声を記憶する記
憶手段を合わせ持ち、階層が変更されたとき、前記単語
選択手段は前記記憶手段に記憶された音声を用いて、決
定された最下層の部分辞書内の単語との一致度を演算
し、出力することができる。
[0017] In addition, when the hierarchy is changed, the word selecting means uses the voice stored in the storage means to store the first input voice. The degree of coincidence with a word in the partial dictionary can be calculated and output.

【0018】[0018]

【作用】階層化した単語辞書を使用し、部分辞書を設定
する際に、初めに所定の最下層の部分辞書を設定するの
で、大語彙認識ができるとともに、1回目の単語を目的
語とすることもできる。そして一回の入力で認識できな
かった場合、階層変更指令で階層変更を行なうので、例
えば現在の部分辞書と目的語のある部分辞書と共通する
上位階層に変更して最下層の部分辞書を決定する段取り
をとることができ、最上位の部分辞書からの入力が必要
でなくなり、効率のよい入力が図られる。また、最初か
ら目的語で入力し、辞書の階層変更が認識できなかった
ときに行なうので、感覚に合致した音声入力が行なえ
る。
When a partial dictionary is set using a hierarchical word dictionary, a predetermined lowermost partial dictionary is first set, so that large vocabulary recognition can be performed and the first word is used as an object. You can also. If it is not recognized by one input, the hierarchy is changed by the hierarchy change command, so for example, change to the upper hierarchy common to the current partial dictionary and the partial dictionary with the object, and determine the lowest partial dictionary This eliminates the need for input from the uppermost partial dictionary, thereby achieving efficient input. In addition, since the input is performed from the beginning with the object and the change in the hierarchy of the dictionary cannot be recognized, the voice input matching the sense can be performed.

【0019】部分辞書決定手段は、上位階層への変更指
示を受けた場合、上位階層の部分辞書を決定し、入力に
より単語選択手段が選択した単語が示す階層に変更し、
該階層下の所定の最下層部分辞書を決定するようにした
場合、階層変更があったも、目的語の部分辞書と上位階
層との間の部分辞書に対する認識が不要で、階層が上位
に変更されても、すぐ目的語の入力となるので、装置使
用者を焦らせることなく、入力できる効果が得られる。
また部分辞書決定手段が初めに決定する部分辞書は前回
使用された最下層の部分辞書となれば、辞書構成を理解
し、階層変更が円滑に行なえる。さらに前回使用された
最下層の部分辞書の代わりに使用頻度の高い最下層の部
分辞書を用いても同じ効果が得られる。
The partial dictionary determining means, when receiving an instruction to change to a higher hierarchical level, determines a partial dictionary in the upper hierarchical level, and changes to a hierarchical level indicated by the word selected by the word selecting means by input.
If the predetermined lowermost sub-dictionary under the hierarchy is determined, even if the hierarchy is changed, the recognition of the partial dictionary between the object partial dictionary and the upper hierarchy is not necessary, and the hierarchy is changed to the higher hierarchy. Even if this is done, the object is immediately input, so that an effect can be obtained that can be input without frustrating the user of the device.
Further, if the partial dictionary determined first by the partial dictionary determining means is the lowest partial dictionary used last time, the dictionary configuration can be understood and the hierarchy can be changed smoothly. Further, the same effect can be obtained by using a lower-level partial dictionary frequently used in place of the lowest-level partial dictionary used last time.

【0020】音声入力装置に表示手段を接続し、単語選
択手段が選択した単語及び上位の階層構造を表示手段で
表示するようにすると、階層変更が必要のときに、表示
された階層構造を頼りに階層変更ができるので、入力負
担が軽減される。表示手段はタッチパネルを合わせ持
ち、使用者によってタッチされた階層に対応した階層変
更指令を前記部分辞書決定部に出力するようにすると、
対話式な入力となり一層の入力負担軽減が図られる。ま
た、タッチ入力があったときに、音声入力によって変更
可能な階層を前記表示手段上で表示した場合、辞書の構
造を表示画面から知ることができ、辞書の構成などの予
備知識が無くても階層変更ができる効果が得られる。
When the display means is connected to the voice input device, and the word selected by the word selection means and the higher hierarchical structure are displayed on the display means, when the hierarchical change is required, the displayed hierarchical structure is relied on. Since the hierarchy can be changed, the input load can be reduced. When the display means has a touch panel and outputs a hierarchy change command corresponding to the hierarchy touched by the user to the partial dictionary determination unit,
The input becomes interactive and the input load is further reduced. In addition, when there is a touch input, when a hierarchy that can be changed by voice input is displayed on the display unit, the structure of the dictionary can be known from the display screen, and even if there is no prior knowledge of the dictionary configuration, etc. The effect that the hierarchy can be changed is obtained.

【0021】前記部分辞書決定手段は、目的語に続き入
力が行なわれたときに、上位階層への変更指示を受けた
とし、設定されている最下層部分辞書より上位の全ての
部分辞書を決定し、単語選択手段が選択した単語が示す
階層に階層変更を行なうようにすると、上位階層の単語
のみで、階層変更ができ入力負担がさらに軽減される。
The partial dictionary determining means determines that when an input is made following the object, an instruction to change to a higher hierarchical level is received, and determines all partial dictionaries higher than the set lowermost partial dictionary. If the hierarchy is changed to the hierarchy indicated by the word selected by the word selection means, the hierarchy can be changed only by the words in the upper hierarchy, and the input load can be further reduced.

【0022】前記部分辞書決定手段は、目的語に続き入
力が行なわれたときに、設定されている最下層部分辞書
及びそれより上位の全ての部分辞書を決定し、単語選択
手段が選択した単語が上位部分辞書内の単語であれば、
上位階層への変更指示を受けたとし、前記単語が示す階
層に階層変更を行なった場合、部分辞書が正確でありな
がら、単語の誤認識による階層変更が防止される。
The sub-dictionary deciding means, when an input is made following the object, decides the lowest sub-dictionary set and all the sub-dictionaries higher than the set sub-dictionary, and selects the word selected by the word selecting means. Is a word in the upper sub-dictionary,
When a change instruction to a higher hierarchy is received and a hierarchy change is performed to the hierarchy indicated by the word, the hierarchy change due to erroneous word recognition is prevented while the partial dictionary is accurate.

【0023】最初に入力された音声を記憶する記憶手段
を合わせ持ち、階層が変更されたとき、前記単語選択手
段は前記記憶手段に記憶された音声を用いて、決定され
た最下層の部分辞書内の単語との一致度を演算し、出力
することで、階層を変更し、最下層の部分辞書が変わっ
ても、目的語の音声を二度と入力する必要がなく、音声
の入力回数がさらに減らされる。
When the hierarchy is changed, the word selection means uses the speech stored in the storage means to determine the partial dictionary of the lowest layer when the storage means for storing the first input speech is stored. By calculating and outputting the degree of matching with the words in the word, even if the hierarchy is changed and the partial dictionary at the bottom layer changes, there is no need to input the target speech again, and the number of times of voice input is further reduced. It is.

【0024】[0024]

【発明の実施の形態】次に、本発明の実施の形態を実施
例により説明する。図1は、本発明の第1の実施例を構
成を示すブロック図である。マイク500により取り込
まれた音声信号はA/Dコンバータ501でディジタル
情報に変えられて信号処理装置5に入力される。スイッ
チ505は音声入力用に装置使用者が発話する直前に操
作され、信号処理装置5に発話音声の検出タイミングを
与えている。外部記憶装置503には単語辞書を記憶さ
せてある。モニタ60はタッチパネル付きモニタで、信
号処理装置5の処理結果を表示するとともに、タッチパ
ネルで部分辞書の変更が行なえるようになっている。
Next, embodiments of the present invention will be described with reference to examples. FIG. 1 is a block diagram showing the configuration of the first embodiment of the present invention. The audio signal captured by the microphone 500 is converted into digital information by the A / D converter 501 and input to the signal processing device 5. The switch 505 is operated just before the device user speaks for voice input, and gives the signal processing device 5 a detection timing of the speech voice. The external storage device 503 stores a word dictionary. The monitor 60 is a monitor with a touch panel, which displays the processing results of the signal processing device 5 and allows the partial dictionary to be changed using the touch panel.

【0025】単語辞書は図2に示すような階層化した単
語辞書が用いられる。図においては部分辞書間の連線が
上下層間の繋がり関係を表示し、枠内に突出したのは下
の階層を代表する単語についての表示である。「住所」
と「施設」の二つの単語で構成する部分辞書aは単語辞
書の最上層にあり、各単語に下位階層辞書として「住
所」からは政令指定都市に従った分類で都道府県名を記
した部分辞書bを設け、各都道府県について、市区名を
記した部分辞書cと区村名を記した部分辞書dが順次に
設けられている。
As the word dictionary, a hierarchical word dictionary as shown in FIG. 2 is used. In the figure, the connecting lines between the partial dictionaries indicate the connection relationship between the upper and lower layers, and what protrudes into the frame is the display of words representing the lower hierarchy. "Street address"
The partial dictionary a, which is composed of two words, “a” and “facility”, is located at the top layer of the word dictionary, and each word is a lower-level dictionary in which the “address” is a part in which the name of the prefecture is written in accordance with the government-designated city. A dictionary b is provided, and for each prefecture, a partial dictionary c in which city and ward names are described and a partial dictionary d in which ward and village names are described are sequentially provided.

【0026】部分辞書aの単語「施設」の下位には
「駅」、「デパート」、「ホテル」のような「施設」の
種類名を記した単語の部分辞書eが続いている。「駅」
の下位に都道府県名を記した部分辞書fが設けられてい
る。部分辞書fからは交通会社名を記した部分辞書g
と、部分辞書gに繋がっている駅名の部分辞書hが都道
府県毎に設けられている。
Below the word "facility" in the partial dictionary a, a partial dictionary e of words describing the type names of "facility" such as "station", "department store" and "hotel" follows. "station"
A sub-dictionary f in which the names of prefectures are described is provided below. From partial dictionary f, partial dictionary g describing transportation company names
And a partial dictionary h of station names connected to the partial dictionary g is provided for each prefecture.

【0027】また、単語「デパート」も都道府県名を記
した下層の部分辞書iと、市名を記した部分辞書jと、
「デパート」名を記した部分辞書kを持ち、都道府県別
に市区の部分辞書と区村の部分辞書に繋がっている。部
分辞書d、部分辞書h、部分辞書kは最下層部分辞書で
あり、目的語が登録され、認識したい単語はここで照合
され認識される。
The word "department store" is also composed of a lower partial dictionary i in which the names of prefectures are written, a partial dictionary j in which the names of cities are written,
It has a partial dictionary k describing the name of "department store", and is linked to a partial dictionary of cities and wards and a partial dictionary of wards and villages by prefecture. The partial dictionary d, the partial dictionary h, and the partial dictionary k are the lowermost partial dictionaries, in which object words are registered, and words to be recognized are collated and recognized here.

【0028】信号処理装置5は、音声入力部53におい
てスイッチ505が操作されるまでマイク500からデ
ィジタル化された音声信号の平均パワーを演算してい
る。その平均パワー値より瞬間パワーが大きくなったと
きに、発話が開始したと判断し、音声信号の取り込みを
開始する。その音声信号は単語選択部52において単語
認識される。この際最初に使われる単語辞書は部分辞書
決定部51が外部記憶装置503から読み込んだ最下位
階層にある部分辞書である。
The signal processor 5 calculates the average power of the audio signal digitized from the microphone 500 until the switch 505 is operated in the audio input unit 53. When the instantaneous power becomes larger than the average power value, it is determined that the utterance has started, and the capture of the audio signal is started. The voice signal is subjected to word recognition in the word selection unit 52. At this time, the word dictionary used first is the partial dictionary in the lowest hierarchy read from the external storage device 503 by the partial dictionary determination unit 51.

【0029】単語出力部54はその認識の結果と単語の
上位階層の名称をモニタ60に一時出力するが、装置と
しての出力は保留する。この間モニタ60上に表示され
た上位階層の名称をタッチ入力すれば部分辞書の変更が
できる。所定時間内にタッチパネルへのタッチが無けれ
ば、単語出力部54は認識の結果を装置の出力として出
力する。
The word output unit 54 temporarily outputs the result of the recognition and the name of the upper layer of the word to the monitor 60, but suspends the output as a device. During this time, the partial dictionary can be changed by touch-inputting the name of the upper hierarchy displayed on the monitor 60. If there is no touch on the touch panel within the predetermined time, the word output unit 54 outputs the recognition result as an output of the device.

【0030】部分辞書の変更があった場合は、部分辞書
決定部51はタッチされた名称の部分辞書を外部記憶装
置503から読み込み、装置使用者はスイッチ505を
操作して下位辞書を決定するための発話をする。部分辞
書決定部51はそれに係わる所定の最下位部分辞書を決
定し、外部記憶装置503から読み込む。その後再び音
声入力をし単語選択部52において単語認識される。最
初に入力するあるいは階層変更後に入力される最下層の
部分辞書は前回使用した部分辞書あるいは使用頻度の最
も高い部分辞書である。
When there is a change in the partial dictionary, the partial dictionary determination unit 51 reads the partial dictionary having the touched name from the external storage device 503, and the device user operates the switch 505 to determine the lower-level dictionary. To speak. The partial dictionary determining unit 51 determines a predetermined lowest-order partial dictionary related thereto and reads it from the external storage device 503. Thereafter, voice input is performed again, and word recognition is performed in the word selection unit 52. The lowest partial dictionary input first or after the hierarchy change is the partial dictionary used last time or the most frequently used partial dictionary.

【0031】次に、図3のフローチャートに従って装置
の作動の流れを説明する。まず、ステップ101におい
て、音声入力の初期状態で、階層の最下位の部分辞書を
最初の認識対象として部分辞書決定部51において設定
する。初期設定として最も使用頻度の高い部分辞書ある
いは前回使用された最下層の部分辞書が用いられるが、
ここでは例えば図2中の部分辞書hが設定されたとす
る。
Next, the operation flow of the apparatus will be described with reference to the flowchart of FIG. First, in step 101, in the initial state of voice input, the partial dictionary determination unit 51 sets the lowest partial dictionary as the first recognition target in the hierarchy. The most frequently used partial dictionary or the lowest partial dictionary used last time is used as the initial setting,
Here, for example, it is assumed that the partial dictionary h in FIG. 2 has been set.

【0032】装置使用者はスイッチ505を操作して、
マイク500に発話をする。ステップ102では、音声
入力部53がスイッチ505が操作されたか否かを検出
し、操作された場合、マイク500が動作してからの音
声の平均パワー値を算出してステップ103へ進む。ス
テップ103では、決定された部分辞書を前記外部記憶
装置503から部分辞書決定部51に読み込む。ここで
は部分辞書hが読み込まれる。
The device user operates the switch 505 to
Speak into the microphone 500. In step 102, the voice input unit 53 detects whether or not the switch 505 has been operated. If the switch 505 has been operated, the average power value of the voice after the microphone 500 has been operated is calculated, and the process proceeds to step 103. In step 103, the determined partial dictionary is read from the external storage device 503 into the partial dictionary determination unit 51. Here, the partial dictionary h is read.

【0033】ステップ104では、使用者の発話音声取
り込みを開始する。すなわち入力した音声信号を絶えず
算出した音声の平均パワーと比べ、瞬間パワーが所定値
以上大きくなったとき、使用者の発話開始と判断し、発
話音声の取り込みを開始する。ここでは、例えば使用者
は目的語として「桜木町」を発話したとする。ステップ
105では、部分辞書決定部51に読み込まれた部分辞
書h内の単語と発話音声との一致度を単語選択部52に
おいて演算する。なお、ここでの処理は音声信号をとり
ながら所定時間で区切った音声区間部分と各単語との比
較が行なわれており、音声信号の取り込みはそれに同時
に進行されている。
In step 104, the utterance voice of the user is started. That is, when the instantaneous power becomes larger than a predetermined value as compared with the average power of the continuously calculated voice signal of the input voice signal, it is determined that the user's utterance has started, and the capturing of the utterance voice starts. Here, for example, it is assumed that the user has spoken “Sakuragicho” as the object. At step 105, the word selection unit 52 calculates the degree of coincidence between the word in the partial dictionary h read by the partial dictionary determination unit 51 and the uttered voice. In this process, a speech section portion divided by a predetermined time is compared with each word while taking a speech signal, and the acquisition of the speech signal is proceeding at the same time.

【0034】ステップ106で、音声信号の瞬間パワー
が平均パワー以下になったとき、使用者が発話完了と判
断し、音声取り込み終了してステップ107へ進む。ス
テップ107では、単語選択部52で演算された一致度
の最も高い単語を認識の結果として単語出力部54に出
力する。
At step 106, when the instantaneous power of the audio signal falls below the average power, the user determines that the utterance has been completed, ends the voice capture, and proceeds to step 107. In step 107, the word having the highest matching degree calculated by the word selection unit 52 is output to the word output unit 54 as a recognition result.

【0035】ステップ108では、単語出力部54はそ
の認識の結果と単語の上位階層の名称をモニタ60に一
時出力し表示させるが、装置としての出力は保留する。
図4はモニタ60の表示画面である。hが選択された単
語の辞書で、a、e、f、gはその単語の上位辞書であ
る。hの枠から表示されているのは目的語であり、その
上の枠に並べられているのが目的語に関連する上位辞書
内の単語で、右から左へ階層順位の増加を表示してい
る。階層を変更する場合に、単語が映っているところを
タッチすれば、階層の変更ができる。そしてステップ1
09で、所定時間内にタッチパネルへのタッチ入力が無
ければ、ステップ110において「桜木町」という単語
の認識結果を装置の出力として出力する。この場合、階
層を変更する必要がなく、発話回数が一回で終了する。
In step 108, the word output unit 54 temporarily outputs the result of the recognition and the name of the upper layer of the word to the monitor 60 for display, but suspends the output as the device.
FIG. 4 shows a display screen of the monitor 60. h is a dictionary of the selected word, and a, e, f, and g are high-order dictionaries of the word. The object displayed from the box of h is the object, and the words arranged in the box above it are words in the upper dictionary related to the object, and the hierarchical order is displayed from right to left. I have. When changing the hierarchy, the user can touch the place where the word is displayed to change the hierarchy. And step 1
In step 09, if there is no touch input to the touch panel within a predetermined time, in step 110, the recognition result of the word "Sakuragicho" is output as an output of the apparatus. In this case, there is no need to change the hierarchy, and the number of utterances ends once.

【0036】もし、使用者がステップ104で「そご
う」と発話していたとすれば、初期の部分辞書がhであ
ったため、他の単語が表示されてしまう。その場合、上
位の階層を変更する必要が生じるため、使用者は図4中
のeをタッチ入力する。
If the user has uttered "SOGO" in step 104, another word will be displayed because the initial partial dictionary is h. In this case, since it is necessary to change the upper layer, the user touches e in FIG.

【0037】そのタッチ入力がステップ109で検出さ
れると処理がステップ111へ進む。ステップ111で
は、タッチされた階層をどの階層に変更可能かモニタ6
0に表示する。ここでは「駅」、「デパート」、「ホテ
ル」を表示する。図5はその表示様子を示している。ス
テップ112では、部分辞書決定部51が新たに部分辞
書eを外部記憶装置503から読み込む。
When the touch input is detected in step 109, the process proceeds to step 111. In step 111, the monitor 6 determines which layer the touched layer can be changed to.
Display at 0. Here, “station”, “department store”, and “hotel” are displayed. FIG. 5 shows the display state. In step 112, the partial dictionary determination unit 51 newly reads the partial dictionary e from the external storage device 503.

【0038】ステップ113では、タッチして表示され
た上位階層の中から選択して音声入力する。「そごう」
の場合は使用者はそれに関連のある「デパート」を音声
で入力し、単語選択部52で認識される。本ステップは
ステップ102からステップ107までと同様の処理を
行なってもよい。ステップ114では、部分辞書決定部
51が、認識された単語「デパート」に繋がる最下層の
部分辞書kを決定する。
In step 113, the user selects from the upper hierarchy displayed by touching and inputs a voice. "Sogo"
In the case of (1), the user inputs the "department store" related thereto by voice and is recognized by the word selection section 52. In this step, the same processing as steps 102 to 107 may be performed. In step 114, the partial dictionary determination unit 51 determines the lowermost partial dictionary k connected to the recognized word “department”.

【0039】ステップ115では、再度音声入力やり直
しを使用者に報知するため、モニタ60に再度入力の指
示マークを表示して、ステップ102に戻る。その後使
用者が「そごう」を発話してステップ102からステッ
プ108までの処理が再び行なわれて、ステップ110
において、「そごう」という単語が出力される。
In step 115, an input instruction mark is displayed again on the monitor 60 in order to notify the user of the voice input retry again, and the process returns to step 102. Thereafter, the user utters “SOGO” and the processing from step 102 to step 108 is performed again.
In, the word “SOGO” is output.

【0040】マイク500、A/Dコンバータ501、
音声入力部53は音声入力手段を構成している。部分辞
書決定部51は辞書決定手段を構成している。単語選択
部52は単語選択手段を構成している。モニタ60は表
示手段を構成している。
Microphone 500, A / D converter 501,
The voice input unit 53 constitutes voice input means. The partial dictionary determination section 51 constitutes a dictionary determination unit. The word selecting section 52 constitutes a word selecting means. The monitor 60 constitutes a display unit.

【0041】本実施例は以上のように構成され、第1の
音声の入力が行なわれた場合、音声信号と所定の最下層
の部分辞書の単語との一致度が演算されるので、第1の
単語を目的語とすることができ、使用者の入力負担が軽
減されることになる。また、部分辞書の変更が必要なと
き、上位の階層の下に複数の階層が存在しても、その最
下層の部分辞書を設定して使用するので、階層変更によ
る入力負担増加が少ない。そして、最初に使用する部分
辞書を使用頻度の高いものあるいは前回使用したものと
しているので、使用者は出力結果を容易に理解し、階層
変更が円滑に行なえる。
The present embodiment is configured as described above. When the first speech is input, the degree of coincidence between the speech signal and a word in a predetermined lowermost partial dictionary is calculated. Can be used as the object, and the input burden on the user is reduced. Further, when the partial dictionary needs to be changed, even if there are a plurality of layers below the upper layer, the lowest part of the partial dictionary is set and used. Then, since the partial dictionary used first is the one frequently used or the one used last time, the user can easily understand the output result and smoothly change the hierarchy.

【0042】また、モニタが単語を表示するときに単語
の上位階層構造を表示するので、どの階層を変更したい
かを容易に判断できる。タッチ入力があったとき、音声
入力によってどのような階層に変更可能かを表示するこ
とで、辞書の構成を知らなくても、どの階層に変更する
かの操作ができる。
Further, since the monitor displays the upper hierarchical structure of the word when displaying the word, it is possible to easily determine which layer the user wants to change. When there is a touch input, by displaying what hierarchy can be changed by voice input, it is possible to operate to change to which hierarchy without knowing the dictionary configuration.

【0043】図6は、第2の実施例の構成を示すブロッ
ク図である。この実施例は、図1に示す第1の実施例に
おける信号処理装置5の代わりに音声記憶部66を設け
た信号処理装置6を用いる。音声信号記憶部66は音声
入力部63の音声信号を単語選択部62に出力するとと
もに、一回目の音声信号は音声記憶部66が記憶する。
部分辞書変更があった場合、音声記憶部66に記憶され
た音声信号を用いて部分辞書との一致度を判定し、目的
語を認識する。その他は第1の実施例と同様である。
FIG. 6 is a block diagram showing the configuration of the second embodiment. In this embodiment, a signal processing device 6 provided with an audio storage unit 66 is used instead of the signal processing device 5 in the first embodiment shown in FIG. The voice signal storage unit 66 outputs the voice signal of the voice input unit 63 to the word selection unit 62, and stores the first voice signal in the voice storage unit 66.
When there is a change in the partial dictionary, the degree of coincidence with the partial dictionary is determined using the voice signal stored in the voice storage unit 66, and the object is recognized. Others are the same as the first embodiment.

【0044】次に、図7のフローチャートに従って装置
の作動の流れを説明する。ステップ201〜ステップ2
15まではスッテプ207を除いて第1の実施例におけ
る図3のフローチャートのステップ101〜ステップ1
14と同様の処理を行なう。すなわちまず、階層の最下
位の部分辞書を最初の認識対象として設定し、外部記憶
装置503から読み込む。使用者の発話音声信号を取り
込みながら設定された部分辞書内の単語と一致度を演算
する。そして音声信号の取り込みが終了すると、ステッ
プ207において、音声記憶部66が音声信号を記憶す
る。一致度の最も高い単語が認識の結果として選択さ
れ、上位階層を含めてモニタ60において表示される。
Next, the operation flow of the apparatus will be described with reference to the flowchart of FIG. Step 201 to Step 2
Steps 15 to 15 in the flowchart of FIG.
The same processing as in step 14 is performed. That is, first, the partial dictionary at the lowest level of the hierarchy is set as the first recognition target, and is read from the external storage device 503. A word in the set partial dictionary and the degree of coincidence are calculated while taking in the user's uttered voice signal. Then, when the capture of the audio signal is completed, in step 207, the audio storage unit 66 stores the audio signal. The word having the highest matching degree is selected as a result of the recognition, and is displayed on the monitor 60 including the upper hierarchy.

【0045】そして、所定時間内にモニタ60のタッチ
パネルに対するタッチ入力が無ければ、選択された単語
を装置の出力として出力するが、タッチ入力があった場
合、タッチされた階層に部分辞書の変更を行なう。階層
変更された部分辞書が新たに外部記憶装置503から読
み込まれる。その後階層変更のための発話をし、音声入
力によって階層を示す単語が特定されると、その階層下
の最下位部分辞書を外部記憶装置503から読み込む。
If there is no touch input on the touch panel of the monitor 60 within a predetermined time, the selected word is output as an output of the device. If there is a touch input, the partial dictionary is changed to the touched hierarchy. Do. The partial dictionary whose hierarchy has been changed is newly read from the external storage device 503. Thereafter, an utterance for changing the hierarchy is made, and when a word indicating the hierarchy is specified by voice input, the lowest partial dictionary under the hierarchy is read from the external storage device 503.

【0046】その後、第1の実施例では再度音声入力や
り直しを使用者に報知するが、本実施例では、ステップ
216において、音声記憶部66から記憶した音声信号
を用い、最下層の部分辞書内の単語との一致度を演算す
る。各単語から一致度の最も高い単語がステップ208
において選択される。そしてステップ210にモニタ6
0へのタッチ入力がないと判定されると、ステップ21
1で認識した単語を装置の出力として出力する。
Thereafter, in the first embodiment, the user is notified again of the re-input of the voice. In the present embodiment, however, in step 216, the voice signal stored from the voice storage unit 66 is used to store the voice in the lowermost partial dictionary. Calculate the degree of matching with the word. The word having the highest matching degree from each word is determined in step 208.
Selected in Then, in step 210, monitor 6
If it is determined that there is no touch input to 0, step 21
The word recognized in 1 is output as an output of the device.

【0047】本実施例によっても、第1の実施例と同様
の効果が得られるとともに、部分辞書の変更後、再度最
初に目的語を発話し音声入力する必要がないので、入力
回数が減少されることになる。
According to this embodiment, the same effect as that of the first embodiment can be obtained. Further, since it is not necessary to first speak and input the object again after the partial dictionary is changed, the number of times of input is reduced. Will be.

【0048】次に、第3の実施例について説明する。図
8は、第3の実施例の構成を示すブロック図である。こ
の実施例は、図6に示す第2の実施例における信号処理
装置6の代わりに信号処理装置7を用いる。音声信号記
憶部66は音声入力部73の音声信号を単語選択部62
に出力するとともに、一回目の音声信号は音声記憶部6
6が記憶する。部分辞書変更がある場合、音声記憶部6
6に記憶された音声信号を用いて認識する。
Next, a third embodiment will be described. FIG. 8 is a block diagram showing the configuration of the third embodiment. In this embodiment, a signal processing device 7 is used instead of the signal processing device 6 in the second embodiment shown in FIG. The voice signal storage unit 66 stores the voice signal of the voice input unit 73 into the word selection unit 62.
The first audio signal is output to the audio storage unit 6
6 memorizes. If there is a partial dictionary change, the voice storage unit 6
6. Recognition is performed by using the audio signal stored in 6.

【0049】第1あるいは第2の実施例では階層の変更
はモニタ60のタッチパネルで行なったが、本実施例は
音声による階層変更を行なう。すなわち第一回目の音声
入力を目的語に対応させ、以後の音声入力は階層変更に
対応させている点が第1あるいは第2の実施例と異な
る。モニタ60の代わりに表示のみのモニタ70を使用
する。その他は第1の実施例と同様である。
In the first or second embodiment, the hierarchy is changed by the touch panel of the monitor 60. In this embodiment, the hierarchy is changed by voice. That is, the first and second embodiments differ from the first and second embodiments in that the first speech input corresponds to the object and the subsequent speech inputs correspond to the hierarchy change. A monitor 70 for display only is used instead of the monitor 60. Others are the same as the first embodiment.

【0050】次に、図9のフローチャートに従って装置
の作動の流れを説明する。ステップ201〜ステップ2
16まではステップ310、ステップ313を除いて第
2の実施例における図7のフローチャートと同様の処理
を行なう。すなわちまず、階層の最下位の部分辞書を最
初の認識対象として設定され、外部記憶装置503から
読み込まれる。使用者の発話音声信号を取り込みながら
設定された部分辞書内の単語と一致度を演算する。そし
て音声信号の取り込みが終了すると、音声記憶部66が
音声信号を記憶する。そしてステップ208において一
致度の最も高い単語が認識の結果として選択され、ステ
ップ209において上位階層を含めてモニタ70におい
て表示されるとステップ310へ進む。
Next, the operation flow of the apparatus will be described with reference to the flowchart of FIG. Step 201 to Step 2
Up to 16, processes similar to those in the flowchart of FIG. 7 in the second embodiment are performed except for steps 310 and 313. That is, first, the lowest partial dictionary in the hierarchy is set as the first recognition target, and is read from the external storage device 503. A word in the set partial dictionary and the degree of coincidence are calculated while taking in the user's uttered voice signal. Then, when the acquisition of the audio signal is completed, the audio storage unit 66 stores the audio signal. Then, in step 208, the word having the highest degree of coincidence is selected as a result of the recognition, and in step 209, when the word including the upper layer is displayed on the monitor 70, the process proceeds to step 310.

【0051】ステップ310では、音声入力部73がス
イッチ505が操作されたか否かを検出し、操作された
場合、音声入力があるため、ステップ313へ進む。ス
テップ313では現在使用中の部分辞書の上位全ての部
分辞書を外部記憶装置503から読み込む。その後第
1、第2の実施例と同様に上位階層の音声入力がなさ
れ、階層を示す単語が特定されると、その階層下の最下
位部分辞書を外部記憶装置503から読み込む。
In step 310, the voice input section 73 detects whether or not the switch 505 has been operated. If the switch 505 has been operated, the flow proceeds to step 313 because there is a voice input. In step 313, all the partial dictionaries higher than the currently used partial dictionary are read from the external storage device 503. After that, as in the first and second embodiments, voice input of the upper hierarchy is performed, and when a word indicating the hierarchy is specified, the lowest partial dictionary under the hierarchy is read from the external storage device 503.

【0052】ステップ216において、音声記憶部66
から記憶した音声信号を用い、部分辞書内の単語との一
致度を演算する。各単語から一致度の最も高い単語がス
テップ208において選択される。ステップ310にお
いてスイッチ505に対する操作がないことが検出され
ると、ステップ211で認識した単語を装置の出力とし
て出力する。このように、例えば最初に部分辞書hが設
定されていて、入力する単語は辞書にない「そごう」の
場合、使用者が「デパート」を発話すれば、図10に示
すように「そごう」の上位辞書を含むハッチングした部
分辞書が自動的に設定されるので、階層変更があっても
煩わしい操作はない。
In step 216, the voice storage unit 66
Then, the degree of coincidence with the word in the partial dictionary is calculated by using the speech signal stored from. From each word, the word with the highest degree of matching is selected in step 208. If it is detected in step 310 that there is no operation on the switch 505, the word recognized in step 211 is output as an output of the device. In this way, for example, when the partial dictionary h is initially set and the word to be input is “Sogo” that is not in the dictionary, if the user speaks “Department store”, “Sogo” as shown in FIG. Since the hatched partial dictionaries including the upper dictionary are automatically set, there is no troublesome operation even if there is a hierarchy change.

【0053】すなわち、最初の発話に対し、装置が使用
者の指定したい階層と異なる階層の単語を出力する場
合、使用者が指定したい階層名称を発話するという手順
をとることで、発話、返答、詳しい説明という自然な対
話を実現でき、最も自然な感覚で音声入力ができる効果
が得られる。各階層の部分辞書が設定されるので、音声
による認識率の低下が懸念されるが、一般に上位の部分
辞書内の単語は、分類のための単語であり、単語数は最
下層の部分辞書内の単語数に比較して極めて少ないた
め、認識率は十分に使用に堪える。
That is, when the apparatus outputs words of a different hierarchy from the hierarchy specified by the user in response to the first utterance, the user can utter the name of the hierarchy desired by the user, so that the utterance, reply, A natural conversation with detailed explanations can be realized, and an effect that voice input can be performed with the most natural feeling can be obtained. Since the partial dictionaries of each hierarchy are set, there is a concern that the recognition rate may decrease due to speech.However, words in the upper partial dictionary are generally words for classification, and the number of words is in the lowest partial dictionary. Since the number of words is extremely small as compared with the number of words, the recognition rate is sufficient for use.

【0054】次に、第3の実施例の変形例について説明
する。前記各実施例では、最下層の部分辞書が正確に設
定されても、騒音などの影響で単語を誤認識することが
ある。この場合すべての上位階層の部分辞書を設定し直
しても辿り着いたのはもとの部分辞書であり、余計な入
力回数を作る結果となる。この変形例はそれに対処する
ためのもので、誤認識があっても、同じ部分辞書内の単
語なら、階層変更せずに処理できるようにしている。
Next, a modification of the third embodiment will be described. In each of the above embodiments, even when the partial dictionary at the lowest level is set correctly, words may be erroneously recognized due to the influence of noise or the like. In this case, even if the partial dictionaries of all the upper layers are reset, the original partial dictionary is reached, which results in an unnecessary input count. This modification is intended to cope with this, and even if there is an erroneous recognition, words in the same partial dictionary can be processed without changing the hierarchy.

【0055】図11は装置作動の流れを示すフローチャ
ートである。このフローチャートはステップ412、ス
テップ413、ステップ414を図9の第3の実施例に
おけるステップ313、ステップ214に置き換えて構
成される。すなわちまず、階層の最下位の部分辞書を最
初の認識対象として設定され、外部記憶装置503から
読み込まれる。使用者の発話音声信号を取りながら設定
された部分辞書内の単語と一致度を演算する。そして音
声信号の取り込みが終了すると、音声記憶部66が音声
信号を記憶する。一致度の最も高い単語が認識の結果と
して選択され、ステップ209において上位階層を含め
てモニタ70において表示される。
FIG. 11 is a flowchart showing the flow of the operation of the apparatus. This flowchart is configured by replacing Steps 412, 413, and 414 with Steps 313 and 214 in the third embodiment of FIG. That is, first, the lowest partial dictionary in the hierarchy is set as the first recognition target, and is read from the external storage device 503. While taking the user's utterance voice signal, the word and the degree of coincidence in the set partial dictionary are calculated. Then, when the acquisition of the audio signal is completed, the audio storage unit 66 stores the audio signal. The word having the highest matching degree is selected as a result of the recognition, and is displayed on the monitor 70 including the upper layer in step 209.

【0056】そして、ステップ310において、音声入
力部73がスイッチ505が操作されたか否かを検出
し、操作された場合、音声入力があるため、ステップ4
12へ進む。ステップ412では、現在使用中の部分辞
書およびその上位全ての部分辞書を読み込む。ステップ
413では、使用者が発話した音声を入力し、読み込ま
れた部分辞書内の単語との一致度を演算する。
In step 310, the voice input unit 73 detects whether or not the switch 505 has been operated. If the switch 505 has been operated, there is a voice input.
Proceed to 12. In step 412, the currently used partial dictionary and all the partial dictionaries thereof are read. In step 413, the speech uttered by the user is input, and the degree of coincidence with the word in the read partial dictionary is calculated.

【0057】ステップ414では一致度の最も高い単語
が上位階層の部分辞書内にあるか否かを判定する。単語
が上位階層の部分辞書内にある場合、階層変更するため
のものであり、ステップ215へ進み、第3の実施例と
同様の処理を行なう。単語が最下層の部分辞書内にある
場合、既に認識された目的語であるためステップ209
へ進み、ステップ310でスイッチに対する新たな操作
が行なわれていないかの判定を経てステップ211にお
いて単語が出力される。
At step 414, it is determined whether or not the word having the highest degree of coincidence exists in the partial dictionary of the upper hierarchy. When the word is in the partial dictionary of the upper hierarchy, it is for changing the hierarchy, and the process proceeds to step 215 to perform the same processing as in the third embodiment. If the word is in the lowermost sub-dictionary, it is an already recognized object, so that step 209 is executed.
The process proceeds to step 310, where it is determined whether a new operation has been performed on the switch. In step 211, a word is output.

【0058】このように、もし最初に部分辞書hが設定
されていて、入力したい単語は「桜木町」の場合、騒音
などによる「桜木町」の音声信号に歪みが生じ、「関
内」と誤って表示される可能性がある。また「桜木町」
を発話するつもりで「関内」と誤って発話することが考
えられる。このような部分辞書が正しい場合でも、図1
2にハッチングで示すように部分辞書hを含む全ての上
位階層の部分辞書が自動的に設定されるので、「桜木
町」と発話し直しすれば、階層変更をせずに音声入力が
できる。これによって階層の変更は必要なときのみ行な
われることになり、音声入力負担がさらに軽減される。
このほか第3の実施例のように発話、返答、詳しい説明
という自然な対話形式で音声入力ができる効果も得られ
る。
As described above, if the partial dictionary h is initially set and the word to be input is "Sakuragicho", the voice signal of "Sakuragicho" is distorted due to noise or the like, and is incorrectly identified as "Kannai". May be displayed. Also "Sakuragicho"
Erroneously speaking "Kannai". Even if such a partial dictionary is correct, FIG.
Since the partial dictionaries of all upper layers including the partial dictionary h are automatically set as indicated by hatching in FIG. 2, if "Sakuragicho" is re-uttered, voice input can be performed without changing the layer. As a result, the hierarchy is changed only when necessary, and the voice input burden is further reduced.
In addition, as in the third embodiment, there is also obtained an effect that a speech input can be performed in a natural dialogue style such as utterance, reply, and detailed explanation.

【0059】[0059]

【発明の効果】以上説明したように、本発明によれば、
1回目の単語を目的語とすることができるので、まず目
的語で入力し、正確でなかったら階層を変更して目的語
を再度入力して認識し、自然な感覚で音声入力ができ
る。また辞書を変更する際、例えば現在の部分辞書から
遡り、目的語と共通する内容の上位部分辞書に変更して
最下層の部分辞書を決定する段取りをとるので、目的語
が認識されるまでの入力回数が少なく、入力負担が軽減
される
As described above, according to the present invention,
Since the first word can be used as the object, the user can first input the object, and if it is not accurate, change the hierarchy and re-enter the object to recognize it, and can input speech with a natural feeling. Also, when changing the dictionary, for example, go back from the current partial dictionary, change to a higher partial dictionary with the same content as the object, and take steps to determine the lowermost partial dictionary, so that until the object is recognized Fewer inputs and less typing

【0060】部分辞書決定手段は、上位階層への変更指
示を受けた場合、上位階層の部分辞書を決定し、音声入
力により単語選択手段が選択した単語が示す階層に変更
し、該階層下の所定の最下層部分辞書を決定するように
した場合、階層変更があったも、目的語の入っている部
分辞書に至るまでの部分辞書に対する認識が不要で、大
きな入力負担とならない効果が得られる。また部分辞書
決定手段が初めに決定する部分辞書は前回使用された最
下層の部分辞書となれば、辞書構成を理解し、階層変更
が円滑に行なえる。さらに前回使用された最下層の部分
辞書の代わりに使用頻度の高い最下層の部分辞書を用い
ても同じ効果が得られる。
When the instruction to change to the upper hierarchy is received, the partial dictionary determination means determines the partial dictionary in the upper hierarchy, changes the hierarchy to the hierarchy indicated by the word selected by the word selection means by voice input, and changes the hierarchy to the lower hierarchy. In the case where the predetermined lower-level partial dictionary is determined, there is no need to recognize the partial dictionary up to the partial dictionary containing the object even if the hierarchy is changed, so that an effect of not causing a large input burden can be obtained. . Further, if the partial dictionary determined first by the partial dictionary determining means is the lowest partial dictionary used last time, the dictionary configuration can be understood and the hierarchy can be changed smoothly. Further, the same effect can be obtained by using a lower-level partial dictionary frequently used in place of the lowest-level partial dictionary used last time.

【0061】音声入力装置に表示手段を接続し、単語選
択手段が選択した単語及び上位の階層構造を表示手段で
表示するようにすると、階層変更が必要のときに、表示
された階層構造を頼りに階層変更ができるので、入力負
担がさらに軽減される。前記表示手段はタッチパネルを
合わせ持ち、使用者によってタッチされた階層に対応し
た階層変更指令を前記部分辞書決定部に出力するように
すると、対話式な入力となり一層の入力負担軽減が図ら
れる。また、タッチ入力があったときに、音声入力によ
って変更可能な階層を前記表示手段上で表示した場合、
辞書の構造を表示画面から知ることができ辞書の構成に
対する予備知識が無くても階層変更ができる効果が得ら
れる。
When the display means is connected to the voice input device and the word selected by the word selection means and the upper hierarchical structure are displayed on the display means, when the hierarchical change is required, the displayed hierarchical structure is relied on. Since the hierarchy can be changed, the input burden can be further reduced. When the display means has a touch panel and outputs a hierarchy change command corresponding to the hierarchy touched by the user to the partial dictionary determination unit, the input becomes interactive and the input load is further reduced. Further, when there is a touch input, when a hierarchy that can be changed by voice input is displayed on the display unit,
The structure of the dictionary can be known from the display screen, and the effect of changing the hierarchy without any prior knowledge of the dictionary configuration can be obtained.

【0062】前記部分辞書決定手段は、目的語に続き音
声入力が行なわれたときに、上位階層への変更指示を受
けたとし、設定されている最下層部分辞書より上位の全
ての部分辞書を決定し、単語選択手段が選択した単語が
示す階層に階層変更を行なうようにすると、階層変更に
対する操作が不要で入力負担が一層軽減される。
The partial dictionary determining means determines that a change instruction to a higher hierarchical level has been received when a voice is input following the object, and determines all partial dictionaries higher than the set lowermost partial dictionary. When the determination is made and the hierarchy is changed to the hierarchy indicated by the word selected by the word selecting means, an operation for the hierarchy change is unnecessary and the input load is further reduced.

【0063】前記部分辞書決定手段は、目的語に続き音
声入力が行なわれたときに、設定されている最下層部分
辞書及びそれより上位の全ての部分辞書を決定し、単語
選択手段が選択した単語が上位部分辞書内の単語であれ
ば、上位階層への変更指示を受けたとし、前記単語が示
す階層に階層変更を行なった場合、部分辞書が正確であ
りながら、単語の誤認識による階層変更が防止され、無
駄な入力を生じさせない効果が得られる。
The sub-dictionary deciding means decides the lowermost sub-dictionary set and all the sub-dictionaries higher than the set sub-dictionary when the speech is input following the object, and the word selecting means selects the sub-dictionary. If the word is a word in the upper partial dictionary, it is assumed that a change instruction to the upper hierarchy has been received, and if the hierarchy is changed to the hierarchy indicated by the word, the hierarchy due to incorrect recognition of the word while the partial dictionary is accurate is changed. The change is prevented, and an effect of not causing useless input is obtained.

【0064】最初に入力された音声を記憶する記憶手段
を合わせ持ち、階層が変更されたとき、前記単語選択手
段は前記記憶手段に記憶された音声を用いて、決定され
た最下層の部分辞書内の単語との一致度を演算し、出力
することで、階層を変更し、最下層の部分辞書が変わっ
ても、目的語の音声入力を二度と行なう必要がなく、音
声の入力回数がさらに減らされる。また、最初の発話に
対し、装置が使用者の指定したい階層と異なる階層の単
語を出力する場合、第2の実施例と同様に使用者が指定
したい階層名称を発話するという手順をとることで、発
話、返答、詳しい説明という自然な対話を実現でき、最
も自然な感覚で音声入力ができる効果が得られる。
When the hierarchy is changed, the word selection means uses the speech stored in the storage means to determine a partial dictionary of the lowest layer when the hierarchy is changed. By calculating the degree of coincidence with the words in the word and outputting it, even if the hierarchy is changed and the partial dictionary at the bottom layer changes, there is no need to input the target word again, and the number of times of voice input is further reduced. It is. Further, when the apparatus outputs words of a different hierarchy from the hierarchy desired by the user in response to the first utterance, the user can utter the hierarchy name desired by the user in the same manner as in the second embodiment. , Utterance, response, and detailed explanation can be realized, and the effect of voice input with the most natural feeling can be obtained.

【図面の簡単な説明】[Brief description of the drawings]

【図1】第1の実施例の構成を示すブロック図である。FIG. 1 is a block diagram showing a configuration of a first embodiment.

【図2】階層化した単語辞書の構成を示す図である。FIG. 2 is a diagram showing a configuration of a hierarchical word dictionary.

【図3】第1の実施例のフローチャートである。FIG. 3 is a flowchart of the first embodiment.

【図4】モニタ表示画面を示すブロック図である。FIG. 4 is a block diagram showing a monitor display screen.

【図5】部分辞書を変更時のモニタ表示画面を示す図で
ある。
FIG. 5 is a diagram showing a monitor display screen when a partial dictionary is changed.

【図6】第2の実施例の構成を示すブロック図である。FIG. 6 is a block diagram showing a configuration of a second embodiment.

【図7】第2の実施例のフローチャートである。FIG. 7 is a flowchart of the second embodiment.

【図8】第3の実施例の構成を示すブロック図である。FIG. 8 is a block diagram showing a configuration of a third embodiment.

【図9】第3の実施例のフローチャートである。FIG. 9 is a flowchart of the third embodiment.

【図10】最下位部分辞書を変更時に仮設定された上位
部分辞書を示す図である。
FIG. 10 is a diagram showing an upper partial dictionary temporarily set when the lowermost partial dictionary is changed.

【図11】第3の実施例の変形例を示すフローチャート
である。
FIG. 11 is a flowchart showing a modification of the third embodiment.

【図12】最下位部分辞書を変更時に仮設定された部分
辞書を示す図である。
FIG. 12 is a diagram illustrating a partial dictionary temporarily set when the lowest partial dictionary is changed.

【図13】従来例の構成を示すブロック図である。FIG. 13 is a block diagram showing a configuration of a conventional example.

【図14】従来例のフローチャートである。FIG. 14 is a flowchart of a conventional example.

【符号の説明】[Explanation of symbols]

5、6、7 信号処理装置 51、71 部分辞書決定部 52、62 単語選択部 53、63、73 音声入力部 54、64 単語出力部 60、70、504 モニタ 66 音声記憶部 500 マイク 501 A/Dコンバータ 502 信号処理装置 503 外部記憶装置 505 スイッチ a、b、c、d、e、f 部分辞書 g、h、i、j、k 部分辞書 5, 6, 7 Signal processing device 51, 71 Partial dictionary determination unit 52, 62 Word selection unit 53, 63, 73 Audio input unit 54, 64 Word output unit 60, 70, 504 Monitor 66 Audio storage unit 500 Microphone 501 A / D converter 502 Signal processing device 503 External storage device 505 Switch a, b, c, d, e, f partial dictionary g, h, i, j, k partial dictionary

Claims (10)

【特許請求の範囲】[Claims] 【請求項1】 発話音声を情報化して入力する音声入力
手段と、複数の単語を含む部分辞書からなるとともに階
層構造を持った単語辞書と、使用する部分辞書を決定す
る部分辞書決定手段と前記音声入力手段からの音声情報
と決定された部分辞書内の単語との一致度を演算する演
算手段と、演算された単語のうち最も一致度の高い単語
を選択して出力する単語選択手段とを有し、前記部分辞
書決定手段は、初めに所定の最下層の部分辞書を決定
し、階層変更指令を受けた場合に階層変更をして部分辞
書を新たに決定することを特徴とする音声入力装置。
1. A speech input means for converting an uttered speech into information and inputting the information, a word dictionary having a hierarchical structure including a partial dictionary including a plurality of words, a partial dictionary determining means for determining a partial dictionary to be used, and Calculating means for calculating the degree of matching between the voice information from the voice input means and the determined word in the partial dictionary; and word selecting means for selecting and outputting the word having the highest matching degree from the calculated words. Wherein the partial dictionary determining means first determines a predetermined lowermost partial dictionary, and when a hierarchical change command is received, changes the hierarchy and newly determines a partial dictionary. apparatus.
【請求項2】 前記部分辞書決定手段は、上位階層への
変更指示を受けた場合、指示された上位階層の部分辞書
を決定し、入力により前記単語選択手段が選択した単語
が示す階層に変更し、該階層下の所定の最下層部分辞書
を決定することを特徴とする請求項1記載の音声入力装
置。
2. When a change instruction to an upper hierarchy is received, the partial dictionary determination means determines a partial dictionary of the specified upper hierarchy, and changes to a hierarchy indicated by the word selected by the word selection means by input. 2. The voice input device according to claim 1, wherein a predetermined lowermost partial dictionary under the hierarchy is determined.
【請求項3】 前記部分辞書決定手段により初めに決定
される部分辞書は前回使用された最下層の部分辞書であ
ることを特徴とする請求項1または2記載の音声入力装
置。
3. The voice input device according to claim 1, wherein the partial dictionary first determined by the partial dictionary determining means is the lowest partial dictionary used last time.
【請求項4】 前記部分辞書決定手段により初めに決定
される部分辞書は使用頻度の高い最下層の部分辞書であ
ることを特徴とする請求項1または2記載の音声入力装
置。
4. The speech input device according to claim 1, wherein the partial dictionary determined first by said partial dictionary determining means is a lower-level partial dictionary frequently used.
【請求項5】 前記音声入力装置には表示手段が接続さ
れ、前記単語選択手段が選択した単語及び上位の階層構
造を表示するようにしたことを特徴とする請求項1また
は2記載の音声入力装置。
5. The voice input device according to claim 1, wherein a display unit is connected to the voice input device, and the word selected by the word selection unit and a higher hierarchical structure are displayed. apparatus.
【請求項6】 前記表示手段はタッチパネルを合わせ持
ち、使用者によってタッチされた階層に対応した階層変
更指令が前記部分辞書決定手段に出力されることを特徴
とする請求項5記載の音声入力装置。
6. The voice input device according to claim 5, wherein the display unit has a touch panel, and a hierarchy change command corresponding to a hierarchy touched by a user is output to the partial dictionary determination unit. .
【請求項7】 前記表示手段はタッチ入力があったとき
に、変更可能な階層を表示することを特徴とする請求項
6記載の音声入力装置。
7. The voice input device according to claim 6, wherein said display means displays a changeable hierarchy when a touch input is made.
【請求項8】 前記部分辞書決定手段は、目的語に続き
入力が行なわれたときに、上位階層への変更指示を受け
たとし、設定されている最下層部分辞書より上位の全て
の部分辞書を決定し、単語選択手段が選択した単語が示
す階層に階層変更を行なうことを特徴とする請求項1ま
たは2記載の音声入力装置。
8. The partial dictionary determining means determines that, when an input is made following an object, an instruction to change to a higher hierarchical level is received, and all partial dictionaries higher than the set lowest partial dictionary are set. 3. The voice input device according to claim 1, wherein the voice input device determines the hierarchy and changes the hierarchy to the hierarchy indicated by the word selected by the word selection unit. 4.
【請求項9】 前記部分辞書決定手段は、目的語に続き
入力が行なわれたときに、設定されている最下層部分辞
書及びそれより上位の全ての部分辞書を決定し、単語選
択手段が選択した単語が上位部分辞書内の単語であれ
ば、上位階層への変更指示を受けたとし、前記単語が示
す階層に階層変更を行なうことを特徴とする請求項1ま
たは2記載の音声入力装置。
9. The partial dictionary determining means, when an input is made following an object, determines a set lowest partial dictionary and all partial dictionaries higher than the set partial dictionary, and the word selecting means selects the partial dictionary. 3. The speech input device according to claim 1, wherein if the word is a word in the upper partial dictionary, a change instruction to a higher hierarchy is received, and the hierarchy is changed to the hierarchy indicated by the word.
【請求項10】 前記音声入力装置は最初に入力された
音声信号を記憶する記憶手段を合わせ持ち、階層が変更
されたとき、前記単語選択手段が前記記憶手段に記憶さ
れた音声信号を用いて、決定された最下層の部分辞書内
の単語との一致度を演算し、出力することを特徴とする
請求項1、2、8または9記載の音声入力装置。
10. The voice input device further has a storage unit for storing a voice signal input first, and when the hierarchy is changed, the word selection unit uses the voice signal stored in the storage unit. 10. The voice input device according to claim 1, wherein the degree of coincidence with the determined word in the lowermost partial dictionary is calculated and output.
JP13579397A 1997-05-12 1997-05-12 Voice input device Expired - Fee Related JP3588975B2 (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP13579397A JP3588975B2 (en) 1997-05-12 1997-05-12 Voice input device

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP13579397A JP3588975B2 (en) 1997-05-12 1997-05-12 Voice input device

Publications (2)

Publication Number Publication Date
JPH10312193A true JPH10312193A (en) 1998-11-24
JP3588975B2 JP3588975B2 (en) 2004-11-17

Family

ID=15159968

Family Applications (1)

Application Number Title Priority Date Filing Date
JP13579397A Expired - Fee Related JP3588975B2 (en) 1997-05-12 1997-05-12 Voice input device

Country Status (1)

Country Link
JP (1) JP3588975B2 (en)

Cited By (6)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2001109492A (en) * 1999-10-07 2001-04-20 Alpine Electronics Inc Voice recognition method
JP2007127813A (en) * 2005-11-02 2007-05-24 Canon Inc Speech recognition apparatus and setting method thereof
JP2008003266A (en) * 2006-06-22 2008-01-10 Alpine Electronics Inc Destination setting device and destination setting method
US7392194B2 (en) 2002-07-05 2008-06-24 Denso Corporation Voice-controlled navigation device requiring voice or manual user affirmation of recognized destination setting before execution
JP2017116714A (en) * 2015-12-24 2017-06-29 日本電信電話株式会社 Speech input device, method thereof, and program
CN110309504A (en) * 2019-05-23 2019-10-08 平安科技(深圳)有限公司 Text handling method, device, equipment and storage medium based on participle

Citations (2)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH08320697A (en) * 1995-05-23 1996-12-03 Hitachi Ltd Voice recognition device
JPH08328584A (en) * 1995-05-29 1996-12-13 Sony Corp Speech recognition device, speech recognition method and navigation device

Patent Citations (2)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH08320697A (en) * 1995-05-23 1996-12-03 Hitachi Ltd Voice recognition device
JPH08328584A (en) * 1995-05-29 1996-12-13 Sony Corp Speech recognition device, speech recognition method and navigation device

Cited By (7)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2001109492A (en) * 1999-10-07 2001-04-20 Alpine Electronics Inc Voice recognition method
US7392194B2 (en) 2002-07-05 2008-06-24 Denso Corporation Voice-controlled navigation device requiring voice or manual user affirmation of recognized destination setting before execution
JP2007127813A (en) * 2005-11-02 2007-05-24 Canon Inc Speech recognition apparatus and setting method thereof
JP2008003266A (en) * 2006-06-22 2008-01-10 Alpine Electronics Inc Destination setting device and destination setting method
JP2017116714A (en) * 2015-12-24 2017-06-29 日本電信電話株式会社 Speech input device, method thereof, and program
CN110309504A (en) * 2019-05-23 2019-10-08 平安科技(深圳)有限公司 Text handling method, device, equipment and storage medium based on participle
CN110309504B (en) * 2019-05-23 2023-10-31 平安科技(深圳)有限公司 Text processing methods, devices, equipment and storage media based on word segmentation

Also Published As

Publication number Publication date
JP3588975B2 (en) 2004-11-17

Similar Documents

Publication Publication Date Title
US8279171B2 (en) Voice input device
US8818816B2 (en) Voice recognition device
US7949524B2 (en) Speech recognition correction with standby-word dictionary
US9239829B2 (en) Speech recognition device
US8751145B2 (en) Method for voice recognition
JP3476007B2 (en) Recognition word registration method, speech recognition method, speech recognition device, storage medium storing software product for registration of recognition word, storage medium storing software product for speech recognition
US7027565B2 (en) Voice control system notifying execution result including uttered speech content
CN107112007B (en) Speech recognition device and speech recognition method
JP2002116793A (en) Data input system and method
US7295923B2 (en) Navigation device and address input method thereof
US6879953B1 (en) Speech recognition with request level determination
JP3588975B2 (en) Voice input device
JP2000322088A (en) Voice recognition microphone, voice recognition system, and voice recognition method
JP5015806B2 (en) Merchandise sales data processing apparatus and program thereof, and merchandise data input apparatus and program thereof
US20120253804A1 (en) Voice processor and voice processing method
JP2010039099A (en) Speech recognition and in-vehicle device
JP3726783B2 (en) Voice recognition device
JP2002278588A (en) Voice recognition device
JP3645104B2 (en) Dictionary search apparatus and recording medium storing dictionary search program
JP2000029486A (en) Speech recognition system and method
JP2002215184A (en) Voice recognition device and program
JP3700533B2 (en) Speech recognition apparatus and processing system
JP4004885B2 (en) Voice control device
JP2006208905A (en) Voice dialogue apparatus and voice dialogue method
JPH11282486A (en) Subword type unspecified speaker speech recognition apparatus and method

Legal Events

Date Code Title Description
A977 Report on retrieval

Free format text: JAPANESE INTERMEDIATE CODE: A971007

Effective date: 20040301

A131 Notification of reasons for refusal

Free format text: JAPANESE INTERMEDIATE CODE: A131

Effective date: 20040309

A521 Request for written amendment filed

Free format text: JAPANESE INTERMEDIATE CODE: A523

Effective date: 20040413

TRDD Decision of grant or rejection written
A01 Written decision to grant a patent or to grant a registration (utility model)

Free format text: JAPANESE INTERMEDIATE CODE: A01

Effective date: 20040727

A61 First payment of annual fees (during grant procedure)

Free format text: JAPANESE INTERMEDIATE CODE: A61

Effective date: 20040809

R150 Certificate of patent or registration of utility model

Free format text: JAPANESE INTERMEDIATE CODE: R150

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20080827

Year of fee payment: 4

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20080827

Year of fee payment: 4

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20090827

Year of fee payment: 5

LAPS Cancellation because of no payment of annual fees