JPS61109100A - Voice recognition system - Google Patents

Voice recognition system

Info

Publication number
JPS61109100A
JPS61109100A JP59228961A JP22896184A JPS61109100A JP S61109100 A JPS61109100 A JP S61109100A JP 59228961 A JP59228961 A JP 59228961A JP 22896184 A JP22896184 A JP 22896184A JP S61109100 A JPS61109100 A JP S61109100A
Authority
JP
Japan
Prior art keywords
gain
recognition
gain value
voice
pattern
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
JP59228961A
Other languages
Japanese (ja)
Inventor
松下 満次
正次 小林
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Oki Electric Industry Co Ltd
Original Assignee
Oki Electric Industry Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Oki Electric Industry Co Ltd filed Critical Oki Electric Industry Co Ltd
Priority to JP59228961A priority Critical patent/JPS61109100A/en
Publication of JPS61109100A publication Critical patent/JPS61109100A/en
Pending legal-status Critical Current

Links

Abstract

(57)【要約】本公報は電子出願前の出願データであるた
め要約のデータは記録されません。
(57) [Summary] This bulletin contains application data before electronic filing, so abstract data is not recorded.

Description

【発明の詳細な説明】 (産業上の利用分野) 本発明は音声認識方式に関し、特に音声認識装置におけ
る入力音声の感度調整方式に関するものである。
DETAILED DESCRIPTION OF THE INVENTION (Field of Industrial Application) The present invention relates to a speech recognition method, and particularly to a method for adjusting the sensitivity of input speech in a speech recognition device.

(従来の技術) 従来の音声認識装置における入力音声の感度調整方式を
大きく2つに分けると固定方式とuf変変成式分けられ
る。固定方式は人力音声を増幅する増幅器の利得を固定
するものである。可変方式には感度を自由に変えられる
ように利得可変機構を付けたものや、AGC(自動利得
制御)機構を利用したものなどがある。
(Prior Art) Sensitivity adjustment methods for input speech in conventional speech recognition devices can be roughly divided into two types: a fixed method and a UF transformation method. The fixed method fixes the gain of the amplifier that amplifies human voice. Variable methods include those equipped with a variable gain mechanism so that the sensitivity can be changed freely, and those using an AGC (automatic gain control) mechanism.

(発明が解決しようとする問題点) しかしながら、前記従来技術の方式では次のような問題
点がある。
(Problems to be Solved by the Invention) However, the prior art system has the following problems.

固定方式の場合、オペレータに依って音量に差が有るた
め最も良い感°度に増幅器の利得を固定することは難し
いという欠点がある。
In the case of a fixed method, there is a drawback that it is difficult to fix the gain of the amplifier to the best sensitivity because the sound volume varies depending on the operator.

可変方式の場合、最良感度に調整した時点の利得を記憶
しておくことができないという欠点がある。従って、標
準パターン作成時と音声認識時において、利得に差が生
じる場合が有り、認識率の低下を招いている。又、オペ
レータが替るたびに感度調整を再度行なわなければなら
ないという欠点があった。
In the case of the variable method, there is a drawback that the gain at the time when the sensitivity is adjusted to the best sensitivity cannot be stored. Therefore, there may be a difference in gain between standard pattern creation and speech recognition, resulting in a reduction in recognition rate. Another disadvantage is that the sensitivity must be adjusted again every time the operator changes.

本発明は前記従来技術の問題点を解決し、音声信号増幅
器の利得の自動設定及び自動記憶する機能を有する音声
認識方式を提供するものである。
The present invention solves the problems of the prior art and provides a speech recognition system having a function of automatically setting and automatically storing the gain of an audio signal amplifier.

C問題点を解決するための手段) 本発明は前記問題点を解決するために、入力音声信号を
音声増幅器で増幅し、増幅された音声信号を音声分析手
段で分析して音声パターンを作成し、該音声パターンと
同様にして予め作成された標準パターンとを照合して音
声認識を行なう音声認識方式において、標準パターンと
該標準パターン作成時の前記音声増幅器の利得値とを記
憶する記憶手段と、前記音声増幅器の利得値を設定する
と共に該利得値を保持する利得制御手段と、パターン照
合して音声認識を行なうと共に前記記憶手段及び利得制
御手段を制御する認識制御手段とから構成され、標準パ
ターン作成時には前記認識ル制御手段は前記利得制御手
段で保持された利得値と標準パターンとを記憶手段に出
力し、パターン照合時には前記認識制御手段は前記記憶
手段に記憶された標準パターンと利得値とを読み出すと
共に該利得値を前記利得制御部を介して前記音声増幅器
に設定する音声認識方式である。
Means for Solving Problem C) In order to solve the above problem, the present invention amplifies an input audio signal with an audio amplifier, and analyzes the amplified audio signal with an audio analysis means to create a audio pattern. , in a speech recognition method in which speech recognition is performed by comparing the speech pattern with a standard pattern created in advance in the same manner, a storage means for storing the standard pattern and a gain value of the audio amplifier at the time of creating the standard pattern; , a gain control means for setting a gain value of the audio amplifier and holding the gain value, and a recognition control means for performing speech recognition by pattern matching and controlling the storage means and the gain control means, When creating a pattern, the recognition control means outputs the gain value and the standard pattern held by the gain control means to the storage means, and when matching the pattern, the recognition control means outputs the standard pattern and gain value stored in the storage means. This is a voice recognition method in which the gain value is read out and set in the voice amplifier via the gain control section.

(作用) 本発明によれば以上のように音声認識方式を構成したの
で技術手段は次のように作用する。標準パターン作成時
には、利得制御手段は入力された適切な利得値を音声増
幅器に設定すると共に、その利得値を保持するように働
き、認識制御手段は音声増幅器で増幅され音声分析手段
で分析され特徴パラメータ化された標準パターンと利得
制御手段に保持された利得値とを記憶手段に出力するよ
うに働く、パターン照合時には、認識制御手段は記憶手
段に記憶された標準パターンとfir得値とを読み出す
と共に、その利得値を利得制御手段を介して音声増幅器
に設定するように働き、認識制御手段は音声増幅器及び
音声分析手段を介し、特徴パラメータ化された音声パタ
ーンと標準パターンとを照合して音声認識するように働
く、従って、前記問題点が解決できるのである。
(Operation) According to the present invention, since the voice recognition system is configured as described above, the technical means operates as follows. When creating a standard pattern, the gain control means sets an appropriate input gain value to the audio amplifier and works to hold that gain value, and the recognition control means sets the input appropriate gain value to the audio amplifier and works to hold the gain value, and the recognition control means sets the input appropriate gain value to the audio amplifier and works to maintain the gain value. The recognition control means operates to output the parameterized standard pattern and the gain value held in the gain control means to the storage means, and during pattern matching, the recognition control means reads out the standard pattern and the fir value stored in the storage means. At the same time, the gain value is set in the audio amplifier through the gain control means, and the recognition control means collates the characteristic parameterized audio pattern with the standard pattern through the audio amplifier and audio analysis means to determine the audio output. Therefore, the above-mentioned problems can be solved.

(実施例) 第1図は本発明による音声認識方式の第1の実施例を示
す図であって、特定話者音声認識装置の構成を示すもの
である。マイクロフォン1は入力音声を電気信号に変換
する。音声信号増幅器2はマイクロフォンlから出力さ
れる音声信号を増幅する。フィルタ3は音声信号増幅器
2から出力される音声信号を高域強調する。フィルタ・
バンク4は高域強調された信号からスペクトル分析を行
ないスペクトルパターンを作成する。認識制御部5は標
準パターン作成時にはフィルタ・バンク4で得られたス
ペクトルパターンを特徴パラメータ化して標準パターン
を作成し、後述の利得制御部9に保持されている音声信
号増幅器2の利得値と共に後述の記憶媒体6に格納する
。パターン照合時には格納された標準パターンと利得値
とを読み出し、利得値を利得制御部9を介して音声信号
増幅器2に設定した後、フィルタ・バンク4で得られた
スペクトルパターンを特徴パラメータ化した被照合の音
声パターンと、読み出した標準パターンとをパターン照
合して音声認識を行なう、その他、認識制御部5は装置
全体を制御する機能を有する。記憶媒体6は認識制御部
5から転送された標準パターンと該標準パターン作成時
に使用した音声信号増幅器2の利得値とを記憶する。認
識結果表示器7は認識間gi5から転送された認識結果
を表示する。利得設定キー8はオペレータが音声信号増
幅器2の利得値を入力することによりその利得値を利得
制御部9へ出力する。利得制御部9は後述するように標
準パターン作成時には利得設定キー8で設定された利得
値を音声信号増幅器2へ出力すると共に保持゛する。ま
た、利得制御部9は認識制御部5から転送され、標準パ
ターン作成時に使用した音声信号増幅器2の利得値を音
声信号増幅器2へ出力する。利得表示器10は利得制御
部9から入力され、音声信号増幅器2に設定されている
利得値を表示する。音声出力レベル表示器11は音声信
号増幅器2の出力レベルを表示する。
(Embodiment) FIG. 1 is a diagram showing a first embodiment of the speech recognition method according to the present invention, and shows the configuration of a specific speaker speech recognition device. Microphone 1 converts input audio into an electrical signal. The audio signal amplifier 2 amplifies the audio signal output from the microphone l. The filter 3 emphasizes the high frequencies of the audio signal output from the audio signal amplifier 2. filter·
Bank 4 performs spectrum analysis on the high frequency emphasized signal to create a spectrum pattern. When creating a standard pattern, the recognition control unit 5 creates a standard pattern by converting the spectral pattern obtained by the filter bank 4 into feature parameters, and uses the gain value of the audio signal amplifier 2 held in the gain control unit 9, which will be described later, as well as the gain value of the audio signal amplifier 2. The data is stored in the storage medium 6 of. During pattern matching, the stored standard pattern and gain value are read out, the gain value is set in the audio signal amplifier 2 via the gain control section 9, and then the spectral pattern obtained by the filter bank 4 is converted into a characteristic parameter. The recognition control unit 5 has the function of performing voice recognition by pattern matching the voice pattern for comparison with the read standard pattern, and controlling the entire apparatus. The storage medium 6 stores the standard pattern transferred from the recognition control unit 5 and the gain value of the audio signal amplifier 2 used when creating the standard pattern. The recognition result display 7 displays the recognition result transferred from the recognition inter-gi5. The gain setting key 8 outputs the gain value to the gain control section 9 when the operator inputs the gain value of the audio signal amplifier 2 . As will be described later, the gain control section 9 outputs the gain value set by the gain setting key 8 to the audio signal amplifier 2 and holds it therein when creating the standard pattern. Further, the gain control unit 9 outputs to the audio signal amplifier 2 the gain value of the audio signal amplifier 2 which is transferred from the recognition control unit 5 and used when creating the standard pattern. The gain display 10 displays the gain value input from the gain control section 9 and set in the audio signal amplifier 2. The audio output level display 11 displays the output level of the audio signal amplifier 2.

音声信号増幅器2の内部構成を第2図に示す。The internal configuration of the audio signal amplifier 2 is shown in FIG.

音声信号増幅器2はオペレーショナルアンプル12.ア
ナログ・マルチプレクサ13.抵抗14(Ro)及び抵
抗列15 (R1〜R8)で構成されるプログラマブル
・ゲインアンプ(公知)である。オペレーショナルアン
プ12の非反転端子はアースされ、反転端子は抵抗14
を介して入力端子2aに接続されると共に抵抗列15及
びアナログマルチプレクサ13を介してオペレーショナ
ルアンプ12の出力端子に接続される。その出力端子は
出力端子2bに接続される0以上の構成により音声信号
増幅器2は反転増幅器となっている。
The audio signal amplifier 2 includes an operational amplifier 12. Analog multiplexer 13. This is a programmable gain amplifier (known) consisting of a resistor 14 (Ro) and a resistor array 15 (R1 to R8). The non-inverting terminal of the operational amplifier 12 is grounded, and the inverting terminal is connected to the resistor 14.
It is connected to the input terminal 2a via the resistor array 15 and the analog multiplexer 13 to the output terminal of the operational amplifier 12. The audio signal amplifier 2 is an inverting amplifier due to the configuration of zero or more output terminals connected to the output terminal 2b.

利得設定キー8の操作又は後述の第2.3の実施例では
認識制御部5(CPU)の信号により利得値が利得制御
部9を介してアナログ自マルチプレクサ13の切換入力
端子13a、 b、 cに入力され、利得値に対応した
抵抗Ri(1=1〜8)が選択される。入力端子2aに
入力された音声信号は抵抗RQと抵抗R; により決定
される増幅率で増幅され、出力端子2bからフィルタ3
へ出力される。また、その大きさは音声出力レベル表示
器11に表示される。
The gain value is changed to the switching input terminals 13a, b, c of the analog self-multiplexer 13 via the gain control unit 9 by operation of the gain setting key 8 or by a signal from the recognition control unit 5 (CPU) in the 2.3 embodiment described later. is input, and a resistor Ri (1=1 to 8) corresponding to the gain value is selected. The audio signal input to the input terminal 2a is amplified by the amplification factor determined by the resistors RQ and R; and is passed from the output terminal 2b to the filter 3.
Output to. Further, its size is displayed on the audio output level display 11.

利得制御部9の内部構成を第3図に示す、利得制御部9
はキーコントローラ21、レジスタ22 、23、デー
タセレクタ24、表示コントローラ25及びドライバー
28から構成される。キーコントローラ21は利得設定
キー8から入力された利得値を示す利得信号Aを入力制
御してレジスタ22に送る。レジスタ22はキーコント
ローラ21からの利得信号Aをデータセレクタ24出力
すると共に保持する。レジスタ23は認識制御部5から
書き込み信号が入力されると、認識制御部5で記憶され
ている利得値を示す利得信号Bをデータセレクタ24へ
出力すると共に保持する。データセレクタ24はキー人
力可能信号により利得信号A、Bのうちどちらか一方を
選択して音声信号増幅器2、表示コントローラ25及び
ドライバー26に出力する。表示コントローラ25はデ
ータセレクタ24から入力される利得信号の示す利得値
を利得表示器12に表示させる。ドライバー26は、認
識制御部5から読み出し信号が入力されるとデータセレ
クタ24から入力される利得信号を認識制御部5へ出力
する。従って、認識制御部5はキー人力可能信号を利得
制御部9のデータセレクタ24に送ることにより利得設
定キー8からの利得値(利得信号A)と認識制御部5か
らの利得値(利得信号B)とを切替えることができる。
The internal configuration of the gain control unit 9 is shown in FIG.
is composed of a key controller 21, registers 22 and 23, a data selector 24, a display controller 25, and a driver 28. The key controller 21 input-controls the gain signal A indicating the gain value input from the gain setting key 8 and sends it to the register 22 . The register 22 outputs the gain signal A from the key controller 21 to the data selector 24 and holds it. When the register 23 receives a write signal from the recognition control section 5, it outputs a gain signal B indicating the gain value stored in the recognition control section 5 to the data selector 24 and holds it. The data selector 24 selects one of the gain signals A and B according to the key input enable signal and outputs it to the audio signal amplifier 2, display controller 25, and driver 26. The display controller 25 causes the gain display 12 to display the gain value indicated by the gain signal input from the data selector 24. The driver 26 outputs the gain signal input from the data selector 24 to the recognition control unit 5 when the read signal is input from the recognition control unit 5 . Therefore, the recognition control section 5 sends the key manual enable signal to the data selector 24 of the gain control section 9, thereby inputting the gain value from the gain setting key 8 (gain signal A) and the gain value from the recognition control section 5 (gain signal B). ) can be switched.

また、認識制御部5は読み出し信号を利得制御部9のド
ライバー26へ送ることにより利得設定キー8で設定さ
れた利得値(利得信号A)を読み出すことができる。
Further, the recognition control section 5 can read out the gain value (gain signal A) set by the gain setting key 8 by sending a readout signal to the driver 26 of the gain control section 9.

次に、第1の実施例の動作を説明する。Next, the operation of the first embodiment will be explained.

標準パターンを新たに作成する場合は次のように動作す
る。マイクロフォン1に音声を入れながら、利得設定キ
ー8を操作して、音声出力レベル表示器11の表示が認
識に適した値となるように音声信号増幅器2の利得値を
設定する。このときの利得値は利得表示器10に表示さ
れると共に利得制御部9に保持される。音声信号増幅器
2で増幅された音声信号はフィルタ3及びフィルタ書バ
ンク4を通してスペクトル分析される。スペトル分析さ
れたスペクトルパターンを特徴パラメータ化し標準パタ
ーンとして認識制御部5は内部のメモリ1B上に作成す
る。次に認識制御部5は利得制御部9の保持している利
得値をメモリ16内の利得エリア17に転送する。この
ときのメモ1月6の使用状態を第5図に示す。次にメモ
リ16の内容を記憶媒体6(フロッピーディスク等)に
転送し、表示パターンの記録を完了する。
When creating a new standard pattern, it operates as follows. While inputting sound into the microphone 1, the gain setting key 8 is operated to set the gain value of the audio signal amplifier 2 so that the display on the audio output level display 11 becomes a value suitable for recognition. The gain value at this time is displayed on the gain display 10 and held in the gain control section 9. The audio signal amplified by the audio signal amplifier 2 is passed through a filter 3 and a filter bank 4 for spectrum analysis. The recognition control unit 5 converts the spectrum pattern subjected to the spectrum analysis into characteristic parameters and creates them as a standard pattern on the internal memory 1B. Next, the recognition control section 5 transfers the gain value held by the gain control section 9 to the gain area 17 in the memory 16. Figure 5 shows how the memo January 6 was used at this time. Next, the contents of the memory 16 are transferred to the storage medium 6 (floppy disk, etc.) to complete recording of the display pattern.

音声認識を実行する場合は次のように動作する。音声認
識に先だち、認識制御部5は記憶媒体6に記憶されてい
るオペークの標準パターン及び利得値をメモリ16上に
転送する。転送が完了したならば利得制御部9は認識制
御部5の制御によリフモリ16上の利得エリアから利得
値を読み出し、その利得値を音声信号増幅器2に設定す
ると共に、利得表示器10に表示する。この利得値が音
声信号増幅器2に設定されたならば音声認識動作を行な
う。即ち、音声をマイクロフォン1に入力すると、音声
信号は標準パターン作成時の利得値で設定された音声信
号増幅器2で増幅されフィルタ3及びフィルタ・バンク
4を通してスペクトルパターンとなる。認識制御部5は
このスペクトルパターンを特徴パラメータ化して被照合
の音声パターンとする。この照合される音声パターンと
、メモリ16上に読み出された標準パターンとをパター
ン照合して音声認識を行なう、この結果を認識制御部5
は認識結果表示器7に表示する。
When performing speech recognition, it operates as follows. Prior to speech recognition, the recognition control unit 5 transfers the opaque standard pattern and gain value stored in the storage medium 6 onto the memory 16. When the transfer is completed, the gain control unit 9 reads the gain value from the gain area on the ref memory 16 under the control of the recognition control unit 5, sets the gain value in the audio signal amplifier 2, and displays it on the gain display 10. do. Once this gain value is set in the audio signal amplifier 2, a speech recognition operation is performed. That is, when a voice is input to the microphone 1, the voice signal is amplified by the voice signal amplifier 2 set at the gain value at the time of creating the standard pattern, and then passed through the filter 3 and filter bank 4 to become a spectral pattern. The recognition control unit 5 converts this spectrum pattern into feature parameters and uses it as a speech pattern to be matched. The speech pattern to be matched and the standard pattern read out on the memory 16 are pattern matched to perform speech recognition, and this result is sent to the recognition control section 5.
is displayed on the recognition result display 7.

第4図は本発明による音声認識方式の第2の実施例を示
す図であり、音声入力日本語処理装置の構成を示すもの
である。第1図と同一の参照符号は同一性のある構成部
分を示す。音声入力装置18は第1の実施例における利
得設定キー8.利得表示器10及び認識結果表示器7を
除いたものである。日本語処理装置19は利得値等を人
力するためのキーボード19a及び利得値及び認識結果
等を表示する表示装置19bを有する。音声入力装置1
8は日本語処理装置19に接続されている。
FIG. 4 is a diagram showing a second embodiment of the speech recognition method according to the present invention, and shows the configuration of a speech input Japanese language processing device. The same reference numerals as in FIG. 1 indicate identical components. The voice input device 18 is the gain setting key 8 in the first embodiment. This figure excludes the gain display 10 and the recognition result display 7. The Japanese language processing device 19 has a keyboard 19a for inputting gain values manually, and a display device 19b for displaying gain values, recognition results, etc. Voice input device 1
8 is connected to a Japanese language processing device 19.

標準パターンを新たに作成する場合は、マイクロフォン
lに音声を入れながら、日本語処理装置18のキーボー
ド19aを操作して音声出力レベル表示器11の表示が
認識に適した値となる様にする。
When creating a new standard pattern, inputting voice into the microphone l, operate the keyboard 19a of the Japanese language processing device 18 so that the display on the voice output level display 11 becomes a value suitable for recognition.

この時、キーボード19aより人力された利得値は1日
本語処理装置13の表示装置19bの表示画面に表示さ
れるとともに、音声入力装置18の認識制御部5内のメ
モリ16の利得エリア17に転送される。認識制御部5
は続いて利得制御部9に利得値を設定する。適当な利得
値が設定できたならば、第1の実施例と同様にして標準
パターンを認識制御部5内のメモリ16上に作成する。
At this time, the gain value entered manually from the keyboard 19a is displayed on the display screen of the display device 19b of the Japanese language processing device 13, and is also transferred to the gain area 17 of the memory 16 in the recognition control unit 5 of the voice input device 18. be done. Recognition control unit 5
Then sets a gain value in the gain control section 9. Once an appropriate gain value has been set, a standard pattern is created on the memory 16 in the recognition control section 5 in the same manner as in the first embodiment.

標準パターンを作成した後、メモリ16の内容を記憶媒
体6(フロッピーディスク等)に転送し、標準パターン
の記録を完了する。
After creating the standard pattern, the contents of the memory 16 are transferred to the storage medium 6 (floppy disk, etc.) to complete recording of the standard pattern.

認識動作を行う場合は、第1の実施例と同様である。The recognition operation is the same as in the first embodiment.

次に、本発明による音声認識方式の第3の実施例を第4
図を用いて説明する。第3の実施例は、第2の実施例の
改良型である。音声入力装置18の上位である日本語処
理装置18は先ず、標準パターン作成モードとなる。日
本語処理装置18は音声入力装置1B内の認識制御部5
に対して、標準パターン作成モードであることを指令し
、認識制御部5は利得制御部9に対して、利得の初期値
を設定する。次にオペレータは日本語処理装置18の表
示装置19bの表示画面のガイダンスにより数個の指示
された音節に対して発声する。認識制御部5は発生され
た音声の音声パワーを分析し、最も適した利得値を算出
する。認識制御部5は最適利得値を利得制御部9、及び
メモリ16内の利得値エリア17に転送する。
Next, the third embodiment of the speech recognition method according to the present invention will be explained in the fourth embodiment.
This will be explained using figures. The third embodiment is an improved version of the second embodiment. The Japanese language processing device 18, which is a higher-level device than the voice input device 18, first enters the standard pattern creation mode. The Japanese language processing device 18 is the recognition control section 5 in the voice input device 1B.
, the recognition control section 5 instructs the gain control section 9 to be in the standard pattern creation mode, and sets the initial value of the gain to the gain control section 9 . Next, the operator utters several designated syllables according to the guidance on the display screen of the display device 19b of the Japanese language processing device 18. The recognition control unit 5 analyzes the voice power of the generated voice and calculates the most suitable gain value. The recognition control unit 5 transfers the optimum gain value to the gain control unit 9 and the gain value area 17 in the memory 16.

最適利得値の設定が完了したならば、日本語処理装置1
8の表示装置19bの表示画面のガイダンスに従い、標
準パターンの作成を行なう、標準パターンの作成が完了
したならば、メモリ16内の標準パターンと利得値を記
録媒体6に記憶する。
Once the setting of the optimal gain value is completed, the Japanese language processing device 1
Following the guidance on the display screen of the display device 19b of 8, a standard pattern is created. Once the creation of the standard pattern is completed, the standard pattern and gain value in the memory 16 are stored in the recording medium 6.

認識動作については、第1の実施例と同様である。The recognition operation is the same as in the first embodiment.

第?、3の実施例においては、記録媒体6は認識制御部
5に接続された形となっているが、日本語処理装置18
に接続した記憶媒体でも行うことが出来る。
No.? , 3, the recording medium 6 is connected to the recognition control section 5, but the Japanese language processing device 18
This can also be done with a storage medium connected to.

本実施例は標準パターン作成時の利得値を標準パターン
と同一の記憶媒体に自動記憶する機能と、認識に先だち
標準パターンを記憶媒体からメモリ上に転送する場合に
利得値をも同時に転送し、入力音声の感度調整を自動的
に完了する機能を有した音声認識方式であるため、オペ
レータが複数の場合、オペレータが替わっても使用する
標準パターンを選択するだけで良く、再度感度調整をす
る必要が無いという利点がある。又、認蟲動作時におい
ては、標準パターン作成時の利得値が自動的に音声信号
増幅器に設定されるため、音声信号の再現性が向上し、
誤認識を低減することができる。
This embodiment has a function that automatically stores the gain value when creating the standard pattern in the same storage medium as the standard pattern, and also transfers the gain value at the same time when the standard pattern is transferred from the storage medium to the memory prior to recognition. Since this is a voice recognition method that has a function that automatically completes the sensitivity adjustment of the input voice, if there are multiple operators, even if the operator changes, you only need to select the standard pattern to use, and there is no need to adjust the sensitivity again. It has the advantage of not having In addition, during the detection operation, the gain value at the time of standard pattern creation is automatically set to the audio signal amplifier, improving the reproducibility of the audio signal.
Misrecognition can be reduced.

(発明の効果) 以上説明したように、本発明によれば、標準パターン作
成時の音声増幅器の利得値をパターン照合時にも使用で
きるので、音声信号の再現性が向上して精度の高い音声
認識を容易に行なうことができる。
(Effects of the Invention) As explained above, according to the present invention, the gain value of the audio amplifier during standard pattern creation can be used during pattern matching, thereby improving the reproducibility of audio signals and achieving highly accurate speech recognition. can be done easily.

【図面の簡単な説明】[Brief explanation of drawings]

第1図は本発明による音声認識方式の第1の実施例を示
す図、第2図は音声信号増幅器を詳細に示す図、第3図
は利得制御部を詳細に示す図、第4図は本発明による第
2及び第3の実施例を示す図、第5図はメモリの使用状
態を示す図である。 2・・・音声信号増幅器、  3・・・フィルタ、4・
・・フィルタ・パンク、  5・・・認識制御部、6・
・・記憶媒体、  9・・・利得制御部。
FIG. 1 is a diagram showing a first embodiment of the speech recognition system according to the present invention, FIG. 2 is a diagram showing the audio signal amplifier in detail, FIG. 3 is a diagram showing the gain control section in detail, and FIG. FIG. 5 is a diagram showing the second and third embodiments of the present invention, and FIG. 5 is a diagram showing how the memory is used. 2... Audio signal amplifier, 3... Filter, 4...
... Filter puncture, 5. Recognition control section, 6.
...Storage medium, 9...Gain control section.

Claims (1)

【特許請求の範囲】[Claims] 入力音声信号を音声増幅器で増幅し、増幅された音声信
号を音声分析手段で分析して音声パターンを作成し、該
音声パターンと同様にして予め作成された標準パターン
とを照合して音声認識を行なう音声認識方式において、
標準パターンと該標準パターン作成時の前記音声増幅器
の利得値とを記憶する記憶手段と、前記音声増幅器の利
得値を設定すると共に該利得値を保持する利得制御手段
と、パターン照合して音声認識を行なうと共に前記記憶
手段及び利得制御手段を制御する認識制御手段とを有し
、標準パターン作成時には前記認識制御手段は前記利得
制御手段で保持された利得値と標準パターンとを記憶手
段に出力し、パターン照合時には前記認識制御手段は前
記記憶手段に記憶された標準パターンと利得値とを読み
出すと共に該利得値を前記利得制御部を介して前記音声
増幅器に設定することを特徴とする音声認識方式。
The input voice signal is amplified by a voice amplifier, the amplified voice signal is analyzed by a voice analysis means to create a voice pattern, and the voice pattern is compared with a standard pattern created in advance in the same manner to perform voice recognition. In the voice recognition method used,
A storage means for storing a standard pattern and a gain value of the audio amplifier at the time of creating the standard pattern, a gain control means for setting the gain value of the audio amplifier and holding the gain value, and pattern matching to perform speech recognition. and a recognition control means for controlling the storage means and the gain control means, and when creating a standard pattern, the recognition control means outputs the gain value held by the gain control means and the standard pattern to the storage means. , a voice recognition system characterized in that during pattern matching, the recognition control means reads out a standard pattern and a gain value stored in the storage means, and sets the gain value in the voice amplifier via the gain control section. .
JP59228961A 1984-11-01 1984-11-01 Voice recognition system Pending JPS61109100A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP59228961A JPS61109100A (en) 1984-11-01 1984-11-01 Voice recognition system

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP59228961A JPS61109100A (en) 1984-11-01 1984-11-01 Voice recognition system

Publications (1)

Publication Number Publication Date
JPS61109100A true JPS61109100A (en) 1986-05-27

Family

ID=16884572

Family Applications (1)

Application Number Title Priority Date Filing Date
JP59228961A Pending JPS61109100A (en) 1984-11-01 1984-11-01 Voice recognition system

Country Status (1)

Country Link
JP (1) JPS61109100A (en)

Similar Documents

Publication Publication Date Title
JP3263392B2 (en) Text processing unit
EP0077194B1 (en) Speech recognition system
US6411928B2 (en) Apparatus and method for recognizing voice with reduced sensitivity to ambient noise
JPS6239746B2 (en)
US5392381A (en) Acoustic analysis device and a frequency conversion device used therefor
US6718217B1 (en) Digital audio tone evaluating system
JPS6238340Y2 (en)
JP3365113B2 (en) Audio level control device
KR0134452B1 (en) Apparatus for marking in a song accompany system
JPH10254493A (en) Method for normalizing volume of recorded voice and apparatus for implementing the method
JP2975808B2 (en) Voice recognition device
JP3864498B2 (en) Analysis equipment
JPS6250850B2 (en)
JP2608702B2 (en) Speech section detection method in speech recognition
JPS6118200B2 (en)
JPS595919B2 (en) Method for detecting abnormality while voice response device is online
Hollien Basic Equipment
Oppenheim Deconvolutlon of Speech
CA1120586A (en) Recording system
JPS637398B2 (en)
Seo New approach to speech compression
JP2712708B2 (en) Voice detection device
JPH01223498A (en) Audio signal writing device
JPS6169296A (en) Voice input circuit
JPS6172299A (en) Voice recognition equipment