JPH0225896A - voice recognition device - Google Patents

voice recognition device

Info

Publication number
JPH0225896A
JPH0225896A JP63176153A JP17615388A JPH0225896A JP H0225896 A JPH0225896 A JP H0225896A JP 63176153 A JP63176153 A JP 63176153A JP 17615388 A JP17615388 A JP 17615388A JP H0225896 A JPH0225896 A JP H0225896A
Authority
JP
Japan
Prior art keywords
dictionary
switch
register
speaker
feature
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
JP63176153A
Other languages
Japanese (ja)
Inventor
Junichiro Fujimoto
潤一郎 藤本
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Ricoh Co Ltd
Original Assignee
Ricoh Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Ricoh Co Ltd filed Critical Ricoh Co Ltd
Priority to JP63176153A priority Critical patent/JPH0225896A/en
Publication of JPH0225896A publication Critical patent/JPH0225896A/en
Pending legal-status Critical Current

Links

Abstract

PURPOSE:To cause plural persons to use one device by providing a display part and displaying the kind of standard patterns used for a recognition or the kind of storing parts to store it. CONSTITUTION:A dictionary for an unspecific speaker is registered to a register 5 beforehand, and a dictionary for a specific speaker to be freely registered by a user is in a register 6. The internal part of the register 6 is divided into several parts, several users use them separately, and it is sufficient for the user to designate the part of the dictionary of himself at the time of using. In order to register it, a switch 4 is made to fall at a B side, and it can be used for the unspecific speaker and for the specific speaker when a switch 7 is made to fall at a C side and a D side next, respectively. For example, when the switch 7 falls at the C side, the fact that the dictionary for the living unspecific speaker is used is displayed on a display part 8. Thus, one device can be used by plural persons.

Description

【発明の詳細な説明】 11分見 本発明は、音声認識装置に関する。[Detailed description of the invention] Watched for 11 minutes The present invention relates to a speech recognition device.

立米l権 現在、音声認識の研究は盛んであるが、音声認識には使
用者が自分の音声を登録してから使用する特定話者方式
と、何の準備もなく使用できる不特定話者方式があり、
前者の方が後者より認識率が高いという特徴がある。一
方、これらを別々の機械とせず、一つの中に両者の機能
を持たせ必要に応じてこれらを使い分けることが考えら
れる。
Research on voice recognition is currently active, but there are two types of speech recognition: the speaker-specific method, in which the user registers his or her own voice, and the unspecified-speaker method, which can be used without any preparation. There is,
The former has a higher recognition rate than the latter. On the other hand, instead of using these as separate machines, it is conceivable to have both functions in one and use them properly as needed.

この場合、共通に使う言葉は不特定話者方式で認識し、
そうでないものは特定話者方式で認識するようにしてお
くと便利である。或は、特定の指定をするためのコマン
ドを不特定にしておき、例えば、自分の名前を言うと、
その人の音声辞書がディスクの中から検索されてロード
されるような使い方が便利である。このような場合、一
つの装置を何人かで使用する訳であるが、ある人が使用
中に席を立ったような場合1次の人が使用する時には前
の人の音声辞書がロードされたままであり、認識しない
という欠点がある。
In this case, commonly used words are recognized using a speaker-independent method,
If this is not the case, it would be convenient to recognize it using the speaker-specific method. Alternatively, you can leave the command to specify a specific specification unspecified, for example, if you say your name,
It is convenient to use it in such a way that the person's speech dictionary is searched from the disk and loaded. In such cases, one device is used by several people, and if one person gets up while using it, the voice dictionary of the previous person will be loaded when the next person uses it. It has the disadvantage that it is not recognized.

且−一五 本発明は、上述のごとき実情に鑑みてなされたもので、
特に、一つの認識装置を複数の人で使う場合、どのよう
な人がどのような時にも使えるような装置を提供するこ
とを目的としてなされたものである。
-15 The present invention was made in view of the above-mentioned circumstances,
In particular, when a single recognition device is used by multiple people, the purpose of this invention is to provide a device that can be used by any person at any time.

藷ニーー成。I'm in love with you.

本発明は、上記目的を達成するために、音声を電気信号
に変換する手段と、この信号から特徴量をとり出す特微
量変換部と、その特徴量を記憶する記憶部とこれとは異
なる第2の記憶部と、特徴量の比較をする比較部と、認
識に係る情報を表示する表示部とを有する音声認識装置
において、認識に使用している標準パターンの種類又は
それを記憶している記憶部の種類を表示部に表示するよ
うにしたことを特徴としたものである。以下5本発明の
実施例に基づいて説明する。
In order to achieve the above object, the present invention provides a means for converting audio into an electrical signal, a feature converter for extracting a feature from this signal, a storage for storing the feature, and a different controller. 2, a comparison unit for comparing feature amounts, and a display unit for displaying information related to recognition; the type of standard pattern used for recognition or the type thereof is stored This device is characterized in that the type of storage section is displayed on the display section. The following will explain based on five embodiments of the present invention.

第1図は、本発明の一実施例を説明するための構成図で
1図中、1はマイクロフォン、2はマイクアンプ、3は
バンドパスフィルタ群、4はスイッチ、5,6はレジス
タ、7はスイッチ、8は表示部、9は比較部、10は結
果出力部で、音響電気変換器としてマイクロフォン1を
使用し又それからの電気信号を増幅する増幅器としてマ
イクアンプ2を使用し、その信号をバンドパスフィルタ
群3に入力して周波数分析する。この時、特徴量はパワ
ースペクトルであるがこれに限るものではなく、LPG
でも零交差数でも良い。又、レジスタ5には不特定話者
用の辞書があらかじめ登録されており、レジスタ6は使
用者が自由に登録することができる特定話者用の辞書が
入る。レジスタ6の中はいくつかに区分されて何人かの
使用者が別々に使用し、使用の際に自分の辞書の部分を
指定すれば良い、これを登録するにはスイッチ4をB側
に倒し、スイッチ7をC側に倒すと不特定話者用として
、また5D側に倒すと特定話者用として使用できる。ス
イッチ7がC側に倒れている時は表示部8のブラウン管
へ現在不特定話者用の辞書を使用していることを表示す
る。表示の仕方は、第2図に示すように現在使用してい
る辞書が不特定用のものであることを表示するものであ
っても良いし、特定話者用の辞書がいくつかに分割され
てその中のAの部分を使っている時は、第3図に示す如
く「A」の文字を反転させても良い、いずれにせよ現在
使用している辞書が明示されれば良い。又、使用時には
スイッチ4をA側に倒し、入力された音声のパターンと
レジスタ内のパターンを比較部9にて比較し類似度を求
め、その大小によって認識するが、ここで比較の仕方は
特に限定するものではない。特定話者方式と不特定話者
方式が辞書パターンの構成だけで区別で遣る方法(「数
理科学J No、287 1987年2月号P、P、6
3〜69.ファジィ音声認識)などが望ましい。
FIG. 1 is a block diagram for explaining one embodiment of the present invention. In the figure, 1 is a microphone, 2 is a microphone amplifier, 3 is a group of band pass filters, 4 is a switch, 5 and 6 are registers, and 7 is a block diagram for explaining an embodiment of the present invention. 8 is a switch, 8 is a display section, 9 is a comparison section, and 10 is a result output section. Microphone 1 is used as an acousto-electrical transducer, and microphone amplifier 2 is used as an amplifier to amplify the electric signal from the microphone. It is input to band pass filter group 3 and frequency analyzed. At this time, the feature quantity is the power spectrum, but it is not limited to this, and the LPG
However, the number of zero crossings is also fine. Further, a dictionary for unspecified speakers is registered in advance in the register 5, and a dictionary for specific speakers, which can be freely registered by the user, is stored in the register 6. The register 6 is divided into several sections, which are used by several users, and all they have to do is specify their own section of the dictionary when using it. To register this, turn the switch 4 to the B side. When the switch 7 is turned to the C side, it can be used for unspecified speakers, and when it is turned to the 5D side, it can be used for a specific speaker. When the switch 7 is turned to the C side, it is displayed on the cathode ray tube of the display section 8 that the dictionary for unspecified speakers is currently being used. The display method may be one that indicates that the dictionary currently in use is for general use, as shown in Figure 2, or one that indicates that the dictionary for specific speakers is divided into several sections. When using the A part of the dictionary, the letter "A" may be reversed as shown in FIG. 3. In any case, the dictionary currently being used should be clearly indicated. In addition, when in use, the switch 4 is turned to the A side, and the comparison section 9 compares the input voice pattern with the pattern in the register to find the degree of similarity, and recognizes it based on the magnitude of the similarity. It is not limited. A method in which the specific speaker method and the non-specific speaker method can be distinguished only by the structure of the dictionary pattern (Mathematical Science J No. 287 February 1987 issue P, P, 6
3-69. Fuzzy speech recognition) etc. are desirable.

また、第4図に示すように、スイッチ11を設け、これ
を押すとスイッチ7がC側つまり不特定話者用の辞書側
へ倒れるようにしておくことが望ましい。不特定話者用
の辞書の中に自分の専用(特定話者用)の辞書を指定す
るコマンド(例えば自分の名前)を入れておき、装置に
向って自分の名前を言うと自分用の特定話者用の辞書が
ロードされる。その状態で特定話者用の装置として使い
、使用状態のまま席を立つと、次の使用者は現在誰の辞
書がロードされているか一目瞭然であり、自分の辞書で
なければスイッチ11を押すとすぐに使うことが出来る
ようになる。
Further, as shown in FIG. 4, it is desirable to provide a switch 11 so that when the switch 11 is pressed, the switch 7 falls to the C side, that is, to the side of the dictionary for unspecified speakers. Insert a command (for example, your name) to specify your own dictionary (for specific speakers) in the dictionary for non-specific speakers, and when you say your name to the device, it will automatically display your specific dictionary. A dictionary for the speaker is loaded. In this state, if you use it as a device for a specific speaker and leave your seat while it is still in use, the next user will be able to see at a glance whose dictionary is currently loaded, and if it is not his/her dictionary, press switch 11. You will be able to use it immediately.

効−一一展 以上の説明から明らかなように、本発明によると、一つ
の装置を複数の人で使うことができ、誰の辞書がロード
されているか調べなければならないというわずらbしさ
から解放される。
As is clear from the above description, the present invention allows one device to be used by multiple people and eliminates the hassle of having to check whose dictionary is loaded. To be released.

【図面の簡単な説明】[Brief explanation of the drawing]

第1図は5本発明の−゛実施例を説明するための構成図
、第2図及び第3図は、表示部の表示例を示す図、第4
図は、本発明の他の実施例を説明するための構成図であ
る。 1・・・マイクロフォン、2・・・マイクアンプ、3・
・・バンドパスフィルタ群、4・・・スイッチ、5,6
・・・レジスタ、7・・・スイッチ、8・・・表示部、
9・・・比較部、10・・・結果出力部、11・・・ス
イッチ。 第 区 第 図 第 図 第 図
FIG. 1 is a configuration diagram for explaining an embodiment of the present invention; FIGS. 2 and 3 are diagrams showing display examples of the display unit;
The figure is a configuration diagram for explaining another embodiment of the present invention. 1...Microphone, 2...Mic amplifier, 3.
... Bandpass filter group, 4... Switch, 5, 6
...Register, 7...Switch, 8...Display section,
9... Comparison section, 10... Result output section, 11... Switch. Ward chart chart chart chart chart

Claims (1)

【特許請求の範囲】[Claims] 1、音声を電気信号に変換する手段と、この信号から特
徴量をとり出す特微量変換部と、その特徴量を記憶する
記憶部とこれとは異なる第2の記憶部と、特徴量の比較
をする比較部と、認識に係る情報を表示する表示部とを
有する音声認識装置において、認識に使用している標準
パターンの種類又はそれを記憶している記憶部の種類を
表示部に表示するようにしたことを特徴とする音声認識
装置。
1. A means for converting audio into an electrical signal, a feature converter that extracts a feature from this signal, a storage that stores the feature, a second storage that is different from this, and a comparison of the feature. In a speech recognition device that has a comparison section that performs the following operations, and a display section that displays information related to recognition, the type of standard pattern used for recognition or the type of storage section that stores it is displayed on the display section. A speech recognition device characterized by:
JP63176153A 1988-07-14 1988-07-14 voice recognition device Pending JPH0225896A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP63176153A JPH0225896A (en) 1988-07-14 1988-07-14 voice recognition device

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP63176153A JPH0225896A (en) 1988-07-14 1988-07-14 voice recognition device

Publications (1)

Publication Number Publication Date
JPH0225896A true JPH0225896A (en) 1990-01-29

Family

ID=16008588

Family Applications (1)

Application Number Title Priority Date Filing Date
JP63176153A Pending JPH0225896A (en) 1988-07-14 1988-07-14 voice recognition device

Country Status (1)

Country Link
JP (1) JPH0225896A (en)

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100423495B1 (en) * 2001-06-21 2004-03-18 삼성전자주식회사 Operation control system by speech recognition for portable device and a method using the same

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100423495B1 (en) * 2001-06-21 2004-03-18 삼성전자주식회사 Operation control system by speech recognition for portable device and a method using the same

Similar Documents

Publication Publication Date Title
US7590538B2 (en) Voice recognition system for navigating on the internet
Hollien et al. Speaker identification by long‐term spectra under normal and distorted speech conditions
CN108962260A (en) A kind of more human lives enable audio recognition method, system and storage medium
US20020002460A1 (en) System method and article of manufacture for a voice messaging expert system that organizes voice messages based on detected emotions
US4718096A (en) Speech recognition system
AU1436792A (en) Acoustic method and apparatus for identifying human sonic sources
US4408096A (en) Sound or voice responsive timepiece
JPH0225896A (en) voice recognition device
JPH02178698A (en) Voice recognition device and telephone set using same
US20020156617A1 (en) Electonic speaking dictionary: a simple device for finding, pronounincing, and defining a word
JP2740866B2 (en) Electronics
EP4544422B1 (en) Method for snore attribution
JPS638798A (en) voice recognition device
JPH0222699A (en) voice recognition device
CN117894348A (en) Recording device, recording method, system and computer readable storage medium
Seaver III et al. Using simultaneous nasometry and standard audio recordings to detect the acoustic onsets and offsets of speech
WO1994009465A1 (en) Exhibition apparatus
JPH01197796A (en) voice recognition device
JPS63136732A (en) announcement device
JPH05134697A (en) Voice recognizing system
JPS63142757A (en) announcement device
JPH0295000A (en) Electronic agreeable hearing device
JP2811196B2 (en) Voice alarm
JPH07219992A (en) Information processing method and device
JPH04200049A (en) Voice dialer