JPH0570868B2 - - Google Patents
Info
- Publication number
- JPH0570868B2 JPH0570868B2 JP60160703A JP16070385A JPH0570868B2 JP H0570868 B2 JPH0570868 B2 JP H0570868B2 JP 60160703 A JP60160703 A JP 60160703A JP 16070385 A JP16070385 A JP 16070385A JP H0570868 B2 JPH0570868 B2 JP H0570868B2
- Authority
- JP
- Japan
- Prior art keywords
- cursor
- characters
- voice
- speech
- display
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Expired - Lifetime
Links
Landscapes
- Document Processing Apparatus (AREA)
Description
【発明の詳細な説明】
(イ) 産業上の利用分野
本発明は、音声入力を用いて作成された文章等
を表示する音声入力文章作成方法に関する。DETAILED DESCRIPTION OF THE INVENTION (A) Field of Industrial Application The present invention relates to a voice input text creation method for displaying sentences created using voice input.
(ロ) 従来の技術
従来のワードプロセツサに於いては、キーボー
ド等により入力された文字を適時かな・漢字変換
処理を行つて文章を作成していたが、人間にとつ
てより自然な情報伝達媒体であるところの音声入
力によつて文章を作成する事を目的として、特開
昭56−114042号公報に開示の如き音声入力ワード
プロセツサ等が開発されている。しかし、入力さ
れた文章を表示するとき、例えば拗音の「きや」
を音声により入力したとすると、表示装置上では
〔き〕と〔や〕とに分割して連続表示され、しか
も入力位置を示すためのカーソルは1文字分の大
きさしかなくてその1文字にしか対応づけられて
いないため、前記の音声により入力された「き
や」をひとつの表音文字として扱うことが実質的
に困難であつた。すなわち、前記の「きや」の表
示〔き〕〔や〕を例えば削除しようとすると、前
記のカーソルを〔き〕の位置に移動させ、該
〔き〕の文字を削除した後、さらに〔や〕の文字
を削除しなければならないと云う煩雑な操作が必
要となる欠点があつた。(b) Conventional technology Conventional word processors create sentences by converting characters entered using a keyboard, etc. into kana and kanji at the appropriate time, but there is a method for conveying information that is more natural for humans. A voice input word processor and the like as disclosed in Japanese Patent Laid-open No. 114042/1983 have been developed for the purpose of creating sentences by voice input as a medium. However, when displaying the input sentence, for example, "Kiya"
If you enter the word by voice, it will be divided into [ki] and [ya] and displayed continuously on the display device, and the cursor used to indicate the input position is only the size of one character, so it will be displayed continuously on the display device. Therefore, it was practically difficult to treat "Kiya" inputted by voice as a single phonetic character. In other words, if you try to delete the above-mentioned display of "Kiya", for example, move the cursor to the position of "Ki", delete the character "Ki", and then press [Kiya] ] had the disadvantage of requiring a complicated operation of deleting the characters.
(ハ) 発明が解決しようとする問題点
本発明は、上記の事情に鑑みてなされたもの
で、音声の入力に対応して表示されているひと組
の文字に対する削除等の編集処理が一度の操作で
行えるようにすることを目的とする。(C) Problems to be Solved by the Invention The present invention has been made in view of the above-mentioned circumstances, and allows editing processing such as deletion of a set of characters displayed in response to voice input to be performed in one step. The purpose is to make it easy to operate.
(ニ) 問題点を解決するための手段
本発明の音声入力文章作成方法は、2文字以上
の表記文字列で表わされる音声を対象とし、入力
された音声を音声認識手段にて認識し、その認識
結果を表示手段で表示しつつ文章を作成する音声
入力文章作成方法に於いて、上記表示手段にて2
文字以上の表記文字列で表わされる音声の認識結
果を表示する場合、該音声の認識結果を特定する
為のカーソルがその表記文字列全体を表示するも
のである。即ち、カーソルは音声単位を示す文字
数に対応して、その表示範囲が可変設定されるの
である。(d) Means for solving the problem The speech input text creation method of the present invention targets speech expressed by a written character string of two or more characters, and recognizes the input speech with a speech recognition means. In the voice input sentence creation method of creating a sentence while displaying the recognition result on the display means, the display means 2
When displaying the recognition result of a voice represented by a written string of characters or more, a cursor for specifying the recognition result of the voice displays the entire written string. In other words, the display range of the cursor is variably set in accordance with the number of characters indicating the phonetic unit.
(ホ) 作用
本発明の音声入力文章作成方法によれば、表示
手段に於いて該カーソルに依つて特定された音声
単位の一連の表記文字列に対して、同時に変更、
削除、消去等の編集処理を行なうものである。(e) Effect: According to the voice input text creation method of the present invention, a series of written character strings of voice units specified by the cursor on the display means can be simultaneously changed,
It performs editing processing such as deletion and deletion.
(ヘ) 実施例
第1図は本発明の音声入力文章作成方法を単音
節音声認識方式にて実現するための一実施例であ
る。1はマイクロホンで、発声された音声を電気
信号に変換するものである。2は特徴抽出部で、
入力された音声波形を例えば高域強調して16チヤ
ンネルのバンドパスフイルタを通すことにより周
波数分析し、パワー情報、スペクトル変化等によ
りセグメンテーシヨンして単音節単位に特徴パタ
ーンを作成するものである。3は標準パターンメ
モリで各単音節毎に抽出された特徴パターンが標
準パターンとして登録されている。4は識別部で
前記特徴抽出部2において作成された特徴パター
ンと前記標準パターンメモリ3内の各単音節の標
準パターンとを比較し、その結果に応じた文字情
報が出力される。5は識別データメモリ部、6は
データポインタで、前記識別部4より出力された
文字情報を識別データメモリ部5内のデータポイ
ンタ6によつて指し示された位置に格納する。7
は操作部で、前記データポインタ6のポインタ値
の変更、及び該データポインタ6が指し示す前記
識別データメモリ部5内の文字情報の操作を行な
う。8は表示部で、識別データメモリ部5内の文
字情報に基づく文字の表示、及びデータポインタ
6が指し示している文字情報に基づく表示文字を
明示するためのカーソル表示を行なう。(F) Embodiment FIG. 1 shows an embodiment for realizing the speech input sentence creation method of the present invention using a monosyllabic speech recognition system. Reference numeral 1 denotes a microphone, which converts the uttered voice into an electrical signal. 2 is the feature extraction part,
The input speech waveform is frequency-analyzed by, for example, emphasizing the high frequencies and passing through a 16-channel bandpass filter, and segmentation is performed using power information, spectral changes, etc. to create characteristic patterns for each single syllable. . 3 is a standard pattern memory in which characteristic patterns extracted for each single syllable are registered as standard patterns. Reference numeral 4 denotes an identification section that compares the characteristic pattern created in the feature extraction section 2 with the standard pattern of each monosyllable in the standard pattern memory 3, and outputs character information according to the result. 5 is an identification data memory section, and 6 is a data pointer, which stores the character information output from the identification section 4 at the position pointed to by the data pointer 6 in the identification data memory section 5. 7
is an operation section for changing the pointer value of the data pointer 6 and operating character information in the identification data memory section 5 pointed to by the data pointer 6. A display section 8 displays characters based on the character information in the identification data memory section 5 and displays a cursor to clearly indicate the displayed character based on the character information pointed to by the data pointer 6.
第2図は前記表示部8を詳細に示したブロツク
図あり、同図によると、識別データメモリ部5か
ら送られてきた文字情報は画面表示部81のキヤ
ラクタジエネレータ及びビデオラムの動作によ
り、対応した文字がCRT82に表示される。一
方前記データポインタ6が指し示す識別データメ
モリ部5内の文字情報をもとに、カーソル制御部
83ではカーソルの表示範囲およびその表示位置
を決定し、前記画面表示部81ではこのカーソル
制御部83からの表示範囲、表示位置の情報、即
ちカーソル情報に基づきCRT82にカーソルを
表示する。 FIG. 2 is a block diagram showing the display unit 8 in detail. According to the figure, character information sent from the identification data memory unit 5 is processed by the operation of the character generator and video ram of the screen display unit 81. , the corresponding characters are displayed on the CRT82. On the other hand, based on the character information in the identification data memory section 5 pointed to by the data pointer 6, the cursor control section 83 determines the cursor display range and its display position, and the screen display section 81 A cursor is displayed on the CRT 82 based on the display range and display position information, that is, cursor information.
第3図はCRT82に表示された文字及びカー
ソルを模式的に示すもので、同図のaは、単音
節単位に認識された「と」「く」「ちよ」「う」
「は」の文字のうち、「く」の部分をカーソル9で
表示している例である。この状態から前記操作部
7を用いて前記データポインタ6の値を変更する
ことによりカーソル9を右に移動させると同図
のbの如くなる。即ち、前記カーソル制御部83
は、カーソル9を右に移動させた後に指し示すべ
き認識単位「ちよ」が「ち」と「よ」の2文字で
構成されていることを検知してカーソル9の表示
範囲が2文字分である事、並びに該カーソルの新
たな位置を示すカーソル情報が前記画面表示部8
1に送られ、CRT82上のカーソル9表示は
「ちよ」の全体を指し示すこととなる。第3図
のbは同図のaにおける状態〔第3図のbと
同じ状態〕のときに前記操作部7を用いて前記カ
ーソルが指し示す「ちよ」を削除した例を示した
もので、前記カーソル制御部83は新たに指し示
すべき認識単位「う」が1文字で構成されている
ことを検知してカーソル9の表示範囲が1文字分
である事並びに該カーソルの新たな位置を示すカ
ーソル情報が前記画面表示部81に送られる。こ
の状態で前記操作部7は識別データメモリ部5内
のデータポインタ6が指し示す「ちよ」に対応し
たデータを削除し、その情報を前記画面表示部8
1に送り、その結果がCRT82に示されるので
ある。又、上述の削除以外の変更や消去等の編集
処理についても、同様に実行できる。 Figure 3 schematically shows the characters and cursor displayed on the CRT 82. In the figure, a indicates the words "to", "ku", "chiyo", and "u" recognized in monosyllable units.
This is an example in which the ``ku'' part of the character ``ha'' is displayed with the cursor 9. From this state, if the cursor 9 is moved to the right by changing the value of the data pointer 6 using the operation section 7, the result will be as shown in b in the figure. That is, the cursor control section 83
After moving the cursor 9 to the right, it detects that the recognition unit "chiyo" to be pointed to consists of two characters "chi" and "yo", and the display range of the cursor 9 is two characters. cursor information indicating the new position of the cursor as well as the new position of the cursor is displayed on the screen display section 8.
1, and the cursor 9 displayed on the CRT 82 points to the entire "Chiyo". FIG. 3b shows an example in which "Chiyo" pointed to by the cursor is deleted using the operation section 7 in the state a of the same figure [the same state as b of FIG. 3]. The cursor control unit 83 detects that the recognition unit "U" to be newly pointed consists of one character, and informs that the display range of the cursor 9 is one character, as well as cursor information indicating the new position of the cursor. is sent to the screen display section 81. In this state, the operation section 7 deletes the data corresponding to "Chiyo" pointed to by the data pointer 6 in the identification data memory section 5, and displays the information on the screen display section 8.
1, and the results are displayed on the CRT82. Further, editing processes such as changes and deletion other than the above-mentioned deletion can be executed in the same manner.
なお、以上の説明は単音節単位の認識の場合を
示したが、単語単位の認識などにおいても本発明
の実施は可能である。また、カーソルはアンダラ
イン方式を用いて図示説明しているが、輝度反転
によるものでも同じである。 Note that although the above explanation has been given for the case of recognition in units of monosyllables, the present invention can also be implemented in recognition of units of words. Further, although the cursor is illustrated and explained using an underline method, the same applies to a method using a luminance inversion method.
(ト) 発明の効果
本発明は以上に説明したように、だく音、半だ
く音、あるいは拗音、長音等、の単音節音声の発
音を2文字以上で表示するときには、表示された
文字を文字単位ではなく単音節単位即ち発音単位
で扱うことにより、又、複数文字からなる単語認
識の場合に於いても表示文字を発音単位で扱うこ
とにより、音声により入力された文章の編集をき
わめて能率よく行うことができる。(G) Effects of the Invention As explained above, when the pronunciation of a monosyllabic sound such as a sound, a half sound, a long sound, a long sound, etc. is displayed using two or more characters, the displayed characters are By handling monosyllable units, that is, pronunciation units, rather than units, and by handling display characters in pronunciation units even in the case of word recognition consisting of multiple characters, it is possible to edit text input by voice extremely efficiently. It can be carried out.
第1図は本発明の音声入力文章作成方法の実施
例構成図、第2図は本発明方法に用いられる表示
部の実施例構成図、第3図及びは夫々本発明
方法の表示例を説明する為の模式図である。
1……マイク、4……識別部、5……識別デー
タメモリ、6……データポインタ、7……操作
部、8……表示部、9……カーソル。
Fig. 1 is a block diagram of an embodiment of the voice input text creation method of the present invention, Fig. 2 is a block diagram of an embodiment of the display unit used in the method of the present invention, and Fig. 3 and 3 illustrate display examples of the method of the present invention, respectively. FIG. 1...Microphone, 4...Identification section, 5...Identification data memory, 6...Data pointer, 7...Operation section, 8...Display section, 9...Cursor.
Claims (1)
対象とし、入力された音声を音声認識手段にて認
識し、その認識結果を表示手段で表示しつつ文章
を作成する音声入力文章作成方法に於いて、上記
表示手段にて2文字以上の表記文字列で表わされ
る音声の認識結果を表示する場合、該音声の認識
結果を特定する為のカーソルがその表記文字列全
体を表示し、該カーソルに依つて特定された音声
単位の表記文字列に対して、同時に変更、削除、
消去等の編集処理を行なう事を特徴とする音声入
力文章作成方法。1. A speech input text creation method that targets speech expressed by a written character string of two or more characters, recognizes the input speech with a speech recognition means, and creates a text while displaying the recognition result on a display means. When displaying the recognition result of a voice represented by a written string of two or more characters on the display means, the cursor for specifying the recognition result of the voice will display the entire written string, and the cursor will At the same time, the written character string of the identified phonetic unit may be changed, deleted, or
A voice input text creation method characterized by performing editing processing such as deletion.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP60160703A JPS6220065A (en) | 1985-07-19 | 1985-07-19 | Text generating method using voice input |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP60160703A JPS6220065A (en) | 1985-07-19 | 1985-07-19 | Text generating method using voice input |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| JPS6220065A JPS6220065A (en) | 1987-01-28 |
| JPH0570868B2 true JPH0570868B2 (en) | 1993-10-06 |
Family
ID=15720642
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP60160703A Granted JPS6220065A (en) | 1985-07-19 | 1985-07-19 | Text generating method using voice input |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPS6220065A (en) |
Family Cites Families (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS57196294A (en) * | 1981-05-28 | 1982-12-02 | Tokyo Shibaura Electric Co | Cursor control system for display unit |
| JPS5968792A (en) * | 1982-10-13 | 1984-04-18 | 電子計算機基本技術研究組合 | Voice recognition equipment |
| JPS60103389A (en) * | 1983-11-11 | 1985-06-07 | キヤノン株式会社 | character processing device |
-
1985
- 1985-07-19 JP JP60160703A patent/JPS6220065A/en active Granted
Also Published As
| Publication number | Publication date |
|---|---|
| JPS6220065A (en) | 1987-01-28 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| JPS62239231A (en) | Speech recognition method by inputting lip picture | |
| JPH0570868B2 (en) | ||
| JPS63157226A (en) | Conversation type sentence reading device | |
| JPS61240361A (en) | Documentation device with hand-written character | |
| JP2580565B2 (en) | Voice information dictionary creation device | |
| JPH08272388A (en) | Speech synthesizer and method thereof | |
| JPH0572620B2 (en) | ||
| JPS62212870A (en) | Sentence reading correcting device | |
| JPH04232997A (en) | System for displaying result of recognition in speech recognition device | |
| JPH0441399Y2 (en) | ||
| JP2603920B2 (en) | Voice recognition device | |
| JPS62154022A (en) | voice typewriter | |
| JPH02308194A (en) | Foreign language learning device | |
| JPH0628396A (en) | Electronic dictionary | |
| JPH0195323A (en) | Voice input device | |
| JPH04235599A (en) | Display process by group of recognition candidate information | |
| JPS61223971A (en) | Sentence generating device | |
| JPH0554115B2 (en) | ||
| JPS63229584A (en) | Character recognition device | |
| JPH06130992A (en) | Voice recognizer | |
| JPH0664571B2 (en) | Character processing method | |
| JPH01113858A (en) | Japanese word processor | |
| JPS62106573A (en) | Electronic dictionary device | |
| JPH0322169A (en) | Chinese input system | |
| JPH01233550A (en) | Display system for chinese language |