JPH06314273A - Document preparing device and candidate output order control method - Google Patents

Document preparing device and candidate output order control method

Info

Publication number
JPH06314273A
JPH06314273A JP5102031A JP10203193A JPH06314273A JP H06314273 A JPH06314273 A JP H06314273A JP 5102031 A JP5102031 A JP 5102031A JP 10203193 A JP10203193 A JP 10203193A JP H06314273 A JPH06314273 A JP H06314273A
Authority
JP
Japan
Prior art keywords
kanji
conversion
input
kana
candidate
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
JP5102031A
Other languages
Japanese (ja)
Inventor
Yukihiro Fukunaga
幸弘 福永
Takeshi Inoue
健 井上
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Toshiba Corp
Toshiba AVE Co Ltd
Original Assignee
Toshiba Corp
Toshiba AVE Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Toshiba Corp, Toshiba AVE Co Ltd filed Critical Toshiba Corp
Priority to JP5102031A priority Critical patent/JPH06314273A/en
Publication of JPH06314273A publication Critical patent/JPH06314273A/en
Pending legal-status Critical Current

Links

Landscapes

  • Document Processing Apparatus (AREA)

Abstract

(57)【要約】 【目的】 本発明は漢字かな混じりかな漢字変換率を向
上させることを目的としている。 【構成】 本発明において、漢字混じりかな漢字変換部
5は変換用バッファ7に入力された文字列を漢字混じり
かな漢字変換し、得られた変換候補を優先順位を付けて
変換候補バッファ11に格納する。入力漢字調査部6は
変換用バッファ7の文字列で漢字表記入力履歴のある漢
字を調査する。漢字表記学習部12は変換用バッファ7
内の入力文字列と、オペレータにより選択された変換候
補文字列を比較して、漢字表記で入力された漢字を漢字
ビットマップ14に記憶する。候補漢字調査部10は漢
字ビットマップ14に記憶済みの漢字を含む変換候補を
候補バッファ11内に見つけて、変換候補制御部9に知
らせる。変換候補制御部9は前記漢字を含む変換候補の
表記がひらがなであった場合、この変換候補の出力優先
順位を下げる制御を行う。
(57) [Summary] [Object] The present invention aims to improve the conversion rate of kanji and kana mixed kanji. According to the present invention, a kanji / kanji / kanji conversion unit 5 converts kanji / kanji / kanji into a character string input to a conversion buffer 7, and stores the obtained conversion candidates in a conversion candidate buffer 11 with priorities. The input kanji character examining unit 6 examines the kanji in the kanji notation input history using the character string in the conversion buffer 7. The kanji notation learning unit 12 uses the conversion buffer 7
The input character string in the above is compared with the conversion candidate character string selected by the operator, and the kanji input in kanji notation is stored in the kanji bitmap 14. The candidate kanji character checking unit 10 finds a conversion candidate including the Kanji stored in the Kanji bit map 14 in the candidate buffer 11 and informs the conversion candidate control unit 9. When the conversion candidate including the Chinese character is written in Hiragana, the conversion candidate control unit 9 controls the output priority of the conversion candidate to be lowered.

Description

【発明の詳細な説明】Detailed Description of the Invention

【0001】[0001]

【産業上の利用分野】本発明は漢字混じりのかな文字列
をスタイラスペンで入力し、文字認識処理した後、これ
を入力文字情報として、漢字混じりかな漢字変換を施
し、漢字混じりの文書を作成する文書作成装置に関わ
り、特に変換候補の出力順位の調整処理に関する。
BACKGROUND OF THE INVENTION The present invention inputs a kana-mixed kana character string with a stylus pen, performs character recognition processing, and performs kana-kanji kana-kanji conversion using this as input character information to create a kanji-mixed document. The present invention relates to a document creation device, and more particularly to adjustment processing of output rank of conversion candidates.

【0002】[0002]

【従来の技術】従来この種の文書作成装置では、入力さ
れたかな文字列を漢字かな混じりかな漢字変換して得ら
れる複数の変換候補の出力優先順位が変換時の文節間接
続点数や単語使用頻度、学習情報などにより決定されて
いた。ところで、手書きワープロなどのオンライン手書
き文字認識手段を持ち漢字混じりかな入力が可能な文書
作成装置では、一般的に簡単な漢字はかな表記よりも漢
字表記のまま入力されることが多い。従って、逆の場
合、即ち、ひらがなで入力された変換結果に通常漢字表
記で入力する漢字が候補として含まれる場合、前記候補
はオペレータが意図した候補とは異なることが多く、正
しい変換候補を選択するために候補選択処理を行う必要
があった。例えば、「きについて」とひらがなで入力さ
れ、その変換候補として「木について」という変換候補
があり、通常、「木」は漢字で入力される場合、オペレ
ータは候補として「木」を意図していないことが多い傾
向にある。このような場合、オペレータは次候補として
例えば「気」を選択する操作を行わなければならなくな
る。
2. Description of the Related Art Conventionally, in this type of document creating apparatus, the output priority of a plurality of conversion candidates obtained by converting an input kana character string into kana / kana mixed kana / kanji character is the number of connection points between phrases and word usage frequency at the time of conversion. It was decided by learning information. By the way, in a document creating apparatus such as a handwriting word processor having an online handwriting character recognition means and capable of inputting kana mixed with kanji, in general, simple kanji is often input in kanji notation rather than kana notation. Therefore, in the opposite case, that is, when the conversion result input in hiragana includes a kanji that is normally input in kanji as a candidate, the candidate is often different from the operator's intended candidate, and the correct conversion candidate is selected. In order to do so, it was necessary to perform candidate selection processing. For example, if you enter hiragana as “Ki About” and there is a conversion candidate “About Tree” as the conversion candidate, and if “Tree” is entered in Kanji, the operator intends “Tree” as a candidate. It tends to be absent. In such a case, the operator has to perform an operation of selecting, for example, "ki" as the next candidate.

【0003】[0003]

【発明が解決しようとする課題】スタイラスペン等で入
力されたかな文字列をオンライン手書き文字認識し、こ
れを漢字かな混じりかな漢字変換する従来の文書作成装
置では、オペレータによりひらがなで入力された変換結
果に通常漢字表記で入力する漢字が候補として含まれる
場合、前記候補はオペレータが意図した候補とは異なる
ことが多く、正しい変換候補を選択するために候補選択
処理を行う必要が生じ、この分、変換率が悪くなるとい
う欠点があった。
SUMMARY OF THE INVENTION In a conventional document creating apparatus for recognizing a kana character string input by a stylus pen or the like on-line handwritten characters and converting the kana character into kana mixed kana kanji, the conversion result input in hiragana by the operator. When the kanji input in the normal kanji notation is included as a candidate in, the candidate is often different from the candidate intended by the operator, and it is necessary to perform candidate selection processing to select the correct conversion candidate. There was a drawback that the conversion rate became worse.

【0004】そこで本発明は上記の欠点を除去し、オペ
レータの意図どおりの漢字かな混じりかな漢字変換が行
われるようにして、その変換率を向上させることができ
る文書作成装置を提供することを目的としている。
SUMMARY OF THE INVENTION Therefore, an object of the present invention is to eliminate the above-mentioned drawbacks and to provide a document creating apparatus capable of performing kanji and kana / kanji mixed kanji conversion as intended by an operator and improving the conversion rate. There is.

【0005】[0005]

【課題を解決するための手段】本発明は、オンライン手
書き文字認識手段を持ち、認識された漢字混じりかな文
字情報を入力文字情報とし、これに漢字混じりかな漢字
変換を行う文書作成装置において、前記入力文字情報と
漢字混じりかな漢字変換後に選択された変換候補の文字
情報を比較して、入力時に漢字表記で入力された文字情
報を記憶する学習手段と、前記漢字混じりかな漢字変換
を行って得られた変換候補の文字情報の中に前記学習手
段に記憶されている漢字が存在するか否かを判定する第
1の判定手段と、前記第1の判定手段によって前記変換
候補の文字情報中に前記学習手段に記憶されている漢字
が存在すると判定された時、変換対象の入力文字情報中
に前記漢字が含まれているか否かを判定する第2の判定
手段と、前記第2の判定手段によって前記漢字が変換対
象文字情報に含まれていない場合に前記漢字を含む変換
候補の出力優先順位を下げる優先順位調整手段とを具備
した構成を有する。
According to the present invention, there is provided an on-line handwritten character recognizing means, wherein the recognized kanji-mixed kana character information is used as input character information, and kanji-mixed kana-kanji conversion is performed on the input character information. A learning means for comparing character information and character information of conversion candidates selected after conversion of kanji mixed with kanji and storing the character information input in kanji notation at the time of input, and conversion obtained by performing kana-kanji conversion with kanji mixed First determining means for determining whether or not there is a kanji stored in the learning means in the candidate character information, and the learning means in the conversion candidate character information by the first determining means. Second judgment means for judging whether or not the input character information to be converted includes the Chinese character when it is determined that the Chinese character stored in the second character exists; Having the configuration and a priority adjustment means for reducing the output priority of conversion candidates the Chinese characters comprising the kanji if not included in the converted character information by the decision means.

【0006】[0006]

【作用】本発明の文書作成装置において、学習手段は入
力文字列と選択された変換候補の文字列を比較して漢字
表記で入力された文字を記憶する。第1の判定手段は前
記漢字混じりかな漢字変換を行って得られた変換候補の
文字列の中に前記学習手段に記憶されている漢字が存在
するかしないかを判定する。第2の判定手段は前記第1
の判定手段によって前記変換候補の文字列中に前記学習
手段に記憶されている漢字が存在すると判定されると、
この時の変換対象の入力文字列中に前記漢字が含まれて
いるかいないかを判定する。優先順位調整手段は前記第
2の判定手段によって前記漢字が変換対象文字列に含ま
れていないことが明らかになった場合に前記漢字を含む
変換候補の出力優先順位を下げる。
In the document creating apparatus of the present invention, the learning means compares the input character string with the selected character string of the conversion candidate and stores the input character in Kanji notation. The first determining means determines whether or not the kanji stored in the learning means exists in the conversion candidate character string obtained by performing the kanji mixed kanji conversion. The second determination means is the first
When it is determined that the kanji stored in the learning means is present in the conversion candidate character string,
At this time, it is determined whether the input character string to be converted contains the Chinese character. The priority adjusting unit lowers the output priority of the conversion candidate including the Chinese character when it is determined by the second determining unit that the Chinese character is not included in the conversion target character string.

【0007】[0007]

【実施例】以下、本発明の一実施例を図面を参照して説
明する。図1は本発明の文書作成装置の一実施例を示し
たブロック図である。1は文字列を利用者がペンで入力
すると、オンライン手書き文字認識処理を行い、対応す
る文字コードを出力する手書き透明タブレット等の入力
装置で、これは後述する表示装置18の上に重ね合わせ
ては位置され、表示情報が見える構造になっている。2
は入力された文字列や設定データ等を必要な部分に振り
分ける入力制御部、3は入力された文字列等を一旦保存
する入力バッファ、4は入力文書の書式を設定する書式
制御部、5は入力される漢字混じりかな文字列を漢字混
じりかな漢字変換する漢字混じりかな漢字変換部、6は
変換候補の漢字コードが過去に漢字表記で入力されたか
否かを調べる入力調査部、7は変換対象の漢字混じりか
な文字列が一旦保存される変換用バッファ、8は漢字混
じりかな漢字変換用辞書、9は漢字混じりかな漢字変換
により得られた変換候補を選択するための各種処理を行
う変換候補制御部、10は全ての変換候補の文字列に漢
字コードがあるかないかを調査すると共に、あった場合
にその位置を調査する候補漢字調査部、11は漢字混じ
りかな漢字変換により得られる変換候補を一旦保存する
変換候補バッファ、12は選択された変換候補中の漢字
が入力状態で漢字/かなのいずれの表記であったかを調
査する漢字表記学習部、13は選択候補の各漢字が入力
状態で漢字/かなのいずれの表記であったかを計数する
表記種別カウンタ、14は選択候補の漢字で漢字表記で
入力されたものを記憶する漢字ビットマップ、15は漢
字混じりかな漢字変換して確定された文字コードを保存
する文書バッファ、16は表示用バッファ16に展開さ
れている表示用データを表示装置18に表示する表示制
御部、17は表示装置18に表示する表示用データを展
開する表示用バッファ、18は文書や図形等を表示する
CRTやLCD等の表示装置である。この表示装置18
は前記入力装置1と一体に重ね合った構造になってい
る。
DETAILED DESCRIPTION OF THE PREFERRED EMBODIMENTS An embodiment of the present invention will be described below with reference to the drawings. FIG. 1 is a block diagram showing an embodiment of a document creating apparatus of the present invention. Reference numeral 1 denotes an input device such as a handwritten transparent tablet that performs online handwritten character recognition processing when a user inputs a character string with a pen and outputs a corresponding character code, which is superimposed on a display device 18 described later. Is located so that the displayed information can be seen. Two
Is an input control unit that sorts input character strings and setting data into necessary parts, 3 is an input buffer that temporarily stores the input character strings, and 4 is a format control unit that sets the format of the input document. Kanji-mixed kana-kanji conversion unit that converts kanji-mixed kana-kana strings to kanji-mixed kana-kanji, 6 is an input research part that checks whether or not the kanji code of the conversion candidate has been input in kanji notation in the past, and 7 is the kanji to be converted A conversion buffer in which mixed kana character strings are once stored, 8 a kanji mixed kanji kanji conversion dictionary, 9 a conversion candidate control unit for performing various processes for selecting conversion candidates obtained by kanji mixed kana kanji conversion, 10 Candidate Kanji Research Department that investigates whether or not there is a Kanji code in the character strings of all conversion candidates, and if so, 11 is a Kanji conversion for mixed Kanji characters. The conversion candidate buffer that temporarily stores the obtained conversion candidates, 12 is a kanji notation learning unit that investigates whether the kanji in the selected conversion candidate is the kanji / kana notation in the input state, and 13 is each of the selection candidates. A notation type counter that counts whether the kanji was written as kanji or kana in the input state, 14 is a kanji bitmap that stores the input kanji in the kanji notation, and 15 is kanji mixed kana kanji conversion A document buffer for storing the confirmed character code, 16 a display control unit for displaying the display data expanded in the display buffer 16 on the display device 18, and 17 expanding the display data for display on the display device 18. A display buffer, 18 is a display device such as a CRT or LCD for displaying documents, figures and the like. This display device 18
Has a structure that is integrally overlapped with the input device 1.

【0008】次に本実施例の動作について説明する。通
常は手書きタブレット等の入力装置1にオペレータがペ
ン等で漢字混じりかな文字列を書くことにより,これを
図示しないオンライン手書き文字認識部が文字認識を行
い、認識文字列を生成し、これを文書として装置に入力
する。漢字混じりかな文字入力/変換はペン入力手段を
持つ文書作成装置において重要な入力/変換手段で、画
数の少ない漢字はその漢字のまま入力でき、複雑で画数
の多い漢字はその読みを入力し、これをかな漢字変換で
きるもので、入力及び変換速度が従来のペン入力手段
(かな文字入力/変換手段のみ)を持つ文書作成装置に
よる文書作成よりも文字認識及び変換効率が高い。この
際、オペレータは表示装置18に表示される各種メッセ
ージ等を参照して対話的に文書作成作業を進めて行く。
入力装置1から入力された文字列は入力制御部2を通し
て、入力バッファ3に変換処理待ちの間一旦保存された
後、再び入力制御部2から漢字混じりかな漢字変換部5
へ送られる。漢字混じりかな漢字変換部5は変換対象と
なる入力文字列が単語や文節等を形成する適当な長さに
なるまで一旦変換用バッファ7に格納する。変換用バッ
ファ7に格納された文字列の長さが句読点の文字コード
の検出によって区切られた長さ、即ち、変換対象の長さ
となった段階で、漢字混じりかな漢字変換部5は漢字混
じりかな漢字変換用辞書9を参照しながら、前記変換対
象文字列の漢字混じりかな漢字変換処理を開始する。
Next, the operation of this embodiment will be described. Usually, an operator writes a kana-mixed kana character string on the input device 1 such as a handwriting tablet with a pen or the like, and an online handwritten character recognition unit (not shown) performs character recognition to generate a recognized character string and document it. As input to the device. Kanji mixed kana character input / conversion is an important input / conversion means in a document creation device having a pen input means. Kanji with a small number of strokes can be input as it is, and kanji with a large number of strokes can be input with its reading, It can convert kana-kanji into characters, and its input and conversion speeds are higher than those of a document creation device having a conventional pen input means (only kana character input / conversion means). At this time, the operator refers to various messages displayed on the display device 18 and interactively proceeds with the document creation work.
The character string input from the input device 1 is temporarily stored in the input buffer 3 through the input control unit 2 while waiting for the conversion process, and then is again input from the input control unit 2 into the kanji / kanji conversion unit 5.
Sent to. The kanji / kanji / kanji conversion unit 5 temporarily stores the input character string to be converted into the conversion buffer 7 until the input character string has an appropriate length for forming a word, a phrase, or the like. At the stage when the length of the character string stored in the conversion buffer 7 is separated by the detection of the punctuation character code, that is, the length of the conversion target, the kanji-mixed kana-kanji conversion unit 5 converts kanji-mixed kana-kanji With reference to the dictionary 9, the kana-kanji conversion processing for the conversion target character string is started.

【0009】上記漢字混じりかな漢字変換部5の変換開
始とともに、入力漢字調査部6は変換対象となっている
変換用バッファ7の文字列を調査し、この文字列の中に
漢字ビットマップ14に記憶されている漢字コードがあ
るかないかを調査すると共に、ある場合はその位置を調
査する。ここで、漢字ビットマップ14とは漢字コード
に対応して最低でもオン・オフの情報が入力できるもの
で、オン状態で記憶済み漢字を表すことができる。即
ち、漢字表記により入力装置1から直接入力された経歴
がある漢字コード、又は、表記種別カウンタを有する場
合は、漢字表記で入力された割合が一定値より高い漢字
コードに対応する漢字ビットマップ14上のビットがオ
ンにされる。更に、このような漢字ビットマップ14は
オン・オフの2段階でなく3段階以上のレベルを入力で
きるものでもよい。
At the same time as the conversion of the Kanji-mixed Kana-Kanji conversion unit 5 is started, the input Kanji-character inspection unit 6 investigates the character string in the conversion buffer 7 to be converted, and stores it in the Kanji bitmap 14 in this character string. Check whether or not there is a kanji code that is displayed, and if there is, check its position. Here, the kanji character bitmap 14 is for inputting at least on / off information corresponding to a kanji code, and can represent a stored kanji in the on state. That is, if there is a kanji code that has a history of being directly input from the input device 1 in kanji notation, or if it has a notation type counter, the kanji bit map 14 corresponding to the kanji code for which the proportion input in kanji notation is higher than a certain value. The upper bit is turned on. Further, such a Kanji bit map 14 may be capable of inputting not only two levels of on / off but three or more levels.

【0010】漢字混じりかな漢字変換処理が終了する
と、漢字混じりかな漢字変換部5は文節間接続点数や単
語使用頻度に基づいて出力優先順位を付加した全ての変
換候補を変換候補制御部9を通して変換候補バッファ1
1に格納する。次に候補漢字調査部10は変換候補バッ
ファ11に格納された全ての変換候補の文字列の中に前
記漢字ビットマップ14に記憶されている漢字コードが
あるかないかを調査すると共に、ある場合はその位置を
調査する。候補漢字調査部10によって漢字ビットマッ
プ14に記憶済みの漢字コードが変換候補中にみつかっ
た場合、変換候補制御部9は入力漢字調査部6の調査結
果から前記みつかった漢字コードが漢字表記で入力され
たか否かを調べる。その結果、変換候補制御部9は前記
漢字コードに対応する文字がかな表記で入力されていた
ことが分かると、変換候補バッファ11に格納されてい
る前記漢字候補を含む変換候補の出力優先度を下げる。
変換候補制御部9は出力優先度の調整が終了すると、出
力優先度が最も高い変換候補を表示制御部16に送って
表示装置18に表示することにより、オペレータに対し
変換候補の選択を促す。その結果、オペレータが選択し
た候補は漢字表記学習部12に送られる。これにより、
漢字表示学習部12は選択候補文字列と変換用バッファ
7に格納されている入力文字列とを比較し、選択候補の
各漢字が入力状態で漢字・かないずれの表記であったか
を調査し、その調査結果を表記種別カウンタ13の該当
漢字のカウンタにセットする。更に、漢字表記学習部1
2は、表記種別カウンタ13にセットされた値より、あ
る漢字がかな表記入力に対し一定以上の割合で漢字表記
入力されていることが分かると、このような漢字のコー
ドに対応する漢字ビットマップ14をオン状態にし、逆
に一定値を下回った場合は前記漢字コードに対応する漢
字ビットマップ14をオフ状態にする。また、これと同
時に、変換候補制御部9は選択された変換候補の文字コ
ードを文書バッファ15に格納すると共に、表示制御部
16を介して、予め書式制御部4に設定された書式に従
って表示用バッファ17に前記文字コードを展開して、
表示装置18で表示する。尚、オペレータが選択処理を
することなく、次の文字入力を開始した場合は、前記表
示した第1候補を選択したものとして、上記と同様の処
理が行われる。
When the kanji-mixed kana-kanji conversion process is completed, the kanji-mixed kana-kanji conversion unit 5 passes all conversion candidates to which the output priority is added based on the number of connection points between phrases and the frequency of word use through the conversion candidate control unit 9. 1
Store in 1. Next, the candidate kanji character checking unit 10 checks whether or not all the conversion candidate character strings stored in the conversion candidate buffer 11 include the kanji code stored in the kanji bitmap 14, and if there is, Investigate its location. When the candidate kanji survey unit 10 finds a Kanji code stored in the Kanji bitmap 14 in the conversion candidates, the conversion candidate control unit 9 inputs the found Kanji code from the survey result of the input Kanji survey unit 6 in Kanji notation. Check whether it has been done. As a result, when the conversion candidate control unit 9 finds that the character corresponding to the Kanji code is input in Kana notation, it outputs the output priority of the conversion candidate including the Kanji candidate stored in the conversion candidate buffer 11. Lower.
After the adjustment of the output priority is completed, the conversion candidate control unit 9 sends the conversion candidate having the highest output priority to the display control unit 16 and displays it on the display device 18, thereby prompting the operator to select the conversion candidate. As a result, the candidates selected by the operator are sent to the kanji writing learning unit 12. This allows
The kanji display learning unit 12 compares the selection candidate character string with the input character string stored in the conversion buffer 7, investigates whether each kanji of the selection candidate is a kanji or kana notation in the input state, and The survey result is set in the corresponding kanji character counter of the notation type counter 13. In addition, Kanji notation learning unit 1
2 indicates that the value set in the notation type counter 13 indicates that a certain kanji is input in kanji notation at a certain ratio or more relative to kana notation input, and the kanji bitmap corresponding to such kanji code. 14 is turned on, and conversely, when it is below a certain value, the kanji bitmap 14 corresponding to the kanji code is turned off. At the same time, the conversion candidate control unit 9 stores the character code of the selected conversion candidate in the document buffer 15 and displays it via the display control unit 16 in accordance with the format preset in the format control unit 4. Expand the character code in the buffer 17,
It is displayed on the display device 18. When the operator starts inputting the next character without performing the selection process, the same process as above is performed assuming that the displayed first candidate is selected.

【0011】図2は図1に示した装置において文字列の
入力から変換候補の選択までの処理の流れを示したフロ
ーチャートである。入力装置1からステップ201にて
入力された文字列は入力制御部2よって一旦入力バッフ
ァ3に格納された後、ステップ202にて漢字混じりか
な漢字変換部5を通して変換用バッファ7に転送され
る。漢字混じりかな漢字変換部5は前記変換用バッファ
7に入力される文字列が単語、文節など適当な長さにな
ったか否かをステップ203にて監視し、適当な長さに
なった段階でステップ204に進む。尚、前記変換用バ
ッファ7内の文字列が変換対象となった場合には変換処
理終了まで入力制御部2からの文字入力を停止する。変
換対象となる文字列が全て変換用バッファ7に格納され
ると、入力漢字調査部6はステップ204にて漢字ビッ
トマップ14を参照し、変換用バッファ7内の文字列中
に漢字表記での入力履歴がある漢字、即ち漢字ビットマ
ップ14がオン状態である漢字が存在するか否かを調査
し、存在する場合にはその文字位置を記憶する。一方、
漢字混じりかな漢字変換部5はステップ205にて変換
用バッファ7内の文字列を漢字混じりかな漢字変換用辞
書8を参照して漢字混じりかな漢字変換する。漢字混じ
りかな漢字変換部5はステップ206にて漢字混じりか
な漢字変換により得られた変換候補を文節間接続点数や
単語使用頻度から求められた出力優先度と共に変換候補
バッファ11に格納する。
FIG. 2 is a flow chart showing the flow of processing from the input of a character string to the selection of conversion candidates in the apparatus shown in FIG. The character string input from the input device 1 in step 201 is temporarily stored in the input buffer 3 by the input controller 2, and then transferred in step 202 to the conversion buffer 7 through the kanji / kanji conversion unit 5. The kana / kanji / kanji conversion unit 5 monitors in step 203 whether or not the character string input to the conversion buffer 7 has an appropriate length such as a word or a phrase, and when the appropriate length is reached, the step is performed. Proceed to 204. When the character string in the conversion buffer 7 is to be converted, the character input from the input control unit 2 is stopped until the conversion process is completed. When all the character strings to be converted are stored in the conversion buffer 7, the input Kanji character checking unit 6 refers to the Kanji bitmap 14 in step 204 to display the Kanji notation in the character string in the conversion buffer 7. It is checked whether or not there is a kanji with an input history, that is, a kanji for which the kanji bitmap 14 is on, and if there is, the character position is stored. on the other hand,
In step 205, the kanji-mixed kana-kanji conversion unit 5 converts the character string in the conversion buffer 7 into kanji-mixed kana-kanji by referring to the kanji-mixed kana-kanji conversion dictionary 8. In step 206, the kanji-mixed kana-kanji conversion unit 5 stores the conversion candidates obtained by the kanji-mixed kana-kanji conversion in the conversion candidate buffer 11 together with the output priority obtained from the number of connection points between phrases and the word usage frequency.

【0012】次に候補漢字調査部10はステップ207
にて全ての変換候補に対し順次処理を行うために、カウ
ンタiを初期化する。候補漢字調査部10はステップ2
08にて漢字ビットマップ14を参照して、変換候補バ
ッファ11内の第i番目の変換候補の文字列に漢字表記
での入力履歴がある漢字が存在するか否か調査し、存在
する場合にはステップ209に進み、ここで、入力漢字
調査部6での調査結果を参照して、該当漢字が入力文字
列中ではひらがなであったか否かを調べる。もし、ひら
がなであった場合、変換候補制御部9はステップ210
にて変換候補バッファ11内のi番目変換候補の出力優
先順位を下げる。その後、候補漢字調査部10はステッ
プ211にて前記カウンタiを1だけ進め、ステップ2
12にて変換候補バッファ11内の全ての変換候補につ
いて上記したステップ208〜210の処理を行ったか
否かを判定し、行った場合はステップ213に進む。上
記した処理により、全ての変換候補に対して出力優先順
位の見直しが終了すると、変換候補制御部9は最も出力
優先順位の高い変換候補を表示制御部16を介して表示
装置18の画面上にステップ213にて表示した後、ス
テップ214にてオペレータの候補選択処理を待つ。
Next, the candidate kanji character investigation unit 10 performs step 207.
In order to sequentially process all conversion candidates, the counter i is initialized. Candidate Kanji survey unit 10 is step 2
At 08, referring to the Kanji bitmap 14, it is checked whether or not there is a Kanji for which an input history in Kanji notation exists in the character string of the i-th conversion candidate in the conversion candidate buffer 11, and if it exists, Advances to step 209, where it is checked whether or not the corresponding Chinese character is a Hiragana character in the input character string by referring to the result of the investigation by the input Chinese character research unit 6. If it is Hiragana, the conversion candidate control unit 9 performs step 210.
At, the output priority of the i-th conversion candidate in the conversion candidate buffer 11 is lowered. After that, the candidate kanji character research unit 10 advances the counter i by 1 in step 211, and then proceeds to step 2
In step 12, it is determined whether or not the above steps 208 to 210 have been performed for all conversion candidates in the conversion candidate buffer 11, and if so, the process proceeds to step 213. When the review of the output priorities for all conversion candidates is completed by the above-described processing, the conversion candidate control unit 9 displays the conversion candidate with the highest output priority on the screen of the display device 18 via the display control unit 16. After displaying in step 213, the operator waits for candidate selection processing in step 214.

【0013】その後、オペレータによって選択された候
補文字列は変換候補制御部9によって変換候補バッファ
11から漢字表記学習部12に送られる。漢字表記学習
部12はステップ215にて前記送られてきた候補文字
列の漢字表記となっている全ての文字の入力時の表記が
漢字であったかひらがなであったかを調べ、漢字表記さ
れた文字毎にその結果を表記種別カウンタ13にセット
する。更に、漢字表記学習部12は表記種別カウンタ1
3をセットした後に、この表記種別カウンタ13漢字表
記及びかな表記入力のカウント数を調査し、漢字表記入
力の割合が一定値以上とステップ216にて判定された
場合は、ステップ217にて該当する漢字ビットマップ
14をオン状態に、即ち通常漢字表記で入力される漢字
であることを記憶し、またステップ216にて一定値未
満であると判定された場合は該当する漢字ビットマップ
14をステップ218にてオフ状態に、即ち通常ひらが
な表記で入力される漢字であることを記憶する。このよ
うにして表記種別カウンタ13のセットが終了すると、
漢字表記学習部12はステップ219にて選択候補文字
列を文書バッファ15に転送して格納する。表示制御部
16は文書バッファ15に格納された文字列を書式設定
部4に予め設定された書式に従って表示用バッファ17
上に展開した後、これを表示装置18上の画面上にステ
ップ220にて表示する。更に、漢字混じりかな漢字変
換部5は次の入力・変換に備えるために変換用バッファ
7をステップ221にて初期化しする。
After that, the candidate character string selected by the operator is sent from the conversion candidate buffer 11 to the Kanji writing learning unit 12 by the conversion candidate control unit 9. In step 215, the kanji notation learning unit 12 checks whether all the characters which are the kanji notation of the sent candidate character string are kanji or hiragana at the time of input, and for each kanji notated character. The result is set in the notation type counter 13. Further, the kanji notation learning unit 12 uses the notation type counter 1
After setting 3, the notation type counter 13 is checked for the number of kanji notation and kana notation input counts, and if it is determined in step 216 that the proportion of kanji notation input is greater than or equal to a certain value, it corresponds in step 217. When the kanji bitmap 14 is turned on, that is, it is stored that the kanji is input in the normal kanji notation, and if it is determined in step 216 that the kanji is less than a certain value, the corresponding kanji bitmap 14 is set in step 218. It is stored in the OFF state, that is, it is a kanji that is normally input in hiragana notation. When the setting of the notation type counter 13 is completed in this way,
The kanji writing learning unit 12 transfers and stores the selection candidate character string in the document buffer 15 in step 219. The display control unit 16 causes the character string stored in the document buffer 15 to be displayed in the display buffer 17 according to the format preset in the format setting unit 4.
After unfolding up, this is displayed on the screen of the display device 18 in step 220. Further, the kanji / kanji / kanji conversion unit 5 initializes the conversion buffer 7 in step 221 in preparation for the next input / conversion.

【0014】図3は図1に示した漢字ビットマップ14
の構造例を示した図である。この例は漢字ビットマップ
に最も単純な2段階のオン・オフ情報だけを入力できる
場合で、文字コードの順番に、1文字に対し1ビットが
割り当てられている。即ち、この漢字ビットマップ14
では、1度でも漢字表記で入力された文字に該当するビ
ットがオン(=1)になる。この例の場合では、「亜」
および「悪」だけが入力段階で直接漢字表記であった履
歴を有することを示している。また、図4は漢字ビット
マップ14が表記種別カウンタ13と連係する2段階構
成となっている例を示しており、漢字・かな表記の行わ
れた回数をカウントとすることが可能な装置を示してい
る。図4(A)に示した漢字ビットマップ14の1段目
の第0ビットが、前記図3における直接漢字表記の入力
が行われた履歴を示す情報である。第1〜fhビットは
表記の回数を記憶するカウンタへのオフセットポインタ
を示す数値である。図4(B)に示すように表記種別カ
ウンタ13はベースアドレスX+オフセットポインタで
示されるアドレス部分にあり、上位・下位各4ビットが
それぞれ漢字表記の回数、ひらがな表記の回数を示して
いる。図4に示すように「亜」は漢字表記で8回入力さ
れ、ひらがな表記で入力後変換処理を行った文字列で漢
字表記が選択された回数は0回であることを表してい
る。漢字表記学習部12は文字の表記状態から表記種別
カウンタ13の値を変化させ、予め設定された割合、例
えば漢字表記が9割以上の場合に漢字表記情報に1をセ
ットする。従って、例えば、漢字表記での入力が1回あ
っても、ひらがな表記での入力が2回ある場合には設定
値を下回るために漢字表記情報は0になる。
FIG. 3 shows the Kanji bitmap 14 shown in FIG.
It is the figure which showed the structural example. In this example, only the simplest two-step on / off information can be input to the Chinese character bitmap, and one bit is assigned to one character in the character code order. That is, this Kanji bitmap 14
Then, the bit corresponding to the character input in Kanji notation is turned on (= 1) even once. In this case, "A"
It is shown that only and "evil" have a history of being directly written in Kanji at the input stage. Further, FIG. 4 shows an example in which the Kanji bitmap 14 has a two-stage configuration in which it is linked to the notation type counter 13, and shows an apparatus capable of counting the number of times Kanji / Kana notation has been performed. ing. The 0th bit in the first row of the Kanji bitmap 14 shown in FIG. 4A is information indicating the history of the direct Kanji notation shown in FIG. The 1st to fh bits are a numerical value indicating an offset pointer to a counter that stores the number of times of notation. As shown in FIG. 4 (B), the notation type counter 13 is located at the address portion indicated by the base address X + offset pointer, and the upper and lower 4 bits respectively indicate the number of Kanji notation and the number of Hiragana notation. As shown in FIG. 4, “A” is input eight times in the Kanji notation, and the number of times the Kanji notation is selected in the character string subjected to the conversion process after the input in the Hiragana notation is 0. The kanji notation learning unit 12 changes the value of the notation type counter 13 from the notation state of characters, and sets 1 in the kanji notation information when a preset ratio, for example, 90% or more kanji notation. Therefore, for example, even if there is one input in Kanji notation, if there are two inputs in Hiragana notation, the Kanji notation information becomes 0 because the value is below the set value.

【0015】図5は、例文として「せいかいでは、」を
変換した際の変換候補バッファ11の格納例を示した図
である。図5は漢字ビットマップ14の格納例を示して
おり、「星」「正」「生」「声」「西」は過去に直接漢
字で入力されたことがあることを示している。図6
(A)は例文「せいかいでは」を変換した時得られる変
換候補を変換候補バッファ11内に格納した際のデータ
例を示した図である。変換候補は候補文字列と共に出力
優先度が付加されて変換候補バッファ11に格納され
る。漢字ビットマップ14による影響がない場合には、
この出力優先度の最も高い候補が第1候補となる。しか
し、本例では図6(A)に示したように漢字ビットマッ
プ14には「正」が過去に漢字で入力されたことが記憶
されており、又、入力文では「正」でなくかな表記の
「せい」で入力されていることから、図6(B)に示す
ように変換候補制御部9は「正解」の出力優先度を下げ
て、結果的に「政界」が第1候補となるように出力優先
度を変更している。本例では出力優先度は半分になって
いるが、表記種別カウンタ13に記憶されている表記種
別の割合や、あるいは漢字ビットマップ14に少なくと
も3段階以上の情報が格納できる場合にはその値によっ
て出力優先度の変動幅を変更することもできる。
FIG. 5 is a diagram showing an example of storage in the conversion candidate buffer 11 when "seikai de," is converted as an example sentence. FIG. 5 shows an example of storing the Kanji bitmap 14, and shows that “star”, “correct”, “raw”, “voice”, and “west” have been directly input in Kanji in the past. Figure 6
(A) is a diagram showing an example of data when a conversion candidate obtained when converting the example sentence “Seikai Ide” is stored in the conversion candidate buffer 11. The conversion candidates are stored in the conversion candidate buffer 11 with the output priority added together with the candidate character strings. If there is no effect from the Kanji bitmap 14,
The candidate with the highest output priority becomes the first candidate. However, in this example, as shown in FIG. 6 (A), it is stored in the Kanji bitmap 14 that “correct” has been input in kanji in the past, and the input sentence may not be “correct”. Since the input is “SEI”, the conversion candidate control unit 9 lowers the output priority of “correct answer” as shown in FIG. 6B, and as a result, “politics” is the first candidate. The output priority is changed so that In this example, the output priority is halved. However, depending on the ratio of the notation type stored in the notation type counter 13 or the value if at least three levels of information can be stored in the Kanji bitmap 14, The fluctuation range of the output priority can be changed.

【0016】本実施例によれば、ある漢字が入力の際に
かな表記であったか漢字表記であったかを記憶しておい
てから、入力文字列を漢字混じりかな漢字変換した際に
得られる変換候補中に、通常漢字表記で入力される漢字
が前記記憶しておいたデータを参照して存在することが
分かり、且つ前記漢字が前記入力文字列ではかな表記で
あった場合、前記漢字表記の変換候補の出力優先順位を
下げることによって、オペレータが意図しない余分な変
換候補の出力を抑制して、漢字混じりかな漢字変換率を
高めることができる。
According to this embodiment, it is memorized whether a certain kanji was in kana or kanji notation at the time of input, and then, in the conversion candidates obtained when kana-mixed kana kanji conversion is performed on the input character string. If it is found that a kanji input in the normal kanji notation exists by referring to the stored data and the kanji is the kana notation in the input character string, the conversion candidate of the kanji notation is selected. By lowering the output priority, it is possible to suppress the output of extra conversion candidates that the operator does not intend, and to increase the kanji conversion rate where kanji is mixed.

【0017】[0017]

【発明の効果】以上記述した如く本発明の文書作成装置
によれば、オペレータの意図どおりの漢字かな混じりか
な漢字変換が行われるようにして、その変換率を向上さ
せることができる。
As described above, according to the document creating apparatus of the present invention, kanji and kana mixed kanji conversion as intended by the operator can be performed, and the conversion rate can be improved.

【図面の簡単な説明】[Brief description of drawings]

【図1】本発明の文書作成装置の一実施例を示したブロ
ック図。
FIG. 1 is a block diagram showing an embodiment of a document creation device of the present invention.

【図2】図1に示し装置の文字列の入力から変換候補の
選択までの処理の流れを示したフローチャート。
FIG. 2 is a flowchart showing a flow of processing from input of a character string of the device shown in FIG. 1 to selection of a conversion candidate.

【図3】図1に示した漢字ビットマップの構造例を示し
た図。
FIG. 3 is a diagram showing an example of the structure of the Kanji bitmap shown in FIG.

【図4】図1に示した漢字ビットマップとこれに連係し
ている表記種別カウンタの構造例を示した図。
FIG. 4 is a diagram showing a structural example of the Kanji bitmap shown in FIG. 1 and a notation type counter associated therewith.

【図5】図1に示した漢字ビットマップの他の構造例を
示した図。
FIG. 5 is a diagram showing another example of the structure of the Kanji bitmap shown in FIG.

【図6】図1に示した変換候補バッファの内容例を示し
た図。
FIG. 6 is a diagram showing an example of contents of a conversion candidate buffer shown in FIG.

【符号の説明】[Explanation of symbols]

1…入力装置 2…入力制御部 3…入力バッファ 4…書式制御部 5…漢字混じりかな漢字変換部 6…入力漢字調
査部 7…変換用バッファ 8…漢字混じり
かな漢字変換辞書 9…変換候補制御部 10…候補漢字
調査部 11…変換候補バッファ 12…漢字表記
学習部 13…表記種別カウンタ 14…漢字ビッ
トマップ 15…文書バッファ 16…表示制御
部 17…表示用バッファ 18…表示装置
DESCRIPTION OF SYMBOLS 1 ... Input device 2 ... Input control unit 3 ... Input buffer 4 ... Format control unit 5 ... Kanji mixed kana / Kanji conversion unit 6 ... Input Kanji research unit 7 ... Conversion buffer 8 ... Kanji mixed Kana / Kanji conversion dictionary 9 ... Conversion candidate control unit 10 ... Candidate Kanji research unit 11 ... Conversion candidate buffer 12 ... Kanji notation learning unit 13 ... Notation type counter 14 ... Kanji bitmap 15 ... Document buffer 16 ... Display control unit 17 ... Display buffer 18 ... Display device

Claims (2)

【特許請求の範囲】[Claims] 【請求項1】 オンライン手書き文字認識手段を持ち、
認識された漢字混じりかな文字情報を入力文字情報と
し、これに漢字混じりかな漢字変換を行う文書作成装置
において、 前記入力文字情報と漢字混じりかな漢字変換後に選択さ
れた変換候補の文字情報を比較して、入力時に漢字表記
で入力された文字情報を記憶する学習手段と、 前記漢字混じりかな漢字変換を行って得られた変換候補
の文字情報の中に前記学習手段に記憶されている漢字が
存在するか否かを判定する第1の判定手段と、 前記第1の判定手段によって前記変換候補の文字情報中
に前記学習手段に記憶されている漢字が存在すると判定
された時、変換対象の入力文字情報中に前記漢字が含ま
れているか否かを判定する第2の判定手段と、 前記第2の判定手段によって前記漢字が変換対象文字情
報に含まれていない場合に前記漢字を含む変換候補の出
力優先順位を下げる優先順位調整手段とを具備したこと
を特徴とする文書作成装置。
1. An online handwritten character recognition means is provided,
In the document creation device that performs the kana-mixed kana-kanji conversion on the recognized kanji-mixed kana-character information, compares the input character information with the conversion candidate character information selected after kanji-mixed kana-kanji conversion. A learning unit that stores the character information input in Kanji notation at the time of input, and whether or not the Kanji stored in the learning unit exists in the character information of the conversion candidates obtained by performing the Kana-mixed Kanji conversion. A first determining means for determining whether or not there is a Chinese character stored in the learning means in the character information of the conversion candidate by the first determining means, and Second determining means for determining whether or not the Chinese character is included in, and the Chinese character if the Chinese character is not included in the conversion target character information by the second determining means. A document creation apparatus comprising: a priority adjustment unit that lowers the output priority of conversion candidates including characters.
【請求項2】 オンライン手書き文字認識手段を持ち、
認識された漢字混じりかな文字情報を入力文字情報と
し、これに漢字混じりかな漢字変換し変換候補を得る文
書作成方法であって、 通常漢字表示で入力される文字が、前記変換候補に含ま
れ、且つ入力文字情報ではひらがな表記であった場合
に、前記変換候補の出力順位を下げることを特徴とする
候補出力順位制御方法。
2. An online handwritten character recognition means is provided,
A method for creating a document in which kana-kana-kana character information that is recognized is used as input character information and kanji-kana kana-kanji conversion is performed to obtain conversion candidates, and characters that are normally input in kanji display are included in the conversion candidates, and A candidate output rank control method, wherein the output rank of the conversion candidates is lowered when the input character information is in Hiragana notation.
JP5102031A 1993-04-28 1993-04-28 Document preparing device and candidate output order control method Pending JPH06314273A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP5102031A JPH06314273A (en) 1993-04-28 1993-04-28 Document preparing device and candidate output order control method

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP5102031A JPH06314273A (en) 1993-04-28 1993-04-28 Document preparing device and candidate output order control method

Publications (1)

Publication Number Publication Date
JPH06314273A true JPH06314273A (en) 1994-11-08

Family

ID=14316395

Family Applications (1)

Application Number Title Priority Date Filing Date
JP5102031A Pending JPH06314273A (en) 1993-04-28 1993-04-28 Document preparing device and candidate output order control method

Country Status (1)

Country Link
JP (1) JPH06314273A (en)

Similar Documents

Publication Publication Date Title
EP0686291B1 (en) Combined dictionary based and likely character string handwriting recognition
US5513278A (en) Handwritten character size determination apparatus based on character entry area
JPH05233630A (en) Method for describing japanese and chinese
US4677585A (en) Method for obtaining common mode information and common field attribute information for a plurality of card images
JP2005508031A (en) Adaptable stroke order system based on radicals
US7911452B2 (en) Pen input method and device for pen computing system
JP2740575B2 (en) Character processor
JPH06314273A (en) Document preparing device and candidate output order control method
JPH0782530B2 (en) Handwriting recognition device
JPH08190603A (en) Character recognition device and its candidate character display method
JP2005050175A (en) Image data document retrieval system
KR100228902B1 (en) Apparatus for emphasiging the first charator of a sentence and method thereof
JPH0895970A (en) Document creation device and kana-kanji conversion method
JPH07210629A (en) Character recognition method
JPH0895971A (en) Document creation device and kana-kanji conversion method
JPH0267676A (en) Kanji numeric conversion processing device
JPH07254044A (en) Handwriting recognition device
JPH0827819B2 (en) Handwritten character recognition method for handwritten character recognition and handwritten character recognition apparatus using the same
JPH0574867B2 (en)
JPH07334497A (en) Document creating device, kanji mixed kana conversion method, kanji input learning method
JPH0736884A (en) Input device for character recognition
JPH1049528A (en) Automatic agate setting method and system therefor
JPS605315A (en) Word processor
JPH0773177A (en) Document creating apparatus and method
JPH04256162A (en) Word processor with learning function