JPH0916721A - Character recognition candidate selector - Google Patents
Character recognition candidate selectorInfo
- Publication number
- JPH0916721A JPH0916721A JP7163172A JP16317295A JPH0916721A JP H0916721 A JPH0916721 A JP H0916721A JP 7163172 A JP7163172 A JP 7163172A JP 16317295 A JP16317295 A JP 16317295A JP H0916721 A JPH0916721 A JP H0916721A
- Authority
- JP
- Japan
- Prior art keywords
- character
- characters
- candidate
- recognition
- recognition candidate
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Pending
Links
Landscapes
- Character Discrimination (AREA)
Abstract
Description
【0001】[0001]
【産業上の利用分野】本発明は、入力された文字のイメ
ージを解析して文字認識を行い一致する文字の候補とな
る文字を選択する文字認識候補選択装置に係わり、特に
文字のイメージを基に選択した所定数の文字の中に入力
した文字と一致する文字が無いとき、文字の“つくり”
や“へん”など文字を構成する各要素を指定して入力し
た文字と一致する文字の候補を選択する文字認識候補選
択装置に関する。BACKGROUND OF THE INVENTION 1. Field of the Invention The present invention relates to a character recognition candidate selection device for analyzing an image of an input character and performing character recognition to select a character that is a candidate for a matching character. When there is no character that matches the input character in the specified number of characters selected in, the character "creation"
The present invention relates to a character recognition candidate selection device that selects a character candidate that matches a character that is input by designating each element that forms a character, such as "Hen".
【0002】[0002]
【従来の技術】コンピュータやワークステーションなど
の情報処理装置は、文字を文字コードとして取り扱って
おり、従来これらの装置への文字の入力はキーボードか
ら行われていた。しかしながら、近年、入力作業の容易
化の要求が高まり、手書き文字を文字認識することで文
字を情報処理装置に取り込むことが盛んに行われてい
る。取り込んだ手書き文字のイメージ情報は、予め登録
してある各文字の文字パターンと比較され、整合度の高
い順に10個程度表示される。オペレータは、これら整
合度の高い文字の中に読み込んだ文字と一致する文字が
存在するときは、その文字を指定して文字の確定を行う
ようになっている。2. Description of the Related Art Information processing devices such as computers and workstations handle characters as character codes, and conventionally, characters have been input to these devices from a keyboard. However, in recent years, there has been an increasing demand for facilitating input work, and characters have been actively incorporated into information processing devices by recognizing handwritten characters. The captured image information of the handwritten character is compared with the character pattern of each character registered in advance, and about 10 characters are displayed in descending order of matching degree. When there is a character that matches the read character among the characters having a high degree of matching, the operator specifies the character and confirms the character.
【0003】整合度の高い順に表示された10個程度の
文字の中に、正解の文字が含まれない場合には、他の多
数の候補文字の中から正解文字を選択する必要があり、
その作業を容易化する種々の提案が成されている。この
ように最初に選ばれた文字の中に正解文字が無いときに
他の多数の候補文字の中から正解文字を選択するための
装置を文字認識候補選択装置と呼ぶことにする。When the correct answer character is not included in about 10 characters displayed in the order of high matching degree, it is necessary to select the correct answer character from many other candidate characters.
Various proposals have been made to facilitate the work. A device for selecting a correct answer character from a large number of other candidate characters when there is no correct answer character among the initially selected characters is called a character recognition candidate selecting device.
【0004】特開昭58−189777号公報には、正
解文字に関する各種情報を入力することによって、読み
込んだ文字と一致する候補文字を選択する文字認識候補
選択装置が開示されている。この装置では、文字の音読
み、訓読み、部首、画数、字種など正解文字のヒント情
報をオペレータが入力することで、候補文字を選択する
ようになっている。この装置では、部首、へん、つくり
など文字を構成する要素(以下、これらを文字パーツと
いう。)のそれぞれに対応付けられた専用キーが設けら
れている。オペレータは、正解文字の持つ部分パターン
に対応する専用キーを押下することで、文字パーツの指
定を行うようになっている。Japanese Unexamined Patent Publication No. 58-189777 discloses a character recognition candidate selection device that selects a candidate character that matches the read character by inputting various information relating to the correct character. In this device, a candidate character is selected by the operator inputting hint information of a correct answer character such as on-phonetic reading, kun-reading, radical, number of strokes, and character type. This device is provided with a dedicated key associated with each of elements (hereinafter referred to as character parts) that form a character, such as a radical, a hem, and making. The operator designates a character part by pressing a dedicated key corresponding to the partial pattern of the correct character.
【0005】また特開平4−205078号公報には、
選択された候補文字が部首ごとに分割できるときは、読
み込んだイメージを部首ごとに分割して各部首の候補を
選択し、これらを部首に持つ漢字を抽出する文字認識候
補選択装置が開示されている。Further, in Japanese Patent Laid-Open No. 205078/1992,
If the selected candidate characters can be divided into radicals, the character recognition candidate selection device that divides the read image into radicals and selects candidates for radicals and extracts Kanji with these radicals is selected. It is disclosed.
【0006】[0006]
【発明が解決しようとする課題】特開昭58−1897
77号公報に開示されている先行技術では、文字パーツ
のリストを予め作成しておき、キーボードの各キーにこ
れらを割り付けている。ここで、漢字を認識の対象とす
ると、文字パーツの数が100個程度になってしまう。
このため、これら多数のキーの中から正解文字の文字パ
ーツを選択することが、オペレータの大きな負担になる
という問題がある。Problems to be Solved by the Invention JP-A-58-1897
In the prior art disclosed in Japanese Patent Publication No. 77, a list of character parts is created in advance and assigned to each key of the keyboard. Here, if a kanji character is to be recognized, the number of character parts will be about 100.
Therefore, there is a problem that it is a heavy burden on the operator to select a character part of the correct answer character from among these many keys.
【0007】また特開平4−205078号公報の場合
では、読み込んだイメージを分割しこれを解析した結果
得られた部首の候補が多数存在する場合には、それらの
うちのいずれかを部首に持つ文字が正解文字の候補とし
て選択されてしまう。このため、部首の候補が多数存在
するときには十分な絞り込みを行うことができない。Further, in the case of Japanese Patent Laid-Open No. 4-205078, if there are many radical candidates obtained as a result of dividing the read image and analyzing the image, one of them is radical. The character with has been selected as a candidate for the correct character. Therefore, when there are many radical candidates, it is not possible to perform sufficient narrowing down.
【0008】そこで本発明の目的は、正解文字の持つ部
首、へん、つくりなどの文字パーツの指定を容易に行う
ことのできる文字認識候補選択装置を提供することにあ
る。SUMMARY OF THE INVENTION An object of the present invention is to provide a character recognition candidate selection device capable of easily designating a character part such as a radical, a hem, or a structure of a correct character.
【0009】[0009]
【課題を解決するための手段】請求項1記載の発明で
は、イメージとして入力された文字を文字認識する際の
比較対象とされる複数の文字と各文字を構成する要素と
を対応付けて記憶した認識対象文字記憶手段と、この認
識対象文字記憶手段に記憶されている文字の中からイメ
ージとして入力された文字と一致する文字の候補になる
複数の認識候補文字を抽出する認識候補文字抽出手段
と、この認識候補文字抽出手段によって抽出された認識
候補文字の中からイメージとして入力された文字との整
合度の高い所定数の文字とその文字を構成する要素とを
表示する候補文字表示手段と、この候補文字表示手段に
よって表示された中から任意の要素を指定する要素指定
手段と、この要素指定手段によって指定された要素を含
む文字を認識候補文字の中から選択して表示する要素一
致文字選択手段とを文字認識候補選択装置に具備させて
いる。According to a first aspect of the present invention, a plurality of characters to be compared when recognizing a character input as an image and elements constituting each character are stored in association with each other. Recognition target character storage means, and a plurality of recognition candidate character extraction means for extracting a plurality of recognition candidate characters that are candidates for a character matching the character input as an image from the characters stored in the recognition target character storage means. And a candidate character display means for displaying a predetermined number of characters having a high degree of matching with the characters input as an image among the recognition candidate characters extracted by the recognition candidate character extraction means and the elements forming the characters. , A candidate character for recognizing a character including an element designated by this element designation means and an element designation means for designating an arbitrary element among those displayed by this candidate character display means And it is provided in the character recognition candidate selection unit to elements matching character selection means for displaying selected from within.
【0010】すなわち請求項1記載の発明では、入力さ
れた文字のイメージを解析して整合度の高い順に選択し
た所定数の文字をその文字の要素と共に表示している。
そして、表示された要素の中から正解文字の持つ要素の
選択を行っている。整合度の高い文字の持つ要素の中か
ら正解文字の持つ要素を選択することで、全ての文字の
要素の中から選択する場合に比べて正解文字の有する文
字の要素を容易に指定することができる。That is, according to the first aspect of the invention, the image of the input character is analyzed and a predetermined number of characters selected in descending order of matching degree are displayed together with the element of the character.
Then, the element of the correct answer character is selected from the displayed elements. By selecting the element that the correct character has from the elements that have a high degree of matching, it is easier to specify the character element that the correct character has than when selecting from the elements of all the characters. it can.
【0011】請求項2記載の発明では、イメージとして
入力された文字を文字認識する際の比較対象とされる複
数の文字と各文字を構成する要素とそれぞれの文字の中
における各要素の位置とを予め対応付けて記憶した認識
対象文字記憶手段と、この認識対象文字記憶手段に記憶
されている文字の中からイメージとして入力された文字
と一致する文字の候補となる複数の認識候補文字を抽出
する認識候補文字抽出手段と、この認識候補文字抽出手
段によって抽出された認識候補文字の中からイメージと
して入力された文字との整合度の高い所定数の文字とそ
の文字を構成する要素とを表示する候補文字表示手段
と、この候補文字表示手段によって表示された中の任意
の要素をその文字内の位置を指示することによって指定
する要素指定手段と、この要素指定手段によって指定さ
れた要素を含む文字を認識候補文字の中から選択して表
示する要素一致文字選択手段とを文字認識候補選択装置
に具備させている。According to the second aspect of the present invention, a plurality of characters to be compared when recognizing a character input as an image, an element forming each character, and a position of each element in each character are described. And a plurality of recognition candidate characters that are candidates for a character that matches the character input as an image from the characters stored in the recognition target character storage unit. And a predetermined number of characters having a high degree of matching with the characters input as an image among the recognition candidate characters extracted by the recognition candidate character extraction unit and the elements forming the characters. A candidate character display means, and an element designating means for designating an arbitrary element displayed by the candidate character display means by designating a position within the character. And it is provided in the character recognition candidate selection unit to elements matching character selection means for selecting and displaying characters from the recognition candidate characters including the specified element by the element designation means.
【0012】すなわち請求項2記載の発明では、入力さ
れた文字のイメージ情報を解析して整合度の高い順に所
定数の文字をその文字の要素と共に表示している。そし
て、文字とその文字の要素とそれら要素の各文字の中で
の位置を予め記憶しておき、表示された中の任意の要素
をその文字内の位置を指示することで指定している。た
とえば、ポインティングデバイスにより表示画面上で文
字の要素を指定すれば、表示された要素の中から必要な
ものを容易かつ迅速に選択することができる。That is, according to the second aspect of the present invention, the image information of the input characters is analyzed and a predetermined number of characters are displayed together with the elements of the characters in descending order of matching. Then, the character, the element of the character, and the position of the element in each character are stored in advance, and any element in the displayed characters is designated by designating the position in the character. For example, if a character element is designated on the display screen with a pointing device, a necessary element can be easily and quickly selected from the displayed elements.
【0013】請求項3記載の発明では、要素一致文字選
択手段は、要素指定手段によって指定された要素を含む
文字を認識対象文字記憶手段に記憶されている文字の中
から抽出する指定要素文字抽出手段と、この指定要素文
字抽出手段によって抽出された文字と認識候補文字とに
重複する文字を選択して表示する重複文字選択手段とを
具備させている。According to the third aspect of the present invention, the element matching character selecting means extracts the designated element character from the characters stored in the recognition target character storing means, the character including the element designated by the element designating means. And a duplicated character selection unit for selecting and displaying a character overlapping with the character extracted by the designated element character extraction unit and the recognition candidate character.
【0014】すなわち請求項3記載の発明では、認識対
象文字記憶手段に記憶されている文字の中から要素指定
手段によって指定された要素を含む文字を全ての抽出
し、これらの文字と認識候補文字とに重複する文字を選
択して表示している。That is, according to the third aspect of the invention, all the characters including the element designated by the element designating means are extracted from the characters stored in the recognition target character storage means, and these characters and the recognition candidate characters are extracted. Characters that overlap with and are selected and displayed.
【0015】請求項4記載の発明では、候補文字表示手
段は、認識対象文字記憶手段に記憶されている各要素の
位置情報を基にして、各文字をその要素ごとに識別可能
に表示するようになっている。According to the fourth aspect of the present invention, the candidate character display means displays each character in a distinguishable manner for each element based on the position information of each element stored in the recognition target character storage means. It has become.
【0016】すなわち請求項4記載の発明では、予め記
憶されている文字の各要素の位置情報を基にして、認識
候補文字をその要素ごとに識別できるように表示してい
る。たとえば、各要素を個別に枠で囲んだり、要素ごと
に表示色を異ならせることにより、オペレータにどの部
分が要素になってるかを認識させることができる。これ
により、文字の要素の指定をより適切かつ容易に行うこ
とができる。That is, in the invention according to claim 4, the recognition candidate character is displayed so that it can be identified for each element based on the position information of each element of the character stored in advance. For example, by enclosing each element in a frame individually or by changing the display color for each element, it is possible to make the operator recognize which part is the element. This makes it possible to more appropriately and easily specify the character element.
【0017】請求項5記載の発明では、要素指定手段
は、候補文字表示手段の表示上で位置指定を行うポイン
ティング手段を備え、このポインティング手段によって
指定された位置の最も近くに表示されている要素を指定
された要素であると判別している。According to a fifth aspect of the invention, the element designating means includes pointing means for designating a position on the display of the candidate character display means, and the element displayed closest to the position designated by the pointing means. Is determined to be the specified element.
【0018】すなわち請求項5記載の発明では、認識候
補文字を表示するとともに、ポインティングデバイスで
表示上の位置を指定することで、要素の指定を行ってい
る。That is, according to the invention of claim 5, the element is designated by displaying the recognition candidate character and designating the position on the display by the pointing device.
【0019】[0019]
【実施例】以下実施例につき本発明を詳細に説明する。DESCRIPTION OF THE PREFERRED EMBODIMENTS The present invention will be described in detail below with reference to embodiments.
【0020】図1は、本発明の一実施例における文字認
識候補選択装置の構成の概要を表わしたものである。文
字認識候補選択装置は、イメージ情報として手書き文字
を入力しその解析を行い文字認識を行う文字認識部11
を備えている。文字認識部11は、入力された手書きの
文字のイメージと予め記憶してある各文字の文字パター
ンとの整合度を求め、その整合度の高い順に順位付けを
行う回路である。候補順位記憶部12は、文字認識部1
1で求めた整合度の高い順に2000個の文字を記憶す
る部分である。ここでは2000個の文字コードで記憶
している。認識候補選択表示部13は、整合度の高い順
に上から5番目までの5個つの文字の表示を行う回路部
分である。FIG. 1 shows an outline of the configuration of a character recognition candidate selection device according to an embodiment of the present invention. The character recognition candidate selection device inputs a handwritten character as image information, analyzes the handwritten character, and performs character recognition to perform character recognition.
It has. The character recognition unit 11 is a circuit that obtains the degree of matching between the image of an input handwritten character and the character pattern of each character stored in advance, and ranks the matching order in descending order. The candidate rank storage unit 12 includes the character recognition unit 1
This is a portion for storing 2000 characters in order of highest matching degree obtained in 1. Here, 2000 character codes are stored. The recognition candidate selection display unit 13 is a circuit unit that displays five characters from the top to the fifth in descending order of matching degree.
【0021】文字パーツデータベース部14は、各文字
コードとその文字コードの文字を構成する文字パーツと
を予め対応つけて記憶している。文字パーツ選択表示部
15は、文字パーツデータベース部14を参照して認識
候補選択表示部13に表示された5個の文字について、
それらの文字の文字パーツを表示する部分である。選択
部16は、認識候補選択表示部13に表示された文字の
中に、読み込んだ文字と一致する正解文字が存在するか
どうかオペレータが判別した結果を入力する入力装置で
ある。また選択部16は、文字パーツ選択表示部15に
表示された文字の要素の中から正解文字に含まれている
文字パーツを選択するための指示を入力する機能も備え
ている。The character parts database section 14 stores each character code and the character parts forming the characters of the character code in association with each other in advance. The character parts selection display unit 15 refers to the character parts database unit 14 and, for the five characters displayed in the recognition candidate selection display unit 13,
This is the part that displays the character parts of those characters. The selection unit 16 is an input device that inputs the result of the operator's determination as to whether or not the correct character that matches the read character exists in the characters displayed on the recognition candidate selection display unit 13. The selection unit 16 also has a function of inputting an instruction for selecting a character part included in the correct character from the character elements displayed on the character part selection display unit 15.
【0022】たとえば、表示された文字パーツの中で、
正解文字に含まれている文字パーツの部分をマウスで指
示することにより、文字パーツの指定が行われる。パー
ツ関連文字選択部17は、選択部16によって文字パー
ツが選択されたとき、候補順位記憶部12に記憶されて
いる2000個の文字の中からその文字パーツを要素に
含む文字を選択する部分である。For example, in the displayed character parts,
The character part is specified by pointing the part of the character part included in the correct character with the mouse. The part-related character selection unit 17 is a unit that, when the character part is selected by the selection unit 16, selects a character including the character part as an element from the 2000 characters stored in the candidate rank storage unit 12. is there.
【0023】文字認識候補選択装置は、各種処理の中枢
的な機能を果たすCPU(中央処理装置)と、CPUの
実行するプログラムや文字パーツデータなど各種固定的
データを記憶するROM(リード・オンリ・メモリ)
と、作業領域としてのRAM(ランダム・アクセス・メ
モリ)を備えている。また選択された文字や各文字パー
ツを表示するためのディスプレイと、選択部16として
のキーボードおよびマウスを具備している。The character recognition candidate selection device includes a CPU (central processing unit) that performs the central functions of various processes, and a ROM (read-only memory) that stores various fixed data such as programs executed by the CPU and character part data. memory)
And a RAM (random access memory) as a work area. Further, it is provided with a display for displaying the selected character and each character part, and a keyboard and a mouse as the selection unit 16.
【0024】図2は、文字パーツデータベース部の記憶
内容の一例を表わしたものである。最左欄は、記憶して
いる各文字の文字コード21を記憶している。ここで
は、説明の便宜上、文字コードの代わりにその文字の文
字パターンを示してある。これら文字コードの右側には
その文字に含まれている文字パーツ22が登録されてい
る。たとえば、文字「問」23は、文字パーツ「門」2
4と「口」25の2つの文字パーツを備えていることが
登録されている。同様に「聞」の文字は「門」と「耳」
の文字パーツから構成されていることが記憶されてい
る。文字パーツデータベース部14は、検索対象とされ
る全ての文字について、その文字コードと文字パーツと
を対応付けてこのような表形式により記憶している。FIG. 2 shows an example of the stored contents of the character parts database section. The leftmost column stores the character code 21 of each stored character. Here, for convenience of description, a character pattern of the character is shown instead of the character code. The character part 22 included in the character is registered on the right side of these character codes. For example, the character "Q" 23 is the character part "Gate" 2
It is registered that it has two character parts of 4 and "mouth" 25. Similarly, the letters "mon" are "gate" and "ear".
It is stored that it is composed of the character parts of. The character parts database unit 14 stores the character codes and the character parts of all the characters to be searched in association with each other in such a table format.
【0025】次に、図1に示した認識文字候補選択装置
の動作について説明する。Next, the operation of the recognized character candidate selection device shown in FIG. 1 will be described.
【0026】文字認識部11に入力された文字のイメー
ジ情報と各文字との整合度を求める。たとえば、イメー
ジとして入力された文字パターンが縦横32画素×32
画素の濃淡値パターンで構成されているものとする。こ
のパターンの所定の隅を原点としたとき、位置座標
(i,j)で表わされる画素の濃度をp(i,j)と表
わす。ただしi、jはそれぞれ1〜32までの正整数と
する。p(i,j)で表わされる各画素の濃度は0〜1
の間の任意の値をとるように正規化されている。The degree of matching between the character image information input to the character recognition unit 11 and each character is calculated. For example, if the character pattern input as an image is 32 pixels in length and width × 32
It is assumed that the pixel is composed of a gray value pattern. When the predetermined corner of this pattern is the origin, the density of the pixel represented by the position coordinates (i, j) is represented by p (i, j). However, i and j are positive integers of 1 to 32, respectively. The density of each pixel represented by p (i, j) is 0 to 1
It has been normalized to take any value between.
【0027】文字認識部11は、認識対象となる文字コ
ードとその代表パターンを記憶している。文字コードを
“Z”としたとき、この文字コードの代表パターンにお
ける位置座標(i,j)における画素の濃度をqz
(i,j)と表わすものとする。入力された文字パター
ンと文字コード“Z”の代表パターンとの整合度をSz
と表わすとすると、Szは次式で求めることができる。The character recognition section 11 stores a character code to be recognized and its representative pattern. When the character code is “Z”, the density of the pixel at the position coordinates (i, j) in the representative pattern of this character code is qz.
Let (i, j) be represented. The degree of matching between the input character pattern and the representative pattern of the character code “Z” is Sz.
Sz can be calculated by the following equation.
【数1】 ここで、(p,qz)は、pとqzを次元数i×jのベ
クトルとした場合の内積計算を表わし、[p][qz]
は、それぞれベクトルのノルムを表わしている。(1)
式によって求めた整合度Szは、0〜1の値を取り、大
きいほど整合性がよい。[Equation 1] Here, (p, qz) represents the inner product calculation when p and qz are vectors of dimension number i × j, and [p] [qz]
Each represent the norm of the vector. (1)
The matching degree Sz obtained by the expression takes a value of 0 to 1, and the larger the matching degree, the better the matching.
【0028】文字認識部11は認識対象となる全ての文
字について整合度Szを求め、Szの大きい順に200
0個の文字コードを選択する。これら2000個の文字
を候補文字と呼ぶことにする。ここで、整合度が上位か
らk(kは1から2000までの任意の正整数)番目の
文字コードをZ(k)と表わす。文字認識部11は、こ
のようにして求めた2000個の文字コードを候補順位
記憶部12に転送する。候補順位記憶部12は、これら
2000個の文字コードをその順位kとの組み合わせて
として蓄積する。また、候補文字のうちその上位から5
番目までを、認識候補表示部13に転送する。これら5
個の候補文字を特に認識候補と呼ぶことにする。The character recognizing unit 11 obtains the matching degrees Sz for all the characters to be recognized, and the matching degree Sz is 200
Select 0 character code. These 2000 characters will be called candidate characters. Here, the k-th character code with the highest degree of matching (k is an arbitrary positive integer from 1 to 2000) is represented as Z (k). The character recognition unit 11 transfers the 2000 character codes thus obtained to the candidate rank storage unit 12. The candidate rank storage unit 12 stores these 2000 character codes as a combination with the rank k. In addition, the upper 5 of the candidate characters
The data up to the th is transferred to the recognition candidate display unit 13. These 5
The candidate characters will be specifically referred to as recognition candidates.
【0029】認識候補選択表示部13では、選択された
5個の表示対象認識候補の文字をその整合度の順に表示
する。オペレータは、これら5個の文字の中で、読み込
んだ文字と一致する正解文字が存在する場合には、その
文字をキーボード等から選択する。たとえば、表示され
た文字にそれぞれ番号を割り付けて表示し、オペレータ
はその番号をキーボードから入力することによって正解
文字の指定を行う。また正解文字が無いときは、“N”
のキーを押下する。正解文字が存在しそれが選択された
場合には、選択された文字コードが正解文字として図示
しない情報処理装置に出力される。The recognition candidate selection / display section 13 displays the five characters of the selected display target recognition candidates in the order of their matching degree. If there is a correct character that matches the read character among these five characters, the operator selects that character from the keyboard or the like. For example, a number is assigned to each of the displayed characters for display, and the operator specifies the correct character by inputting the number from the keyboard. If there is no correct character, "N"
Press the key. When the correct character exists and is selected, the selected character code is output to the information processing device (not shown) as the correct character.
【0030】表示された5個の文字のなかに正解文字が
無いときは、文字パーツ選択表示15が起動される。文
字パーツ選択表示部15は、図2に示した形式で登録さ
れている文字パーツデータデース部14から当該5個の
文字の文字パーツを検索する。文字パーツ選択表示部1
5は、これら5個の文字について文字パーツを表示す
る。When there is no correct answer character among the displayed five characters, the character parts selection display 15 is activated. The character part selection display unit 15 retrieves the character parts of the five characters from the character part data database unit 14 registered in the format shown in FIG. Character parts selection display section 1
5 displays the character parts for these 5 characters.
【0031】図3は、文字認識される手書き文字の一例
を表わしたものである。読み込まれた文字「路」は、手
書きのためその形が若干いびつになっている。FIG. 3 shows an example of handwritten characters for character recognition. The shape of the read character "Michi" is slightly distorted because it is handwritten.
【0032】図4は、文字パーツ選択表示部の表示内容
の一例を表わしたものである。この図は、図3に示した
手書き文字「路」を解析した結果の表示画面を表わして
いる。最左列の欄31は、認識された5個の認識候補の
文字パターンを表示する欄である。手書き文字「路」を
認識した結果は、「賂」、「踊」、「跨」、「跳」、
「隅」の5つである。各文字パターンの左側には、それ
ぞれの文字を構成する文字パーツを表示する欄32が設
けられている。認識候補「賂」に含まれる文字パーツ
は、「貝」と「各」である。また認識候補「踊」に含ま
れる文字パーツは「足」や「用」などであることが表わ
されている。このように文字パーツは必ずしも、部首、
へん、つくりではなく、それらの一部である場合もあ
る。FIG. 4 shows an example of the display contents of the character parts selection display section. This figure shows a display screen as a result of analysis of the handwritten character "road" shown in FIG. The column 31 in the leftmost column is a column for displaying the five recognized recognition candidate character patterns. As a result of recognizing the handwritten character "Ro", "Ren", "Dance", "Cross", "Jump",
There are five "corners". On the left side of each character pattern, there is provided a column 32 for displaying character parts forming each character. The character parts included in the recognition candidate "賂" are "shellfish" and "each". Further, it is shown that the character parts included in the recognition candidate "dance" are "feet" and "for". In this way character parts are not always radicals,
Hen, it may be a part of them rather than making.
【0033】オペレータは、このような表示を基にし
て、選択部16により表示された中の1つの文字パーツ
を選択する。パーツ関連文字選択部17は、指定された
文字パーツを含む文字を、候補順位記憶12に記憶され
ている2000個の文字集合の中から選択する。つま
り、選択された文字パーツを含む文字の集合を文字パー
ツデータベース部14の中から選択する。この文字の集
合を集合Aとする。次に、候補順位記憶部12に記憶さ
れている文字の集合を集合Bとし、集合Aと集合Bの交
わり集合を求める。これを集合Cとする。その後候補順
位記憶部12に記憶されている集合Bを交わり集合Cと
入れ換える。これにより、指定された文字パーツを含む
文字だけが候補文字として選択される。候補順位記憶手
段12には、2000個の文字の文字コードしか登録さ
れていないので、指定された文字パーツを有する文字の
集合Aは、文字パーツを記憶している文字パーツデータ
ベース部14から求めている。Based on such a display, the operator selects one of the character parts displayed by the selection unit 16. The parts-related character selection unit 17 selects a character including the specified character part from the 2000 character sets stored in the candidate rank storage 12. That is, a set of characters including the selected character part is selected from the character part database unit 14. This set of characters is set A. Next, a set of characters stored in the candidate rank storage unit 12 is set as a set B, and an intersection set of the set A and the set B is obtained. This is set C. After that, the set B stored in the candidate rank storage unit 12 is replaced with the intersecting set C. As a result, only the character including the designated character part is selected as the candidate character. Since only 2000 character codes are registered in the candidate rank storage means 12, the character set A having the designated character parts is obtained from the character part database section 14 storing the character parts. There is.
【0034】オペレータがさらに別の文字パーツを指定
した場合には、指定された文字パーツを含む文字の集合
を文字パーツデータベース14から抽出し、これを新た
な文字集合Aとする。そして、候補順位記憶部12に記
憶されている入れ換え後の文字集合と集合Aとの交わり
集合をとり、これを新たな集合Cとして求める。これを
候補順位記憶部12の記憶内容と入れ換える。この処理
を文字パーツが指示される限り繰り返すことで、指定さ
れた全ての文字パーツを有する文字の集合がもとまる。
このようにして次第に正解文字を含む集合へと絞られ
る。When the operator designates another character part, a set of characters including the designated character part is extracted from the character part database 14 and is set as a new character set A. Then, the intersection set of the character set after the replacement stored in the candidate rank storage unit 12 and the set A is taken, and this is obtained as a new set C. This is replaced with the contents stored in the candidate rank storage unit 12. By repeating this processing as long as the character parts are designated, a set of characters having all the designated character parts is obtained.
In this way, it is gradually narrowed down to a set including correct characters.
【0035】たとえば、認識候補「賂」33の文字パー
ツ「各」34が指定されたときは、文字パーツデータベ
ース部14の中で「各」を含む文字を全て選択して集合
Aとする。そして、候補順位記憶部12に記憶されてい
る文字集合Bと集合Aとの交わり集合Cを求めてこれを
集合Bと入れ換える。次に、オペレータによりへん
「足」が指定されると文字パーツデータベース部14の
中の「足」を含む文字を全て選択しこれを集合Aとす
る。集合Aと集合Bの交わり集合を求め、得られた集合
を集合Bと置き換え、候補順位記憶部12に登録する。For example, when the character parts "respective" 34 of the recognition candidate "case" 33 are specified, all the characters including "respective" are selected in the character part database unit 14 to be set A. Then, the intersection set C of the character set B and the set A stored in the candidate rank storage unit 12 is obtained, and this is replaced with the set B. Next, when the operator specifies "foot", all the characters including "foot" in the character parts database unit 14 are selected and set as set A. The intersection set of the set A and the set B is obtained, the obtained set is replaced with the set B, and registered in the candidate rank storage unit 12.
【0036】得られた集合Bを図4と同一の表示形式
で、整合度の順位が上位のものから順に表示する。この
場合には、集合Bは、「路」だけになるが、集合Bの中
に複数の文字が存在する場合には、それらの文字を表示
した状態で、正解文字の指定、あるいは別の文字パーツ
の指定を待つ。この例では、2回の指示により認識候補
文字の選択が行われたが、文字パーツを指定する回数
は、正解文字が表示されるまでの任意の回数となる。The obtained set B is displayed in the same display format as in FIG. 4 in the order of highest matching degree. In this case, the set B is only “road”. However, when there are a plurality of characters in the set B, the correct character is designated or another character is displayed in the state where those characters are displayed. Wait for parts to be specified. In this example, the recognition candidate character is selected by two instructions, but the number of times the character parts are designated is any number of times until the correct character is displayed.
【0037】変形例 Modification
【0038】次に、文字パーツが各文字のどの部分に配
置されているかを表わした位置情報を文字パーツデータ
ベース部14が記憶している場合について説明する。変
形例では、この位置情報を基に、認識候補文字を表示す
る際に、その文字のどの部分が要素になっているかをオ
ペレータが容易に認識できるようになっている。また、
位置情報を基にしてオペレータが容易に文字パーツを指
定できるようになっている。Next, a case will be described in which the character parts database section 14 stores position information indicating in which part of each character the character parts are arranged. In the modification, when displaying the recognition candidate character based on this position information, the operator can easily recognize which part of the character is the element. Also,
The operator can easily specify the character parts based on the position information.
【0039】図5は、文字パーツの位置情報を備えた文
字パーツデータベース部の記憶内容の一例を表わしたも
のである。図中、最左の欄41は、文字コードを表わし
ている。ここでは説明の便宜上、文字コードの代わりに
対応する文字パターンを表示してある。各文字コードの
右側には、その文字を構成する文字パーツ42、43
と、各文字パーツに文字内における位置情報44、45
が登録されている。たとえば、文字「問」46は、文字
パーツ「門」47と「口」48とから構成されているこ
とを表わしている。また、位置情報は、各文字の左上隅
を、2次元平面の原点(0,0)とし、右下隅を座標
(1,1)として表わしている。すなわち、この座標上
で各文字パーツの存在する領域を矩形で囲んだときその
左上隅と右下隅の2つの頂点の座標により配置される位
置を表わしている。FIG. 5 shows an example of the stored contents of the character parts database section provided with the position information of the character parts. The leftmost column 41 in the figure represents the character code. Here, for convenience of description, a corresponding character pattern is displayed instead of the character code. On the right side of each character code, the character parts 42, 43 that make up the character
And the position information 44, 45 in each character for each character part
Is registered. For example, the character “question” 46 is shown to be composed of character parts “gate” 47 and “mouth” 48. In the position information, the upper left corner of each character is the origin (0,0) of the two-dimensional plane, and the lower right corner is the coordinate (1,1). That is, when the area where each character part is present is surrounded by a rectangle on this coordinate, the position is defined by the coordinates of the two vertices of the upper left corner and the lower right corner.
【0040】たとえば、文字パーツ「門」47の位置情
報49は、(0,0)−(1,1)と表わされ、(0,
0)を左上頂点、(1,1)を右下頂点とする矩形の中
に配置されていることを示している。同様に文字パーツ
「口」48の位置情報51は(0.2,0.4)−
(0.8,1)であり、横方向に“0.2”〜“0.
8”、および縦方向に“0.4”から“1”の範囲で示
される矩形領域の中に配置されていることを表わしてい
る。この図では、1つの文字に含まれる文字パーツの数
は2個であるが、文字パーツの数は、その文字によって
異なる。そして、文字パーツデータベースに登録する場
合、その文字パーツの数を2個あるいは3個など固定値
としても良いし、文字により登録できる文字パーツの数
を可変にしてもよい。For example, the position information 49 of the character part "gate" 47 is represented as (0,0)-(1,1), and (0,0)
It is arranged in a rectangle whose upper left corner is (0) and lower right corner is (1, 1). Similarly, the position information 51 of the character part "mouth" 48 is (0.2, 0.4)-
(0.8, 1), and "0.2" to "0.
8 ", and that it is arranged in a rectangular area shown in the range of" 0.4 "to" 1 "in the vertical direction. In this figure, the number of character parts included in one character. Is 2, but the number of character parts differs depending on the character, and when registering in the character part database, the number of character parts may be a fixed value such as 2 or 3, or registered by character. The number of character parts that can be created may be variable.
【0041】実施例と同様に、読み込んだイメージとの
整合度を基にして2000個の文字が抽出され、その整
合度の順位と共に候補順位記憶部12に記憶される。そ
の内の上位5個について、文字パーツの位置情報を基に
して文字パーツ選択表示部15は各文字の文字パーツが
どの部分であるかをオペレータが容易に認識できるに表
示する。Similar to the embodiment, 2000 characters are extracted based on the matching degree with the read image and stored in the candidate ranking storage unit 12 together with the ranking of the matching degree. For the top five of them, the character part selection display unit 15 displays which part the character part of each character is based on the positional information of the character parts so that the operator can easily recognize it.
【0042】図6は、認識候補の文字を文字パーツの位
置情報を基にした表示の一例を表わしたものである。図
3に示した手書き文字「路」を入力文字としたとき、
「賂」、「踊」、「跨」、「跳」、「隅」が整合性の高
い文字として選択されたものとする。文字「賂」に対応
する文字パーツ「貝」の位置情報を基にしてこれを囲む
矩形61が、また、文字パーツ「各」の位置情報により
矩形62がそれぞれ表示されている。同様に認識候補
「踊」に対応する文字パーツ「足」などがそれぞれ矩形
63、64により囲まれて表示されている。もちろん、
図6に示した文字「問」が選択されたときには、認識候
補「問」全体を(0,0)−(1,1)とした場合に、
文字パーツ「口」は(0.2,0.4)、(0.8,
1)を対角頂点とする矩形によって囲まれることにな
る。FIG. 6 shows an example of the display of the recognition candidate characters based on the position information of the character parts. When the handwritten character “Michi” shown in FIG. 3 is used as the input character,
It is assumed that "賂", "dance", "crossover", "jump", and "corner" are selected as highly consistent characters. Based on the position information of the character part "shell" corresponding to the character "賂", a rectangle 61 surrounding it is displayed, and a rectangle 62 is displayed by the position information of the character part "each". Similarly, the character parts “legs” and the like corresponding to the recognition candidate “dance” are displayed surrounded by rectangles 63 and 64, respectively. of course,
When the character “question” shown in FIG. 6 is selected and the entire recognition candidate “question” is (0,0) − (1,1),
The character part "mouth" is (0.2, 0.4), (0.8,
It will be surrounded by a rectangle whose diagonal vertex is 1).
【0043】次に、文字パーツの指定の仕方について説
明する。Next, a method of designating character parts will be described.
【0044】まず、マウスなどのポインティングデバイ
スが無い場合には、キーボードにより正解文字に関連す
る文字パーツの選択を行う。この際、図6に示したよう
な位置情報を備えている場合には、表示された順に、そ
の文字の有する文字パーツに通し番号を割り付ける。た
とえば、1つの認識候補に2つづつの文字パーツが有る
場合には、第1の認識候補に含まれている文字パーツに
“1”、“2”の番号を与え、第2の認識候補にふくま
れている文字パーツに“3”、“4”と番号を付ける。
第n番目の文字には、“2n−1”、“2n”の番号を
割り振る。個々の認識候補が持つ文字パーツの数が異な
る場合でも同様に、それらの文字の文字パーツに通し番
号を与える。First, when there is no pointing device such as a mouse, character parts related to the correct answer character are selected by the keyboard. At this time, when the positional information as shown in FIG. 6 is provided, serial numbers are assigned to the character parts of the character in the displayed order. For example, when one recognition candidate has two character parts, the character parts included in the first recognition candidate are given numbers “1” and “2”, and the second recognition candidate is covered. Number the rare character parts with "3" and "4".
The numbers "2n-1" and "2n" are assigned to the nth character. Even when the number of character parts of each recognition candidate is different, serial numbers are similarly given to the character parts of those characters.
【0045】たとえば、図6に示した形式で、各文字パ
ーツを枠で囲んで表示するときは、まず、最初に番号
“1”が付された文字パーツの枠の色を他の枠の色と変
えて表示する。たとえば番号“1”が割り当てられてい
る枠62だけを赤色で表示し、他の枠を黒色で表示す
る。このほか枠の太さや、実線と破線のように線種を異
ならせてもよく、1つの枠だけを他の枠と区別できるよ
うに表示できればよい。オペレータが“N”のキーを入
力するたびに、割り付けられた番号順に表示色の異なる
枠が移動する。そして、リターンキーを押下したとき、
表示色の異なる枠で囲まれている文字パーツが指定され
る。このようにポインティングデバイスが無くても、文
字パーツの指定を容易に行うことができる。For example, in the case of displaying each character part by enclosing it in a frame in the format shown in FIG. 6, first, the color of the frame of the character part to which the number "1" is first attached is changed to the color of other frames. And display it. For example, only the frame 62 to which the number “1” is assigned is displayed in red, and the other frames are displayed in black. In addition, the thickness of the frame and the line type such as a solid line and a broken line may be different as long as only one frame can be displayed so as to be distinguishable from other frames. Each time the operator inputs the "N" key, the frames with different display colors move in the order of the assigned numbers. And when you press the return key,
Character parts surrounded by frames with different display colors are specified. As described above, it is possible to easily specify the character parts without using the pointing device.
【0046】その後、指定された文字パーツを含む認識
候補が、候補順位記憶部12に記憶されている文字の中
から選択される。すなわち、選択された文字パーツを含
む文字の集合を集合A、候補順位記憶部12に記憶され
ている文字の集合を集合Bとして、これら集合Aと集合
Bの交わり集合を求める。その結果の集合を集合Bと入
れ換える。別の文字パーツが指定されたときはこのよう
な処理が繰り返し行われ、次第に正解文字を含む候補に
絞られる。After that, a recognition candidate including the designated character part is selected from the characters stored in the candidate rank storage unit 12. That is, the set of characters including the selected character parts is set as set A, and the set of characters stored in the candidate rank storage unit 12 is set as set B, and the intersection set of these sets A and B is obtained. Swap the resulting set with set B. When another character part is specified, such a process is repeated, and the candidates are gradually narrowed down to include correct characters.
【0047】次に、マウスなどのポインティングデバイ
スによって文字パーツを指定する場合について説明す
る。Next, the case of designating a character part with a pointing device such as a mouse will be described.
【0048】図5に示したように、各文字パーツの配置
を文字内での正規化された座標として備えている場合に
は、まず、認識候補として表示すべき文字の各文字パー
ツの座標をその文字パーツの表示されている表示画面上
での位置座標(x1,y1)−(x2,y2)に変換す
る。このように表示画面上において各文字パーツに外接
する枠の左上隅と,右下隅の座標として表わすための変
換式は次式で表わされる。As shown in FIG. 5, when the arrangement of each character part is provided as the normalized coordinates within the character, first, the coordinates of each character part of the character to be displayed as a recognition candidate are determined. It is converted into position coordinates (x1, y1)-(x2, y2) on the display screen where the character parts are displayed. In this way, the conversion equation for expressing the coordinates of the upper left corner and the lower right corner of the frame circumscribing each character part on the display screen is expressed by the following equation.
【数2】 ここで、Sは文字の表示サイズを、x、yは表示画面上
での各文字の左上隅の座標をそれぞれ表わしている。(Equation 2) Here, S represents the display size of the character, and x and y represent the coordinates of the upper left corner of each character on the display screen.
【0049】図5に示した位置座標を(2)式の結果得
られた座標に置き換え、これを文字パーツの表示位置と
して記憶する。ここでは、表示位置(x,y)は(10
0,100×j)になっている。ここで、jは第j番目
の認識候補を表わしている。。この場合には、文字全体
を囲む4つの辺は、(x1,y1)−(x1,y2)、
(x1,y2)−(x2,y2)、(x2,y2)−
(x2,y1)、(x2,y1)−(x1,y1)で表
わされる。文字パーツデータベース部14から検索した
全ての文字パーツの個数がn個の場合に、これら文字パ
ーツを囲む枠のそれぞれについて4辺の線分の座標を求
めて記憶するものとする。このとき文字パーツの枠のデ
ータ数は、全部で4n個になる。これら文字パーツの枠
の辺の集合に通し番号を付けて、第j番目の辺の表示画
面上での座標を(Xj1,Yj1)−(Xj2,Yj
2)とする。ここでjは1〜4nまでの自然数とする。The position coordinates shown in FIG. 5 are replaced with the coordinates obtained as a result of the equation (2), and this is stored as the display position of the character part. Here, the display position (x, y) is (10
0,100 × j). Here, j represents the jth recognition candidate. . In this case, the four sides surrounding the entire character are (x1, y1)-(x1, y2),
(X1, y2)-(x2, y2), (x2, y2)-
It is represented by (x2, y1) and (x2, y1)-(x1, y1). When the number of all character parts retrieved from the character part database unit 14 is n, the coordinates of the line segments on the four sides are obtained and stored for each of the frames surrounding these character parts. At this time, the total number of data in the character part frame is 4n. A serial number is assigned to the set of sides of the frame of these character parts, and the coordinates of the j-th side on the display screen are (Xj1, Yj1)-(Xj2, Yj
2). Here, j is a natural number from 1 to 4n.
【0050】ポインティングデバイスにより指示された
画面上の座標を(u,v)とすると、指示された座標か
ら第j番目の辺までの距離djは次式により求めること
ができる。Assuming that the coordinates on the screen designated by the pointing device are (u, v), the distance dj from the designated coordinate to the j-th side can be calculated by the following equation.
【数3】 ただし、ABS(x)はxの絶対値をもとめる関数を表
わし、SQRT(x)は、xの平方根を求める関数を表
わしている。(Equation 3) However, ABS (x) represents a function for obtaining the absolute value of x, and SQRT (x) represents a function for obtaining the square root of x.
【0051】全ての辺について距離dを求め、その中で
距離が最小の辺を選択するとともに、その辺のを含む枠
に対応する文字パーツを求め、これを指定された文字パ
ーツと判別する。選択した以後の処理については、先に
説明したものと同様であり、その記載を省略する。The distance d is calculated for all sides, the side having the smallest distance is selected, and the character part corresponding to the frame including the side is calculated, and this is discriminated as the designated character part. The processing after the selection is the same as that described above, and the description thereof will be omitted.
【0052】次に、各文字パーツの近傍にそれぞれ入力
バーを表示し、この入力バーの任意の点をマウスで指示
することによって文字パーツを選択する例について説明
する。Next, an example in which an input bar is displayed near each character part and a character part is selected by pointing an arbitrary point on the input bar with a mouse will be described.
【0053】図7は、認識文字とともに入力バーを表示
した場合における表示画面の一例を表わしたものであ
る。図示のように各認識候補の周辺には入力バー71〜
74が表示される。この図では、4つの認識候補が表示
されている。1つの認識候補についてそれぞれ上下左右
に1つずつ入力バーが表示される。それぞれの入力バー
は、その左上隅の座標と右下隅の座標によって記述され
る。手書き文字「路」75の文字認識を行った結果とし
て4つの認識候補が選択された場合、第j番目の文字に
おける第i番目の入力バーをBjiと表わすことにす
る。その左上隅の座標を(Xji1,Yji1)と、右
下隅の座標を(Xji2,Yji2)と表わすものとす
る。さらに入力バーの代表点として(Xjic,Yji
c)を定義する。代表点の座標は、((Xji1+Xj
i2)/2,(Yji1+Yji2)/2)として定め
ている。代表点の定め方は他の方法でも良いことは言う
までもない。FIG. 7 shows an example of a display screen when the input bar is displayed together with the recognized characters. As shown in the figure, input bars 71 to 71 are provided around each recognition candidate.
74 is displayed. In this figure, four recognition candidates are displayed. One input bar is displayed on each of the upper, lower, left and right sides of one recognition candidate. Each input bar is described by the coordinates of its upper left corner and its lower right corner. When four recognition candidates are selected as a result of character recognition of the handwritten character "Michi" 75, the i-th input bar in the j-th character is represented as Bji. The coordinates of the upper left corner are represented by (Xji1, Yji1), and the coordinates of the lower right corner are represented by (Xji2, Yji2). Further, as a representative point of the input bar (Xjic, Yji
Define c). The coordinates of the representative point are ((Xji1 + Xj
i2) / 2, (Yji1 + Yji2) / 2). It goes without saying that the representative point may be determined by other methods.
【0054】パーツ関連文字選択部17では、全ての文
字候補の有する全ての文字パーツの位置を(2)式によ
り求めて、個々の文字パーツの代表点としてその中心位
置を記憶してある。文字パーツの位置が(x1,y1)
−(x2,y2)の場合に、中心座標は((x1+x
2)/2,(y1+y2)/2)で求める。In the parts-related character selection section 17, the positions of all the character parts possessed by all the character candidates are obtained by the equation (2), and the central position thereof is stored as the representative point of each character part. The position of the character part is (x1, y1)
In the case of − (x2, y2), the center coordinates are ((x1 + x
2) / 2, (y1 + y2) / 2).
【0055】文字パーツの選択には、対象とする文字パ
ーツに最も近い入力バーをマウスなどのポインティング
デバイスでクリックすることによって選択する。マウス
により指定された画面上の座標を(x,y)とする。表
示されている全ての入力バーについて、入力された座標
(x,y)をその内部に含んでいるかどうかを判定す
る。以下の4つの条件式が全て満足されているとき、第
j候補文字の第i番目の入力バーに座標(x,y)が含
まれると判定する。To select a character part, the input bar closest to the target character part is clicked with a pointing device such as a mouse. The coordinates on the screen designated by the mouse are (x, y). For all displayed input bars, it is determined whether or not the input coordinates (x, y) are included therein. When all of the following four conditional expressions are satisfied, it is determined that the i-th input bar of the j-th candidate character includes the coordinates (x, y).
【数4】 (Equation 4)
【0056】この条件を満足する入力バーが存在しない
場合には、入力バーの指示が行われなかったものとす
る。条件式を満足する入力バーを検出した場合には、そ
の入力バーに最も近い文字文字パーツを探索する。全て
の認識候補の全ての文字パーツに通し番号を与えた場合
における最大番号をnとする。その場合、第k番目の文
字パーツの代表点を(xk,yk)とする。選択した入
力バーの代表点(Xjic,Yjic)と第k番目の文
字パーツの代表点との距離を求める。ここではユーグリ
ッドを用いたがこれに限られない。全ての文字パーツに
ついて距離を求め、最小距離を与える文字パーツを選択
する。これをオペレータにより指定された文字パーツと
して認識する。文字パーツを選択した後の処理について
先に説明したものとは同様であるのでその記載を省略す
る。If there is no input bar satisfying this condition, it is assumed that the input bar is not instructed. When an input bar that satisfies the conditional expression is detected, the character part closest to the input bar is searched. When the serial numbers are given to all character parts of all recognition candidates, the maximum number is n. In this case, the representative point of the kth character part is (xk, yk). The distance between the representative point (Xjic, Yjic) of the selected input bar and the representative point of the kth character part is calculated. Although Eugrid is used here, it is not limited to this. Find the distances for all the character parts and select the character part that gives the minimum distance. This is recognized as a character part designated by the operator. Since the processing after selecting the character parts is the same as that described above, the description thereof will be omitted.
【0057】以上説明した実施例および変形例では、
(1)式によって整合度を求めたが、各文字ごとに整合
度を与えるものであればこれに限られるものではない。
また整合度が0〜1の範囲に収まらない場合であって
も、これを線型変換によって0〜1に写像すればよい。
また、整合度を基にして文字認識部で抽出する文字の数
を2000個とし、認識候補をその上位から5個とした
が、これらの数はこの値に限定されるものではない。In the embodiment and the modification described above,
Although the degree of matching is calculated by the equation (1), the degree of matching is not limited to this as long as the degree of matching is given to each character.
Even if the degree of matching does not fall within the range of 0 to 1, it may be mapped to 0 to 1 by linear conversion.
Further, the number of characters extracted by the character recognition unit based on the degree of matching is 2000 and the recognition candidates are 5 from the top, but these numbers are not limited to this value.
【0058】また文字パーツデータベース部で記憶する
各文字の文字パーツの数は任意でよい。このほかキーボ
ードによって文字パーツを選択する場合に、“n”キー
を押下するごとに表示色の異なる枠が順次変更されるも
のとしたが、押下するキーは“n”キー以外であっても
よい。さらに矢印キーにより枠の位置が変更されるよう
にしてもよい。The number of character parts of each character stored in the character part database section may be arbitrary. In addition, when character parts are selected with the keyboard, the frames with different display colors are sequentially changed each time the "n" key is pressed, but the pressed key may be other than the "n" key. . Further, the position of the frame may be changed by using the arrow keys.
【0059】また距離の求め方は(3)式に限られな
い。さらに、図7では入力バーを4つ表示するようにし
たが、入力バーの表示は各文字パーツを個別に指定可能
なものであれば任意でよい。The method of obtaining the distance is not limited to the equation (3). Further, although four input bars are displayed in FIG. 7, the input bar may be displayed as long as each character part can be designated individually.
【0060】[0060]
【発明の効果】以上説明したように請求項1記載の発明
によれば、整合度の高い文字の持つ要素を表示しその中
から正解文字の持つ要素を選択したので、全ての文字の
要素の中から選択する場合に比べて容易に文字の要素を
指定することができる。これにより文字の要素を指示す
る際のオペレータの負担を軽減することができる。As described above, according to the first aspect of the present invention, since the elements of the character having a high degree of matching are displayed and the element of the correct character is selected from among them, the elements of all the characters are selected. Character elements can be specified more easily than when selecting from the inside. As a result, it is possible to reduce the burden on the operator when designating a character element.
【0061】また請求項2記載の発明によれば、文字と
その文字の要素とそれら要素の各文字の中での位置を予
め記憶しておき、表示された整合度の高い文字の要素を
その文字内の位置により指定している。たとえば、ポイ
ンティングデバイスにより表示画面上での位置指定によ
り文字の要素を指定すれば、要素の選択におけるオペレ
ータの負担をより一層軽減することができる。According to the second aspect of the invention, the characters, the elements of the characters and the positions of those elements in each character are stored in advance, and the displayed elements of the character having a high degree of matching are stored. It is specified by the position within the character. For example, if a character element is specified by specifying the position on the display screen with a pointing device, the operator's burden in selecting the element can be further reduced.
【0062】さらに請求項3記載の発明によれば、認識
対象文字記憶手段に記憶されている文字の中から要素指
定手段によって指定された要素を含む文字を全ての抽出
し、これらの文字と認識候補文字とに重複する文字を選
択して表示している。これにより認識候補文字に対応付
けて文字の要素を別途記憶する必要がなく、記憶領域を
少なくすることができる。Further, according to the third aspect of the invention, all the characters including the element designated by the element designating means are extracted from the characters stored in the recognition target character storage means and recognized as these characters. Characters that overlap with the candidate characters are selected and displayed. As a result, it is not necessary to separately store the character element in association with the recognition candidate character, and the storage area can be reduced.
【0063】また請求項4記載の発明によれば、予め記
憶されている文字の各要素の位置情報を基にして、認識
候補文字をその要素ごとに識別できるように表示してた
ので、文字の要素の指定をより適切かつ容易に行うこと
ができる。Further, according to the invention of claim 4, the recognition candidate character is displayed so that it can be identified for each element based on the position information of each element of the character stored in advance. The element of can be specified more appropriately and easily.
【0064】さらに請求項5記載の発明によれば、認識
候補の文字を表示するとともに、ポインティングデバイ
スにより正解文字の要素の指定を行っったので、文字の
要素の指定を容易に行うことができ、オペレータの負担
を軽減することができる。Further, according to the invention of claim 5, the character of the recognition candidate is displayed and the element of the correct character is specified by the pointing device, so that the element of the character can be easily specified. It is possible to reduce the burden on the operator.
【図1】本発明の一実施例における文字認識候補選択装
置の構成の概要を表わしたブロック図である。FIG. 1 is a block diagram showing an outline of a configuration of a character recognition candidate selection device in an embodiment of the present invention.
【図2】文字パーツデータベース部の記憶内容の一例を
表わした説明図である。FIG. 2 is an explanatory diagram showing an example of stored contents of a character parts database unit.
【図3】手書き文字の一例を表わしたものである。FIG. 3 illustrates an example of handwritten characters.
【図4】文字パーツ選択表示部の表示内容の一例を表わ
した説明図である。FIG. 4 is an explanatory diagram showing an example of display contents of a character parts selection display unit.
【図5】文字パーツの位置情報を備えた文字パーツデー
タベース部の記憶内容の一例を表わした説明図である。FIG. 5 is an explanatory diagram showing an example of stored contents of a character parts database unit including position information of character parts.
【図6】認識候補の文字を文字パーツの位置情報を基に
した表示の一例を表わした説明図である。FIG. 6 is an explanatory diagram showing an example of display of recognition candidate characters based on position information of character parts.
【図7】認識文字とともに入力バーを表示した場合にお
ける表示画面の一例を表わした説明図である。FIG. 7 is an explanatory diagram showing an example of a display screen when an input bar is displayed together with a recognized character.
11 文字認識部 12 候補順位記憶部 13 認識候補選択表示部 14 文字パーツデータベース部 15 文字パーツ選択表示部 16 選択部 17 パーツ関連文字選択部 11 Character Recognition Section 12 Candidate Order Storage Section 13 Recognition Candidate Selection Display Section 14 Character Parts Database Section 15 Character Parts Selection Display Section 16 Selection Section 17 Parts-Related Character Selection Section
Claims (5)
識する際の比較対象とされる複数の文字と各文字を構成
する要素とを対応付けて記憶した認識対象文字記憶手段
と、 この認識対象文字記憶手段に記憶されている文字の中か
らイメージとして入力された文字と一致する文字の候補
になる複数の認識候補文字を抽出する認識候補文字抽出
手段と、 この認識候補文字抽出手段によって抽出された認識候補
文字の中からイメージとして入力された文字との整合度
の高い所定数の文字とその文字を構成する要素とを表示
する候補文字表示手段と、 この候補文字表示手段によって表示された中から任意の
要素を指定する要素指定手段と、 この要素指定手段によって指定された要素を含む文字を
前記認識候補文字の中から選択して表示する要素一致文
字選択手段とを具備することを特徴とする文字認識候補
選択装置。1. A recognition target character storage unit that stores a plurality of characters to be compared when recognizing a character input as an image and an element forming each character, and a recognition target character storage unit. A recognition candidate character extraction unit that extracts a plurality of recognition candidate characters that are candidates for a character that matches a character input as an image from the characters stored in the storage unit, and the recognition candidate character extraction unit extracts the recognition candidate character. A candidate character display means for displaying a predetermined number of characters having a high degree of matching with the character input as an image from among the recognition candidate characters and the elements constituting the character, and among the candidates displayed by the candidate character display means. Element designating means for designating an arbitrary element and element matching for selecting and displaying a character including the element designated by the element designating means from the recognition candidate characters Character recognition candidate selection apparatus characterized by comprising a shaped selection means.
識する際の比較対象とされる複数の文字と各文字を構成
する要素とそれぞれの文字の中における各要素の位置と
を予め対応付けて記憶した認識対象文字記憶手段と、 この認識対象文字記憶手段に記憶されている文字の中か
らイメージとして入力された文字と一致する文字の候補
となる複数の認識候補文字を抽出する認識候補文字抽出
手段と、 この認識候補文字抽出手段によって抽出された認識候補
文字の中からイメージとして入力された文字との整合度
の高い所定数の文字とその文字を構成する要素とを表示
する候補文字表示手段と、 この候補文字表示手段によって表示された中の任意の要
素をその文字内の位置を指示することによって指定する
要素指定手段と、 この要素指定手段によって指定された要素を含む文字を
前記認識候補文字の中から選択して表示する要素一致文
字選択手段とを具備することを特徴とする文字認識候補
選択装置。2. A plurality of characters to be compared when recognizing a character input as an image, an element forming each character, and a position of each element in each character are stored in association with each other in advance. Recognition target character storage means, and a recognition candidate character extraction means for extracting a plurality of recognition candidate characters that are candidates for a character matching the character input as an image from the characters stored in the recognition target character storage means And a candidate character display means for displaying a predetermined number of characters having a high degree of matching with the characters input as an image among the recognition candidate characters extracted by the recognition candidate character extraction means and the elements forming the characters. , An element designating means for designating an arbitrary element displayed by this candidate character display means by designating a position in the character, and an element designating means Character recognition candidate selection device characterized by comprising the elements matched character selection means for displaying selected from among the recognized candidate characters characters containing the specified element I.
指定手段によって指定された要素を含む文字を前記認識
対象文字記憶手段に記憶されている文字の中から抽出す
る指定要素文字抽出手段と、この指定要素文字抽出手段
によって抽出された文字と前記認識候補文字とに重複す
る文字を選択して表示する重複文字選択手段とを具備す
ることを特徴とする請求項1または請求項2記載の文字
認識候補選択装置。3. The element matching character selecting means extracts a character including an element designated by the element designating means from a character stored in the recognition target character storage means, and a designated element character extracting means. 3. The character according to claim 1 or 2, further comprising: duplicated character selecting means for selecting and displaying a character overlapping with the character extracted by the designated element character extracting means and the recognition candidate character. Recognition candidate selection device.
文字記憶手段に記憶されている各要素の位置情報を基に
して、各文字をその要素ごとに識別可能に表示すること
を特徴とする請求項2記載の文字認識候補選択装置。4. The candidate character display means displays each character in a distinguishable manner for each element based on the position information of each element stored in the recognition target character storage means. The character recognition candidate selection device according to claim 2.
手段の表示上で位置指定を行うポインティング手段を備
え、このポインティング手段によって指定された位置の
最も近くに表示されている要素を指定された要素である
と判別することを特徴とする請求項2記載の文字認識候
補選択装置。5. The element designating means comprises pointing means for designating a position on the display of the candidate character display means, and the element displayed closest to the position designated by the pointing means is designated. The character recognition candidate selection device according to claim 2, wherein the character recognition candidate selection device determines that the character recognition candidate is an element.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP7163172A JPH0916721A (en) | 1995-06-29 | 1995-06-29 | Character recognition candidate selector |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP7163172A JPH0916721A (en) | 1995-06-29 | 1995-06-29 | Character recognition candidate selector |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| JPH0916721A true JPH0916721A (en) | 1997-01-17 |
Family
ID=15768616
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP7163172A Pending JPH0916721A (en) | 1995-06-29 | 1995-06-29 | Character recognition candidate selector |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPH0916721A (en) |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2011128688A (en) * | 2009-12-15 | 2011-06-30 | Fujitsu Ltd | Character identification device and character identification method |
Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS60142478A (en) * | 1983-12-28 | 1985-07-27 | Fujitsu Ltd | Character correction method in character recognition device |
| JPH0696266A (en) * | 1992-09-11 | 1994-04-08 | Hitachi Ltd | Correction support method for character recognition results |
| JPH06195324A (en) * | 1992-12-24 | 1994-07-15 | Canon Inc | Character input method and device |
-
1995
- 1995-06-29 JP JP7163172A patent/JPH0916721A/en active Pending
Patent Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS60142478A (en) * | 1983-12-28 | 1985-07-27 | Fujitsu Ltd | Character correction method in character recognition device |
| JPH0696266A (en) * | 1992-09-11 | 1994-04-08 | Hitachi Ltd | Correction support method for character recognition results |
| JPH06195324A (en) * | 1992-12-24 | 1994-07-15 | Canon Inc | Character input method and device |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2011128688A (en) * | 2009-12-15 | 2011-06-30 | Fujitsu Ltd | Character identification device and character identification method |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US5852448A (en) | Stroke-based font generation independent of resolution | |
| US6501475B1 (en) | Glyph-based outline font generation independent of resolution | |
| US6661417B1 (en) | System and method for converting an outline font into a glyph-based font | |
| CN115812221B (en) | Image generation and coloring method and apparatus | |
| US11783625B2 (en) | Method for verifying the identity of a user by identifying an object within an image that has a biometric characteristic of the user and separating a portion of the image comprising the biometric characteristic from other portions of the image | |
| CN114359038A (en) | Multi-style dynamic word forming method based on generation of confrontation network | |
| JP3319203B2 (en) | Document filing method and apparatus | |
| JPH03119387A (en) | Method and apparatus for forming contour of digital type surface | |
| JP6533395B2 (en) | Character search method and system | |
| US6483943B1 (en) | Feature value extraction methods and apparatus for image recognition and storage medium for storing image analysis program | |
| JP2024117033A (en) | How to assist with data entry in forms | |
| JP3448606B2 (en) | Method and apparatus for generating stroke-based characters in full resolution space | |
| JP4164976B2 (en) | Character recognition device | |
| CN102096814A (en) | Font element determining device and font element determining method | |
| JP2853167B2 (en) | Dictionary creation device for pattern recognition | |
| JP5176390B2 (en) | Character input device and computer program | |
| JP2853169B2 (en) | Pattern recognition device | |
| JPH01290090A (en) | Dictionary production method | |
| JP2937607B2 (en) | Layout creation device | |
| JP2002236877A (en) | Character string recognizing method, character recognizing device and program | |
| JPH1021325A (en) | Method for recognizing character | |
| JPS5875279A (en) | Character classification system | |
| JPH0362186A (en) | Generation method for online handwriting character recognition dictionary | |
| JPH02176973A (en) | Drawing read processing method | |
| JPH09147125A (en) | Contour line extraction method and extraction device |