JPH02220186A - Character reader - Google Patents

Character reader

Info

Publication number
JPH02220186A
JPH02220186A JP1044581A JP4458189A JPH02220186A JP H02220186 A JPH02220186 A JP H02220186A JP 1044581 A JP1044581 A JP 1044581A JP 4458189 A JP4458189 A JP 4458189A JP H02220186 A JPH02220186 A JP H02220186A
Authority
JP
Japan
Prior art keywords
character
characters
unreadable
similar
memory
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
JP1044581A
Other languages
Japanese (ja)
Inventor
Yuji Shinozaki
祐司 篠崎
Toshifumi Yamauchi
山内 俊史
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
NEC Corp
Original Assignee
NEC Corp
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by NEC Corp filed Critical NEC Corp
Priority to JP1044581A priority Critical patent/JPH02220186A/en
Publication of JPH02220186A publication Critical patent/JPH02220186A/en
Pending legal-status Critical Current

Links

Landscapes

  • Character Discrimination (AREA)

Abstract

PURPOSE:To reduce the number of times of correction by means of an operator when similarly deformed characters cannot be read for many times by correcting the similar illegible characters in a lump. CONSTITUTION:A correcting part 31 displays a character image described in a character frame 2 as the character image of the illegible character on a CRT 61, inputs a corrected character code '8' from an operator, compares characteristics data of the illegible characters stored and displayed on a memory 42 with the characteristic data of the characters in character frames 6, 9 and 11 as the other illegible characters in the same slip, decides whether or not they are the similar characters, and obtains a result that the all characters described in the character frames 6, 9 and 11 are similar to the character described in the character frame 2. The corrected part 31 collectively replaces the read result of the characters described in the character frames 2, 6, 9 and 11 stored into the memory 42, obtains the correction result, outputs the correction result to a floppy disk driver 81, and then outputs it to a floppy disk. Thus when many similar deformed characters are generated, the number of times of the correction by the operator can be reduced.

Description

【発明の詳細な説明】[Detailed description of the invention] 【産業上の利用分野〕[Industrial application field]

本発明は文字読取装置に関し、特に不読文字の修正方法
に関する。 〔従来の技術〕 従来の文字読取装置において、不読文字の修正を行なう
場合は、操作者が全ての不読文字に対し1文字ずつ文字
コードを入力する必要があった。 [発明が解決しようとする課題] 上述した従来の文字読取装置は、不読文字の修正を行な
う場合、操作者が全ての不読文字に対し1文字ずつ文字
コードを入力する必要があったため、類似した変形文字
の不読が多数発生した場合にそれに伴ない操作者による
修正回数も増加するという欠点がある。 例えば第2図の帳票の読み取りを行ない表1の判定結果
となった場合、操作者が文字枠■、■。 ■、■に記入された類似した字形不読文字4文字に対し
4回の同様な修正操作が必要であった。 表1 【課題を解決するための手段】 本発明の文字読取装置は、イメージメモリと、帳票を1
枚毎に走査し前記イメージメモリにイメージデータな格
納する手段と、認識結果格納メモリと、前記イメージメ
モリに格納されているイメージデータを読出し、1文字
ずつ切出し認識し、文字認識結果として、認識文字コー
ドまたは予め定義されている不読文字コードと認識に付
随し得られる不読文字の特徴データを認識結果格納メモ
リに格納する手段と、操作者が不読となった文字に対す
るための修正文字コードを入力する修正文字コード入力
手段と、認識結果中の不読文字コードを前記修正文字コ
ード入力手段により入力された文字コードに置換する修
正手段と、前記認識結果メモリを参照し前記修正手段に
より修正された不読文字と類似した特徴を持つ類似不読
文字を検索する手段と、前記類似不読文字の不読文字コ
ードを前記指定文字コードに一括置換する手段と、前記
一括置換手段により置換された認識結果を出力す盃手段
とを有している。 〔作 用〕 類似不読文字を一括して修正するので、類似した変形文
字の不読が多数発生した場合に操作者による修正の回数
が減少する。 〔実施例] 次に、本発明の実施例について図面を参照して説明する
。 第1図は本発明の一実施例である文字読取装置の概略構
成図である。 イメージ入力部11は帳票のイメージデータを帳票1枚
毎に操作しイメージデータなバス51を介しメモリ41
に格納する。認識部21はバス51を介しメモリ41に
格納されたイメージデータを読出し、文字を1文字ずつ
切出し認識する。 この時文字認識結果として候補文字が1つに定まった場
合は該当文字コードを、それ以外、すなわち不読の場合
は予め定義されている不読文字コードと認識に付随して
得られる該当不読文字の特徴データおよび不読となった
文字の文字イメージをバス52を介しメモリ42に格納
する。修正部31は操作者が修正する文字のコードを入
力するためにキーボード71とCRT61が、また読取
結果を出力するためのフロッピーディスクドライバ81
が備えられている。修正部31はメモリ42に格納され
た不読文字のイメージをCRT61に表示し、操作者か
ら修正文字コードを入力する。この時メモリ42に格納
された表示した不読文字の特徴データと同−帳票内の他
の不読文字の特徴データを検索、比較し、表示された不
読文字と類似の不読文字が存在するかどうかを判定する
。類似不読文字が存在しない場合は表示した不読文字に
ついてのみ、類似不読文字が存在する場合は表示した不
読文字と類似不読文字全てについてメ七す42に格納さ
れた認識結果を操作者から入力された文字コードに置換
する。修正部31は認識部21が帳票1枚分を処理完了
する毎に動作し、帳票毎に修正語の認識結果をフロッピ
ーディスクドライバ81に出力することによりフロッピ
ーディスクに出力する。 次に、本実施例の装置の動作について第2図の帳票を入
力した場合について説明する。イメージ入力部11は第
2図の帳票を走査しイメージデータを入力しメモリ41
に格納する。認識部21はメモリ41に格納された帳票
のイメージデータから文字枠■から■に記入された文字
イメージを順次切り出し認識を行ない、結果として表1
の読取結果を得る。このとき文字枠■、■、■、■。 ■、■、[相]に記入された文字については読取結果に
対応する文字コードを、また認識した結果が不読となっ
た文字枠■、■、■、■に記入された文字の特徴データ
と文字イメージおよび予め定義されている不読文字コー
ドをメモリ42に格納する。修正部31は不読文字の文
字イメージとして文字枠■に記入された文字イメージを
CRT61に表示し操作者から修正文字コード”8“を
入力すると共にメモリ42に格納された表示した不読文
字の特徴データと同−帳票内の他の不読文字として文字
枠■、■、■の文字の特徴データを比較し類似文字か否
かを判定し、文字枠■、■、■に記入された文字全てが
文字枠■に記入された文字と類似しているという結論を
得る。このため修正部31はメモリ42に格納された文
字枠■、■。 ■、■に記入された文字の読取結果を一括して置換し表
2の修正結果を得、この修正結果をフロッピーディスク
ドライバ81に出力することによりフロッピーディスク
に出力する。 表2 以上のように、本実施例の場合、操作者が不読文字4文
字に対し1回の修正操作で修正を行ない結果を出力する
ことができる。 なお、本実施例では修正部31の類似文字検索範囲を同
−帳票内としたが、複数帳票に同一記入者が記入する運
用で使用する装置の場合などに検索範囲を複数帳票間に
拡張することが可能であることは明らかである。
The present invention relates to a character reading device, and more particularly to a method for correcting unreadable characters. [Prior Art] In conventional character reading devices, when correcting unreadable characters, an operator has to input character codes for all unreadable characters one by one. [Problems to be Solved by the Invention] In the conventional character reading device described above, when correcting unreadable characters, the operator had to input character codes for all unreadable characters one by one. There is a drawback that when a large number of similar deformed characters are misread, the number of corrections made by the operator increases accordingly. For example, when the form shown in Figure 2 is read and the judgment result shown in Table 1 is obtained, the operator selects the character frames ■ and ■. Similar correction operations were required four times for the four illegible characters with similar glyph shapes written in ■ and ■. Table 1 [Means for solving the problem] The character reading device of the present invention has an image memory and a document.
means for scanning each sheet and storing image data in the image memory; a recognition result storage memory; and a means for reading out the image data stored in the image memory, cutting out and recognizing characters one by one, and outputting recognized characters as character recognition results. A means for storing a code or a predefined unreadable character code and characteristic data of unreadable characters obtained upon recognition in a recognition result storage memory, and a modified character code for characters that are unreadable by an operator. a correction means for replacing an unreadable character code in the recognition result with the character code input by the correction character code input means; and correction means for referring to the recognition result memory and correcting it by the correction means means for searching for similar unreadable characters having similar characteristics to the unreadable characters, means for collectively replacing unreadable character codes of the similar unreadable characters with the specified character codes, and means for replacing the unreadable characters by the collective replacing means. and a cup means for outputting the recognized recognition results. [Function] Since similar illegible characters are corrected all at once, the number of corrections by the operator is reduced when a large number of similar deformed characters are illegible. [Example] Next, an example of the present invention will be described with reference to the drawings. FIG. 1 is a schematic diagram of a character reading device according to an embodiment of the present invention. The image input unit 11 operates the image data of the form for each form and inputs the image data to the memory 41 via the image data bus 51.
Store in. The recognition unit 21 reads the image data stored in the memory 41 via the bus 51, cuts out and recognizes characters one by one. At this time, if one candidate character is determined as a character recognition result, the corresponding character code is used, and if the other character is unreadable, the predefined unreadable character code and the corresponding unreadable character code obtained along with the recognition. The characteristic data of the characters and the character images of the illegible characters are stored in the memory 42 via the bus 52. The correction section 31 has a keyboard 71 and a CRT 61 for inputting the code of the character to be corrected by the operator, and a floppy disk driver 81 for outputting the reading result.
is provided. The correction unit 31 displays the image of the unreadable characters stored in the memory 42 on the CRT 61, and inputs a correction character code from the operator. At this time, the feature data of the displayed unreadable character stored in the memory 42 is searched and compared with the feature data of other unreadable characters in the form, and an unreadable character similar to the displayed unreadable character is found. Determine whether or not. If there are no similar unreadable characters, operate the recognition results stored in the menu 42 only for the displayed unreadable characters, or if similar unreadable characters exist, operate the recognition results stored in the menu 42 for the displayed unreadable characters and all similar unreadable characters. Replace with the character code input by the user. The correction unit 31 operates every time the recognition unit 21 completes processing one form, and outputs the recognition result of the corrected word for each form to the floppy disk driver 81, thereby outputting it to a floppy disk. Next, the operation of the apparatus of this embodiment will be described in the case where the form shown in FIG. 2 is input. The image input section 11 scans the form shown in FIG. 2, inputs image data, and stores it in the memory 41.
Store in. The recognition unit 21 sequentially extracts and recognizes the character images written in the character frames ■ to ■ from the image data of the form stored in the memory 41, and as a result, Table 1 is obtained.
Obtain the reading result. At this time, the character frame ■, ■, ■, ■. For the characters written in ■, ■, [phase], the character code corresponding to the reading result, and the characteristic data of the character written in the character frame ■, ■, ■, ■ whose recognition result was illegible. , a character image, and a predefined unreadable character code are stored in the memory 42 . The correction unit 31 displays the character image written in the character frame ■ as a character image of the unreadable character on the CRT 61, inputs the correction character code "8" from the operator, and at the same time inputs the character image of the displayed unreadable character stored in the memory 42. Same as the characteristic data - Compare the characteristic data of the characters in the character boxes ■, ■, ■ as other unreadable characters in the form to determine whether they are similar characters, and then compare the characteristic data of the characters written in the character frames ■, ■, ■. It is concluded that all the characters are similar to the characters written in the character box ■. Therefore, the correction unit 31 edits the character frames ■ and ■ stored in the memory 42. The reading results of the characters written in (1) and (2) are replaced all at once to obtain the correction results shown in Table 2, and the correction results are output to the floppy disk driver 81 to be output to the floppy disk. Table 2 As described above, in the case of this embodiment, the operator can correct four unreadable characters in one correction operation and output the result. In this embodiment, the similar character search range of the correction unit 31 is within the same form, but the search range may be expanded between multiple forms in the case of a device used in an operation where the same person fills in multiple forms. It is clear that this is possible.

【発明の効果】【Effect of the invention】

以上説明したように本発明は、類似不読文字を一括して
修正することにより、類似した変形文字の不読が多数発
生した場合に操作者による修正の回数を減少させる効果
があり、特に記入者に文字記入に関する充分な訓練を行
なうことのできない場合、例えば不特定多数の人が記入
し、くせ字、変形文字が多い帳票を読取る場合などに操
作者による不読文字の修正回数を大幅に減少させること
ができる効果がある。
As explained above, by collectively correcting similar illegible characters, the present invention has the effect of reducing the number of corrections made by the operator when a large number of similar deformed characters are illegible. In cases where it is not possible to provide sufficient training to operators on character entry, for example, when reading forms that have been filled out by an unspecified number of people and have many curly or distorted characters, the number of times operators must correct unreadable characters may be significantly reduced. There are effects that can be reduced.

【図面の簡単な説明】[Brief explanation of the drawing]

第1図は本発明の一実施例を示す文字読取装置の概略図
構成図、第2図は文字読取装置で読み取る帳票の例を示
す図である。 11・・・・・・・・・イメージ入力部、21・・・・
・・・・・認識部、 31・・・・・・・・・修正部、 41.42−−−−−−メモリ、 51、52・・・・・・バス、 61・・・・・・・・・CRT。 71・・・・・・・・・キーボード、 81・・・・・・・・・フロッピーディスクドライバ。 第1図
FIG. 1 is a schematic block diagram of a character reading device showing an embodiment of the present invention, and FIG. 2 is a diagram showing an example of a form read by the character reading device. 11... Image input section, 21...
...Recognition unit, 31...Modification section, 41.42--Memory, 51, 52...Bus, 61... ...CRT. 71...Keyboard, 81...Floppy disk driver. Figure 1

Claims (1)

【特許請求の範囲】 1、イメージメモリと、 帳票を1枚毎に走査し前記イメージメモリにイメージデ
ータを格納する手段と、 認識結果格納メモリと、 前記イメージメモリに格納されているイメージデータを
読出し、1文字ずつ切出し認識し、文字認識結果として
、認識文字コードまたは予め定義されている不読文字コ
ードと認識に付随し得られる不読文字の特徴データを認
識結果格納メモリに格納する手段と、 操作者が不読となった文字に対する修正文字コードを入
力するための修正文字コード入力手段と、 認識結果中の不読文字コードを前記修正文字コード入力
手段により入力された文字コードに置換する修正手段と
、 前記認識結果メモリを参照し前記修正手段により修正さ
れた不読文字と類似した特徴を持つ類似不読文字を検索
する手段と、 前記類似不読文字の不読文字コードを前記指定文字コー
ドに一括置換する手段と、 前記一括置換手段により置換された認識結果を出力する
手段とを有する文字読取装置。
[Claims] 1. An image memory, means for scanning a form one by one and storing image data in the image memory, a recognition result storage memory, and reading the image data stored in the image memory. , means for cutting out and recognizing characters one by one and storing, as a character recognition result, a recognized character code or a predefined unreadable character code and feature data of the unreadable character obtained along with the recognition in a recognition result storage memory; A corrected character code input means for inputting a corrected character code for an illegible character by an operator; and a correction for replacing the unreadable character code in the recognition result with the character code input by the corrected character code input means. means for referring to the recognition result memory and searching for similar unreadable characters having similar characteristics to the unreadable character corrected by the correcting means; and converting the unreadable character code of the similar unreadable character into the specified character. A character reading device comprising: means for collectively replacing a code; and means for outputting a recognition result replaced by the batch replacing means.
JP1044581A 1989-02-22 1989-02-22 Character reader Pending JPH02220186A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP1044581A JPH02220186A (en) 1989-02-22 1989-02-22 Character reader

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP1044581A JPH02220186A (en) 1989-02-22 1989-02-22 Character reader

Publications (1)

Publication Number Publication Date
JPH02220186A true JPH02220186A (en) 1990-09-03

Family

ID=12695460

Family Applications (1)

Application Number Title Priority Date Filing Date
JP1044581A Pending JPH02220186A (en) 1989-02-22 1989-02-22 Character reader

Country Status (1)

Country Link
JP (1) JPH02220186A (en)

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2002207960A (en) * 2001-01-12 2002-07-26 Nippon Digital Kenkyusho:Kk Method and program for recognized character correction

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2002207960A (en) * 2001-01-12 2002-07-26 Nippon Digital Kenkyusho:Kk Method and program for recognized character correction

Similar Documents

Publication Publication Date Title
US5048107A (en) Table region identification method
US20010014176A1 (en) Document image processing device and method thereof
US5509092A (en) Method and apparatus for generating information on recognized characters
JP3319203B2 (en) Document filing method and apparatus
JPH0423185A (en) Table reader provided with automatic cell attribution deciding function
JPH02220186A (en) Character reader
JPH0388062A (en) Device for preparing document
JPH0157837B2 (en)
JP3221968B2 (en) Character recognition device
JPH0384681A (en) Input processing method for business card information
JPH11232381A (en) Character reader
JP2606560B2 (en) Document image storage device
JP2976990B2 (en) Character recognition device
JPH06251187A (en) Method and device for correcting character recognition error
JPS61133487A (en) Character recognizing device
JP2990734B2 (en) Character recognition device output control method for character recognition device
JPH05303661A (en) Acquring/displaying device for partial image data
JPH03161866A (en) Device for recognizing contents
JPH0721303A (en) Character recognition device
JPH01189788A (en) Character reader
JPH07120396B2 (en) Document reader
JPS63143685A (en) Method for displaying recognized result in character recognizing device
JPH01134584A (en) Device for recognizing character
JPS60160490A (en) Character reader
JPH02271470A (en) Recognition result display device