JPH0475557B2 - - Google Patents
Info
- Publication number
- JPH0475557B2 JPH0475557B2 JP59160567A JP16056784A JPH0475557B2 JP H0475557 B2 JPH0475557 B2 JP H0475557B2 JP 59160567 A JP59160567 A JP 59160567A JP 16056784 A JP16056784 A JP 16056784A JP H0475557 B2 JPH0475557 B2 JP H0475557B2
- Authority
- JP
- Japan
- Prior art keywords
- character
- recognition result
- similar
- recognition
- type
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Expired - Lifetime
Links
Landscapes
- Character Discrimination (AREA)
Description
【発明の詳細な説明】
[発明の技術分野]
本発明は、類似文字群の文字認識処理を行なう
ことができる光学的文字読取装置に関する。DETAILED DESCRIPTION OF THE INVENTION [Technical Field of the Invention] The present invention relates to an optical character reading device that can perform character recognition processing for a group of similar characters.
[発明の技術的背景とその問題点]
近年、光学的文字読取装置(OCR)では、認
識対象の文字種が英数字、片仮名文字、平仮名文
字及び漢字のように広範囲になりつつある。この
ようなOCRでは、例えば数字の「0」と英字の
「O」または数字の「5」と英字の「S」のよう
な類似文字の認識が困難である。[Technical Background of the Invention and Problems Therewith] In recent years, optical character reading devices (OCR) are recognizing a wide range of character types, including alphanumeric characters, katakana characters, hiragana characters, and kanji characters. With such OCR, it is difficult to recognize similar characters, such as the number "0" and the alphabet "O" or the number "5" and the alphabet "S".
従来、上記のような類似文字には、用紙に記入
する際に例えば英字の「O」の上にラインを付加
したり、英字の「S」の左下部にセラフ(serif)
を付加するようなことが行われている。これによ
り、類似文字に対する認識率を向上することがで
きるが、認識対象文字の中に漢字が含まれると、
類似文字の数は飛躍的に増大する。このため、用
紙に記入する際に、類似文字間を区別するように
記入することは実際上困難である。このため、従
来のOCRでは、漢字を含む類似文字を確実に認
識することは極めて困難であつた。 Conventionally, when writing similar characters like the ones above, for example, a line was added above the letter "O", or a serif was added at the bottom left of the letter "S".
Things are being done to add . This can improve the recognition rate for similar characters, but if kanji are included in the characters to be recognized,
The number of similar characters increases dramatically. For this reason, when writing on a form, it is actually difficult to distinguish between similar characters. For this reason, it has been extremely difficult for conventional OCR to reliably recognize similar characters including kanji.
[発明の目的]
本発明は上記の点に鑑みてなされたもので、そ
の目的は、漢字を含む類似文字の認識を確実に実
行できる光学的文字読取装置を提供することにあ
る。[Object of the Invention] The present invention has been made in view of the above points, and an object thereof is to provide an optical character reading device that can reliably recognize similar characters including Chinese characters.
[発明の概要]
本発明は、文字認識手段の認識結果が類似文字
種であるか否かを判定する類似文字判定手段を備
えている。この判定結果により認識結果が類似文
字種に属する場合、文字種判別手段により、認識
結果がその前後の文字の文字種と同一である否か
が判別される。文字種が同一である場合、上記前
後の文字の文字種に対応する認識結果の文字コー
ドを最終的認識結果として出力する出力手段が設
けられている。また、文字種が同一でない場合、
予め用意された文字種判定テーブルにより上記認
識結果の文字種を決定する決定手段が設けられて
いる。[Summary of the Invention] The present invention includes similar character determining means for determining whether the recognition result of the character recognizing means is a similar character type. If the recognition result belongs to a similar character type, the character type determining means determines whether the recognition result is the same as the character type of the characters before and after it. If the character types are the same, output means is provided for outputting the character code of the recognition result corresponding to the character type of the preceding and succeeding characters as the final recognition result. Also, if the character types are not the same,
Determination means is provided for determining the character type of the recognition result using a character type determination table prepared in advance.
このような構成により、類似文字に対する文字
認識処理を確実に実行できる。 With such a configuration, character recognition processing for similar characters can be reliably executed.
[発明の実施例]
以下図面を参照して本発明の一実施例を説明す
る。第1図は一実施例に係わるOCRの構成を示
すブロツク図である。第1図において、文字認識
部10は、光電変換された文字パターンに対する
文字認識処理を実行し、認識結果Rをバツフアメ
モリ11に出力する。バツフアメモリ11は、通
常1行分の認識結果Rを記憶する。類似文字判別
部12は、バツフアメモリ11から出力された認
識結果Rが類似文字種に属するか否かを判別す
る。認識結果Rが類似文字種に属さない場合、類
似文字判別部12は認識結果Rを答出力部17に
出力する。[Embodiment of the Invention] An embodiment of the present invention will be described below with reference to the drawings. FIG. 1 is a block diagram showing the configuration of an OCR according to an embodiment. In FIG. 1, a character recognition unit 10 executes character recognition processing on a photoelectrically converted character pattern, and outputs a recognition result R to a buffer memory 11. The buffer memory 11 normally stores recognition results R for one line. The similar character determination unit 12 determines whether the recognition result R output from the buffer memory 11 belongs to a similar character type. If the recognition result R does not belong to the similar character type, the similar character discrimination section 12 outputs the recognition result R to the answer output section 17.
前後文字判別部13は、類似文字種に属する認
識結果Rを受信すると、バツフアメモリ11から
その認識結果Rの前後に存在する文字の認識結果
FBを読出し、認識結果Rが認識結果FBの文字種
と同一であるか否かを判別する。最終コード決定
部14は、前後文字判別部13の判別結果が同一
である場合、認識結果FBの文字種に対応する認
識結果Rの文字コードを最終的コードとして決定
し、答出力部17に出力する。最終コード判定部
15は、前後文字判別部13の判別結果が同一で
ない場合、予めテーブルメモリ16に記憶された
文字種判定テーブルを利用して認識結果Rの文字
種を判定し、その文字種に対応する文字コードを
答出力部17に出力する。 When receiving the recognition result R belonging to the similar character type, the preceding and following character discrimination unit 13 extracts the recognition results of the characters existing before and after the recognition result R from the buffer memory 11.
FB is read and it is determined whether the recognition result R is the same as the character type of the recognition result FB. When the discrimination results of the preceding and following character discriminating section 13 are the same, the final code determination section 14 determines the character code of the recognition result R corresponding to the character type of the recognition result FB as the final code, and outputs it to the answer output section 17. . If the discrimination results of the preceding and following character discrimination sections 13 are not the same, the final code determination section 15 determines the character type of the recognition result R using a character type determination table stored in the table memory 16 in advance, and determines the character type corresponding to the character type. The code is output to the answer output section 17.
このような構成のOCRにおいて、一実施例に
係わる動作を説明する。先ず、用紙に記録された
文字が光電変換された後に、文字認識部10に与
えられる。文字認識部10では、文字パターンに
対する文字認識処理が実行される。これにより得
られた認識結果Rが文字コードとしてバツフアメ
モリ11に格納される。このとき、認識対象の文
字が類似文字種である場合、漢字等に対応する文
字コードが取敢えずバツフアメモリ11に格納さ
れる。例えば認識結果Rが「−」である場合、長
音を示す文字コードがバツフアメモリ11に格納
される。類似文字判別部12では、バツフアメモ
リ11からの認識結果Rが類似文字種であるか否
かの判別がなされる。ここで、類似文字判別部1
2は、予め例えば第2図に示すような類似文字群
(第1水準の漢字を含む)からなるテーブルを記
憶している。認識結果Rが第2図の類似文字群に
属していなければ、認識結果Rはそのままの文字
コードで答出力部17に出力される。 In the OCR having such a configuration, the operation according to one embodiment will be explained. First, characters recorded on paper are subjected to photoelectric conversion and then provided to the character recognition section 10. The character recognition unit 10 performs character recognition processing on character patterns. The recognition result R obtained thereby is stored in the buffer memory 11 as a character code. At this time, if the character to be recognized is a similar character type, the character code corresponding to the Chinese character or the like is temporarily stored in the buffer memory 11. For example, if the recognition result R is "-", a character code indicating a long sound is stored in the buffer memory 11. The similar character determination unit 12 determines whether the recognition result R from the buffer memory 11 is a similar character type. Here, similar character discriminator 1
2 stores in advance a table consisting of a group of similar characters (including first level kanji) as shown in FIG. 2, for example. If the recognition result R does not belong to the similar character group shown in FIG. 2, the recognition result R is output to the answer output section 17 with the same character code.
一方、認識結果Rが類似文字群に属する場合、
前後文字判別部13では認識結果Rの該当文字の
前後の文字の文字種が判別される。この場合、前
後文字判別部13はバツフアメモリ11から前後
の文字に対応する認識結果FBを読出し、文字種
の判別を実行する。この判別結果において、同一
文字種である場合にはその文字種に対応する認識
結果Rの文字コードが最終コード決定部14に出
力される。例えば、認識結果Rが「ハ」のとき、
前後の認識結果FBが片仮名文字であれば、片仮
名文字の「ハ」に対応する文字コードが最終コー
ド決定部14に出力される。最終コード決定部1
4では前後文字判別部13から出力された文字コ
ードを最終的な文字コードとして決定され、その
文字コードが答出力部17に出力される。 On the other hand, if the recognition result R belongs to a similar character group,
The preceding and following character determination unit 13 determines the character types of the characters before and after the corresponding character in the recognition result R. In this case, the preceding and following character discrimination section 13 reads out the recognition results FB corresponding to the preceding and following characters from the buffer memory 11, and executes character type discrimination. As a result of this discrimination, if the characters are of the same type, the character code of the recognition result R corresponding to the character type is output to the final code determining section 14. For example, when the recognition result R is "Ha",
If the preceding and succeeding recognition results FB are katakana characters, the character code corresponding to the katakana character "ha" is output to the final code determination unit 14. Final code determination section 1
In step 4, the character code output from the preceding and succeeding character discriminating section 13 is determined as the final character code, and the character code is output to the answer output section 17.
また、認識結果Rの文字種が前後の文字の場合
と異なるとき、最終コード判定部15において認
識結果Rの文字コードが判定される。最終コード
判定部15では、例えば第3図に示すような文字
種判定テーブルを利用して、文字コードの判定が
実行される。ここで、第3図において、長は長
音、ダはダツシユー、−は漢数字、マはマイナ
ス、?は判定不能を示すものとする。例えば、認
識結果Rが未知なる「−」である場合、前後文字
判別部13の判別結果により前後の文字がそれぞ
れ漢字であるとする。この場合には、第3図のテ
ーブルにより、認識結果Rは漢数字の「−」であ
ると判定される。そして、最終コード判定部15
で判定された文字コードが、認識結果Rの最終的
コードとして答出力部17に出力される。尚、最
終コード判定部15において、判定結果が「?」
の場合には、文字種候補が2種以上かまたは頻度
が希の場合と判定し、判定歩能としてリジエクト
される。 Further, when the character type of the recognition result R is different from that of the preceding and succeeding characters, the character code of the recognition result R is determined in the final code determining section 15. The final code determination section 15 executes character code determination using, for example, a character type determination table as shown in FIG. Here, in Figure 3, long is a long sound, da is datsushiyu, - is a Chinese numeral, ma is a minus, ? indicates that it cannot be determined. For example, when the recognition result R is an unknown "-", it is assumed that the preceding and succeeding characters are respectively Chinese characters based on the discrimination results of the preceding and following character discriminating unit 13. In this case, the recognition result R is determined to be the Chinese numeral "-" according to the table shown in FIG. Then, the final code determination section 15
The character code determined in is outputted to the answer output section 17 as the final code of the recognition result R. In addition, the final code determination unit 15 determines that the determination result is "?"
In this case, it is determined that there are two or more character type candidates or the frequency is rare, and the character type candidate is rejected as a judged gait performance.
このようにして、認識対象の文字が類似文字群
に属する場合、その文字の前後の文字の文字種に
基づいて文字種を判定する。この判定結果によ
り、決定された文字種に対応する文字コードを最
終的文字認識結果として出力する。したがつて、
用紙に予め類似文字間を区別する記入をする必要
がなく、漢字を含む類似文字の認識を確実に行な
うことができる。 In this way, when a character to be recognized belongs to a similar character group, the character type is determined based on the character types of the characters before and after the character. Based on this determination result, the character code corresponding to the determined character type is output as the final character recognition result. Therefore,
There is no need to write in advance on paper to distinguish between similar characters, and similar characters including kanji can be reliably recognized.
[発明の効果]
以上詳述したように本発明によれば、漢字を含
む類似文字の認識を、簡単な構成で確実に行なう
ことができる。したがつて、広い範囲の文字種に
属する認識対象文字に対する認識率を大幅に向上
することができるものである。[Effects of the Invention] As described in detail above, according to the present invention, similar characters including Chinese characters can be reliably recognized with a simple configuration. Therefore, the recognition rate for characters to be recognized belonging to a wide range of character types can be greatly improved.
第1図は本発明の一実施例に係わるOCRの構
成を示すブロツク図、第2図は第1図の類似文字
判別部に記憶されたテーブルの一例を示す図、第
3図は第1図のテーブルメモリに記憶された文字
種判定テーブルの一例を示す図である。
10……文字認識部、12……類似文字判別
部、13……前後文字判別部、14……最終コー
ド決定部、15……最終コード判定部、16……
テーブルメモリ。
FIG. 1 is a block diagram showing the configuration of an OCR according to an embodiment of the present invention, FIG. 2 is a diagram showing an example of a table stored in the similar character discriminator shown in FIG. 1, and FIG. 3 is a diagram similar to the one shown in FIG. FIG. 2 is a diagram showing an example of a character type determination table stored in a table memory of FIG. 10...Character recognition unit, 12...Similar character discrimination unit, 13...Next and subsequent character discrimination unit, 14...Final code determination unit, 15...Final code determination unit, 16...
table memory.
Claims (1)
識処理を行なう文字認識手段と、 この文字認識手段による認識結果が類似文字群
に属するか否かを、予め用意された類似文字群テ
ーブルを参照して判別する類似文字判別手段と、 この類似文字判別手段の判別結果により前記認
識結果が類似文字群に属する場合に、前記認識結
果の文字種と前記認識結果の前後に存在する各文
字の文字種とが同一であるか否かを判別する前後
文字判別手段と、 この前後文字判別手段の判別結果により前記認
識結果が前記各文字の文字種と同一文字種の場合
に、その文字種に対応する前記認識結果の文字コ
ードを最終的文字認識結果として出力する最終コ
ード決定手段と、 前記前後文字判別手段の判別結果により前記認
識結果の文字種と前記各文字の文字種とが異なる
場合に、予め用意された文字種判定テーブルを参
照して、前記各文字の文字種に基づいて前記認識
結果の文字コードを決定し、この文字コードを最
終的文字認識結果として出力する最終コード判定
手段とを具備したことを特徴とする光学的文字読
取装置。[Scope of Claims] 1. A character recognition means that performs character recognition processing on a photoelectrically converted character pattern, and a similar character group table prepared in advance to determine whether or not a recognition result by this character recognition means belongs to a similar character group. and a similar character discriminating means that discriminates by referring to the similar character discriminating means, and when the recognition result belongs to a similar character group according to the discrimination result of the similar character discriminating means, the character type of the recognition result and each character existing before and after the recognition result. a preceding and following character discriminating means for discriminating whether or not the preceding and following character types are the same; and when the recognition result of the preceding and following character discriminating means is the same character type as the character type of each character, the recognition corresponding to the character type; a final code determination means for outputting the resulting character code as a final character recognition result; and a character type prepared in advance when the character type of the recognition result is different from the character type of each character according to the determination result of the preceding and following character determining means. The present invention is characterized by comprising a final code determination means for referring to a determination table, determining a character code of the recognition result based on the character type of each character, and outputting this character code as a final character recognition result. Optical character reader.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP16056784A JPS6139175A (en) | 1984-07-31 | 1984-07-31 | Optical character reading device |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP16056784A JPS6139175A (en) | 1984-07-31 | 1984-07-31 | Optical character reading device |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| JPS6139175A JPS6139175A (en) | 1986-02-25 |
| JPH0475557B2 true JPH0475557B2 (en) | 1992-12-01 |
Family
ID=15717764
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP16056784A Granted JPS6139175A (en) | 1984-07-31 | 1984-07-31 | Optical character reading device |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPS6139175A (en) |
Families Citing this family (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS6330991A (en) * | 1986-07-25 | 1988-02-09 | Matsushita Electric Ind Co Ltd | Character recognizing device |
| JPH0290384A (en) * | 1988-09-28 | 1990-03-29 | Ricoh Co Ltd | Character recognition device post-processing method |
| US5048113A (en) * | 1989-02-23 | 1991-09-10 | Ricoh Company, Ltd. | Character recognition post-processing method |
Family Cites Families (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS5927381A (en) * | 1982-08-07 | 1984-02-13 | Fujitsu Ltd | Character recognizing system |
-
1984
- 1984-07-31 JP JP16056784A patent/JPS6139175A/en active Granted
Also Published As
| Publication number | Publication date |
|---|---|
| JPS6139175A (en) | 1986-02-25 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| JPH0475557B2 (en) | ||
| JPS56149676A (en) | Pattern recognizer | |
| JPS59158482A (en) | Character recognizing device | |
| JP2732593B2 (en) | Character reading system | |
| JPS6160184A (en) | Optical character reader | |
| JP2686745B2 (en) | Character reader | |
| JPS61114388A (en) | Character input device | |
| JPS6095689A (en) | Optical character reader | |
| JPS59128682A (en) | Character reader | |
| JPH0576674B2 (en) | ||
| JPH0259504B2 (en) | ||
| JPS60254388A (en) | Optical character reader | |
| JPS58101378A (en) | Manuscript document reading method | |
| JPH02242389A (en) | Zip code reader | |
| JPS59109978A (en) | Optical character reader | |
| JPS60138689A (en) | Character recognizing method | |
| JPS63188284A (en) | Character reader | |
| JPS6115288A (en) | Optical character reader | |
| JPH01316888A (en) | Zip code reader | |
| JPH01265378A (en) | European character recognizing system | |
| JPS59149569A (en) | Optical character reader | |
| JPS62140188A (en) | Character recognition post-processing method | |
| JPS59121476A (en) | Optical character reader | |
| JPS59128681A (en) | Character reader | |
| JPS6363954B2 (en) |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| EXPY | Cancellation because of completion of term |