JPH0199189A - Character segmenting system - Google Patents

Character segmenting system

Info

Publication number
JPH0199189A
JPH0199189A JP62257812A JP25781287A JPH0199189A JP H0199189 A JPH0199189 A JP H0199189A JP 62257812 A JP62257812 A JP 62257812A JP 25781287 A JP25781287 A JP 25781287A JP H0199189 A JPH0199189 A JP H0199189A
Authority
JP
Japan
Prior art keywords
frame
character
auxiliary frame
image data
null
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
JP62257812A
Other languages
Japanese (ja)
Inventor
Tatsunosuke Iwahara
岩原 達之助
Hirobumi Okazaki
岡崎 博文
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Sanyo Electric Co Ltd
Original Assignee
Sanyo Electric Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Sanyo Electric Co Ltd filed Critical Sanyo Electric Co Ltd
Priority to JP62257812A priority Critical patent/JPH0199189A/en
Publication of JPH0199189A publication Critical patent/JPH0199189A/en
Pending legal-status Critical Current

Links

Landscapes

  • Character Input (AREA)

Abstract

PURPOSE:To accurately segment character, and to improve the recognition rate by executing searches for vacant null and null dot processings in both up-and-down directions and right-and-left direction from an inner auxiliary frame to an outer auxiliary frame. CONSTITUTION:In image data obtained by scanning an input sheet, the outer auxiliary frame 12 which is generated by extending a logical frame 11 corresponding to a writing frame outward for a prescribed number of dots, as well as the inner auxiliary frame 13 generated by shrinking inward for a prescribed number of dots, are set, and null lines are searched in up-and-down direction and right-and-left direction from the frame 13 to the frame 12. When a null line is searched, the image data from this line to the frame 13 is processed to be null, and image data surrounded by the frame 13 is segmented as the pattern of one character. As a result, even if a part of a character pattern to be recognized overflows from the logical frame 11, the character can be segment without any lacked part only if the character is within the outer auxiliary frame, also, even if the neighboring character intrudes the logical frame, it is eliminated if the intruding is no longer deeper than the inner auxiliary frame, hence an accurate segmenting of a character can be executed.

Description

【発明の詳細な説明】 (イ)産業上の利用分野 本発明は、手書き文字の認識装置に係り、文字を1文字
づつ切出す文字切出し方式に関する。
DETAILED DESCRIPTION OF THE INVENTION (A) Field of Industrial Application The present invention relates to a handwritten character recognition device, and more particularly to a character cutting method for cutting out characters one by one.

(ロ)従来の技術 手書き文字を認識する装置においては、一般に、文字を
記入すべき位置及び大きさを筆者に知らしめるため、入
力シートに文字の記入枠が文字単位にドロップアウトカ
ラーで印刷されている。
(b) Conventional technology In devices that recognize handwritten characters, generally a character entry frame is printed in a dropout color on the input sheet for each character in order to inform the writer of the position and size at which the character should be entered. ing.

そして、従来は、入力シートを走査して得られる画像デ
ータから文字を切出すには、タイミングマーク等を基準
にして記入枠に応じた理論枠を設定し、この理論枠で囲
まれた画像データを1文字のパターンとして切出してい
た。このような技術は、例えば、特公昭5B−3162
8号公報に開示されている。
Conventionally, in order to cut out characters from image data obtained by scanning an input sheet, a theoretical frame corresponding to the entry frame is set based on timing marks, etc., and the image data surrounded by this theoretical frame is was cut out as a single character pattern. Such technology is known, for example, from Japanese Patent Publication No. 5B-3162.
It is disclosed in Publication No. 8.

(ハ〉発明が解決しようとする問題点 文字を手書きする場合、記入枠からのはみだしや隣接文
字の入り込みを完全には回避することができず、この場
合、記入枠に応じた理論枠内には、認識しようとする文
字の一部が欠落した文字パターンや、不要な隣接文字の
一部を含む文字パターンが得られることとなる。従って
、従来の如く、理論枠で囲まれた画像データを1文字の
バターンとして切出すと、正確な切出しが行えず認識率
の低下を招く。
(C) Problems to be Solved by the Invention When writing characters by hand, it is impossible to completely avoid protruding from the writing frame or entering adjacent characters, and in this case, it is impossible to completely avoid writing characters outside the writing frame or entering the adjacent characters. This results in a character pattern in which part of the character to be recognized is missing, or a character pattern including part of unnecessary adjacent characters.Therefore, as in the past, image data surrounded by a theoretical frame is If the pattern is cut out as a single character, accurate cutout cannot be performed, resulting in a decrease in the recognition rate.

又、仮に記入枠内にはみださないように文字を手書きし
ても入力シート自体が傾くと、理論枠内から認識しよう
とする文字パターンの一部がはみだしたり、隣接文字の
一部が入り込む現象が起き、この場合も従来の切出し方
式では認識率が低下する。
Also, even if you handwrite the characters so that they do not protrude into the writing frame, if the input sheet itself is tilted, part of the character pattern to be recognized may protrude from within the theoretical frame, or some adjacent characters may This phenomenon occurs, and in this case too, the recognition rate decreases with the conventional extraction method.

(ニ)問題点を解決するための手段 本発明は、入力シートを走査して得られる画像データに
、記入枠に応じた理論枠より外方へ所定ドツト数分拡げ
た外側補助枠及び内方へ所定ドツト数分狭めた内側補助
枠とを設定し、該内側補助枠から外側補助枠に向って上
下左右各方向に空白ラインの探索を行い、探索されたと
きは該空白ラインから前記外側補助枠までの画像データ
を空白とする処理を行い、前記外側補助枠で囲まれた画
像データを1文字のパターンとして切出すようにしたも
のである。
(d) Means for Solving the Problems The present invention provides image data obtained by scanning an input sheet with an outer auxiliary frame that extends outward by a predetermined number of dots from the theoretical frame corresponding to the entry frame, and an inner auxiliary frame. An inner auxiliary frame narrowed by a predetermined number of dots is set to the inner auxiliary frame, and a blank line is searched in each direction from the inner auxiliary frame to the outer auxiliary frame. The image data up to the frame is blanked out, and the image data surrounded by the outer auxiliary frame is cut out as a pattern of one character.

(ネ)作用 本発明では、認識しようとする文字パターンの一部が理
論枠からはみだしても、外側補助枠内に入ってさえいれ
ば欠落することなく文字を切出せ、又、理論枠内に隣接
文字が入り込んでも内側補助枠外までであれば、これら
の不要な隣接文字パターンは排除きれてしまい、正確な
文字切出しが行われる。
(N) Function In the present invention, even if a part of the character pattern to be recognized protrudes from the theoretical frame, as long as it falls within the outer auxiliary frame, the character can be cut out without being lost. Even if adjacent characters enter, as long as they are outside the inner auxiliary frame, these unnecessary adjacent character patterns can be eliminated and accurate character extraction can be performed.

(へ)実施例 第2図は、本発明の文字切出し方式を実現する文字認識
装置全体の構成を示すブロック図であり、(1)は文字
単位の記入枠(2)がドロップアウトカラーで印刷され
、各行の左端に黒色のタイミングマーク(3)が、印刷
きれた入力シート、(4)は入力シートを走査して白黒
2値の画像データを得る文字観測部、(5)は入力シー
ト(1)の複数性分の容量を有し得られた画像データを
記憶する画像メモリ1.そして、(6)が1文字分の画
像データを記憶するための文字メモリ(7)を有し、画
像メモリ(5)の画像データから1文字づつ文字パター
ンを切出す文字切出し部であって、切出された文字パタ
ーンは認識部(8)へ送出され、ここで、特徴パターン
が抽出され、標準パターンメモリ(9)に予め記憶され
ている標準パターンと照合されて、文字の認識が行われ
る。
(v) Embodiment Figure 2 is a block diagram showing the overall configuration of a character recognition device that realizes the character extraction method of the present invention, in which (1) the entry frame (2) for each character is printed in dropout color. A black timing mark (3) is placed at the left end of each line on the fully printed input sheet, (4) is the character observation unit that scans the input sheet to obtain black and white binary image data, and (5) is the input sheet ( An image memory 1 which has a capacity corresponding to the plurality of characteristics of 1) and stores the obtained image data. and (6) is a character cutting unit which has a character memory (7) for storing image data for one character and cuts out a character pattern one character at a time from the image data of the image memory (5), The extracted character pattern is sent to the recognition unit (8), where characteristic patterns are extracted and compared with standard patterns stored in advance in the standard pattern memory (9) to perform character recognition. .

以下、文字切出し部(6)での処理内容を、第1図及び
第3図〜第4図のフローチャートを参照しながら詳細に
説明する。
Hereinafter, the processing contents of the character cutting section (6) will be explained in detail with reference to the flowcharts of FIG. 1 and FIGS. 3 to 4.

先ず、画像メモリ(5)中の画像データにおいて、認識
しようとする行の読取ったタイミングマーク(10)の
座標(To、■、)を求める。次に、入力シート(1)
のタイミングマーク(3)から記入枠(2)までの相対
距離1. 、1.、及び、記入枠(2)の大きさ即ち幅
w1及び高さちと、求めたタイミングマークの座標(1
,、T、)に基づいて、記入枠(2)に応じた理論枠(
11)(実線)を求める。更に、予め定められた外側X
方向余裕度n、及び外側y方向余裕度n、を用いて、理
論枠(11)より外方へ所定ドツト数分拡げた外側補助
枠(12)(1点鎖線)を設定し、この外側補助枠(1
2〉で囲まれた画像データを文字メモリ(7)に移す。
First, in the image data in the image memory (5), the coordinates (To, ■,) of the read timing mark (10) of the line to be recognized are determined. Next, input sheet (1)
Relative distance from timing mark (3) to entry frame (2) 1. , 1. , and the size of the entry frame (2), that is, the width w1 and height, and the coordinates (1
,,T,), the theoretical frame (
11) Find (solid line). Furthermore, a predetermined outer X
Using the direction margin n and the outer y direction margin n, an outer auxiliary frame (12) (dotted chain line) is set that is expanded outward by a predetermined number of dots from the theoretical frame (11), and this outer auxiliary frame Frame (1
Move the image data surrounded by 2> to the character memory (7).

ここで、外側補助枠(12)の4点A、B、C,Dの画
像メモリ(5〉上での座標は、各々、(T、+1. 1
1. t T、+1.  nt) F (T、+1゜1
1+ * IF” IF” nt + My) + (
Tz+1. + nI+ W、 、 T、 +1y−n
t ) + (T t ” 1 x + n 1 + 
W、 + T、” 1y + n、 + W、 )であ
り、文字メモリ(7)上ではA点が原点となる。
Here, the coordinates of the four points A, B, C, and D of the outer auxiliary frame (12) on the image memory (5) are (T, +1. 1
1. t T, +1. nt) F (T, +1゜1
1+ * IF"IF" nt + My) + (
Tz+1. + nI+ W, , T, +1y-n
t ) + (T t ” 1 x + n 1 +
W, + T, 1y + n, + W, ), and point A is the origin on the character memory (7).

次に、文字メモリ(7)上で、予め定められた内側X方
向余裕度m□及び内側y方向余裕度m、を用いて、理論
枠(11)より内方へ所定ドツト数分狭めた内側補助枠
(13) (破線)を設定し、空白ライン探索及び空白
処理を実行する。この探索及び処理は、第4図のフロー
チャートに詳述するように、文字メモリ(7)で、先ず
、内側補助枠(13)の上方ライン(0+ ax + 
m@  l )〜(w、+2n+  1 、 n、+m
2−1〉から、外側補助枠(12)の上方ライン(0,
0)〜(%I、+2n、−1,0)に向って、1ドツト
ラインづつ空白ラインであるかチエツクし、空白ライン
であればその空白ラインから外側補助枠の上方ラインま
での全てのドツトラインを空白にし、同様の処理を、内
側補助枠(13)の下方、右方、左方についても行う。
Next, on the character memory (7), using the predetermined inner X-direction margin m A supplementary frame (13) (broken line) is set, and blank line search and blank processing are performed. As detailed in the flowchart of FIG. 4, this search and processing begins with the upper line (0+ ax +
m@l ) ~ (w, +2n+ 1, n, +m
2-1>, the upper line (0,
0) to (%I, +2n, -1,0), check each dot line to see if it is a blank line, and if it is a blank line, check all dot lines from that blank line to the upper line of the outer auxiliary frame. The blank is left blank, and the same process is performed for the lower, right, and left sides of the inner auxiliary frame (13).

以上の処理を施した文字メモリ(7)内の画像データを
、1文字のパターンとして切出して認識部(8〉へ送出
し、同一行の次文字についても同様の処理を繰り返す。
The image data in the character memory (7) subjected to the above processing is cut out as a pattern of one character and sent to the recognition unit (8>), and the same processing is repeated for the next character in the same line.

1行の全文字について切出しが終了したら、次行へ移り
、その行のタイミングマークの座標検出から同様の処理
を開始する。
When all the characters in one line have been cut out, the process moves to the next line and starts the same process from detecting the coordinates of the timing mark in that line.

第5図は、本発明の方式を用いて実際に文字を切出した
例を示す図であり、第1図と同様、実線が理論枠(11
)、1点鎖線が外側補助枠(12)、破線が内側補助枠
(13)である。
FIG. 5 is a diagram showing an example of actually cutting out characters using the method of the present invention. Similar to FIG. 1, the solid line indicates the theoretical frame (11
), the dashed line is the outer auxiliary frame (12), and the broken line is the inner auxiliary frame (13).

第5図(a)の如く理論枠(11)内に文字がきれいに
納まっているときだけでなく、第5図(b)のように、
文字の一部が理論枠(11)からはみ出たり、第5図(
C)のように、隣接文字の一部が理論枠(11)内に入
り込んだ場合でも、正確に文字が切出される。
Not only when the characters fit neatly within the theoretical frame (11) as shown in Figure 5(a), but also when the characters fit neatly within the theoretical frame (11) as shown in Figure 5(b).
Some of the letters may protrude from the theoretical frame (11) or
Even if part of the adjacent character falls within the theoretical frame (11) as in C), the character is accurately cut out.

(ト)発明の効果 本発明に依れば、文字の記入の仕方が悪かったり、ある
いは、入力シートの傾きによって、記入枠に応じた理論
枠から文字の一部がはみ出したり、隣接文字の一部が入
り込んでも、正確に文字を切出すことができ、従って、
認識率が向上する。
(G) Effects of the Invention According to the present invention, if characters are entered incorrectly or due to the inclination of the input sheet, some of the characters may protrude from the theoretical frame corresponding to the entry frame, or some of the adjacent characters may Even if the parts get stuck in, the characters can be cut out accurately.
Recognition rate improves.

【図面の簡単な説明】[Brief explanation of the drawing]

第1図は本発明の切出し方式を説明するための説明図、
第2UgJは本発明方式を用いた文字認識装置全体の構
成を示すブロック図、第3図は本発明方式の処理内容を
示すフローチへ・−ト、第4図は本発明方式における要
部処理内容を詳述したフローチャート、第5図は本発明
方式の切出し例を示す図である。 (1)・・・入力シート、 (2)・・・記入枠、(3
)・・・タイミングマーク、 <5)・・・画像メモリ
、(6)・・・文字切出し部、 (7〉・・・文字メモ
リ。
FIG. 1 is an explanatory diagram for explaining the cutting method of the present invention,
2nd UgJ is a block diagram showing the overall configuration of a character recognition device using the method of the present invention, FIG. 3 is a flowchart showing the processing contents of the method of the present invention, and FIG. 4 is the main processing contents of the method of the present invention. FIG. 5 is a flowchart illustrating an example of the extraction method of the present invention. (1)...Input sheet, (2)...Entry frame, (3
)...Timing mark, <5)...Image memory, (6)...Character cutting section, (7>...Character memory.

Claims (1)

【特許請求の範囲】[Claims] (1)文字を記入するための記入枠が予め印刷された入
力シートを用いる文字認識装置において、前記入力シー
トを走査して得られる画像データに、前記記入枠に応じ
た理論枠より外方へ所定ドット数分拡げた外側補助枠及
び内方へ所定ドット数分狭めた内側補助枠とを設定し、
該内側補助枠から外側補助枠に向って上下左右各方向に
空白ラインの探索を行い、探索されたときは該空白ライ
ンから前記外側補助枠までの画像データを空白とする処
理を行い、前記外側補助枠で囲まれた画像データを1文
字のパターンとして切出すことを特徴とする文字切出し
方式。
(1) In a character recognition device that uses an input sheet on which a writing frame for writing characters is printed in advance, the image data obtained by scanning the input sheet is added to the outside of the theoretical frame corresponding to the writing frame. Setting an outer auxiliary frame that is expanded by a predetermined number of dots and an inner auxiliary frame that is narrowed inward by a predetermined number of dots,
A search for a blank line is performed in each direction from the inner auxiliary frame to the outer auxiliary frame, and when a blank line is found, the image data from the blank line to the outer auxiliary frame is blanked, and the image data from the outer auxiliary frame is A character cutting method characterized by cutting out image data surrounded by an auxiliary frame as a single character pattern.
JP62257812A 1987-10-13 1987-10-13 Character segmenting system Pending JPH0199189A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP62257812A JPH0199189A (en) 1987-10-13 1987-10-13 Character segmenting system

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP62257812A JPH0199189A (en) 1987-10-13 1987-10-13 Character segmenting system

Publications (1)

Publication Number Publication Date
JPH0199189A true JPH0199189A (en) 1989-04-18

Family

ID=17311463

Family Applications (1)

Application Number Title Priority Date Filing Date
JP62257812A Pending JPH0199189A (en) 1987-10-13 1987-10-13 Character segmenting system

Country Status (1)

Country Link
JP (1) JPH0199189A (en)

Similar Documents

Publication Publication Date Title
CA1160347A (en) Method for recognizing a machine encoded character
JP2822189B2 (en) Character recognition apparatus and method
US5075895A (en) Method and apparatus for recognizing table area formed in binary image of document
JPS6132187A (en) Character recognition system
JPH06187489A (en) Character recognizing device
JP2957729B2 (en) Line direction determination device
JPS615383A (en) Character pattern separating device
JP3095470B2 (en) Character recognition device
JPH0713994A (en) Character recognition device
JPH0388085A (en) Optical character reader
JP3077929B2 (en) Character extraction method
JP2003016385A (en) Image processing apparatus, method, program, and storage medium
JP2025091589A (en) Image processing device, image processing system and program
JPH0697470B2 (en) Character string extractor
JP3566738B2 (en) Shaded area processing method and shaded area processing apparatus
JPH11250256A (en) Graphic recognition processing method and recording medium recording the program
JP2957740B2 (en) Line direction determination device
JP3127413B2 (en) Character recognition device
JPH06274690A (en) Character recognition device and optical character reader
JPH04156694A (en) Character recognition system
JPH0632079B2 (en) Character recognition device
JPS63195783A (en) Character segmenting system
JPH0746371B2 (en) Character reader
JPH05274472A (en) Image recognizing device
JPS63239569A (en) character recognition device