JPS59112367A - Reading method of character - Google Patents
Reading method of characterInfo
- Publication number
- JPS59112367A JPS59112367A JP57222489A JP22248982A JPS59112367A JP S59112367 A JPS59112367 A JP S59112367A JP 57222489 A JP57222489 A JP 57222489A JP 22248982 A JP22248982 A JP 22248982A JP S59112367 A JPS59112367 A JP S59112367A
- Authority
- JP
- Japan
- Prior art keywords
- character
- turn
- chisel
- pattern
- clusters
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
Landscapes
- Character Input (AREA)
- Character Discrimination (AREA)
Abstract
Description
【発明の詳細な説明】
定で々い文書の文字を精度よくかつ高速に読取ることの
できる文字読取方式に関するものである。DETAILED DESCRIPTION OF THE INVENTION The present invention relates to a character reading method that can read characters in fixed documents with high precision and high speed.
従来の文字読取装置では第1図に示すように走査・光電
変換した帳票上の文字・ξターンを信号入力端子1を介
してノミターンメモリ2に一旦格納し、該・ξターンメ
モリ2上の文字行に対し文字切出し部3により文字列の
先頭から予め定められた一定間隔で一文字ずつ切出し、
その文字ノミターンを識別部4へ転送してそれが何とい
う文字であるかを判定し,更にその識別結果を出力端子
5より出力するようになしていた。このため帳票」二の
文字ピッチが一定でないとまったく読取れないとーう欠
点があった。In a conventional character reading device, as shown in FIG. The character cutting section 3 cuts out each character from the beginning of the character string at predetermined intervals from the character line,
The character number turn is transferred to the identification section 4 to determine what character it is, and the identification result is outputted from the output terminal 5. For this reason, there was a drawback that unless the character pitch of the form was constant, it could not be read at all.
また上記文字読取装置の文字切出しg++ :i K補
助識別部6を設け、幅の狭いパターンのみを該補助識別
部6にて一文字であるがあるいは文字の一部であるかを
判定し、一文字のパターンと判定された場合には該幅の
狭いパターンを一文字とみなして文字切出しを行ない、
また文字の一部と判定されれば次のノミターンに加えて
文字切出しを行なうようにして文字ピッチが一定で々−
帳票より文字を読取れるようにしたものは既に提案され
て層だ。In addition, the character cutting g++:iK auxiliary identification section 6 of the character reading device is provided, and the auxiliary identification section 6 judges whether only a narrow pattern is one character or a part of a character. If it is determined to be a pattern, the narrow pattern is regarded as one character and character extraction is performed,
Also, if it is determined to be part of a character, the character is cut out in addition to the next chisel turn, so that the character pitch remains constant.
There have already been proposals to make text more readable than on forms.
しかしながら上記装置では幅の狭いノミターンが他の文
字パターンと同様なパターンを成す(例えば横書きの帳
票において漢字「化」のようにその偏「イ」がカタカナ
の「イ」と同じ字形を成すような)場合や、文字に切れ
が生じて一文字が多くの,2ターンに分かれたり、連続
l2て書かれた2以−Lの文字が一文字として判定され
る場合等において十分な切出し精度を得ることが困難で
あった。捷だ補助識別部6が動作する場合は、その判定
結果が出されるまで文字切出[、部3の動作が停止する
ため読取速度が遅くなるという欠点があった。また、こ
の読取速度の低下は補助識別部6で対処しなければ々ら
ないノミターンが増える程、顕著になる欠点があった。However, in the above device, the narrow chimi-turn forms a pattern similar to other character patterns (for example, in a horizontally written form, the kanji ``ka'', whose partial ``i'' forms the same character shape as the katakana ``i'', ), when one character is divided into many two turns due to a break in the character, and when characters 2 or more -L written consecutively are judged as one character, it is difficult to obtain sufficient cutting accuracy. It was difficult. When the auxiliary identification section 6 is operated, the operation of the character extraction section 3 is stopped until the determination result is output, which has the disadvantage of slowing down the reading speed. In addition, this reduction in reading speed becomes more noticeable as the number of chisel turns that must be dealt with by the auxiliary identification section 6 increases.
本発明は上記従来の欠点を除去するため、帳票上の文字
行におりで予め定められた一定区間内に複数のノミター
ンが存在する場合は該ノクターンを順次組合せて文字切
出しを行ない、それらをすべて文字識別し、その結果か
ら、より文字らしいものを選択するように々したもので
、その目的とするところは文字ピッチの一定でない帳票
等の文書の文字を精度よくかつ高速に読取ることのでき
る文字読取方式を提供することにある。以下図面につい
て詳細に説明する。In order to eliminate the above-mentioned conventional drawbacks, the present invention, when a plurality of chisel turns exist within a predetermined interval in a character line on a form, cuts out the characters by sequentially combining the nocturnes and cuts out all of them. It is a system that identifies characters and selects those that are more likely to be characters based on the results.The purpose is to identify characters that can be read accurately and quickly in documents such as forms that do not have a constant character pitch. The objective is to provide a reading method. The drawings will be explained in detail below.
第3図乃至第8図は本発明の一実施例を示すもので、図
中、11は入力端子、12はノミターンメモl1、13
は文字切出し部、14は識別部、15は文字決定部、1
6は出力端子である。3 to 8 show an embodiment of the present invention, in which 11 is an input terminal, 12 is a chime-turn memory l1, 13
1 is a character cutting section, 14 is an identification section, 15 is a character determination section, 1
6 is an output terminal.
これを動作するには、捷ず帳票上の文字を光電変換装置
(図示せず)によりパターンデータに変換l,、これを
入力端子11を介1,て・Qターンメモリ12に一旦蓄
える。文字切出1−7部13は該ノミターンメモリ12
より第4図に示すような一行分の文字を含む行パターン
20を切出1−、これを行方向(図中、矢印X方向)に
走査していき列方向(図中、矢印Y方向)に・ぐターン
の存在する部分を黒(2准将号では「1」)、存在しな
い部分を白(2准将号では「0」)で表示したデータ(
以下、これを点列データと称す。)30を取り出す。更
に該文字切出し部12は点列データ30に基づいて後述
する処理を実行し行ノぐターン20より個別ノミターン
21を切出し、個々の個別パターン21とその切出しに
関する情報とを一対の個別データとして識別部】4に順
次送出する。識別部14は上記個別データのうち個別ノ
ミターン21のみを順次文字識別し、その識別結果(例
えば文字コー¥)と上記切出しに関する情報とを一対の
データとして文字決定部15に順次送出する。文字決定
部15は該データに後述する処理を施こし正l〜い文字
の識別結果のみを出力端子I6よりllliQ次出力す
る。To operate this, the characters on the unshuffled form are converted into pattern data by a photoelectric conversion device (not shown), and this data is temporarily stored in the Q-turn memory 12 via the input terminal 11. The character cutout 1-7 section 13 is the chisel turn memory 12.
Cut out a row pattern 20 containing one line of characters as shown in FIG. The data shows the part where the second turn exists in black ("1" for the 2nd Brigadier's rank), and the part where it does not exist in white ("0" for the 2nd Brigadier's rank).
Hereinafter, this will be referred to as point sequence data. ) Take out 30. Furthermore, the character cutting unit 12 executes a process described later based on the point sequence data 30 to cut out individual chisel turns 21 from the line-cut turns 20, and identifies each individual pattern 21 and information regarding its cutting as a pair of individual data. Section] 4 is sent sequentially. The identification unit 14 sequentially identifies the characters of only the individual nomiturns 21 of the individual data, and sequentially sends the identification result (for example, the character ``co\'') and the information regarding the cutout to the character determining unit 15 as a pair of data. The character determination unit 15 performs processing to be described later on the data and outputs only the identification results of correct characters from the output terminal I6.
文字切出し部】3における個別パターン21の切出しは
、点列データ30の先頭を開始点(以下、基準位置と称
す。)と17で予め設定された一定区間α内に存在する
黒の部分の集合(以下、これを点列の塊りと称す。)の
個数を調べ、−個の場合はその区間を一文字の個別ツク
ターンとみなし7、複数個存在する場合は連続する点列
の塊りを順次−個−Cつ増して組合わせた複数個のツク
ターンをそれぞれ一文字の個別ツクターンとみなすとと
もに該複数個の点列の塊りのうち先頭の塊りを除いた位
置を次の一定区間αの基準位置とする如くなっている。[Character Cutting Section] To cut out the individual pattern 21 in step 3, use the beginning of the point sequence data 30 as a starting point (hereinafter referred to as the reference position) and a set of black parts existing within a certain interval α preset in step 17. (Hereinafter, this is referred to as a cluster of point sequences.) If there are -, the section is considered as an individual character 7, and if there are multiple clusters, consecutive clusters of point sequences are - A plurality of tscutans combined by increasing number of -C are each regarded as an individual tscutane of one character, and the position excluding the first cluster of the plurality of point sequences is the standard of the next fixed interval α. It's like a position.
次に第5図に示すフローチャートに従って詳細に説明す
るが、図中DNOは一定区間α内の点列の塊り数Nを検
出するだめの動作が何回繰返し生じたかを表わす動作番
号、またPNOは点列の塊りの組合わせによる個別ツク
ターン(以下、これを組合わせ)ξターンと称す。)を
作成する際のツクターンの順番を示すツクターン番号で
あり、該動作番号DNOと・ξターン番号PNOは切出
i−に関する情報を構成する。Next, a detailed explanation will be given according to the flow chart shown in FIG. is an individual turn (hereinafter referred to as a combination) ξ turn by a combination of clusters of point sequences. ) is a turn number indicating the order of turns when creating a turn, and the operation number DNO and .xi. turn number PNO constitute information regarding cutting i-.
捷ず点列データ30の先頭すなわち文字切出しの開始位
置から一定区間α内に存在する点列の塊りの数Nを計数
し1、N=Gのときはその区間がスペースであればDN
O=1 、 PNO=]を付与してスペースツクターン
を識別部14へ送出し、区間が全て点列の塊りであれば
次の点列の塊りの終了の位置を検出L7、その区間を接
触文字とみな[、て強制分離を行ない、それぞれの個別
)ξターンにDNO=1. 、 PNO=1を付与して
識別部14へ送出する。またN=]のときはその区間が
一文字の個別ツクターンであるとみなしてその個別パタ
ーンVCT)NO−2、PNO=1を付与して識別部J
4へ送出する。Count the number N of clusters of point sequences that exist within a certain interval α from the beginning of the uncut point sequence data 30, that is, the starting position of character extraction, and calculate 1, and if N=G, if the interval is a space, DN
O=1, PNO=] and sends the space cut turn to the identification unit 14, and if the section is all a cluster of point sequences, detect the end position of the next cluster of point sequences L7, and select the section. is regarded as a contact character [, and forced separation is performed, and DNO=1. , and sends it to the identification unit 14 with PNO=1 assigned thereto. In addition, when N= ], the section is considered to be an individual character pattern, and the individual pattern VCT) NO-2, PNO=1 is added to the identification part J.
Send to 4.
更にN)]のときは黒列の塊りの出現順序を変えること
なく先頭から現われる点列の塊りを順次組合わせ、N個
の絹合せツクターンを作成し動作番号DNOと1からN
−iでのパターン番号PNOを付与して識別部14へ送
出する9例えばN=3のとき、点列の塊りをa、l)、
Cとすると、最初の処理ではDNO=1でPNO=1の
、oターン「a」、DNO=1でPNO=2のツクター
ンrabJ、DNO=1 でP N O= 3のノぐ
ターンrabcJの3個の切出しに関する情報付きの組
合わせ・ぐターンを識別部】4へ送出する。次に先頭の
点列の塊り「a」を除きDNO=2として処理を繰返し
、DNO=2でPNO=1のパターン「l)」、DNO
=2でPNO=2のツクターン「l)c」を識別部14
へ送出し、史に先頭の点列の塊り[1]」を除き1)N
O−2として処理を繰返し、DNO=3でPNO=1の
AターンrcJを識別部14へ送出する如くなって因る
。Furthermore, when [N)], sequentially combine the clusters of dots that appear from the beginning without changing the appearance order of the clusters of black rows, create N silk-combining turns, and use the operation number DNO and 1 to N.
9. For example, when N=3, a cluster of point sequences a, l),
C, in the first process, DNO = 1 and PNO = 1, o-turn "a", DNO = 1 and PNO = 2, tsuk-turn rabJ, DNO = 1, PNO = 3, o-turn rabcJ, 3. The combination/gram with information regarding the extraction of the individual pieces is sent to the identification unit 4. Next, the process is repeated with DNO=2 except for the first point sequence block "a", and the pattern "l)" with DNO=2 and PNO=1, DNO
= 2, the identification unit 14
1)N except for the cluster of points at the beginning [1]
The process is repeated as O-2, and A-turn rcJ with DNO=3 and PNO=1 is sent to the identification unit 14.
文字決定部15では切出しに関する情報より一定区間α
内のツクターンが一文字の個別ツクターンとみなされて
いる場合にはその識別結果をその一!ま出力し、複数個
のパターンとみなされている場合にはその複数個の組合
わせパターンの各々の識別結果の中からりジェツトを除
いて該組合わせパターンの内で坪4ノξターン幅が最も
長いものを正しい識別結果として出力するとともに該区
間内の点列の塊りをその一々ターン内に含む後続の識別
結果を排除する如くなって因る。The character determination unit 15 selects a certain interval α based on the information regarding cutting out.
If the tsukutan within is considered to be an individual tscutan of one character, the identification result is that one! If the pattern is considered to be multiple patterns, then from among the identification results of each of the multiple combination patterns, excluding the jet, among the combination patterns, the turn width is 4 square meters. The longest one is output as the correct identification result, and subsequent identification results that include a cluster of point sequences within the section within each turn are excluded.
第6図は文字決定部15での処理の詳細を示すフローチ
ャートで、切出L7/eターンに関する情報す々わち動
作番号]) N Oとパターン番号PNOから読取結果
として出力するための対象区間の組合わせツクターンで
あるか否かを判定して識別結果をバッファに格納し格納
したバッファの中から雇員の糾合わせパターンで識別で
きたものを読取結果として出力する。FIG. 6 is a flowchart showing the details of the processing in the character determining unit 15, in which information regarding the cutting L7/e turn, that is, the operation number]) is used to determine the target section for outputting as a reading result from NO and the pattern number PNO. It is determined whether or not it is a combination of patterns, the identification result is stored in a buffer, and those that can be identified based on the employee's combination pattern are output as reading results from the stored buffer.
次にm lr図の行ノ々ターン20を例にとって文字切
出しと文字決定の過程を説明する。行、oターン20の
うちのツクターン「べ」、「りJ 、 r l−Jにつ
いてはその点列データ30中の一定区間α内における点
列の塊り数が一個であるから、それぞれ−文字毎の個別
ツクターン21として切出され、その識別結果がその捷
ま出力端子】6に送出される。次のツクターン「ル」を
含む一定区間α(ここでは対象区間■と称す6 )K
は点列の塊りが2個存午するため、文字切出し部13i
d該2個のパターンを順次組合わせた個別ノミターン「
ノ」及び「ル」とその切出しに関する情報を識別部14
に送出するとともに、該対象区間■における点列の塊り
のうぢの先頭の塊シ「ノ」を除いた位置を次の対象区間
■の基準位置として設定する。ここでは該対象区間■に
おいても2個の点列の塊りが検出され上記同様に組合わ
せパターンとその切出しに関する情報が送出され、見、
下対象区間■、■においても同様となる。識別部14で
は対象区間■のパターン「ノ」に対して「ノ」や「1」
などの文字を識別結果として出力し、パターン「ル」に
対して「ル」の文字を出力する。文字決定部15ではパ
ターン幅が最も太きくて識別結果の確度が高いもの、対
象区間■では「ル」を読取結果として出力し、同時にパ
ターン「し」を含む対象区間■の識別結果を排除し、対
象区間■の識別結果から次の文字決定を行なう。該対象
区間■の識別結果からは「化」の文字が読取結果として
出力され、次の区間■は排除される。以下のパターン「
を」。Next, the process of character extraction and character determination will be explained using the line turn 20 of the mlr diagram as an example. For the tsukturns ``be'', ``riJ, r l-J'' in the row and o-turns 20, the number of clusters of point sequences within a certain interval α in the point sequence data 30 is one, so each - character The identification result is sent to the cutout output terminal 6. A certain section α (herein referred to as target section 6) that includes the next cut turn "ru" is cut out as an individual cut turn 21.
Since there are two clusters of dot sequences, the character cutting part 13i
dIndividual chisel turns that sequentially combine the two patterns
The identification unit 14 collects information regarding ``ノ'' and ``ru'' and their extraction.
At the same time, the position excluding the first block ``NO'' of the point sequence block U in the target section (2) is set as the reference position of the next target section (2). Here, a cluster of two point sequences is also detected in the target section (■), and information regarding the combination pattern and its extraction is sent out in the same way as above.
The same applies to the lower target sections ■ and ■. The identification unit 14 distinguishes “No” and “1” from the pattern “No” in the target section ■.
It outputs characters such as "ru" as the recognition result, and outputs the character "ru" for the pattern "ru". The character determining unit 15 outputs "ru" as the reading result in the target section ■, which has the widest pattern width and has the highest accuracy of identification result, and at the same time eliminates the identification result of the target section ■ that includes the pattern "shi". , the next character is determined based on the identification result of the target section ■. From the identification result of the target section (2), the character "C" is output as a reading result, and the next section (2) is excluded. The following pattern
of".
「進」等については上記同様に一文字として吊゛lコ取
られる。第7図は上記説明した第4図の行パターンの切
出し、識別、文字決定処理の実行のようすを示したもの
で、また第8図はその処理の流れを示したものである。Regarding ``Shin'' and the like, the ``l'' is taken as a single character in the same way as above. FIG. 7 shows how the line pattern cutout, identification, and character determination processing of FIG. 4 explained above is executed, and FIG. 8 shows the flow of the processing.
このように上記実施例によれば、一定区間σ内の点列の
塊り数に基づいて一文字のパターンか、そうでないかを
区別するようになしたため、−文字として切出す区間と
複数の組合わせパターンを構成すべき区間とを確実に区
別することができ、1だ複数個の点列の塊りが一定区間
α内に存在した場合は先頭の塊りを除いた位置を次の区
間の基準位置となしたため、考え得る全ての組合わせパ
ターンを取出すことができ、読取精度を上げることがで
きる。まだ文字切出し部では点列の塊9数に従って機械
的に・ξターンを切出すのみでよいから従来例の如く補
助識別部の識別結果を待つ必要がなく、この処理全体を
パイプライン構成とすることもでき、処理の高速化がは
かれる。In this way, according to the above embodiment, it is possible to distinguish between a single character pattern and a non-character pattern based on the number of clusters of point sequences within a certain interval σ. It is possible to reliably distinguish between the sections that should constitute the matching pattern, and if a cluster of one or more point sequences exists within a certain interval α, the position excluding the first cluster can be used to Since it is set as a reference position, all possible combination patterns can be extracted, and reading accuracy can be improved. Since the character cutting section only needs to mechanically cut out ξ turns according to the number of clusters of 9 points, there is no need to wait for the identification result of the auxiliary identification section as in the conventional case, and this entire process is configured as a pipeline. You can also speed up the processing.
以上説明したように本発明によれば、帳宗上の文章を走
査光電変換し得られた文字行のノミターンから一文字ず
つ切出して文字認識を行なう文字読取方式において、文
字行上の予め定められた一定区間内に存在する点列の塊
りの個数を調べ、−個の場合はその区間を一文字のパタ
ーンとみなして切出し、被数個の場合は該点列の塊りを
順次適宜に組合わせた複数の組合わせパターンをそれぞ
れ一文字のパターンとみなして切出し、該切出しだパタ
ーンとその切出しに関する情報を出力する切出し工程と
、該切出しだ・ξターンの識別結果とその切出しに関す
る情報とより一文字のパターンとみなされている場合は
その識別結果をそのまま出力し、被数個の・ξターンと
みガされている場合はその複数の組合わせノミターンの
各々の識別結果の中から最もパターン幅の長い組合わせ
パターンに対応する識別結果を出力する文字決定工程と
をSするため、分離文字や切れが生じた文字を含み文字
ピッチが一定でない文書からの文字切出しを複雑々識別
や判定を必要とすることなく、一義的に行なうことがで
き処理の高速化がはかれるとともに、文字の一部が他の
文字と同様な場合であっても正しく読取ることができ、
また複数個の点列の塊りが一定区間内に存在する場合に
連続する点列の塊りを順次−個ずつ増して組合わせたノ
ミターンをそれぞれ一文字のパターンとみなして切出す
とともに該複数個の点列の塊りのうち先頭の塊りを除い
た位置を次の一定区間の基準位置となしたものでは考え
得る全ての組合わせパターンを取出すことができ読取精
度を上げることができ、従って読取対象を拡大できる等
の利点がある。As explained above, according to the present invention, in a character reading method in which character recognition is performed by cutting out each character from the chisel turn of a character line obtained by scanning and photoelectrically converting a text on a book, a predetermined Check the number of clusters of point sequences that exist within a certain interval, and if there are - pieces, consider that interval as a pattern of one character and cut it out, and if there are a number of digits, combine the clusters of the point sequence sequentially as appropriate. A cutting step in which each of the plurality of combination patterns is regarded as a pattern of one character and is cut out, and the cutout pattern and information related to the cutout are outputted; If it is regarded as a pattern, the identification result is output as is, and if it is recognized as a pattern, the set with the longest pattern width is output from among the identification results of the multiple combinations of nomiturns. In order to perform a character determination process that outputs identification results corresponding to matching patterns, it is necessary to perform complex identification and judgment on character extraction from a document that includes separated characters and cut characters and whose character pitch is not constant. It can be done unambiguously, speeding up processing, and even if some of the characters are similar to other characters, they can be read correctly.
In addition, when a plurality of clusters of point sequences exist within a certain interval, the number of consecutive clusters of dot sequences is sequentially increased by - pieces, and the combined chimiturns are regarded as one character pattern and cut out, and the plurality of clusters are By setting the position excluding the first cluster of points in the cluster as the reference position for the next certain section, all possible combination patterns can be extracted and reading accuracy can be improved. This has advantages such as being able to expand the range of objects to be read.
図面は本発明の説明に供するもので、第1図は従来の文
字読取装置を示すブロック図、第2図は従来の他の文字
読取装置を示すブロック図、第3図は本発明方式を適用
した文字読取装置の一実施例を示すブロック図、第4図
ハ行パターン及びその点列データの一例を示す説明図、
第5図は文字切出部13のフローチャート、第6図は文
字決定部15のフローチャート、第7図は第4図の行パ
ターンに対する切出し、識別、文字決定処理の実行のよ
うすを示す説明図、第8図は第7図の処理の流れを示す
説明図である。
11・・・入力端子、12・・・パターンメモリ、13
・・・文字切出し部、14・・・識別部、15・・・文
字決定部、16・・・出力端子
特許出願人 日本電信電話公社
代理人 弁理士 吉 1)精 孝
第1図
第2図
第3図
第4図
379−
第5図
−J−統r山王書(自発)
昭和59年 3月 7日
特許庁長官 若 杉 和 夫 殿
1事件の表示
昭和57年特許願第222489号
2発明の名称
文字読取方式
3補正をする者
事件との関係 特許出願人
住 所 東京都千代田区内幸町1丁目1番6号名 称
(422)日本電信電話公社
代表者 真藤 恒
4代理人 〒105 電(03150B−9866自
発
6?1n正の対象
1図 面」
7補1■の内賽
別紙のとおり
7補正の内容
(1)図中、第5図を別紙のとおり補正する。The drawings are for explaining the present invention, and FIG. 1 is a block diagram showing a conventional character reading device, FIG. 2 is a block diagram showing another conventional character reading device, and FIG. 3 is a block diagram showing a conventional character reading device. A block diagram showing an example of a character reading device according to the present invention; FIG.
5 is a flowchart of the character extraction section 13, FIG. 6 is a flowchart of the character determination section 15, and FIG. 7 is an explanatory diagram showing how the extraction, identification, and character determination processing is executed for the line pattern of FIG. 4. FIG. 8 is an explanatory diagram showing the flow of the process shown in FIG. 7. 11... Input terminal, 12... Pattern memory, 13
...Character cutting section, 14...Identification section, 15...Character determination section, 16...Output terminal Patent applicant Nippon Telegraph and Telephone Public Corporation agent Patent attorney Yoshi 1) Takashi Sei Figure 1 Figure 2 Figure 3 Figure 4 379- Figure 5 - J-R Sannosho (spontaneous) March 7, 1980 Director-General of the Patent Office Kazuo Wakasugi 1 Display of case 1989 Patent Application No. 222489 2 Invention Relationship with the name character reading method 3 amendment case Patent applicant address 1-1-6 Uchisaiwai-cho, Chiyoda-ku, Tokyo Name
(422) Nippon Telegraph and Telephone Public Corporation Representative Tsune Shindo 4 Agent 105 Telephone (03150B-9866)
Contents of 7th Amendment (1) Figure 5 of the drawings will be amended as shown in the attached sheet.
Claims (2)
、oターンから一文字ずつ切出して文字認識を行なう文
字読取方式において、文字行上の予め定められた一定区
間内に存在する点列の塊りの個数を調べ、−個の場合は
その区間を一文字のパターンとみなして切出し、複数個
の場合は該点列の塊りを順次適宜に組合わせた複数の組
合わせノミターンをそれぞれ一文字のノミターンとみな
して切出し、該切出したノミターンとその切出しに関す
る情報を出力する切出し工程と、該切出したノミターン
の識別結果とその切出しに関する情報とより一文字のノ
ミターンとみなされている場合はその識別結果をそのま
ま出力し、複数個のパターンとみなされている場合はそ
の複数の組合わせパターンの各々の識別結果の中から最
もパターン幅の長い組合わせノミターンに対応する識別
結果を出力する文字決定工程とを廟することを特徴とす
る文字読取方式。(1) In a character reading method that performs character recognition by cutting out each character from an O-turn in a character line obtained by scanning and photoelectrically converting text on a document, points that exist within a predetermined interval on the character line Check the number of clusters in the sequence, and if there are -, consider the interval as a single character pattern and cut it out, and if there are multiple clusters, create multiple combinations of nomiturns by suitably combining the clusters in the sequence. A cutting step in which the chisel turn is regarded as one character and is cut out, and the cut out chisel turn and information about the cut out are outputted; and the identification result of the cut out chisel turn and the information about the cut out, if it is considered to be one character chisel turn, the identification. A character determination step that outputs the result as it is, and if it is considered to be a plurality of patterns, outputs the identification result corresponding to the combination nomiturn with the longest pattern width from among the identification results of each of the plurality of combination patterns. A character reading method characterized by the use of .
の、oターンから一文字ずつ切出して文字認識を行なう
文字読取方式において、文字行上の予め定められた一定
区間内に存在する点列の塊りの個数を調べ、−個の場合
はその区間を一文字のノミターンとみなして切出し、複
数個の場合は連続する点列の塊りを順次−個ずつ増して
組合わせたノミターンをそれぞれ一文字のノミターンと
みなして切出すとともに該複数個の点列の塊りのうち先
頭の塊りを除いた位置を次の一定区間の基準位置とし、
該切出しだ、oターンとその切出しに関する情報を出力
する切出し工程と、該切出しだノミターンの識別結果と
その切出しに関する情報とより一文字のノミターンとみ
なされている場合はその識別結果をそのま捷出力し、複
数個のノミターンとみなされている場合はその複数の組
合わせノミターンの各々の識別結果の中から最もパター
ン幅の長い組合わせノミターンに対応する識別結果を出
力する文字決定工程とを消寸5ことを特徴とする文字読
取方式。(2) In a character reading method that performs character recognition by cutting out each character from the O-turn in a character line obtained by scanning and photoelectrically converting text on a form, the characters that exist within a predetermined interval on the character line Check the number of clusters of point sequences, and if there are - pieces, consider that section as a chisel turn for one character and cut it out, and if there are multiple clusters, sequentially increase the clusters of consecutive point sequences by - pieces and combine them to create a chisel turn. Each of them is treated as a chisel turn of one character and cut out, and the position excluding the first cluster of the plurality of point sequences is used as the reference position of the next certain section,
A cutting step that outputs the cutout, o-turn, and information regarding the cutting, and the identification result of the cutout chisel turn and the information regarding the cutting, and if it is considered to be a chisel turn of one character, the identification result is output as is. However, if a plurality of chisel turns are considered, the character determination step outputs the identification result corresponding to the combination chisel turn with the longest pattern width among the identification results of each of the plurality of combination chisel turns. A character reading method characterized by 5 things.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP57222489A JPS59112367A (en) | 1982-12-18 | 1982-12-18 | Reading method of character |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP57222489A JPS59112367A (en) | 1982-12-18 | 1982-12-18 | Reading method of character |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| JPS59112367A true JPS59112367A (en) | 1984-06-28 |
| JPH0210472B2 JPH0210472B2 (en) | 1990-03-08 |
Family
ID=16783225
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP57222489A Granted JPS59112367A (en) | 1982-12-18 | 1982-12-18 | Reading method of character |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPS59112367A (en) |
Cited By (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS6195481A (en) * | 1984-10-17 | 1986-05-14 | Hitachi Ltd | Pattern extraction and recognition method |
| JPS63223890A (en) * | 1987-03-12 | 1988-09-19 | Toshiba Corp | Drawing reader |
| JPH01303586A (en) * | 1988-05-31 | 1989-12-07 | Ricoh Co Ltd | How to cut out characters |
| JP2013047887A (en) * | 2011-08-29 | 2013-03-07 | Fuji Xerox Co Ltd | Image processor and image processing program |
-
1982
- 1982-12-18 JP JP57222489A patent/JPS59112367A/en active Granted
Cited By (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS6195481A (en) * | 1984-10-17 | 1986-05-14 | Hitachi Ltd | Pattern extraction and recognition method |
| JPS63223890A (en) * | 1987-03-12 | 1988-09-19 | Toshiba Corp | Drawing reader |
| JPH01303586A (en) * | 1988-05-31 | 1989-12-07 | Ricoh Co Ltd | How to cut out characters |
| JP2013047887A (en) * | 2011-08-29 | 2013-03-07 | Fuji Xerox Co Ltd | Image processor and image processing program |
Also Published As
| Publication number | Publication date |
|---|---|
| JPH0210472B2 (en) | 1990-03-08 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| GB2208735A (en) | Character recognition system | |
| US4850026A (en) | Chinese multifont recognition system based on accumulable stroke features | |
| JPS5827551B2 (en) | Online handwritten character recognition method | |
| JPH076206A (en) | Automatic sorting device of character | |
| US5708731A (en) | Pattern matching method with pixel vectors | |
| JPS60153574A (en) | Character reading system | |
| JPS59112367A (en) | Reading method of character | |
| US4887301A (en) | Proportional spaced text recognition apparatus and method | |
| JPS60153575A (en) | Character reading system | |
| JPS59158482A (en) | Character recognizing device | |
| JPS62121589A (en) | Character segmenting system | |
| JPS6139175A (en) | Optical character reading device | |
| JPS6343788B2 (en) | ||
| JPH02230484A (en) | Character recognizing device | |
| KR100241447B1 (en) | English writing/number recognition method using outline information | |
| JPH0259504B2 (en) | ||
| JPH06162263A (en) | Device and method for recognizing character | |
| JPH0969139A (en) | Optical character reading method and its device | |
| JPH10154207A (en) | Method for segmenting character, and device therefor | |
| JPH0378892A (en) | Recognizing device for tabular document | |
| JP2000029982A (en) | Character recognition device and character recognition result output method | |
| JPH0934992A (en) | On-line handwritten character string segmenting device | |
| KR19990052967A (en) | Korean Recognition Method Using Window and Projection Information | |
| JPH08243506A (en) | Address reading device and method | |
| JPH0392989A (en) | Character recognizer |