JPH0362285A - Document reading system - Google Patents
Document reading systemInfo
- Publication number
- JPH0362285A JPH0362285A JP1198774A JP19877489A JPH0362285A JP H0362285 A JPH0362285 A JP H0362285A JP 1198774 A JP1198774 A JP 1198774A JP 19877489 A JP19877489 A JP 19877489A JP H0362285 A JPH0362285 A JP H0362285A
- Authority
- JP
- Japan
- Prior art keywords
- character
- underline
- information
- character line
- signal
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Pending
Links
- 238000000926 separation method Methods 0.000 claims abstract description 16
- 238000001514 detection method Methods 0.000 claims description 20
- 238000010586 diagram Methods 0.000 abstract description 13
- 230000010365 information processing Effects 0.000 abstract description 4
- 230000006870 function Effects 0.000 abstract description 3
- 230000019771 cognition Effects 0.000 abstract 1
- 238000000034 method Methods 0.000 description 13
- 238000000605 extraction Methods 0.000 description 9
- 239000000284 extract Substances 0.000 description 2
- 230000010354 integration Effects 0.000 description 2
- 238000007781 pre-processing Methods 0.000 description 1
- 238000003672 processing method Methods 0.000 description 1
Landscapes
- Character Input (AREA)
- Character Discrimination (AREA)
- Document Processing Apparatus (AREA)
Abstract
Description
【発明の詳細な説明】
(産業上の利用分野)
本発明は文書読取方式に関し、特に文書の自動読取およ
び自WJllI集に有益な文書読取方式に関する。DETAILED DESCRIPTION OF THE INVENTION (Field of Industrial Application) The present invention relates to a document reading method, and particularly to a document reading method useful for automatic document reading and for own WJllI collection.
(従来の技術)
情報処理システムの多様化に伴い、様々なデータ入力方
法が要求されており、文字認識技術も有力なデータ入力
方法として実用化が進められている。しかし、現、在の
文字認識技術は、文字の読み取り性能の点で人間に比べ
てはるかに劣っており、種々の改良が行われている。改
良の方向の一つとしては個別文字の認識性能を向上させ
ることがあげられるが、もう一つの方向は文書単位の読
み取り能力の改良である。(Prior Art) With the diversification of information processing systems, various data input methods are required, and character recognition technology is also being put into practical use as a powerful data input method. However, current character recognition technology is far inferior to humans in terms of character reading performance, and various improvements are being made. One direction of improvement is to improve the recognition performance of individual characters, and another direction is to improve the reading ability of individual documents.
近年、日本語ワードプロセッサの普及にともない、オフ
ィスにおいて様々な印刷文書が作成されるようになって
きている9人間はその様々な文書からテキスト部、図表
等を容易に判断して読みとっている。In recent years, with the spread of Japanese word processors, a variety of printed documents have been created in offices.9 People can easily determine and read text parts, figures, tables, etc. from the various documents.
しかし、現状の文字認識装置(以後OCR:0ptic
al character Readerの略:と呼ぶ
)ではあらかじめ文書の形式を定めるか、またはテキス
トの部分を操作者が指定し、指定されたテキスト部だけ
を読みとり、文字だけをコード化して電子化文書を作成
するものに限られている。However, the current character recognition device (hereinafter referred to as OCR: 0ptic)
(abbreviation for character reader)), the format of the document is determined in advance, or the operator specifies the text part, reads only the specified text part, and encodes only the characters to create an electronic document. limited to things.
従って、文書の重要な情報である下線を情報処理システ
ムに入力するためにはコード化された電子化文書を編集
して下線情報を付加するか、あるいは下線の引かれた箇
所は操作者が別に入力する必要が生じている。Therefore, in order to input underlines, which are important information in a document, into an information processing system, the coded electronic document must be edited and the underline information added, or the underlined parts must be manually edited by the operator. I need to enter it.
(発明が解決しようとする課l1l)
上述の従来の文書読取方式は操作者にとって煩わしい操
作が必要になっている。(Problems to be Solved by the Invention l1l) The above-described conventional document reading method requires cumbersome operations for the operator.
そこで、本発明の目的は、文字分離と同時に下線情報も
抽出して、下線情報を認識結果に付加することによって
、電子化文書に下線を付加するためには操作者が編集ま
たは一部入力をしなければならないという問題を解決し
5.操作者の負担を軽減し、文書の自動読取機能の向上
を図るための文書読取方式を提供することにある。Therefore, an object of the present invention is to extract underline information at the same time as character separation and add the underline information to the recognition result, so that an operator can edit or input part of the information in order to add an underline to an electronic document. 5. Solve the problem that you have to do. An object of the present invention is to provide a document reading method that reduces the burden on an operator and improves automatic document reading functions.
(課題を解決するための手段)
前述の課題を解決するために本発明による文書読取方式
は、二次元格子状の配列として与えられる文書画像を格
納する文書画像記憶部と、前記文書画像記憶部から前記
文書画像信号を読み込み、下線も含んだ文字行の分離を
行う文字行処理部と、前記文字行処理部で分離された各
文字行画像を格納する文字行画像記憶部と、前記文字行
画像から文字行領域と下線部領域を検出する下線検出部
と、前記文字行画像から文字分離を行って個別文字パタ
ーンを抽出するとともに各文字パターンに下線が引かれ
ているか否かを検出し、文字パターンおよび下線の有無
を含む文字情報を出力する文字分離部と、前記各文字パ
ターンを認識処理して認識結果および文字情報を出力す
る文字認識部とを備えて構成される。(Means for Solving the Problems) In order to solve the above-mentioned problems, a document reading method according to the present invention includes a document image storage unit that stores document images provided as a two-dimensional grid array, and a document image storage unit that stores document images provided as a two-dimensional grid array. a character line processing unit that reads the document image signal from the character line and separates character lines including underlines; a character line image storage unit that stores each character line image separated by the character line processing unit; and a character line image storage unit that stores each character line image separated by the character line processing unit; an underline detection unit that detects a character line area and an underlined area from an image, and performs character separation from the character line image to extract individual character patterns, and detects whether each character pattern is underlined; The apparatus includes a character separation section that outputs character information including a character pattern and the presence or absence of an underline, and a character recognition section that performs recognition processing on each of the character patterns and outputs a recognition result and character information.
(作用〉
以下、本発明の原理、作用について図面を用いて説明す
る。(Operation) The principle and operation of the present invention will be explained below with reference to the drawings.
第2図は下線を含んだ印刷文書画像を示している。従来
はこの下線は不用な情報として画像処理の前処理によっ
て削除されていた。その−例を示しているのが第3図(
b)である。FIG. 2 shows a printed document image that includes an underline. Conventionally, this underline was considered unnecessary information and was deleted by preprocessing of image processing. An example of this is shown in Figure 3 (
b).
第3図(a)に示すように文書画像から水平方向の投影
情報を求めることによって文書中の文字行の位置を求め
る手がかりが得られる。投影情報中の大きなピークは文
字行を示しており、幅の狭い比較的低いピークは下線の
存在を示している。As shown in FIG. 3(a), by obtaining horizontal projection information from the document image, clues for determining the position of character lines in the document can be obtained. Large peaks in the projection information indicate character lines, and narrow, relatively low peaks indicate the presence of an underline.
従来は、ピークの幅を検出することによって、第3図(
b)に示すように、下線を除いて文字行を抽出している
。これに対して、本発明においては、第3図(c)に示
すように下線を含んだ状態で文字行を抽出する。Conventionally, by detecting the width of the peak, the method shown in Figure 3 (
As shown in b), character lines are extracted without underlines. In contrast, in the present invention, character lines are extracted with underlines included, as shown in FIG. 3(c).
下線部領域の検出方法は、第4図に示すように、下線を
含んだ文字行から投影のピークの幅から文字行の領域S
と下線の領域Tとを分離することで実現できる。The method for detecting the underlined area is as shown in Fig. 4.
This can be realized by separating the underlined region T from the underlined region T.
第5図は、各個別文字と下線との関係を示している。下
線を含まない文字行Sから、例えば、水平軸上への投影
を用いて文字分離を行い、分離位置を求め、さらにその
分離位置を用いて下線の領域Tも分離する0分離された
下線の領域の個々の領域について下線の有無の判定を行
って、下線の引かれた文字を判別する。第5図ではA、
B、Cの3文字が下線の引かれた文字である。この後、
下線を含まない文字パターンの認識を行い、認識結果と
下線の有無等を含む文字情報を出力する。FIG. 5 shows the relationship between each individual character and the underline. From a character line S that does not contain an underline, for example, perform character separation using projection onto the horizontal axis, find the separation position, and use that separation position to also separate the underline area T. The presence or absence of an underline is determined for each area to determine which character is underlined. In Figure 5, A,
The three characters B and C are underlined. After this,
Recognizes character patterns that do not include underlines, and outputs recognition results and character information including the presence or absence of underlines.
この結果文書の文字コードのみならず、下線の情報が情
報処理システムに自動入力できる。As a result, not only the character code of the document but also the underlined information can be automatically input into the information processing system.
(実施例〉 次に本発明について図面を参照して説明する。(Example> Next, the present invention will be explained with reference to the drawings.
第1図は本発明の一実施例を示すブロック図である。FIG. 1 is a block diagram showing one embodiment of the present invention.
文書画像記憶部1は、画像入力手段等によって電気信号
に変換された文書画像を格納する。The document image storage unit 1 stores document images converted into electrical signals by an image input means or the like.
文字行処理部2は、文書画像記憶部1から信号11とし
て文書画像信号を読み込み、作用の項で説明したように
下線を含んだ文字行領域を抽出し、文字行画像信号12
として出力する。その詳細は後述する。The character line processing unit 2 reads the document image signal as the signal 11 from the document image storage unit 1, extracts the character line area including the underline as explained in the operation section, and converts it into the character line image signal 12.
Output as . The details will be described later.
文字行画像記憶13は、文字行画像信号12を格納する
。The character line image storage 13 stores the character line image signal 12.
下線検出部4は、文字行画像記憶部3から文字行画像を
信号13として読み込み、下線の有無を調べて、下線が
存在するときには作用の項で説明した文字行と下線の分
離を行い、下線の位置を検出し、下線の始端と終端の位
置座標を信号14として出力し、下線が存在しないとき
には下線が存在しないことを示すフラッグ信号を出力す
るもので、詳細は後述する。The underline detection unit 4 reads a character line image from the character line image storage unit 3 as a signal 13, checks whether there is an underline, and when an underline is present, separates the character line and the underline as explained in the section of the operation, and detects the underline. It detects the position of the underline, outputs the position coordinates of the start and end of the underline as a signal 14, and when the underline does not exist, outputs a flag signal indicating that the underline does not exist, the details of which will be described later.
文字分離部5は、文字行画像記憶M3から文字行画像信
号13を読み込み、下線検出部4から下線の始端および
終端位置信号14を読み込み、作用の項で説明したよう
に文字分離を行い、文字分離結果として得られる位rI
l座標で下線の領域の分離を行って分離された個々の領
域に下線が含まれているか否かを判定し、各個別文字パ
ターンと下線の有無の情報、位置情報等を信号15とし
て出力するもので、詳細は後述する。The character separation unit 5 reads the character line image signal 13 from the character line image storage M3, reads the underline start and end position signals 14 from the underline detection unit 4, performs character separation as explained in the operation section, and separates the characters. The position rI obtained as a result of separation
Separate the underlined areas using the l coordinate, determine whether or not each separated area contains an underline, and output each individual character pattern, information on the presence or absence of an underline, position information, etc. as a signal 15. The details will be explained later.
文字認識部6は、信号15として送られてきた各個別文
字パターンと下線の有無の情報、位置情報等のうちの個
別文字パターンに対して文字認識処理を行い、得られた
文字認識結果または候補字種と前記下線の有無の情報、
位置情報等とを信号16として出力するもので、公知の
技術で実現できる。尚、個別文字認識処理方法は種々の
ものが提案されており、本発明においては特定のものに
限るものではない。The character recognition unit 6 performs character recognition processing on each individual character pattern sent as a signal 15, information on the presence or absence of underlining, position information, etc., and generates character recognition results or candidates. Information on the character type and the presence or absence of the underline,
It outputs position information and the like as a signal 16, and can be realized using known technology. Note that various individual character recognition processing methods have been proposed, and the present invention is not limited to any particular method.
第6図は文字行処理部2の構成の一例を詳細に説明する
ための図である。FIG. 6 is a diagram for explaining in detail an example of the configuration of the character line processing section 2. As shown in FIG.
水平方向投影抽出部21は、文書画像信号11を入力し
て水平方向に走査して得られる水平方向投影ヒストグラ
ムを信号121として出力する。The horizontal projection extraction unit 21 inputs the document image signal 11 and outputs a horizontal projection histogram obtained by scanning in the horizontal direction as a signal 121.
投影記憶部22は、信号121として出力された投影ヒ
ストグラムを格納する。The projection storage unit 22 stores the projection histogram output as the signal 121.
極大・極小検出部23は、投影記憶部22から信号12
2として水平方向投影ヒストグラムを読み込み、大局的
な範囲内の極大値・極小値およびそれらの位置座標を検
出し、極大値・極小値およびそれらの位置座標の上から
下への系列を信号123として順次出力するもので公知
技術で実現できる。The local maximum/minimum detection unit 23 receives the signal 12 from the projection storage unit 22.
2, read the horizontal projection histogram, detect local maximum values, local minimum values, and their position coordinates within a global range, and detect the local maximum value, local minimum value, and their position coordinates from top to bottom as a signal 123. It outputs sequentially and can be realized using known technology.
ピーク幅検出部24は、信号123を読み込み、極大値
・極小値の系列からピークを検出し、各ピークの上端の
座標および下端の座標からピークの幅を求め、各ピーク
の上端、下端の座標および幅を上から順に信号124と
して出力するもので公知技術で実現できる。The peak width detection unit 24 reads the signal 123, detects a peak from the series of maximum values and minimum values, calculates the width of the peak from the coordinates of the upper end and the coordinate of the lower end of each peak, and calculates the coordinates of the upper end and lower end of each peak. and the width are outputted as a signal 124 in order from the top, and can be realized by known technology.
文字行検出部25は、各ピークの上端、下端の座標およ
びピークの幅を信号124として読み込み、幅の大きい
ピークの下に小さいピークが隣接するときには大きいピ
ークの上端座標と小さいピークの゛下端座標を文字行の
上端座標・下端座標と決定して信号125として出力し
、大きいピークの下に大きいピークが隣接するときは上
の大きいピークの上端・下端の座標を文字行の上端Ji
JI(、下端Jijl[と決定して信号125として出
力し、この処理を上から下までのすべてのピークについ
て順に行う、尚、大きいピークか小さいピークの判定は
しきい値処理でよい。The character line detection unit 25 reads the coordinates of the upper and lower ends of each peak and the width of the peak as a signal 124, and when a small peak is adjacent to a large peak, the character line detection unit 25 reads the upper end coordinates of the larger peak and the lower end coordinates of the smaller peak. are determined as the upper and lower coordinates of the character line and output as a signal 125, and when a large peak is adjacent to a large peak, the coordinates of the upper and lower ends of the upper large peak are determined as the upper and lower coordinates of the character line.
JI(, lower end Jijl[) is determined and output as a signal 125, and this process is performed sequentially for all peaks from top to bottom. Note that threshold processing may be used to determine whether a peak is a large peak or a small peak.
文字行抽出部2゛6は、文書画像信号11と各文字行の
上端座標・下端座標を示す信号125を読み込み、各文
字行の上端座標と下端座標で定まる矩形領域を文字行画
像信号12として順次出力するもので実現できる。The character line extraction unit 2'6 reads the document image signal 11 and the signal 125 indicating the upper and lower coordinates of each character line, and extracts a rectangular area defined by the upper and lower coordinates of each character line as the character line image signal 12. This can be achieved by sequential output.
第7図は下線検出部4の構成の一例を詳細に説明するた
めの図である。FIG. 7 is a diagram for explaining in detail an example of the configuration of the underline detection section 4. As shown in FIG.
水平方向投影抽出部41は、文字行画像信号13を読み
込み、水平方向に走査して得られる水平方向投影ヒスト
グラムを信号141として出力するもので、水平方向抽
出部21と同様の機能をもつ。The horizontal projection extraction section 41 reads the character line image signal 13, scans it in the horizontal direction, and outputs a horizontal projection histogram obtained as a signal 141, and has the same function as the horizontal extraction section 21.
投影記憶部42は、信号141として出力された投影ヒ
ストグラムを格納する。The projection storage unit 42 stores the projection histogram output as the signal 141.
ピーク検出部43は、極大・極小検出g23とピーク幅
検出部24と同様の処理を行って小さいピークの有無を
調べ、小さいピークが存在するときには上端座標および
下端座標を信号143として出力し、小さいピークが存
在しないときには存在しないことを示すフラッグを信号
14として出力する。The peak detection unit 43 performs the same processing as the local maximum/minimum detection g23 and the peak width detection unit 24 to check for the presence or absence of a small peak, and when a small peak exists, outputs the upper end coordinate and the lower end coordinate as a signal 143, and detects a small peak. When the peak does not exist, a flag indicating that the peak does not exist is outputted as a signal 14.
第8図は文字分離部5の構成の一例を詳細に説明するた
めの図である。FIG. 8 is a diagram for explaining in detail an example of the configuration of the character separating section 5. As shown in FIG.
領域分離部51は、文字行画像信号13と下線の有無ま
たは位置座標に関する信号14を読み込み、下線が存在
するときには文字行領域と下線領域とに分割して、文字
行領域画像信号150と下線領域画像信号154を出力
し、下線が存在しないときには文字行画像全体を文字行
領域画像信号150として出力し、画素の値がすべてク
リアされた画像を下線領域画像信号154として出力す
る。The area separation unit 51 reads the character line image signal 13 and the signal 14 regarding the presence or absence of an underline or the position coordinates, and when an underline exists, divides it into a character line area and an underline area, and separates the character line area image signal 150 and the underline area. An image signal 154 is output, and when there is no underline, the entire character line image is output as a character line area image signal 150, and an image in which all pixel values are cleared is output as an underline area image signal 154.
文字行領域画像記憶部52は、文字行領域画像信号15
0を格納する。The character line area image storage unit 52 stores the character line area image signal 15.
Store 0.
文字切り出し部53は、文字行領域画像記憶部52から
文字行領域画像信号151を読み出し、個別文字パター
ンの切り出し処理を行い、個別文字パターンを信号15
2として、個別文字パターンの左端および右端の位置座
標を信号153として出力する。The character cutting unit 53 reads the character line area image signal 151 from the character line area image storage unit 52, performs a process of cutting out individual character patterns, and converts the individual character pattern into a signal 15.
2, the position coordinates of the left end and right end of the individual character pattern are output as a signal 153.
文字パターン記憶部54は、個別文字パターンを格納す
る。The character pattern storage section 54 stores individual character patterns.
文字パターン座標記憶部55は、信号153で表される
個別文字パターンの左端および右端の位置座標を格納す
る。The character pattern coordinate storage section 55 stores the position coordinates of the left end and right end of the individual character pattern represented by the signal 153.
下線領域画像記憶部56は、下線領域画像を記憶する。The underlined area image storage unit 56 stores underlined area images.
下線情報検出部57は、下線領域画像記憶部56から信
号155として下線領域画像を読み込み、文字パターン
座標記憶部55から信号156として各個別文字パター
ンの左端と右端の位置座標を読み込み、第4図で示した
ように各個別文字パターンに対応する位置に下線が引か
れているか否かを判定し、下線の有無を信号157とし
て出力する。The underline information detection unit 57 reads the underline area image as a signal 155 from the underline area image storage unit 56, reads the position coordinates of the left end and right end of each individual character pattern as a signal 156 from the character pattern coordinate storage unit 55, and reads the position coordinates of the left end and right end of each individual character pattern as a signal 156. As shown in , it is determined whether or not an underline is drawn at a position corresponding to each individual character pattern, and the presence or absence of an underline is outputted as a signal 157.
文字パターン情報統合部58は、文字パターン記憶部5
4から信号158として個別文字パターンを読み込み、
文字パターン座標記憶部55から個別文字パターンの位
置情報を信号156として読み込み、下線情報検出部5
7から下線の有無の情報を信号157として読み込み、
これらをまとめて信号15として出力するもので容易に
実現できる。The character pattern information integration unit 58 includes the character pattern storage unit 5
4 reads the individual character pattern as signal 158,
The position information of the individual character pattern is read from the character pattern coordinate storage unit 55 as a signal 156, and the underline information detection unit 5
7 reads information on the presence or absence of an underline as a signal 157,
This can be easily realized by outputting these signals together as a signal 15.
(発明の効果)
以上説明したように、本発明によれば、文字の位置情報
と認識結果のみならず、下線の有無の情報も抽出するこ
とができるので、下線を含んだ文書の自動読取の性能向
上に大きく役立つ。(Effects of the Invention) As explained above, according to the present invention, it is possible to extract not only character position information and recognition results, but also information on the presence or absence of underlines, which makes it possible to automatically read documents containing underlines. This greatly helps improve performance.
第1図は本発明の一実施例を示すブロック図、第2図は
下線を含んだ印刷文書画像の一例を示す図、第3図(a
)〜(c)は下線を含んだ文書の処理を示す図、第4図
は文字行画像の文字行領域部と下線領域部を投影ヒスト
グラムで分離できる例を示す図、第5図は文字切り出し
結果を下線領域部に適用して各個別文字パターンの下線
の有無を検出する開を示す図、第6図は文字行処理部2
の構成の一例を詳細に説明するためのブロック図、第7
図は下線検出部4の構成の一例を詳細に説明するための
ブロック図、第8図は文字分離部5の構成の一例を詳細
に説明するためのブロック図である。
1・・・文書画像記憶部、2・・・文字行処理部、3・
・・文字行画像記憶部、4・・・下線抽出部、5・・・
文字分離部、6・・・文字認識部、21・・・水平方向
投影抽出部、22・・・投影記憶部、23・・・極大・
極小検出部、24・・・ピーク幅検出部、25・・・文
字行検出部、26・・・文字行抽出部、41・・・水平
方向投影抽出部、42・・・投影記憶部、43・・・ピ
ーク検出部、51・・・領域分離部、52・・・文字行
領域画像記憶部、53・・・文字切り出し部、54・・
・文字パターン記憶部、55・・・文字パターン座標記
憶部、56・・・下線領域画像記憶部、57・・・下線
情報検出部、58・・・文字パターン情報統合部。FIG. 1 is a block diagram showing an embodiment of the present invention, FIG. 2 is a diagram showing an example of a printed document image including underlines, and FIG.
) to (c) are diagrams showing the processing of a document containing underlines, Figure 4 is a diagram showing an example in which the character line area and underline area of a character line image can be separated using a projection histogram, and Figure 5 is character extraction. A diagram showing the process of applying the result to the underlined area to detect the presence or absence of an underline in each individual character pattern.
Block diagram 7 for explaining in detail an example of the configuration of
8 is a block diagram for explaining in detail an example of the configuration of the underline detection unit 4, and FIG. 8 is a block diagram for explaining in detail an example of the configuration of the character separation unit 5. 1... Document image storage section, 2... Character line processing section, 3.
...Character line image storage section, 4...Underline extraction section, 5...
Character separation unit, 6... Character recognition unit, 21... Horizontal projection extraction unit, 22... Projection storage unit, 23... Maximum...
Minimum detection unit, 24... Peak width detection unit, 25... Character line detection unit, 26... Character line extraction unit, 41... Horizontal projection extraction unit, 42... Projection storage unit, 43 . . . Peak detection unit, 51 . . . Region separation unit, 52 .
Character pattern storage section, 55... Character pattern coordinate storage section, 56... Underline area image storage section, 57... Underline information detection section, 58... Character pattern information integration section.
Claims (1)
る文書画像記憶部と、前記文書画像記憶部から前記文書
画像信号を読み込み、下線も含んだ文字行の分離を行う
文字行処理部と、前記文字行処理部で分離された各文字
行画像を格納する文字行画像記憶部と、前記文字行画像
から文字行領域と下線部領域を検出する下線検出部と、
前記文字行画像から文字分離を行って個別文字パターン
を抽出するとともに各文字パターンに下線が引かれてい
るか否かを検出し、文字パターンおよび下線の有無を含
む文字情報を出力する文字分離部と、前記各文字パター
ンを認識処理して認識結果および文字情報を出力する文
字認識部とを備えて構成されることを特徴とする文書読
取方式。a document image storage unit that stores document images given as a two-dimensional grid array; a character line processing unit that reads the document image signal from the document image storage unit and separates character lines including underlines; a character line image storage unit that stores each character line image separated by the character line processing unit; an underline detection unit that detects a character line area and an underlined area from the character line image;
a character separation unit that performs character separation from the character line image to extract individual character patterns, detects whether or not each character pattern is underlined, and outputs character information including the character pattern and the presence or absence of an underline; , and a character recognition unit that recognizes each of the character patterns and outputs recognition results and character information.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP1198774A JPH0362285A (en) | 1989-07-31 | 1989-07-31 | Document reading system |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP1198774A JPH0362285A (en) | 1989-07-31 | 1989-07-31 | Document reading system |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| JPH0362285A true JPH0362285A (en) | 1991-03-18 |
Family
ID=16396704
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP1198774A Pending JPH0362285A (en) | 1989-07-31 | 1989-07-31 | Document reading system |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPH0362285A (en) |
-
1989
- 1989-07-31 JP JP1198774A patent/JPH0362285A/en active Pending
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| EP0481979B1 (en) | Document recognition and automatic indexing for optical character recognition | |
| JP2940936B2 (en) | Tablespace identification method | |
| EP0843277A3 (en) | Page analysis system | |
| JPH08235341A (en) | Document filing apparatus and method | |
| JPH08180068A (en) | Electronic filing equipment | |
| JPH0548510B2 (en) | ||
| JPH0362286A (en) | Document reading system | |
| JPS58197581A (en) | Method and device for recognizing character and figure | |
| JPH0564396B2 (en) | ||
| JPH01119885A (en) | Document reader | |
| JPS5949671A (en) | Optical character reader | |
| JP2722549B2 (en) | Optical character reader | |
| JP3160458B2 (en) | Character reading device and character reading method | |
| JPH03160582A (en) | Method for separating ruled line and character in document picture data | |
| JP3006294B2 (en) | Optical character reader | |
| JP3060237B2 (en) | Japanese character recognition device | |
| JP2823350B2 (en) | Multimedia input device | |
| JPH08202824A (en) | Document image recognition device | |
| JPH02187883A (en) | Document reader | |
| JP2023034823A (en) | Image processing apparatus, and control method, and program for image processing apparatus | |
| JPS63184181A (en) | Optical character recognition device | |
| JPH0728932A (en) | Image processing device | |
| JPH08190606A (en) | Optical character reader | |
| JPH0773273A (en) | Pattern cutting and recognition method and its system | |
| JPH0334081A (en) | Drawing reader |