CN1838112B - translation device, translation method - Google Patents
translation device, translation method Download PDFInfo
- Publication number
- CN1838112B CN1838112B CN2005100928181A CN200510092818A CN1838112B CN 1838112 B CN1838112 B CN 1838112B CN 2005100928181 A CN2005100928181 A CN 2005100928181A CN 200510092818 A CN200510092818 A CN 200510092818A CN 1838112 B CN1838112 B CN 1838112B
- Authority
- CN
- China
- Prior art keywords
- text
- area
- data
- translated
- text data
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Expired - Fee Related
Links
Images
Landscapes
- Machine Translation (AREA)
- Document Processing Apparatus (AREA)
Abstract
Description
技术领域 technical field
本发明涉及一种用于在将文档中包含的文本从一种语言翻译成另一种语言后生成翻译数据的技术。 The present invention relates to a technique for generating translation data after translating text contained in a document from one language to another. the
背景技术Background technique
已经提出了各种类型的翻译设备,其接收有图或无图的文档的图像数据、翻译该图像数据的文本区中包含的文本、并且生成包含翻译后文本的经翻译的文档或生成包含翻译后文本和原始图形的文档。 Various types of translation devices have been proposed which receive image data of a document with or without a picture, translate text contained in a text area of the image data, and generate a translated document containing the translated text or generate a document containing the translated text. Documentation of post text and original graphics. the
已知提出了一种技术,即采用布局分析使输入数据的文本区和图形区分开,并且识别文本区中的字符以进行翻译。然后将获得的翻译后的文本量与现存的文本区的大小进行比较,从而可以根据比较结果再形成文本区。然而,作为对文本区再形成的结果,如果不能再将图形区分配在同一页上,则将图形区分配在下一页中。因此,由于文本区和图形区分配的改变,读者可能难于阅读翻译后的文档。 It is known to propose a technique of using layout analysis to distinguish text areas and graphics of input data, and to recognize characters in the text areas for translation. Then the obtained translated text volume is compared with the size of the existing text area, so that the text area can be re-formed according to the comparison result. However, as a result of reformatting the text area, if the graphics area can no longer be allocated on the same page, the graphics area is allocated in the next page. Therefore, it may be difficult for readers to read the translated document due to the changed allocation of text area and graphics area. the
此外,由于常用的翻译设备在同一页的不同区域或不同页中输出原始文档和翻译后的文档,因而用户通常很难找到原始文本和翻译后文本之间的对应。已知提出了一种技术,即在原始文本的行间布置翻译后的文本,从而减少了用户寻找原始文本和翻译后文本之间的对应所引起的的麻烦。 In addition, since commonly used translation devices output the original document and the translated document in different regions of the same page or in different pages, it is often difficult for users to find the correspondence between the original text and the translated text. It is known to propose a technique of arranging translated text between lines of an original text, thereby reducing the trouble caused by the user to find a correspondence between the original text and the translated text. the
然而,翻译后文本中所包含的字符的数量和类型会与原始文本中的不同;结果,一行翻译后文本中包含的字符串的长度与一行原始文本所占据的长度不相等。 However, the number and type of characters contained in the translated text will be different from those in the original text; as a result, the length of the character string contained in a line of translated text is not equal to the length occupied by a line of original text. the
发明内容Contents of the invention
鉴于上述情况提出了本发明,本发明提供了一种系统,用于保留文本区和图形区的原始布局地生成包含翻译后文本部分和原始图形部分的翻译数据。此外,本发明提供了一种系统,用于使得能够生成翻译数据,通过使用该系统,用户很容易将原始文档与翻译后文档相联系,提高了可阅览性。 The present invention has been made in view of the above circumstances, and provides a system for generating translation data including a translated text portion and an original graphic portion while preserving the original layout of a text area and a graphic area. Furthermore, the present invention provides a system for enabling translation data to be generated, and by using the system, a user can easily associate an original document with a translated document, improving viewability. the
一方面,本发明提供了一种翻译装置,包括:字符识别单元,用于识别输入图像文本区中的文本数据;翻译器,用于翻译文本区中的文本数据;以及布局结构处理器,用于通过改变翻译后文本数据的字符大小来生成包含文本区翻译后文本数据和输入图像中的图形的数据,其中在由该布局结构处理器生成的数据的图像布局中保持输入图像的布局,其中所述布局结构处理器通过改变图形区的大小来生成所述数据。 In one aspect, the present invention provides a translation device, comprising: a character recognition unit for recognizing text data in a text area of an input image; a translator for translating the text data in the text area; and a layout structure processor for for generating data including translated text data of a text area and graphics in an input image by changing the character size of the translated text data, wherein the layout of the input image is maintained in the image layout of the data generated by the layout structure processor, wherein The layout structure processor generates the data by changing the size of the graphics area. the
根据本发明的实施例,可以生成其文本部分经翻译的翻译后文本与图形数据,同时保持输入数据的文本区和图形区的布局。 According to an embodiment of the present invention, it is possible to generate translated text and graphic data whose text parts are translated while maintaining the layout of the text area and the graphic area of the input data. the
另一方面,本发明提供了一种翻译方法,所述方法包括以下步骤:字符识别步骤,识别输入图像文本区中的文本数据;翻译步骤,翻译所述文本区中的文本数据;以及布局结构处理步骤,通过改变翻译后文本数据的字符大小,生成包含文本区翻译后文本数据和输入图像中的图形的数据,从而生成的数据的图像布局保持输入图像的布局,其中通过改变所述图形区的大小来生成所述数据。 In another aspect, the present invention provides a translation method comprising the steps of: a character recognition step of recognizing text data in a text area of an input image; a translation step of translating the text data in the text area; and a layout structure a processing step of generating data including the translated text data of the text area and graphics in the input image by changing the character size of the translated text data, whereby the image layout of the generated data maintains the layout of the input image, wherein by changing the graphics area to generate the data. the
根据本发明的实施例,由于翻译后文本的字符串与原始文本的字符串长度相同,且两者平行放置,所以用户可以很容易地进行原始文本和翻译后文本之间的对应。结果,大大提高了可阅览性。 According to the embodiment of the present invention, since the character strings of the translated text and the character strings of the original text have the same length and are placed in parallel, the user can easily make the correspondence between the original text and the translated text. As a result, readability is greatly improved. the
附图说明 Description of drawings
下面将根据附图对本发明的实施例进行详细说明,在附图中: Embodiments of the present invention will be described in detail below according to the accompanying drawings, in the accompanying drawings:
图1是示出了根据本发明第一实施例的翻译装置的结构的框图; Fig. 1 is a block diagram showing the structure of a translation device according to a first embodiment of the present invention;
图2是示出了根据第一实施例的有图文档翻译程序的过程的流程图; Fig. 2 is a flowchart showing the process of the document translation program according to the first embodiment;
图3的图示出了根据第一实施例的有图文档翻译程序的过程的一部分; The diagram of Fig. 3 shows a part of the process of the document translation program according to the first embodiment;
图4A至4D的图与传统技术相比较地描述了第一实施例的效果; The figure of Fig. 4A to 4D has described the effect of the first embodiment compared with conventional technology;
图5是示出了根据本发明第二实施例的有图文档翻译程序的过程的流程图; Fig. 5 is a flowchart showing the process of a document translation program according to a second embodiment of the present invention;
图6的图描述了根据另一实施例的处理; The diagram of Figure 6 describes processing according to another embodiment;
图7的图描述了根据另一实施例的处理; The diagram of Figure 7 describes processing according to another embodiment;
图8是示出了根据本发明第三实施例的图像生成装置的框图; Fig. 8 is a block diagram showing an image generating device according to a third embodiment of the present invention;
图9是示出了根据第三实施例的图像生成装置中翻译处理单元的结构的框图; 9 is a block diagram showing the structure of a translation processing unit in an image generating device according to a third embodiment;
图10是示出了由翻译处理单元进行的处理的流程图; Fig. 10 is a flowchart showing the processing performed by the translation processing unit;
图11A和11B的图示出了说明在翻译处理单元中进行的处理的图像;以及 The diagrams of Figures 11A and 11B show images illustrating the processing performed in the translation processing unit; and
图12A和12B的图示出了说明校正字符串长度的处理(即在翻译处理单元中进行的处理)的字符串。 12A and 12B are diagrams showing character strings illustrating processing of correcting the length of character strings (ie, processing performed in the translation processing unit). the
具体实施方式 Detailed ways
下面,将参照附图说明本发明的实施例。 Hereinafter, embodiments of the present invention will be described with reference to the drawings. the
第一实施例 first embodiment
图1是示出了根据本发明第一实施例的翻译装置的基本结构的框图。该翻译装置的结构与结合了扫描功能、复印功能、打印功能和传真功能的多功能机相类似。该翻译装置具有带ADF(自动送纸器)的图像读取装置1、打印装置2、通信接口3、显示单元4、操作单元5、易失性存储器6、非易失性存储器7以及控制上述各单元的CPU8。在CPU8的控制下,在该翻译装置中实现了多种功能:通过打印装置2对由图像读取装置1读取的图像进行打印来实现复印功能、将图像通过通信接口3和网络传输到相应的传真机来实现传真功能等。
FIG. 1 is a block diagram showing the basic structure of a translation apparatus according to a first embodiment of the present invention. The structure of the translation device is similar to a multifunction machine that combines scanning functions, copying functions, printing functions, and facsimile functions. This translation device has an image reading device 1 with an ADF (automatic document feeder), a
在非易失性存储器7中存储有本实施例的专用程序以及用于使得CPU8进行控制以实现此类多功能装置设有的多种功能的控制程序,所述特有程序为将从装置外扫描到的图像中包含的文本翻译成另一种语言的有图文档翻译程序,并且该专用程序生成用于输出的带有图形的翻译数据,所述翻译数据包含取代了表示原始文本的图像的翻译后文本数据。在图2的流程图中示出了有图文档翻译程序进行的典型过程。为了避免重复说明,将在本实施例后面部分提供对本实施例的操作的说明时再说明该过程的细节。 In the non-volatile memory 7, a dedicated program of this embodiment and a control program for making the CPU 8 perform control to realize various functions provided by this type of multifunctional device are stored. document translation program for translating text contained in received images into another language, and the specialized program generates for output translation data with graphics containing translations in place of images representing original text post text data. A typical process performed by a document translation program with pictures is shown in the flowchart of FIG. 2 . In order to avoid redundant descriptions, details of this process will be described later in this embodiment when a description of the operation of this embodiment is provided.
根据本实施例的翻译技术具有这样一种结构使得可以以下述方式实现有图文档翻译程序: The translation technique according to the present embodiment has such a structure that a document translation program with pictures can be realized in the following manner:
a.在图像读取装置1中,读取包含以一种语言撰写的文本和图形的文档的图像,并将该图像存储在易失性存储器6中,以便使用有图文档翻译程序对该图像进行处理并使用打印装置2将通过该处理获得的翻译后文本与图形数据作为图像输出。另选地,采用通信接口3通过传真或电子邮件,将通过该处理获得的翻译后文本与图形数据传输到需要该翻译后数据的用户;以及
a. In the image reading device 1, an image of a document containing text and graphics written in one language is read, and the image is stored in the
b.通过通信接口3接收包含以一种语言撰写的文本和图形的文档的图像,并将该图像存储在易失性存储器6中,以便使用有图文档翻译程序对该图像进行处理并使用打印装置2将通过该处理获得的翻译后文本与图形数据作为图像输出。另选地,采用通信接口3通过传真或电子邮件,将通过该处理获得的翻译后文本与图形数据传输到需要该翻译后数据的用户。
b. Receive an image of a document containing text and graphics written in a language via the communication interface 3 and store the image in the
在非易失性存储器7中存储有使得能够传输图像(其为有图文档翻译程序的输入信息)和翻译后文本与图形数据(其为有图文档翻译程序的输出信息)的控制程序。通过经由操作单元5或通信接口3提供的命令,提供详细指令,使控制程序进行信息传输。
In the nonvolatile memory 7 is stored a control program that enables transmission of images, which are input information of the document translation program with pictures, and translated text and graphic data, which are output information of the document translation program with pictures. Through commands provided via the
下面,将给出对本实施例操作的说明。一旦通过图像读取装置1和通信接口3输入了具有一页或多页的待处理图像,则随后将其存储在易失性存储器6中,CPU8执行有图文档翻译程序,其流程示于图2中。在布局分析处理101(其为有图文档翻译程序的第一处理)中,CPU8分析存储在易失性存储器6中的输入图像的布局。然后CPU8获得如图3中所示的文本区201、图形区202和余白区203,并将具有与所获区域相同区域的翻译后文本与图形数据存储在易失性存储器6的工作区。在这一阶段,文本区201和图形区202为空。
Next, a description will be given of the operation of this embodiment. Once the image to be processed with one or more pages is input through the image reading device 1 and the communication interface 3, it is then stored in the
随后CPU8顺序进行字符识别处理102、翻译处理103和文本量计算处理104。在字符识别处理102中,对包含在输入图像200的文本区201中的图像进行字符识别,从而生成文本区文本数据。该文本区文本数 据包含有关经识别的字符的类型的信息和诸如字符大小、行距、页边距等的设置信息。在字符识别处理102中,还对图形区202中包含的图像进行处理以生成图形区文本数据。该图形区文本数据包含有关图形区202中包含的字符的类型、位置和大小的信息。
The CPU 8 then sequentially performs
在翻译处理103中,将文本区文本数据和图形区文本数据翻译成另一种语言,以生成文本区翻译后的文本数据204和图形区翻译后的文本数据205,分别将其分配到已存储在工作区中的翻译后文本与图形数据的文本区201和图形区202。该文本区翻译后的文本数据204包含有关形成翻译后文本的字符的类型的信息以及从原始文本区文本数据继承来的设置信息;并且图形区翻译后的文本数据205包含表示形成翻译后文本的字符的类型的信息和表示从原始图形区文本数据继承来的字符的位置和大小的信息。通过操作单元5或通信接口3提供的命令分别指定包含在输入图像200中的文本的语言和翻译后文本的语言,并且根据该命令在翻译处理103中进行翻译。
In the
在图2中,由短划线进行的分组形成了布局结构处理器500,用于生成翻译后文本与图形数据,该数据在其文本区包含文本区翻译后的文本数据204并在其图形区包含图形区翻译后的文本数据205以及输入图像200的图形区202中的数据。在文本量计算处理104中,计算由文本区翻译后的文本数据204表示的翻译后文本的数据量并将其存储在非易失性存储器7中。在区域大小计算处理105中计算文本区201和图形区202的大小(即面积)并将其存储在易失性存储器6中。可以在布局分析处理101之后和文本量/区域大小比较处理106之前的任何时刻进行区域大小计算处理105。
In FIG. 2, the grouping by dashed lines forms a layout structure processor 500 for generating translated text and graphics data containing translated text data 204 in its text area and translated text data 204 in its graphics area. Contains the translated text data 205 of the graphics area and the data in the
在文本量/区域大小比较处理106中,将由文本量计算处理104获得的翻译后文本的字符量与由区域大小计算处理105获得的文本区201的大小进行比较,并且将比较结果存储在易失性存储器6中。具体而言,在文本量/区域大小比较处理106中,计算对由文本区翻译后的文本数据204表示的翻译后文本成像获得的图像所占据的面积与要容纳该图像的文本区201的大小之间的比值,并将该计算出的比值存储在易失性存储 器6中。
In the text amount/area
CPU8基于进行文本量/区域大小比较处理106所获得的结果,进行文本大小缩放处理107或图形缩放处理108。文本大小缩放处理107作为设置控制装置对在翻译后文本与图形数据中的文本区翻译后的文本数据204的设置进行控制,从而继承了输入图像200的布局作为翻译后文本与图形数据的文本区201和图形区202的布局。具体而言,在文本大小缩放处理107中,根据目前翻译后文本占据的面积和文本区的大小之间的比值(该比值已经在文本量/区域大小比较处理106中获得)计算翻译后文本的各字符的大小,使得当对文本区翻译后的文本数据204成像时,该图像可以容纳在文本区201中。
The CPU 8 performs text
当判定如果缩小(或放大)图形区202的大小则无需改变字符大小就可以将翻译后文本的图像容纳到文本区201中时,则进行图形缩放处理108,以在最大允许限度内放大(或缩小)文本区201的大小。根据在文本量/区域大小比较处理106中获得的比值作出该判定。在图形缩放处理108中,获得在不改变设置(诸如形成文本区翻译后的文本数据204的字符大小)的情况下容纳通过对数据204进行成像获得的图像所需的文本区201的大小。此外,在根据所获得的大小改变文本区201的大小的情况下,还获得图形区202的大小,使得在输入图像200的同一页中的文本区201和图形区202仍可以容纳在翻译后文本与图形数据中的同一页中。然后,获得对图形的缩放因子,该缩放因子为改变前后的大小的比值。
When it is determined that the image of the translated text can be accommodated in the
在重构处理109中,使用进行文本大小缩放处理107和图形缩放处理108获得的结果,重构翻译后文本与图形数据,将该数据存储在易失性存储器6的工作区中。
In the
在进行了文本大小缩放处理107而未进行图形缩放处理108的情况下,在重构处理109中,首先将输入图像200不包括文本图像的图形区202的图像存储在工作区中翻译后文本与图形数据的图形区202中。接着,将从输入图像200的图形区202中获得的图形区翻译后的文本数据205存储在翻译后文本与图形数据的图形区202中。当将存储在图形区202 中的数据重现为图像时,图形区202中的翻译后文本将与输入图像200的原始文本具有相同的大小并占据相同的位置。随后,将从输入图像200的文本区201中获得的文本区翻译后的文本数据204存储在翻译后文本与图形数据的文本区201中。指定包含在文本区翻译后的文本数据204中的字符大小的信息示出了应用文本大小缩放处理107之后的大小。结果,当对文本区翻译后的文本数据204进行成像时,该图像准确适合于文本区201。
In the case that the text
另一方面,在重构处理109中,在不进行文本大小缩放处理107而进行图形缩放处理108的情况下,根据执行图形缩放处理108而获得的结果来改变翻译后文本与图形数据的文本区201和图形区202。根据图形缩放处理108获得的缩放因子,对除去字符图像后的输入图像200的图形区202的图像进行放大或缩小,以存储在翻译后文本与图形数据的图形区202中。
On the other hand, in the
此外,对表示图形区翻译后的文本数据205中所包含的字符大小的信息进行修改以示出通过乘以图形缩放处理108获得的缩放因子而获得的值。再根据该缩放因子,修改显示图形区翻译后的文本数据205中的字符位置的信息。进行该修改使得当对图形区翻译后的文本数据205成像时,在图形区202中,翻译后文本的字符的图像占据与输入图像的原始字符相同的位置。接着,将文本区翻译后的文本数据204存储在翻译后文本与图形数据的文本区201中。不对指定包含在文本区翻译后的文本数据204中的字符大小的信息进行文本大小缩放处理107,而是通过图形缩放处理108改变文本区201。结果,当对文本区翻译后的文本数据204进行成像时,翻译后文本的图像准确适合于文本区201。
Also, the information indicating the character size contained in the graphics area translated text data 205 is modified to show a value obtained by multiplying the scaling factor obtained by the
因此,在以翻译后文本取代了包含在原始输入图像中的文本后,将翻译后文本与图形数据存储在易失性存储器6的工作区中。然后通过打印装置2将翻译后文本与图形数据打印到记录纸上,或通过通信接口3将其传输到需要该翻译结果的外部用户处。
Therefore, the translated text and graphics data are stored in the work area of the
图4A至4B的图与传统技术相比较地说明了本发明的效果。在该示例中,如图4A所示,将外文的有图文档的输入图像存储在易失性存储器 6中,并将输入图像的文本部分翻译成日语。
4A to 4B are diagrams illustrating the effects of the present invention in comparison with conventional techniques. In this example, as shown in FIG. 4A, an input image of a document with pictures in a foreign language is stored in the
当采用传统技术时,在输入图像的文本区中的字符串的翻译后文本不再能够容纳在原始文本区中的情况下,如图4B所示,翻译后文本被置于被放大了的文本区中,结果图形区被移到了原始页的下一页。结果,在采用传统技术输出的翻译后文本与图形中,文本区和图形区的布局看上去与输入图像的外文有图原始文档差异很大。因此导致翻译后文本很难阅读。 When using the conventional technique, in the case where the translated text of the character string in the text area of the input image can no longer be accommodated in the original text area, as shown in Figure 4B, the translated text is placed in the enlarged text area, the result graphics area is moved to the next page of the original page. As a result, in translated text and graphics output using conventional techniques, the layout of the text and graphics areas looks very different from the original foreign-language document with the input image. This makes the translated text difficult to read. the
相反,在本实施例中,获得了如图4C或图4D所示的翻译后文本与图形。图4C示出了通过进行文本大小缩放处理107而未进行图形缩放处理108获得的翻译后文本与图形;而图4D示出了通过未进行文本大小缩放处理107而进行了图形缩放处理108获得的翻译后文本与图形。如这些图中所示,根据本实施例获得的翻译后文本与图形的文本区和图形区布局与输入图像的布局相同(图4C)或与输入图像的布局差异不是很大(图4D)。因此,当与采用传统技术的翻译后文本与图形相比时,根据本实施例获得的翻译后文本与图形很容易被用户所理解。
On the contrary, in this embodiment, the translated text and graphics as shown in FIG. 4C or 4D are obtained. Figure 4C shows the translated text and graphics obtained by performing the text
第二实施例 Second embodiment
第一实施例对于要处理有多个页的输入图像的情况也有效。在要处理有多页的输入图像的情况下,根据第一实施例,在尽可能保持各页的文本区、图形区、和余白的同时,生成翻译后文本与图形数据,并且将从输入图像的各页的文本区获得的文本区翻译后的文本数据存储在翻译后文本与图形数据的与输入图像相同的页的文本区中。然而,当严格应用该规则时,可能会引起在翻译后文本和图形数据的各页的文本区中的翻译后文本密度在各页间不一致。在本实施例中,可以将文本区翻译后的文本数据在最大允许限度内在各页之间传送以降低翻译后文本密度的不一致。换句话说,例如,在翻译后文本的文本量相对于特定页的文本区的大小很大,但是在下页中翻译后文本的文本量相对于其文本区大小很小的情况下,可以将在最后部分中的字符串(其不太可能适合于前一页)传送到后一页。相反,在翻译后文本的文本量相对于特定页的文本区的大小很小,但在下页中翻译后文本的文本量相对于其文本区的大小 很大的情况下,可以将后一页中的前端部分中的字符串发送到前一页。 The first embodiment is also effective for a case where an input image having a plurality of pages is to be processed. In the case where an input image having multiple pages is to be processed, according to the first embodiment, while maintaining the text area, graphic area, and margin of each page as much as possible, translated text and graphic data are generated and translated from the input image The translated text data is stored in the text area of each page of the translated text and graphics data in the text area of the same page as the input image. However, when this rule is strictly applied, it may cause the translated text density in the text area of each page of translated text and graphic data to be inconsistent from page to page. In this embodiment, the translated text data of the text area can be transferred between pages within the maximum allowable limit to reduce the inconsistency of the translated text density. In other words, for example, in the case where the text volume of the translated text is large relative to the size of the text area on a particular page, but the text volume of the translated text on the next page is small relative to the size of its text area, the Strings in the last part (which are unlikely to fit in the previous page) are passed to the next page. Conversely, in the case where the text volume of the translated text is small relative to the size of the text area of a particular page, but the text volume of the translated text on the next page is large relative to the size of its text area, the The string in the front section of the is sent to the previous page. the
因此降低了翻译后文本在不同页之间密度的不一致;然而,这会带来另一个问题。即,传送到另一页中的字符串包含该经传送的字符串以前位于的页中包含的图形的标号的情况下,用户必须很麻烦地翻到前一页以确认包含该标号的文本中引用的图形。本实施例还防止了给用户带来的这种不便。 Inconsistencies in the density of the translated text between different pages are thus reduced; however, this creates another problem. That is, when a character string transmitted to another page contains a label of a graphic contained in the page where the transmitted character string was previously located, the user must troublesomely turn to the previous page to confirm that the text containing the label Referenced graphics. This embodiment also prevents such inconvenience to the user. the
图5是示出了根据本实施例进行的有图文档翻译程序的过程的流程图。在该图中,对应于图2中所示处理的那些处理采用相同的标号。在根据本实施例的有图文档翻译程序中,还提供了同一图形标号搜索处理110,并且,根据执行同一图形标号搜索处理110获得的结果,判定是否进行由短划线归在一起的文本量计算处理104、文本量/区域大小比较处理106、文本大小缩放处理107以及图形缩放处理108(下文中为了方便称为组处理111)。换句话说,同一图形标号搜索处理110包括:翻译后文本数据传送控制手段,用于将文本区翻译后的文本数据在最大允许限度内在翻译后文本与图形数据的不同页的文本区之间传送;和传送控制/设置控制切换手段,用于在以下情况下,使设置控制手段控制所讨论页的设置,来取代由翻译后文本数据传送控制手段传送翻译后文本数据:a)标识图形的图形标识信息包含在图形区翻译后的文本数据中;b)在翻译后文本数据传送控制手段对翻译后文本数据进行传送之前,与该图形区翻译后的文本数据中包含的图形标识信息同一页中的文本区翻译后的文本数据中包含与该图形区翻译后的文本数据中包含的图形标识信息相同的图形标识信息;并且c)如果翻译后文本数据被翻译后文本数据传送控制手段传送,则图形标识信息将传送到不同页中。下面,将对此进行详细说明。
FIG. 5 is a flowchart showing the procedure of the document translation program with pictures according to the present embodiment. In this figure, processes corresponding to those shown in FIG. 2 are given the same reference numerals. In the document translation program according to the present embodiment, the same figure
当可能进行将翻译后文本的字符串在不同页之间传送时,执行同一图形标号搜索处理110。具体而言,在特定字符串可能从其溢出到另一页的各页中进行同一图形标号搜索处理110,并判定是否允许这种流出。
The same figure
在从特定页的文本区流出的字符串不包含诸如标号和标题的任何图形标识信息、或即使包含这种图形标识信息,但其并不对应于相同页的 图形区中的图形标识信息的情况下,允许字符串这样流到另一页中。相反,在从特定页的文本区中流出的字符串包含图形标识信息并且其对应于相同页的图形区中的图形标识信息的情况下,不允许字符串流到另一页。 In the case where a character string flowing from a text area of a particular page does not contain any graphic identification information such as a label and a title, or even if it contains such graphic identification information, it does not correspond to graphic identification information in a graphic area of the same page Next, allow strings to flow like this into another page. Conversely, in the case where a character string flowing out from a text area of a certain page contains graphic identification information and it corresponds to graphic identification information in a graphic area of the same page, the character string is not allowed to flow to another page. the
在同一图形标号搜索处理110中,对在不同页之间传送时,翻译后文本的字符串从其传送或传送到此的各页进行上述判定,以判定是否允许对这些字符的传送。仅对被允许的页进行这些字符的传送,并且确定待分配到各页的文本区的翻译后文本。结果,一些页的文本区可能由于不允许字符串流出到另一页,而被填满溢出,或由于不允许字符串从另一页中流入,一些页的文本区中可能包含空白。在同一图形标号搜索处理110中,判定应该对一些这种页进行组处理111并且不需要对其他页进行组处理111。
In the same figure
仅当上述同一图形标号搜索处理110判定需要执行组处理111时才执行组处理111。在第一实施例中描述了组处理111和重构处理的细节,因此将略去重复的说明。
The
根据本实施例,仅当在使标号与该标号所引用的图形位于相同页中为优选的情况下,由于翻译,文本区中的该图形标号可能被传送到另一页时,才进行文本大小缩放处理107和图形缩放处理108,以防止将标号传送到标号引用的图形并不位于其中的另一页中。因此,由于当参照图形阅读正文时,用户不必参照不同页,并且由于除必要时不改变字符和/或图形的大小,所以,用户将很容易阅读根据本发明获得的翻译后文本与图形。
According to this embodiment, the text size is only done if it is preferable to have the label in the same page as the figure it refers to, and the figure label in the text area may be transferred to another page due to
其他实施例 other embodiments
前面已经说明了第一实施例和第二实施例,但是本发明还可以有下面的实施例。 The first embodiment and the second embodiment have been described above, but the present invention can also have the following embodiments. the
1)在第一和第二实施例中,可以改变行距使得不用进行文本大小缩放处理107而使翻译后文本适合于文本区。
1) In the first and second embodiments, the line spacing can be changed so that the translated text fits in the text area without performing the text
2)在第一和第二实施例中,可以通过放大或缩小余白的大小来调整文本区的大小使得可以将翻译后文本容纳于其中。 2) In the first and second embodiments, the size of the text area can be adjusted by enlarging or reducing the size of the margin so that the translated text can be accommodated therein.
3)可以进行图形缩放处理108和文本大小缩放处理107两者来重构布局。在这种情况下可以设想一些模式。在第一种模式中,可以提供字符大小缩放因子的上限和下限,并且在所提供的限度内进行文本大小缩放处理107。在即使根据限度内的缩放因子缩小或放大字符大小,仍不能将文本区中的翻译后文本的图像容纳入与原始页相同的页中的情况下,进而进行图形缩放处理108以通过缩小或放大图形区来保留容纳文本区翻译后的文本数据的图像所需的面积。在第二种模式中,提供图形区缩放因子的上限和下限,并且在所提供的限度内进行图形缩放处理108。在即使根据限度内的缩放因子缩小或放大图形区后,仍不能将文本区中翻译后的文本数据的图像容纳在与原始页相同的页中的情况下,进行文本大小缩放处理107使得文本区翻译后的文本数据的图像能够容纳入文本区中。
3) Both
4)在一些情况下,用户可以指定所谓的“N-Up”打印,其中“N-Up”将N个页的图像作为一组合图像输出在一页上。在这种情况下,同一图形标号搜索处理110可以进行以下判定。下文将给出对”2-Up”打印的示例的描述。
4) In some cases, the user can designate so-called "N-Up" printing in which images of N pages are output on one page as a combined image. In this case, the same figure
如图6中所示,假设这样一种情况:将生成翻译后文本与图形数据400-1,其为将输出在同一页面上的输入图像200-1和200-2的2-Up模式数据。在图6所示的示例中,输入图像200-1在其文本区201包含对应于包含在其图形区202中的图形标识信息的图形标识信息。在翻译后文本与图形数据400-1中,在一些情况下,由于翻译改变了字符数,对对应于输入图像200-1的页的文本区201中包含的图形标识信息“图1”的翻译可能不能容纳在相同文本区201中。在这种情况下,由于由图形标识信息“图1”标识的图形存在于相同页面中,在同一图形标号搜索处理110中允许将对“图1”的翻译从对应于输入图像200-1的页的文本区传送到对应于输入图像200-2的页的文本区201。进而在同一图形标号搜索处理110中判定是否需要对最初容纳翻译“图1”的页进行组处理111。
As shown in FIG. 6, assume a case where translated text and graphics data 400-1, which is 2-Up mode data of input images 200-1 and 200-2 to be output on the same page, will be generated. In the example shown in FIG. 6 , the input image 200 - 1 contains in its
如图7所示,假设这样一种情况:将生成翻译后文本数据400-1(其使对应于输入图像200-1和200-2的数据以2-Up模式输出在同一页面中) 和翻译后文本数据400-2(其使得对应于输入图像200-3和200-4的数据以2-Up模式输出在同一页面中)。在图7所示的示例中,对于输入图像200-1和200-2,对应于输入图像200-1的图形区202中存在的图形标识信息的图形标识信息存在于输入图像200-2的文本区201中。在翻译后文本与图形数据400-1中,在某些情况下,由于通过翻译处理改变了字符数,因而如果字符大小相同,则对对应于输入图像200-2的页的文本区201中容纳的图形标识信息“图1”的翻译不能容纳在文本区201中。如果允许将对图形标识信息“图1”的翻译传送到下一页(即对应于输入图像200-3的页),则由图形标识信息“图1”标识的图形将不能与该图形标识信息存在于相同的页中。为了避免这种结果,在同一图形标号搜索处理110中,不允许传送字符串,并且判定应该对对应于输入图像200-2的数据进行组处理111。根据本实施例可以获得与上述第二实施例相同的效果。
As shown in FIG. 7 , assume a case where translated text data 400-1 (which causes data corresponding to input images 200-1 and 200-2 to be output in the same page in 2-Up mode) and translated text data 400-1 will be generated. Post text data 400-2 (which causes data corresponding to the input images 200-3 and 200-4 to be output in the same page in 2-Up mode). In the example shown in FIG. 7, for the input images 200-1 and 200-2, the graphic identification information corresponding to the graphic identification information existing in the
5)在上述各实施例中,将翻译后文本数据作为文本数据存储在翻译后文本与图形数据的文本区和图形区中,使得生成翻译后文本与图形数据,该数据同时具有文本数据和图像数据。另选地,可以对翻译后文本数据进行成像以将经成像的翻译后文本映射到文本区和图形区,从而生成翻译后文本与图形数据,该数据包含全部图像数据。 5) In each of the above-mentioned embodiments, the translated text data is stored as text data in the text area and the graphic area of the translated text and graphic data, so that the translated text and graphic data are generated, and the data has both text data and images data. Alternatively, the translated text data may be imaged to map the imaged translated text to a text area and a graphics area, thereby generating translated text and graphics data, the data comprising the entire image data. the
6)在上述各实施例中,在多功能机中安装有图文档翻译程序。然而,本发明的实施例并不限于这种模式,而可以将上述实施例的有图文档翻译程序存储在诸如CD-ROM的计算机可读取记录介质中以发给普通用户。另选地,可以通过网络将有图文档翻译程序发给普通用户。 6) In each of the above-mentioned embodiments, a picture-to-file translation program is installed in the multi-function machine. However, the embodiments of the present invention are not limited to this mode, but the document translation program of the above-described embodiments may be stored in a computer-readable recording medium such as a CD-ROM for distribution to general users. Alternatively, the document translation program with pictures can be sent to ordinary users through the network. the
第三实施例 third embodiment
下面将参照附图给出对本发明第三实施例的描述。 A description will be given below of a third embodiment of the present invention with reference to the drawings. the
图8为示出了根据本发明实施例的图像形成装置100的框图。如图所示,图像形成装置100具有翻译处理单元1、操作单元2、网络I/F单元3、存储器存储单元4、打印单元5和图像读取单元6。
FIG. 8 is a block diagram showing an
打印单元5具有感光器、曝光单元、传送单元和定影单元(fixingunit)。打印单元5根据翻译处理单元1提供的图像数据生成色粉图像以 并将该色粉图像定影在一页纸(其为记录介质)上。操作单元2具有由液晶显示器(未示出)形成的显示单元和各种按钮,从而可以接收来自用户的指令。用户使用操作单元2选择要使用的纸、输入各种打印设置等。
The
图像读取单元6扫描源文档的数据,并且作为图像数据输出。存储器存储单元4存储由图像读取单元6扫描的图像数据以及其他数据。通过网络I/F单元3使得能够进行翻译处理单元1、操作单元2、存储器存储单元4、打印单元5以及图像读取单元6之间的数据通信。
The
如图9所示,翻译处理单元1包括CPU(中央处理单元)11、RAM(随机访问存储器)12以及ROM(只读存储器)13。翻译处理单元1控制图像形成装置100的各单元,并执行各图像处理以及输入图像数据的翻译处理所需的计算。图像数据暂时存储在RAM12中以待处理。在ROM13中存储有图像数据处理和翻译处理所需的各种图像处理程序PRG和翻译程序PRG。
As shown in FIG. 9 , the translation processing unit 1 includes a CPU (Central Processing Unit) 11 , a RAM (Random Access Memory) 12 , and a ROM (Read Only Memory) 13 . The translation processing unit 1 controls each unit of the
接下来将根据图10所示的流程图,给出由具有上述结构的图像形成装置100进行的操作的示例。在该示例中,由图像形成装置100读取源文档的图像,并且将包含在该图像中的日语翻译成英语。
Next, an example of operations performed by the
将源文档置于图像读取单元6的扫描板(未示出)上,然后使图像读取单元6开始读取该文档。结果,图像读取单元6读取存在于扫描区的图像(步骤S01)。
A source document is placed on a scanning plate (not shown) of the
图11A示出了由图像读取单元6读取的图像的一个示例。如图所示,扫描图像G为包括包含日语的原始文本J(包括由图中圆圈所示的部分)和图形Z的图像。
FIG. 11A shows an example of an image read by the
用户参照示出了诸如图11A中所示的图像的操作单元2的显示屏,操作诸如鼠标或键盘的输入仪器来移动指针或光标,以指定扫描图像G用于翻译的区域(步骤S02)。在该示例中,图像G的全部区域(图11A所示的图像G的全部)被指定为翻译区域。
Referring to the display screen of the
当用户指定全部区域为翻译区域时,翻译处理单元1识别在图像G中指定的全部区域中的图像数据的文本。将识别出的文本从全部图像数 据中提取出(步骤S03)。 When the user specifies the entire area as the translation area, the translation processing unit 1 recognizes the text of the image data in the entire area specified in the image G. The recognized text is extracted from all image data (step S03). the
然后翻译处理单元1对包含在所提取出的文本的区域中的图像数据进行OCR(光学字符阅读器)处理,将该区域中的文本的图像数据转变为文本数据,并且读取经转变的文本数据作为原始文本J的文本数据(步骤S04)。 The translation processing unit 1 then performs OCR (Optical Character Reader) processing on the image data contained in the extracted text area, converts the image data of the text in the area into text data, and reads the converted text The data is used as text data of the original text J (step S04). the
随后翻译处理单元1的CPU11获得存储在ROM13中的语言信息并将语言信息与原始文本J的文本数据进行比较以识别出原始文本J的语言(步骤S05)。在该示例中,在将该文本与ROM13中的语言信息比较后,CPU11识别原始文本J的语言为日语。
The CPU 11 of the translation processing unit 1 then obtains the language information stored in the
此外,翻译处理单元1的CPU11从存储在ROM13中的翻译处理程序PGM中运行并执行日译英的翻译程序,从而生成英语的翻译后文本E的文本数据(步骤S06)。
Further, the CPU 11 of the translation processing unit 1 runs and executes the translation program from Japanese to English from the translation processing program PGM stored in the
在根据原始文本J生成翻译后文本E后,翻译处理单元1比较翻译后数据E和原始文本J的字符串的长度(步骤S07)。 After generating the translated text E from the original text J, the translation processing unit 1 compares the lengths of the character strings of the translated data E and the original text J (step S07). the
在本实施例中,确定原始文本J的语义部分(例如分句或整句)的图像的横向长度为字符串的长度;并且通过将翻译后文本E的各字符在横向上的点数和横向上各字符间的间距的点数求和获得的长度确定为字符串的长度。将所确定的翻译后文本E的的字符串的长度与原始文本J的相比较。另选地,可以将原始文本J的字符串长度确定为语义部分横向上字符的点数之和。 In this embodiment, determine the horizontal length of the image of the semantic part (such as a clause or a whole sentence) of the original text J as the length of the character string; The length obtained by summing the points of the spaces between the characters is determined as the length of the character string. The determined length of the character string of the translated text E is compared with that of the original text J. Alternatively, the character string length of the original text J may be determined as the sum of the points of the characters in the horizontal direction of the semantic part. the
翻译处理单元1根据原始文本J和翻译后文本E的字符串长度之间的比较结果,从诸如改变字体和/或点、均等间距等的校正处理中确定适当的处理,作为要应用到翻译后文本E的校正处理。翻译处理单元1然后进行经确定的校正处理以将翻译后文本E的字符串长度变为与原始文本J的长度相等(步骤S08)。 The translation processing unit 1 determines appropriate processing from correction processing such as changing the font and/or point, equal spacing, etc., as a result of the comparison between the character string lengths of the original text J and the translated text E, as to be applied to the translated text E. Correction processing of text E. The translation processing unit 1 then performs the determined correction processing to make the character string length of the translated text E equal to the length of the original text J (step S08). the
应当注意的是,当翻译后文本E的字符串长度落在预定点数范围(在视觉上,用户可认为是相同)内时,翻译处理单元1将翻译后文本E和原始文本J的字符串长度确定为彼此相等。 It should be noted that when the character string length of the translated text E falls within a predetermined range of points (visually, the user can regard it as the same), the translation processing unit 1 compares the character string lengths of the translated text E and the original text J determined to be equal to each other. the
图12A示出了日语原始文本J的字符串和翻译后文本E的字符串。 如图所示,翻译后文本E的字符串长于原始文本J的字符串。在这种情况下,翻译处理单元1根据原始文本J和翻译后文本E之间的字符串长度的比较结果,改变字体以及减小字符大小的点数作为校正处理。翻译处理单元1改变字体和字符大小点数,以减小字符串长度,从而如图12B所示,将翻译后文本E的字符串长度改变为与原始文本J的字符串长度相同的预定点数范围内。 FIG. 12A shows a character string of the Japanese original text J and a character string of the translated text E. FIG. As shown, the string of the translated text E is longer than the string of the original text J. In this case, the translation processing unit 1 changes the font and reduces the number of points of the character size according to the comparison result of the character string length between the original text J and the translated text E as correction processing. The translation processing unit 1 changes the font and character size points to reduce the character string length, thereby changing the character string length of the translated text E to within the same predetermined point range as the character string length of the original text J as shown in FIG. 12B . the
接着,翻译处理单元1将进行了校正处理的翻译后文本E的文本数据加入到整个图像G的图像数据,从而字符串长度经过校正的翻译后文本E平行于并位于原始文本J的下面(步骤S09)。 Next, the translation processing unit 1 adds the text data of the translated text E that has been corrected to the image data of the entire image G, so that the translated text E whose character string length is corrected is parallel to and located below the original text J (step S09). the
当用户通过操作单元2输入了打印指令时,将在翻译处理单元1中处理过的图像数据输出到打印单元5,并且在打印单元5中,将表示图像数据的图像打印到纸上(步骤S10)。
When the user has input a printing instruction through the
结果,如图11B所示,在限定范围的整个部分中,字符串长度合适的英语的翻译后文本E(在图中以×标记示出)位于现存的日语原始文本J的下面。 As a result, as shown in FIG. 11B , the translated text E in English (shown with an X mark in the figure) of an appropriate character string length is located below the existing Japanese original text J in the entire portion of the limited range. the
因此,根据本实施例,通过改变翻译后文本E的字符的点数和字体,可以将翻译后文本E平行于原始文本J放置,且其字符串的长度与原始文本J的相同。结果,可以根据以这种方式放置的表示翻译后文本E和原始文本J的打印件,很容易地识别原始文本J和翻译后文本E之间的对应;因此,大大提高了打印件的可阅览性。 Therefore, according to this embodiment, by changing the points and fonts of the characters of the translated text E, the translated text E can be placed parallel to the original text J with the same character string length as that of the original text J. As a result, the correspondence between the original text J and the translated text E can be easily recognized based on the printouts representing the translated text E and the original text J placed in this manner; thus, the viewability of the printouts is greatly improved sex. the
此外,由于在翻译前提取出图像数据中的文本数据,因而可以确保仅对包含图形的图像的字符或文本进行翻译。 Furthermore, since the text data in the image data is taken out before translation, only the characters or text of the image including graphics can be surely translated. the
在上述实施例中,给出了对翻译后文本的字符串长于原始文本的情况的说明。在翻译后文本的字符串短于原始文本的情况下,翻译处理单元1进行与上述相同的处理,使翻译后文本与原始文本相关地等间距排列,使得翻译后文本的长度与原始文本的相等,并且翻译后文本和原始文本彼此平行放置。 In the above-mentioned embodiments, an explanation has been given of the case where the character string of the translated text is longer than that of the original text. In the case where the character string of the translated text is shorter than the original text, the translation processing unit 1 performs the same processing as above to arrange the translated text at equal intervals relative to the original text so that the length of the translated text is equal to that of the original text , and the translated and original texts are placed parallel to each other. the
此外,在上述实施例中,由于通过改变翻译后文本中字符的点数或字体来校正字符串的长度,因而另选地可以将翻译后文本的字符改为注 音型字符(ruby-type character)以校正翻译后文本的字符串长度。 Furthermore, in the above-described embodiment, since the length of the character string is corrected by changing the points or font of the characters in the translated text, it is alternatively possible to change the characters of the translated text into ruby-type characters to correct the string length of the translated text. the
在上述实施例中,给出了使字符串长度适合于原始文本的过程的示例的说明。显然,可以改变原始文本的字符串长度或原始文本和翻译后文本的字符串长度使其具有彼此相等的长度。在改变原始文本的字符串长度的情况下,可以在将原始文本转变为字符后,改变字符的字体或点数,或者可以将转变后的字符与翻译后文本相关地等间距排列。另选地,可以放大或缩小原始文本的图像数据的大小来进行调整。 In the above-described embodiments, a description has been given of an example of the process of adapting the length of a character string to an original text. Obviously, the character string length of the original text or the character string lengths of the original text and the translated text may be changed to have lengths equal to each other. In the case of changing the character string length of the original text, after converting the original text into characters, the font or point of the characters may be changed, or the converted characters may be arranged at equal intervals relative to the translated text. Alternatively, the size of the image data of the original text may be enlarged or reduced for adjustment. the
只要字符横向上和字符之间间距中的点数是预定的,则不仅可以通过对点数求和,而且可以通过使用字符数来确定字符串的长度。另选地,可以根据打印输出的长度(毫米)来确定字符串的长度。 As long as the number of points in the lateral direction of characters and in the space between characters is predetermined, the length of the character string can be determined not only by summing the number of points but also by using the number of characters. Alternatively, the length of the character string may be determined according to the length (mm) of the printout. the
显然,原始文本和翻译后文本的语言并不限于上述实施例中所描述的那些,可以是德语、法语、俄语、西班牙语、汉语、韩国语和其他语种。 Obviously, the languages of the original text and the translated text are not limited to those described in the above embodiments, and may be German, French, Russian, Spanish, Chinese, Korean and other languages. the
在上述实施例中,给出了假设本发明应用于图像形成装置100中的说明,但本发明不限于此。例如还可以提供一种仅具有上述图像形成装置100中翻译处理单元1的功能的翻译装置或图像处理装置。在这种情况下,翻译装置或图像处理装置可以是具有上述翻译处理单元1的功能的ASIC(特定用途集成电路)。此外,还可以通过将用于进行上述翻译处理的翻译处理程序PRG记录在诸如磁盘、软盘、CD(光盘)、DVD(多功能数码光盘)和RAM等的记录介质上来提供本发明。
In the above-described embodiments, descriptions have been given assuming that the present invention is applied to the
如上所述,本发明提供了一种翻译装置,包括:字符识别单元,用于识别输入图像文本区的文本数据;翻译器,用于翻译所述文本区的文本数据;以及布局结构处理器,用于生成包含文本区翻译后文本数据和输入图像中图形的数据,其中在由所述布局结构处理器生成的数据的图像布局中保持输入图像的布局。 As described above, the present invention provides a translation device including: a character recognition unit for recognizing text data in a text area of an input image; a translator for translating the text data in the text area; and a layout structure processor, for generating data including translated text data of a text area and graphics in an input image, wherein the layout of the input image is maintained in the image layout of the data generated by said layout structure processor. the
根据本发明实施例,所述布局结构处理器可以根据翻译后文本数据量和文本区大小,改变翻译后文本数据的设置。 According to an embodiment of the present invention, the layout structure processor can change the setting of the translated text data according to the amount of the translated text data and the size of the text area. the
根据另一实施例,所述布局结构处理器可以通过改变翻译后文本数据的字符的大小来生成翻译后文本数据。 According to another embodiment, the layout structure processor may generate the translated text data by changing the size of characters of the translated text data.
根据另一实施例,所述布局结构处理器可以通过改变翻译后文本数据的行距来生成翻译后文本数据。 According to another embodiment, the layout structure processor may generate the translated text data by changing the line spacing of the translated text data. the
根据另一实施例,所述布局结构处理器可以通过改变翻译后文本数据周围的余白来生成翻译后文本数据。 According to another embodiment, the layout structure processor may generate translated text data by changing margins around the translated text data. the
根据另一实施例,所述布局结构处理器可以通过改变图形区的大小来生成数据。 According to another embodiment, the layout structure processor may generate data by changing the size of the graphics area. the
根据本发明的又一实施例,所述翻译装置还可以包括图形再形成单元,用于在以下情况下再形成图形区:通过在预定区域内再形成文本区,可以将翻译后文本数据以与该文本数据相同的字符大小容纳在再形成的文本区中。 According to still another embodiment of the present invention, the translation device may further include a graphic reformation unit for reforming the graphic area in the following cases: by reforming the text area in a predetermined area, the translated text data can be combined with The same character size of the text data is accommodated in the reformulated text area. the
根据另一实施例,所述翻译装置还包括图形再形成单元,用于在以预定方式改变翻译后文本数据的设置,还不能将翻译后文本数据容纳在文本区中的情况下再形成文本区和图形区。 According to another embodiment, the translating apparatus further includes a graphic reforming unit for reforming the text area when the setting of the translated text data is changed in a predetermined manner and the translated text data cannot be accommodated in the text area. and graphics area. the
根据另一实施例,所述字符识别单元还可以对输入图像图形区的图像执行字符识别处理,以输出表示图形区字符的类型、位置和大小的图形区文本数据;所述翻译处理器还可以对图形区文本数据执行翻译处理,以输出要容纳到图形区的表示字符的类型、位置和大小的图形区翻译后文本数据;并且所述布局结构处理器还包括:翻译后数据传送控制器,用于将文本区翻译后文本数据在预定方式内在翻译后文本与图形数据的不同页的文本区之间进行传送;以及切换单元,用于在以下情况下对页的设置进行控制,代替对标识图形的图形标识信息进行传送:所述图形标识信息包含在该页的图形区翻译后文本数据中;在翻译后数据传送控制器对翻译后文本数据进行传送之前,包含该被包含图形标识信息的同一页的文本区翻译后文本数据中包含与该被包含图形标识信息相同的图形标识信息;并且如果所述翻译后数据传送控制器对翻译后文本数据进行传送,则该图形标识信息将被传送到另一页的文本区中。 According to another embodiment, the character recognition unit may also perform character recognition processing on the image of the graphic area of the input image to output graphic area text data representing the type, position and size of the character in the graphic area; the translation processor may also performing a translation process on the graphic area text data to output graphic area translated text data representing the type, position and size of characters to be accommodated in the graphic area; and the layout structure processor further includes: a translated data transfer controller, for transferring the translated text data of the text area between text areas of different pages of the translated text and graphic data in a predetermined manner; and a switching unit for controlling the setting of the page in the case of The graphic identification information of the graphics is transmitted: the graphic identification information is included in the translated text data of the graphic area of the page; before the translated data transmission controller transmits the translated text data, the included graphic identification information The translated text data in the text area of the same page contains the same graphic identification information as the included graphic identification information; and if the translated data transmission controller transmits the translated text data, the graphic identification information will be transmitted into a text area on another page. the
根据另一实施例,字符识别单元还可以对输入图像的图形区图像执行字符识别处理,以输出表示图形区中字符的类型、位置和大小的图形区文本数据;所述翻译处理器还可以对图形区文本数据执行翻译处理, 以输出表示要容纳到所述图形区的字符的类型、位置和大小的图形区翻译后文本数据;并且其中所述布局结构处理器可以包括:翻译后数据传送控制器,用于以预定方式将文本区翻译后文本数据在翻译后文本与图形数据的不同页的文本区之间进行传送;以及切换单元,用于在生成具有将多个页布置在同一页面的N-up结构的翻译后文本与图形数据时,在以下情况下对页的设置进行控制,代替对标识图形的图形标识信息进行传送:所述图形标识信息包含在布置在页面中的图形区翻译后文本数据中;在翻译后数据传送控制器对翻译后文本数据进行传送之前,包含该被包含图形标识信息的同一页面的文本区翻译后文本数据中包含与该被包含图形标识信息相同的图形标识信息;并且如果所述翻译后数据传送控制器对翻译后文本数据进行传送,则该图形标识信息将被传送到另一页面的文本区中。 According to another embodiment, the character recognition unit can also perform character recognition processing on the graphic area image of the input image to output graphic area text data representing the type, position and size of the character in the graphic area; The graphic area text data performs translation processing to output the graphic area translated text data representing the type, position and size of characters to be accommodated in the graphic area; and wherein the layout structure processor may include: translated data transfer control a device for transferring the translated text data of the text area between text areas of different pages of the translated text and graphic data in a predetermined manner; In the case of translated text and graphic data of N-up structure, the setting of the page is controlled in the following cases, instead of transmitting the graphic identification information of the identification graphic: said graphic identification information is included in the translation of the graphic area arranged in the page In the post-text data; before the translated data transmission controller transmits the translated text data, the text area of the same page containing the contained graphic identification information contains the same graphics as the contained graphic identification information in the translated text data identification information; and if the translated data transmission controller transmits the translated text data, the graphic identification information will be transmitted to the text area of another page. the
如上所述,本发明提供了一种翻译装置,该翻译装置包括翻译处理器,用于翻译原始文本并进行校正使得原始文本字符串的长度和翻译后文本字符串的长度彼此基本相同,其中原始文本和翻译后文本平行放置。 As described above, the present invention provides a translation device including a translation processor for translating an original text and correcting so that the lengths of the original text character string and the length of the translated text character string are substantially the same as each other, wherein the original Text and translated text are placed in parallel. the
根据本发明的实施例,所述翻译处理器可以改变原始文本和翻译后文本中至少之一的字符长度。 According to an embodiment of the present invention, the translation processor may change a character length of at least one of the original text and the translated text. the
根据另一实施例,所述翻译处理器可以通过对字符间的空间和字符的点数求和来确定字符串的长度。 According to another embodiment, the translation processor may determine the length of the character string by summing the spaces between characters and the points of the characters. the
根据另一实施例,所述翻译处理器可以改变字符的点大小。 According to another embodiment, the translation processor may change the point size of the characters. the
根据另一实施例,所述翻译处理器可以改变字符字体。 According to another embodiment, the translation processor may change character fonts. the
根据另一实施例,所述翻译处理器可以将字符改为注音型字符。 According to another embodiment, the translation processor may change the characters into phonetic characters. the
一方面,本发明提供了一种翻译方法,包括以下步骤:识别输入图像的文本区的文本数据;翻译文本区的文本数据;并且生成包含文本区翻译后文本数据和输入图像中的图形的数据,其中在生成步骤中生成的数据的图像布局中保持输入图像的布局。 In one aspect, the present invention provides a translation method comprising the steps of: identifying text data in a text area of an input image; translating the text data in the text area; and generating data comprising the translated text data in the text area and graphics in the input image , where the layout of the input image is maintained in the image layout of the data generated in the generation step. the
另一方面,本发明还提供了翻译方法,包括以下步骤:翻译原始文本,并进行校正以使得原始文本的字符串的长度和翻译后文本的字符串长度彼此基本相同,其中原始文本和翻译后文本平行放置。 On the other hand, the present invention also provides a translation method, comprising the steps of: translating the original text, and correcting so that the length of the character string of the original text and the length of the character string of the translated text are substantially the same as each other, wherein the original text and the translated text The text is placed in parallel.
又一方面,本发明还提供了一种计算机可读取存储介质,所述存储介质存储这样的程序,其指令可由计算机执行,以实现翻译功能,所述功能包括:字符识别功能,用于识别输入图像文本区中的文本数据;翻译处理功能,用于翻译文本区的文本数据;以及布局结构处理功能,用于生成包含文本区翻译后文本数据和输入图像中的图形的数据,其中在由所述布局结构处理功能生成的数据的图像布局中保持输入图像的布局。 In yet another aspect, the present invention also provides a computer-readable storage medium, which stores such a program whose instructions can be executed by a computer to realize the translation function, and the function includes: a character recognition function for recognizing Text data in the input image text area; a translation processing function for translating the text data in the text area; and a layout structure processing function for generating data including the translated text data in the text area and graphics in the input image, wherein The layout of the input image is maintained in the image layout of the data generated by the layout structure processing function. the
此外,本发明还提供了一种计算机可读取存储介质,所述存储介质存储这样的程序,其指令可由计算机执行,以实现翻译功能,所述功能包括:翻译处理功能,用于翻译原始文本,并进行校正以使得原始文本的字符串的长度和翻译后文本的字符串长度彼此基本相同,其中原始文本和翻译后文本平行放置。 In addition, the present invention also provides a computer-readable storage medium, which stores a program whose instructions can be executed by a computer to realize a translation function, and the function includes: a translation processing function for translating an original text , and correct so that the lengths of the character strings of the original text and the character strings of the translated text are substantially the same as each other, wherein the original text and the translated text are placed in parallel. the
为了解释和说明的目的,给出了对于本发明实施例的上述说明。这不是详尽的,也并不希望将本发明限于所公开的具体形式。显然,对于本领域技术人员来说,进行多种修改和变型是显而易见的。选择并说明了上述实施例是为了最好地说明本发明的原理和其实际应用,从而使得本领域其他技术人员能够理解本发明的各种实施例,并且在用于所设想的特定用途时进行各种适当的修改。本发明的范围由后面的权利要求书及其等同物来限定。 The foregoing description of the embodiments of the invention has been presented for purposes of illustration and description. It is not exhaustive, nor is it intended to limit the invention to the precise forms disclosed. Obviously, various modifications and variations will be apparent to those skilled in the art. The above-described embodiments were chosen and described in order to best explain the principles of the invention and its practical application, thereby enabling others skilled in the art to understand various embodiments of the invention and to understand the various embodiments of the invention for the particular use contemplated. Various appropriate modifications. The scope of the invention is defined by the following claims and their equivalents. the
在此通过引用并入2005年3月22日提交的日本专利申请No.2005-82047和2005年3月25日提交的日本专利申请No.2005-90179的全部内容,包括说明书、权利要求书、附图和摘要。 The entire contents of Japanese Patent Application No. 2005-82047 filed on March 22, 2005 and Japanese Patent Application No. 2005-90179 filed on March 25, 2005, including specification, claims, Figures and abstract.
Claims (10)
Applications Claiming Priority (6)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP2005082047 | 2005-03-22 | ||
| JP2005082047A JP2006268150A (en) | 2005-03-22 | 2005-03-22 | Translation device, method and program, and storage medium stored with the program |
| JP2005-082047 | 2005-03-22 | ||
| JP2005090179 | 2005-03-25 | ||
| JP2005090179A JP2006276905A (en) | 2005-03-25 | 2005-03-25 | Translation device, image processing device, image forming device, and translation method and program |
| JP2005-090179 | 2005-03-25 |
Related Child Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| CN2010102937538A Division CN101923541B (en) | 2005-03-22 | 2005-08-18 | Translation device and translation method |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| CN1838112A CN1838112A (en) | 2006-09-27 |
| CN1838112B true CN1838112B (en) | 2012-05-30 |
Family
ID=37015510
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| CN2005100928181A Expired - Fee Related CN1838112B (en) | 2005-03-22 | 2005-08-18 | translation device, translation method |
Country Status (2)
| Country | Link |
|---|---|
| JP (1) | JP2006268150A (en) |
| CN (1) | CN1838112B (en) |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN105159868A (en) * | 2015-09-01 | 2015-12-16 | 广东欧珀移动通信有限公司 | Text display method and system |
Families Citing this family (16)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2008299780A (en) | 2007-06-04 | 2008-12-11 | Fuji Xerox Co Ltd | Image processing device and program |
| JP4483909B2 (en) * | 2007-08-24 | 2010-06-16 | 富士ゼロックス株式会社 | Translation apparatus and program |
| JP2009053932A (en) | 2007-08-27 | 2009-03-12 | Fuji Xerox Co Ltd | Document image processor and document image processing program |
| JP4998176B2 (en) * | 2007-09-27 | 2012-08-15 | 富士ゼロックス株式会社 | Translation apparatus and program |
| JP4569622B2 (en) | 2007-12-18 | 2010-10-27 | 富士ゼロックス株式会社 | Image processing apparatus and image processing program |
| JP2009193283A (en) * | 2008-02-14 | 2009-08-27 | Fuji Xerox Co Ltd | Document image processing apparatus and document image processing program |
| JP4544324B2 (en) | 2008-03-25 | 2010-09-15 | 富士ゼロックス株式会社 | Document processing apparatus and program |
| JP5126018B2 (en) * | 2008-11-25 | 2013-01-23 | 富士ゼロックス株式会社 | Document image processing apparatus and program |
| JP5211193B2 (en) * | 2010-11-10 | 2013-06-12 | シャープ株式会社 | Translation display device |
| JP2012173785A (en) * | 2011-02-17 | 2012-09-10 | Nec Corp | Translation result display method, translation result display system, translation result creation device and translation result display program |
| JP5285727B2 (en) * | 2011-02-22 | 2013-09-11 | シャープ株式会社 | Image forming apparatus and image forming method |
| JP2013020585A (en) * | 2011-07-14 | 2013-01-31 | Sharp Corp | Device control program, recording medium, device control method and controller |
| US20140006004A1 (en) * | 2012-07-02 | 2014-01-02 | Microsoft Corporation | Generating localized user interfaces |
| CN102833449B (en) * | 2012-07-27 | 2015-06-10 | 富士施乐实业发展(中国)有限公司 | Automatic document processing method based on multifunctional machine |
| CN113297859A (en) * | 2020-10-19 | 2021-08-24 | 阿里巴巴集团控股有限公司 | Table information translation method and device and electronic equipment |
| CN115759006A (en) * | 2022-11-21 | 2023-03-07 | 网易有道信息技术(北京)有限公司 | Document rendering method, device, electronic device and storage medium |
-
2005
- 2005-03-22 JP JP2005082047A patent/JP2006268150A/en active Pending
- 2005-08-18 CN CN2005100928181A patent/CN1838112B/en not_active Expired - Fee Related
Non-Patent Citations (2)
| Title |
|---|
| JP特开平9114835A 1997.05.02 |
| JP特开昭60-20283A 1985.02.01 |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN105159868A (en) * | 2015-09-01 | 2015-12-16 | 广东欧珀移动通信有限公司 | Text display method and system |
Also Published As
| Publication number | Publication date |
|---|---|
| CN1838112A (en) | 2006-09-27 |
| JP2006268150A (en) | 2006-10-05 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US7865353B2 (en) | Translation device, image processing device, translation method, and recording medium | |
| CN1838112B (en) | translation device, translation method | |
| US20090180126A1 (en) | Information processing apparatus, method of generating document, and computer-readable recording medium | |
| JP5594269B2 (en) | File name creation device, image forming device, and file name creation program | |
| JP2000267829A (en) | Information processing apparatus, control method therefor, and storage medium | |
| JP2006093917A (en) | Image reading apparatus and image processor, and image forming apparatus | |
| JP2008236250A (en) | Image processing apparatus, program, and image processing method | |
| JP2006252048A (en) | Translation apparatus, translation program, and translation method | |
| JP2019029823A (en) | Image processing apparatus, image forming apparatus, and program | |
| JP4992216B2 (en) | Translation apparatus and program | |
| JP2018151699A (en) | Information processing device and program | |
| JP2009223363A (en) | Document processor and document processing program | |
| JP6205973B2 (en) | Change history output device, program | |
| JP2007148486A (en) | Method for supporting document browsing, system for the same, document processor, and program | |
| JP4797507B2 (en) | Translation apparatus, translation system, and program | |
| JP2006276905A (en) | Translation device, image processing device, image forming device, and translation method and program | |
| JP2023011397A (en) | Translation support device and image forming device | |
| JP2001202362A (en) | Character editing processor | |
| JP2026042577A (en) | Image forming apparatus and image forming system | |
| US20050134897A1 (en) | Image forming apparatus | |
| US20230325126A1 (en) | Information processing apparatus and method and non-transitory computer readable medium | |
| JP2020005061A (en) | Image processing device and program | |
| JP4111202B2 (en) | Image forming apparatus | |
| JP3424942B2 (en) | Bilingual image forming device | |
| JP2025166395A (en) | Information processing device, control method for information processing device, program, and information processing system |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| C06 | Publication | ||
| PB01 | Publication | ||
| C10 | Entry into substantive examination | ||
| SE01 | Entry into force of request for substantive examination | ||
| C14 | Grant of patent or utility model | ||
| GR01 | Patent grant | ||
| CF01 | Termination of patent right due to non-payment of annual fee | ||
| CF01 | Termination of patent right due to non-payment of annual fee |
Granted publication date: 20120530 Termination date: 20170818 |