JPH01114971A - Device for supporting document formation/calibration - Google Patents
Device for supporting document formation/calibrationInfo
- Publication number
- JPH01114971A JPH01114971A JP62274154A JP27415487A JPH01114971A JP H01114971 A JPH01114971 A JP H01114971A JP 62274154 A JP62274154 A JP 62274154A JP 27415487 A JP27415487 A JP 27415487A JP H01114971 A JPH01114971 A JP H01114971A
- Authority
- JP
- Japan
- Prior art keywords
- value
- evaluation
- evaluated
- parameter
- values
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Pending
Links
Landscapes
- Machine Translation (AREA)
- Document Processing Apparatus (AREA)
Abstract
Description
【発明の詳細な説明】
〈産業上の利用分野〉
本発明は、日本語文章の中から、誤りや不適当な表現の
部分を抽出したり、ガイダンスや修正情報を搗示したり
、文書作成時に必要な!!鴬やその他の情報を提供する
ことにより、文書を作成する文書作成・校正支援装置に
関するものである。また、かな漢字交じり文を言語解析
し、かな文字列を得、その結果を利用して音声合成を行
う、音声読みあげ装置としての利用も可能である。[Detailed Description of the Invention] <Industrial Application Field> The present invention is useful for extracting errors and inappropriate expressions from Japanese texts, providing guidance and correction information, and necessary! ! The present invention relates to a document creation/proofreading support device that creates documents by providing information such as information and other information. It can also be used as a speech-to-speech device that performs language analysis on sentences containing kana and kanji, obtains kana character strings, and performs speech synthesis using the results.
〈従来の技術〉
現在、日本語ワードプロセッサ(以下、ワープロと略す
る)が実用化されており、それに関連した、日本語の入
出力、編集、かな漢字変換アルゴリズム、辞書の技術な
どの基本技術が確立している。<Conventional technology> Currently, Japanese word processors (hereinafter referred to as word processors) are in practical use, and related basic technologies such as Japanese input/output, editing, kana-kanji conversion algorithms, and dictionary technology have been established. are doing.
かな漢字変換方式のワープロは、変換用の辞書が必要で
あり、ユーザが独自に単語を追加できるユーザ辞書の技
術が確立している。Word processors that use the kana-kanji conversion method require a dictionary for conversion, and user dictionary technology has been established that allows users to add their own words.
また、日本語処理技術では、形態素解析、構文解析、意
味解析などの基本的な技術が知られている。Furthermore, basic techniques such as morphological analysis, syntactic analysis, and semantic analysis are known in Japanese language processing technology.
欧米ではワープロが早くから普及したため関連技術が進
んでおり、スペルのチエツク、コレクトの機能を持った
装置が実用化されている。In Europe and the United States, word processors became popular early on, so related technology has advanced, and devices with spell check and correct functions have been put into practical use.
また、現在のところ開発中あるいは試作の段階であるが
、形態素情報を用いることにより、文章の読みやすさを
装置が診断し、使用者に知らせることで分かりやすい文
章を書くことに対する支援を目指した装置も発表されて
いる。In addition, although it is currently in the development or prototype stage, the device uses morphological information to diagnose the readability of sentences, and aims to support users in writing easy-to-understand sentences by notifying the user. The device has also been announced.
欧米の言葉が単語単位に句切られて記述されるのに比べ
、日本語は句切りのない漢字仮名交り文で記述されるの
が通常であり、更に、正書法が徹底していないことも加
わって、解析が難しく校正を自動化する装置は実用化さ
れていない。Compared to Western words, which are written with punctuation into each word, Japanese words are usually written in kanji, kana, and kanji letters without punctuation, and the orthography is not very strict. In addition, analysis is difficult and no equipment for automating calibration has been put into practical use.
従来、正確な日本語を扱うことが要求される場合、曳敢
の人が対になり読み合わせをして問題のある部分を抽出
したり、あるいは校正の専門的な知識を持った人が逐次
照合を加え、校正する方法などが用いられている。Traditionally, when accurate Japanese was required, a pair of translators would work together to read the text and extract problematic parts, or someone with specialized proofreading knowledge would cross-check the text one by one. Methods such as adding and calibrating are used.
最近、このような校正作業を支援するための装置が開発
されつつある。Recently, devices for supporting such calibration work are being developed.
以上、言語処理を中心とした従来技術について述べたが
、該技術以外では、ワークステーションに関連した技術
が確立しており、作業の効率を上げるためのマルヂウイ
ンドウを用いた装置が実用化されている。Above, we have described conventional technologies centered on language processing, but in addition to these technologies, technologies related to workstations have been established, and devices using multi-windows to improve work efficiency have been put into practical use. ing.
〈発明が解決しようとする問題点〉
近年日本語のワープロが普及し、該装置で作成した文書
が多くなっている。ワープロでは、入力の簡便な、かな
漢字変換方式(以下特に断りがない場合、ローマ字漢字
変換方式を含む)を採用した機種が多くなっている。<Problems to be Solved by the Invention> In recent years, Japanese word processors have become widespread, and an increasing number of documents are created using these devices. Many word processors are now using the kana-kanji conversion method (hereinafter, unless otherwise specified, this includes the romaji-kanji conversion method), which makes input easier.
かな漢字変換のアルゴリズムは、かなを漢字に変換する
過程で単語辞書、文法などの言語的な情報、単語の出現
頻度などの確率的な情報を利用するのが一般的であり、
常に正しい変換結果が出力される訳ではない。Kana-Kanji conversion algorithms generally use word dictionaries, linguistic information such as grammar, and probabilistic information such as word frequency in the process of converting kana to kanji.
Correct conversion results are not always output.
このため、ワープロなどの使用者が間違った変換結果に
気付かなかったり、間違いには気付いても修正の作業を
失念したり、あるいは単1こ人力の間違いを起こしたり
することにより、間違った文章を作成することが生じる
場合がある。For this reason, users of word processors, etc., may not notice incorrect conversion results, may forget to correct the errors even after noticing them, or may make a single human error, resulting in incorrect sentences. It may occur that you create one.
このような間違いの部分を抽出し、修正の情報を提示し
たり、自動的に修正するのが文書作成・校正支援装置で
ある。A document creation/proofreading support device extracts such mistakes, presents correction information, and automatically corrects them.
該装置の一つの機能として、形態素の情報を用いて、数
値あるいは相対的な値で文章の読みやすさを表し、その
結果を装置の使用者に提示することにより、文章の改良
を行う方法が機業されている。現在、知られている方法
は、読みやすさのパラメータと評価文でのパラメータの
測定値と、全体的な読みやすさの指標を提示する方法で
ある。One of the functions of this device is a method for improving sentences by expressing the readability of sentences using morpheme information in numerical or relative values and presenting the results to the device user. It is machined. A currently known method is to present readability parameters, measured values of the parameters in evaluation sentences, and an overall readability index.
ところが従来の方法は、数値が表示されるだけであり、
その数値の持つ意味を理解しにくいという欠点があった
。However, conventional methods only display numerical values,
The drawback was that it was difficult to understand the meaning of the numbers.
また、評価文章で、パラメータの値が超過しているのか
、不足しているのか、適当であるかの判別がしにくいと
ういう欠点があった。Another drawback is that it is difficult to determine whether a parameter value exceeds, falls short of, or is appropriate in the evaluation text.
また、文章全体に対し、どの項目に問題があるが鳥かん
しにくという欠点があった。In addition, there was a drawback that it was difficult to understand which items had problems in relation to the entire text.
更に、問題点が明確に表示されないため、文章を評価し
てもどのような部分を修正していけば良いかの指針が得
られないという欠点を有していた。Furthermore, since problems are not clearly displayed, there is a drawback that even if the text is evaluated, it is not possible to obtain guidelines as to what parts should be corrected.
本発明は、文章評価のパラメータの値を、グラフなどで
視覚化するとともに評価の程度に応じて区別して表示す
ることにより、かかる問題を解決しようとするものであ
る。The present invention attempts to solve this problem by visualizing the values of text evaluation parameters using a graph or the like and displaying them separately according to the degree of evaluation.
く問題点を解決するための手段〉
本発明は、日本語を入力・編集する手段と、該入力され
た日本論を記憶する手段と、辞書を記憶する手段と、文
法を記憶する手段と、該入力されたかな文字列を漢字交
じり文に変換したり、編集したりするマイクロプロセッ
サなどの制御手段と、文字・記号列などを表示する手段
と、校正すべき文字・記号列がある場合に該文字列を修
正する手段から構成される。Means for Solving Problems> The present invention provides a means for inputting and editing Japanese, a means for storing the input Japanese theory, a means for storing a dictionary, a means for storing grammar, A control means such as a microprocessor that converts and edits the input kana character string into a sentence mixed with kanji, a means for displaying a character/symbol string, and a character/symbol string that needs to be proofread. It consists of means for modifying the character string.
〈作用〉
本発明は、評価対象文を文字コードや言語解析機能で分
析し、その結果を表示するときに、評価パラメータとそ
の実測値をグラフとともに表示し、見易くするように作
用する。<Operation> The present invention operates so that when a sentence to be evaluated is analyzed using a character code or a language analysis function and the results are displayed, the evaluation parameters and their actual measured values are displayed together with a graph for easy viewing.
また、グラフを、評価パラメータの推奨値との比較によ
って区分することにより、文章中の間厘点をユーザに分
かりやすく提示するように作用する。Furthermore, by dividing the graph by comparing it with the recommended value of the evaluation parameter, it is possible to present the intermediate score in the sentence in an easy-to-understand manner to the user.
〈実施例〉
以下図に基づいて本発明の詳細な説明する。第1図は本
発明に係わる文書作成・校正支援装置のブロック構成図
である。<Example> The present invention will be described in detail below based on the drawings. FIG. 1 is a block diagram of a document creation/proofreading support device according to the present invention.
図においてlは日本語の文字列を入力・編集するキーボ
ードなどの手段である。この中には、現在では周知の事
実になっているかなを漢字に変えるかな漢字変換機能、
ある文字列を指定する機能も含まれる。In the figure, l is a means such as a keyboard for inputting and editing Japanese character strings. This includes a kana-kanji conversion function that turns kana into kanji, which is now a well-known fact.
It also includes the ability to specify a certain string.
2は該入力手段により入力された日本語の文字列を記憶
する手段である。入力手段は通常キーボードが用いられ
るが逐次的に入力を行なわないで、たとえばフロッピー
ディスク、磁気テープなどのように入力した日本語の文
字列を記憶する外部記憶手段で代用することも可能であ
る。即ち、lの入力手段が省略された構成も存在しうる
。2 is a means for storing the Japanese character string inputted by the input means. A keyboard is usually used as the input means, but it is also possible to use an external storage means such as a floppy disk or magnetic tape for storing input Japanese character strings without sequential input. That is, there may also be a configuration in which the l input means is omitted.
3は上記2に蓄積された日本語の文字・記号列を解析す
るための辞書を記憶する手段であり、この中にはユーザ
が定義でき登録、消去の出来るユーザ辞書を記憶する手
段も含まれる。3 is a means for storing a dictionary for analyzing the Japanese character/symbol strings accumulated in 2 above, and this also includes a means for storing a user dictionary that can be defined, registered, and deleted by the user. .
4は文法、その他の文章を解析するための規則類を記憶
する手段である。4 is a means for storing rules for analyzing grammar and other sentences.
5は2に蓄えられた文字列の中の一部分を抽出したり、
途中結果を記憶したり、表示の司令などを行ったりする
制御手段である。該制御手段には制御によって得られる
結果を記憶する手段を含む。5 extracts a part of the string stored in 2,
This is a control means that stores intermediate results and commands display. The control means includes means for storing results obtained by the control.
6は入力された文字列、照合の途中結果、校正すべき文
字列、KWIC(キーワードイン コンチクスト)など
を表示するCRTなどの表示の手段である。Reference numeral 6 denotes display means such as a CRT for displaying input character strings, intermediate results of verification, character strings to be proofread, KWIC (Keyword In Context), and the like.
7は6によって表示された校正すべき部分に対し修正を
加えた結果を原文中に正しく反映するための校正手段で
ある。Reference numeral 7 denotes a proofreading means for correctly reflecting the results of corrections made to the portion to be proofread indicated by 6 in the original text.
以下、具体的な説明を行うために、文章の読みやすさの
評価を行う場合について述べる。ただし、本発明は文章
の読みやすさの評価に止どまらず、広く、基阜値が設定
された色々な評価を行うときに適用可能なものである。In order to provide a concrete explanation, a case will be described below in which the readability of a text is evaluated. However, the present invention is not limited to evaluating the readability of sentences, but can be broadly applied to various evaluations for which reference values are set.
第2図は該装置の表示手段に表示された校正の解析をし
た結果を表す図である。8は読みやすさの評価対象とな
る文であり、図中斜線を施している部分は該装置が誤り
の可能性のある場所として抽出したことを示している。FIG. 2 is a diagram showing the results of analyzing the calibration displayed on the display means of the apparatus. 8 is a sentence to be evaluated for readability, and the shaded portion in the figure indicates that the device has extracted a location where there is a possibility of an error.
図では説明のために代表的な例を上げている。9は誤字
の例であり、かな漢字変換ソフトウェアの不備などによ
り起こる同音異義語の間違いの例である。10は同音の
かなの間違いの例であり、11は送り仮名の間違いの例
であり、12は脱字の例であり、13は不要文字の挿入
の例である。The figure shows a typical example for explanation. 9 is an example of a typographical error, which is an example of a homonym error caused by a defect in the kana-kanji conversion software. 10 is an example of a mistake in the same sound kana, 11 is an example of a mistake in okurikana, 12 is an example of an omission, and 13 is an example of insertion of an unnecessary character.
第3図は読みやすさを表すパラメータを格納するバッフ
ァの構造の例を表した図である。14は文章の読みやす
さを表すパラメータ、15は評価する文章の各パラメー
タの測定値、16には各パラメータの推奨値、17には
各パラメータの評価値が入る。FIG. 3 is a diagram showing an example of the structure of a buffer that stores parameters representing readability. Reference numeral 14 contains a parameter representing the readability of the text, 15 contains the measured value of each parameter of the text to be evaluated, 16 contains the recommended value of each parameter, and 17 contains the evaluation value of each parameter.
第3図のパラメータの中で、18は形態素情報を表し、
19は校正情報を表している。18の例として、第3図
では、含有漢字の割合、平均の文の長さ、平均の句読点
の頻度を上げ、また、19の例として校正文節出現比を
上げている。Among the parameters in FIG. 3, 18 represents morpheme information,
19 represents calibration information. As an example of No. 18, in FIG. 3, the proportion of Chinese characters included, the average sentence length, and the average frequency of punctuation marks are increased, and as an example of No. 19, the ratio of occurrences of proofread clauses is increased.
ここで、第3図のパラメータの測定値15について具体
的に説明する。評価対象文章8の全文字数は句読点を含
めて45文字である。全文字数として、句読点を除外し
て計算する方法もあるが、これから述べる方法とおなし
ようなやりかたで値を求めることができる。8の中の漢
字は20文字であり、含有漢字の割合は少数第2位まで
出すと、20/45=0.444となり、44.4%が
値となる。文の長さの平均は45/2=22.5、同様
に平均の句読点の頻度は、45/2=22゜5となり、
校正文節出現比=5/14=0.36である。Here, the parameter measurement value 15 in FIG. 3 will be specifically explained. The total number of characters in the evaluation target sentence 8 is 45 including punctuation marks. Although there is a method to calculate the total number of characters excluding punctuation marks, you can calculate the value using a method similar to the method described below. There are 20 kanji in 8, and the percentage of kanji included is 20/45=0.444, which is 44.4%, if you calculate it to the second decimal place. The average sentence length is 45/2 = 22.5, and the average punctuation frequency is 45/2 = 22°5.
The proof clause appearance ratio = 5/14 = 0.36.
次に推奨値について、説明する。推奨値は、個人によっ
て若干異なりはあるが、経験的に、民族の常識的な値と
して決定されることが多い。これは、−膜内にわれわれ
が用いている言葉は最初に文法、用語を決定してから実
際の言語の運用を行うのではなく、民族の意志のまとま
りとして言語の体系が順次できるのに似ている。Next, the recommended values will be explained. Although the recommended value differs slightly depending on the individual, it is often determined empirically as a common sense value for each ethnic group. This is similar to how a language system is created sequentially as a group of people's will, rather than first determining the grammar and terminology for the words we use within the membrane and then proceeding with the actual use of the language. ing.
そのような値として、漢字比率、平均の文の長さ、平均
の句読点の頻度、校正抽出個数の各推奨値RISR2、
R3、R4とするとき15のように
25≦R1≦45、25≦R2≦45
7 ≦R3≦ 15 R4≦ 0.1を得る。ただし
、本発明は、この値の絶対値そのものを云々するもので
はない。Such values include the recommended values of kanji ratio, average sentence length, average punctuation frequency, and number of proofreading extractions, RISR2,
When R3 and R4 are 15, we obtain 25≦R1≦45, 25≦R2≦45 7≦R3≦15 R4≦0.1. However, the present invention does not refer to the absolute value of this value itself.
次に、評価値は、推奨値と評価対象文の実測値との差を
何等かの評価関数で評価したものである。Next, the evaluation value is obtained by evaluating the difference between the recommended value and the actual value of the sentence to be evaluated using some evaluation function.
説明を簡単にするため、ここでは、実測値が推奨値の値
であれば100、そうでなければ、0という極端な例を
上げる。第3図の17はそのようにして出された値であ
る。To simplify the explanation, an extreme example will be given here in which the value is 100 if the actual measured value is the recommended value, and 0 otherwise. 17 in FIG. 3 is the value obtained in this way.
評価値の値の出しかたの別の例は、台形型の近似である
。これは指定された区域の間を評価値100とし、指定
された値から外れる処は線形近似をとるやり方である。Another example of how to calculate the evaluation value is trapezoidal approximation. This is a method in which the evaluation value is set to 100 between designated areas, and linear approximation is applied to areas that deviate from the designated values.
当然、外れた部分に対し、非線形の関数を割り当てる方
法もある。本発明は、近似する関数そのものを限定して
いないので、各種の近似関数が適用できることのみの言
及に止どめておく。Naturally, there is also a method of assigning a nonlinear function to the deviant portion. Since the present invention does not limit the approximation function itself, it will only be mentioned that various approximation functions can be applied.
第4図は本発明に係わる文章の読みやすさの評価結果を
表した図である。FIG. 4 is a diagram showing the results of evaluating the readability of sentences according to the present invention.
縦軸は評価パラメータを、横軸はパラメータの相対的な
値を示している。20は具体的なパラメータの例を表し
、21はパラメータの値を示している。The vertical axis shows the evaluation parameters, and the horizontal axis shows the relative values of the parameters. 20 represents a specific example of a parameter, and 21 represents a value of the parameter.
22は21の値に対応する棒グラフとなっている。パラ
メータの値は、0から100までに各パラメータの推奨
値が来るように正規化されている。22 is a bar graph corresponding to the value of 21. The parameter values are normalized so that the recommended value for each parameter ranges from 0 to 100.
実際の表示の際は、パラメータの値が100を越える場
合は100を指し、0以下の場合は0を指すように設計
されている。In actual display, the parameter value is designed to indicate 100 if it exceeds 100, and to indicate 0 if it is less than or equal to 0.
第4図は棒グラフで読みやすさの評価結果を表したが、
棒グラフでなくても本発明には影響しない。たとえば、
他の表現の例は、レーダチャートである。Figure 4 shows the readability evaluation results using a bar graph.
The present invention is not affected even if the graph is not a bar graph. for example,
Another example of a representation is a radar chart.
以上の準備をしたところで、本発明の核心部分について
述べる。Now that the above preparations have been made, the core part of the present invention will be described.
第5図(a)は、第4図の含有漢字比率の部分のみを取
り出した図である。また、第5図(b)、(C)はそれ
ぞれ、別の評価文を評価し、その値が50%であるとき
、10%であるときに対応した図である。混同を避ける
ため、棒グラフの長さと、数値に番号を振っておく。2
3は、評価文8を評価したときの含有漢字比率の棒グラ
フの長さであり、24は数値である。25は、含有漢字
比率が50の別の文章を評価したときの棒グラフの長さ
であり、26はその値である。27は含有漢字比率がl
Oの別の文章を評価したときの棒グラフの長さであり、
28はその値である。FIG. 5(a) is a diagram in which only the portion of the content ratio of kanji characters in FIG. 4 is extracted. Moreover, FIGS. 5(b) and 5(C) are diagrams corresponding to when different evaluation sentences are evaluated and the value is 50% and 10%, respectively. To avoid confusion, number the lengths of the bar graphs and numbers. 2
3 is the length of the bar graph of the proportion of kanji characters included when evaluating the evaluation sentence 8, and 24 is a numerical value. 25 is the length of a bar graph when another sentence containing 50 kanji characters is evaluated, and 26 is its value. 27 has a kanji ratio of l
It is the length of the bar graph when evaluating another sentence of O,
28 is its value.
文章の読みやすさのパラメータを、棒グラフと実際の値
を併記して、視覚的に見やすくしたことは本発明の一つ
の特徴である。One of the features of the present invention is that the readability parameters of sentences are shown in bar graphs and actual values to make them easier to visually see.
第6図(a)、(b)、(C)はそれぞれ、第5図(a
)、(b)、(C)に対応したもので棒グラフがその値
に応じて、区別して表示されている。29は含有漢字比
率が44.4%の文章8の時の、欅グラフの長さであり
、30はその値である。31は含有漢字比率が50%の
文章の棒グラフの長さであり、32はその値である。3
3は含有漢字比率が10%の文章の棒グラフの長さであ
り、34はその値である。Figures 6(a), (b), and (C) are respectively shown in Figure 5(a).
), (b), and (C), and bar graphs are displayed differently according to their values. 29 is the length of the Keyaki graph for sentence 8 with a kanji content ratio of 44.4%, and 30 is its value. 31 is the length of a bar graph of a sentence containing 50% kanji, and 32 is its value. 3
3 is the length of a bar graph of a sentence containing 10% of kanji characters, and 34 is its value.
ここで、このような表示をいかにして実現しているかに
ついて説明する。今、例にあげている含有漢字比率の推
奨値R1は
25≦rtt≦45
で表される。Here, a description will be given of how such a display is realized. The recommended value R1 of the ratio of kanji included in the example given now is expressed as 25≦rtt≦45.
rttと評価文章における含有漢字比率め値とを比較す
ることにより、次の3つに分類できる。即ち、含有漢字
比較が推奨値より大きい、等しい、小さいの3つである
。By comparing the rtt and the percentage of kanji included in the evaluated text, it can be classified into the following three types. That is, the included kanji comparison is greater than, equal to, or less than the recommended value.
この3種類の値に応じて区別して表示すれば、パラメー
タが推奨値に対しどのような状態にあるかを判別しやす
くなる。If these three types of values are distinguished and displayed, it becomes easier to determine what state the parameter is in with respect to the recommended value.
第6図は含有漢字比率が推奨値より 小さければ璽震を 等しければ を 大きければ4HHHを用いた場合を示したものである。Figure 6 shows the proportion of kanji included is higher than the recommended value. If it is small, use a seal If they are equal, The larger value indicates the case where 4HHH is used.
第6図では、区別情報を網掛けのg類で表示したが、区
別できればこれにこだわる必要はなく、たとえば、色で
区別したりすることでも可能である。In FIG. 6, the distinction information is displayed by shaded G types, but there is no need to stick to this as long as it can be distinguished; for example, it is also possible to distinguish by color.
パラメータの推奨値が不等号1個しか持たないような場
合は、区別の種類は2つになるが、3種類の場合と同様
の処理で本発明を実現できる。In a case where the recommended value of a parameter has only one inequality sign, there are two types of distinctions, but the present invention can be implemented by the same processing as in the case of three types.
第7図は本発明の別の表現例である。これも、文章の読
みやすさの中の含有漢字比率の場合を例にとって説明す
る。35は推奨値より小さい値の範囲であることを示し
た領域であり、36は推奨値と等しい領域であり、37
は推奨値より大きい領域であり、38は実際の評価値を
示す記号であり、39はその数値である。FIG. 7 is another example of expression of the present invention. This will also be explained using the case of the ratio of kanji included in the readability of a text as an example. 35 is an area indicating a range of values smaller than the recommended value, 36 is an area equal to the recommended value, and 37
is an area larger than the recommended value, 38 is a symbol indicating the actual evaluation value, and 39 is its numerical value.
この図の場合は、パラメータの推奨値の範囲が区分され
ており、また評価値も明確に示されるため評価文章が適
当であるか、あるいは、どのような問題を含んでいるの
かの判断がしやすくなっている。In the case of this figure, the range of recommended values for the parameters is divided and the evaluation values are also clearly shown, making it easy to judge whether the evaluation text is appropriate or what kind of problems it contains. It's getting easier.
第8図は、本発明の別の表示の例である。これは、第5
図(a)に、パラメータの推奨値を並行して表示した図
である。比較のために第5図(a)の番号も併せて示し
ている。40はパラメータの推奨値を表わしている。FIG. 8 is an example of another display of the present invention. This is the fifth
FIG. 3A is a diagram in which recommended values of parameters are displayed in parallel with FIG. For comparison, the numbers in FIG. 5(a) are also shown. 40 represents recommended values of parameters.
第9図は本発明の該略フロー図である。種々の表示の方
法を提案してきたが、ここでは、その中から第6図(a
)の場合を例にとって説明する。FIG. 9 is a schematic flow diagram of the present invention. We have proposed various display methods, but here we will introduce the method shown in Figure 6 (a).
) will be explained using the case as an example.
まず、該装置の表示装置6へ手段2に記憶された評価文
章8を呼び出す。もし、文章がない場合は手段lにより
入力する。この評価文章表示処理ブロックを41とする
。First, the evaluation text 8 stored in the means 2 is called up on the display device 6 of the device. If there is no text, input it using means 1. This evaluation text display processing block is designated as 41.
次に、文章の読みやすさを表すためのパラメータになっ
ている言語情報を出すための言語解析処理を行う。この
言語情報作成のブロックを42とする。Next, language analysis processing is performed to generate linguistic information that is a parameter for expressing the readability of the text. This linguistic information creation block is designated as 42.
次に、1,3.4.5の各手段を用いて、評価文章を解
析し、校正の可能性のある場所を抽出する。この処理に
は、校正解析処理が実行される。Next, the evaluation text is analyzed using each of the methods described in 1, 3, 4, and 5, and locations where proofreading is possible are extracted. In this process, a calibration analysis process is executed.
この校正解析処理ブロックを43とする。This calibration analysis processing block is designated as 43.
次に、形態素解析、校正解析の処理結果から、読みやす
さを表す各パラメータの実測値を求め、制御手段の中の
バッファに格納する。この処理ブロックを44とする。Next, from the processing results of the morphological analysis and the calibration analysis, actual measured values of each parameter representing readability are obtained and stored in a buffer in the control means. This processing block is assumed to be 44.
次に、バッファの中の評価値を第4図の棒グラフで表現
するとともに、評価値そのものを表示する。この処理ブ
ロックを45とする。Next, the evaluation values in the buffer are expressed in the bar graph of FIG. 4, and the evaluation values themselves are displayed. This processing block is assumed to be 45.
次に、評価パラメータの実測値と推奨値との比較を行う
。この処理ブロックを46とする。この比較により、3
つに分岐する〇
それぞれの分岐に対し、棒グラフの表示方法をたとえば
、A、B、Cのように変える。Next, the actual measured values and recommended values of the evaluation parameters are compared. This processing block is assumed to be 46. This comparison shows that 3
Branch into 〇 For each branch, change the bar graph display method, for example, A, B, C.
このA、BSCの各区分表示を、それぞれ47.48.
49とする。The A and BSC classification display is 47.48.
49.
この区分の表示方法は、制御手段の中に手続きとして記
述しておいても良いし、テーブル形式にして区分表示の
種類ASBSCをキーとして手続きを検索する方法で実
現できる。This method of displaying the classification may be described as a procedure in the control means, or it can be realized by searching for a procedure in a table format using the classification display type ASBSC as a key.
50は終了処理のブロックである−0この処理は、次文
章の評価の準備の処理も含めてあり、評価処理のモード
から抜は出すか、評価する文章が無くなるまで繰り返さ
れた後に実行される。50 is the end processing block - 0 This processing also includes processing to prepare for the evaluation of the next sentence, and is executed after being repeated until the evaluation processing mode is removed or there are no more sentences to evaluate. .
上の、フローと異なるフローの例を以下に述べる。上で
は、文章の解析を42.43の順に行うようした。これ
は、校正の解析に言語解析の結果を使用する場合が多い
ことから決定したもので、パラメータの値が求まればい
ずれから行っても良い。An example of a flow different from the above flow is described below. Above, the sentences are analyzed in the order of 42.43. This was decided because the results of language analysis are often used for proofreading analysis, and any parameter value can be used.
〈発明の効果〉
本発明の効果は、推奨できるパラメータの値を有する文
章の評価を行うときに、評価値のみでなく、グラフ化し
て表示されるので見易い点にある。<Effects of the Invention> An advantage of the present invention is that when evaluating a text having recommendable parameter values, not only the evaluation value but also the evaluation value is displayed in the form of a graph, making it easy to see.
また、評価値と推奨値との比較により、パラメータに対
し、評価文章がどのような質的な評価を受けているかが
簡単に確認できる点でも効果がある。It is also effective in that by comparing the evaluation value and the recommended value, it is possible to easily check what kind of qualitative evaluation the evaluation text has received with respect to the parameter.
これらの評価結果を基に、文章のどのパラメータの問題
あるいは推奨の程度が簡単に判断でき、それに合わせて
文章の修正、追加などができる点でも効果がある。Based on these evaluation results, it is possible to easily determine which parameter of the text is a problem or the degree of recommendation, and it is also effective in that the text can be modified or added accordingly.
第1図は本発明装置の購成ブロック図、第2図は評価文
章の例を示す図、第3図バッファの構造及び各項目の値
を示す図、第4図は文章の読みやすさを評価した結果を
表示を示す図、第5図(a)、(b)、(c)はそれぞ
れ、読みやすさパラメータの中の含何漢字比率の違う文
章を評価した結果の例を示す図、第6図(a)、(b)
、(c)はそれぞれ、評価値によりパラメータの程度を
区分して表示した例を示す図、第7図及び第8図は本発
明の他の実施例の表示例を示す図、第9図は本発明の概
略フロー図である。Figure 1 is a purchasing block diagram of the device of the present invention, Figure 2 is a diagram showing an example of an evaluation text, Figure 3 is a diagram showing the structure of the buffer and the values of each item, and Figure 4 is a diagram showing the readability of the text. Figures 5(a), 5(b), and 5(c) are diagrams showing examples of the results of evaluating sentences with different proportions of kanji in the readability parameter. Figure 6 (a), (b)
, (c) are diagrams showing examples of displaying the degree of parameters according to evaluation values, FIGS. 7 and 8 are diagrams showing display examples of other embodiments of the present invention, and FIG. FIG. 1 is a schematic flow diagram of the present invention.
Claims (1)
記憶する手段と、辞書を記憶する手段と、文法を記憶す
る手段と、該入力された日本語の中から校正すべき文字
・記号列を抽出する手段と、文章及び該候補文字・記号
列などを表示する手段と、校正すべき文字・記号列があ
る場合に該文字を修正する手段を有する文書処理システ
ムにおいて、文書を評価するときに、パラメータの基準
値と実測値との相対関係により、実測データを区別して
表示することを特徴とする文書作成・校正支援装置。A means for inputting/editing Japanese, a means for storing the input Japanese, a means for storing a dictionary, a means for storing grammar, and a method for editing characters and characters to be proofread from the input Japanese. Evaluate a document in a document processing system that has a means for extracting a symbol string, a means for displaying the text and the candidate character/symbol string, and a means for correcting the character/symbol string when there is a character/symbol string to be proofread. 1. A document creation/proofreading support device that distinguishes and displays measured data based on the relative relationship between reference values and measured values of parameters.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP62274154A JPH01114971A (en) | 1987-10-28 | 1987-10-28 | Device for supporting document formation/calibration |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| JP62274154A JPH01114971A (en) | 1987-10-28 | 1987-10-28 | Device for supporting document formation/calibration |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| JPH01114971A true JPH01114971A (en) | 1989-05-08 |
Family
ID=17537781
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP62274154A Pending JPH01114971A (en) | 1987-10-28 | 1987-10-28 | Device for supporting document formation/calibration |
Country Status (1)
| Country | Link |
|---|---|
| JP (1) | JPH01114971A (en) |
Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS5773417A (en) * | 1980-10-27 | 1982-05-08 | Mitsubishi Electric Corp | Numerical controller |
| JPS57201908A (en) * | 1981-06-04 | 1982-12-10 | Mitsubishi Electric Corp | Operation monitoring device for plant equipment or the like |
| JPS58114126A (en) * | 1981-12-28 | 1983-07-07 | Canon Inc | printer |
| JPS60254367A (en) * | 1984-05-31 | 1985-12-16 | Fujitsu Ltd | Sentence analyzer |
| JPS6126137A (en) * | 1984-07-16 | 1986-02-05 | Toshiro Kutsuwa | 2's complement displayed parallel multiplication/division system |
-
1987
- 1987-10-28 JP JP62274154A patent/JPH01114971A/en active Pending
Patent Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JPS5773417A (en) * | 1980-10-27 | 1982-05-08 | Mitsubishi Electric Corp | Numerical controller |
| JPS57201908A (en) * | 1981-06-04 | 1982-12-10 | Mitsubishi Electric Corp | Operation monitoring device for plant equipment or the like |
| JPS58114126A (en) * | 1981-12-28 | 1983-07-07 | Canon Inc | printer |
| JPS60254367A (en) * | 1984-05-31 | 1985-12-16 | Fujitsu Ltd | Sentence analyzer |
| JPS6126137A (en) * | 1984-07-16 | 1986-02-05 | Toshiro Kutsuwa | 2's complement displayed parallel multiplication/division system |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| Anderson et al. | A cross-linguistic database of phonetic transcription systems | |
| US7110939B2 (en) | Process of automatically generating translation-example dictionary, program product, computer-readable recording medium and apparatus for performing thereof | |
| JPH07325824A (en) | Grammar check system | |
| KR20000057355A (en) | Method and system for unambiguous braille input and conversion | |
| JPH08235182A (en) | Text processing method and device | |
| JPH01114972A (en) | Device for supporting document formation/calibration | |
| Li et al. | Transbench: Benchmarking machine translation for industrial-scale applications | |
| JPH10301933A (en) | Document processor, its method and recording medium | |
| JPH0916597A (en) | Text reviewing device and method | |
| JPH01114974A (en) | Device for supporting document formation/calibration | |
| Garabík et al. | A cross linguistic database of children's printed words in three Slavic languages | |
| JPH06266769A (en) | Synonym information creation device | |
| JP4229457B2 (en) | Data display device and data display method | |
| Karim | Arabic Tā'MarbŭṭAh in Latin Transliteration for Digital Communication: A New Proposed Character | |
| JPS63118868A (en) | Proofreading device for japanese sentence | |
| JP2818185B2 (en) | Document creation support device | |
| Gui et al. | Chinese Braille Translation System Based on Bidirectional Long Short-Term Memory-Conditional Random Field Algorithm | |
| JPS63163956A (en) | Document preparation and correction supporting device | |
| JPH0696117A (en) | Document change support system | |
| JPH05204299A (en) | Teaching device for language | |
| JPH033064A (en) | Character processor | |
| JPH0628396A (en) | Electronic dictionary | |
| JP3220133B2 (en) | Kana-Kanji conversion device | |
| JPH01114973A (en) | Device for supporting document formation/calibration | |
| JPS63229561A (en) | Back-up device for production/correction of document |