JPH0481198B2 - - Google Patents

Info

Publication number
JPH0481198B2
JPH0481198B2 JP59062806A JP6280684A JPH0481198B2 JP H0481198 B2 JPH0481198 B2 JP H0481198B2 JP 59062806 A JP59062806 A JP 59062806A JP 6280684 A JP6280684 A JP 6280684A JP H0481198 B2 JPH0481198 B2 JP H0481198B2
Authority
JP
Japan
Prior art keywords
frame
silent
signal
audio
audio signal
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Expired - Lifetime
Application number
JP59062806A
Other languages
Japanese (ja)
Other versions
JPS60205598A (en
Inventor
Takashi Tojama
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
NEC Corp
Original Assignee
Nippon Electric Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Nippon Electric Co Ltd filed Critical Nippon Electric Co Ltd
Priority to JP59062806A priority Critical patent/JPS60205598A/en
Publication of JPS60205598A publication Critical patent/JPS60205598A/en
Publication of JPH0481198B2 publication Critical patent/JPH0481198B2/ja
Granted legal-status Critical Current

Links

Description

【発明の詳細な説明】 〔発明の属する技術分野〕 本発明は、情報処理装置に使われる音声応答装
置のデイジタル音声信号の記憶装置に関する。特
に、文章を単語や文節のフレーズに分解してフレ
ーズ単位で個別に記憶装置に書込む音声信号の記
憶装置に関する。
DETAILED DESCRIPTION OF THE INVENTION [Field of the Invention] The present invention relates to a storage device for digital voice signals of a voice response device used in an information processing device. In particular, the present invention relates to an audio signal storage device that breaks down a sentence into words and phrases and writes each phrase individually into a storage device.

〔従来技術の説明〕[Description of prior art]

音声応答装置ではあらかじめ登録された単語や
文節を組合わせて一連の文章として再生する方式
のものがある。この単語や文節の前縁や後縁に適
当な量の無声部分を持つようにして一連の文章と
して組合わせ再生したときに自然な音声となるよ
うにしている。
Some voice response devices combine pre-registered words and phrases and reproduce them as a series of sentences. An appropriate amount of unvoiced parts are provided at the leading and trailing edges of these words and phrases, so that the sounds will be natural when combined and reproduced as a series of sentences.

従来、このような単語や文節を書込む装置にお
いては、単語や文節の前縁後縁に無音部分を十分
付加した形で記憶装置に書込み、モニタしながら
前縁後縁の無音部分が適当な長さになるように切
取り作業を繰返していた。しかしこの従来の方法
によると無音部分の切取り聴覚にたよる手作業と
なるために、文章として組合せて再生したときに
自然な音声となるように適当な切取りを行うに
は、熟練と多くの時間とを必要とする欠点があつ
た。また、手作業によるため前縁と後縁とに残さ
れた無音部分の長さがまちまちで一定品質のもの
が得られにくい欠点もあつた。
Conventionally, in devices for writing such words and phrases, the words and phrases are written to the storage device in a form with sufficient silent parts added to the leading and trailing edges, and while being monitored, the silent parts at the leading and trailing edges are adjusted as appropriate. The cutting process was repeated until the length was reached. However, with this conventional method, the silent parts are cut out manually and rely on auditory sense, so it takes skill and a lot of time to cut out the silent parts appropriately so that the sound sounds natural when combined and played back. There was a drawback that it required In addition, since it was done by hand, the length of the silent portions left on the leading and trailing edges varied, making it difficult to obtain products of consistent quality.

〔発明の目的〕[Purpose of the invention]

本発明は、上記欠点を除去し、音声信号の前縁
後縁に指定された量の無音部分を確実に確保でき
るデイジタル音声信号の記憶装置を提供すること
を目的とする。
SUMMARY OF THE INVENTION It is an object of the present invention to provide a digital audio signal storage device that eliminates the above-mentioned drawbacks and can reliably ensure a specified amount of silence at the leading and trailing edges of the audio signal.

〔発明の特徴〕[Features of the invention]

本発明は、単語や文節の前縁後縁に付加された
無音部分の不必要分を切取る際に、まず単語や文
節の有音部分の最初および最後の位置を判定し、
次に有音部分の前後に指定した量だけの無音部分
を残すように無音部分の切取りを行うことを特徴
とする。
The present invention first determines the beginning and end positions of the voiced parts of a word or phrase when cutting out unnecessary silent parts added to the leading and trailing edges of the word or phrase,
Next, the silent part is cut out so that a specified amount of silent part is left before and after the sound part.

すなわち、本発明は、入力されるアナログ音声
をデイジタル音声信号に変換し、フレーム単位で
記憶装置へ書込む手段と、外部からの書込み開始
および停止指示により、上記記憶装置への書込み
を制御し、デイジタル音声信号の書込み開始アド
レスと停止アドレスとを記憶する手段と、上記書
込んだデイジタル音声信号の所定数のフレーム中
に無音を示すパターンを設定数以上持つときこの
フレームを無音フレームと決定する手段と、デイ
ジタル音声信号の前縁にNフレーム後縁にMフレ
ーム無音フレームを残すことを指示する入力手段
とを備え、上記デイジタル音声信号の開始アドレ
スより順に1フレームずつ無音フレームの判定を
行い、無音フレームと決定されない最初のフレー
ム位置をVS、また停止アドレスより逆上つて1
フレームずつ無音フレームの判定を行い、無音フ
レームと決定されない最初のフレーム位置をVE
とし、VS−NフレームからVE+Mフレームを有
効な音声フレーズとすることを特徴とする。
That is, the present invention includes means for converting input analog audio into digital audio signals and writing them into the storage device in frame units, and controlling writing to the storage device by external writing start and stop instructions. means for storing a write start address and stop address of the digital audio signal; and means for determining a frame as a silent frame when a predetermined number of frames of the written digital audio signal include a predetermined number or more of patterns indicating silence. and an input means for instructing to leave N frames at the leading edge of the digital audio signal and M frames at the trailing edge of the digital audio signal, and determines the silent frame frame by frame sequentially from the start address of the digital audio signal, VS the first frame position that is not determined to be a frame, and 1 from the stop address
Determine silent frames frame by frame, and VE the first frame position that is not determined to be a silent frame.
The system is characterized in that the VS-N frame to the VE+M frame are valid voice phrases.

なお、ここでフレームとは、デイジタル音声信
号を一定バイト単位で区切つた場合の音声信号の
単位をいう。
Note that a frame here refers to a unit of an audio signal when a digital audio signal is divided into fixed byte units.

〔実施例による説明〕[Explanation based on examples]

本発明の実施例について図面を参照して説明す
る。第1図は本発明一実施例デイジタル音声信号
記憶装置のブロツク構成図である。第1図におい
て、外部から入力される単語または文節であるア
ナログ音声信号S1がアナログデイジタル変換部1
に接続される。アナログデイジタル変換部1から
変換されたデイジタル音声信号がフレーム単位で
音声メモリ部2に与えられる。
Embodiments of the present invention will be described with reference to the drawings. FIG. 1 is a block diagram of a digital audio signal storage device according to an embodiment of the present invention. In FIG. 1, an analog audio signal S1, which is a word or phrase inputted from the outside, is sent to an analog-to-digital converter 1.
connected to. The digital audio signal converted from the analog-to-digital converter 1 is provided to the audio memory unit 2 in units of frames.

ここで本発明の特徴とするところは、一点鎖線
で囲むデイジタル音声信号の記憶部分である。す
なわち、外部から書込み開始信号S2がコントロー
ル部3に入力され、コントロール部3から書込み
開始アドレス、インクリメント指示などの制御信
号が信号線l1を介してカウンタ4に接続される。
書込み停止信号S3がコントロール部3に入力さ
れ、コントロール部3からインクリメント停止指
示の制御信号が信号線l1を介してカウンタ4に接
続される。カウンタ4からの出力信号が信号線l2
を介して音声メモリ部2に接続され、デイジタル
音声信号が蓄積される。またカウンタ4からの出
力信号は分岐されてコントロール部3に接続され
る。音声メモリ部2から蓄積されたデイジタル音
声信号が無音パターン検出部5に接続され、無音
パターン検出信号がコントロール部3に接続され
る。コントロール部3において無音パターン検出
信号がカウントされて、無音フレームが判定さ
れ、最初の有音フレームと最後の有音フレームと
が決定される。外部から無音フレーム指定信号S4
がコントロール部3に入力され、コントロール部
3から指定数の無音フレームを最初の有音フレー
ムの前縁および最後の有音フレームの後縁に残す
制御信号が信号線l1を介してカウンタ4に接続さ
れる。カウンタ4から信号線l2を介して出力信号
が音声メモリ部2およびコントロール部3に接続
される。コントロール部3から制御信号が外部の
記憶装置に接続され、音声メモリ部2から有音フ
レームの前縁および後縁に指定数の無音フレーム
が付加された音声フレーズが外部の記憶装置に接
続される。
Here, the feature of the present invention is the storage portion of the digital audio signal, which is surrounded by a dashed line. That is, a write start signal S2 is input from the outside to the control unit 3, and control signals such as a write start address and an increment instruction are connected from the control unit 3 to the counter 4 via the signal line l1.
A write stop signal S3 is input to the control section 3, and a control signal instructing to stop incrementing is connected from the control section 3 to the counter 4 via the signal line l1. The output signal from counter 4 is connected to signal line l 2
It is connected to the audio memory unit 2 via the audio memory section 2, and digital audio signals are stored therein. Further, the output signal from the counter 4 is branched and connected to the control section 3. The digital audio signal stored from the audio memory section 2 is connected to the silence pattern detection section 5, and the silence pattern detection signal is connected to the control section 3. The control unit 3 counts the silent pattern detection signal, determines a silent frame, and determines the first sound frame and the last sound frame. External silent frame designation signal S 4
is input to the control unit 3, and a control signal is sent from the control unit 3 to the counter 4 via the signal line l1 to leave a designated number of silent frames at the leading edge of the first sound frame and the trailing edge of the last sound frame. Connected. An output signal from the counter 4 is connected to the audio memory section 2 and the control section 3 via a signal line l2 . A control signal is connected from the control unit 3 to an external storage device, and an audio phrase in which a specified number of silent frames are added to the leading and trailing edges of a sound frame is connected from the audio memory unit 2 to the external storage device. .

このような構成のデイジタル信号の記憶装置の
動作について説明する。第2図は音声メモリ部2
に格納されたデイジタル音声信号のメモリマツプ
図である。第1図において外部より入力される単
語や文節であるアナログ音声信号S1はアナログデ
イジタル変換部1によりデイジタル音声信号に変
換され、フレーム単位で音声メモリ部2に蓄積さ
れる。このときに、外部からの書込み開始信号S2
によりコントロール部3は音声メモリ部2の開始
アドレスを決定し、カウンタ4に書込み開始アド
レスを設定し、その後は外部からの書込み停止信
号S3があるまでカウンタ4へ一定周期でインクリ
メント指示を送出する。外部からの書込み停止信
号S3があればコントロール部3は上記インクリメ
ント指示の送出を停止し音声メモリ部2への書込
み動作を終了する。
The operation of the digital signal storage device having such a configuration will be explained. Figure 2 shows the voice memory section 2.
FIG. 3 is a memory map diagram of digital audio signals stored in the . In FIG. 1, an analog audio signal S1 , which is a word or phrase inputted from the outside, is converted into a digital audio signal by an analog-to-digital converter 1, and is stored in an audio memory 2 frame by frame. At this time, the external write start signal S2
The control unit 3 determines the start address of the audio memory unit 2, sets the write start address in the counter 4, and thereafter sends an increment instruction to the counter 4 at a constant cycle until a write stop signal S3 is received from the outside. . If there is a write stop signal S3 from outside, the control section 3 stops sending out the above-mentioned increment instruction and ends the write operation to the audio memory section 2.

このときの音声メモリ部2の蓄積状態を第2図
を参照して説明する。第2図aに示すSTARTは
書込み開始アドレスであり、STOPは書込み停止
アドレスである。また斜線部は有音のデイジタル
音声信号および空白部は無音のデイジタル音声信
号を示している。
The storage state of the audio memory section 2 at this time will be explained with reference to FIG. START shown in FIG. 2a is a write start address, and STOP is a write stop address. Further, the shaded area indicates a digital audio signal with sound, and the blank area indicates a digital audio signal with no sound.

無音フレームの判定方法について説明すると、
第1図に示す音声メモリ部2に蓄積されたデイジ
タル音声信号を書込み開始アドレスSTARTから
所定数のフレームを読み出し無音パターン検出部
5に供給する。コントロール部3では無音パター
ン検出部5から出力される無音パターン検出信号
をカウントし、あらかじめ定められた設定数と比
較する。カウント結果が多い場合には、この所定
数のフレームを無音フレームと判定する。同様に
こんどは書込み停止アドレスSTOPから逆上つて
所定数のバイト中の無音パターンの数を判定し無
音フレームの判定を行う。以上の動作を順次進め
て行き無音フレームでないと判定されるフレーム
を求めることにより有音部の最初と最後とのフレ
ーム位置を決定する。
To explain how to determine silent frames,
A predetermined number of frames are read out from the write start address START of the digital audio signal stored in the audio memory section 2 shown in FIG. 1 and supplied to the silent pattern detection section 5. The control unit 3 counts the silence pattern detection signal output from the silence pattern detection unit 5 and compares it with a predetermined set number. If the count result is large, this predetermined number of frames are determined to be silent frames. Similarly, next time, from the write stop address STOP, the number of silent patterns in a predetermined number of bytes is determined, and a silent frame is determined. By sequentially performing the above operations and finding frames that are determined to be non-silent frames, the first and last frame positions of the sound portion are determined.

このようにして決定された有音信号部に外部か
らの無音フレーム指定信号S4で指定されたフレー
ム数だけ最初の有音フレーム位置を前縁のフレー
ム位置へ、最後のフレーム位置を後縁のフレーム
位置へ移動することにより、前縁後縁に外部から
の指定フレーム数だけ無音フレームを付加し目的
の音声フレーズを生成する。
In the thus determined sound signal section, the first sound frame position is moved to the leading edge frame position, and the last frame position is moved to the trailing edge frame position by the number of frames specified by the external silent frame designation signal S4 . By moving to the frame position, silent frames are added to the leading edge and the trailing edge by the number of externally specified frames to generate the target audio phrase.

第2図aを用いてさらに詳細に説明する。まず
この例では無音パターンが80%以上のときは無音
フレームとし、かつ有音部の前縁後縁に各々1フ
レームずつの無音フレームを残すような音声フレ
ーズを作成するものとする。書込み開始点のフレ
ームAから無音フレームの判定を行う。フレーム
AおよびフレームBは完全に無音であり、無音パ
ターンが100%をしめるため無音フレームとする。
続くフレームCは1/3が無音パターンであり80%
末満であるため無音フレームでないと判定され
る。同様に書込み停止点フレームzからフレーム
Y、フレームXと逆上つて判定するとフレームX
の無音パターンが80%未満のため無音フレームで
ないと判定される。このことからフレームCおよ
びフレームXが両端の有音フレームであることが
わかる。次に外部から1フレームずつ無音フレー
ムを残すように指定されているために、フレーム
Cより1フレーム前のフレームBと、フレームX
より1フレーム後のフレームYとを含む第2図b
に示すようにフレームBからフレームYを有効な
音声フレーズと決定する。
This will be explained in more detail using FIG. 2a. First, in this example, a voice phrase is created in which a silent frame is created when the silent pattern is 80% or more, and one silent frame is left at each of the leading and trailing edges of the sound part. A silent frame is determined from frame A, which is the writing start point. Frame A and frame B are completely silent and have a silent pattern of 100%, so they are considered silent frames.
In the following frame C, 1/3 is a silent pattern and 80%
Since the frame is full, it is determined that it is not a silent frame. Similarly, if it is determined that the write stop point frame z goes up from frame Y to frame X, frame X
Since the silence pattern of is less than 80%, it is determined that the frame is not a silent frame. From this, it can be seen that frame C and frame X are the sound frames at both ends. Next, since it is specified from the outside to leave a silent frame one frame at a time, frame B, which is one frame before frame C, and frame X
Figure 2b, which includes frame Y one frame later than
As shown in the figure, frames B to Y are determined to be valid voice phrases.

以上のようにして作成された音声フレーズは、
第1図に示すコントロール部3により音声メモリ
部2から読出され、外部の記憶装置へ送出され
る。
The audio phrase created as above is
The control unit 3 shown in FIG. 1 reads out the audio data from the audio memory unit 2 and sends it to an external storage device.

〔発明の効果〕〔Effect of the invention〕

以上説明したように、本発明は、コントロール
部、カウンタ部および無音パターン検出部を設け
ることにより、音声応答装置に用いる単語や文節
などの音声フレーズを記憶装置に書込む際に、自
動的に有音信号の前縁および後縁を判定し、指定
された量の無音信号を有音信号の前縁後縁に付加
することができる優れた効果がある。したがつ
て、容易にしかも文章を構成する音声フレーズの
前縁後縁の無音部分の長さが同一条件のもとで決
定されるため、組合せて文章として再生したとき
に、極めて自然な音となるような音声フレーズを
作成できる利点がある。
As explained above, by providing a control section, a counter section, and a silence pattern detection section, the present invention automatically enables effective use when writing voice phrases such as words and phrases used in a voice response device into a storage device. It is advantageous to determine the leading edge and trailing edge of the sound signal, and to add a specified amount of silence signal to the leading edge and trailing edge of the sound signal. Therefore, since the lengths of the silent parts at the leading and trailing edges of the audio phrases that make up a sentence can be determined easily and under the same conditions, when they are combined and played back as a sentence, they can produce extremely natural sounds. It has the advantage of being able to create voice phrases that sound like

【図面の簡単な説明】[Brief explanation of drawings]

第1図は本発明一実施例デイジタル音声信号の
記憶装置のブロツク図。第2図は音声メモリに格
納された音声データのメモリマツプ図。 1……アナログデイジタル変換部、2……音声
メモリ、3……コントロール部、4……カウン
タ、5……無音パターン検出部、l1,l2……信号
線、S1……アナログ音声信号、S2……書込み開始
信号、S3……書込み停止信号、S4……無音フレー
ム指示信号。
FIG. 1 is a block diagram of a digital audio signal storage device according to an embodiment of the present invention. FIG. 2 is a memory map diagram of audio data stored in the audio memory. DESCRIPTION OF SYMBOLS 1...Analog-digital converter, 2...Audio memory, 3...Control unit, 4...Counter, 5...Silent pattern detection unit, l1 , l2 ...Signal line, S1 ...Analog audio signal , S2 ...Writing start signal, S3 ...Writing stop signal, S4 ...Silent frame instruction signal.

Claims (1)

【特許請求の範囲】 1 入力されるアナログ音声信号をデイジタル音
声信号に変換するアナログデイジタル変換部と、 このアナログデイジタル変換部からのデイジタ
ル音声信号をフレーム単位で記憶する音声メモリ
部と、 この音声メモリ部の書込みおよび読出しを制御
する制御手段と を備えた音声応答装置のデイジタル音声信号の記
憶装置において、 上記音声メモリ部の読出し出力に接続され、そ
の読出し出力が無音パターンであるときに検出出
力を送出する無音パターン検出部を備え、 上記制御手段には、 上記音声メモリ部に書込まれたデイジタル音声
信号を一時的に読出して上記音声パターン検出部
に与える手段と、 この無音パターン検出部の出力が1フレーム中
で一定割合を超えたフレームを無音フレームと判
定する手段と、 この手段の判定に基づき有音のデイジタル音声
信号フレームの前および後に無音フレームを予め
設定された数だけ存在するように無音フレームを
削除して音声フレーズを編集する手段と を備えたことを特徴とするデイジタル音声信号の
記憶装置。
[Claims] 1. An analog-digital converter that converts an input analog audio signal into a digital audio signal, an audio memory unit that stores the digital audio signal from the analog-digital converter in units of frames, and the audio memory. A digital voice signal storage device for a voice response device, which is connected to a readout output of the voice memory unit and outputs a detection output when the readout output is a silent pattern. The control means includes a means for temporarily reading out the digital audio signal written in the audio memory section and providing it to the audio pattern detection section, and an output of the silence pattern detection section. means for determining a frame in which a frame exceeds a certain percentage in one frame as a silent frame; A storage device for a digital audio signal, comprising means for editing audio phrases by deleting silent frames.
JP59062806A 1984-03-30 1984-03-30 Memory unit for digital voice signal Granted JPS60205598A (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP59062806A JPS60205598A (en) 1984-03-30 1984-03-30 Memory unit for digital voice signal

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP59062806A JPS60205598A (en) 1984-03-30 1984-03-30 Memory unit for digital voice signal

Publications (2)

Publication Number Publication Date
JPS60205598A JPS60205598A (en) 1985-10-17
JPH0481198B2 true JPH0481198B2 (en) 1992-12-22

Family

ID=13210944

Family Applications (1)

Application Number Title Priority Date Filing Date
JP59062806A Granted JPS60205598A (en) 1984-03-30 1984-03-30 Memory unit for digital voice signal

Country Status (1)

Country Link
JP (1) JPS60205598A (en)

Family Cites Families (4)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPS5579500A (en) * 1978-12-11 1980-06-14 Hitachi Ltd Speech answering system
JPS5733000U (en) * 1980-07-30 1982-02-20
JPS5774796A (en) * 1980-10-28 1982-05-11 Citizen Watch Co Ltd Voice recorder
JPS58205196A (en) * 1982-05-25 1983-11-30 東芝エンジニアリング株式会社 Automatic editting of voice information for voice processor

Also Published As

Publication number Publication date
JPS60205598A (en) 1985-10-17

Similar Documents

Publication Publication Date Title
TW347619B (en) A communication system and method using a speaker dependent time-scaling technique a method for time-scale modification of speech using a modified version of the Waveform Similarity based Overlap-Add technique (WSOLA).
US5810600A (en) Voice recording/reproducing apparatus
JPS58205196A (en) Automatic editting of voice information for voice processor
JPS60205598A (en) Memory unit for digital voice signal
JPH0258639B2 (en)
JPH0210959B2 (en)
JPS6199198A (en) Voice analyzer/synthesizer
JPS6014360B2 (en) voice response device
JPH0368399B2 (en)
JP2532052B2 (en) Audio storage device and audio storage / playback device
JPS62994A (en) Pcm voice signal memory
JPS59160193A (en) Voice data editing system
JPH0251200A (en) Audio signal continuous writing device
JPS633320B2 (en)
JPS62125577A (en) Voice storing and reproducing device
JPS59123889A (en) Voice editing/synthesization processing system
JPS58223197A (en) Voice editing system
JPS58198100A (en) Correction system for connection rest time
JP2511299Y2 (en) Voice output learning machine
JPH0287199A (en) System and device for sounding actuation for voice
JPS58123595A (en) Outputting of voice data
JPS5918720B2 (en) audio output device
JPH01163800A (en) Digital voice service apparatus
JPH0410820A (en) Voice file
JPS5868799A (en) Voice synthesizer