JPH11259090A - Sound pickup device - Google Patents

Sound pickup device

Info

Publication number
JPH11259090A
JPH11259090A JP10061518A JP6151898A JPH11259090A JP H11259090 A JPH11259090 A JP H11259090A JP 10061518 A JP10061518 A JP 10061518A JP 6151898 A JP6151898 A JP 6151898A JP H11259090 A JPH11259090 A JP H11259090A
Authority
JP
Japan
Prior art keywords
sound source
frequency component
microphone
target signal
noise
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Granted
Application number
JP10061518A
Other languages
Japanese (ja)
Other versions
JP3435687B2 (en
Inventor
Tomohiro Takano
智大 高野
Hiroyuki Matsui
弘行 松井
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
NTT Inc
Original Assignee
Nippon Telegraph and Telephone Corp
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Nippon Telegraph and Telephone Corp filed Critical Nippon Telegraph and Telephone Corp
Priority to JP06151898A priority Critical patent/JP3435687B2/en
Publication of JPH11259090A publication Critical patent/JPH11259090A/en
Application granted granted Critical
Publication of JP3435687B2 publication Critical patent/JP3435687B2/en
Anticipated expiration legal-status Critical
Expired - Fee Related legal-status Critical Current

Links

Abstract

PROBLEM TO BE SOLVED: To obtain a target sound wave signal with noises suppressed even if the target sound wave signal picked up with air conductive microphones is buried in the ambient noises. SOLUTION: Sound wave signals, which are picked up with microphones 1, 2 installed at a position near a sound source of a target sound wave signal and at another position farther than the aforementioned position from the sound source of the target sound wave signal and contain ambient noises, are converted into spectra, respectively (3, 4); a predominant frequency component of the target sound wave signal contained in the microphone 1 is extracted (6) from the level differences (5) of the amplitude spectra of each frequency component; as to unextracted frequency components, a noise level is estimated (8) from the spectral intensities of the unextracted frequency components and attenuated by a noise suppression quantity determined according to the estimated noise level quantity; and such a processed spectrum is converted into a time waveform (11), and consequently, the target sound wave signal is obtained with the noises suppressed.

Description

【発明の詳細な説明】DETAILED DESCRIPTION OF THE INVENTION

【0001】[0001]

【発明の属する技術分野】この発明は、周囲騒音が混在
した音源信号に対して、周囲騒音成分を抑圧し、目的信
号を抽出する機能を有する収音装置に関するものであ
る。
BACKGROUND OF THE INVENTION 1. Field of the Invention The present invention relates to a sound pickup apparatus having a function of suppressing an ambient noise component from a sound source signal in which ambient noise is mixed and extracting a target signal.

【0002】[0002]

【従来の技術】騒音下で、SN比より目的信号を抽出す
る従来技術として、目的信号の音場分布差を利用し、周
波数軸上で目的信号が支配的な周波数成分を抽出する手
法を、特願平10−39206号「収音装置」で提案し
た。前記提案した技術の収音系では、目的信号の音源近
くに主入力マイクロホンが、その主入力マイクロホンよ
りも前記音源から離れた位置に補助入力マイクロホンが
設置される。そして、これら2つのマイクロホン間に生
じるレベル差の特性が騒音と目的信号で異なる、一般に
前者の方が後者より小さいことに着目して、目的信号が
支配的な周波数成分の抽出を実現している。
2. Description of the Related Art As a conventional technique for extracting a target signal from an SN ratio under noise, there is a method of extracting a frequency component in which the target signal is dominant on a frequency axis using a sound field distribution difference of the target signal. It was proposed in Japanese Patent Application No. Hei 10-39206 "Sound collection device". In the sound collection system of the proposed technique, a main input microphone is installed near a sound source of a target signal, and an auxiliary input microphone is installed at a position farther from the sound source than the main input microphone. By paying attention to the fact that the characteristic of the level difference generated between these two microphones differs between the noise and the target signal, generally, the former is smaller than the latter, thereby realizing the extraction of the frequency component in which the target signal is dominant. .

【0003】図10には前記提案した技術の一例を示
す。主入力マイクロホン1(近接音源用)は目的信号の
音源に近い位置に配され、このマイクロホン1より目的
信号の音源から離されて補助入力マイクロホン2(遠隔
音源用)が配される。これらマイクロホン1,2の出力
信号はスペクトル変換部3,4においてそれぞれ周波数
成分L(ωh ),R(ωh ),(h=1,2,…,n)
に分解される。レベル差算出部5では分解された周波数
成分L(ωh ),R(ωh )のレベル差が、外部より設
定したしきい値よりも大きい場合、音源周波数成分選択
部6において、目的信号が支配的な周波数成分と見な
し、スペクトル変換部3の出力から、対応する周波数成
分L(ωm )(m=i,j,…)を抽出し、設定しきい
値よりも小さい場合には、騒音周波数成分抑圧部10に
おいて、スペクトル変換部3の出力中の、対応する周波
数成分L(ωm )(m=k,l,…)を減衰させる。音
源周波数成分抽出部7および騒音周波数成分抑圧部10
で、スペクトル変換部3の出力スペクトルに対し、この
ような処理を行ったものを、時間波形変換部11におい
て時間波形に変換して出力する。このような処理によれ
ば、時間とともに非定常に変化する騒音に対しても適用
できる騒音抑圧が実現される。
FIG. 10 shows an example of the proposed technique. The main input microphone 1 (for the near sound source) is arranged at a position close to the sound source of the target signal, and the auxiliary input microphone 2 (for the remote sound source) is arranged apart from the sound source of the target signal. The output signals of these microphones 1 and 2 are frequency components L (ω h ), R (ω h ), (h = 1, 2,...
Is decomposed into When the level difference between the frequency components L (ω h ) and R (ω h ) decomposed by the level difference calculation unit 5 is larger than a threshold value set from outside, the sound source frequency component selection unit 6 determines whether the target signal is Assuming that the frequency component is dominant, the corresponding frequency component L (ω m ) (m = i, j,...) Is extracted from the output of the spectrum converter 3. The frequency component suppression unit 10 attenuates the corresponding frequency component L (ω m ) (m = k, 1,...) In the output of the spectrum conversion unit 3. Sound source frequency component extraction unit 7 and noise frequency component suppression unit 10
Then, the output spectrum of the spectrum conversion unit 3 that has been subjected to such processing is converted into a time waveform by the time waveform conversion unit 11 and output. According to such processing, noise suppression that can be applied to noise that changes unsteadily with time is realized.

【0004】[0004]

【発明が解決しようとする課題】しかしながら、前記技
術では、目的信号が支配的でない周波数成分の減衰量が
適切でないと処理信号の品質が劣化するという問題があ
る。特に、周囲が静かな騒音下では、上記の周波数成分
の減衰量が大きすぎると、目的信号が支配的でない周波
数成分に重畳した目的信号の欠落が目立ち、音質が聴感
上著しく劣化しやすいという問題があった。
However, the above technique has a problem that the quality of a processed signal is deteriorated if the attenuation of a frequency component in which a target signal is not dominant is not appropriate. In particular, under the noise of quiet surroundings, if the attenuation of the above frequency component is too large, the loss of the target signal superimposed on the frequency component in which the target signal is not dominant is conspicuous, and the sound quality is apt to be significantly deteriorated in audibility. was there.

【0005】この発明の目的は、騒音の状況に応じて騒
音抑圧量を制御することにより、静かな騒音下で、目的
信号が支配的でない周波数成分に重畳した目的信号の欠
落が目立つことなく、音質劣化が少ない、つまり先に提
案した技術における処理信号の品質を改善する収音装置
を提供することにある。
An object of the present invention is to control the amount of noise suppression in accordance with the situation of noise so that the target signal superimposed on a frequency component in which the target signal is not dominant is not noticeably lost under quiet noise. It is an object of the present invention to provide a sound pickup device which has little sound quality deterioration, that is, improves the quality of a processed signal in the technique proposed above.

【0006】[0006]

【課題を解決するための手段】請求項1記載の発明で
は、目的信号の音源に近い位置に設置された第1マイク
ロホンと、前記位置より目的信号の音源から離れた位置
に設置された第2マイクロホンとの各々の出力信号を振
幅スペクトルと位相スペクトルに第1、第2スペクトル
変換手段で変換し、前記複数のスペクトル変換手段から
出力される各周波数成分ごとの振幅スペクトルについ
て、前記複数のスペクトル変換手段間のレベル差をレベ
ル差算出手段で計算し、前記レベル差算出手段により出
力される各周波数成分ごとのレベル差と、予め設定され
たしきい値とを比較し、目的信号が支配的な周波数成分
か否かを音源周波数成分選択手段で判定し、前記第1マ
イクロホンの出力信号の振幅スペクトルから、前記音源
周波数成分選択手段により目的信号が支配的と判定され
た周波数成分を音源周波数成分抽出手段で抽出し、前記
第1マイクロホンの出力信号の振幅スペクトルから、前
記音源周波数成分選択手段により目的信号が支配的と判
定されなかった周波数成分を抽出し、その振幅スペクト
ルから目的信号以外の騒音の振幅スペクトルあるいは出
力レベルを騒音レベル推定手段で推定し、前記騒音レベ
ル推定手段より出力される騒音の振幅スペクトルあるい
は出力レベルに応じて騒音の抑圧量を騒音抑圧量算出手
段で決定し、前記第1マイクロホンの出力信号の振幅ス
ペクトルにおいて、前記音源周波数成分選択手段におい
て目的信号が支配的と判定されなかった周波数成分に対
して前記騒音抑圧量算出手段で決定した減衰を、騒音周
波数成分抑圧手段で行い、前記音源周波数成分抽出手段
および前記騒音周波数成分抑圧手段より出力される振幅
スペクトルを前記第1スペクトル変換手段により算出さ
れる位相スペクトルを用いて時間波形に時間波形変換手
段で変換する。
According to the first aspect of the present invention, the first microphone installed at a position close to the sound source of the target signal and the second microphone installed at a position farther from the sound source of the target signal than the position. First and second spectrum conversion means convert each output signal from the microphone into an amplitude spectrum and a phase spectrum, and the plurality of spectrum conversions are performed on the amplitude spectrum for each frequency component output from the plurality of spectrum conversion means. The level difference between the means is calculated by the level difference calculating means, and the level difference for each frequency component output by the level difference calculating means is compared with a preset threshold value. The sound source frequency component selection means determines whether the frequency component is present or not, and from the amplitude spectrum of the output signal of the first microphone, the sound source frequency component selection means The frequency component for which the target signal is determined to be dominant is extracted by the sound source frequency component extracting means, and from the amplitude spectrum of the output signal of the first microphone, the target signal is not determined to be dominant by the sound source frequency component selecting means. The frequency components extracted are extracted from the amplitude spectrum, and the amplitude spectrum or output level of noise other than the target signal is estimated by the noise level estimating means from the amplitude spectrum, and according to the amplitude spectrum or output level of the noise output from the noise level estimating means. The noise suppression amount is determined by the noise suppression amount calculation means, and the noise component is determined for the frequency component in which the target signal is not determined to be dominant by the sound source frequency component selection means in the amplitude spectrum of the output signal of the first microphone. The attenuation determined by the suppression amount calculation means is performed by the noise frequency component suppression means, and the sound source frequency Convert time waveform converting means to a time waveform with a phase spectrum calculated amplitude spectrum output from component extraction means and the noise frequency component suppressing means by said first orthogonal transform means.

【0007】請求項2記載の発明は、請求項1記載の収
音装置において、前記音源周波数成分抽出手段で目的信
号が支配的と判定された周波数成分の振幅スペクトルの
大きさと、予め設定された無音区間判定しきい値とを比
較し、前記振幅スペクトルが前記無音区間判定しきい値
よりも小さい時目的信号の音源が無音状態であると音源
無音区間判定手段で判定し、前記音源無音区間判定手段
により目的信号の音源が無音状態と判定された場合にお
いて、前記レベル差算出手段より出力されるレベル差以
上となるように前記音源周波数成分選択手段に用いるし
きい値を、しきい値算出手段で算出し更新する。
According to a second aspect of the present invention, in the sound collecting apparatus according to the first aspect, the magnitude of the amplitude spectrum of the frequency component whose target signal is determined to be dominant by the sound source frequency component extracting means is set in advance. The sound source silent section determination means determines that the sound source of the target signal is in a silent state when the amplitude spectrum is smaller than the silent section determination threshold value. When the sound source of the target signal is determined to be in a silent state by the means, the threshold value used for the sound source frequency component selection means is set to be equal to or larger than the level difference output from the level difference calculation means. Is calculated and updated.

【0008】請求項3記載の発明は、請求項1又は2の
収音装置において、前記音源周波数成分抽出手段におい
て、目的信号が支配的と判定された周波数成分の振幅ス
ペクトルの大きさと、予め設定された無音区間判定しき
い値とを比較し、前記振幅スペクトルが前記無音区間判
定しきい値よりも小さいとき、目的信号の音源が無音状
態であると音源無音区間判定手段で判定し、前記音源無
音区間判定手段により目的信号の音源が無音状態と判定
された場合において「前記時間波形変換手段の出力」ま
たは「前記音源周波数成分抽出手段と前記騒音周波数成
分抑圧手段の出力」を音源無音区間減衰手段で減衰させ
る。
According to a third aspect of the present invention, in the sound pickup apparatus according to the first or second aspect, the magnitude of the amplitude spectrum of the frequency component whose target signal is determined to be dominant by the sound source frequency component extracting means is set in advance. The sound source silent section determination means determines that the sound source of the target signal is in a silent state when the amplitude spectrum is smaller than the silent section determination threshold value. When the sound source of the target signal is determined to be in a silent state by the silent section determining means, the "output of the time waveform converting means" or the "output of the sound source frequency component extracting means and the noise frequency component suppressing means" is attenuated by the sound source silent section. Attenuate by means.

【0009】請求項4記載の発明は、請求項1、2、ま
たは3記載の収音装置において、前記第1マイクロホン
が第2マイクロホンに比べて使用時に口元に近い位置に
なるようにハンドセット、ヘッドセット、イヤーマイク
セットに組み込まれたものである。作用 請求項1記載の発明の構成によれば第1、第2マイクロ
ホンの振幅スペクトルのレベル差によって目的信号が支
配的な周波数成分か否かの判定が行われ、更に目的信号
について、その音源と第1、第2マイクロホンの位置関
係はほとんど変化せずそれらの距離も短いため、2つの
マイクロホンの間で安定したレベル差が生じる。一方、
騒音については、その音源と第1、第2マイクロホンと
の間の距離は、目的信号の音源と第1、第2マイクロホ
ンとの間の距離に比べて長くなると考えてよい。このた
め、目的信号によって生じるレベル差は、騒音によって
生じるレベル差よりも常に大きくなると考えられる。
According to a fourth aspect of the present invention, in the sound collection device according to the first, second or third aspect, the handset and the head are arranged such that the first microphone is closer to the mouth when used than the second microphone. Set and ear microphone set. First According to the structure of the invention acts according to claim 1, wherein, a determination is made whether the target signal is dominant frequency component by the level difference of the amplitude spectrum of the second microphone, the further object signals, and the sound source Since the positional relationship between the first and second microphones hardly changes and their distance is short, a stable level difference occurs between the two microphones. on the other hand,
Regarding noise, it may be considered that the distance between the sound source and the first and second microphones is longer than the distance between the sound source of the target signal and the first and second microphones. For this reason, it is considered that the level difference caused by the target signal is always larger than the level difference caused by the noise.

【0010】この発明では、上記のように2つのマイク
ロホンに生じるレベル差が目的信号と騒音とで異なる点
に着目して目的信号が支配的な周波数成分の抽出処理を
行う。また、目的信号が支配的と判定されなかった周波
数成分、すなわち騒音の重畳が無視できない周波数成分
については、その振幅スペクトルより騒音の振幅スペク
トルあるいは出力レベルを推定し、推定された周囲騒音
の振幅スペクトルあるいは出力レベルに応じた減衰処理
を行う。このような処理によれば、非定常な騒音に対し
ても適用でき、かつ騒音の状況に応じて騒音抑圧量を制
御できる、騒音抑圧機能を有する収音装置が実現でき
る。
According to the present invention, the frequency component in which the target signal is dominant is extracted by focusing on the fact that the level difference between the two microphones differs between the target signal and the noise as described above. For frequency components for which the target signal is not determined to be dominant, that is, for frequency components for which noise superimposition is not negligible, the amplitude spectrum or output level of the noise is estimated from its amplitude spectrum, and the estimated amplitude spectrum of the ambient noise is obtained. Alternatively, an attenuation process according to the output level is performed. According to such processing, it is possible to realize a sound collection device having a noise suppression function, which can be applied to non-stationary noise and can control the amount of noise suppression according to the state of the noise.

【0011】請求項2記載の発明においては、請求項1
記載の発明において目的音源が無音状態と判定された場
合において目的信号が支配的な周波数成分か否かを判定
するためのしきい値を算出し、更新することによって、
音源周波数成分選択手段において目的信号が支配的な周
波数成分の判定精度を向上させ、音質を向上させる。請
求項3記載の発明においては、請求項1又は2記載の発
明において目的信号の音源が無音状態と判定された場合
において、「前記時間波形変換手段の出力」または、
「前記音声周波数成分抽出手段と前記騒音周波数成分抑
圧手段の出力」を減衰させることによって、目的音源が
無音状態であるときの騒音抑圧効果が向上する。
According to the second aspect of the present invention, the first aspect is provided.
By calculating and updating a threshold value for determining whether the target signal is a dominant frequency component in the case where the target sound source is determined to be in a silent state in the described invention,
In the sound source frequency component selection means, the accuracy of determining the frequency component in which the target signal is dominant is improved, and the sound quality is improved. According to the third aspect of the present invention, when the sound source of the target signal is determined to be in a silent state in the first or second aspect of the present invention, the "output of the time waveform conversion means" or
By attenuating "the output of the audio frequency component extracting means and the noise frequency component suppressing means", the noise suppressing effect when the target sound source is in a silent state is improved.

【0012】請求項4記載の発明においては、請求項
1、請求項2、または請求項3記載の発明において、第
1マイクロホンが第2マイクロホンに比べて口元に近い
位置になるようにハンドセット、ヘッドセット、イヤー
マイクセットを組み込むことによって、各々の送受話器
において送話信号の耐騒音性能を向上させることが可能
となる。
According to a fourth aspect of the present invention, in the first, second or third aspect of the present invention, the handset and the head are arranged such that the first microphone is closer to the mouth than the second microphone. By incorporating a set and an ear microphone set, it becomes possible to improve the noise resistance of the transmission signal in each handset.

【0013】[0013]

【発明の実施の形態】実施例1 図1は請求項1の発明の実施例を示すブロック図であ
る。図10に示したものに対し、騒音レベル推定部8と
騒音抑圧量算出部9とが付け加えられる。請求項1に示
した実施例の処理手順を図4の流れ図を参照して説明す
る。まず、マイクロホン1,2に騒音が重畳した目的信
号が各々取り込まれ、それらをディジタル信号として読
み込む(S02)。読み込まれたマイクロホン1,2の
信号を以下では、L,Rとする。
DESCRIPTION OF THE PREFERRED EMBODIMENTS Embodiment 1 FIG. 1 is a block diagram showing an embodiment of the present invention. A noise level estimating unit 8 and a noise suppression amount calculating unit 9 are added to those shown in FIG. The processing procedure of the first embodiment will be described with reference to the flowchart of FIG. First, target signals in which noise is superimposed on the microphones 1 and 2 are fetched, and read as digital signals (S02). Hereinafter, the read signals of the microphones 1 and 2 are referred to as L and R, respectively.

【0014】スペクトル変換部3,4では、取り込んだ
信号L,RをスペクトルL(ωh ),R(ωh )(h=
1,2,…,n)に変換する(S03)。この変換は、
例えば離散的フーリエ変換によって実行される。レベル
差算出部5では、L(ωh ),R(ωh )の各周波数成
分について、以下の式で与えられるレベル差ΔLR(ω
h )を計算する(S04)。
The spectrum converters 3 and 4 convert the captured signals L and R into spectra L (ω h ) and R (ω h ) (h =
1, 2,..., N) (S03). This conversion is
For example, this is performed by a discrete Fourier transform. The level difference calculator 5 calculates a level difference ΔLR (ω given by the following equation for each frequency component of L (ω h ) and R (ω h ).
h ) is calculated (S04).

【0015】ΔLR(ωh )=20log 10(|L
(ωh )|/|R(ωh )|) 上式中のωh は周波数(h=1,2,…,n),|L
(ωh )|,|R(ωh )|は、各々L,R信号の振幅
スペクトル成分を表わす。音源周波数成分選択部6で
は、各周波数成分についてΔLR(ωh )と予め設定さ
れたしきい値Th(ωh )の大小関係より、目的信号が
支配的な周波数成分の選択を行う。目的信号が支配的な
周波数成分か否かの判定条件は例えば以下の式によって
決定される(S05)。
ΔLR (ω h ) = 20 log 10 (| L
(Ω h) | / | R (ω h) |) ω h in the above formula is the frequency (h = 1,2, ..., n ), | L
(Ω h ) | and | R (ω h ) | represent the amplitude spectrum components of the L and R signals, respectively. The sound source frequency component selection unit 6 selects a frequency component in which the target signal is dominant based on the magnitude relationship between ΔLR (ω h ) and a preset threshold value Th (ω h ) for each frequency component. The condition for determining whether or not the target signal is a dominant frequency component is determined by, for example, the following equation (S05).

【0016】 ΔLR(ωh )>Th(ωh ) → 目的信号が支配的 ΔLR(ωh )≦Th(ωh ) → 目的信号が支配的でない 音源周波数成分抽出部7では、L(ωh )について目的
信号が支配的な周波数成分L(ωm )(m=i,j,
…)をスペクトル成分格納部(図示せず)にS(ωm )
として格納する(S06)。
[0016] ΔLR (ω h)> Th in (ω h) → target signal is dominant ΔLR (ω h) ≦ Th ( ω h) → sound source frequency component extraction unit 7 target signal is not dominant, L (ω h ), The frequency component L (ω m ) (m = i, j,
Spectral component storage unit a ...) (not shown) to S (omega m)
(S06).

【0017】 S(ωm )=L(ωm )(m=i,j,…) 音声周波数成分選択部6において、目的信号が支配的で
ないと判定された周波数ωm (m=k,l,…)につい
ては以下の騒音抑圧処理(S07)〜(S10)を行
う。まず、騒音レベル推定部8で、L(ωh )(h=
1,2,…,n)より目的信号が支配的でない周波数成
分L(ωm )(m=k,l,…)を抽出する(S0
7)。このL(ωm )(m=k,l,…)より騒音の全
帯域にわたる出力レベルLvを推定する(S08)。推
定の方法としては例えば以下の式が考えられる。
S (ω m ) = L (ω m ) (m = i, j,...) In the audio frequency component selection unit 6, the frequency ω m (m = k, l) at which the target signal is determined not to be dominant ,...) Perform the following noise suppression processing (S07) to (S10). First, the noise level estimating unit 8 calculates L (ω h ) (h =
The frequency component L (ω m ) (m = k, l,...) In which the target signal is not dominant is extracted from 1, 2,.
7). From this L (ω m ) (m = k, 1,...), An output level Lv over the entire noise band is estimated (S08). As an estimation method, for example, the following equation can be considered.

【0018】 Lv=20log 10((n/q)×Σ|L(ωm )|) ここで、qは目的信号が支配的でないと判定された周波
数成分の個数、和Σは目的信号が支配的でない周波数ω
m (m=k,1,…)に対応するものについてとる。騒
音抑圧量算出部9では、目的信号が支配的でない周波数
成分に乗ずる重み係数w(ωm )(m=k,l,…)を
算出する(S09)。w(ωm )の算出には例えば次式
を用いる。
Lv = 20 log 10 ((n / q) × Σ | L (ω m ) |) Here, q is the number of frequency components determined that the target signal is not dominant, and the sum Σ is the target signal. Untargeted frequency ω
m (m = k, 1,...). The noise suppression amount calculation unit 9 calculates a weight coefficient w (ω m ) (m = k, l,...) By which the frequency component in which the target signal is not dominant is multiplied (S09). For example, the following equation is used to calculate w (ω m ).

【0019】 w(ωm )=C (Lv<Lv1) C((Lvh−Lv)/(Lvh−Lv1))npw (Lv1≦Lv≦Lvh) 0 (Lv>Lvh) ここで、Cは0≦C≦1を満たす定数、Lvhは騒音抑
圧を充分に行う必要があるような大きい騒音レベルの目
安、Lv1は騒音抑圧をそれほど行う必要がない程度の
小さい騒音レベルの目安、C,npwはw(ωm )を変
化させる勾配を決める定数である。
W (ω m ) = C (Lv <Lv1) C ((Lvh−Lv) / (Lvh−Lv1)) npw (Lv1 ≦ Lv ≦ Lvh) 0 (Lv> Lvh) where C is 0 ≦ A constant that satisfies C ≦ 1, Lvh is a measure of a large noise level that requires sufficient noise suppression, Lv1 is a measure of a small noise level that does not need to perform noise suppression so much, and C and npw are w ( ω m ).

【0020】図6にC=1としたときの上式のw
(ωm )−Lv特性を示す。この図が示すように、騒音
が小さいときには重み係数w(ωm )は1に近づく。こ
の場合には、騒音抑圧量は小さくなるため処理後の信号
の劣化や残留雑音の問題が克服される。また、高騒音下
においては、重み係数w(ωm )は0に近づくため、騒
音抑圧量が大きくなり、処理後の信号の明瞭性を向上さ
せることができる。
FIG. 6 shows w in the above equation when C = 1.
(Ω m ) -Lv characteristics are shown. As shown in this figure, when the noise is low, the weight coefficient w (ω m ) approaches 1. In this case, since the amount of noise suppression is reduced, the problems of signal degradation after processing and residual noise are overcome. Further, under high noise, the weight coefficient w (ω m ) approaches 0, so that the noise suppression amount increases, and the clarity of the processed signal can be improved.

【0021】騒音周波数成分抑圧部10では、騒音抑圧
量算出部9で計算された重み係数を目的信号が支配的で
ない周波数成分L(ωm )(m=k,1,…)に乗じた
値を騒音抑圧処理後のスペクトル成分格納部(図示せ
ず)にS(ωm )(m=k,1,…)として格納する
(S10)。 S(ωm )=w(ωm )×L(ωm ) そして、(S10)の出力および(S06)の出力を合
成したS(ωh )(h=1,2,…)を時間波形変換部
11において信号Lの位相スペクトルΦ(ωh)を用い
て時間波形に変換し、時間波形信号を出力する(S1
1)。
The noise frequency component suppression unit 10 multiplies the weight coefficient calculated by the noise suppression amount calculation unit 9 by a frequency component L (ω m ) (m = k, 1,...) Where the target signal is not dominant. Is stored as S (ω m ) (m = k, 1,...) In a spectrum component storage unit (not shown) after the noise suppression processing (S10). S (ω m ) = w (ω m ) × L (ω m ) Then, S (ω h ) (h = 1, 2,...) Obtained by synthesizing the output of (S10) and the output of (S06) is a time waveform. The conversion unit 11 converts the signal L into a time waveform using the phase spectrum Φ (ω h ) and outputs a time waveform signal (S1).
1).

【0022】以上の処理はフレーム処理を基本とし、
(S02)で読み込んだ信号の時間長をシフトして重ね
合わせる方法で行う。例えば、時間長40msのときシ
フト幅を1/2にすればフレーム周期20msで上記
(S02)〜(S12)の処理がくり返されることにな
る。なお、(S09)で算出される重み係数は騒音の全
帯域における出力レベルに応じて算出されるが、これら
の値は、騒音を複数のサブ帯域に分けて求めることによ
り、各サブ帯域ごとの騒音の出力レベルに応じた値とし
て求めることができる。また、(S09)のような重み
係数による騒音抑圧ではなく、(S07)の出力が形成
する振幅スペクトル包絡より、騒音スペクトルを推定し
て、信号Lの振幅スペクトルから差引くスペクトルサブ
トラクション処理を適用することも可能である。実施例2 請求項1記載の発明では音源周波数成分選択部6におい
て、ある周波数成分が目的信号が支配的であるか否かを
判定するしきい値Th(ωh )を外部より設定してい
る。請求項2の発明では、目的信号の音源が無音状態で
あるときに周囲騒音に生じているマイクロホン1,2間
の各周波数成分におけるレベル差を利用して、しきい値
Th(ωh )を算出し、修正することにより音源周波数
成分選択部6において目的信号が支配的であるか否かの
判定精度を向上させ、音質を向上させるものである。
The above processing is based on frame processing.
This is performed by a method in which the time lengths of the signals read in (S02) are shifted and superimposed. For example, if the shift width is halved when the time length is 40 ms, the processes of (S02) to (S12) are repeated at a frame period of 20 ms. Note that the weighting factor calculated in (S09) is calculated according to the output level of the noise in all the bands, and these values are obtained by dividing the noise into a plurality of sub-bands, so that It can be obtained as a value corresponding to the noise output level. Also, instead of noise suppression using a weighting coefficient as in (S09), a noise spectrum is estimated from the amplitude spectrum envelope formed by the output of (S07), and a spectral subtraction process of subtracting from the amplitude spectrum of the signal L is applied. It is also possible. Second Embodiment In the invention according to the first embodiment, the sound source frequency component selection unit 6 externally sets a threshold value Th (ω h ) for determining whether or not a certain frequency component is dominant in a target signal. . According to the second aspect of the present invention, the threshold value Th (ω h ) is determined by utilizing the level difference between the frequency components of the microphones 1 and 2 generated in the ambient noise when the sound source of the target signal is in a silent state. By calculating and correcting, the sound source frequency component selection unit 6 improves the accuracy of determining whether or not the target signal is dominant, thereby improving the sound quality.

【0023】この請求項2の実施例を図2に示し、この
例では、図1に対し、音源無音区間判定部12、しきい
値算出部13が付け加えられ、その他は図1と同じ動作
である。以下では、この請求項2の実施例を示す図5を
用いて音源無音区間判定部12、およびしきい値算出部
13における処理について説明する。音源無音区間判定
部12では、まず第一に目的信号が支配的な振幅スペク
トルの和Pを求め(S07)、Pと外部より設定したし
きい値PThとの大小関係より目的信号の音源の無音状
態を検出する(S12)。
FIG. 2 shows a second embodiment of the present invention. In this embodiment, a sound source silent section determination unit 12 and a threshold value calculation unit 13 are added to FIG. is there. Hereinafter, the processing in the sound source silent section determination unit 12 and the threshold value calculation unit 13 will be described with reference to FIG. First, the sound source silence section determination unit 12 calculates the sum P of the amplitude spectrum in which the target signal is dominant (S07), and determines the silence of the sound source of the target signal based on the magnitude relationship between P and a threshold value PTh set from outside. The state is detected (S12).

【0024】 P>PTh → 目的信号の音源が有音状態 P≦PTh → 目的信号の音源が無音状態 音源無音区間判定部12において、目的信号の音源が無
音状態と判定された場合には、しきい値算出部13にお
いてしきい値Th(ωh )(h=1,2,…,n)を算
出する。例えば、新しいしきい値を以下の式により算出
する(S13,S14)。
P> PTh → the sound source of the target signal is in a sound state P ≦ PTh → the sound source of the target signal is in a silent state When the sound source silent section determination unit 12 determines that the sound source of the target signal is in a silent state, The threshold value calculation unit 13 calculates a threshold value Th (ω h ) (h = 1, 2,..., N). For example, a new threshold is calculated by the following equation (S13, S14).

【0025】Th(ωh )=ΔLR(ωh ) (ΔLR(ωh )>Th(ωh )のときのみ)実施例3 請求項3記載の発明は、請求項1または請求項2記載の
発明において音源無音区間判定部12により目的信号の
音源が無音状態と判定された場合に、「時間波形変換部
11の出力」または、「音源周波数成分抽出部7と騒音
周波数成分抑圧部10の出力」を減衰させ、騒音抑圧効
果を向上させるものである。
[0025] Th (ω h) = ΔLR ( ω h) the invention of Example 3 according to claim 3, wherein (ΔLR (ω h)> Th (ω h) when only) is as claimed in claim 1 or claim 2, wherein In the present invention, when the sound source of the target signal is determined to be in a silent state by the sound source silent section determination unit 12, the “output of the time waveform conversion unit 11” or the “output of the sound source frequency component extraction unit 7 and the noise frequency component suppression unit 10” Is attenuated to improve the noise suppression effect.

【0026】図2中に破線で示すように請求項3の実施
例は、請求項2の発明の実施例に対し、音源無音区間減
衰部14が付加される。この音源無音区間減衰部14の
動作を除けば請求項2の動作と同じであり、請求項3の
実施例を図5の破線枠で示される部分を用いて音源無音
区間減衰部14における処理について説明する。
As shown by a broken line in FIG. 2, the third embodiment has a sound source silent section attenuating section 14 added to the second embodiment. Except for the operation of the sound source silent section attenuator 14, the operation of the third embodiment is the same as that of the second embodiment. explain.

【0027】音源無音区間減衰部14では、音源無音区
間判定部12において目的信号の音源が無音状態と判定
された場合には(S16)、時間波形変換部11の出力
信号S(th )を減衰させる(S17)。なお、音源無
音区間減衰部14の処理は、音源周波数成分抽出部7と
騒音周波数成分抑圧部10の出力であるS(ωh )(h
=1,2,…,n)に対して行ってもよく、その効果は
(S17)の処理による効果と同等である。実施例4 図3に請求項4の実施例を示す。図3Aはハンドセット
21にマイクロホン1とマイクロホン2を取付けた場合
である。ハンドセット21の使用状態においてマイクロ
ホン1はその使用者の口22、つまり目的信号の音源近
くに位置され、マイクロホン2はハンドセット21の受
話器部分、つまり耳23の近くに位置するようにされて
いる。
[0027] In the sound source silent interval damping unit 14, when the sound source object signal is determined to silence the sound source silent section determining unit 12 (S16), the output signal S of the time waveform conversion section 11 (t h) It is attenuated (S17). The processing of the sound source silent section attenuating unit 14 is performed by the output of the sound source frequency component extracting unit 7 and the noise frequency component suppressing unit 10 as S (ω h ) (h
= 1, 2,..., N), and the effect is the same as the effect of the processing of (S17). Fourth Embodiment FIG. 3 shows a fourth embodiment. FIG. 3A shows a case where the microphone 1 and the microphone 2 are attached to the handset 21. When the handset 21 is in use, the microphone 1 is positioned near the user's mouth 22, that is, near the sound source of the target signal, and the microphone 2 is positioned near the handset 21 of the handset 21, that is, near the ear 23.

【0028】図3Bはヘッドセット25にマイクロホン
1,2を取付けた場合でヘッドセット25を使用者の頭
部26に装着した使用状態で、その耳23に対接される
受話器27の部分にマイクロホン2が取付けられ、この
受話器27の部分から、支持アーム28が延長され、支
持アーム28の遊端部が口22の近くに位置し、ここに
マイクロホン1が取付けられる。
FIG. 3B shows a case where the microphones 1 and 2 are attached to the headset 25 and the microphone 27 is attached to a part of the receiver 27 which is in contact with the ear 23 when the headset 25 is mounted on the head 26 of the user. 2, a support arm 28 is extended from the receiver 27, and the free end of the support arm 28 is located near the mouth 22, where the microphone 1 is mounted.

【0029】図3Cはイヤーマイクセット31に取付け
た場合で、イヤーマイクセット31が耳23の部分に取
付けられた状態で、アーム32が口22側に延長され、
これにマイクロホン1が取付けられ、このアーム32と
反対にアーム33が延長され、これにマイクロホン2が
取付けられる。この図に示したようにマイクロホン1,
2をハンドセット、ヘッドセット、イヤーマイクセット
の組み込み、実施例1から3の処理を実現する装置を構
成することによって、各々の送受話器において送話信号
の耐騒音性能を向上させることが可能となる。実験例 請求項1記載の発明を適用した実験例を以下に示す。目
的信号は音声、騒音信号は駅のホームでの周囲騒音を用
い、マイクロホン1とマイクロホン2の入力信号は、図
7に示すように目的信号の音源41よりの目的信号がマ
イクロホン1には直接入力され、マイクロホン2には抵
抗素子42により6dB(電力で半分に)減衰されて入
力され、騒音源43よりの騒音はマイクロホン1,2に
同一レベルで入力されるように計算機上で作成した。S
/N比は目的信号の平均電力と騒音信号の平均電力の比
で定義し、マイクロホン1におけるその値を5dB、−
5dBとしたものについて各々処理を行った。信号のス
ペクトル分解における周波数分解能は22Hz、分析フ
レームの時間長は46ms、フレーム周期は23msと
した。
FIG. 3C shows a case where the ear microphone set 31 is attached to the ear 23 and the arm 32 is extended toward the mouth 22 with the ear microphone set 31 attached to the ear 23.
The microphone 1 is attached thereto, the arm 33 is extended opposite to the arm 32, and the microphone 2 is attached thereto. As shown in FIG.
By incorporating a handset, a headset, and an ear microphone set into the apparatus 2 and configuring an apparatus that realizes the processing of the first to third embodiments, it is possible to improve the noise resistance of the transmission signal in each handset. . Experimental Example An experimental example to which the invention of claim 1 is applied is shown below. The target signal uses voice and the noise signal uses the ambient noise at the platform of the station. The input signals of the microphone 1 and the microphone 2 are input directly from the sound source 41 of the target signal to the microphone 1 as shown in FIG. The microphone 2 was attenuated by 6 dB (halved by electric power) by the resistance element 42 and input, and the noise from the noise source 43 was created on the computer so that it was input to the microphones 1 and 2 at the same level. S
The / N ratio is defined as the ratio of the average power of the target signal to the average power of the noise signal, and the value of the microphone 1 is 5 dB, −
The processing was performed for each of 5 dB. The frequency resolution in the spectral decomposition of the signal was 22 Hz, the time length of the analysis frame was 46 ms, and the frame period was 23 ms.

【0030】図8A,B、図9A,Bは、それぞれS/
Nが5dBの時と−5dBの時のマイクロホン1の処理
前の目的信号、騒音信号、図8C、図9CはS/N=5
dB、S/N=−5dBの時のそれぞれの騒音信号+目
的信号、図8D、図9DはそれぞれS/N=5dB、S
/N=−5dBの時の処理後の信号である。これらの図
から、SN比が5dB、−5dBの条件下のいずれにお
いても、処理後の信号(図8D、図9D)が処理前の目
的信号図8A、図9Aをよく復元していることが確認で
きる。
FIGS. 8A and 8B and FIGS. 9A and 9B show S /
The target signal and the noise signal before the processing of the microphone 1 when N is 5 dB and -5 dB, S / N = 5 in FIGS. 8C and 9C.
dB, S / N = −5 dB, each noise signal + target signal, FIGS. 8D and 9D show S / N = 5 dB, S
This is the processed signal when / N = -5 dB. From these figures, it can be seen that the signal after processing (FIGS. 8D and 9D) well restores the target signal before processing to FIG. 8A and FIG. 9A under the condition that the SN ratio is 5 dB and −5 dB. You can check.

【0031】また、ヘッドホン受聴により、SN比5d
Bの処理信号では歪みのない音声が得られ、SN比−5
dBの処理信号では充分な騒音抑圧効果が得られている
ことが確認できた。このことは、騒音抑圧量の制御が良
好に動作していることを示している。
Also, by listening to the headphones, the SN ratio is 5d.
A sound without distortion is obtained with the processed signal of B, and the SN ratio is -5.
It was confirmed that a sufficient noise suppression effect was obtained with the dB processed signal. This indicates that the control of the noise suppression amount is operating well.

【0032】[0032]

【発明の効果】以上、説明したように、請求項1記載の
発明では、目的信号の音源に近い位置に設置された第1
マイクロホンと、前記位置より目的信号の音源から離れ
た位置に設置された第2マイクロホンとの各出力信号を
第1、第2スペクトル変換手段で振幅スペクトルと位相
スペクトルに変換し、これらスペクトル変換手段の出力
中の各周波数成分ごとの振幅スペクトルについてレベル
差算出手段で前記第1、第2スペクトル変換手段間のレ
ベル差を計算し、前記レベル差算出手段により出力され
る各周波数成分ごとのレベル差と、予め設定されたしき
い値とを音源周波数成分選択手段で比較し、目的信号が
支配的な周波数成分か否かを判定し、前記第1マイクロ
ホンの出力信号の振幅スペクトルから、前記音源周波数
成分選択手段により目的信号が支配的と判定された周波
数成分を音源周波数成分抽出手段で抽出し、第1マイク
ロホンの出力信号の振幅スペクトルから、前記音源周波
数成分選択手段により目的信号が支配的と判定されなか
った周波数成分を抽出し、かつその振幅スペクトルから
目的信号以外の騒音の振幅スペクトルあるいは出力レベ
ルを騒音レベル推定手段で推定し、前記騒音レベル推定
手段より出力される騒音の振幅スペクトルあるいは出力
レベルに応じて騒音の抑圧量を騒音抑圧量算出手段で決
定し、前記第1マイクロホンの出力信号の振幅スペクト
ル中の、前記音源周波数成分選択手段において目的信号
が支配的と判定されなかった周波数成分に対して前記騒
音抑圧量算出手段で決定した減衰を騒音周波数成分抑圧
手段で行い、前記音源周波数成分抽出手段および前記騒
音周波数成分抑圧手段より出力される振幅スペクトルを
前記第1マイクロホンの前記第1スペクトル変換手段に
より算出される位相スペクトルを用いて時間波形に時間
波形変換手段で変換することにより、非定常な騒音に対
しても有効に動作し、かつ騒音の状況に応じて騒音抑圧
量を制御できる、新しい騒音抑圧処理機能を有する収音
装置を提供できる。
As described above, according to the first aspect of the present invention, the first signal generator installed at a position close to the sound source of the target signal is used.
The output signals of the microphone and the second microphone installed at a position distant from the sound source of the target signal from the position are converted into an amplitude spectrum and a phase spectrum by first and second spectrum conversion means. A level difference calculator calculates a level difference between the first and second spectrum converters for an amplitude spectrum of each frequency component being output, and calculates a level difference between each frequency component output by the level difference calculator. The sound source frequency component selection means compares the predetermined frequency with a preset threshold value to determine whether or not the target signal is a dominant frequency component. From the amplitude spectrum of the output signal of the first microphone, the sound source frequency component The frequency component for which the target signal is determined to be dominant by the selection means is extracted by the sound source frequency component extraction means, and the output signal of the first microphone is extracted. From the amplitude spectrum, frequency components for which the target signal is not determined to be dominant by the sound source frequency component selection means are extracted, and the amplitude spectrum or output level of noise other than the target signal is estimated from the amplitude spectrum by the noise level estimation means. The noise suppression amount is determined by the noise suppression amount calculation means in accordance with the amplitude spectrum or the output level of the noise output from the noise level estimation means, and the sound source in the amplitude spectrum of the output signal of the first microphone is determined. The frequency components for which the target signal is not determined to be dominant by the frequency component selection means are subjected to attenuation determined by the noise suppression amount calculation means by the noise frequency component suppression means, and the sound source frequency component extraction means and the noise frequency component The amplitude spectrum output from the suppression means is transmitted to the first microphone of the first microphone. By using the phase spectrum calculated by the vector conversion means to convert it to a time waveform by the time waveform conversion means, it operates effectively even for unsteady noise and controls the amount of noise suppression according to the noise situation It is possible to provide a sound collection device having a new noise suppression processing function.

【0033】請求項2記載の発明では、請求項1記載の
収音装置において、前記音源周波数成分抽出手段によ
り、目的信号が支配的と判定された周波数成分の振幅ス
ペクトルの大きさと、予め設定された無音区間判定しき
い値とを比較し、前記振幅スペクトルが前記無音区間判
定しきい値よりも小さいとき、目的信号の音源が無音状
態であると音源無音区間判定手段で判定し、この判定が
目的信号の音源が無音状態と判定されると、前記レベル
差算出手段より出力されるレベル差以上となるように前
記音源周波数成分選択手段に用いるしきい値を、しきい
値算出手段で算出更新することにより、音源周波数成分
選択手段において目的信号が支配的な周波数成分抽出精
度を向上させ、処理後の信号の品質向上が可能な収音装
置を提供できる。
According to a second aspect of the present invention, in the sound pickup apparatus of the first aspect, the magnitude of the amplitude spectrum of the frequency component determined as the dominant signal of the target signal by the sound source frequency component extracting means is set in advance. The amplitude spectrum is smaller than the silent section determination threshold, and the sound source silent section determining section determines that the sound source of the target signal is in a silent state when the amplitude spectrum is smaller than the silent section determination threshold. When the sound source of the target signal is determined to be in the silent state, the threshold value used by the sound source frequency component selection means is updated by the threshold value calculation means so that the threshold value becomes equal to or more than the level difference output from the level difference calculation means. By doing so, it is possible to provide a sound pickup device capable of improving the frequency component extraction accuracy in which the target signal is dominant in the sound source frequency component selection means and improving the quality of the processed signal.

【0034】請求項3記載の発明は、請求項1の収音装
置において、前記音源周波数成分抽出手段で目的信号が
支配的と判定された周波数成分の振幅スペクトルの大き
さと、予め設定された無音区間判定しきい値とを比較
し、前記振幅スペクトルが前記無音区間判定しきい値よ
りも小さいとき目的信号の音源が無音状態であると音源
無音区間判定手段で判定し、前記音源無音区間判定手段
により目的信号の音源が無音状態と判定された場合にお
いて「前記時間波形変換手段の出力」、または「前記音
源周波数成分抽出手段と前記騒音周波数成分抑圧手段の
出力」を音源無音区間減衰手段で減衰させることによ
り、目的信号の音源が無音状態のときは信号は減衰さ
れ、これにより騒音が抑圧され、さらに騒音の少ない収
音装置が提供される。
According to a third aspect of the present invention, in the sound collection device of the first aspect, the magnitude of the amplitude spectrum of the frequency component for which the target signal is determined to be dominant by the sound source frequency component extracting means, and the preset silence. Comparing with a section determination threshold, when the amplitude spectrum is smaller than the silent section determination threshold, the sound source silent section determining means determines that the sound source of the target signal is in a silent state, and the sound source silent section determining means When the sound source of the target signal is determined to be in a silent state, the "output of the time waveform converting means" or the "output of the sound source frequency component extracting means and the noise frequency component suppressing means" is attenuated by the sound source silent section attenuating means. By doing so, when the sound source of the target signal is in a silent state, the signal is attenuated, whereby the noise is suppressed, and a sound pickup device with less noise is provided.

【0035】請求項4記載の発明は、請求項1、2、ま
たは3記載の収音装置において、前記目的信号の音源に
近い位置に設置された第1マイクロホンと前記目的信号
の音源から離れた位置に設置された第2マイクロホンの
うち、前者のマイクロホンが後者のマイクロホンに比べ
て使用時に口元に近い位置になるようにハンドセット、
ヘッドセット、イヤーマイクセットに組み込まれている
ことにより従来のハンドセット、ヘッドセット、イヤー
マイクセットにおいて送話信号の耐騒音性能を向上させ
ることが可能となる。従来、耐騒音性に優れた送話信号
を得るイヤーマイクセットとして骨導マイクロホンとレ
シーバを一体化したものがある。しかし骨導マイクロホ
ンによって収音された音声は周波数成分が低周波成分に
偏っており、高周波成分が少ないため、音質が悪い。ま
た骨導マイクロホンとレシーバとの音響結合の問題もあ
る。この発明では、気導音をベースとした収音であり、
レシーバとマイクロホン間の距離も確保できるため、上
記の問題を持たないイヤーマイクセットの提供が可能と
なる。
According to a fourth aspect of the present invention, in the sound collection device according to the first, second, or third aspect, the first microphone installed at a position close to the sound source of the target signal is separated from the sound source of the target signal. A handset such that, of the second microphones installed at the positions, the former microphone is closer to the mouth when used than the latter microphone,
By being incorporated in the headset and the ear microphone set, it becomes possible to improve the noise resistance of the transmission signal in the conventional handset, headset and ear microphone set. 2. Description of the Related Art Conventionally, as an ear microphone set for obtaining a transmission signal excellent in noise resistance, there is an ear microphone set in which a bone conduction microphone and a receiver are integrated. However, the sound picked up by the bone conduction microphone has a low frequency component and a low frequency component, so that the sound quality is poor. There is also a problem of acoustic coupling between the bone conduction microphone and the receiver. In the present invention, the sound is collected based on the air conduction sound,
Since the distance between the receiver and the microphone can be ensured, an ear microphone set that does not have the above-described problem can be provided.

【0036】なお、以上の説明で使用したマイクロホン
は、無指向性マイクロホンに限定されるものではなく、
例えば、マイクロホン1は、目的信号の音源の方向に指
向性を有するマイクロホンを使用し、マイクロホン2
は、目的信号の音源と反対の方向に指向性を有するマイ
クロホンを使用してもよい。この場合、目的信号の音源
方向のみに鋭い指向性を有する超指向性マイクロホンと
して利用できる。
The microphone used in the above description is not limited to a non-directional microphone.
For example, the microphone 1 uses a microphone having directivity in the direction of the sound source of the target signal, and the microphone 2
May use a microphone having directivity in the direction opposite to the sound source of the target signal. In this case, it can be used as a super-directional microphone having sharp directivity only in the direction of the sound source of the target signal.

【0037】この発明は、騒音抑圧が必要な各種収音装
置のほか、通話を目的とした電話装置や、音声認識の入
力装置にも利用できる。
The present invention can be used not only for various sound collection devices that require noise suppression, but also for telephone devices for speech communication and voice recognition input devices.

【図面の簡単な説明】[Brief description of the drawings]

【図1】請求項1の発明の実施例の機能的構成を示すブ
ロック図。
FIG. 1 is a block diagram showing a functional configuration of an embodiment of the present invention.

【図2】請求項2及び3の各発明の実施例の機能的構成
を示すブロック図。
FIG. 2 is a block diagram showing a functional configuration of an embodiment of each of the second and third aspects of the present invention.

【図3】請求項4の発明の各種実施例を示す側面図。FIG. 3 is a side view showing various embodiments of the invention of claim 4;

【図4】図1に示した実施例の動作手順を示す流れ図。FIG. 4 is a flowchart showing an operation procedure of the embodiment shown in FIG. 1;

【図5】図2に示した実施例の動作手順を示す流れ図。FIG. 5 is a flowchart showing an operation procedure of the embodiment shown in FIG. 2;

【図6】図1に示した実施例におけるw(ωm )のLv
に対する特性例を示す図。
FIG. 6 shows Lv of w (ω m ) in the embodiment shown in FIG.
The figure which shows the characteristic example with respect to FIG.

【図7】この発明の実験例に用いたマイクロホン入力信
号の発生例を示す図。
FIG. 7 is a diagram showing a generation example of a microphone input signal used in an experimental example of the present invention.

【図8】この発明の実験例に適用した処理前の目的信
号、騒音信号、騒音+目的信号(SN比=5dB)、及
び処理後の信号をそれぞれ示す図。
FIG. 8 is a diagram showing a target signal before processing, a noise signal, a noise + target signal (SN ratio = 5 dB), and a signal after processing applied to the experimental example of the present invention.

【図9】この発明の実験例に適用した処理前の目的信
号、騒音信号、騒音+目的信号(SN比=−5dB)、
及び処理後の信号をそれぞれ示す図。
FIG. 9 shows a target signal, a noise signal, a noise + target signal (SN ratio = −5 dB) before processing applied to an experimental example of the present invention,
FIG. 3 is a diagram showing signals after processing.

【図10】先に提案した技術の機能構成例を説明するブ
ロック図。
FIG. 10 is a block diagram illustrating an example of a functional configuration of the technology proposed earlier.

Claims (4)

【特許請求の範囲】[Claims] 【請求項1】 目的信号の音源に近い位置に設置された
第1マイクロホンと、 前記位置より目的信号の音源から離れた位置に設置され
た第2マイクロホンと、 前記第1、第2マイクロホンの各々の出力信号を振幅ス
ペクトルと位相スペクトルに変換する第1、第2スペク
トル変換手段と、 前記第1、第2スペクトル変換手段から出力される各周
波数成分ごとの振幅スペクトルについて、相互のレベル
差を計算するレベル差算出手段と、 前記レベル差算出手段により出力される各周波数成分ご
とのレベル差と、予め設定されたしきい値とを比較し、
目的信号が支配的な周波数成分か否かを判定する音源周
波数成分選択手段と、 前記第1マイクロホンの出力信号の振幅スペクトルか
ら、前記音源周波数成分選択手段により目的信号が支配
的と判定された周波数成分を抽出する音源周波数成分抽
出手段と、 前記第1マイクロホンの出力信号の振幅スペクトルか
ら、前記音源周波数成分選択手段により目的信号が支配
的と判定されなかった周波数成分を抽出し、その振幅ス
ペクトルから目的信号以外の騒音の振幅スペクトルある
いは出力レベルを推定する騒音レベル推定手段と、 前記騒音レベル推定手段より出力される騒音の振幅スペ
クトルあるいは出力レベルに応じて騒音の抑圧量を決定
する騒音抑圧量算出手段と、 前記第1マイクロホンの出力信号の振幅スペクトルにお
いて、前記音源周波数成分選択手段において目的信号が
支配的と判定されなかった周波数成分に対して、前記騒
音抑圧量算出手段で決定した減衰を行う騒音周波数成分
抑圧手段と、 前記音源周波数成分抽出手段および前記騒音周波数成分
抑圧手段より出力される振幅スペクトルを前記第1スペ
クトル変換手段により算出される位相スペクトルを用い
て時間波形に変換する時間波形変換手段とを有すること
を特徴とする収音装置。
A first microphone installed at a position close to a sound source of a target signal; a second microphone installed at a position farther from the sound source of the target signal than the position; and each of the first and second microphones A first and a second spectrum converting means for converting the output signal of FIG. 1 into an amplitude spectrum and a phase spectrum; and calculating a level difference between the amplitude spectrums of the respective frequency components output from the first and the second spectrum converting means. Level difference calculation means, and the level difference for each frequency component output by the level difference calculation means, and compares a predetermined threshold value,
Sound source frequency component selecting means for determining whether or not the target signal is a dominant frequency component; and a frequency at which the target signal is determined to be dominant by the sound source frequency component selecting means from an amplitude spectrum of an output signal of the first microphone. Sound source frequency component extracting means for extracting a component; and extracting, from the amplitude spectrum of the output signal of the first microphone, a frequency component in which the target signal is not determined to be dominant by the sound source frequency component selecting means, and from the amplitude spectrum. Noise level estimation means for estimating the amplitude spectrum or output level of noise other than the target signal; and noise suppression amount calculation for determining the noise suppression amount according to the amplitude spectrum or output level of the noise output from the noise level estimation means. Means, in the amplitude spectrum of the output signal of the first microphone, A noise frequency component suppressing unit that performs attenuation determined by the noise suppression amount calculating unit on frequency components for which the target signal is not determined to be dominant by the wave number component selecting unit; the sound source frequency component extracting unit and the noise frequency And a time waveform converting means for converting the amplitude spectrum output from the component suppressing means into a time waveform using the phase spectrum calculated by the first spectrum converting means.
【請求項2】 請求項1記載の収音装置において、 前記音源周波数成分抽出手段で目的信号が支配的と判定
された周波数成分の振幅スペクトルの大きさと、予め設
定された無音区間判定しきい値とを比較し、前記振幅ス
ペクトルが前記無音区間判定しきい値よりも小さいと
き、目的信号の音源が無音状態であると判定する音源無
音区間判定手段と、 前記音源無音区間判定手段により目的信号の音源が無音
状態と判定された場合において、前記レベル差算出手段
より出力されるレベル差以上となるように前記音源周波
数成分選択手段に用いるしきい値を算出し更新するしき
い値算出手段を具備することを特徴とする収音装置。
2. The sound collection apparatus according to claim 1, wherein the magnitude of the amplitude spectrum of the frequency component for which the target signal is determined to be dominant by the sound source frequency component extracting means, and a preset silent section determination threshold value. When the amplitude spectrum is smaller than the silent section determination threshold, the sound source silent section determining means for determining that the sound source of the target signal is in a silent state; and A threshold calculator for calculating and updating a threshold used for the sound source frequency component selector so as to be equal to or more than the level difference output from the level difference calculator when the sound source is determined to be in a silent state; A sound pickup device.
【請求項3】 請求項1又は2の収音装置において、 前記音源周波数成分抽出手段において目的信号が支配的
と判定された周波数成分の振幅スペクトルの大きさと、
予め設定された無音区間判定しきい値とを比較し、前記
振幅スペクトルが前記無音区間判定しきい値よりも小さ
いとき目的信号の音源が無音状態であると判定する音源
無音区間判定手段と、 前記音源無音区間判定手段により目的信号の音源が無音
状態と判定された場合において前記時間波形変換手段の
出力、または前記音源周波数成分抽出手段と前記騒音周
波数成分抑圧手段の両出力を減衰させる音源無音区間減
衰手段を具備することを特徴とする収音装置。
3. The sound pickup device according to claim 1, wherein said sound source frequency component extracting means has a magnitude of an amplitude spectrum of a frequency component whose target signal is determined to be dominant;
Sound source silent section determining means for comparing with a preset silent section determination threshold value, and determining that the sound source of the target signal is in a silent state when the amplitude spectrum is smaller than the silent section determination threshold value; When the sound source of the target signal is determined to be in a silent state by the sound source silent section determining means, the output of the time waveform converting means, or the sound source silent section in which both outputs of the sound source frequency component extracting means and the noise frequency component suppressing means are attenuated. A sound pickup device comprising a damping means.
【請求項4】 請求項1、2、または3記載の収音装置
において、 前記第1マイクロホンと前記第2マイクロホンのうち、
前記第1マイクロホンが前記第2マイクロホンに比べて
使用時に口元に近い位置になるようにハンドセット、ヘ
ッドセット、イヤーマイクセットに組み込まれたことを
特徴とする収音装置。
4. The sound pickup device according to claim 1, wherein the first microphone and the second microphone are selected from the group consisting of the first microphone and the second microphone.
A sound pickup device, wherein the first microphone is incorporated in a handset, a headset, and an ear microphone set such that the first microphone is closer to a mouth when used than the second microphone.
JP06151898A 1998-03-12 1998-03-12 Sound pickup device Expired - Fee Related JP3435687B2 (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
JP06151898A JP3435687B2 (en) 1998-03-12 1998-03-12 Sound pickup device

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
JP06151898A JP3435687B2 (en) 1998-03-12 1998-03-12 Sound pickup device

Publications (2)

Publication Number Publication Date
JPH11259090A true JPH11259090A (en) 1999-09-24
JP3435687B2 JP3435687B2 (en) 2003-08-11

Family

ID=13173402

Family Applications (1)

Application Number Title Priority Date Filing Date
JP06151898A Expired - Fee Related JP3435687B2 (en) 1998-03-12 1998-03-12 Sound pickup device

Country Status (1)

Country Link
JP (1) JP3435687B2 (en)

Cited By (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2009134102A (en) * 2007-11-30 2009-06-18 Kobe Steel Ltd Object sound extraction apparatus, object sound extraction program and object sound extraction method
JP2010020165A (en) * 2008-07-11 2010-01-28 Fujitsu Ltd Noise suppressing device, mobile phone, noise suppressing method and computer program
US9368097B2 (en) 2011-11-02 2016-06-14 Mitsubishi Electric Corporation Noise suppression device
JP2017538151A (en) * 2014-11-12 2017-12-21 シラス ロジック、インコーポレイテッド Adaptive channel-to-channel discriminative rescaling filter
JP2019009800A (en) * 2018-08-30 2019-01-17 日本電信電話株式会社 headset

Family Cites Families (2)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2863214B2 (en) 1989-10-05 1999-03-03 株式会社リコー Noise removal device and speech recognition device using the device
JP3355598B2 (en) 1996-09-18 2002-12-09 日本電信電話株式会社 Sound source separation method, apparatus and recording medium

Cited By (7)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2009134102A (en) * 2007-11-30 2009-06-18 Kobe Steel Ltd Object sound extraction apparatus, object sound extraction program and object sound extraction method
JP2010020165A (en) * 2008-07-11 2010-01-28 Fujitsu Ltd Noise suppressing device, mobile phone, noise suppressing method and computer program
US9135924B2 (en) 2008-07-11 2015-09-15 Fujitsu Limited Noise suppressing device, noise suppressing method and mobile phone
US9368097B2 (en) 2011-11-02 2016-06-14 Mitsubishi Electric Corporation Noise suppression device
DE112011105791B4 (en) 2011-11-02 2019-12-12 Mitsubishi Electric Corporation Noise suppression device
JP2017538151A (en) * 2014-11-12 2017-12-21 シラス ロジック、インコーポレイテッド Adaptive channel-to-channel discriminative rescaling filter
JP2019009800A (en) * 2018-08-30 2019-01-17 日本電信電話株式会社 headset

Also Published As

Publication number Publication date
JP3435687B2 (en) 2003-08-11

Similar Documents

Publication Publication Date Title
JP3565226B2 (en) Noise reduction system, noise reduction device, and mobile radio station including the device
US6717991B1 (en) System and method for dual microphone signal noise reduction using spectral subtraction
US6549586B2 (en) System and method for dual microphone signal noise reduction using spectral subtraction
JP3373306B2 (en) Mobile radio device having speech processing device
JP4402295B2 (en) Signal noise reduction by spectral subtraction using linear convolution and causal filtering
US6487257B1 (en) Signal noise reduction by time-domain spectral subtraction using fixed filters
US8194872B2 (en) Multi-channel adaptive speech signal processing system with noise reduction
HUP0101288A2 (en) Method of noise suppression in telecommunication for tansmitting acoustic efficient signals, inparticular speech
JPH09503590A (en) Background noise reduction to improve conversation quality
JP6073456B2 (en) Speech enhancement device
JP4836720B2 (en) Noise suppressor
US20230352039A1 (en) Audio signal processing method, electronic device and storage medium
KR20020018625A (en) Process and Apparatus for Eliminating Loudspeaker Interference from Microphone Signals
US6507623B1 (en) Signal noise reduction by time-domain spectral subtraction
JPH11265199A (en) Transmitter
JP2020202448A (en) Acoustic device and acoustic processing method
US8275147B2 (en) Selective shaping of communication signals
JP3756828B2 (en) Reverberation elimination method, apparatus for implementing this method, program, and recording medium therefor
JP3435687B2 (en) Sound pickup device
US7130794B2 (en) Received speech signal processing apparatus and received speech signal reproducing apparatus
JPH09311696A (en) Automatic gain adjustment device
US20250316256A1 (en) Device for reducing noise during the reproduction of an audio signal using a headphone or hearing aid, and corresponding method
JP2020160290A (en) Signal processing equipment, signal processing system and signal processing method
JPH07146700A (en) Pitch emphasizing method and device and hearing compensator
JP2003044087A (en) Noise suppression device, noise suppression method, voice identification device, communication device, and hearing aid

Legal Events

Date Code Title Description
FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20090606

Year of fee payment: 6

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20090606

Year of fee payment: 6

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20100606

Year of fee payment: 7

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20100606

Year of fee payment: 7

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20110606

Year of fee payment: 8

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20120606

Year of fee payment: 9

FPAY Renewal fee payment (event date is renewal date of database)

Free format text: PAYMENT UNTIL: 20130606

Year of fee payment: 10

LAPS Cancellation because of no payment of annual fees