WO2016167215A1 - 線形予測符号化装置、線形予測復号装置、これらの方法、プログラム及び記録媒体 - Google Patents

線形予測符号化装置、線形予測復号装置、これらの方法、プログラム及び記録媒体 Download PDF

Info

Publication number
WO2016167215A1
WO2016167215A1 PCT/JP2016/061682 JP2016061682W WO2016167215A1 WO 2016167215 A1 WO2016167215 A1 WO 2016167215A1 JP 2016061682 W JP2016061682 W JP 2016061682W WO 2016167215 A1 WO2016167215 A1 WO 2016167215A1
Authority
WO
WIPO (PCT)
Prior art keywords
coefficient
linear
linear prediction
converted
unit
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/JP2016/061682
Other languages
English (en)
French (fr)
Inventor
守谷 健弘
優 鎌本
登 原田
弘和 亀岡
亮介 杉浦
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
University of Tokyo NUC
NTT Inc
Original Assignee
Nippon Telegraph and Telephone Corp
University of Tokyo NUC
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Nippon Telegraph and Telephone Corp, University of Tokyo NUC filed Critical Nippon Telegraph and Telephone Corp
Priority to US15/562,689 priority Critical patent/US10325609B2/en
Priority to KR1020177028710A priority patent/KR102061300B1/ko
Priority to EP16780006.9A priority patent/EP3270376B1/en
Priority to JP2017512523A priority patent/JP6517924B2/ja
Priority to CN201680021332.5A priority patent/CN107408390B/zh
Publication of WO2016167215A1 publication Critical patent/WO2016167215A1/ja
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Images

Classifications

    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/08Determination or coding of the excitation function; Determination or coding of the long-term prediction parameters
    • G10L19/09Long term prediction, i.e. removing periodical redundancies, e.g. by using adaptive codebook or pitch predictor
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/06Determination or coding of the spectral characteristics, e.g. of the short-term prediction coefficients
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/08Determination or coding of the excitation function; Determination or coding of the long-term prediction parameters
    • G10L19/12Determination or coding of the excitation function; Determination or coding of the long-term prediction parameters the excitation function being a code excitation, e.g. in code excited linear prediction [CELP] vocoders
    • G10L19/13Residual excited linear prediction [RELP]
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/26Pre-filtering or post-filtering
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/02Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
    • G10L19/032Quantisation or dequantisation of spectral components
    • G10L19/038Vector quantisation, e.g. TwinVQ audio
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/06Determination or coding of the spectral characteristics, e.g. of the short-term prediction coefficients
    • G10L19/07Line spectrum pair [LSP] vocoders
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L2019/0001Codebooks
    • G10L2019/0007Codebook element generation
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L2019/0001Codebooks
    • G10L2019/0016Codebook for LPC parameters

Definitions

  • This invention relates to a technique for encoding or decoding a coefficient that can be converted into a linear prediction coefficient.
  • Non-Patent Document 1 As a technique for quantizing an LSP parameter that is one of coefficients that can be converted into a linear prediction coefficient, a technique such as vector quantization is known (for example, see Non-Patent Document 1).
  • This parameter ⁇ is an encoding method for arithmetic coding in a coding scheme that arithmetically encodes a quantized value of a frequency domain coefficient using a linear prediction envelope as used in, for example, 3GPP EVS (Enhanced Voice Services) standard. It is a shape parameter that determines the probability distribution to which the object belongs.
  • the parameter ⁇ is related to the distribution of the encoding target, and if the parameter ⁇ is appropriately determined, efficient encoding and decoding can be performed.
  • the parameter ⁇ can be an index representing the characteristics of the time series signal. For this reason, if the parameter ⁇ is appropriately used, it is possible to efficiently encode and decode a coefficient that can be converted into a linear prediction coefficient such as an LSP parameter.
  • An object of the present invention is to provide a linear prediction encoding apparatus, a linear prediction decoding apparatus, a method, a program, and a recording medium for encoding or decoding a coefficient that can be converted into a linear prediction coefficient using a parameter ⁇ . To do.
  • the parameter ⁇ is a positive number
  • the parameter ⁇ corresponding to the time-series signal is set to the absolute value of the absolute value of the frequency domain sample sequence corresponding to the time-series signal to the ⁇ th power.
  • the shape parameter of the generalized Gaussian distribution that approximates the histogram of the whitened spectrum sequence which is a sequence obtained by dividing the frequency domain sample sequence by the spectrum envelope estimated by considering the power spectrum
  • ⁇ 1 is a predetermined parameter ⁇ Linear prediction analysis using a pseudo-correlation function signal sequence obtained by performing an inverse Fourier transform assuming that the absolute value ⁇ 1 of the frequency domain sample sequence corresponding to the time series signal is a power spectrum.
  • N is an integer of 1 or more
  • N codebooks are stored, and each codebook is stored in a codebook storage unit in which a plurality of coefficient candidates that can be converted into linear prediction coefficients corresponding to the respective parameters ⁇ are stored, and in the codebook storage unit A plurality of candidates for coefficients that can be converted into linear prediction coefficients stored in the codebook, and a coefficient that can be converted into linear prediction coefficients obtained by the linear prediction analysis unit, and a matching unit that adapts the value of ⁇ , A coefficient that can be converted into a linear prediction coefficient obtained by the linear prediction analysis unit is obtained by using a plurality of candidates of coefficients that can be converted into a linear prediction coefficient to which the value of ⁇ is adapted and a coefficient that can be converted into a linear prediction coefficient.
  • An encoding unit that obtains a corresponding linear prediction coefficient code.
  • the parameter ⁇ is a positive number
  • the parameter ⁇ corresponding to the time-series signal is set to the absolute value of the absolute value of the frequency domain sample sequence corresponding to the time-series signal to the ⁇ th power.
  • the shape parameter of the generalized Gaussian distribution that approximates the histogram of the whitened spectrum sequence which is a sequence obtained by dividing the frequency domain sample sequence by the spectrum envelope estimated by considering the power spectrum
  • ⁇ 1 is a predetermined parameter ⁇ Linear prediction analysis using a pseudo-correlation function signal sequence obtained by performing an inverse Fourier transform assuming that the absolute value ⁇ 1 of the frequency domain sample sequence corresponding to the time series signal is a power spectrum.
  • a linear prediction analyzer for obtaining a performs linear prediction coefficients can be converted into coefficients, and codebook codebook storage unit stored, based on eta 1 input
  • a matching unit that adapts at least one of a codebook stored in the codebook storage unit and a coefficient that can be converted into a linear prediction coefficient, and a coefficient that can be converted into a linear prediction coefficient using the codebook or the adapted codebook
  • an encoding unit that encodes a coefficient that can be converted into an adapted linear prediction coefficient.
  • a codebook storage unit codebook is stored, the eta 1 as a positive number, on the basis of eta 1 entered, stored in the codebook storage unit And at least a coefficient candidate that can be converted into a linear prediction coefficient corresponding to the input linear prediction coefficient code, among the coefficient candidates that can be converted into a plurality of linear prediction coefficients stored in the codebook.
  • Non-smoothing is provided that has a matching unit that adapts one of the coefficients, and the coefficients that can be converted into linear prediction coefficients are 1 / ⁇ 1 powers of the series of amplitude spectrum envelopes corresponding to the coefficients that can be converted into linear prediction coefficients Used to obtain a spectral envelope sequence.
  • Coefficients that can be converted into linear prediction coefficients can be encoded or decoded using the parameter ⁇ .
  • the block diagram for demonstrating the example of a linear prediction encoding apparatus The block diagram for demonstrating the example of a linear prediction encoding apparatus.
  • the block diagram for demonstrating the example of a linear prediction encoding apparatus The flowchart for demonstrating the example of the linear prediction encoding method.
  • the block diagram for demonstrating the example of a linear prediction decoding apparatus The flowchart for demonstrating the example of the linear prediction decoding method.
  • the block diagram for demonstrating the example of an encoding apparatus The flowchart for demonstrating the example of the encoding method.
  • the block diagram for demonstrating the example of an encoding part The block diagram for demonstrating the example of an encoding part.
  • the flowchart for demonstrating the example of a process of an encoding part The block diagram for demonstrating the example of a decoding apparatus.
  • the flowchart for demonstrating the example of a decoding method The flowchart for demonstrating the example of a process of a decoding part.
  • the block diagram for demonstrating the example of an encoding apparatus The flowchart for demonstrating the example of an encoding apparatus.
  • the flowchart for demonstrating the example of the encoding method The block diagram for demonstrating the example of a parameter determination apparatus.
  • the flowchart for demonstrating the example of the parameter determination method The figure for demonstrating generalized Gaussian distribution.
  • the block diagram for demonstrating the example of a linear prediction encoding apparatus The flowchart for demonstrating the example of the linear prediction encoding method.
  • the block diagram for demonstrating the example of a linear prediction decoding apparatus The flowchart for demonstrating the example of the linear prediction decoding method.
  • the block diagram for demonstrating the example of a linear prediction encoding apparatus The block diagram for demonstrating the example of a linear prediction encoding apparatus.
  • the block diagram for demonstrating the example of a linear prediction encoding apparatus The block diagram for demonstrating the example of a linear prediction decoding apparatus.
  • Linear predictive coding apparatus linear predictive decoding apparatus, and methods thereof
  • a linear prediction encoding apparatus linear prediction decoding apparatus, an encoding apparatus using these methods, a decoding apparatus, and examples of these methods will be described.
  • the linear prediction encoding apparatus of the first embodiment includes, for example, a linear prediction analysis unit 221, a codebook storage unit 222, an encoding unit 224, and a linear conversion unit 225.
  • the frequency domain transform unit 220 is provided outside the linear predictive coding device, but the linear predictive coding device may further include the frequency domain transform unit 220.
  • Each part of the linear predictive coding apparatus performs each process illustrated in FIG. 4 to realize the linear predictive coding method.
  • the time domain sound signal which is a time-series signal, is input to the frequency domain transform unit 220.
  • the frequency domain conversion unit 41 converts the input time domain sound signal into N frequency MDCT coefficient sequences X (0), X (1),..., X (N ⁇ Convert to 1). N is a positive integer.
  • the obtained MDCT coefficient sequence X (0), X (1),..., X (N-1) is output to the linear prediction analysis unit 221.
  • the subsequent processing is performed in units of frames.
  • the frequency domain transforming unit 220 obtains a frequency domain sample sequence corresponding to the time series signal, for example, an MDCT coefficient sequence.
  • the linear prediction analysis unit 221 receives, for example, a frequency domain sample sequence which is an MDCT coefficient sequence X (0), X (1),..., X (N-1) and a parameter ⁇ 1 corresponding to the frequency domain sample sequence. Is done.
  • the parameter ⁇ 1 is a positive number.
  • the parameter ⁇ 1 is determined by, for example, parameter determination units 27 and 27 ′ described later.
  • the parameter ⁇ 1 is used to encode an arithmetic code in an encoding method that arithmetically encodes a quantized value of a frequency domain coefficient using a linear prediction envelope such as that used in the 3GPP EVS (Enhanced Voice Services) standard, for example.
  • This is a parameter ⁇ that determines the probability distribution to which the object belongs.
  • the parameter ⁇ can be an index representing the characteristics of the time series signal.
  • the parameters ⁇ 2 and ⁇ 3 appearing later are also the parameter ⁇ . It can be said that ⁇ 1 , ⁇ 2 , and ⁇ 3 are predetermined values of the parameter ⁇ .
  • information about the parameter ⁇ 1 is transmitted to the linear prediction decoding apparatus.
  • a parameter code representing the parameter ⁇ 1 is transmitted to the linear predictive decoding device.
  • the linear prediction analysis unit 221 uses the MDCT coefficient sequence X (0), X (1),..., X (N-1) and ⁇ 1 and is defined by the following equation (A7) to R (0). , ⁇ R (1), ..., ⁇ R (N-1) are used to perform linear prediction analysis to generate coefficients that can be converted into linear prediction coefficient coefficients (step DE1).
  • the coefficient that can be converted into the generated linear prediction coefficient coefficient is output to the encoding unit 224.
  • the linear prediction analysis unit 22 firstly performs an inverse Fourier which considers the absolute value ⁇ 1 of the MDCT coefficient sequence X (0), X (1),..., X (N ⁇ 1) as the power spectrum. Time corresponding to the absolute value of the MDCT coefficient sequence X (0), X (1), ..., X (N-1) to the ⁇ 1 power by performing the operation corresponding to the conversion, that is, the operation of the equation (A7)
  • the pseudo-correlation function signal sequence ⁇ R (0), ⁇ R (1), ..., ⁇ R (N-1) which is the signal sequence of the region, is obtained.
  • the linear prediction analysis unit 22 performs linear prediction analysis using the obtained pseudo correlation function signal sequence ⁇ R (0), ⁇ R (1), ..., ⁇ R (N-1) to obtain a linear prediction coefficient. Generate coefficients that can be converted to coefficients.
  • the linear prediction analysis unit 221 performs inverse Fourier transform, assuming that ⁇ 1 is a positive number, and that the absolute value ⁇ 1 of the frequency domain sample sequence corresponding to the time-series signal is regarded as the power spectrum. Using the pseudo-correlation function signal sequence obtained by the above, linear prediction analysis is performed to obtain coefficients that can be converted into linear prediction coefficients.
  • the coefficients that can be converted into linear prediction coefficients are, for example, LSP, PARCOR coefficient, ISP, and the like.
  • the coefficient that can be converted into the linear prediction coefficient may be the linear prediction coefficient itself.
  • p is a predetermined positive number
  • the possible degree of the linear prediction coefficient is p-th order.
  • the codebook storage unit 222 stores a codebook in which a plurality of coefficient candidates that can be converted into linear prediction coefficients corresponding to the parameter ⁇ 2 are stored.
  • a pair of a coefficient candidate that can be converted into a linear prediction coefficient and a code corresponding to the coefficient candidate that can be converted into the linear prediction coefficient will be referred to as a candidate code pair.
  • a plurality of candidate code pairs are stored in the codebook. In other words, when N is a predetermined number of 2 or more, N candidate pairs are stored in the codebook.
  • a predetermined number of bits are assigned to each code corresponding to a coefficient candidate that can be converted into a linear prediction coefficient. Each code is represented by a predetermined number of assigned bits.
  • each coefficient candidate that can be converted into the linear prediction coefficient is composed of p values.
  • a candidate coefficient that can be converted to a linear prediction coefficient corresponding to parameter ⁇ 2 is optimal for encoding a coefficient that can be converted to a linear prediction coefficient corresponding to a frequency domain sample sequence having a parameter ⁇ value of ⁇ 2 This is a coefficient candidate that can be converted into a linear prediction coefficient.
  • the linear conversion unit 225 receives a coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 and a parameter ⁇ 1 corresponding to the coefficient that can be converted into the linear prediction coefficient.
  • the parameter ⁇ 1 is determined by, for example, parameter determination units 27 and 27 ′ described later.
  • the linear conversion unit 225 includes at least one of a first linear conversion unit 2251 and a second linear conversion unit 2252.
  • the linear conversion unit 225 includes the first linear conversion unit 2251 as shown in FIG. 1
  • the linear conversion unit 225 is the second case as shown in FIG.
  • the case where the linear conversion unit 2252 is provided is the second case
  • Each case will be described as the case 3.
  • the first linear conversion unit 2251 of the linear conversion unit 225 has at least input parameters for coefficient candidates that can be converted into linear prediction coefficients stored in the codebook storage unit 222.
  • a first linear transformation corresponding to ⁇ 1 is performed (step DE2).
  • the first linear conversion unit 2251 performs the first linear conversion according to the input parameter ⁇ 1 and the parameter ⁇ 2 corresponding to the coefficient candidate that can be converted into the linear prediction coefficient stored in the codebook storage unit 222.
  • the coefficient candidate that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 2 read from the codebook storage unit 222 is converted into the coefficient candidate that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 1 .
  • Coefficient candidates that can be converted to linear prediction coefficients corresponding to parameter ⁇ 1 are optimal for encoding coefficients that can be converted to linear prediction coefficients corresponding to the frequency domain sample sequence whose parameter ⁇ is ⁇ 1 This is a coefficient candidate that can be converted into a linear prediction coefficient.
  • the candidate coefficients that can be converted into the linear prediction coefficients after the first linear conversion are output to the encoding unit 224.
  • the first linear conversion unit 2251 may not perform the first linear conversion.
  • the first linear transformation unit 2251 of the linear transformation section 225 in accordance with the input parameter eta 1, as parameter eta 1 is small with the input can be converted into linear prediction coefficients after the first linear transformation First linear conversion is performed on the coefficient candidates that can be converted into the linear prediction coefficients read from the codebook storage unit 222 so that the amplitude spectrum envelope series corresponding to the coefficient candidates becomes flat, and the converted linear Coefficient candidates that can be converted into prediction coefficients are output.
  • the coefficient that can be converted into the linear prediction coefficient is LSP
  • Fig. 5 shows an example of LSP parameter values when parameter ⁇ takes various values.
  • the horizontal axis in FIG. 5 is the parameter ⁇ , and the vertical axis is the LSP parameter.
  • FIG. 5 shows that the smaller the parameter ⁇ , the closer the LSP parameter approaches a value obtained by equally dividing 0 to ⁇ .
  • the second linear conversion unit 2252 of the linear conversion unit 225 uses at least the input parameter ⁇ 1 for the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221.
  • the second linear transformation is performed according to (step DE2).
  • the second linear conversion unit 2252 can convert a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 obtained by the linear prediction analysis unit 221 into a linear prediction coefficient stored in the codebook storage unit 222. In order to correspond to a candidate of a new coefficient, the second linear conversion is performed to a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 2 .
  • the coefficient that can be converted into the linear prediction coefficient after the second linear conversion is output to the encoding unit 224.
  • the second linear conversion unit 2252 may not perform the second linear conversion.
  • the second linear transformation unit 2252 of the linear transformation section 225 in accordance with the input parameter eta 1, the smaller the inputted parameter eta 1, can be converted into linear prediction coefficients after the second linear transformation Second linear transformation is performed on the coefficients that can be converted to linear prediction coefficients so that the series of amplitude spectrum envelopes corresponding to the coefficients is flat, and the coefficients that can be converted to linear prediction coefficients after conversion are output. To do.
  • the first linear conversion unit 2251 of the linear conversion unit 225 sets at least the parameter ⁇ 3 to the coefficient candidates that can be converted into the linear prediction coefficients stored in the codebook storage unit 222.
  • a first linear transformation is performed.
  • the parameter ⁇ 3 is a positive value, and a value different from the parameter ⁇ 2 is determined in advance or input from the outside of the linear prediction coefficient encoding device.
  • the first linear transformation unit 2251 a first linear transformation in accordance with the parameter eta 2 corresponding to the candidate parameter eta 3 and codebook can be converted to a linear prediction coefficient stored in the storage unit 222 coefficients, code A coefficient candidate that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 2 read from the book storage unit 222 is converted into a coefficient candidate that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 3 .
  • Coefficient candidates that can be converted to linear prediction coefficients corresponding to parameter ⁇ 3 are optimal for encoding coefficients that can be converted to linear prediction coefficients corresponding to the frequency domain sample sequence whose parameter ⁇ is ⁇ 3 This is a coefficient candidate that can be converted into a linear prediction coefficient.
  • the candidate coefficients that can be converted into the linear prediction coefficients after the first linear conversion are output to the encoding unit 224.
  • the first linear conversion unit 2251 may not perform the first linear conversion.
  • the first linear conversion unit 2251 of the linear conversion unit 225 has a flatter amplitude spectrum envelope corresponding to a coefficient candidate that can be converted into a linear prediction coefficient after the first linear conversion, as the parameter ⁇ 3 is smaller.
  • the first linear conversion is performed on the coefficient candidates that can be converted into the linear prediction coefficients read from the codebook storage unit 222, and the coefficient candidates that can be converted into the converted linear prediction coefficients are output.
  • the second linear conversion unit 2252 of the linear conversion unit 225 has a second value corresponding to at least the parameter ⁇ 1 with respect to the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221. Perform linear transformation.
  • the second linear conversion unit 2252 converts the coefficient that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 1 obtained by the linear prediction analysis unit 221 into the coefficient that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 3. Second linear transformation.
  • the candidate coefficients that can be converted into the linear prediction coefficients after the second linear conversion are output to the encoding unit 224.
  • second linear conversion unit 2252 may not perform the second linear conversion.
  • the second linear transformation unit 2252 of the linear transformation section 225 in accordance with the input parameter eta 1, the smaller the inputted parameter eta 1, can be converted into linear prediction coefficients after the second linear transformation
  • the second linear transformation is performed on the input coefficient that can be converted into the linear prediction coefficient so that the amplitude spectrum envelope corresponding to the coefficient becomes flat, and the coefficient that can be converted into the converted linear prediction coefficient is output.
  • the linear conversion unit 225 performs the first linear operation according to ⁇ 3 on the coefficient candidates that can be converted into the linear prediction coefficients stored in the codebook storage unit 222. At least one of the conversion and the second linear conversion corresponding to ⁇ 3 is performed on the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 (step DE2).
  • ⁇ Encoding unit 224 The processing of the encoding unit 224 differs depending on the configuration of the linear conversion unit 225. Therefore, the processing of the encoding unit 224 when the linear conversion unit 225 is (1) the first case, (2) the second case, and (3) the third case will be described below.
  • the encoding unit 224 includes a coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 and a linear conversion unit. Coefficient candidates that can be converted into linear prediction coefficients after the first linear conversion obtained by the first linear conversion unit 2251 of 225 are input.
  • the encoding unit 224 encodes a coefficient that can be converted into a linear prediction coefficient using a coefficient candidate that can be converted into a linear prediction coefficient after the first linear conversion to obtain a linear prediction coefficient code (step DE3).
  • the encoding unit 224 selects a candidate closest to a coefficient that can be converted into a linear prediction coefficient from among a plurality of coefficient candidates that can be converted into a linear prediction coefficient after the first linear conversion.
  • the code corresponding to the selected candidate is a linear prediction coefficient code.
  • the obtained linear prediction coefficient code is output to the decoding device.
  • the encoding unit 224 can convert the linear prediction coefficient obtained by the second linear conversion unit 2252 of the linear prediction analysis unit 221. And a coefficient candidate that can be converted into a linear prediction coefficient stored in the codebook storage unit 222 is input.
  • the encoding unit 224 encodes a coefficient that can be converted into a linear prediction coefficient after the second linear conversion using a coefficient candidate that can be converted into a linear prediction coefficient to obtain a linear prediction coefficient code (step DE3).
  • the encoding unit 224 selects a candidate closest to the coefficient that can be converted into the linear prediction coefficient after the second linear conversion from among a plurality of coefficient candidates that can be converted into the linear prediction coefficient.
  • the code corresponding to the selected candidate is a linear prediction coefficient code.
  • the obtained linear prediction coefficient code is output to the decoding device.
  • the encoding unit 224 can convert the linear prediction coefficient obtained by the second linear conversion unit 2252 of the linear prediction analysis unit 221. And a coefficient candidate that can be converted into a linear prediction coefficient obtained by the first linear conversion unit 2251 of the linear prediction analysis unit 221 is input.
  • the encoding unit 224 encodes a coefficient that can be converted into the linear prediction coefficient after the second linear conversion using a coefficient candidate that can be converted into the linear prediction coefficient after the first linear conversion to obtain a linear prediction coefficient code. (Step DE3).
  • the encoding unit 224 converts the coefficient that can be converted into the linear prediction coefficient after the second linear conversion, from among a plurality of coefficient candidates that can be converted into the linear prediction coefficient after the first linear conversion. The closest one is selected, and the code corresponding to the selected candidate is set as the linear prediction coefficient code.
  • the obtained linear prediction coefficient code is output to the decoding device.
  • the parameter ⁇ corresponding to the coefficient that can be converted into a linear prediction coefficient and the linear prediction coefficient are used.
  • the linear predictive decoding apparatus includes, for example, a codebook storage unit 311, a decoding unit 313, and a linear conversion unit 314.
  • Each unit of the linear predictive decoding apparatus performs each process illustrated in FIG. 7 to realize the linear predictive decoding method.
  • the codebook storage unit 311 stores the same codebook as the codebook stored in the codebook storage unit 222. That is, the codebook storage unit 311 stores a codebook in which a plurality of coefficient candidates that can be converted into linear prediction coefficients corresponding to the parameter ⁇ 2 are stored.
  • the decoding unit 313 receives the linear prediction coefficient code output from the linear prediction encoding apparatus.
  • the decoding unit 313 is a coefficient candidate that can be converted into a linear prediction coefficient corresponding to the input linear prediction coefficient code, among coefficient candidates that can be converted into a plurality of linear prediction coefficients stored in the codebook storage unit 311. Are obtained as coefficients that can be converted into linear prediction coefficients (step DD1).
  • the obtained coefficient that can be converted into the linear prediction coefficient is output to the linear conversion unit 314.
  • the obtained coefficient that can be converted into a linear prediction coefficient is any one of coefficient candidates that can be converted into a plurality of linear prediction coefficients corresponding to the parameter ⁇ 2 stored in the codebook storage unit 311. For this reason, the coefficient that can be converted into the linear prediction coefficient obtained by the decoding unit 313 is a coefficient that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 2 .
  • ⁇ Linear conversion unit 314 A coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 2 obtained by the decoding unit 313 and the parameter ⁇ 1 are input to the linear conversion unit 314.
  • This parameter ⁇ 1 is obtained, for example, by decoding a parameter code received from the linear predictive coding apparatus.
  • the linear conversion unit 314 performs a linear conversion according to at least the parameter ⁇ 1 on the coefficient that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 2 to obtain a coefficient that can be converted into a linear prediction coefficient after the linear conversion. .
  • the linear conversion unit 314 can convert a linear prediction coefficient corresponding to the parameter ⁇ 2 by linear conversion according to the input parameter ⁇ 1 and the parameter ⁇ 2 corresponding to the coefficient that can be converted into the linear prediction coefficient.
  • the coefficient is converted into a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 .
  • the obtained coefficient that can be converted into the linear prediction coefficient after the linear conversion is output as a decoding result by the linear prediction decoding apparatus or method.
  • the linear conversion unit 314 may not perform linear conversion.
  • the linear conversion unit 3144 when obtaining the linear convertible to prediction coefficients coefficients corresponding coefficients that can be converted to a linear prediction coefficient corresponding to the parameter eta 2 to be a linear transformation parameters eta 1, parameter eta 1 both A configuration may be adopted in which linear transformation is performed a plurality of times using a parameter ⁇ 4 different from the parameter ⁇ 2 .
  • the linear conversion unit 314 linearly converts a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 2 to obtain a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 4 . Further, the linear conversion unit 314 linearly converts the obtained coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 4 to obtain a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 .
  • the linear conversion unit 225 of the linear prediction coefficient encoding device in the third case is converted into two linear conversions.
  • the same linear transformation can be used as the linear transformation that obtains the coefficient that can be transformed into the linear prediction coefficient corresponding to the parameter ⁇ 3 by converting the coefficient that can be transformed into the linear prediction coefficient corresponding to the parameter ⁇ 1. .
  • the linear conversion unit 314 performs linear prediction corresponding to the parameter ⁇ 2 by combining one linear conversion obtained by combining the linear conversion from the parameter ⁇ 2 to the parameter ⁇ 3 and the linear conversion from the parameter ⁇ 3 to the parameter ⁇ 1 .
  • a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 may be obtained by using a coefficient that can be converted into a coefficient.
  • the obtained coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 is output as a decoding result by the linear prediction decoding apparatus or method.
  • the linear conversion unit 31 like the linear conversion unit 225 of the linear predictive encoding device, the amplitude spectrum corresponding to the coefficient that can be converted into the linear prediction coefficient after the linear conversion as the input ⁇ 1 is smaller.
  • a coefficient that can be converted into a linear prediction coefficient after linear conversion may be obtained by linearly converting the coefficient that can be converted into the linear prediction coefficient obtained by the decoding unit 313 so that the envelope becomes flat.
  • the coefficient that can be converted into the linear prediction coefficient after the linear conversion obtained by the linear conversion unit 314 is the amplitude spectrum envelope series corresponding to the coefficient that can be converted into the linear prediction coefficient obtained by the linear conversion unit 314. used to obtain the non-smoothed spectral envelope sequence is 1 squared series.
  • the 1st linear transformation part 2251, the 2nd linear transformation part 2252, the inverse linear transformation part 226, and the linear transformation part 314 perform the linear transformation shown, for example by the following formula
  • x 1 , x 2 , ... x p , y 1 , y 2 , ... y p-1 , z 2 , z 3 , ... z p are predetermined non-negative numbers and y 1 , y 2 , ... y p-1, z 2, z 3, ... at least one of z p is assumed to be the number of a predetermined positive, x 1, x 2 and K, ... x p, y 1 , y 2, ... y p-1 , z 2 , z 3 ,... z p is a matrix whose elements are 0.
  • x 1 , x 2 , ... x p , y 1 , y 2 , ... y p-1 , z 2 , z 3 , ... z p are specific values that can be converted to linear prediction coefficients before linear transformation
  • parameter ⁇ hereinafter referred to as parameter ⁇ A before linear conversion
  • X 1 , x 2 , ... x p , y 1 , y 2 , ... y p-1 , z 2 , z 3 corresponding to a plurality of different sets of pre-linear transformation parameter ⁇ A and post-linear transformation parameter ⁇ B ,... Z p is stored in advance in a storage unit (not shown).
  • the first linear conversion unit 2251, the second linear conversion unit 2252, the inverse linear conversion unit 226, and the linear conversion unit 314 perform linear conversion, the pre-linear conversion parameter ⁇ A and the post-linear conversion parameter ⁇ B in the linear conversion are performed.
  • X 1 , x 2 ,... x p , y 1 , y 2 ,... y p ⁇ 1 , z 2 , z 3 ,... z p corresponding to the pair and read these values May be used to perform linear transformation according to the above equation.
  • the first linear conversion unit 2251 of the linear conversion unit 225 performs the first linear conversion so that the order of the coefficient candidates that can be converted into the linear prediction coefficient after the first linear conversion becomes smaller as the parameter ⁇ 1 is smaller. You may go.
  • the linear conversion unit 314 may perform linear conversion so that the order of the coefficient that can be converted into the linear prediction coefficient after the linear conversion becomes smaller as the parameter ⁇ 1 is smaller.
  • the coefficients that can be converted to the linear prediction coefficients before the linear conversion before the linear conversion or the candidate orders of the coefficients that can be converted to the linear prediction coefficients, and the coefficients or the linear prediction that can be converted to the linear prediction coefficients after the linear conversion Linear conversion may be performed so that the order of coefficient candidates that can be converted into coefficients is different.
  • the first linear conversion unit 2251 reduces the order of candidate coefficients that can be converted into linear prediction coefficients after linear conversion after performing linear conversion in which the order before linear conversion is the same as the order after linear conversion. May be. In addition, the first linear conversion unit 2251 performs linear conversion in which the order before linear conversion and the order after linear conversion are the same after reducing the order of coefficient candidates that can be converted into linear prediction coefficients after linear conversion. May be.
  • the linear conversion unit 314 may reduce the order of coefficients that can be converted into linear prediction coefficients after linear conversion after performing linear conversion in which the order before linear conversion is the same as the order after linear conversion. .
  • the linear conversion unit 314 may perform linear conversion in which the order before linear conversion and the order after linear conversion are the same after reducing the order of coefficients that can be converted into linear prediction coefficients after linear conversion.
  • the first linear transformation unit 2251 if the parameter eta 1 is small, by integrating a plurality of candidates of convertible coefficient to the linear predictive coefficients after the linear transformation, after the linear transformation as the parameter eta 1 is small The number of candidates for coefficients that can be converted to the linear prediction coefficient may be reduced.
  • the linear predictive encoding apparatus includes, for example, a linear predictive analysis unit 221, a codebook storage unit 222, a codebook selection unit 223, and an encoding unit 224.
  • the frequency domain transform unit 220 is provided outside the linear predictive coding device, but the linear predictive coding device may further include the frequency domain transform unit 220.
  • Each part of the linear predictive coding apparatus performs each process illustrated in FIG. 22 to realize the linear predictive coding method.
  • the time domain sound signal which is a time-series signal, is input to the frequency domain transform unit 220.
  • the frequency domain conversion unit 41 converts the input time domain sound signal into N frequency MDCT coefficient sequences X (0), X (1),..., X (N ⁇ Convert to 1). N is a positive integer.
  • the obtained MDCT coefficient sequence X (0), X (1),..., X (N-1) is output to the linear prediction analysis unit 221.
  • the subsequent processing is performed in units of frames.
  • the frequency domain transforming unit 220 obtains a frequency domain sample sequence corresponding to the time series signal, for example, an MDCT coefficient sequence.
  • the linear prediction analysis unit 221 receives a frequency domain sample sequence, for example, an MDCT coefficient sequence X (0), X (1),..., X (N-1) and a parameter ⁇ corresponding to the frequency domain sample sequence.
  • a frequency domain sample sequence for example, an MDCT coefficient sequence X (0), X (1),..., X (N-1) and a parameter ⁇ corresponding to the frequency domain sample sequence.
  • the parameter ⁇ is a positive number.
  • the parameter ⁇ is determined by, for example, parameter determination units 27 and 27 'described later.
  • the parameter ⁇ is the encoding target of the arithmetic code in the encoding scheme that arithmetically encodes the quantized value of the frequency domain coefficient using the linear prediction envelope as used in the 3GPP3EVS (Enhanced Voice Services) standard, for example.
  • the linear prediction analysis unit 221 is defined by the following equation (A7) using the MDCT coefficient sequence X (0), X (1),..., X (N-1) and ⁇ .
  • a coefficient that can be converted into a linear prediction coefficient is generated by performing linear prediction analysis using ⁇ R (0), ⁇ R (1), ..., ⁇ R (N-1) (step DE1).
  • the coefficient that can be converted into the generated linear prediction coefficient is output to the encoding unit 224.
  • the linear prediction analysis unit 22 firstly performs an inverse Fourier transform in which the absolute value of the MDCT coefficient sequence X (0), X (1),. , That is, in the time domain corresponding to the absolute value of MDCT coefficient sequence X (0), X (1), ..., X (N-1) to the ⁇ th power A pseudo-correlation function signal sequence ⁇ R (0), ⁇ R (1), ..., ⁇ R (N-1) which is a signal string is obtained. Then, the linear prediction analysis unit 22 performs linear prediction analysis using the obtained pseudo correlation function signal sequence ⁇ R (0), ⁇ R (1), ..., ⁇ R (N-1) to obtain a linear prediction coefficient. Generate coefficients that can be converted to.
  • the linear prediction analysis unit 221 obtains ⁇ as a positive number and performs an inverse Fourier transform assuming that the absolute value of the ⁇ power of the frequency domain sample sequence corresponding to the time-series signal is a power spectrum.
  • a coefficient that can be converted into a linear prediction coefficient is obtained by performing a linear prediction analysis using the generated pseudo correlation function signal sequence.
  • the coefficients that can be converted into linear prediction coefficients are, for example, LSP, PARCOR coefficient, ISP, and the like.
  • the coefficient that can be converted into the linear prediction coefficient may be the linear prediction coefficient itself.
  • p is a predetermined positive number
  • the possible degree of the linear prediction coefficient is p-th order.
  • the codebook storage unit 222 stores a plurality of codebooks.
  • each codebook stores a plurality of candidate code pairs.
  • I is a predetermined number of 2 or more and N i is a predetermined number of 2 or more determined according to i
  • a predetermined number of bits are assigned to each code corresponding to a coefficient candidate that can be converted into a linear prediction coefficient.
  • Each code is represented by a predetermined number of assigned bits.
  • each coefficient candidate that can be converted into the linear prediction coefficient is composed of p values.
  • the plurality of codebooks stored in the codebook storage unit 222 differ depending on the codebook selection method of the codebook selection unit 223. Therefore, an example of a plurality of code books stored in the code book storage unit 222 will be described together with an example of a code book selection unit 223 described later.
  • the code book selection unit 223 receives the parameter ⁇ .
  • the codebook selection unit 223 selects a codebook from a plurality of codebooks stored in the codebook storage unit 222 according to ⁇ input (step DE2). Information about the selected codebook is output to the encoding unit 224.
  • the codebook storage unit 222 stores a plurality of codebooks with different numbers of coefficient candidates that can be converted into linear prediction coefficients. Also, the codebook selection unit 223 selects a codebook having a larger number of coefficient candidates that can be converted into linear prediction coefficients from a plurality of codebooks stored in the codebook storage unit 222 as the parameter ⁇ increases. .
  • the parameter ⁇ when the parameter ⁇ is small, the range of coefficients that can be converted to linear prediction coefficients tends to be narrow, so it can be converted to linear prediction coefficients with a small number of coefficient candidates that can be converted to linear prediction coefficients. It is possible to express various coefficients. For this reason, when the parameter is small, even if encoding and decoding are performed using a codebook with a small number of coefficient candidates that can be converted into linear prediction coefficients, the quantization distortion is small, so the encoding and decoding accuracy is not so high. It doesn't get worse.
  • the codebook selection unit 223 increases the number of coefficient candidates that can be converted into linear prediction coefficients from among a plurality of codebooks stored in the codebook storage unit 222 as the parameter ⁇ increases. Select a codebook with many.
  • the determination about the size of the parameter ⁇ in other words, selection of an appropriate codebook can be made based on a threshold value. For example, it is assumed that the number of coefficient candidates that can be converted into linear prediction coefficients in the first codebook is smaller than the number of coefficient candidates that can be converted into linear prediction coefficients in the second codebook. In this case, one threshold value of the parameter ⁇ is determined in advance, and when the input parameter ⁇ is smaller than the threshold value, it is determined that the parameter ⁇ is small and the first codebook is selected. If the input parameter ⁇ is greater than or equal to the threshold, it is determined that the parameter ⁇ is large and the second codebook is selected. When the number of codebooks is 3 or more, the codebook may be selected in the same manner using the threshold value of the number obtained by subtracting 1 from the number of codebooks.
  • the parameter ⁇ when the parameter ⁇ is large, the first layer and the second layer are used, and when the parameter ⁇ is small, only the first layer is used.
  • the determination of whether the parameter ⁇ is large or small can be made based on the threshold value as described above.
  • the coefficient that can be converted to the linear prediction coefficient of the first layer and the code corresponding to the coefficient that can be converted to the input linear prediction coefficient are selected.
  • the one closest to the subtraction value and the corresponding code are selected.
  • the two codes selected in the first layer and the second layer are linear prediction coefficient codes. That is, the linear prediction coefficient code is expressed by 15 bits.
  • the sum of the coefficient candidates that can be converted into the linear prediction coefficients selected in the first layer and the second layer is the quantization result of the coefficients that can be converted into the input linear prediction coefficients.
  • the coefficient that can be converted to the linear prediction coefficient of the first layer and the code corresponding to the coefficient that can be converted to the input linear prediction coefficient are selected.
  • the code selected in the first layer is the linear prediction coefficient code. That is, the linear prediction coefficient code is expressed by 10 bits.
  • the coefficient candidates that can be converted into the linear prediction coefficients selected in the first layer are the quantization results of the coefficients that can be converted into the input linear prediction coefficients.
  • this example is also an example of (1) the first method.
  • the search range of candidate code pairs in one codebook is variable.
  • the search range of candidate code pairs may be narrowed as the parameter ⁇ is small.
  • the codebook storage unit 222 stores, in the 1 / ⁇ power, a series of amplitude spectrum envelopes corresponding to coefficient candidates that can be converted into linear prediction coefficients stored in the codebook.
  • a plurality of codebooks having different degrees of flatness of the non-smoothed spectrum envelope sequence, which is the obtained sequence, are stored.
  • the codebook selection unit 223 corresponds to a coefficient candidate that can be converted into a linear prediction coefficient stored in the codebook from a plurality of codebooks stored in the codebook storage unit 222 as ⁇ is smaller.
  • a codebook is selected in which the non-smoothed spectrum envelope sequence, which is a sequence obtained by raising the amplitude spectrum envelope sequence to the power of 1 / ⁇ , is flatter.
  • the coefficient that can be converted into a linear prediction coefficient is LSP
  • Fig. 5 shows an example of LSP parameter values when parameter ⁇ takes various values.
  • the horizontal axis in FIG. 5 is the parameter ⁇ , and the vertical axis is the LSP parameter.
  • FIG. 5 shows that the smaller the parameter ⁇ , the closer the LSP parameter approaches a value obtained by equally dividing 0 to ⁇ .
  • the coefficients that can be converted into linear prediction coefficients are ISP parameters. That is, when the coefficient that can be converted into the linear prediction coefficient is an ISP parameter, the smaller the parameter ⁇ , the closer the coefficient that can be converted into the linear prediction coefficient that is the ISP parameter is closer to a value obtained by equally dividing 0 to ⁇ .
  • the coefficient that can be converted into the linear prediction coefficient is a PARCOR coefficient
  • the smaller the parameter ⁇ the smaller the coefficient that can be converted into the linear prediction coefficient that is the PARCOR coefficient as a whole.
  • the second method uses these tendencies, and encodes and decodes using candidate coefficients that can be converted into linear prediction coefficients corresponding to the case where the unsmoothed spectrum envelope sequence is flatter as the parameter ⁇ is smaller. By trying to improve the quantization performance.
  • Coefficients that can be converted into linear prediction coefficients corresponding to the flattest non-smoothed spectral envelope are denoted as ⁇ F [1], ⁇ F [2], ..., ⁇ F [p].
  • selection of an appropriate codebook may be performed based on a threshold value.
  • the non-smoothed spectrum envelope sequence which is a sequence obtained by raising the amplitude spectrum envelope sequence corresponding to the candidate coefficients that can be converted into the linear prediction coefficients of the first codebook to the 1 / ⁇ power, is linear in the second codebook.
  • the amplitude spectrum envelope sequence corresponding to the coefficient candidates that can be converted into prediction coefficients is flatter than the unsmoothed spectrum envelope sequence that is a sequence obtained by raising the power to 1 / ⁇ .
  • one threshold value of the parameter ⁇ is determined in advance, and when the input parameter ⁇ is smaller than the threshold value, it is determined that the parameter ⁇ is small and the first codebook is selected. If the input parameter ⁇ is greater than or equal to the threshold, it is determined that the parameter ⁇ is large and the second codebook is selected.
  • the codebook may be selected in the same manner using the threshold value of the number obtained by subtracting 1 from the number of codebooks.
  • the codebook storage unit 222 stores a plurality of codebooks having different intervals between candidate coefficients that can be converted into linear prediction coefficients. Further, the codebook selection unit 223 selects a codebook having a smaller interval between coefficient candidates that can be converted into linear prediction coefficients from a plurality of codebooks stored in the codebook storage unit 222 as ⁇ decreases. To do.
  • the interval between coefficient candidates that can be converted into linear prediction coefficients is any index that represents the width of the interval between coefficient candidates that can be converted into linear prediction coefficients included in the codebook. May be.
  • the interval between coefficient candidates that can be converted to a linear prediction coefficient is the coefficient candidate that can be converted into a certain linear prediction coefficient and the coefficient candidate that can be converted into another linear prediction coefficient. May be an average value of the distance to the distance, or may be a maximum value, a minimum value, or a median value of the distance.
  • the third method uses this tendency.
  • the interval between coefficient candidates that can be converted into linear prediction coefficients is the average value of the distances between coefficient candidates that can be converted into two adjacent linear prediction coefficients included in the codebook. It may be.
  • an appropriate codebook may be selected based on a threshold value. For example, it is assumed that the interval between coefficient candidates that can be converted into linear prediction coefficients in the first codebook is narrower than the interval between coefficient candidates that can be converted into linear prediction coefficients in the second codebook.
  • one threshold value of the parameter ⁇ is determined in advance, and when the input parameter ⁇ is smaller than the threshold value, it is determined that the parameter ⁇ is small and the first codebook is selected. If the input parameter ⁇ is greater than or equal to the threshold, it is determined that the parameter ⁇ is large and the second codebook is selected.
  • the codebook may be selected in the same manner using the threshold value of the number obtained by subtracting 1 from the number of codebooks.
  • Coding section 224 receives information about the coefficients that can be converted into linear prediction coefficients obtained by linear prediction analysis section 221 and the selected codebook obtained by codebook selection section 223.
  • the encoding unit 224 encodes a coefficient that can be converted into a linear prediction coefficient using the selected codebook to obtain a linear prediction coefficient code (step DE3).
  • the obtained linear prediction coefficient code is output to the decoding device.
  • the linear predictive decoding apparatus includes, for example, a codebook storage unit 311, a codebook selection unit 312 and a decoding unit 313.
  • Each unit of the linear predictive decoding apparatus performs each process illustrated in FIG. 24 to realize the linear predictive decoding method.
  • the code book storage unit 311 stores a plurality of code books.
  • Each codebook stores a plurality of candidate code pairs.
  • Candidate pairs are stored.
  • a predetermined number of bits are assigned to each code corresponding to a coefficient candidate that can be converted into a linear prediction coefficient.
  • Each code is represented by a predetermined number of assigned bits.
  • the coefficient candidates that can be converted into linear prediction coefficients are composed of p values.
  • the plurality of codebooks stored in the codebook storage unit 311 differ depending on the codebook selection method of the codebook selection unit 312. Therefore, an example of a plurality of code books stored in the code book storage unit 311 will be described together with an example of a code book selection unit 312 described later.
  • the codebook storage unit 311 stores the same codebook as the plurality of codebooks stored in the codebook storage unit 222.
  • the code book selection unit 312 receives the parameter ⁇ .
  • the parameter ⁇ is obtained by decoding the parameter code.
  • the parameter ⁇ may be the same number predetermined by the encoding device and the decoding device.
  • the codebook selection unit 312 selects a codebook according to ⁇ input from among a plurality of codebooks stored in the codebook storage unit 311 (step DD1). Information about the selected codebook is output to the decoding unit 313.
  • the codebook storage unit 311 stores the same codebook as a plurality of codebooks stored in the codebook storage unit 222. Further, it is assumed that the same selection criteria as the codebook selection criteria by the codebook selection unit 223 of the encoding device are set in the codebook selection unit 312 in advance. As a result, a codebook having the same contents as the codebook selected on the code side is also selected on the decoding side.
  • the decoding unit 313 receives the linear prediction coefficient code output from the encoding device and information on the selected codebook obtained by the codebook selection unit 312. The decoding unit 313 reads the code book specified by the information about the selected code book by the code book storage unit 311.
  • the decoding unit 313 obtains a coefficient that can be converted into a linear prediction coefficient by decoding the linear prediction coefficient code by using the selected codebook (step DD2).
  • the coefficient that can be converted into a linear prediction coefficient is used to obtain a non-smoothed spectrum envelope sequence that is a series obtained by raising the amplitude spectrum envelope sequence corresponding to the coefficient that can be converted into a linear prediction coefficient to the power of 1 / ⁇ .
  • the adapting unit 22A includes at least one of the codebook selecting unit 223 and the linear conversion unit 225, the adapting unit 22A Is adapted to match at least one of the codebook stored in the codebook storage unit 222 and the coefficient that can be converted into the linear prediction coefficient generated by the linear prediction analysis unit 221 based on the inputted ⁇ 1 . It can be said.
  • the matching unit 22A uses a plurality of candidates for coefficients that can be converted into linear prediction coefficients stored in the codebook stored in the codebook storage unit 22 and the linear prediction coefficients obtained by the linear prediction analysis unit 221. It can be said that the convertible coefficient is matched with the value of ⁇ .
  • the matching unit 22A may include, for example, a codebook stored in the codebook storage unit 222 before matching, that is, a parameter ⁇ value corresponding to a plurality of candidates for coefficients that can be converted into linear prediction coefficients, and linear prediction analysis.
  • the adapting unit 22A performs the adaptation so that the values of the two parameters ⁇ are substantially the same after the adaptation.
  • the processing of the first linear conversion unit 2251 of the linear conversion unit 225 described in the first embodiment and the processing of the code book selection unit 223 described in the second embodiment are adapted to the code book stored in the code book storage unit 222. It is an example.
  • the processing of the second linear conversion unit 2252 of the linear conversion unit 225 described in the second embodiment is an example of adaptation of coefficients that can be converted into linear prediction coefficients generated by the linear prediction analysis unit 221.
  • the encoding unit 224 performs encoding using at least one codebook adapted by the adaptation unit 22A and a coefficient that can be converted into a linear prediction coefficient.
  • the encoding unit 224 uses the codebook selected by the codebook selection unit 223 or the codebook adapted by the adaptation unit 22A, and the coefficient or adaptation that can be converted into the linear prediction coefficient by the linear prediction analysis unit 221. It can be said that the coefficient that can be converted into the linear prediction coefficient adapted by the unit 22A is encoded.
  • the encoding unit 224 uses a plurality of candidates for coefficients that can be converted into linear prediction coefficients to which the value of ⁇ is adapted and a coefficient that can be converted into linear prediction coefficients, and then uses the linear prediction analysis unit 221. It can be said that the linear prediction coefficient code corresponding to the coefficient that can be converted into the linear prediction coefficient obtained is obtained.
  • the adapting unit 22A in the first case performs a first linear transformation corresponding to ⁇ 1 on a coefficient candidate that can be transformed into a linear prediction coefficient stored in the codebook storage unit 222. It can be said that a linear conversion unit 225 that obtains a plurality of candidates of coefficients that can be converted into linear prediction coefficients after the first linear conversion is provided.
  • the encoding unit 224 includes a plurality of coefficients that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 and the coefficient that can be converted into the linear prediction coefficient after the first linear conversion obtained by the adaptation unit 22A. It can be said that the linear prediction coefficient code corresponding to the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 is obtained using the candidates.
  • the adapting unit 22A in the second case (2) of the first embodiment performs the second linear transformation according to ⁇ 1 on the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221; It can be said that the linear conversion part 225 which obtains the coefficient which can be converted into the linear prediction coefficient after 2nd linear conversion is provided.
  • the encoding unit 224 has a plurality of candidates for the coefficient that can be converted into the linear prediction coefficient after the second linear conversion obtained by the adaptation unit 22A and the coefficient that can be converted into the linear prediction coefficient stored in the codebook.
  • the linear prediction coefficient code corresponding to the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 is obtained.
  • the adapting unit 22A in the third case of the first embodiment assumes that the codebook storage unit 222 stores the codebook corresponding to ⁇ 2 , and stores the linear data stored in the codebook storage unit 222.
  • a plurality of candidates for coefficients that can be converted into prediction coefficients are subjected to a first linear transformation according to ⁇ 3 to obtain a plurality of candidates for coefficients that can be converted into linear prediction coefficients after the first linear transformation,
  • a coefficient that can be converted into a linear prediction coefficient obtained by the linear prediction analysis unit 221 is subjected to a second linear conversion according to ⁇ 3 to obtain a coefficient that can be converted into a linear prediction coefficient after the second linear conversion. It can be said.
  • the encoding unit 224 can convert the coefficient that can be converted into the linear prediction coefficient after the second linear transformation obtained by the adaptation unit 22A and the linear prediction coefficient after the first linear transformation obtained by the adaptation unit 22A. It can be said that a linear prediction coefficient code corresponding to a coefficient that can be converted into a linear prediction coefficient obtained by the linear prediction analysis unit is obtained using a plurality of coefficient candidates.
  • the adaptation unit 22A may perform codebook adaptation using, for example, the codebook selection unit 223 and the second linear conversion unit 2252 illustrated in FIG.
  • the code book selection unit 223 selects a code book from a plurality of code books stored in the code book storage unit 222 according to the parameter ⁇ 2 .
  • the second linear conversion unit 2252 performs the second linear conversion according to ⁇ 2 on the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221.
  • the encoding unit 224 encodes the coefficient that can be converted into the linear prediction coefficient after the second linear conversion using the selected codebook to obtain a linear prediction coefficient code.
  • the adaptation unit 22A may perform codebook adaptation using, for example, the codebook selection unit 223 and the first linear conversion unit 2251 illustrated in FIG.
  • the code book selection unit 223 selects a code book from a plurality of code books stored in the code book storage unit 222 according to the parameter ⁇ 2 .
  • the first linear conversion unit 2251 performs a first linear conversion corresponding to ⁇ 1 on a plurality of candidates for coefficients that can be converted into linear prediction coefficients stored in the selected codebook.
  • the encoding unit 224 encodes the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221 using the coefficient candidates that can be converted into the linear prediction coefficient after the first linear conversion. Obtain the linear prediction coefficient code.
  • the adaptation unit 22A may perform codebook adaptation using, for example, the codebook selection unit 223, the first linear conversion unit 2251, and the second conversion unit 2252 shown in FIG.
  • the code book selection unit 223 selects a code book from a plurality of code books stored in the code book storage unit 222 according to the parameter ⁇ 3. To do.
  • the first linear conversion unit 2251 performs the first linear conversion corresponding to ⁇ 2 on a plurality of candidates for coefficients that can be converted into linear prediction coefficients stored in the selected codebook.
  • the second linear conversion unit 2252 performs the second linear conversion according to ⁇ 2 on the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 221.
  • the encoding unit 224 encodes the coefficients that can be converted into the linear prediction coefficients after the second linear conversion using the coefficient candidates that can be converted into the linear prediction coefficients after the first linear conversion, and linear prediction coefficients. Get the sign.
  • the adaptation unit 31A includes at least one of the codebook selection unit 312 and the linear conversion unit 314 and the decoding unit 313, the adaptation unit 31A is, eta 1 as a positive number, on the basis of eta 1 inputted, the codebook stored in the codebook storage unit 311, convertible coefficients into a plurality of linear prediction coefficients stored in the codebook Among the candidates, it can be said that at least one of the candidates of coefficients that can be converted into linear prediction coefficients corresponding to the input linear prediction coefficient code is adapted.
  • the adaptation unit 31A may perform the adaptation process in both the codebook selection unit 312 and the linear conversion unit 314 illustrated in FIG.
  • the code book selection unit 312 selects a code book from a plurality of code books stored in the code book storage unit 311 according to the parameter ⁇ 2 .
  • the linear conversion unit 314 can convert the coefficient that can be converted into the linear prediction coefficient obtained by the decoding unit 313 into a linear prediction coefficient by performing linear conversion according to ⁇ 1 that is a predetermined positive number. Get a good coefficient.
  • the encoding apparatus according to the first embodiment includes a frequency domain transform unit 21, a linear prediction analysis unit 22, a non-smoothed amplitude spectrum envelope sequence generation unit 23, and a smoothed amplitude spectrum envelope sequence generation.
  • a unit 24, an envelope normalization unit 25, an encoding unit 26, and a parameter determination unit 27 are provided.
  • An example of each process of the encoding method of the first embodiment realized by this encoding apparatus is shown in FIG.
  • any one of a plurality of parameters ⁇ can be selected by the parameter determination unit 27 for each predetermined time interval.
  • the parameter determination unit 27 stores a plurality of parameters ⁇ as parameters ⁇ candidates.
  • the parameter determination unit 27 sequentially reads one parameter ⁇ among the plurality of parameters, and outputs it to the linear prediction analysis unit 22, the unsmoothed amplitude spectrum envelope sequence generation unit 23, and the decoding unit 26 (step A0).
  • the frequency domain transform unit 21, the linear prediction analysis unit 22, the unsmoothed amplitude spectrum envelope sequence generation unit 23, the smoothed amplitude spectrum envelope sequence generation unit 24, the envelope normalization unit 25, and the encoding unit 26 include a parameter determination unit 27.
  • processing from step A1 to step A6 described below is performed to generate a code for the frequency domain sample sequence corresponding to the time-series signal in the same predetermined time interval.
  • two or more codes may be obtained for frequency domain sample sequences corresponding to time-series signals in the same predetermined time interval.
  • the codes for the frequency domain sample sequences corresponding to the time-series signals in the same predetermined time section are a combination of these two or more obtained codes.
  • the code is a combination of a linear prediction coefficient code, a gain code, and an integer signal code.
  • the parameter determination unit 27 selects one code from the codes obtained for each parameter ⁇ with respect to the frequency domain sample sequence corresponding to the time-series signal in the same predetermined time interval. Then, the parameter ⁇ corresponding to the selected code is determined (step A7). This determined parameter ⁇ becomes the parameter ⁇ for the frequency domain sample sequence corresponding to the time-series signal in the same predetermined time interval. Then, the parameter determining unit 27 outputs the selected code and the code representing the determined parameter ⁇ to the decoding device. Details of the process of step A7 by the parameter determination unit 27 will be described later.
  • the frequency domain converter 21 receives a sound signal that is a time-series signal in the time domain.
  • sound signals are voice digital signals or acoustic digital signals.
  • the frequency domain transform unit 21 converts the input time domain sound signal into N frequency MDCT coefficient sequences X (0), X (1),..., X (N ⁇ 1) (step A1). N is a positive integer.
  • the obtained MDCT coefficient sequences X (0), X (1),..., X (N-1) are output to the linear prediction analysis unit 22 and the envelope normalization unit 25.
  • the subsequent processing is performed in units of frames.
  • the frequency domain conversion unit 21 obtains a frequency domain sample sequence corresponding to the sound signal, for example, an MDCT coefficient sequence.
  • the linear prediction analysis unit 22 receives the MDCT coefficient sequence X (0), X (1),..., X (N-1) obtained by the frequency domain conversion unit 21.
  • the linear prediction analysis unit 22 is the linear prediction encoding device of any one of FIGS. 1 to 3 and FIG. 21 described in [Linear prediction encoding device, linear prediction decoding device and methods thereof]. In [Encoder, Decoder, and These Methods] and FIG. 8, the linear prediction of any of FIGS. 1 to 3 and FIG. 21 described in [Linear Predictive Encoder, Linear Predictive Decoder, and These Methods].
  • the encoding device is expressed as “linear prediction analysis unit 22”. Note that the linear prediction analysis unit 22 may be any one of the linear prediction encoding apparatuses shown in FIGS.
  • the linear prediction analysis unit 22 performs the same process as described in [Linear prediction encoding apparatus, linear prediction decoding apparatus and their methods], for example, the absolute value of the absolute value of the frequency domain sample sequence that is an MDCT coefficient sequence to the ⁇ 1 power Can be converted into linear prediction coefficients by performing linear prediction analysis using the pseudo correlation function signal sequence obtained by performing inverse Fourier transform assuming that the power spectrum is a power spectrum.
  • a linear prediction coefficient code is obtained by encoding a simple coefficient.
  • the obtained linear prediction coefficient code is output to the parameter determination unit 27 and the decoding device.
  • linear conversion unit 225 of the linear prediction encoding apparatus (1) is the first case, conversion to a linear prediction coefficient corresponding to the parameter ⁇ 1 corresponding to the linear prediction coefficient code obtained by the encoding unit 224 is performed. possible coefficients quantized linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., as ⁇ beta p, is output to the non-smoothed spectral envelope sequence generation unit 23 and smoothed amplitude spectrum envelope sequence generator 24.
  • the linear conversion unit 225 of the linear prediction encoding apparatus can convert the linear prediction coefficient corresponding to the parameter ⁇ 2 corresponding to the linear prediction coefficient code obtained by the encoding unit 224.
  • the coefficient is input to the inverse linear transformation unit 226 indicated by a broken line in FIG.
  • the inverse linear transformation unit 226 performs inverse linear transformation of the second linear transformation performed by the second linear transformation unit 2252 on the coefficient that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 2 , corresponding to the linear prediction coefficient code. To obtain a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 .
  • the coefficients that can be converted into the linear prediction coefficients corresponding to the parameter ⁇ 1 are the quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p , and the unsmoothed spectrum envelope sequence generation unit 23 and the smoothed amplitude. It is output to the spectrum envelope sequence generation unit 24.
  • the inverse linear conversion unit 226 may not perform linear conversion.
  • the linear conversion unit 225 of the linear prediction encoding apparatus can convert the linear prediction coefficient corresponding to the parameter ⁇ 3 corresponding to the linear prediction coefficient code obtained by the encoding unit 224.
  • the coefficient is input to the inverse linear transformation unit 226 indicated by a broken line in FIG.
  • the inverse linear transformation unit 226 performs inverse linear transformation of the second linear transformation performed by the second linear transformation unit 2252 on the coefficient that can be converted into the linear prediction coefficient corresponding to the parameter ⁇ 3 , corresponding to the linear prediction coefficient code. To obtain a coefficient that can be converted into a linear prediction coefficient corresponding to the parameter ⁇ 1 .
  • the coefficients that can be converted into the linear prediction coefficients corresponding to the parameter ⁇ 1 are the quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p , and the unsmoothed spectrum envelope sequence generation unit 23 and the smoothed amplitude. It is output to the spectrum envelope sequence generation unit 24.
  • the inverse linear conversion unit 226 does not have to perform linear conversion.
  • the energy ⁇ 2 of the prediction residual is calculated in the course of the linear prediction analysis process.
  • the calculated energy ⁇ 2 of the prediction residual is output to the variance parameter determining unit 268 of the encoding unit 26.
  • the unsmoothed amplitude spectrum envelope sequence generation unit 23 receives the quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p generated by the linear prediction analysis unit 22.
  • Textured amplitude spectral envelope sequence generating unit 23 the quantized linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., ⁇ ⁇ is the sequence of the amplitude spectrum envelope corresponding to p textured amplitude spectral envelope sequence ⁇ H ( 0), ⁇ H (1),..., ⁇ H (N-1) are generated (step A3).
  • the generated non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1) is output to the encoding unit 26.
  • Textured amplitude spectral envelope sequence generating unit 23 the quantized linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., using the ⁇ beta p, unsmoothed amplitude spectral envelope sequence ⁇ H (0), ⁇ H ( 1), ..., ⁇ H (N-1), the unsmoothed amplitude spectrum envelope sequence defined by equation (A2) ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1) Is generated.
  • the unsmoothed amplitude spectrum envelope sequence generation unit 23 is a sequence obtained by raising the amplitude spectrum envelope sequence corresponding to the coefficient that can be converted into the linear prediction coefficient generated by the linear prediction analysis unit 22 to the 1 / ⁇ 1 power.
  • the spectral envelope is estimated by obtaining the unsmoothed spectral envelope sequence.
  • the sequence obtained by raising c to a power of a sequence composed of a plurality of values, where c is an arbitrary number is a sequence composed of values obtained by raising each of the plurality of values to the c-th power.
  • a series obtained by raising the amplitude spectrum envelope series to the 1 / ⁇ 1 power is a series composed of values obtained by raising each coefficient of the amplitude spectrum envelope to the 1 / ⁇ 1 power.
  • the processing of the 1 / ⁇ 1 power by the non-smoothed amplitude spectrum envelope sequence generation unit 23 is caused by the processing in which the absolute value ⁇ 1 power of the frequency domain sample sequence performed by the linear prediction analysis unit 22 is regarded as the power spectrum. To do. That, 1 / eta 1 square of processing by the non-smoothed amplitude spectrum envelope sequence generating unit 23, and the eta 1 square of the absolute value of the frequency domain sample sequences performed by the linear prediction analyzer 22 regarded as a power spectral processing Is performed to return the value raised to the power of ⁇ 1 to the original value.
  • ⁇ Smoothing Amplitude Spectrum Envelope Sequence Generation Unit 24 Quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p generated by the linear prediction analysis unit 22 are input to the smoothed amplitude spectrum envelope sequence generation unit 24.
  • the generated smoothed amplitude spectrum envelope sequences ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N ⁇ 1) are output to the envelope normalization unit 25 and the encoding unit 26.
  • the smoothed amplitude spectrum envelope sequence generation unit 24 uses the quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p and the correction coefficient ⁇ to smooth the smoothed amplitude spectrum envelope sequence ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N-1), the smoothed amplitude spectrum envelope sequence defined by equation (A3) ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N-1) is generated.
  • the correction coefficient ⁇ is a predetermined constant less than 1, and the amplitude unevenness of the unsmoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1)
  • the coefficient for blunting in other words, the coefficient for smoothing the unsmoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1).
  • the envelope normalization unit 25 includes the MDCT coefficient sequence X (0), X (1),..., X (N-1) obtained by the frequency domain conversion unit 21 and the smoothed amplitude spectrum envelope generation unit 24. ⁇ H ⁇ (0), ⁇ H ⁇ (1), ..., ⁇ H ⁇ (N-1) are input.
  • the envelope normalization unit 25 converts each coefficient of the MDCT coefficient sequence X (0), X (1),..., X (N-1) into a corresponding smoothed amplitude spectrum envelope sequence ⁇ H ⁇ (0), ⁇ H. Normalized MDCT coefficient sequence X N (0), X N (1), ..., X N (N-1 by normalizing with each value of ⁇ (1), ..., ⁇ H ⁇ (N-1) ) Is generated (step A5).
  • the generated normalized MDCT coefficient sequence is output to the encoding unit 26.
  • the encoding unit 26 includes normalized MDCT coefficient sequences X N (0), X N (1),..., X N (N ⁇ 1) generated by the envelope normalization unit 25, an unsmoothed amplitude spectrum envelope generation unit. 23, the non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1), and the smoothed amplitude spectrum envelope sequence generated by the smoothed amplitude spectrum envelope generation unit 24 ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N ⁇ 1) and the prediction residual energy ⁇ 2 calculated by the linear prediction analysis unit 22 are input.
  • the encoding unit 26 performs encoding, for example, by performing the processing from step A61 to step A65 shown in FIG. 12 (step A6).
  • the encoding unit 26 obtains a global gain g corresponding to the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) (step A61), and the normalized MDCT coefficient sequence Quantized normalized coefficient series X, which is a series of integer values obtained by quantizing the result of dividing each coefficient of X N (0), X N (1), ..., X N (N-1) by global gain g Q (0), X Q (1), ..., X Q (N-1) is obtained (step A62), and the quantized normalized coefficient series X Q (0), X Q (1), ..., X Q Dispersion parameters ⁇ (0), ⁇ (1), ..., ⁇ (N-1) corresponding to each coefficient of (N-1) are set to global gain g and unsmoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1) and smoothed amplitude spectrum envelope series ⁇ H ⁇ (0), ⁇
  • the normalized amplitude spectrum envelope sequence in the above formula (A1) ⁇ H N (0 ), ⁇ H N (1), ..., ⁇ H N is unsmoothed amplitude spectral envelope sequence ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1) values are converted into corresponding smoothed amplitude spectrum envelope sequences ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N- Divided by each value of 1), that is, obtained by the following equation (A8).
  • the generated integer signal code and gain code are output to the parameter determination unit 27 as codes corresponding to the normalized MDCT coefficient sequence.
  • step A61 to step A65 the encoding unit 26 determines a global gain g such that the number of bits of the integer signal code is equal to or smaller than the allocated bit number B, which is the number of bits allocated in advance, and as large as possible.
  • a function of generating a gain code corresponding to the determined global gain g and an integer signal code corresponding to the determined global gain g is realized.
  • step A63 the characteristic processing is included in step A63, where the global gain g and the quantized normalized coefficient series X Q (0), X Q (1 ),..., X Q (N-1) are encoded to obtain a code corresponding to the normalized MDCT coefficient sequence.
  • the encoding process itself includes various techniques including those described in Non-Patent Document 1. Known techniques exist. Two specific examples of the encoding process performed by the encoding unit 26 will be described below.
  • FIG. 10 shows a configuration example of the encoding unit 26 of the first specific example.
  • the encoding unit 26 of the first specific example includes a gain acquisition unit 261, a quantization unit 262, a dispersion parameter determination unit 268, an arithmetic encoding unit 269, and a gain encoding unit 265.
  • a gain acquisition unit 261 As shown in FIG. 10, the encoding unit 26 of the first specific example includes a gain acquisition unit 261, a quantization unit 262, a dispersion parameter determination unit 268, an arithmetic encoding unit 269, and a gain encoding unit 265.
  • the gain acquisition unit 261 receives the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) generated by the envelope normalization unit 25.
  • Gain acquisition unit 261 the normalized MDCT coefficients X N (0), X N (1), ..., from X N (N-1), the number of bits of the integer signal code is the number of bits in advance allocation
  • a global gain g that is equal to or less than the number of allocated bits B and that is as large as possible is determined and output (step S261).
  • the gain acquisition unit 261 has, for example, a negative correlation between the square root of the total energy of the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) and the allocated bit number B.
  • the multiplication value with a certain constant is obtained as the global gain g and output.
  • the gain acquisition unit 261 calculates the total energy of the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1), the number of allocated bits B, and the global gain g. , And a global gain g may be obtained and output by referring to the table.
  • the gain acquisition unit 261 obtains a gain for dividing all samples of the normalized frequency domain sample sequence, which is a normalized MDCT coefficient sequence, for example.
  • the obtained global gain g is output to the quantization unit 262 and the dispersion parameter determination unit 268.
  • the quantization unit 262 includes the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) generated by the envelope normalization unit 25 and the global obtained by the gain acquisition unit 261. Gain g is input.
  • the quantization unit 262 is a series of integer parts as a result of dividing each coefficient of the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) by the global gain g. quantized normalized haze coefficient sequence X Q (0), X Q (1), ..., X Q (N-1) the obtained output (step S262).
  • the quantization unit 262 divides each sample of the normalized frequency domain sample sequence, which is a normalized MDCT coefficient sequence, for example, by the gain and quantizes it to obtain a quantized normalized coefficient sequence.
  • the obtained quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N ⁇ 1) are output to the arithmetic coding unit 269.
  • the dispersion parameter determination unit 268 includes the parameter ⁇ 1 read by the parameter determination unit 27, the global gain g obtained by the gain acquisition unit 261, and the unsmoothed amplitude spectrum envelope sequence generated by the unsmoothed amplitude spectrum envelope generation unit 23 H (0), ⁇ H (1), ..., ⁇ H (N-1), the smoothed amplitude spectrum envelope sequence generated by the smoothed amplitude spectrum envelope generator 24 ⁇ H ⁇ (0), ⁇ H ⁇ (1 ),..., ⁇ H ⁇ (N ⁇ 1) and the energy ⁇ 2 of the prediction residual obtained by the linear prediction analysis unit 22 are input.
  • the dispersion parameter determination unit 268 calculates the global gain g, the unsmoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1), and the smoothed amplitude spectrum envelope sequence ⁇ H. ⁇ (0), ⁇ H ⁇ (1), ..., ⁇ H ⁇ (N-1) and, from the energy sigma 2 Metropolitan prediction residual, the above formula (A1), the dispersion parameter sequence by formula (A8) phi Each of the dispersion parameters (0), ⁇ (1),..., ⁇ (N ⁇ 1) is obtained and output (step S268).
  • the obtained dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N ⁇ 1) are output to the arithmetic coding unit 269.
  • the arithmetic encoding unit 269 includes the parameter ⁇ 1 read by the parameter determination unit 27 and the quantized normalized coefficient series X Q (0), X Q (1),..., X Q ( N-1) and the dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N-1) obtained by the dispersion parameter determination unit 268 are input.
  • the arithmetic coding unit 269 uses a dispersion parameter sequence ⁇ (0) as a dispersion parameter corresponding to each coefficient of the quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N ⁇ 1). ), ⁇ (1), ..., ⁇ (N-1) using the respective dispersion parameters, the quantized normalized coefficient series X Q (0), X Q (1), ..., X Q (N-1 ) Is arithmetically encoded to obtain and output an integer signal code (step S269).
  • the arithmetic coding unit 269 performs generalized Gaussian distribution on each coefficient of the quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N ⁇ 1) during arithmetic coding.
  • ⁇ (k), ⁇ 1 ) is configured, and encoding is performed using the arithmetic code based on this configuration.
  • the expected value of the bit allocation to each coefficient of the quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N-1) is expressed as the dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N ⁇ 1).
  • the obtained integer signal code is output to the parameter determination unit 27.
  • Quantized normalized Haze coefficient sequence X Q (0), X Q (1), ..., arithmetic coding may be performed over a plurality of coefficients in X Q (N-1).
  • the dispersion parameters of the dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N-1) are unsmoothed amplitude spectrum envelopes as can be seen from equations (A1) and (A8). Since it is based on the sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1), the arithmetic coding unit 269 is based on the estimated spectral envelope (unsmoothed amplitude spectral envelope). Thus, it can be said that encoding is performed in which the bit allocation is substantially changed.
  • the gain encoder 265 receives the global gain g obtained by the gain acquisition unit 261.
  • the gain encoding unit 265 encodes the global gain g to obtain and output a gain code (step S265).
  • the generated integer signal code and gain code are output to the parameter determination unit 27 as codes corresponding to the normalized MDCT coefficient sequence.
  • Steps S261, S262, S268, S269, and S265 of this specific example 1 correspond to the above steps A61, A62, A63, A64, and A65, respectively.
  • FIG. 11 shows a configuration example of the encoding unit 26 of the specific example 2.
  • the encoding unit 26 of the specific example 2 includes a gain acquisition unit 261, a quantization unit 262, a dispersion parameter determination unit 268, an arithmetic encoding unit 269, a gain encoding unit 265, For example, a determination unit 266 and a gain update unit 267 are provided.
  • a gain acquisition unit 261 the encoding unit 26 of the specific example 2
  • a quantization unit 262 includes a quantization unit 262, a dispersion parameter determination unit 268, an arithmetic encoding unit 269, a gain encoding unit 265,
  • a determination unit 266 and a gain update unit 267 are provided.
  • the gain unit 261 receives the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) generated by the envelope normalization unit 25.
  • Gain acquisition unit 261 the normalized MDCT coefficients X N (0), X N (1), ..., from X N (N-1), the number of bits of the integer signal code is the number of bits in advance allocation
  • a global gain g that is equal to or less than the number of allocated bits B and that is as large as possible is determined and output (step S261).
  • the gain acquisition unit 261 has, for example, a negative correlation between the square root of the total energy of the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) and the allocated bit number B.
  • the multiplication value with a certain constant is obtained as the global gain g and output.
  • the obtained global gain g is output to the quantization unit 262 and the dispersion parameter determination unit 268.
  • the global gain g obtained by the gain acquisition unit 261 is an initial value of the global gain used by the quantization unit 262 and the dispersion parameter determination unit 268.
  • the quantization unit 262 includes a normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) generated by the envelope normalization unit 25 and a gain acquisition unit 261 or a gain update unit.
  • the global gain g obtained by 267 is input.
  • the quantization unit 262 is a series of integer parts as a result of dividing each coefficient of the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1) by the global gain g. quantized normalized haze coefficient sequence X Q (0), X Q (1), ..., X Q (N-1) the obtained output (step S262).
  • the global gain g used when the quantization unit 262 is executed for the first time is the global gain g obtained by the gain acquisition unit 261, that is, the initial value of the global gain.
  • the global gain g used when the quantizing unit 262 is executed for the second time or later is the global gain g obtained by the gain updating unit 267, that is, the updated value of the global gain.
  • the obtained quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N ⁇ 1) are output to the arithmetic coding unit 269.
  • the dispersion parameter determination unit 268 includes the parameter ⁇ 1 read by the parameter determination unit 27, the global gain g obtained by the gain acquisition unit 261 or the gain update unit 267, and the non-smoothing generated by the non-smoothed amplitude spectrum envelope generation unit 23.
  • smoothed amplitude spectrum envelope sequence generated by the smoothed amplitude spectrum envelope generator 24 ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N ⁇ 1) and the energy ⁇ 2 of the prediction residual obtained by the linear prediction analysis unit 22 are input.
  • the dispersion parameter determination unit 268 calculates the global gain g, the unsmoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1), and the smoothed amplitude spectrum envelope sequence ⁇ H. ⁇ (0), ⁇ H ⁇ (1), ..., ⁇ H ⁇ (N-1) and, from the energy sigma 2 Metropolitan prediction residual, the above formula (A1), the dispersion parameter sequence by formula (A8) phi Each of the dispersion parameters (0), ⁇ (1),..., ⁇ (N ⁇ 1) is obtained and output (step S268).
  • the global gain g used when the dispersion parameter determination unit 268 is executed for the first time is the global gain g obtained by the gain acquisition unit 261, that is, the initial value of the global gain.
  • the global gain g used when the dispersion parameter determination unit 268 is executed for the second time or later is the global gain g obtained by the gain update unit 267, that is, the updated value of the global gain.
  • the obtained dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N ⁇ 1) are output to the arithmetic coding unit 269.
  • the arithmetic encoding unit 269 includes the parameter ⁇ 1 read by the parameter determination unit 27 and the quantized normalized coefficient series X Q (0), X Q (1),..., X Q ( N-1) and the dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N-1) obtained by the dispersion parameter determination unit 268 are input.
  • the arithmetic coding unit 269 uses a dispersion parameter sequence ⁇ (0) as a dispersion parameter corresponding to each coefficient of the quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N ⁇ 1). ), ⁇ (1), ..., ⁇ (N-1) using the respective dispersion parameters, the quantized normalized coefficient series X Q (0), X Q (1), ..., X Q (N-1 ) Are arithmetically encoded to obtain and output an integer signal code and a consumed bit number C that is the number of bits of the integer signal code (step S269).
  • the arithmetic coding unit 269 performs generalized Gaussian distribution on each coefficient of the quantized normalized coefficient series X Q (0), X Q (1),..., X Q (N ⁇ 1) during arithmetic coding. Bit allocation that is optimal when obeying f GG (X
  • the obtained integer signal code and the number C of consumed bits are output to the determination unit 266.
  • Quantized normalized Haze coefficient sequence X Q (0), X Q (1), ..., arithmetic coding may be performed over a plurality of coefficients in X Q (N-1).
  • the dispersion parameters of the dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N-1) are unsmoothed amplitude spectrum envelopes as can be seen from equations (A1) and (A8). Since it is based on the sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1), the arithmetic coding unit 269 is based on the estimated spectral envelope (unsmoothed amplitude spectral envelope). Thus, it can be said that encoding is performed in which the bit allocation is substantially changed.
  • ⁇ Determining unit 266 The integer signal code obtained by the arithmetic coding unit 269 is input to the determination unit 266.
  • the determination unit 266 outputs an integer signal code when the number of gain updates is a predetermined number, and also instructs the gain encoding unit 265 to encode the global gain g obtained by the gain updating unit 267.
  • the gain update count is less than the predetermined count, the consumed bit count C measured by the arithmetic encoding section 264 is output to the gain update section 267 (step S266).
  • the gain updating unit 267 receives the number of consumed bits C measured by the arithmetic coding unit 264.
  • the gain updating unit 267 updates the global gain g value when the consumed bit number C is greater than the allocated bit number B, and outputs the updated value.
  • the gain g is updated to a smaller value, and the updated global gain g is output (step S267).
  • the updated global gain g obtained by the gain update unit 267 is output to the quantization unit 262 and the gain encoding unit 265.
  • the gain encoding unit 265 receives the output instruction from the determination unit 266 and the global gain g obtained by the gain update unit 267.
  • the gain encoder 265 encodes the global gain g according to the instruction signal to obtain and output a gain code (step 265).
  • the integer signal code output from the determination unit 266 and the gain code output from the gain encoding unit 265 are output to the parameter determination unit 27 as codes corresponding to the normalized MDCT coefficient sequence.
  • step S267 performed last corresponds to the above step A61
  • steps S262, S263, S264, and S265 correspond to the above steps A62, A63, A64, and A65, respectively.
  • the encoding unit 26 may perform encoding that changes the bit allocation based on the estimated spectral envelope (non-smoothed amplitude spectral envelope), for example, by performing the following processing.
  • the encoding unit 26 first obtains a global gain g corresponding to the normalized MDCT coefficient sequence X N (0), X N (1),..., X N (N ⁇ 1), and normalizes the MDCT coefficient sequence X N. (0), X N (1), ..., X N (N-1) coefficients divided by the global gain g Quantized normalized coefficient series X Q ( Find 0), X Q (1), ..., X Q (N-1).
  • the quantized bit corresponding to each coefficient of this quantized normalized coefficient series X Q (0), X Q (1), ..., X Q (N-1) has a range in which X Q (k) is distributed.
  • the range can be determined from the envelope estimate.
  • the encoding unit 26 for example, the value of the normalized amplitude spectrum envelope sequence based on linear prediction as in the following equation (A9) ⁇ H N ( k) can be used to determine the range of X Q (k).
  • the encoding unit 26 determines the number of allocated bits by collecting a plurality of samples instead of assigning each sample, and the quantization unit 26 does not perform scalar quantization for each sample but also a vector for each vector including a plurality of samples. It is also possible to quantize.
  • X Q (k) can be changed from -2 b (k) -1 to 2 b (k ) Can take 2 b (k) types of integers up to -1 .
  • the encoding unit 26 encodes each sample with b (k) bits to obtain an integer signal code.
  • the generated integer signal code is output to the decoding device.
  • the encoding unit 26 encodes the global gain g to obtain and output a gain code.
  • the encoding unit 26 may perform encoding other than arithmetic encoding.
  • ⁇ Parameter determining unit 27 Through the processing from step A1 to step A6, codes generated for each parameter ⁇ 1 for frequency domain sample sequences corresponding to time-series signals in the same predetermined time interval (in this example, linear prediction coefficient code, gain The code and the integer signal code) are input to the parameter determination unit 27.
  • codes generated for each parameter ⁇ 1 for frequency domain sample sequences corresponding to time-series signals in the same predetermined time interval (in this example, linear prediction coefficient code, gain The code and the integer signal code) are input to the parameter determination unit 27.
  • the parameter determination unit 27 selects one code from the codes obtained for each parameter ⁇ 1 for the frequency domain sample sequence corresponding to the time-series signal in the same predetermined time interval, and selects the selected code Is determined (step A7). This determined parameter ⁇ becomes the parameter ⁇ for the frequency domain sample sequence corresponding to the time-series signal in the same predetermined time interval. Then, the parameter determining unit 27 outputs the selected code and the parameter code representing the determined parameter ⁇ to the decoding device. The selection of the code is performed based on at least one of the code amount of the code and the coding distortion corresponding to the code. For example, the code with the smallest code amount or the code with the smallest coding distortion is selected.
  • the coding distortion is an error between the frequency domain sample sequence obtained from the input signal and the frequency domain sample sequence obtained by locally decoding the generated code.
  • the encoding apparatus may include an encoding distortion calculation unit for calculating encoding distortion.
  • the encoding distortion calculation unit includes a decoding unit that performs processing similar to that of the decoding device described below, and locally decodes the code generated by the decoding unit. Thereafter, the coding distortion calculation unit calculates an error between the frequency domain sample sequence obtained from the input signal and the frequency domain sample sequence obtained by local decoding, and obtains the coding distortion.
  • FIG. 13 shows a configuration example of a decoding device corresponding to the encoding device.
  • the decoding device according to the first embodiment includes a linear prediction coefficient decoding unit 31, a non-smoothed amplitude spectrum envelope sequence generating unit 32, a smoothed amplitude spectrum envelope sequence generating unit 33, and a decoding unit 34. And an envelope denormalization unit 35, a time domain conversion unit 36, and a parameter decoding unit 37, for example.
  • FIG. 13 shows a configuration example of a decoding device corresponding to the encoding device.
  • the decoding device includes a linear prediction coefficient decoding unit 31, a non-smoothed amplitude spectrum envelope sequence generating unit 32, a smoothed amplitude spectrum envelope sequence generating unit 33, and a decoding unit 34.
  • an envelope denormalization unit 35, a time domain conversion unit 36, and a parameter decoding unit 37 for example.
  • the decoding apparatus receives at least the parameter code, the code corresponding to the normalized MDCT coefficient sequence, and the linear prediction coefficient code output from the encoding apparatus.
  • ⁇ Parameter decoding unit 37> The parameter code output from the encoding device is input to the parameter decoding unit 37.
  • the parameter decoding unit 37 obtains a decoding parameter ⁇ by decoding the parameter code.
  • the obtained decoding parameter ⁇ is output to the linear prediction coefficient decoding unit 31, the unsmoothed amplitude spectrum envelope sequence generation unit 32, the smoothed amplitude spectrum envelope sequence generation unit 33, and the decoding unit 34.
  • the parameter decoding unit 37 stores a plurality of decoding parameters ⁇ as candidates.
  • the parameter decoding unit 37 obtains a decoding parameter ⁇ candidate corresponding to the parameter code as a decoding parameter ⁇ .
  • the plurality of decoding parameters ⁇ stored in the parameter decoding unit 37 are the same as the plurality of parameters ⁇ stored in the parameter determining unit 27 of the encoding device.
  • the linear prediction coefficient decoding unit 31 receives the linear prediction coefficient code output from the encoding device and the decoding parameter ⁇ obtained by the parameter decoding unit 37.
  • the linear prediction coefficient decoding unit 31 is the linear prediction decoding device described above with reference to FIGS. 6 and 21 described in [Linear prediction encoding device, linear prediction decoding device and their methods]. [Encoding device, decoding device and their methods] and FIG. 13 show the linear prediction encoding device of FIGS. 6 and 21 described in [Linear prediction encoding device, linear prediction decoding device and these methods]. This is expressed as “linear prediction coefficient decoding unit 31”.
  • the linear prediction coefficient decoding unit 31 may be the linear prediction decoding device in FIG.
  • the linear prediction coefficient decoding unit 31 uses the same linear prediction coefficient code as the process described in [Linear prediction encoding apparatus, linear prediction decoding apparatus and their methods] with the decoding parameter ⁇ as the parameter ⁇ 1. , Decoded linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p that are coefficients that can be converted into decoded linear prediction coefficients are obtained (step B1).
  • the obtained decoded linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p are output to the non-smoothed amplitude spectrum envelope sequence generation unit 32 and the non-smoothed amplitude spectrum envelope sequence generation unit 33.
  • the unsmoothed amplitude spectrum envelope sequence generation unit 32 includes the decoding parameter ⁇ obtained by the parameter decoding unit 37 and the decoded linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,. Is entered.
  • Textured amplitude spectral envelope sequence generating unit 32 decodes the linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., ⁇ ⁇ unsmoothed amplitude spectrum is a series of amplitude spectrum envelope corresponding to p envelope sequence ⁇ H (0 ), ⁇ H (1),..., ⁇ H (N-1) are generated by the above equation (A2) (step B2).
  • the generated non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1) is output to the decoding unit 34.
  • the unsmoothed amplitude spectrum envelope sequence generation unit 32 converts the amplitude spectrum envelope sequence corresponding to the coefficient that can be converted into the linear prediction coefficient generated by the linear prediction coefficient decoding unit 31 to 1 / ⁇ .
  • a non-smoothed spectral envelope sequence which is a raised sequence is obtained.
  • the smoothed amplitude spectrum envelope sequence generation unit 33 receives the decoding parameter ⁇ obtained by the parameter decoding unit 37 and the decoded linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p obtained by the linear prediction coefficient decoding unit 31. Entered.
  • Smoothing the amplitude spectral envelope sequence generating unit 33 decodes the linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., smoothing the amplitude is a sequence blunted amplitude of irregularities of the amplitude spectral envelope of the sequence corresponding to the ⁇ beta p spectral envelope sequence ⁇ H ⁇ (0), ⁇ H ⁇ (1), ..., ⁇ H ⁇ a (N-1) produced by the equation a (3) above (step B3).
  • the generated smoothed amplitude spectrum envelope sequences ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N-1) are output to the decoding unit 34 and the envelope denormalization unit 35.
  • the decoding unit 34 includes a decoding parameter ⁇ obtained by the parameter decoding unit 37, a code corresponding to the normalized MDCT coefficient sequence output by the encoding device, and a non-smoothed amplitude spectrum generated by the non-smoothed amplitude spectrum envelope generating unit 32.
  • Envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1) and smoothed amplitude spectrum envelope sequence generated by the smoothed amplitude spectrum envelope generator 33 ⁇ H ⁇ (0), ⁇ H ⁇ (1), ..., ⁇ H ⁇ (N-1) is input.
  • the decryption unit 34 includes a dispersion parameter determination unit 342.
  • the decoding unit 34 performs decoding by performing, for example, the processing from step B41 to step B44 shown in FIG. 15 (step B4). That is, the decoding unit 34 decodes the gain code included in the code corresponding to the input normalized MDCT coefficient sequence for each frame to obtain the global gain g (step B41).
  • the dispersion parameter determination unit 342 of the decoding unit 34 includes a global gain g, a non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1) and a smoothed amplitude spectrum envelope sequence.
  • Step B42 The decoding unit 34 converts the integer signal code included in the code corresponding to the normalized MDCT coefficient sequence to arithmetic corresponding to each dispersion parameter of the dispersion parameter sequence ⁇ (0), ⁇ (1),..., ⁇ (N ⁇ 1).
  • arithmetic decoding is performed to obtain decoded normalized coefficient series ⁇ X Q (0), ⁇ X Q (1), ..., ⁇ X Q (N-1) (step B43), and decoding normalized Coefficient sequence ⁇ X Q (0), ⁇ X Q (1), ..., ⁇ X Q (N-1) is multiplied by global gain g and decoded normalized MDCT coefficient sequence ⁇ X N (0), ⁇ X N (1),..., ⁇ X N (N-1) are generated (step B44).
  • the decoding unit 34 may perform decoding of the input integer signal code according to bit allocation that substantially changes based on the non-smoothed spectrum envelope sequence.
  • the decoding unit 34 When encoding is performed by the process described in [Modification of Encoding Unit 26], the decoding unit 34 performs, for example, the following process.
  • the decoding unit 34 decodes the gain code included in the code corresponding to the input normalized MDCT coefficient sequence for each frame to obtain the global gain g.
  • the dispersion parameter determination unit 342 of the decoding unit 34 includes a non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1),...
  • the decoding unit 34 obtains b (k) by Expression (A10) based on each dispersion parameter ⁇ (k) of the dispersion parameter series ⁇ (0), ⁇ (1),..., ⁇ (N ⁇ 1).
  • XQ (k) can be sequentially decoded with the number of bits b (k) and the normalized normalized coefficient sequence ⁇ X Q (0), ⁇ X Q (1),..., ⁇ X Q (N -1) is obtained, and the coefficients of the decoded normalized coefficient series ⁇ X Q (0), ⁇ X Q (1), ..., ⁇ X Q (N-1) are multiplied by the global gain g to obtain the decoding normal MDCT coefficient sequence ⁇ X N (0), ⁇ X N (1), ..., ⁇ X N (N-1) is generated.
  • the decoding unit 34 may perform decoding of the input integer signal code in accordance with bit allocation that changes based on the non-smoothed spectrum envelope sequence.
  • the generated decoded normalized MDCT coefficient sequence ⁇ X N (0), ⁇ X N (1),..., ⁇ X N (N ⁇ 1) is output to the envelope denormalization unit 35.
  • the envelope denormalization unit 35 includes a smoothed amplitude spectrum envelope sequence ⁇ H ⁇ (0), ⁇ H ⁇ (1), ..., ⁇ H ⁇ (N-1) generated by the smoothed amplitude spectrum envelope generation unit 33.
  • the decoding normalization MDCT coefficient sequence ⁇ X N (0), ⁇ X N (1),..., ⁇ X N (N-1) generated by the decoding unit 34 is input.
  • the envelope denormalization unit 35 uses the smoothed amplitude spectrum envelope sequence ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N-1) to decode the normalized MDCT coefficient sequence ⁇ X
  • N (0), ⁇ X N (1), ..., ⁇ X N (N-1) the decoded MDCT coefficient sequence ⁇ X (0), ⁇ X (1), ..., ⁇ X (N-1) is generated (step B5).
  • the generated decoded MDCT coefficient sequence ⁇ X (0), ⁇ X (1), ..., ⁇ X (N-1) is output to the time domain conversion unit 36.
  • the envelope inverse normalization unit 35, k 0, 1, ..., a N-1, decoding the normalized MDCT coefficients ⁇ X N (0), ⁇ X N (1), ..., ⁇ X N (N -1) for each coefficient ⁇ X N (k), the smoothed amplitude spectrum envelope series ⁇ H ⁇ (0), ⁇ H ⁇ (1),..., ⁇ H ⁇ (N-1) envelope values ⁇ H
  • the time domain transform unit 36 receives the decoded MDCT coefficient sequence ⁇ X (0), ⁇ X (1),..., ⁇ X (N-1) generated by the envelope denormalization unit 35.
  • the time domain transform unit 36 transforms the decoded MDCT coefficient sequence ⁇ X (0), ⁇ X (1), ..., ⁇ X (N-1) obtained by the envelope denormalization unit 35 into the time domain for each frame.
  • a sound signal (decoded sound signal) in units of frames is obtained (step B6).
  • the decoding device obtains a time-series signal by decoding in the frequency domain.
  • the encoding apparatus and method according to the first embodiment generate a code by performing encoding for each of a plurality of parameters ⁇ , select an optimal code from the codes generated for each parameter ⁇ , and select the selected code. And a parameter code corresponding to the selected code.
  • the parameter determination unit 27 first determines the parameter ⁇ , performs encoding based on the determined parameter ⁇ , generates a code, and outputs it. .
  • the parameter ⁇ is made variable by the parameter determination unit 27 for each predetermined time interval.
  • the parameter ⁇ being variable for each predetermined time interval means that the parameter ⁇ can be changed if the predetermined time interval is changed, and the value of the parameter ⁇ is not changed in the same time interval.
  • the encoding device includes a frequency domain transform unit 21, a linear prediction analysis unit 22, a non-smoothed amplitude spectrum envelope sequence generation unit 23, a smoothed amplitude spectrum envelope sequence generation unit 24, and an envelope
  • a normalization unit 25 an encoding unit 26, and a parameter determination unit 27 ′ are provided.
  • An example of each process of the encoding method realized by this encoding apparatus is shown in FIG.
  • a time domain sound signal which is a time-series signal, is input to the parameter determination unit 27 ′.
  • sound signals are voice digital signals or acoustic digital signals.
  • the parameter determining unit 27 ′ determines the parameter ⁇ by a process described later based on the input time series signal (step A7 ′).
  • the parameter ⁇ determined by the parameter determination unit 27 ′ is referred to as parameter ⁇ 1 .
  • ⁇ 1 determined by the parameter determination unit 27 ′ is output to the linear prediction analysis unit 22, the non-smoothed amplitude spectrum envelope estimation unit 23, the smoothed amplitude spectrum envelope estimation unit 24, and the encoding unit 26.
  • the parameter determination unit 27 ′ generates a parameter code by encoding the determined ⁇ 1 .
  • the generated parameter code is transmitted to the decoding device.
  • the frequency domain transform unit 21, the linear prediction analysis unit 22, the unsmoothed amplitude spectrum envelope sequence generation unit 23, the smoothed amplitude spectrum envelope sequence generation unit 24, the envelope normalization unit 25, and the encoding unit 26 include a parameter determination unit 27.
  • a code is generated by the same processing as in the first embodiment (step A1 to step A6).
  • the code is a combination of a linear prediction coefficient code, a gain code, and an integer signal code.
  • the generated code is transmitted to the decoding device.
  • FIG. 18 shows a configuration example of the parameter determination unit 27 '.
  • the parameter determination unit 27 ′ includes, for example, a frequency domain conversion unit 41, a spectrum envelope estimation unit 42, a whitened spectrum sequence generation unit 43, and a parameter acquisition unit 44.
  • the spectrum envelope estimation unit 42 includes, for example, a linear prediction analysis unit 421 and a non-smoothed amplitude spectrum envelope sequence generation unit 422.
  • FIG. 19 shows an example of each process of the parameter determination method realized by the parameter determination unit 27 '.
  • the time domain sound signal which is a time series signal, is input to the frequency domain transform unit 41.
  • Examples of sound signals are voice digital signals or acoustic digital signals.
  • the frequency domain conversion unit 41 converts the input time domain sound signal into N frequency MDCT coefficient sequences X (0), X (1),..., X (N ⁇ Convert to 1). N is a positive integer.
  • the obtained MDCT coefficient sequences X (0), X (1),..., X (N-1) are output to the spectrum envelope estimation unit 42 and the whitened spectrum sequence generation unit 43.
  • the subsequent processing is performed in units of frames.
  • the frequency domain conversion unit 41 obtains a frequency domain sample sequence corresponding to the sound signal, for example, an MDCT coefficient sequence (step C41).
  • the spectrum envelope estimation unit 42 receives the MDCT coefficient sequence X (0), X (1),..., X (N ⁇ 1) obtained by the frequency domain conversion unit 21.
  • the spectrum envelope estimation unit 42 Based on the parameter ⁇ 0 determined by a predetermined method, the spectrum envelope estimation unit 42 performs spectrum envelope estimation using the absolute value ⁇ 0 of the frequency domain sample sequence corresponding to the time-series signal as a power spectrum ( Step C42).
  • the estimated spectrum envelope is output to the whitened spectrum sequence generation unit 43.
  • the spectrum envelope estimation unit 42 estimates the spectrum envelope by generating a non-smoothed amplitude spectrum envelope sequence, for example, by processing of a linear prediction analysis unit 421 and a non-smoothed amplitude spectrum envelope sequence generation unit 422 described below. .
  • the parameter ⁇ 0 is determined by a predetermined method.
  • ⁇ 0 is a predetermined number greater than zero.
  • ⁇ 0 1.
  • the frame before the frame for which the current parameter ⁇ is to be obtained (hereinafter referred to as the current frame) is, for example, a frame before the current frame and in the vicinity of the current frame.
  • the frame in the vicinity of the current frame is, for example, a frame immediately before the current frame.
  • ⁇ Linear prediction analysis unit 421 MDCT coefficient sequences X (0), X (1),..., X (N ⁇ 1) obtained by the frequency domain transform unit 41 are input to the linear prediction analysis unit 421.
  • the linear prediction analysis unit 421 uses the MDCT coefficient sequence X (0), X (1),..., X (N-1) to define ⁇ R (0), ⁇ R defined by the following equation (C1). (1),..., ⁇ R (N-1) are used to generate linear prediction coefficients ⁇ 1 , ⁇ 2 ,..., ⁇ p subjected to linear prediction analysis, and the generated linear prediction coefficients ⁇ 1 , ⁇ 2 , ..., ⁇ p are encoded and linear prediction coefficient codes and quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p , which are quantized linear prediction coefficients corresponding to the linear prediction coefficient codes, are obtained. Generate.
  • the generated quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p are output to the non-smoothed spectrum envelope sequence generation unit 422.
  • the linear prediction analyzer 421 first MDCT coefficients X (0), X (1 ), ..., X (N-1) of the inverse Fourier that the eta 0 squared regarded as a power spectrum of the absolute value
  • the linear prediction analysis unit 421 performs linear prediction analysis using the obtained pseudo correlation function signal sequence ⁇ R (0), ⁇ R (1), ..., ⁇ R (N-1) to obtain a linear prediction coefficient. ⁇ 1 , ⁇ 2 ,..., ⁇ p are generated. Then, the linear prediction analysis unit 421 encodes the generated linear prediction coefficients ⁇ 1 , ⁇ 2 ,..., ⁇ p so as to encode a linear prediction coefficient code and a quantized linear prediction coefficient corresponding to the linear prediction coefficient code. ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p are obtained.
  • Linear prediction coefficients ⁇ 1, ⁇ 2, ..., ⁇ p is, MDCT coefficient sequence X (0), X (1 ), ..., and the eta 0 square of the absolute value of X (N-1) was regarded as a power spectrum It is a linear prediction coefficient corresponding to the time domain signal.
  • the generation of the linear prediction coefficient code by the linear prediction analysis unit 421 is performed by, for example, a conventional encoding technique.
  • the conventional encoding technique is, for example, an encoding technique in which a code corresponding to the linear prediction coefficient itself is a linear prediction coefficient code, and a code corresponding to the LSP parameter by converting the linear prediction coefficient into an LSP parameter.
  • an encoding technique for converting a linear prediction coefficient into a PARCOR coefficient and a code corresponding to the PARCOR coefficient as a linear prediction coefficient code for example, an encoding technique for converting a linear prediction coefficient into a PARCOR coefficient and a code corresponding to the PARCOR coefficient as a linear prediction coefficient code.
  • the linear prediction analysis unit 42 for example, a pseudo correlation function signal sequence obtained by performing an inverse Fourier transform in which the absolute value ⁇ 0 of the frequency domain sample sequence that is an MDCT coefficient sequence is regarded as a power spectrum. Is used to generate a coefficient that can be converted into a linear prediction coefficient (step C421).
  • the linear prediction analysis unit 421 obtains a linear prediction coefficient code by the method described in the section of [Linear prediction encoding apparatus, linear prediction decoding apparatus and their methods], and corresponds to the obtained linear prediction coefficient code.
  • Coefficients that can be converted into linear prediction coefficients to be used may be quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p .
  • ⁇ Non-smoothed Amplitude Spectrum Envelope Sequence Generation Unit 422 Quantized linear prediction coefficients ⁇ ⁇ 1 , ⁇ ⁇ 2 ,..., ⁇ ⁇ p generated by the linear prediction analysis unit 421 are input to the unsmoothed amplitude spectrum envelope sequence generation unit 422.
  • Textured amplitude spectral envelope sequence generation unit 422 the quantized linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., ⁇ ⁇ is the sequence of the amplitude spectrum envelope corresponding to p textured amplitude spectral envelope sequence ⁇ H ( 0), ⁇ H (1), ..., ⁇ H (N-1) are generated.
  • the generated non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1) is output to the whitened spectrum sequence generation unit 43.
  • Textured amplitude spectral envelope sequence generation unit 422 the quantized linear prediction coefficient ⁇ ⁇ 1, ⁇ ⁇ 2, ..., using the ⁇ beta p, unsmoothed amplitude spectral envelope sequence ⁇ H (0), ⁇ H ( 1),..., ⁇ H (N-1) as unsmoothed amplitude spectrum envelope sequence defined by equation (C2) ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1) Is generated.
  • the unsmoothed amplitude spectrum envelope sequence generation unit 422 performs linear prediction analysis on the unsmoothed spectrum envelope sequence that is a sequence obtained by raising the amplitude spectrum envelope sequence corresponding to the pseudo correlation function signal sequence to the 1 / ⁇ 0 power.
  • the spectral envelope is estimated by obtaining the coefficient based on the coefficient that can be converted into the linear prediction coefficient generated by the unit 421 (step C422).
  • the whitened spectrum sequence generation unit 43 includes an MDCT coefficient sequence X (0), X (1),..., X (N-1) obtained by the frequency domain conversion unit 41 and a non-smoothed amplitude spectrum envelope generation unit 422.
  • the generated non-smoothed amplitude spectrum envelope sequence ⁇ H (0), ⁇ H (1), ..., ⁇ H (N-1) is input.
  • the whitened spectrum sequence generation unit 43 converts each coefficient of the MDCT coefficient sequence X (0), X (1),..., X (N-1) into a corresponding non-smoothed amplitude spectrum envelope sequence ⁇ H (0), By dividing each value of ⁇ H (1), ..., ⁇ H (N-1), the whitened spectrum series X W (0), X W (1), ..., X W (N-1) Generate.
  • the generated whitening spectrum series X W (0), X W (1),..., X W (N ⁇ 1) are output to the parameter acquisition unit 44.
  • k the coefficients X (()) of the MDCT coefficient sequence X (0), X (1),.
  • k the coefficients X (()) of the MDCT coefficient sequence X (0), X (1),.
  • ⁇ H (0), ⁇ H (1),..., ⁇ H (N-1) values ⁇ H (k) the whitened spectrum sequence X
  • the whitened spectrum sequence generation unit 43 obtains a whitened spectrum sequence that is a sequence obtained by dividing a frequency domain sample sequence that is an MDCT coefficient sequence, for example, by a spectrum envelope that is an unsmoothed amplitude spectrum envelope sequence, for example ( Step C43).
  • the parameter acquisition unit 44 receives the whitened spectrum series X W (0), X W (1),..., X W (N ⁇ 1) generated by the whitened spectrum series generating unit 43.
  • the parameter acquisition unit 44 approximates the histogram of the whitened spectrum series X W (0), X W (1),..., X W (N ⁇ 1) with the generalized Gaussian distribution having the parameter ⁇ as a shape parameter. Is obtained (step C44).
  • the parameter acquisition unit 44 is a distribution of histograms in which the generalized Gaussian distribution having the parameter ⁇ as a shape parameter is a whitened spectrum series X W (0), X W (1), ..., X W (N-1).
  • the parameter ⁇ that is close to is determined.
  • the generalized Gaussian distribution with the parameter ⁇ as a shape parameter is defined as follows, for example.
  • is a gamma function.
  • is a predetermined number greater than zero.
  • may be a predetermined number other than 2 that is greater than 0.
  • may be a predetermined positive number less than 2.
  • is a parameter corresponding to the variance.
  • ⁇ obtained by the parameter acquisition unit 44 is defined by the following equation (C3), for example.
  • F ⁇ 1 is an inverse function of the function F. This equation is derived by the so-called moment method.
  • the parameter acquisition unit 44 inputs the value of m 1 / ((m 2 ) 1/2 ) into the formulated inverse function F ⁇ 1 .
  • the parameter ⁇ can be obtained by calculating the output value.
  • the parameter acquisition unit 44 calculates, for example, the first method or the second method described below in order to calculate the value of ⁇ defined by the equation (C3).
  • the parameter ⁇ may be obtained by
  • a first method for obtaining the parameter ⁇ will be described.
  • the parameter obtaining unit 44 based on the whitened spectrum sequence to calculate the m 1 / ((m 2) 1/2), a plurality of different which had been prepared beforehand, corresponding to the eta F ⁇ corresponding to F ( ⁇ ) closest to the calculated m 1 / ((m 2 ) 1/2 ) is obtained with reference to the pair of ( ⁇ ).
  • a plurality of different pairs of F ( ⁇ ) corresponding to ⁇ prepared in advance are stored in advance in the storage unit 441 of the parameter acquisition unit 44.
  • the parameter acquisition unit 44 refers to the storage unit 441, finds F ( ⁇ ) closest to the calculated m 1 / ((m 2 ) 1/2 ), and stores ⁇ corresponding to the found F ( ⁇ ). Read from the unit 441 and output.
  • the approximate curve function of the inverse function F ⁇ 1 is set as, for example, ⁇ F ⁇ 1 represented by the following formula (C3 ′), and the parameter acquisition unit 44 uses m 1 / ((m 2 ) 1/2 ) is calculated, and ⁇ is calculated by calculating the output value when m 1 / ((m 2 ) 1/2 ) calculated in the approximate curve function ⁇ F -1 is input.
  • the approximate curve function ⁇ F -1 may be a monotonically increasing function whose output is a positive value in the domain to be used.
  • ⁇ obtained by the parameter acquisition unit 44 is not an expression (C3) but an expression (C3) using positive integers q1 and q2 determined in advance as in an expression (C3 ′′) (where q1 ⁇ q2). It may be defined by a generalized formula.
  • can be obtained by the same method as that when ⁇ is defined by equation (C3). That is, the parameter acquisition unit 44 calculates a value m q1 / ((m q2 ) q1 / q2 ) based on the q 1st moment m q1 and the q 2nd moment m q2 based on the whitened spectrum series. Then, for example, as in the first and second methods described above, the calculated m q1 / ((() by referring to a plurality of different pairs of F ′ ( ⁇ ) corresponding to ⁇ prepared in advance.
  • is a value based on two different moments m q1 and m q2 having different dimensions.
  • the value of the moment with the lower dimension or a value based on this (hereinafter referred to as the former) and the value of the moment with the higher dimension or ⁇ may be obtained based on the value of the ratio based on the value (hereinafter referred to as the latter), the value based on the value of this ratio, or the value obtained by dividing the former by the latter.
  • the value based on the moment for example, is that the m Q a Q to the moment and m as a given real number.
  • may be obtained by inputting these values into the approximate curve function ⁇ F- 1 .
  • the approximate curve function to F ′ ⁇ 1 may be a monotonically increasing function whose output is a positive value in the domain to be used, as described above.
  • the parameter determination unit 27 ′ may obtain the parameter ⁇ by loop processing. That is, the parameter determination unit 27 ′ sets the parameter ⁇ obtained by the parameter acquisition unit 44 as the parameter ⁇ 0 determined by a predetermined method, and performs processing by the spectrum envelope estimation unit 42, the whitened spectrum sequence generation unit 43, and the parameter acquisition unit 44. May be performed once more.
  • the parameter ⁇ obtained by the parameter acquisition unit 44 is output to the spectrum envelope estimation unit 42.
  • the spectrum envelope estimation unit 42 estimates the spectrum envelope by performing the same process as described above using ⁇ obtained by the parameter acquisition unit 44 as the parameter ⁇ 0 .
  • the whitened spectrum sequence generation unit 43 Based on the newly estimated spectrum envelope, the whitened spectrum sequence generation unit 43 generates a whitened spectrum sequence by performing the same process as described above.
  • the parameter acquisition unit 44 performs a process similar to the process described above based on the newly generated whitened spectrum sequence to obtain the parameter ⁇ .
  • the processing of the spectrum envelope estimation unit 42, the whitened spectrum series generation unit 43, and the parameter acquisition unit 44 may be further performed a predetermined number of times ⁇ .
  • the spectrum envelope estimation unit 42 performs the spectrum envelope estimation unit 42, the whitened spectrum sequence generation unit 43, and the parameter until the absolute value of the difference between the parameter ⁇ obtained this time and the parameter ⁇ obtained last time is equal to or less than a predetermined threshold. You may repeat the process of the acquisition part 44. FIG.
  • the spectrum envelope estimation unit 2A is a frequency domain that is, for example, an MDCT coefficient sequence corresponding to a time series signal. the eta 1 square of the absolute value of the sample sequence it can be said that we estimated spectral envelope that is regarded as a power spectrum (unsmoothed amplitude spectral envelope sequence).
  • “considered as a power spectrum” means that a spectrum of ⁇ 1 is used where a power spectrum is normally used.
  • the linear prediction analysis unit 22 of the spectrum envelope estimation unit 2A performs, for example, a pseudo Fourier transform obtained by performing an inverse Fourier transform in which the absolute value ⁇ 1 of the frequency domain sample sequence that is an MDCT coefficient sequence is regarded as a power spectrum. It can be said that a coefficient that can be converted into a linear prediction coefficient is obtained by performing a linear prediction analysis using the correlation function signal sequence.
  • the non-smoothed amplitude spectrum envelope sequence generation unit 23 of the spectrum envelope estimation unit 2A converts the amplitude spectrum envelope sequence corresponding to the coefficient that can be converted into the linear prediction coefficient obtained by the linear prediction analysis unit 22 to 1 / ⁇ 1. It can be said that the spectrum envelope is estimated by obtaining a non-smoothed spectrum envelope sequence which is a raised sequence.
  • the encoding unit 2B is a spectrum estimated by the spectrum envelope estimation unit 2A. Coding for changing the bit allocation based on the envelope (non-smoothed amplitude spectrum envelope sequence) or changing the bit allocation substantially for each coefficient of the frequency domain sample sequence corresponding to the time-series signal, for example, MDCT coefficient sequence It can be said that it is going.
  • the decoding unit 3A is input according to a bit allocation that changes based on a non-smoothed spectrum envelope sequence or a bit allocation that changes substantially. It can be said that the frequency domain sample sequence corresponding to the time-series signal is obtained by decoding the integer signal code.
  • the encoding unit 2B may perform encoding other than the arithmetic encoding described above if the bit allocation is changed based on the spectral envelope (unsmoothed amplitude spectral envelope sequence) or the bit allocation is changed substantially. Processing may be performed.
  • the decoding unit 3A performs a decoding process corresponding to the encoding process performed by the encoding unit 2B.
  • the encoding unit 2B may perform Golomb-Rice encoding on the frequency domain sample sequence using the Rice parameter determined based on the spectrum envelope (unsmoothed amplitude spectrum envelope sequence).
  • the decoding unit 3A may perform Golomb-Rice decoding using the Rice parameter determined based on the spectrum envelope (unsmoothed amplitude spectrum envelope sequence).
  • the encoding device may not perform the encoding process to the end when determining the parameter ⁇ .
  • the parameter determination unit 27 may determine the parameter ⁇ based on the estimated code amount.
  • the encoding unit 2B uses each of the plurality of parameters ⁇ to estimate the code obtained by the same encoding process as described above for the frequency domain sample sequence corresponding to the time-series signal in the same predetermined time interval. Get quantity.
  • the parameter determination unit 27 selects one of a plurality of parameters ⁇ based on the obtained estimated code amount. For example, the parameter ⁇ having the smallest estimated code amount is selected.
  • the encoding unit 2B obtains and outputs a code by performing the same encoding process as described above using the selected parameter ⁇ .
  • the processing described above is not only executed in time series in the order described, but may also be executed in parallel or individually as required by the processing capability of the apparatus that executes the processing.
  • the program describing the processing contents can be recorded on a computer-readable recording medium.
  • a computer-readable recording medium for example, any recording medium such as a magnetic recording device, an optical disk, a magneto-optical recording medium, and a semiconductor memory may be used.
  • this program is distributed by selling, transferring, or lending a portable recording medium such as a DVD or CD-ROM in which the program is recorded. Further, the program may be distributed by storing the program in a storage device of the server computer and transferring the program from the server computer to another computer via a network.
  • a computer that executes such a program first stores a program recorded on a portable recording medium or a program transferred from a server computer in its storage unit. When executing the process, this computer reads the program stored in its own storage unit and executes the process according to the read program.
  • a computer may read a program directly from a portable recording medium and execute processing according to the program. Further, each time a program is transferred from the server computer to the computer, processing according to the received program may be executed sequentially.
  • the program is not transferred from the server computer to the computer, and the above-described processing is executed by a so-called ASP (Application Service Provider) type service that realizes a processing function only by an execution instruction and result acquisition. It is good.
  • the program includes information provided for processing by the electronic computer and equivalent to the program (data that is not a direct command to the computer but has a property that defines the processing of the computer).
  • each device is configured by executing a predetermined program on a computer, at least a part of these processing contents may be realized by hardware.

Landscapes

  • Engineering & Computer Science (AREA)
  • Physics & Mathematics (AREA)
  • Computational Linguistics (AREA)
  • Signal Processing (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Spectroscopy & Molecular Physics (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)

Abstract

 線形予測符号化装置は、時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析部221と、符号帳記憶部222に記憶された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、線形予測分析部221が得た線形予測係数に変換可能な係数と、のηの値を適合させる適合部22Aと、ηの値が適合された線形予測係数に変換可能な係数の複数個の候補と線形予測係数に変換可能な係数とを用いて、線形予測分析部221が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る符号化部224と、を備えている。

Description

線形予測符号化装置、線形予測復号装置、これらの方法、プログラム及び記録媒体
 この発明は、線形予測係数に変換可能な係数を符号化又は復号する技術に関する。
 線形予測係数に変換可能な係数の1つであるLSPパラメータの量子化技術として、ベクトル量子化等の手法が知られている(例えば、非特許文献1参照)。
 ところで、公知とはなっていないが、発明者によりパラメータηが提案されている。このパラメータηは、例えば3GPP EVS(Enhanced Voice Services)規格で使われているような線形予測包絡を利用する周波数領域の係数の量子化値を算術符号化する符号化方式において、算術符号の符号化対象の属する確率分布を定める形状パラメータである。パラメータηは、符号化対象の分布と関連性を有しており、パラメータηを適宜定めると効率の良い符号化及び復号を行うことが可能である。
 また、パラメータηは、時系列信号の特徴を表す指標と成り得る。このため、パラメータηを適宜用いると、LSPパラメータ等の線形予測係数に変換可能な係数を効率の良く符号化及び復号を行うことが可能である。
守谷健弘,「高圧縮音声符号化の必須技術:線スペクトル対(LSP)」,NTT技術ジャーナル,2014年9月,P.58-60
 しかしながら、パラメータηを用いた線形予測係数に変換可能な係数の符号化及び復号技術は知られていなかった。
 本発明は、パラメータηを用いて線形予測係数に変換可能な係数の符号化又は復号を行う線形予測符号化装置、線形予測復号装置、これらの方法、プログラム及び記録媒体を提供することを目的とする。
 本発明の一態様による線形予測符号化装置によれば、パラメータηを正の数として、時系列信号に対応するパラメータηを、その時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做すことにより推定されたスペクトル包絡で周波数領域サンプル列を除算した系列である白色化スペクトル系列のヒストグラムを近似する一般化ガウス分布の形状パラメータとし、η1はパラメータηの所定の値であるとして、時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析部と、N種類(Nは1以上の整数)のパラメータηのそれぞれに対応するN個の符号帳が記憶され、各符号帳にはそれぞれのパラメータηに対応する線形予測係数に変換可能な係数の候補が複数個格納された符号帳記憶部と、符号帳記憶部に記憶された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、線形予測分析部が得た線形予測係数に変換可能な係数と、のηの値を適合させる適合部と、ηの値が適合された線形予測係数に変換可能な係数の複数個の候補と線形予測係数に変換可能な係数とを用いて、線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る符号化部と、を備えている。
 本発明の一態様による線形予測符号化装置によれば、パラメータηを正の数として、時系列信号に対応するパラメータηを、その時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做すことにより推定されたスペクトル包絡で周波数領域サンプル列を除算した系列である白色化スペクトル系列のヒストグラムを近似する一般化ガウス分布の形状パラメータとし、η1はパラメータηの所定の値であるとして、時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析部と、符号帳が記憶された符号帳記憶部と、入力されたη1に基づいて、符号帳記憶部に記憶された符号帳と線形予測係数に変換可能な係数との少なくとも一方を適合させる適合部と、符号帳又は適合された符号帳を用いて、線形予測係数に変換可能な係数又は適合された線形予測係数に変換可能な係数を符号化する符号化部と、を備えている。
 本発明の一態様による線形予測復号装置によれば、符号帳が記憶された符号帳記憶部と、η1を正の数として、入力されたη1に基づいて、符号帳記憶部に記憶された符号帳と、符号帳に格納された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補との少なくとも一方を適合させる適合部を備えており、線形予測係数に変換可能な係数は、線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列を得るために用いられる。
 パラメータηを用いて線形予測係数に変換可能な係数の符号化又は復号を行うことができる。
線形予測符号化装置の例を説明するためのブロック図。 線形予測符号化装置の例を説明するためのブロック図。 線形予測符号化装置の例を説明するためのブロック図。 線形予測符号化方法の例を説明するためのフローチャート。 LSPパラメータとηとの関係の例を説明するための図。 線形予測復号装置の例を説明するためのブロック図。 線形予測復号方法の例を説明するためのフローチャート。 符号化装置の例を説明するためのブロック図。 符号化方法の例を説明するためのフローチャート。 符号化部の例を説明するためのブロック図。 符号化部の例を説明するためのブロック図。 符号化部の処理の例を説明するためのフローチャート。 復号装置の例を説明するためのブロック図。 復号方法の例を説明するためのフローチャート。 復号部の処理の例を説明するためのフローチャート。 符号化装置の例を説明するためのブロック図。 符号化方法の例を説明するためのフローチャート。 パラメータ決定装置の例を説明するためのブロック図。 パラメータ決定方法の例を説明するためのフローチャート。 一般化ガウス分布を説明するための図。 線形予測符号化装置の例を説明するためのブロック図。 線形予測符号化方法の例を説明するためのフローチャート。 線形予測復号装置の例を説明するためのブロック図。 線形予測復号方法の例を説明するためのフローチャート。 線形予測符号化装置の例を説明するためのブロック図。 線形予測符号化装置の例を説明するためのブロック図。 線形予測符号化装置の例を説明するためのブロック図。 線形予測復号装置の例を説明するためのブロック図。
 [線形予測符号化装置、線形予測復号装置及びこれらの方法]
 以下、線形予測符号化装置、線形予測復号装置及びこれらの方法を用いた符号化装置、復号装置及びこれらの方法の例について説明する。
 [線形予測符号化装置、線形予測復号装置及びこれらの方法の第一実施形態]
 (符号化)
 第一実施形態の線形予測符号化装置及び方法の一例について説明する。
 第一実施形態の線形予測符号化装置は、図1、図2又は図3に示すように、線形予測分析部221、符号帳記憶部222、符号化部224及び線形変換部225を例えば備えている。図1、図2又は図3の例では線形予測符号化装置の外部に周波数領域変換部220が設けられているが、線形予測符号化装置が周波数領域変換部220を更に備えていてもよい。線形予測符号化装置の各部が、図4に例示する各処理を行うことにより線形予測符号化方法が実現される。
 <周波数領域変換部220>
 周波数領域変換部220には、時系列信号である時間領域の音信号が入力される。
 周波数領域変換部41は、所定の時間長のフレーム単位で、入力された時間領域の音信号を周波数領域のN点のMDCT係数列X(0),X(1),…,X(N-1)に変換する。Nは正の整数である。
 得られたMDCT係数列X(0),X(1),…,X(N-1)は、線形予測分析部221に出力される。
 特に断りがない限り、以降の処理はフレーム単位で行われるものとする。
 このようにして、周波数領域変換部220は、時系列信号に対応する、例えばMDCT係数列である周波数領域サンプル列を求める。
 <線形予測分析部221>
 線形予測分析部221には、例えばMDCT係数列X(0),X(1),…,X(N-1)である周波数領域サンプル列及びその周波数領域サンプル列に対応するパラメータη1が入力される。
 パラメータη1は、正の数である。パラメータη1は、例えば、後述するパラメータ決定部27,27’により決定される。パラメータη1は、例えば3GPP EVS(Enhanced Voice Services)規格で使われているような線形予測包絡を利用する周波数領域の係数の量子化値を算術符号化する符号化方式において、算術符号の符号化対象の属する確率分布を定めるパラメータηである。パラメータηは、時系列信号の特徴を表す指標と成り得るものである。後に出てくるパラメータη23も、パラメータηである。η123は、パラメータηの所定の値とも言える。
 なお、パラメータη1についての情報は、線形予測復号装置に送信されるとする。例えば、パラメータη1を表すパラメータ符号が線形予測復号装置に送信される。
 線形予測分析部221は、MDCT係数列X(0),X(1),…,X(N-1)及びη1を用いて、以下の式(A7)により定義される~R(0),~R(1),…,~R(N-1)を用いて線形予測分析を行って線形予測係数係数に変換可能な係数を生成する(ステップDE1)。
Figure JPOXMLDOC01-appb-M000003
 生成された線形予測係数係数に変換可能な係数は、符号化部224に出力される。
 具体的には、線形予測分析部22は、まずMDCT係数列X(0),X(1),…,X(N-1)の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換に相当する演算、すなわち式(A7)の演算を行うことにより、MDCT係数列X(0),X(1),…,X(N-1)の絶対値のη1乗に対応する時間領域の信号列である擬似相関関数信号列~R(0),~R(1),…,~R(N-1)を求める。そして、線形予測分析部22は、求まった擬似相関関数信号列~R(0),~R(1),…,~R(N-1)を用いて線形予測分析を行って、線形予測係数係数に変換可能な係数を生成する。
 このようにして、線形予測分析部221は、η1を正の数として、時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る。
 線形予測係数に変換可能な係数とは、例えばLSP、PARCOR係数、ISP等である。線形予測係数に変換可能な係数は、線形予測係数自体であってもよい。
 pを所定の正の数とし、線形予測係数に可能な係数の次数をp次とする。
 <符号帳記憶部222>
 符号帳記憶部222には、パラメータη2に対応する線形予測係数に変換可能な係数の候補が複数個格納された符号帳が記憶されている。
 以下、線形予測係数に変換可能な係数の候補と、その線形予測係数に変換可能な係数の候補に対応する符号とのペアを、候補符号ペアと呼ぶことにする。符号帳には、複数個の候補符号ペアが記憶されている。言い換えると、Nを所定の2以上の数とすると、符号帳には、N個の候補ペアが記憶されている。線形予測係数に変換可能な係数の候補に対応する符号のそれぞれには、所定の数のビットが割り当てられている。各符号は、割り当てられた所定の数のビットで表現される。
 線形予測係数に変換可能な係数の次数がpであるため、線形予測係数に変換可能な係数の各候補はp個の値から構成される。
 パラメータη2に対応する線形予測係数に変換可能な係数の候補とは、パラメータηの値がη2である周波数領域サンプル列に対応する線形予測係数に変換可能な係数を符号化するために最適化された線形予測係数に変換可能な係数の候補である。
 <線形変換部225>
 線形変換部225には、線形予測分析部221が得た線形予測係数に変換可能な係数と、その線形予測係数に変換可能な係数に対応するパラメータη1とが入力される。パラメータη1は、例えば、後述するパラメータ決定部27,27’により決定される。
 線形変換部225は、第一線形変換部2251及び第二線形変換部2252の少なくとも一方を備えている。
 以下、(1)図1に示すように線形変換部225が第一線形変換部2251を備えている場合を第1の場合とし、(2)図2に示すように線形変換部225が第二線形変換部2252を備えている場合を第2の場合とし、(3)図3に示すように線形変換部225が第一線形変換部2251及び第二線形変換部2252を備えている場合を第3の場合として、各場合について説明する。
 (1)第1の場合
 この場合、線形変換部225の第一線形変換部2251は、符号帳記憶部222に記憶された線形予測係数に変換可能な係数の候補に対し、少なくとも入力されたパラメータη1に応じた第一線形変換を行う(ステップDE2)。
 例えば、第一線形変換部2251は、入力されたパラメータη1と符号帳記憶部222に格納された線形予測係数に変換可能な係数の候補に対応するパラメータη2とに応じた第一線形変換により、符号帳記憶部222から読み込んだパラメータη2に対応する線形予測係数に変換可能な係数の候補を、パラメータη1に対応する線形予測係数に変換可能な係数の候補に変換する。
 パラメータη1に対応する線形予測係数に変換可能な係数の候補とは、パラメータηの値がη1である周波数領域サンプル列に対応する線形予測係数に変換可能な係数を符号化するために最適化された線形予測係数に変換可能な係数の候補である。
 第一線形変換後の線形予測係数に変換可能な係数の候補は、符号化部224に出力される。
 なお、パラメータη1の値とパラメータη2の値とが同一である場合には、第一線形変換部2251は、第一線形変換をしなくてもよい。
 また、例えば、線形変換部225の第一線形変換部2251は、入力されたパラメータη1に応じて、入力されたパラメータη1が小さいほど、第一線形変換後の線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列が平坦になるように、符号帳記憶部222から読み込んだ線形予測係数に変換可能な係数の候補に対して第一線形変換を行い、変換後の線形予測係数に変換可能な係数の候補を出力する。
 一般にパラメータηが小さいほど、非平滑化スペクトル包絡系列は平坦になる傾向があり、線形予測係数に変換可能な係数はより同じような値を取る傾向がある。例えば線形予測係数に変換可能な係数がLSPである場合には、パラメータηが小さいほど、LSPである線形予測係数に変換可能な係数は0からπまでを均等分割した値により近づく傾向がある。
 図5に、パラメータηが各値を取るときのLSPパラメータの値の例を示す。図5の横軸はパラメータηであり、縦軸はLSPパラメータである。図5をみると、パラメータηが小さいほどLSPパラメータは0からπまでを均等分割した値に近づく傾向があることがわかる。
 この傾向を用いて、パラメータηが小さいほど、非平滑化スペクトル包絡系列がより平坦な場合に対応するように線形予測係数に変換可能な係数の候補を変換したものを用いて符号化及び復号を行うことにより量子化性能を向上させることができる。
 (2)第2の場合
 この場合、線形変換部225の第二線形変換部2252は、線形予測分析部221で得られた線形予測係数に変換可能な係数に対し、少なくとも入力されたパラメータη1に応じた第二線形変換を行う(ステップDE2)。
 例えば、第二線形変換部2252は、線形予測分析部221で得られたパラメータη1に対応する線形予測係数に変換可能な係数を、符号帳記憶部222に格納された線形予測係数に変換可能な係数の候補に対応するようにするために、パラメータη2に対応する線形予測係数に変換可能な係数に、第二線形変換する。
 第二線形変換後の線形予測係数に変換可能な係数は、符号化部224に出力される。
 なお、パラメータη1の値とパラメータη2の値とが同一である場合には、第二線形変換部2252は、第二線形変換をしなくてもよい。
 または、例えば、線形変換部225の第二線形変換部2252は、入力されたパラメータη1に応じて、入力されたパラメータη1が小さいほど、第二線形変換後の線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列が平坦になるように、入力された線形予測係数に変換可能な係数に対して第二線形変換を行い、変換後の線形予測係数に変換可能な係数を出力する。
 (3)第3の場合
 この場合、線形変換部225の第一線形変換部2251は、符号帳記憶部222に記憶された線形予測係数に変換可能な係数の候補に対し、少なくともパラメータη3に応じた第一線形変換を行う。パラメータη3は、正の値であり、パラメータη2とは異なる値を予め定めておくか、線形予測係数符号化装置の外部から入力されるものである。
 例えば、第一線形変換部2251は、パラメータη3と符号帳記憶部222に格納された線形予測係数に変換可能な係数の候補に対応するパラメータη2とに応じた第一線形変換により、符号帳記憶部222から読み込んだパラメータη2に対応する線形予測係数に変換可能な係数の候補を、パラメータη3に対応する線形予測係数に変換可能な係数の候補に変換する。
 パラメータη3に対応する線形予測係数に変換可能な係数の候補とは、パラメータηの値がη3である周波数領域サンプル列に対応する線形予測係数に変換可能な係数を符号化するために最適化された線形予測係数に変換可能な係数の候補である。
 第一線形変換後の線形予測係数に変換可能な係数の候補は、符号化部224に出力される。
 なお、パラメータη2の値とパラメータη3の値とが同一である場合には、第一線形変換部2251は、第一線形変換をしなくてもよい。
 また、例えば、線形変換部225の第一線形変換部2251は、パラメータη3が小さいほど、第一線形変換後の線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡が平坦になるように、符号帳記憶部222から読み込んだ線形予測係数に変換可能な係数の候補に対して第一線形変換を行い、変換後の線形予測係数に変換可能な係数の候補を出力する。
 また、この第3の場合、線形変換部225の第二線形変換部2252は、線形予測分析部221で得られた線形予測係数に変換可能な係数に対し、少なくともパラメータη1に応じた第二線形変換を行う。
 例えば、第二線形変換部2252は、線形予測分析部221で得られたパラメータη1に対応する線形予測係数に変換可能な係数を、パラメータη3に対応する線形予測係数に変換可能な係数に、第二線形変換する。
 第二線形変換後の線形予測係数に変換可能な係数の候補は、符号化部224に出力される。
 なお、パラメータη1の値とパラメータη3の値とが同一である場合には、第二線形変換部2252は、第二線形変換をしなくてもよい。
 または、例えば、線形変換部225の第二線形変換部2252は、入力されたパラメータη1に応じて、入力されたパラメータη1が小さいほど、第二線形変換後の線形予測係数に変換可能な係数に対応する振幅スペクトル包絡が平坦になるように、入力された線形予測係数に変換可能な係数に対して第二線形変換を行い、変換後の線形予測係数に変換可能な係数を出力する。
 このようにして、(3)第3の場合には、線形変換部225は、符号帳記憶部222に記憶された線形予測係数に変換可能な係数の候補に対する、η3に応じた第一線形変換と、線形予測分析部221で得られた線形予測係数に変換可能な係数に対する、η3に応じた第二線形変換との少なくとも一方を行う(ステップDE2)。
 <符号化部224>
 符号化部224の処理は、線形変換部225の構成に応じて異なる。このため、線形変換部225が(1)第1の場合、(2)第2の場合及び(3)第3の場合のそれぞれ場合の符号化部224の処理について以下に説明する。
 (1)第1の場合
 線形変換部22が(1)第1の場合には、符号化部224には、線形予測分析部221が得た線形予測係数に変換可能な係数と、線形変換部225の第一線形変換部2251が得た第一線形変換後の線形予測係数に変換可能な係数の候補とが入力される。
 符号化部224は、線形予測係数に変換可能な係数について、第一線形変換後の線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る(ステップDE3)。
 具体的には、符号化部224は、複数個の、第一線形変換後の線形予測係数に変換可能な係数の候補の中で、線形予測係数に変換可能な係数に最も近いものを選択し、その選択された候補に対応する符号を線形予測係数符号とする。
 得られた線形予測係数符号は、復号装置に出力される。
 (2)第2の場合
 線形変換部22が(2)第2の場合には、符号化部224には、線形予測分析部221の第二線形変換部2252が得た線形予測係数に変換可能な係数と、符号帳記憶部222に記憶された線形予測係数に変換可能な係数の候補とが入力される。
 符号化部224は、第二線形変換後の線形予測係数に変換可能な係数について、線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る(ステップDE3)。
 具体的には、符号化部224は、複数個の、線形予測係数に変換可能な係数の候補の中で、第二線形変換後の線形予測係数に変換可能な係数に最も近いものを選択し、その選択された候補に対応する符号を線形予測係数符号とする。
 得られた線形予測係数符号は、復号装置に出力される。
 (3)第3の場合
 線形変換部22が(3)第3の場合には、符号化部224には、線形予測分析部221の第二線形変換部2252が得た線形予測係数に変換可能な係数と、線形予測分析部221の第一線形変換部2251が得た線形予測係数に変換可能な係数の候補とが入力される。
 符号化部224は、第二線形変換後の線形予測係数に変換可能な係数について、第一線形変換後の線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る(ステップDE3)。
 具体的には、符号化部224は、複数個の、第一線形変換後の線形予測係数に変換可能な係数の候補の中で、第二線形変換後の線形予測係数に変換可能な係数に最も近いものを選択し、その選択された候補に対応する符号を線形予測係数符号とする。
 得られた線形予測係数符号は、復号装置に出力される。
 このように、線形予測係数に変換可能な係数を線形予測係数に変換可能な係数の候補を用いて符号化する際に、線形予測係数に変換可能な係数に対応するパラメータηと線形予測係数に変換可能な係数の候補に対応するパラメータηとが同じ値または近い値となるように、線形予測係数に変換可能な係数と線形予測係数に変換可能な係数の候補の少なくとも何れかに対して線形変換を行ったものを符号化に用いることにより、符号化歪を小さくすることができる及び/又は線形予測係数符号の符号量を小さくすることができる。
 (復号)
 第一実施形態の線形予測復号装置及び方法の一例について説明する。
 第一実施形態の線形予測復号装置は、図6に示すように、符号帳記憶部311、復号部313及び線形変換部314を例えば備えている。線形予測復号装置の各部が、図7に例示する各処理を行うことにより線形予測復号方法が実現される。
 <符号帳記憶部311>
 符号帳記憶部311には、符号帳記憶部222に記憶されている符号帳と同じ符号帳が記憶されている。すなわち、符号帳記憶部311には、パラメータη2に対応する線形予測係数に変換可能な係数の候補が複数個格納された符号帳が記憶されている。
 <復号部313>
 復号部313には、線形予測符号化装置が出力した線形予測係数符号が入力される。
 復号部313は、符号帳記憶部311に記憶された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補を線形予測係数に変換可能な係数として得る(ステップDD1)。
 得られた線形予測係数に変換可能な係数は、線形変換部314に出力される。
 得られた線形予測係数に変換可能な係数は、符号帳記憶部311に記憶されたパラメータη2に対応する複数個の線形予測係数に変換可能な係数の候補の何れか1つである。このため、復号部313で得られた線形予測係数に変換可能な係数は、パラメータη2に対応する線形予測係数に変換可能な係数となる。
 <線形変換部314>
 線形変換部314には、復号部313で得られたパラメータη2に対応する線形予測係数に変換可能な係数と、パラメータη1とが入力される。このパラメータη1は、例えば線形予測符号化装置から受信したパラメータ符号を復号することにより得られるものである。
 線形変換部314は、パラメータη2に対応する線形予測係数に変換可能な係数に対して、少なくともパラメータη1に応じた線形変換をして線形変換後の線形予測係数に変換可能な係数を得る。
 例えば、線形変換部314は、入力されたパラメータη1と線形予測係数に変換可能な係数に対応するパラメータη2とに応じた線形変換により、パラメータη2に対応する線形予測係数に変換可能な係数を、パラメータη1に対応する線形予測係数に変換可能な係数に変換する。
 得られた線形変換後の線形予測係数に変換可能な係数は、線形予測復号装置又は方法による復号結果として出力される。
 なお、パラメータη1の値とパラメータη2の値とが同一である場合には、線形変換部314は、線形変換をしなくてもよい。
 また、線形変換部314は、パラメータη2に対応する線形予測係数に変換可能な係数を線形変換してパラメータη1に対応する線形予測係数に変換可能な係数を得る際に、パラメータη1ともパラメータη2とも異なるパラメータη4を用いて、線形変換を複数回行う構成としてもよい。
 例えば、線形変換を2回行う場合について説明する。この場合、線形変換部314は、パラメータη2に対応する線形予測係数に変換可能な係数を線形変換してパラメータη4に対応する線形予測係数に変換可能な係数を得る。また、線形変換部314は、得られたパラメータη4に対応する線形予測係数に変換可能な係数を線形変換してパラメータη1に対応する線形予測係数に変換可能な係数を得る。ここで、パラメータη4を線形予測係数符号化装置が用いたパラメータη3と同一の値とすれば、2つの線形変換に、線形予測係数符号化装置の線形変換部225の第3の場合におけるパラメータη2に対応する線形予測係数に変換可能な係数の候補からパラメータη3に対応する線形予測係数に変換可能な係数の候補を得る線形変換と、線形予測係数符号化装置の線形変換部225の第3の場合におけるパラメータη1に対応する線形予測係数に変換可能な係数をパラメータη3に対応する線形予測係数に変換可能な係数を得る線形変換と、同一の線形変換を用いることができる。
 なお、線形変換部314は、パラメータη2からパラメータη3への線形変換と、パラメータη3からパラメータη1への線形変換とを合成した1つの線形変換を、パラメータη2に対応する線形予測係数に変換可能な係数に対してすることにより、パラメータη1に対応する線形予測係数に変換可能な係数を得てもよい。
得られたパラメータη1に対応する線形予測係数に変換可能な係数は、線形予測復号装置又は方法による復号結果として出力される。
 また、例えば、線形変換部314は、線形予測符号化装置の線形変換部225と同様に、入力されたη1が小さいほど、線形変換後の線形予測係数に変換可能な係数に対応する振幅スペクトル包絡が平坦になるように、復号部313で得られた線形予測係数に変換可能な係数を線形変換して線形変換後の線形予測係数に変換可能な係数を得てもよい。
 これは、一般にパラメータηが小さいほど、非平滑化スペクトル包絡系列は平坦になるという傾向に基づくものである。
 線形変換部314で得られた線形変換後の線形予測係数に変換可能な係数は、線形変換部314で得られた線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列を得るために用いられる。
 [線形変換]
 以下、第一線形変換及び第二線形変換等の線形変換の例について説明する。
 線形変換前の線形予測係数に変換可能な係数又は線形予測係数に変換可能な係数の候補を^ω[k][k=1,2,…,p]とし、線形変換後の線形予測係数に変換可能な係数又は上記線形予測係数に変換可能な係数の候補を~ω[k][k=1,2,…,p]とする。また、線形変換前の線形予測係数に変換可能な係数はLSPであるとする。このとき、第一線形変換部2251、第二線形変換部2252、逆線形変換部226及び線形変換部314は、例えば以下の式に示される線形変換を行う。
Figure JPOXMLDOC01-appb-M000004
 ここで、x1,x2,…xp,y1,y2,…yp-1,z2,z3,…zpを所定の非負の数とし、y1,y2,…yp-1,z2,z3,…zpの少なくとも1つは所定の正の数であるとし、Kをx1,x2,…xp,y1,y2,…yp-1,z2,z3,…zp以外の要素が0である行列とする。
 x1,x2,…xp,y1,y2,…yp-1,z2,z3,…zpの具体的な値は、線形変換前の線形予測係数に変換可能な係数又は線形予測係数に変換可能な係数の候補に対応するパラメータη(以下、線形変換前パラメータηAとする)の値と、線形変換後の線形予測係数に変換可能な係数又は線形予測係数に変換可能な係数の候補に対応するパラメータη(以下、線形変換後パラメータηBとする)の値とに基づいて適宜定まるものである。
 異なる複数の、線形変換前パラメータηAと線形変換後パラメータηBとの組に対応するx1,x2,…xp,y1,y2,…yp-1,z2,z3,…zpの具体的な値を図示していない記憶部に予め記憶しておく。第一線形変換部2251、第二線形変換部2252、逆線形変換部226及び線形変換部314は、線形変換をするときに、その線形変換における線形変換前パラメータηAと線形変換後パラメータηBとの組に対応するx1,x2,…xp,y1,y2,…yp-1,z2,z3,…zpの具体的な値を読み込み、読み込んだこれらの値を用いて上記式による線形変換を行えばよい。
 ところで、パラメータη1が大きい場合には、線形予測係数に変換可能な係数を使って計算したスペクトル包絡の変動は大きい傾向がある。このため、次数が大きい線形予測係数に変換可能な係数の候補を用いて符号化及び復号をすることが望ましい。
 逆に、パラメータη1が小さい場合には、線形予測係数に変換可能な係数を使って計算したスペクトル包絡の変動は小さい傾向がある。このため、次数が小さい線形予測係数に変換可能な係数の候補を用いて符号化及び復号をしても量子化歪は小さいため符号化及び復号の精度はそれほど悪くならない。
 このため、線形変換部225の第一線形変換部2251は、パラメータη1が小さいほど第一線形変換後の線形予測係数に変換可能な係数の候補の次数が小さくなるように第一線形変換を行ってもよい。
 同様に、線形変換部314は、パラメータη1が小さいほど線形変換後の線形予測係数に変換可能な係数の次数が小さくなるように線形変換を行ってもよい。
 このように、線形変換前の線形変換前の線形予測係数に変換可能な係数又は線形予測係数に変換可能な係数の候補の次数と、線形変換後の線形予測係数に変換可能な係数又は線形予測係数に変換可能な係数の候補の次数とが異なるように線形変換が行われてもよい。
 なお、第一線形変換部2251は、線形変換前の次数と線形変換後の次数とが同じである線形変換を行った後に線形変換後の線形予測係数に変換可能な係数の候補の次数を減らしてもよい。また、第一線形変換部2251は、線形変換後の線形予測係数に変換可能な係数の候補の次数を減らした後に線形変換前の次数と線形変換後の次数とが同じである線形変換を行ってもよい。
 同様に、線形変換部314は、線形変換前の次数と線形変換後の次数とが同じである線形変換を行った後に線形変換後の線形予測係数に変換可能な係数の次数を減らしてもよい。また、線形変換部314は、線形変換後の線形予測係数に変換可能な係数の次数を減らした後に線形変換前の次数と線形変換後の次数とが同じである線形変換を行ってもよい。
 また、第一線形変換部2251は、パラメータη1が小さい場合には、線形変換後の線形予測係数に変換可能な係数の複数の候補を統合することにより、パラメータη1が小さいほど線形変換後の線形予測係数に変換可能な係数の複数の候補数を減らしてもよい。
 [線形予測符号化装置、線形予測復号装置及びこれらの方法の第二実施形態]
 (符号化)
 第二実施形態の線形予測符号化装置及び方法の一例について説明する。
 第二実施形態の線形予測符号化装置は、図21に示すように、線形予測分析部221、符号帳記憶部222、符号帳選択部223及び符号化部224を例えば備えている。図21の例では線形予測符号化装置の外部に周波数領域変換部220が設けられているが、線形予測符号化装置が周波数領域変換部220を更に備えていてもよい。線形予測符号化装置の各部が、図22に例示する各処理を行うことにより線形予測符号化方法が実現される。
 第二実施形態では、「パラメータη1」のことを「パラメータη」と表記する。
 <周波数領域変換部220>
 周波数領域変換部220には、時系列信号である時間領域の音信号が入力される。
 周波数領域変換部41は、所定の時間長のフレーム単位で、入力された時間領域の音信号を周波数領域のN点のMDCT係数列X(0),X(1),…,X(N-1)に変換する。Nは正の整数である。
 得られたMDCT係数列X(0),X(1),…,X(N-1)は、線形予測分析部221に出力される。
 特に断りがない限り、以降の処理はフレーム単位で行われるものとする。
 このようにして、周波数領域変換部220は、時系列信号に対応する、例えばMDCT係数列である周波数領域サンプル列を求める。
 <線形予測分析部221>
 線形予測分析部221には、例えばMDCT係数列X(0),X(1),…,X(N-1)である周波数領域サンプル列及びその周波数領域サンプル列に対応するパラメータηが入力される。
 パラメータηは、正の数である。パラメータηは、例えば、後述するパラメータ決定部27,27’により決定される。パラメータηは、例えば3GPP EVS(Enhanced Voice Services)規格で使われているような線形予測包絡を利用する周波数領域の係数の量子化値を算術符号化する符号化方式において、算術符号の符号化対象の属する確率分布を定める形状パラメータである。パラメータηは、時系列信号の特徴を表す指標と成り得るものである。
 線形予測分析部221は、線形予測分析部22は、MDCT係数列X(0),X(1),…,X(N-1)及びηを用いて、以下の式(A7)により定義される~R(0),~R(1),…,~R(N-1)を用いて線形予測分析行って線形予測係数に変換可能な係数を生成する(ステップDE1)。
Figure JPOXMLDOC01-appb-M000005
 生成された線形予測係数に変換可能な係数は、符号化部224に出力される。
 具体的には、線形予測分析部22は、まずMDCT係数列X(0),X(1),…,X(N-1)の絶対値のη乗をパワースペクトルと見做した逆フーリエ変換に相当する演算、すなわち式(A7)の演算を行うことにより、MDCT係数列X(0),X(1),…,X(N-1)の絶対値のη乗に対応する時間領域の信号列である擬似相関関数信号列~R(0),~R(1),…,~R(N-1)を求める。そして、線形予測分析部22は、求まった擬似相関関数信号列~R(0),~R(1),…,~R(N-1)を用いて線形予測分析を行って、線形予測係数に変換可能な係数を生成する。
 このようにして、線形予測分析部221は、ηを正の数として、時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る。
 線形予測係数に変換可能な係数とは、例えばLSP,PARCOR係数、ISP等である。線形予測係数に変換可能な係数は、線形予測係数自体であってもよい。
 pを所定の正の数とし、線形予測係数に可能な係数の次数をp次とする。
 <符号帳記憶部222>
 符号帳記憶部222には、複数の符号帳が記憶されている。
 以下、線形予測係数に変換可能な係数の候補と、その線形予測係数に変換可能な係数の候補に対応する符号とのペアを、候補符号ペアと呼ぶことにする。各符号帳には、複数の候補符号ペアが記憶されている。言い換えると、Iを所定の2以上の数として、Niをiに応じて定まる所定の2以上の数とすると、符号帳i(i=1,2,…,I)のそれぞれには、Ni個の候補ペアが記憶されている。線形予測係数に変換可能な係数の候補に対応する符号のそれぞれには、所定の数のビットが割り当てられている。各符号は、割り当てられた所定の数のビットで表現される。
 線形予測係数に変換可能な係数の次数がpであるため、、線形予測係数に変換可能な係数の各候補はp個の値から構成される。
 符号帳記憶部222に記憶されている複数の符号帳は、符号帳選択部223の符号帳の選択方法によって異なる。このため、符号帳記憶部222に記憶されている複数の符号帳の例は、後述する符号帳選択部223の例と合わせて説明する。
 <符号帳選択部223>
 符号帳選択部223には、パラメータηが入力される。
 符号帳選択部223は、符号帳記憶部222に記憶された複数の符号帳の中から入力されたηに応じて符号帳を選択する(ステップDE2)。選択された符号帳についての情報は、符号化部224に出力される。
 以下、符号帳記憶部222に記憶された複数の符号帳の例及び符号帳選択部223による符号帳の選択基準の例について説明する。
 (1)第一の方法
 第一の方法では、符号帳記憶部222には、線形予測係数に変換可能な係数の候補数が異なる複数の符号帳が記憶されている。また、符号帳選択部223は、パラメータηが大きいほど、符号帳記憶部222に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補数が多い符号帳を選択する。
 パラメータηが大きい場合には、線形予測係数に変換可能な係数の取り得る範囲は広い傾向があるため、線形予測係数に変換可能な係数を表現するために必要な線形予測係数に変換可能な係数の候補数は多くなる。このため、パラメータηが大きい場合には、線形予測係数に変換可能な係数の候補数が多い符号帳を用いて符号化及び復号をすることが望ましい。
 逆に、パラメータηが小さい場合には、線形予測係数に変換可能な係数の取り得る範囲は狭い傾向があるため、少ない個数の線形予測係数に変換可能な係数の候補で線形予測係数に変換可能な係数を表現することができる。このため、パラメータが小さい場合には、線形予測係数に変換可能な係数の候補数が少ない符号帳を用いて符号化及び復号をしても量子化歪は小さいため符号化及び復号の精度はそれほど悪くならない。
 このため、第一の方法では、符号帳選択部223は、パラメータηが大きいほど、符号帳記憶部222に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補数が多い符号帳を選択する。
 パラメータηの大きさについての判断は、言い換えれば適切な符号帳の選択は、閾値に基づいて行うことができる。例えば、第一符号帳の線形予測係数に変換可能な係数の候補数の方が、第二符号帳の線形予測係数に変換可能な係数の候補数よりもよりも少ないとする。この場合、パラメータηの閾値を1つ予め定めておき、入力されたパラメータηが閾値よりも小さい場合はパラメータηが小さいと判断し第一符号帳を選択する。入力されたパラメータηが閾値以上である場合はパラメータηが大きいと判断し第二符号帳を選択する。符号帳の数が3以上である場合には、符号帳の数から1を減算した値の個数の閾値を用いてこれと同様に符号帳を選択すればよい。
 なお、符号帳が多層構造を有しており、パラメータηに応じてどの層まで用いるのかを決定してもよい。例えば、p=16であり、16次の線形予測係数に変換可能な係数を2層の符号帳で符号化する例について説明する。この符号帳の第一層には10ビット、第二層には5ビットの量子化ビット数が割り当てられているとする。これにより、第一層には210=1024個の、線形予測係数に変換可能な係数の候補である16次元ベクトルとその候補に対応する符号とのペアが格納され、第二層には25=32個の、線形予測係数に変換可能な係数の候補である16次元ベクトルとその候補に対応する符号とのペアが格納されているとする。
 この場合、パラメータηが大きい場合には、第一層及び第二層を用いることにし、パラメータηが小さい場合には第一層のみを用いることにする。パラメータηが大きいか小さいかの判断は、上記と同様に閾値に基づいて行うことができる。
 パラメータηが大きい場合には、まず第一層の線形予測係数に変換可能な係数の候補の中で、入力された線形予測係数に変換可能な係数に最も近いもの及び対応する符号を選択する。次に選択された線形予測係数に変換可能な係数の候補の値を入力された線形予測係数に変換可能な係数から減算し、第二層の線形予測係数に変換可能な係数の候補の中で、その減算値と最も近いもの及び対応する符号を選択する。この場合、第一層及び第二層で選択された2個の符号が線形予測係数符号となる。すなわち、線形予測係数符号は15ビットで表現される。また、第一層及び第二層で選択された線形予測係数に変換可能な係数の候補の和が、入力された線形予測係数に変換可能な係数の量子化結果となる。
 パラメータηが小さい場合には、第一層の線形予測係数に変換可能な係数の候補の中で、入力された線形予測係数に変換可能な係数に最も近いもの及び対応する符号を選択する。この場合、第一層で選択された符号が線形予測係数符号となる。すなわち、線形予測係数符号は10ビットで表現される。また、第一層で選択された線形予測係数に変換可能な係数の候補が、入力された線形予測係数に変換可能な係数の量子化結果となる。
 第一層から構成される符号帳と、第一層及び第二層から構成される符号帳とを異なる符号帳と考えると、この例も(1)第一の方法の一例と言える。
 この多層構造を有する符号帳の例にように、1つの符号帳の中の候補符号ペアの数が可変である場合には、言い換えれば1つの符号帳の中の候補符号ペアの探索範囲が可変である場合には、パラメータηが小さいほど、候補符号ペアの探索範囲を狭くしてもよい。探索範囲が異なる候補符号ペアの集合を異なる符号帳と考えれば、この例も(1)第一の方法の一例と言える。
 (2)第二の方法
 第二の方法では、符号帳記憶部222には、符号帳に記憶された線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η乗した系列である非平滑化スペクトル包絡系列の平坦度合いが異なる複数の符号帳が記憶されている。また、符号帳選択部223は、ηが小さいほど、符号帳記憶部222に記憶された複数の符号帳の中から、符号帳に記憶された線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η乗した系列である非平滑化スペクトル包絡系列がより平坦である符号帳を選択する。
 一般にパラメータηが小さいほど、非平滑化スペクトル包絡系列は平坦になる傾向があり、線形予測係数に変換可能な係数はより同じような値を取る傾向がある。例えば線形予測係数に変換可能な係数がLSPである場合には、パラメータηが小さいほど、LSPパラメータである線形予測係数に変換可能な係数は0からπまでを均等分割した値により近づく傾向がある。
 図5に、パラメータηが各値を取るときのLSPパラメータの値の例を示す。図5の横軸はパラメータηであり、縦軸はLSPパラメータである。図5をみると、パラメータηが小さいほどLSPパラメータは0からπまでを均等分割した値に近づく傾向があることがわかる。
 線形予測係数に変換可能な係数がISPパラメータの場合にも、同様の傾向がある。すなわち、線形予測係数に変換可能な係数がISPパラメータの場合、パラメータηが小さいほど、ISPパラメータである線形予測係数に変換可能な係数は0からπまでを均等分割した値により近づく傾向がある。
線形予測係数に変換可能な係数がPARCOR係数の場合には、パラメータηが小さいほど、PARCOR係数である線形予測係数に変換可能な係数は全体的に値が小さくなる傾向がある。
 第二の方法は、これらの傾向を用いて、パラメータηが小さいほど、非平滑化スペクトル包絡系列がより平坦な場合に対応する線形予測係数に変換可能な係数の候補を用いて符号化及び復号を行うことにより量子化性能を向上させようとするものである。
 線形予測係数に変換可能な係数がLSP又はPARCOR係数であるとして、符号帳i(i=1,2,…,I)の線形予測係数に変換可能な係数の候補を^ωn[1],^ωn[2],…,^ωn[p](n=1,2,…,Ni)と表記する。また、非平滑化スペクトル包絡が最も平坦な場合に対応する線形予測係数に変換可能な係数をωF[1],ωF[2],…,ωF[p]と表記する。
 この場合、第二の方法は、例えば、符号帳記憶部222には、以下のSi 1の値が異なる複数の符号帳i(i=1,2,…,I)が記憶されているとし、符号帳選択部223が、ηが小さいほど、以下のSi 1の値が小さい符号帳iを選択することにより実現される。
 Si 1=(1/pNin=1 NiΣk=1 p|^ωn[k]-ωF[k]|
 第二の方法においても、適切な符号帳の選択を閾値に基づいて行ってもよい。例えば、第一符号帳の線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η乗した系列である非平滑化スペクトル包絡系列の方が、第二符号帳の線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η乗した系列である非平滑化スペクトル包絡系列よりも平坦であるとする。この場合、パラメータηの閾値を1つ予め定めておき、入力されたパラメータηが閾値よりも小さい場合はパラメータηが小さいと判断し第一符号帳を選択する。入力されたパラメータηが閾値以上である場合はパラメータηが大きいと判断し第二符号帳を選択する。符号帳の数が3以上である場合には、符号帳の数から1を減算した値の個数の閾値を用いてこれと同様に符号帳を選択すればよい。
 (3)第三の方法
 第三の方法では、符号帳記憶部222には、線形予測係数に変換可能な係数の候補間の間隔が異なる複数の符号帳が記憶されている。また、符号帳選択部223は、ηが小さいほど、符号帳記憶部222に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補間の間隔が狭い符号帳を選択する。
 線形予測係数に変換可能な係数の候補間の間隔とは、その符号帳に含まれる線形予測係数に変換可能な係数の候補間の間隔の広さを表す指標であればどのようなものであってもよい。例えば、線形予測係数に変換可能な係数の候補間の間隔は、その符号帳に含まれる、ある線形予測係数に変換可能な係数の候補と、別のある線形予測係数に変換可能な係数の候補との距離の平均値であってもよいし、その距離の最大値、最小値又は中央値であってもよい。
 第一の方法で述べたように、パラメータηが大きい場合には、線形予測係数に変換可能な係数の変動は大きい傾向がある。このため、線形予測係数に変換可能な係数の候補間の間隔が広い符号帳を用いて符号化及び復号をすることが望ましい。
 逆に、パラメータηが小さい場合には、線形予測係数に変換可能な係数の変動は小さい傾向がある。このため、線形予測係数に変換可能な係数の候補間の間隔が狭い符号帳を用いて符号化及び復号をしても量子化歪は小さいため符号化及び復号の精度はそれほど悪くならない。
 第三の方法は、この傾向を利用したものである。
 符号帳i(i=1,2,…,I)の線形予測係数に変換可能な係数の候補を^ωn[1],^ωn[2],…,^ωn[p](n=1,2,…,Ni)と表記する。
 この場合、第三の方法は、例えば、符号帳記憶部222には、以下のSi 2の値が異なる複数の符号帳i(i=1,2,…,I)が記憶されているとし、符号帳選択部223が、ηが小さいほど、以下のSi 2の値が小さい符号帳iを選択することにより実現される。
 Si 2=(1/Nin=1 Ni-1k=1 p(^ωn[k]-^ωn+1[k])|2)1/2
 この例のように、また、線形予測係数に変換可能な係数の候補間の間隔は、その符号帳に含まれる、隣接する2個の線形予測係数に変換可能な係数の候補の距離の平均値であってもよい。
 第三の方法においても、適切な符号帳の選択を閾値に基づいて行ってもよい。例えば、第一符号帳の線形予測係数に変換可能な係数の候補間の間隔の方が、第二符号帳の線形予測係数に変換可能な係数の候補間の間隔よりも狭いとする。この場合、パラメータηの閾値を1つ予め定めておき、入力されたパラメータηが閾値よりも小さい場合はパラメータηが小さいと判断し第一符号帳を選択する。入力されたパラメータηが閾値以上である場合はパラメータηが大きいと判断し第二符号帳を選択する。符号帳の数が3以上である場合には、符号帳の数から1を減算した値の個数の閾値を用いてこれと同様に符号帳を選択すればよい。
 <符号化部224>
 符号化部224には、線形予測分析部221が得た線形予測係数に変換可能な係数及び符号帳選択部223が得た選択された符号帳についての情報が入力される。
 符号化部224は、選択された符号帳を用いて、線形予測係数に変換可能な係数を符号化して線形予測係数符号を得る(ステップDE3)。得られた線形予測係数符号は、復号装置に出力される。
 (復号)
 第二実施形態の線形予測復号装置及び方法の一例について説明する。
 第二実施形態の線形予測復号装置は、図23に示すように、符号帳記憶部311、符号帳選択部312及び復号部313を例えば備えている。線形予測復号装置の各部が、図24に例示する各処理を行うことにより線形予測復号方法が実現される。
 第二実施形態では、「パラメータη1」のことを「パラメータη」と表記する。
 <符号帳記憶部311>
 符号帳記憶部311には、複数の符号帳が記憶されている。
 以下、線形予測係数に変換可能な係数の候補と、その線形予測係数に変換可能な係数の候補に対応する符号とのペアを、候補符号ペアと呼ぶことにする。各符号帳には、複数の候補符号ペアが記憶されている。言い換えると、Iを所定の2以上の数として、Niをiに応じて定まる所定の2以上の数とすると、符号帳i(i=1,2,…,I)には、Ni個の候補ペアが記憶されている。線形予測係数に変換可能な係数の候補に対応する符号のそれぞれには、所定の数のビットが割り当てられている。各符号は、割り当てられた所定の数のビットで表現される。
 pを所定の正の数とし、線形予測係数に変換可能な係数の次元がpであるとすると、各線形予測係数に変換可能な係数の候補はp個の値から構成される。
 符号帳記憶部311に記憶されている複数の符号帳は、符号帳選択部312の符号帳の選択方法によって異なる。このため、符号帳記憶部311に記憶されている複数の符号帳の例は、後述する符号帳選択部312の例と合わせて説明する。
 なお、符号帳記憶部311には、符号帳記憶部222に記憶されている複数の符号帳と同じ符号帳が記憶されている。
 <符号帳選択部312>
 符号帳選択部312には、パラメータηが入力される。パラメータηは、パラメータ符号を復号することにより得られる。パラメータηは、符号化装置及び復号装置で予め定められた同一の数であってもよい。
 符号帳選択部312は、符号帳記憶部311に記憶された複数の符号帳の中から入力されたηに応じて符号帳を選択する(ステップDD1)。選択された符号帳についての情報は、復号部313に出力される。
 符号帳記憶部311には、符号帳記憶部222に記憶された複数の符号帳と同じ符号帳が記憶されているとする。また、符号帳選択部312には、符号化装置の符号帳選択部223による符号帳の選択基準と同じ選択基準が予め定められているとする。これにより、符号側で選択される符号帳と同じ内容の符号帳が復号側でも選択されることになる。
 符号帳の選択基準については、符号化側で説明したため、ここでは重複説明を省略する。
 <復号部313>
 復号部313には、符号化装置が出力した線形予測係数符号及び符号帳選択部312が得た選択された符号帳についての情報が入力される。また、復号部313は、選択された符号帳についての情報により特定される符号帳を符号帳記憶部311により読み込む。
 復号部313は、選択された符号帳を用いて、線形予測係数符号を復号して線形予測係数に変換可能な係数を得る(ステップDD2)。
 線形予測係数に変換可能な係数は、線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η乗した系列である非平滑化スペクトル包絡系列を得るために用いられる。
 [線形予測符号化装置、線形予測復号装置及びこれらの方法の変形例]
 図1から図3、図21及び図25から図27に一点鎖線で示すように、適合部22Aが符号帳選択部223及び線形変換部225の少なくとも一方から構成されているとすると、適合部22Aは、入力されたη1に基づいて、符号帳記憶部222に記憶された符号帳と、線形予測分析部221により生成された線形予測係数に変換可能な係数との少なくとも一方を適合させていると言える。言い換えれば、適合部22Aは、符号帳記憶部22に記憶された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、線形予測分析部221が得た線形予測係数に変換可能な係数と、のηの値を適合させていると言える。適合部22Aは、例えば、適合前の「符号帳記憶部222に記憶されている符号帳、つまり線形予測係数に変換可能な係数の複数個の候補に対応するパラメータηの値と、線形予測分析部221により生成された線形予測係数に変換可能な係数に対応するパラメータηの値との差」に比べて、適合後の2つのパラメータηの値の差が小さくなるように、少なくとも一方の線形予測係数に変換可能な係数を変形しているとも言える。なお、適合部22Aは、適合後には2つのパラメータηの値がほぼ同じ値になるように適合を行っているとも言える。。第一実施形態で説明した線形変換部225の第一線形変換部2251の処理及び第二実施形態で説明した符号帳選択部223の処理は、符号帳記憶部222に記憶された符号帳の適合の一例である。第二実施形態で説明した線形変換部225の第二線形変換部2252の処理は、線形予測分析部221により生成された線形予測係数に変換可能な係数の適合の一例である。
 この場合、符号化部224は、適合部22Aにより適合された少なくとも一方の符号帳及び線形予測係数に変換可能な係数を用いて、符号化を行っていると言える。言い換えれば、符号化部224は、符号帳選択部223で選択された符号帳又は適合部22Aにより適合された符号帳を用いて、線形予測分析部221により線形予測係数に変換可能な係数又は適合部22Aにより適合された線形予測係数に変換可能な係数を符号化していると言える。さらに、言い換えれば、符号化部224は、ηの値が適合された線形予測係数に変換可能な係数の複数個の候補と線形予測係数に変換可能な係数とを用いて、線形予測分析部221が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得ていると言える。
 第一実施形態の(1)第1の場合の適合部22Aは、符号帳記憶部222に記憶された線形予測係数に変換可能な係数の候補に対して、η1に応じた第一線形変換を行い、第一線形変換後の線形予測係数に変換可能な係数の複数個の候補を得る線形変換部225を備えていると言える。この場合、符号化部224は、線形予測分析部221が得た線形予測係数に変換可能な係数と、適合部22Aが得た第一線形変換後の線形予測係数に変換可能な係数の複数個の候補と、を用いて、線形予測分析部221が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得ていると言える。
 第一実施形態の(2)第2の場合の適合部22Aは、線形予測分析部221が得た線形予測係数に変換可能な係数に対して、η1に応じた第二線形変換を行い、第二線形変換後の線形予測係数に変換可能な係数を得る線形変換部225を備えていると言える。この場合、符号化部224は、適合部22Aが得た第二線形変換後の線形予測係数に変換可能な係数と、符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、を用いて、線形予測分析部221が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得ていると言える。
 第一実施形態の(3)第3の場合の適合部22Aは、符号帳記憶部222には、η2に対応する符号帳が記憶されているとして、符号帳記憶部222に記憶された線形予測係数に変換可能な係数の複数個の候補に対して、η3に応じた第一線形変換を行い、第一線形変換後の線形予測係数に変換可能な係数の複数個の候補を得、線形予測分析部221が得た線形予測係数に変換可能な係数に対して、η3に応じた第二線形変換を行い、第二線形変換後の線形予測係数に変換可能な係数を得ていると言える。この場合、符号化部224は、適合部22Aが得た第二線形変換後の線形予測係数に変換可能な係数と、適合部22Aが得た第一線形変換後の線形予測係数に変換可能な係数の複数個の候補と、を用いて、線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得ていると言える。
 適合部22Aは、例えば図25に示す符号帳選択部223及び第二線形変換部2252により、符号帳の適合を行ってもよい。例えば、パラメータη2は所定のパラメータηであるとして、符号帳選択部223は、符号帳記憶部222に記憶された複数の符号帳の中からパラメータη2に応じて符号帳を選択する。そして、第二線形変換部2252は、線形予測分析部221で得られた線形予測係数に変換可能な係数に対する、η2に応じた第二線形変換を行う。この場合、符号化部224は、第二線形変換後の線形予測係数に変換可能な係数について、選択された符号帳を用いて符号化して線形予測係数符号を得る。
 適合部22Aは、例えば図26に示す符号帳選択部223及び第一線形変換部2251により、符号帳の適合を行ってもよい。例えば、パラメータη2は所定のパラメータηであるとして、符号帳選択部223は、符号帳記憶部222に記憶された複数の符号帳の中からパラメータη2に応じて符号帳を選択する。そして、第一線形変換部2251は、選択された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補に対する、η1に応じた第一線形変換を行う。この場合、符号化部224は、線形予測分析部221で得られた線形予測係数に変換可能な係数について、第一線形変換後の線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る。
 適合部22Aは、例えば図27に示す符号帳選択部223、第一線形変換部2251及び第二変換部2252により、符号帳の適合を行ってもよい。例えば、パラメータη23は所定のパラメータηであるとして、符号帳選択部223は、符号帳記憶部222に記憶された複数の符号帳の中からパラメータη3に応じて符号帳を選択する。そして、第一線形変換部2251は、選択された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補に対する、η2に応じた第一線形変換を行う。そして、第二線形変換部2252は、線形予測分析部221で得られた線形予測係数に変換可能な係数に対する、η2に応じた第二線形変換を行う。この場合、符号化部224は、第二線形変換後の線形予測係数に変換可能な係数について、第一線形変換後の線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る。
 図6、図23及び図28に一点鎖線で示すように、適合部31Aが符号帳選択部312及び線形変換部314の少なくとも一方と、復号部313とから構成されているとすると、適合部31Aは、η1を正の数として、入力されたη1に基づいて、符号帳記憶部311に記憶された符号帳と、符号帳に格納された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補との少なくとも一方を適合させていると言える。
 適合部31Aは、例えば図28に示す符号帳選択部312及び線形変換部314の両方において適合の処理を行ってもよい。例えば、η2を正の数として、符号帳選択部312は、符号帳記憶部311に記憶された複数の符号帳の中からパラメータηに応じて符号帳を選択する。そして、線形変換部314は、復号部313で得られた線形予測係数に変換可能な係数に対して、所定の正の数であるη1に応じた線形変換をして線形予測係数に変換可能な係数を得る。
 [符号化装置、復号装置及びこれらの方法]
 以下、線形予測符号化装置、線形予測復号装置及びこれらの方法を用いた符号化装置、復号装置及びこれらの方法の例について説明する。
 [符号化装置、復号装置及びこれらの方法の第一実施形態]
 (符号化)
 第一実施形態の符号化装置の構成例を図8に示す。第一実施形態の符号化装置は、図8に示すように、周波数領域変換部21と、線形予測分析部22と、非平滑化振幅スペクトル包絡系列生成部23と、平滑化振幅スペクトル包絡系列生成部24と、包絡正規化部25と、符号化部26と、パラメータ決定部27とを例えば備えている。この符号化装置により実現される第一実施形態の符号化方法の各処理の例を図9に示す。
 以下、図8の各部について説明する。
 <パラメータ決定部27>
 第一実施形態では、所定の時間区間ごとに複数のパラメータηの何れかがパラメータ決定部27により選択可能とされている。
 パラメータ決定部27には、複数のパラメータηがパラメータηの候補として記憶されているとする。パラメータ決定部27は、複数のパラメータの中の1つのパラメータηを順次読み出し、線形予測分析部22、非平滑化振幅スペクトル包絡系列生成部23及び復号化部26に出力する(ステップA0)。
 周波数領域変換部21、線形予測分析部22、非平滑化振幅スペクトル包絡系列生成部23、平滑化振幅スペクトル包絡系列生成部24、包絡正規化部25及び符号化部26は、パラメータ決定部27が順次読み出した各パラメータηに基づいて、例えば以下に説明するステップA1からステップA6の処理を行い同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対して符号を生成する。一般に、パラメータηを所与として、同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対して2個以上の符号が得られる場合がある。この場合、同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対する符号は、これらの得られた2個以上の符号をまとめたものである。この例では、符号は、線形予測係数符号と、利得符号と、整数信号符号とを合わせたものである。これにより、同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対する各パラメータηごとの符号が得られる。
 ステップA6の処理の後に、パラメータ決定部27は、同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対して各パラメータηごとに得られた符号の中から1つの符号を選択し、選択された符号に対応するパラメータηを決定する(ステップA7)。この決定されたパラメータηが、その同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対するパラメータηとなる。そして、パラメータ決定部27は、選択された符号及び決定されたパラメータηを表す符号を復号装置に出力する。パラメータ決定部27によるステップA7の処理の詳細については後述する。
 以下では、パラメータ決定部27により1つのパラメータη1が読み出されており、この読み出された1つのパラメータη1について処理が行われるとする。
 <周波数領域変換部21>
 周波数領域変換部21には、時間領域の時系列信号である音信号が入力される。音信号の例は、音声ディジタル信号又は音響ディジタル信号である。
 周波数領域変換部21は、所定の時間長のフレーム単位で、入力された時間領域の音信号を周波数領域のN点のMDCT係数列X(0),X(1),…,X(N-1)に変換する(ステップA1)。Nは正の整数である。
 得られたMDCT係数列X(0),X(1),…,X(N-1)は、線形予測分析部22と包絡正規化部25に出力される。
 特に断りがない限り、以降の処理はフレーム単位で行われるものとする。
 このようにして、周波数領域変換部21は、音信号に対応する、例えばMDCT係数列である周波数領域サンプル列を求める。
 <線形予測分析部22>
 線形予測分析部22には、周波数領域変換部21が得たMDCT係数列X(0),X(1),…,X(N-1)が入力される。
 線形予測分析部22は、[線形予測符号化装置、線形予測復号装置及びこれらの方法]で説明した図1から図3、図21の何れかの線形予測符号化装置である。[符号化装置、復号装置及びこれらの方法]及び図8では、[線形予測符号化装置、線形予測復号装置及びこれらの方法]で説明した図1から図3、図21の何れかの線形予測符号化装置のことを「線形予測分析部22」と表記する。なお、線形予測分析部22は、図25から図27の何れかの線形予測符号化装置であってもよい。
 線形予測分析部22は、[線形予測符号化装置、線形予測復号装置及びこれらの方法]で説明した処理と同様の処理により、例えばMDCT係数列である周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得て、得られた線形予測係数に変換可能な係数を符号化して線形予測係数符号を得る。
 得られた線形予測係数符号は、パラメータ決定部27及び復号装置に出力される。
 また、線形予測符号化装置の線形変換部225が(1)第1の場合には、符号化部224で得られた線形予測係数符号に対応する、パラメータη1に対応する線形予測係数に変換可能な係数が量子化線形予測係数^β1,^β2,…,^βpとして、非平滑化スペクトル包絡系列生成部23と平滑化振幅スペクトル包絡系列生成部24に出力される。
 線形予測符号化装置の線形変換部225が(2)第2の場合には、符号化部224で得られた線形予測係数符号に対応する、パラメータη2に対応する線形予測係数に変換可能な係数が、図2に破線で示す逆線形変換部226に入力される。逆線形変換部226は、線形予測係数符号に対応する、パラメータη2に対応する線形予測係数に変換可能な係数に対して第二線形変換部2252が行った第二線形変換の逆の線形変換を行い、パラメータη1に対応する線形予測係数に変換可能な係数とする。このパラメータη1に対応する線形予測係数に変換可能な係数が、量子化線形予測係数^β1,^β2,…,^βpとして、非平滑化スペクトル包絡系列生成部23と平滑化振幅スペクトル包絡系列生成部24に出力される。なお、パラメータη1の値とパラメータη2の値とが同一である場合には、逆線形変換部226は、線形変換をしなくてもよい。
 線形予測符号化装置の線形変換部225が(3)第3の場合には、符号化部224で得られた線形予測係数符号に対応する、パラメータη3に対応する線形予測係数に変換可能な係数が、図3に破線で示す逆線形変換部226に入力される。逆線形変換部226は、線形予測係数符号に対応する、パラメータη3に対応する線形予測係数に変換可能な係数に対して第二線形変換部2252が行った第二線形変換の逆の線形変換を行い、パラメータη1に対応する線形予測係数に変換可能な係数とする。このパラメータη1に対応する線形予測係数に変換可能な係数が、量子化線形予測係数^β1,^β2,…,^βpとして、非平滑化スペクトル包絡系列生成部23と平滑化振幅スペクトル包絡系列生成部24に出力される。なお、パラメータη1の値とパラメータη3の値とが同一である場合には、逆線形変換部226は、線形変換をしなくてもよい。
 なお、線形予測分析処理の過程で予測残差のエネルギーσ2が算出される。この場合、算出された予測残差のエネルギーσ2は、符号化部26の分散パラメータ決定部268に出力される。
 <非平滑化振幅スペクトル包絡系列生成部23>
 非平滑化振幅スペクトル包絡系列生成部23には、線形予測分析部22が生成した量子化線形予測係数^β1,^β2,…,^βpが入力される。
 非平滑化振幅スペクトル包絡系列生成部23は、量子化線形予測係数^β1,^β2,…,^βpに対応する振幅スペクトル包絡の系列である非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)を生成する(ステップA3)。
 生成された非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)は、符号化部26に出力される。
 非平滑化振幅スペクトル包絡系列生成部23は、量子化線形予測係数^β1,^β2,…,^βpを用いて、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)として、式(A2)により定義される非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)を生成する。
Figure JPOXMLDOC01-appb-M000006
 このようにして、非平滑化振幅スペクトル包絡系列生成部23は、線形予測分析部22により生成された線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列を得ることによりスペクトル包絡の推定を行う。ここで、cを任意の数として、複数の値から構成される系列をc乗した系列とは、複数の値のそれぞれをc乗した値から構成される系列のことである。例えば、振幅スペクトル包絡の系列を1/η1乗した系列とは、振幅スペクトル包絡の各係数を1/η1乗した値から構成される系列のことである。
 非平滑化振幅スペクトル包絡系列生成部23による1/η1乗の処理は、線形予測分析部22で行われた周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した処理に起因するものである。すなわち、非平滑化振幅スペクトル包絡系列生成部23による1/η1乗の処理は、線形予測分析部22で行われた周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した処理によりη1乗された値を元の値に戻すために行われる。
 <平滑化振幅スペクトル包絡系列生成部24>
 平滑化振幅スペクトル包絡系列生成部24には、線形予測分析部22が生成した量子化線形予測係数^β1,^β2,…,^βpが入力される。
 平滑化振幅スペクトル包絡系列生成部24は、量子化線形予測係数^β1,^β2,…,^βpに対応する振幅スペクトル包絡の系列の振幅の凸凹を鈍らせた系列である平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)を生成する(ステップA4)。
 生成された平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)は、包絡正規化部25及び符号化部26に出力される。
 平滑化振幅スペクトル包絡系列生成部24は、量子化線形予測係数^β1,^β2,…,^βpと補正係数γを用いて、平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)として、式(A3)により定義される平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)を生成する。
Figure JPOXMLDOC01-appb-M000007
 ここで、補正係数γは予め定められた1未満の定数であり非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)の振幅の凹凸を鈍らせる係数、言い換えれば非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)を平滑化する係数である。
 <包絡正規化部25>
 包絡正規化部25には、周波数領域変換部21が得たMDCT係数列X(0),X(1),…,X(N-1)及び平滑化振幅スペクトル包絡生成部24が生成した平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)が入力される。
 包絡正規化部25は、MDCT係数列X(0),X(1),…,X(N-1)の各係数を、対応する平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)の各値で正規化することにより、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)を生成する(ステップA5)。
 生成された正規化MDCT係数列は、符号化部26に出力される。
 包絡正規化部25は、例えば、k=0,1,…,N-1として、MDCT係数列X(0),X(1),…,X(N-1)の各係数X(k)を平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)で除算することにより、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)の各係数XN(k)を生成する。すなわち、k=0,1,…,N-1として、XN(k)=X(k)/^Hγ(k)である。
 <符号化部26>
 符号化部26には、包絡正規化部25が生成した正規化MDCT係数列XN(0),XN(1),…,XN(N-1)、非平滑化振幅スペクトル包絡生成部23が生成した非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)、平滑化振幅スペクトル包絡生成部24が生成した平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)及び線形予測分析部22が算出した予測残差のエネルギーσ2が入力される。
 符号化部26は、図12に示すステップA61からステップA65の処理を例えば行うことにより符号化を行う(ステップA6)。
 符号化部26は、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)に対応するグローバルゲインgを求め(ステップA61)、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)の各係数をグローバルゲインgで割り算した結果を量子化した整数値による系列である量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を求め(ステップA62)、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数に対応する分散パラメータφ(0),φ(1),…,φ(N-1)をグローバルゲインgと非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)と平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)と平均残差のエネルギーσ2とから式(A1)により求め(ステップA63)、分散パラメータφ(0),φ(1),…,φ(N-1)を用いて量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を算術符号化して整数信号符号を得(ステップA64)、グローバルゲインgに対応する利得符号を得る(ステップA65)。
Figure JPOXMLDOC01-appb-M000008
 ここで、上記の式(A1)における正規化振幅スペクトル包絡系列^HN(0),^HN(1),…,^HNは、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)の各値を、対応する平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)の各値で除算したもの、すなわち、以下の式(A8)により求まるものである。
Figure JPOXMLDOC01-appb-M000009
 生成された整数信号符号と利得符号は正規化MDCT係数列に対応する符号として、パラメータ決定部27に出力される。
 符号化部26は、ステップA61からステップA65により、整数信号符号のビット数が、予め配分されたビット数である配分ビット数B以下、かつ、なるべく大きな値となるようなグローバルゲインgを決定し、決定されたグローバルゲインgに対応する利得符号と、この決定されたグローバルゲインgに対応する整数信号符号とを生成する機能を実現している。
 符号化部26が行うステップA61からステップA65のうち、の特徴的な処理が含まれるのはステップA63であり、グローバルゲインgと量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)のそれぞれを符号化することにより正規化MDCT係数列に対応する符号を得る符号化処理自体には、非特許文献1に記載された技術を含む様々な公知技術が存在する。以下では符号化部26が行う符号化処理の具体例を2つ説明する。
 [符号化部26が行う符号化処理の具体例1]
 符号化部26が行う符号化処理の具体例1として、ループ処理を含まない例について説明する。
 具体例1の符号化部26の構成例を図10に示す。具体例1の符号化部26は、図10に示すように、利得取得部261と、量子化部262と、分散パラメータ決定部268と、算術符号化部269と、利得符号化部265とを例えば備えている。以下、図10の各部について説明する。
 <利得取得部261>
 利得取得部261には、包絡正規化部25が生成した正規化MDCT係数列XN(0),XN(1),…,XN(N-1)が入力される。
 利得取得部261は、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)から、整数信号符号のビット数が、予め配分されたビット数である配分ビット数B以下、かつ、なるべく大きな値となるようなグローバルゲインgを決定して出力する(ステップS261)。利得取得部261は、例えば、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)のエネルギーの合計の平方根と配分ビット数Bと負の相関のある定数との乗算値をグローバルゲインgとして得て出力する。または、利得取得部261は、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)のエネルギーの合計と、配分ビット数Bと、グローバルゲインgと、の関係を予めテーブル化しておき、そのテーブルを参照することによりグローバルゲインgを得て出力してもよい。
 このようにして、利得取得部261は、例えば正規化MDCT係数列である正規化周波数領域サンプル列の全サンプルを除算するための利得を得る。
 得られたグローバルゲインgは、量子化部262及び分散パラメータ決定部268に出力される。
 <量子化部262>
 量子化部262には、包絡正規化部25が生成した正規化MDCT係数列XN(0),XN(1),…,XN(N-1)及び利得取得部261が得たグローバルゲインgが入力される。
 量子化部262は、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)の各係数をグローバルゲインgで割り算した結果の整数部分による系列である量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を得て出力する(ステップS262)。
 このようにして、量子化部262は、例えば正規化MDCT係数列である正規化周波数領域サンプル列の各サンプルを、利得で除算するとともに量子化して量子化正規化済係数系列を求める。
 得られた量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)は、算術符号化部269に出力される。
 <分散パラメータ決定部268>
 分散パラメータ決定部268には、パラメータ決定部27が読み出したパラメータη1、利得取得部261が得たグローバルゲインg、非平滑化振幅スペクトル包絡生成部23が生成した非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)、平滑化振幅スペクトル包絡生成部24が生成した平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)及び線形予測分析部22が得た予測残差のエネルギーσ2が入力される。
 分散パラメータ決定部268は、グローバルゲインgと、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)と、平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)と、予測残差のエネルギーσ2とから、上記の式(A1),式(A8)により分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータを得て出力する(ステップS268)。
 得られた分散パラメータ系列φ(0),φ(1),…,φ(N-1)は、算術符号化部269に出力される。
 <算術符号化部269>
 算術符号化部269には、パラメータ決定部27が読み出したパラメータη1、量子化部262が得た量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)及び分散パラメータ決定部268が得た分散パラメータ系列φ(0),φ(1),…,φ(N-1)が入力される。
 算術符号化部269は、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数に対応する分散パラメータとして分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータを用いて、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を算術符号化して整数信号符号を得て出力する(ステップS269)。
 算術符号化部269は、算術符号化の際に、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数が一般化ガウス分布fGG(X|φ(k),η1)に従うときに最適になるような算術符号を構成し、この構成に基づく算術符号により符号化を行う。この結果、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数へのビット割り当ての期待値が分散パラメータ系列φ(0),φ(1),…,φ(N-1)で決定されることになる。
 得られた整数信号符号は、パラメータ決定部27に出力される。
 量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の中の複数の係数に跨って算術符号化が行われてもよい。この場合、分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータは、式(A1),式(A8)からわかるように、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)に基づいているため、算術符号化部269は、推定されたスペクトル包絡(非平滑化振幅スペクトル包絡)を基に実質的にビット割り当てが変わる符号化を行っていると言える。
 <利得符号化部265>
 利得符号化部265には、利得取得部261が得たグローバルゲインgが入力される。
 利得符号化部265は、グローバルゲインgを符号化して利得符号を得て出力する(ステップS265)。
 生成された整数信号符号と利得符号は正規化MDCT係数列に対応する符号として、パラメータ決定部27に出力される。
 本具体例1のステップS261,S262,S268,S269,S265がそれぞれ上記のステップA61,A62,A63,A64,A65に対応する。
 [符号化部26が行う符号化処理の具体例2]
 符号化部26が行う符号化処理の具体例2として、ループ処理を含む例について説明する。
 具体例2の符号化部26の構成例を図11に示す。具体例2の符号化部26は、図11に示すように、利得取得部261と、量子化部262と、分散パラメータ決定部268と、算術符号化部269と、利得符号化部265と、判定部266と、利得更新部267とを例えば備えている。以下、図11の各部について説明する。
 <利得取得部261>
 利得部261には、包絡正規化部25が生成した正規化MDCT係数列XN(0),XN(1),…,XN(N-1)が入力される。
 利得取得部261は、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)から、整数信号符号のビット数が、予め配分されたビット数である配分ビット数B以下、かつ、なるべく大きな値となるようなグローバルゲインgを決定して出力する(ステップS261)。利得取得部261は、例えば、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)のエネルギーの合計の平方根と配分ビット数Bと負の相関のある定数との乗算値をグローバルゲインgとして得て出力する。
 得られたグローバルゲインgは、量子化部262及び分散パラメータ決定部268に出力される。
 利得取得部261が得たグローバルゲインgは、量子化部262及び分散パラメータ決定部268で用いられるグローバルゲインの初期値となる。
 <量子化部262>
 量子化部262には、包絡正規化部25が生成した正規化MDCT係数列XN(0),XN(1),…,XN(N-1)及び利得取得部261又は利得更新部267が得たグローバルゲインgが入力される。
 量子化部262は、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)の各係数をグローバルゲインgで割り算した結果の整数部分による系列である量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を得て出力する(ステップS262)。
 ここで、量子化部262が初回に実行される際に用いられるグローバルゲインgは、利得取得部261が得たグローバルゲインg、すなわちグローバルゲインの初期値である。また、量子化部262が2回目以降に実行される際に用いられるグローバルゲインgは、利得更新部267が得たグローバルゲインg、すなわちグローバルゲインの更新値である。
 得られた量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)は、算術符号化部269に出力される。
 <分散パラメータ決定部268>
 分散パラメータ決定部268には、パラメータ決定部27が読み出したパラメータη1、利得取得部261又は利得更新部267が得たグローバルゲインg、非平滑化振幅スペクトル包絡生成部23が生成した非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)、平滑化振幅スペクトル包絡生成部24が生成した平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)及び線形予測分析部22が得た予測残差のエネルギーσ2が入力される。
 分散パラメータ決定部268は、グローバルゲインgと、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)と、平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)と、予測残差のエネルギーσ2とから、上記の式(A1),式(A8)により分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータを得て出力する(ステップS268)。
 ここで、分散パラメータ決定部268が初回に実行される際に用いられるグローバルゲインgは、利得取得部261が得たグローバルゲインg、すなわちグローバルゲインの初期値である。また、分散パラメータ決定部268が2回目以降に実行される際に用いられるグローバルゲインgは、利得更新部267が得たグローバルゲインg、すなわちグローバルゲインの更新値である。
 得られた分散パラメータ系列φ(0),φ(1),…,φ(N-1)は、算術符号化部269に出力される。
 <算術符号化部269>
 算術符号化部269には、パラメータ決定部27が読み出したパラメータη1、量子化部262が得た量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)及び分散パラメータ決定部268が得た分散パラメータ系列φ(0),φ(1),…,φ(N-1)が入力される。
 算術符号化部269は、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数に対応する分散パラメータとして分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータを用いて、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を算術符号化して、整数信号符号と整数信号符号のビット数である消費ビット数Cとを得て出力する(ステップS269)。
 算術符号化部269は、算術符号化の際に、量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数が一般化ガウス分布fGG(X|φ(k),η1)に従うときに最適になるようなビット割り当てを算術符号により行い、行われたビット割り当てに基づく算術符号により符号化を行う。
 得られた整数信号符号及び消費ビット数Cは、判定部266に出力される。
 量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の中の複数の係数に跨って算術符号化が行われてもよい。この場合、分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータは、式(A1),式(A8)からわかるように、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)に基づいているため、算術符号化部269は、推定されたスペクトル包絡(非平滑化振幅スペクトル包絡)を基に実質的にビット割り当てが変わる符号化を行っていると言える。
 <判定部266>
 判定部266には、算術符号化部269が得た整数信号符号が入力される。
 判定部266は、利得の更新回数が予め定めた回数の場合には、整数信号符号を出力するとともに、利得符号化部265に対し利得更新部267が得たグローバルゲインgを符号化する指示信号を出力し、利得の更新回数が予め定めた回数未満である場合には、利得更新部267に対し、算術符号化部264が計測した消費ビット数Cを出力する(ステップS266)。
 <利得更新部267>
 利得更新部267には、算術符号化部264が計測した消費ビット数Cが入力される。
 利得更新部267は、消費ビット数Cが配分ビット数Bより多い場合にはグローバルゲインgの値を大きな値に更新して出力し、消費ビット数Cが配分ビット数Bより少ない場合にはグローバルゲインgの値を小さな値に更新し、更新後のグローバルゲインgの値を出力する(ステップS267)。
 利得更新部267が得た更新後のグローバルゲインgは、量子化部262及び利得符号化部265に出力される。
 <利得符号化部265>
 利得符号化部265には、判定部266からの出力指示及び利得更新部267が得たグローバルゲインgが入力される。
 利得符号化部265は、指示信号に従って、グローバルゲインgを符号化して利得符号を得て出力する(ステップ265)。
 判定部266が出力した整数信号符号と、利得符号化部265が出力した利得符号は、正規化MDCT係数列に対応する符号として、パラメータ決定部27に出力される。
 すなわち、本具体例2においては、最後に行われたステップS267が上記のステップA61に対応し、ステップS262,S263,S264,S265がそれぞれ上記のステップA62,A63,A64,A65に対応する。
 なお、符号化部26が行う符号化処理の具体例2については、国際公開公報WO2014/054556などに更に詳細に説明されている。
 [符号化部26の変形例]
 符号化部26は、例えば以下の処理を行うことにより、推定されたスペクトル包絡(非平滑化振幅スペクトル包絡)を基にビット割り当てを変える符号化を行ってもよい。
 符号化部26は、まず、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)に対応するグローバルゲインgを求め、正規化MDCT係数列XN(0),XN(1),…,XN(N-1)の各係数をグローバルゲインgで割り算した結果を量子化した整数値による系列である量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)を求める。
 この量子化正規化済係数系列XQ(0),XQ(1),…,XQ(N-1)の各係数に対応する量子化ビットは、XQ(k)の分布がある範囲内で一様であると仮定して、その範囲を包絡の推定値から決めることができる。複数のサンプルごとの包絡の推定値を符号化することもできるが、符号化部26は、例えば以下の式(A9)のように線形予測に基づく正規化振幅スペクトル包絡系列の値^HN(k)を使用してXQ(k)の範囲を決めることができる。
Figure JPOXMLDOC01-appb-M000010
あるkにおけるXQ(k)を量子化するときに、XQ(k)の二乗誤差を最小とするために
Figure JPOXMLDOC01-appb-M000011
の制約のもとに、割り当てるビット数b(k)
Figure JPOXMLDOC01-appb-M000012
を設定することができる。Bは予め定められた正の整数である。この際にb(k)が整数となるように四捨五入するとか、0より小さくなる場合にはb(k)=0とするなどして、b(k)の再調整の処理を符号化部26は行ってもよい。
 また、符号化部26は、サンプルごとの割り当てでなく、複数のサンプルをまとめて配分ビット数を決めて、量子化にもサンプルごとのスカラ量子化でなく、複数のサンプルをまとめたベクトルごとの量子化をすることも可能である。
 サンプルkのXQ(k)の量子化ビット数b(k)が上記で与えられ、サンプルごとに符号化するとすると、XQ(k)は-2b(k)-1から2b(k)-1までの2b(k)種類の整数を取り得る。符号化部26は、b(k)ビットで各サンプルを符号化して整数信号符号を得る。
 生成された整数信号符号は、復号装置に出力される。例えば、生成されたXQ(k)に対応するb(k)ビットの整数信号符号は、k=0から順次復号装置に出力される。
 もし、XQ(k)が上記の-2b(k)-1から2b(k)-1までの範囲をこえる場合には最大値、または最小値に置き換える。
 gが小さすぎるとこの置き換えで量子化歪が発生し、gが大きすぎると量子化誤差は大きくなり、XQ(k)のとりうる範囲がb(k)に比べて小さすぎて、情報の有効利用ができないことになる。このため、gの最適化を行ってもよい。
 符号化部26は、グローバルゲインgを符号化して利得符号を得て出力する。
 この符号化部26の変形例のように、符号化部26は算術符号化以外の符号化を行ってもよい。
 <パラメータ決定部27>
 ステップA1からステップA6の処理により、同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対して各パラメータη1ごとに生成された符号(この例では、線形予測係数符号、利得符号及び整数信号符号)は、パラメータ決定部27に入力される。
 パラメータ決定部27は、同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対して各パラメータη1ごとに得られた符号の中から1つの符号を選択し、選択された符号に対応するパラメータηを決定する(ステップA7)。この決定されたパラメータηが、その同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対するパラメータηとなる。そして、パラメータ決定部27は、選択された符号及び決定されたパラメータηを表すパラメータ符号を復号装置に出力する。符号の選択は、符号の符号量及び符号に対応する符号化歪の少なくとも一方に基づいて行われる。例えば、符号量が最も小さい符号又は符号化歪が最も小さい符号が選択される。
 ここで、符号化歪みとは、入力信号から得られる周波数領域サンプル列と、生成された符号をローカルデコードすることにより得られる周波数領域サンプル列との誤差のことである。符号化装置は、符号化歪みを計算するための符号化歪計算部を備えていてもよい。この符号化歪計算部は、以下に述べる復号装置と同様の処理を行う復号部を備え、この復号部が生成された符号をローカルデコードする。その後、符号化歪計算部は、入力信号から得られる周波数領域サンプル列と、ローカルデコードすることにより得られた周波数領域サンプル列との誤差を計算し、符号化歪とする。
 (復号)
 符号化装置に対応する復号装置の構成例を図13に示す。第一実施形態の復号装置は、図13に示すように、線形予測係数復号部31と、非平滑化振幅スペクトル包絡系列生成部32と、平滑化振幅スペクトル包絡系列生成部33と、復号部34と、包絡逆正規化部35と、時間領域変換部36と、パラメータ復号部37とを例えば備えている。この復号装置により実現される第一実施形態の復号方法の各処理の例を図14に示す。
 復号装置には、符号化装置が出力した、パラメータ符号、正規化MDCT係数列に対応する符号及び線形予測係数符号が少なくとも入力される。
 以下、図13の各部について説明する。
 <パラメータ復号部37>
 パラメータ復号部37には、符号化装置が出力したパラメータ符号が入力される。
 パラメータ復号部37は、パラメータ符号を復号することにより復号パラメータηを求める。求まった復号パラメータηは、線形予測係数復号部31、非平滑化振幅スペクトル包絡系列生成部32、平滑化振幅スペクトル包絡系列生成部33及び復号部34に出力される。パラメータ復号部37には、複数の復号パラメータηが候補として記憶されいる。パラメータ復号部37は、パラメータ符号に対応する復号パラメータηの候補を復号パラメータηとして求める。パラメータ復号部37に記憶されている複数の復号パラメータηは、符号化装置のパラメータ決定部27に記憶された複数のパラメータηと同じである。
 <線形予測係数復号部31>
 線形予測係数復号部31には、符号化装置が出力した線形予測係数符号及びパラメータ復号部37により得られた復号パラメータηが入力される。
 線形予測係数復号部31は、[線形予測符号化装置、線形予測復号装置及びこれらの方法]で説明した図6、図21を用いて上記説明した線形予測復号装置である。[符号化装置、復号装置及びこれらの方法]及び図13では、[線形予測符号化装置、線形予測復号装置及びこれらの方法]で説明した図6、図21の線形予測符号化装置のことを「線形予測係数復号部31」と表記する。なお、線形予測係数復号部31は、図28の線形予測復号装置であってもよい。
 線形予測係数復号部31は、復号パラメータηをパラメータη1とする[線形予測符号化装置、線形予測復号装置及びこれらの方法]で説明した処理と同様の処理により、入力された線形予測係数符号を復号することにより、復号された線形予測係数に変換可能な係数である復号線形予測係数^β1,^β2,…, ^βpを得る(ステップB1)。
 得られた復号線形予測係数^β1,^β2,…, ^βpは、非平滑化振幅スペクトル包絡系列生成部32及び非平滑化振幅スペクトル包絡系列生成部33に出力される。
 <非平滑化振幅スペクトル包絡系列生成部32>
 非平滑化振幅スペクトル包絡系列生成部32には、パラメータ復号部37が求めた復号パラメータη及び線形予測係数復号部31が得た復号線形予測係数^β1,^β2,…,^βpが入力される。
 非平滑化振幅スペクトル包絡系列生成部32は、復号線形予測係数^β1,^β2,…,^βpに対応する振幅スペクトル包絡の系列である非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)を上記の式(A2)により生成する(ステップB2)。
 生成された非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)は、復号部34に出力される。
 このようにして、非平滑化振幅スペクトル包絡系列生成部32は、線形予測係数復号部31により生成された線形予測係数に変換可能な係数に対応するに対応する振幅スペクトル包絡の系列を1/η乗した系列である非平滑化スペクトル包絡系列を得る。
 <平滑化振幅スペクトル包絡系列生成部33>
 平滑化振幅スペクトル包絡系列生成部33には、パラメータ復号部37が求めた復号パラメータη及び線形予測係数復号部31が得た復号線形予測係数^β1,^β2,…,^βpが入力される。
 平滑化振幅スペクトル包絡系列生成部33は、復号線形予測係数^β1,^β2,…,^βpに対応する振幅スペクトル包絡の系列の振幅の凹凸を鈍らせた系列である平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)を上記の式A(3)により生成する(ステップB3)。
 生成された平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)は、復号部34及び包絡逆正規化部35に出力される。
 <復号部34>
 復号部34には、パラメータ復号部37が求めた復号パラメータη、符号化装置が出力した正規化MDCT係数列に対応する符号、非平滑化振幅スペクトル包絡生成部32が生成した非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)及び平滑化振幅スペクトル包絡生成部33が生成した平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)が入力される。
 復号部34は、分散パラメータ決定部342を備えている。
 復号部34は、図15に示すステップB41からステップB44の処理を例えば行うことにより復号を行う(ステップB4)。すなわち、復号部34は、フレームごとに、入力された正規化MDCT係数列に対応する符号に含まれる利得符号を復号してグローバルゲインgを得る(ステップB41)。復号部34の分散パラメータ決定部342は、グローバルゲインgと非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)と平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)とパラメータηとから上記の式(A1)により分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータを求める(ステップB42)。復号部34は、正規化MDCT係数列に対応する符号に含まれる整数信号符号を分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータに対応する算術復号の構成に従い、算術復号して復号正規化済係数系列^XQ(0),^XQ(1),…,^XQ(N-1)を得(ステップB43)、復号正規化済係数系列^XQ(0),^XQ(1),…,^XQ(N-1)の各係数にグローバルゲインgを乗算して復号正規化MDCT係数列^XN(0),^XN(1),…,^XN(N-1)を生成する(ステップB44)。このように、復号部34は、非平滑化スペクトル包絡系列に基づいて実質的に変わるビット割り当てに従って、入力された整数信号符号の復号を行ってもよい。
 なお、[符号化部26の変形例]に記載された処理により符号化が行われた場合には、復号部34は例えば以下の処理を行う。復号部34は、フレームごとに、入力された正規化MDCT係数列に対応する符号に含まれる利得符号を復号してグローバルゲインgを得る。復号部34の分散パラメータ決定部342は、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)と平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)とから上記の式(A9)により分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータを求める。復号部34は、分散パラメータ系列φ(0),φ(1),…,φ(N-1)の各分散パラメータφ(k)に基づいて式(A10)によりb(k)を求めることができ、XQ(k)の値をそのビット数b(k)で順次復号して、復号正規化済係数系列^XQ(0),^XQ(1),…,^XQ(N-1)を得て、復号正規化済係数系列^XQ(0),^XQ(1),…,^XQ(N-1)の各係数にグローバルゲインgを乗算して復号正規化MDCT係数列^XN(0),^XN(1),…,^XN(N-1)を生成する。このように、復号部34は、非平滑化スペクトル包絡系列に基づいて変わるビット割り当てに従って、入力された整数信号符号の復号を行ってもよい。
 生成された復号正規化MDCT係数列^XN(0),^XN(1),…,^XN(N-1)は、包絡逆正規化部35に出力される。
 <包絡逆正規化部35>
 包絡逆正規化部35には、平滑化振幅スペクトル包絡生成部33が生成した平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)及び復号部34が生成した復号正規化MDCT係数列^XN(0),^XN(1),…,^XN(N-1)が入力される。
 包絡逆正規化部35は、平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)を用いて、復号正規化MDCT係数列^XN(0),^XN(1),…,^XN(N-1)を逆正規化することにより、復号MDCT係数列^X(0),^X(1),…,^X(N-1)を生成する(ステップB5)。
 生成された復号MDCT係数列^X(0),^X(1),…,^X(N-1)は、時間領域変換部36に出力される。
 例えば、包絡逆正規化部35は、k=0,1,…,N-1として、復号正規化MDCT係数列^XN(0),^XN(1),…,^XN(N-1)の各係数^XN(k)に、平滑化振幅スペクトル包絡系列^Hγ(0),^Hγ(1),…,^Hγ(N-1)の各包絡値^Hγ(k)を乗じることにより復号MDCT係数列^X(0),^X(1),…,^X(N-1)を生成する。すなわち、k=0,1,…,N-1として、^X(k)=^XN(k)×^Hγ(k)である。
 <時間領域変換部36>
 時間領域変換部36には、包絡逆正規化部35が生成した復号MDCT係数列^X(0),^X(1),…,^X(N-1)が入力される。
 時間領域変換部36は、フレームごとに、包絡逆正規化部35が得た復号MDCT係数列^X(0),^X(1),…,^X(N-1)を時間領域に変換してフレーム単位の音信号(復号音信号)を得る(ステップB6)。
 このようにして、復号装置は、周波数領域での復号により時系列信号を得る。
 [符号化装置、復号装置及びこれらの方法の第二実施形態]
 第一実施形態の符号化装置及び方法は、複数のパラメータηのそれぞれについて符号化を行い符号を生成し、パラメータηごとに生成された符号の中から最適な符号を選択し、選択された符号及び選択された符号に対応するパラメータ符号を出力するものであった。
 これに対して、第二実施形態の符号化装置及び方法は、まずパラメータ決定部27がパラメータηを決定し、決定されたパラメータηに基づいて符号化を行い符号を生成し出力するものである。第二実施形態では、所定の時間区間ごとにパラメータηがパラメータ決定部27により可変とされている。ここで、所定の時間区間ごとにパラメータηが可変とは、所定の時間区間が変わればパラメータηも変わり得ることを意味し、同一の時間区間ではパラメータηの値は変わらないとする。
 以下、第一実施形態と異なる部分を中心に説明する。第一実施形態と同様の部分については重複説明を省略する。
 (符号化)
 第二実施形態の符号化装置の構成例を図16に示す。符号化装置は、図16に示すように、周波数領域変換部21と、線形予測分析部22と、非平滑化振幅スペクトル包絡系列生成部23と、平滑化振幅スペクトル包絡系列生成部24と、包絡正規化部25と、符号化部26と、パラメータ決定部27’とを例えば備えている。この符号化装置により実現される符号化方法の各処理の例を図17に示す。
 以下、図16の各部について説明する。
 <パラメータ決定部27’>
 パラメータ決定部27’には、時系列信号である時間領域の音信号が入力される。音信号の例は、音声ディジタル信号又は音響ディジタル信号である。
 パラメータ決定部27’は、入力された時系列信号に基づいて、後述する処理により、パラメータηを決定する(ステップA7’)。以下、パラメータ決定部27’により決定されたパラメータηをパラメータη1とする。
 パラメータ決定部27’により決定されたη1は、線形予測分析部22、非平滑化振幅スペクトル包絡推定部23、及び平滑化振幅スペクトル包絡推定部24及び符号化部26に出力される。
 また、パラメータ決定部27’は、決定されたη1を符号化することによりパラメータ符号を生成する。生成されたパラメータ符号は、復号装置に送信される。
 パラメータ決定部27’の詳細については後述する。
 周波数領域変換部21、線形予測分析部22、非平滑化振幅スペクトル包絡系列生成部23、平滑化振幅スペクトル包絡系列生成部24、包絡正規化部25及び符号化部26は、パラメータ決定部27が決定したパラメータη1に基づいて、第一実施形態と同様の処理により符号を生成する(ステップA1からステップA6)。この例では、符号は、線形予測係数符号と、利得符号と、整数信号符号とを合わせたものである。生成された符号は、復号装置に送信される。
 パラメータ決定部27’の構成例を図18に示す。パラメータ決定部27’は、図18に示すように、周波数領域変換部41と、スペクトル包絡推定部42と、白色化スペクトル系列生成部43と、パラメータ取得部44とを例えば備えている。スペクトル包絡推定部42は、線形予測分析部421及び非平滑化振幅スペクトル包絡系列生成部422を例えば備えている。例えばこのパラメータ決定部27’により実現されるパラメータ決定方法の各処理の例を図19に示す。
 以下、図18の各部について説明する。
 <周波数領域変換部41>
 周波数領域変換部41には、時系列信号である時間領域の音信号が入力される。音信号の例は、音声ディジタル信号又は音響ディジタル信号である。
 周波数領域変換部41は、所定の時間長のフレーム単位で、入力された時間領域の音信号を周波数領域のN点のMDCT係数列X(0),X(1),…,X(N-1)に変換する。Nは正の整数である。
 得られたMDCT係数列X(0),X(1),…,X(N-1)は、スペクトル包絡推定部42及び白色化スペクトル系列生成部43に出力される。
 特に断りがない限り、以降の処理はフレーム単位で行われるものとする。
 このようにして、周波数領域変換部41は、音信号に対応する、例えばMDCT係数列である周波数領域サンプル列を求める(ステップC41)。
 <スペクトル包絡推定部42>
 スペクトル包絡推定部42には、周波数領域変換部21が得たMDCT係数列X(0),X(1),…,X(N-1)が入力される。
 スペクトル包絡推定部42は、所定の方法で定められるパラメータη0に基づいて、時系列信号に対応する周波数領域サンプル列の絶対値のη0乗をパワースペクトルとして用いたスペクトル包絡の推定を行う(ステップC42)。
 推定されたスペクトル包絡は、白色化スペクトル系列生成部43に出力される。
 スペクトル包絡推定部42は、例えば以下に説明する線形予測分析部421及び非平滑化振幅スペクトル包絡系列生成部422の処理により、非平滑化振幅スペクトル包絡系列を生成することによりスペクトル包絡の推定を行う。
 パラメータη0は所定の方法で定められるとする。例えば、η0を0より大きい所定の数とする。例えば、η0=1とする。また、現在パラメータηを求めようとしているフレームよりも前のフレームで求まったηを用いてもよい。現在パラメータηを求めようとしているフレーム(以下、現フレームとする。)よりも前のフレームとは、例えば現フレームのよりも前のフレームであって現フレームの近傍のフレームである。現フレームの近傍のフレームは、例えば現フレームの直前のフレームである。
 <線形予測分析部421>
 線形予測分析部421には、周波数領域変換部41が得たMDCT係数列X(0),X(1),…,X(N-1)が入力される。
 線形予測分析部421は、MDCT係数列X(0),X(1),…,X(N-1)を用いて、以下の式(C1)により定義される~R(0),~R(1),…,~R(N-1)を用いて線形予測分析を行った線形予測係数β12,…,βpを生成し、生成された線形予測係数β12,…,βpを符号化して線形予測係数符号と線形予測係数符号に対応する量子化された線形予測係数である量子化線形予測係数^β1,^β2,…,^βpとを生成する。
Figure JPOXMLDOC01-appb-M000013
 生成された量子化線形予測係数^β1,^β2,…,^βpは、非平滑化スペクトル包絡系列生成部422に出力される。
 具体的には、線形予測分析部421は、まずMDCT係数列X(0),X(1),…,X(N-1)の絶対値のη0乗をパワースペクトルと見做した逆フーリエ変換に相当する演算、すなわち式(C1)の演算を行うことにより、MDCT係数列X(0),X(1),…,X(N-1)の絶対値のη0乗に対応する時間領域の信号列である擬似相関関数信号列~R(0),~R(1),…,~R(N-1)を求める。そして、線形予測分析部421は、求まった擬似相関関数信号列~R(0),~R(1),…,~R(N-1)を用いて線形予測分析を行って、線形予測係数β12,…,βpを生成する。そして、線形予測分析部421は、生成された線形予測係数β12,…,βpを符号化することにより、線形予測係数符号と、線形予測係数符号に対応する量子化線形予測係数^β1,^β2,…,^βpとを得る。
 線形予測係数β12,…,βpは、MDCT係数列X(0),X(1),…,X(N-1)の絶対値のη0乗をパワースペクトルと見做したときの時間領域の信号に対応する線形予測係数である。
 線形予測分析部421による線形予測係数符号の生成は、例えば従来的な符号化技術によって行われる。従来的な符号化技術とは、例えば、線形予測係数そのものに対応する符号を線形予測係数符号とする符号化技術、線形予測係数をLSPパラメータに変換してLSPパラメータに対応する符号を線形予測係数符号とする符号化技術、線形予測係数をPARCOR係数に変換してPARCOR係数に対応する符号を線形予測係数符号とする符号化技術などである。
 このようにして、線形予測分析部421は、例えばMDCT係数列である周波数領域サンプル列の絶対値のη0乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を生成する(ステップC421)。
 なお、線形予測分析部421は、[線形予測符号化装置、線形予測復号装置及びこれらの方法]の欄で説明した方法により、線形予測係数符号を得て、得られた線形予測係数符号に対応する線形予測係数に変換可能な係数を量子化線形予測係数^β1,^β2,…,^βpとしてもよい。
 <非平滑化振幅スペクトル包絡系列生成部422>
 非平滑化振幅スペクトル包絡系列生成部422には、線形予測分析部421が生成した量子化線形予測係数^β1,^β2,…,^βpが入力される。
 非平滑化振幅スペクトル包絡系列生成部422は、量子化線形予測係数^β1,^β2,…,^βpに対応する振幅スペクトル包絡の系列である非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)を生成する。
 生成された非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)は、白色化スペクトル系列生成部43に出力される。
 非平滑化振幅スペクトル包絡系列生成部422は、量子化線形予測係数^β1,^β2,…,^βpを用いて、非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)として、式(C2)により定義される非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)を生成する。
Figure JPOXMLDOC01-appb-M000014
 このようにして、非平滑化振幅スペクトル包絡系列生成部422は、疑似相関関数信号列に対応する振幅スペクトル包絡の系列を1/η0乗した系列である非平滑化スペクトル包絡系列を線形予測分析部421により生成された線形予測係数に変換可能な係数に基づいて得ることによりスペクトル包絡の推定を行う(ステップC422)。
 <白色化スペクトル系列生成部43>
 白色化スペクトル系列生成部43には、周波数領域変換部41が得たMDCT係数列X(0),X(1),…,X(N-1)及び非平滑化振幅スペクトル包絡生成部422が生成した非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)が入力される。
 白色化スペクトル系列生成部43は、MDCT係数列X(0),X(1),…,X(N-1)の各係数を、対応する非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)の各値で除算することにより、白色化スペクトル系列XW(0),XW(1),…,XW(N-1)を生成する。
 生成された白色化スペクトル系列XW(0),XW(1),…,XW(N-1)は、パラメータ取得部44に出力される。
 白色化スペクトル系列生成部43は、例えば、k=0,1,…,N-1として、MDCT係数列X(0),X(1),…,X(N-1)の各係数X(k)を非平滑化振幅スペクトル包絡系列^H(0),^H(1),…,^H(N-1)の各値^H(k)で除算することにより、白色化スペクトル系列XW(0),XW(1),…,XW(N-1)の各値XW(k)を生成する。すなわち、k=0,1,…,N-1として、XW(k)=X(k)/^H(k)である。
 このようにして、白色化スペクトル系列生成部43は、例えば非平滑化振幅スペクトル包絡系列であるスペクトル包絡で例えばMDCT係数列である周波数領域サンプル列を除算した系列である白色化スペクトル系列を得る(ステップC43)。
 <パラメータ取得部44>
 パラメータ取得部44には、白色化スペクトル系列生成部43が生成した白色化スペクトル系列XW(0),XW(1),…,XW(N-1)が入力される。
 パラメータ取得部44は、パラメータηを形状パラメータとする一般化ガウス分布が白色化スペクトル系列XW(0),XW(1),…,XW(N-1)のヒストグラムを近似するパラメータηを求める(ステップC44)。言い換えれば、パラメータ取得部44は、パラメータηを形状パラメータとする一般化ガウス分布が白色化スペクトル系列XW(0),XW(1),…,XW(N-1)のヒストグラムの分布に近くなるようなパラメータηを決定する。
 パラメータηを形状パラメータとする一般化ガウス分布は、例えば以下のように定義される。Γは、ガンマ関数である。
Figure JPOXMLDOC01-appb-M000015
 一般化ガウス分布は、形状パラメータであるηを変えることにより、図20のようにη=1の時はラプラス分布、η=2の時はガウス分布、といったように様々な分布を表現することができるものである。ηは、0より大きい所定の数である。ηは、0より大きい2以外の所定の数であってもよい。具体的には、ηは、2未満の所定の正の数であってよい。φは分散に対応するパラメータである。
 ここで、パラメータ取得部44が求めるηは、例えば以下の式(C3)により定義される。F-1は、関数Fの逆関数である。この式は、いわゆるモーメント法により導出されるものである。
Figure JPOXMLDOC01-appb-M000016
 逆関数F-1が定式化されている場合には、パラメータ取得部44は、定式化された逆関数F-1にm1/((m2)1/2)の値を入力したときの出力値を計算することによりパラメータηを求めることができる。
 逆関数F-1が定式化されていない場合には、パラメータ取得部44は、式(C3)で定義されるηの値を計算するために、例えば以下に説明する第一方法又は第二方法によりパラメータηを求めてもよい。
 パラメータηを求めるための第一方法について説明する。第一の方法では、パラメータ取得部44は、白色化スペクトル系列に基づいてm1/((m2)1/2)を計算し、予め用意しておいた異なる複数の、ηと対応するF(η)のペアを参照して、計算されたm1/((m2)1/2)に最も近いF(η)に対応するηを取得する。
 予め用意しておいた異なる複数の、ηと対応するF(η)のペアは、パラメータ取得部44の記憶部441に予め記憶しておく。パラメータ取得部44は、記憶部441参照して、計算されたm1/((m2)1/2)に最も近いF(η)を見つけ、見つかったF(η)に対応するηを記憶部441から読み込み出力する。
 計算されたm1/((m2)1/2)に最も近いF(η)とは、計算されたm1/((m2)1/2)との差の絶対値が最も小さくなるF(η)のことである。
 パラメータηを求めるための第二方法について説明する。第二の方法では、逆関数F-1の近似曲線関数を例えば以下の式(C3’)で表される~F-1として、パラメータ取得部44は、白色化スペクトル系列に基づいてm1/((m2)1/2)を計算し、近似曲線関数~F-1に計算されたm1/((m2)1/2)を入力したときの出力値を計算することによりηを求める。この近似曲線関数~F-1は使用する定義域において出力が正値となる単調増加関数であればよい。
Figure JPOXMLDOC01-appb-M000017
 なお、パラメータ取得部44が求めるηは、式(C3)ではなく、式(C3'')のように予め定めた正の整数q1及びq2を用いて(ただしq1<q2)式(C3)を一般化した式により定義されてもよい。
Figure JPOXMLDOC01-appb-M000018
 なお、ηが式(C3'')により定義される場合も、ηが式(C3)により定義されている場合と同様の方法により、ηを求めることができる。すなわち、パラメータ取得部44が、白色化スペクトル系列に基づいてそのq1次モーメントであるmq1とそのq2次モーメントであるmq2とに基づく値mq1/((mq2)q1/q2)を計算した後、例えば上記の第一及び第二の方法と同様、予め用意しておいた異なる複数の、ηと対応するF’(η)のペアを参照して、計算されたmq1/((mq2)q1/q2)に最も近いF’(η)に対応するηを取得するか、逆関数F’-1の近似曲線関数を~F’-1として、近似曲線関数~F-1に計算されたmq1/((mq2)q1/q2)を入力したときの出力値を計算してηを求めることができる。
 このようにηは次元が異なる2つの異なるモーメントmq1,mq2に基づく値であるとも言える。例えば、次元が異なる2つの異なるモーメントmq1,mq2のうち、次元が低い方のモーメントの値又はこれに基づく値(以下、前者とする。)と次元が高い方のモーメントの値又はこれに基づく値(以下、後者とする)との比の値、この比の値に基づく値、又は、前者を後者で割って得られる値に基づき、ηを求めてもよい。モーメントに基づく値とは、例えば、そのモーメントをmとしQを所定の実数としてmQのことである。また、これらの値を近似曲線関数~F-1に入力してηを求めてもよい。この近似曲線関数~F’-1は上記同様、使用する定義域において出力が正値となる単調増加関数であればよい。
 パラメータ決定部27’は、ループ処理によりパラメータηを求めてもよい。すなわち、パラメータ決定部27’は、パラメータ取得部44で求まるパラメータηを所定の方法で定められるパラメータη0とする、スペクトル包絡推定部42、白色化スペクトル系列生成部43及びパラメータ取得部44の処理を更に1回以上行ってもよい。
 この場合、例えば、図18で破線で示すように、パラメータ取得部44で求まったパラメータηは、スペクトル包絡推定部42に出力される。スペクトル包絡推定部42は、パラメータ取得部44で求まったηをパラメータη0として用いて、上記説明した処理と同様の処理を行いスペクトル包絡の推定を行う。白色化スペクトル系列生成部43は、新たに推定されたスペクトル包絡に基づいて、上記説明した処理と同様の処理を行い白色化スペクトル系列を生成する。パラメータ取得部44は、新たに生成された白色化スペクトル系列に基づいて、上記説明した処理と同様の処理を行いパラメータηを求める。
 例えば、スペクトル包絡推定部42、白色化スペクトル系列生成部43及びパラメータ取得部44の処理は、所定の回数であるτ回だけ更に行われてもよい。τは所定の正の整数であり、例えばτ=1又はτ=2である。
 また、スペクトル包絡推定部42は、今回求まったパラメータηと前回求まったパラメータηとの差の絶対値が所定の閾値以下となるまで、スペクトル包絡推定部42、白色化スペクトル系列生成部43及びパラメータ取得部44の処理を繰り返してもよい。
 (復号)
 第二実施形態の復号装置及び方法は、第一実施形態と同様であるため重複説明を省略する。
 [符号化装置、復号装置及びこれらの方法の変形例]
 線形予測分析部22及び非平滑化振幅スペクトル包絡系列生成部23を1つのスペクトル包絡推定部2Aとして捉えると、このスペクトル包絡推定部2Aは、時系列信号に対応する例えばMDCT係数列である周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做したスペクトル包絡(非平滑化振幅スペクトル包絡系列)の推定を行っていると言える。ここで、「パワースペクトルと見做した」とは、パワースペクトルを通常用いるところに、η1乗のスペクトルを用いることを意味する。
 この場合、スペクトル包絡推定部2Aの線形予測分析部22は、例えばMDCT係数列である周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得ていると言える。また、スペクトル包絡推定部2Aの非平滑化振幅スペクトル包絡系列生成部23は、線形予測分析部22により得られた線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列を得ることによりスペクトル包絡の推定を行っていると言える。
 また、平滑化振幅スペクトル包絡系列生成部24、包絡正規化部25及び符号化部26を1つの符号化部2Bとして捉えると、この符号化部2Bは、スペクトル包絡推定部2Aにより推定されたスペクトル包絡(非平滑化振幅スペクトル包絡系列)を基にビット割り当てを変える又は実質的にビット割り当てが変わる符号化を時系列信号に対応する例えばMDCT係数列である周波数領域サンプル列の各係数に対して行っていると言える。
 復号部34及び包絡逆正規化部35を1つの復号部3Aとして捉えると、この復号部3Aは、非平滑化スペクトル包絡系列に基づいて変わるビット割り当て又は実質的に変わるビット割り当てに従って、入力された整数信号符号の復号を行うことにより時系列信号に対応する周波数領域サンプル列を得ていると言える。
 符号化部2Bは、スペクトル包絡(非平滑化振幅スペクトル包絡系列)を基にビット割り当てを変える又は実質的にビット割り当てが変わる符号化を行うのであれば、上記説明した算術符号化以外の符号化処理を行ってもよい。この場合、復号部3Aは、符号化部2Bが行った符号化処理に対応する復号処理を行う。
 例えば、符号化部2Bは、スペクトル包絡(非平滑化振幅スペクトル包絡系列)に基づいて決定されたRiceパラメータを用いて周波数領域サンプル列をGolomb-Rice符号化してもよい。この場合、復号部3Aは、スペクトル包絡(非平滑化振幅スペクトル包絡系列)に基づいて決定されたRiceパラメータを用いてGolomb-Rice復号してもよい。
 第一実施形態において、符号化装置は、パラメータηを決定する際に符号化処理を最後まで行わなくてもよい。言い換えれば、パラメータ決定部27は、推定符号量に基づいてパラメータηを決定してもよい。この場合、符号化部2Bは、複数のパラメータηのそれぞれを用いて同一の所定の時間区間の時系列信号に対応する周波数領域サンプル列に対する上記と同様の符号化処理により得られる符号の推定符号量を得る。パラメータ決定部27は、得られた推定符号量に基づいて複数のパラメータηの何れか1つを選択する。例えば、推定符号量が最も小さいパラメータηを選択する。符号化部2Bは、選択されたパラメータηを用いて上記と同様の符号化処理を行うことにより符号を得て出力する。
 上記説明した処理は、記載の順にしたがって時系列に実行されるのみならず、処理を実行する装置の処理能力あるいは必要に応じて並列的にあるいは個別に実行されてもよい。
 [プログラム及び記録媒体]
 また、各装置又は各方法における各部をコンピュータによって実現してもよい。その場合、各装置又は各方法の処理内容はプログラムによって記述される。そして、このプログラムをコンピュータで実行することにより、各装置又は各方法における各部がコンピュータ上で実現される。
 この処理内容を記述したプログラムは、コンピュータで読み取り可能な記録媒体に記録しておくことができる。コンピュータで読み取り可能な記録媒体としては、例えば、磁気記録装置、光ディスク、光磁気記録媒体、半導体メモリ等どのようなものでもよい。
 また、このプログラムの流通は、例えば、そのプログラムを記録したDVD、CD-ROM等の可搬型記録媒体を販売、譲渡、貸与等することによって行う。さらに、このプログラムをサーバコンピュータの記憶装置に格納しておき、ネットワークを介して、サーバコンピュータから他のコンピュータにそのプログラムを転送することにより、このプログラムを流通させてもよい。
 このようなプログラムを実行するコンピュータは、例えば、まず、可搬型記録媒体に記録されたプログラムもしくはサーバコンピュータから転送されたプログラムを、一旦、自己の記憶部に格納する。そして、処理の実行時、このコンピュータは、自己の記憶部に格納されたプログラムを読み取り、読み取ったプログラムに従った処理を実行する。また、このプログラムの別の実施形態として、コンピュータが可搬型記録媒体から直接プログラムを読み取り、そのプログラムに従った処理を実行することとしてもよい。さらに、このコンピュータにサーバコンピュータからプログラムが転送されるたびに、逐次、受け取ったプログラムに従った処理を実行することとしてもよい。また、サーバコンピュータから、このコンピュータへのプログラムの転送は行わず、その実行指示と結果取得のみによって処理機能を実現する、いわゆるASP(Application Service Provider)型のサービスによって、上述の処理を実行する構成としてもよい。なお、プログラムには、電子計算機による処理の用に供する情報であってプログラムに準ずるもの(コンピュータに対する直接の指令ではないがコンピュータの処理を規定する性質を有するデータ等)を含むものとする。
 また、コンピュータ上で所定のプログラムを実行させることにより、各装置を構成することとしたが、これらの処理内容の少なくとも一部をハードウェア的に実現することとしてもよい。

Claims (30)

  1.  パラメータηを正の数として、時系列信号に対応するパラメータηを、その時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做すことにより推定されたスペクトル包絡で上記周波数領域サンプル列を除算した系列である白色化スペクトル系列のヒストグラムを近似する一般化ガウス分布の形状パラメータとし、η1はパラメータηの所定の値であるとして、
     時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析部と、
     N種類(Nは1以上の整数)のパラメータηのそれぞれに対応するN個の符号帳が記憶され、各符号帳にはそれぞれのパラメータηに対応する線形予測係数に変換可能な係数の候補が複数個格納された符号帳記憶部と、
     上記符号帳記憶部に記憶された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、上記線形予測分析部が得た線形予測係数に変換可能な係数と、のηの値を適合させる適合部と、
     上記ηの値が適合された線形予測係数に変換可能な係数の複数個の候補と線形予測係数に変換可能な係数とを用いて、上記線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る符号化部と、
     を含む線形予測符号化装置。
  2.  請求項1の線形予測符号化装置であって、
     上記適合部は、上記符号帳記憶部に記憶された線形予測係数に変換可能な係数の候補に対して、η1に応じた第一線形変換を行い、第一線形変換後の線形予測係数に変換可能な係数の複数個の候補を得る線形変換部を含み、
     上記符号化部は、上記線形予測分析部が得た線形予測係数に変換可能な係数と、上記適合部が得た上記第一線形変換後の線形予測係数に変換可能な係数の複数個の候補と、を用いて、上記線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る、
     線形予測符号化装置。
  3.  請求項1の線形予測符号化装置であって、
     上記適合部は、上記線形予測分析部が得た線形予測係数に変換可能な係数に対して、η1に応じた第二線形変換を行い、第二線形変換後の線形予測係数に変換可能な係数を得る線形変換部を含み、
     上記符号化部は、上記適合部が得た上記第二線形変換後の線形予測係数に変換可能な係数と、上記符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、を用いて、上記線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る、
     線形予測符号化装置。
  4.  請求項1の線形予測符号化装置であって、
     η23はパラメータηの所定の値であるとして、
     上記符号帳記憶部には、η2に対応する符号帳が記憶されており、
     上記適合部は、上記符号帳記憶部に記憶された線形予測係数に変換可能な係数の複数個の候補に対して、η3に応じた第一線形変換を行い、第一線形変換後の線形予測係数に変換可能な係数の複数個の候補を得、上記線形予測分析部が得た線形予測係数に変換可能な係数に対して、η3に応じた第二線形変換を行い、第二線形変換後の線形予測係数に変換可能な係数を得る、線形変換部であり、
     上記符号化部は、上記適合部が得た上記第二線形変換後の線形予測係数に変換可能な係数と、上記適合部が得た上記第一線形変換後の線形予測係数に変換可能な係数の複数個の候補と、を用いて、上記線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る、
     線形予測符号化装置。
  5.  請求項1の線形予測符号化装置であって、
     η2はパラメータηの所定の値であるとして、
     上記符号帳記憶部には、複数の符号帳が記憶されており、
     上記適合部は、上記符号帳記憶部に記憶された複数の符号帳の中から上記η2に応じて符号帳を選択する符号帳選択部と、上記線形予測分析部で得られた線形予測係数に変換可能な係数に対する、η2に応じた第二線形変換を行う線形変換部とであり、
     上記符号化部は、上記第二線形変換後の線形予測係数に変換可能な係数について、上記選択された符号帳を用いて符号化して線形予測係数符号を得る、
     線形予測符号化装置。
  6.  請求項1の線形予測符号化装置であって、
     η2はパラメータηの所定の値であるとして、
     上記符号帳記憶部には、複数の符号帳が記憶されており、
     上記適合部は、上記符号帳記憶部に記憶された複数の符号帳の中から上記η2に応じて符号帳を選択する符号帳選択部と、上記選択された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補に対する、η1に応じた第一線形変換を行う線形変換部とであり、
     上記符号化部は、上記線形予測分析部で得られた線形予測係数に変換可能な係数について、上記第一線形変換後の線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る、
     線形予測符号化装置。
  7.  請求項1の線形予測符号化装置であって、
     η23はパラメータηの所定の値であるとして、
     上記符号帳記憶部には、複数の符号帳が記憶されており、
     上記適合部は、上記符号帳記憶部に記憶された複数の符号帳の中から上記η3に応じて符号帳を選択する符号帳選択部と、上記選択された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補に対する、η2に応じた第一線形変換を行うと共に、上記線形予測分析部で得られた線形予測係数に変換可能な係数に対する、η2に応じた第二線形変換を行う線形変換部とであり、
     上記符号化部は、上記第二線形変換後の線形予測係数に変換可能な係数について、上記第一線形変換後の線形予測係数に変換可能な係数の候補を用いて符号化して線形予測係数符号を得る、
     線形予測符号化装置。
  8.  請求項2の線形予測符号化装置において、
     上記線形変換部は、上記η1が小さいほど上記第一線形変換後の線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列が平坦になるように上記第一線形変換を行う、
     線形予測符号化装置。
  9.  請求項2から8の何れかの線形予測符号化装置において、
     pを線形予測係数に変換可能な係数の次数とし、上記線形予測係数に変換可能な係数又は上記線形予測係数に変換可能な係数の候補を^ω[k][k=1,2,…,p]とし、上記第一線形変換後及び上記第二線形変換後の線形予測係数に変換可能な係数又は上記線形予測係数に変換可能な係数の候補を~ω[k][k=1,2,…,p]とし、x1,x2,…xp,y1,y2,…yp-1,z2,z3,…zpを所定の非負の数とし、y1,y2,…yp-1,z2,z3,…zpの少なくとも1つは所定の正の数であるとし、Kをx1,x2,…xp,y1,y2,…yp-1,z2,z3,…zp以外の要素が0である行列として、
     上記線形変換部は、下記式により上記第一線形変換と上記第二線形変換との少なくとも一方を行う、
    Figure JPOXMLDOC01-appb-M000001

     線形予測符号化装置。
  10.  請求項2の線形予測符号化装置において、
     上記線形変換部は、上記η1が小さいほど上記第一線形変換後の線形予測係数に変換可能な係数の候補の次数が小さくなるように上記第一線形変換を行う、
     線形予測符号化装置。
  11.  請求項1,2,3,4の何れかの線形予測符号化装置であって、
     上記符号帳記憶部には、複数の符号帳が記憶されており、
     上記適合部は、上記符号帳記憶部に記憶された複数の符号帳の中から上記η1に応じて符号帳を選択する符号帳選択部を含み、
     上記符号化部は、上記線形予測分析部が得た線形予測係数に変換可能な係数と、上記適合部が得た線形予測係数に変換可能な係数の複数個の候補と、を用いて、上記線形予測分析部が得た上記線形予測係数に変換可能な係数に対応する線形予測係数符号を得る、
     線形予測符号化装置
  12.  請求項11の線形予測符号化装置において、
     上記符号帳記憶部には、線形予測係数に変換可能な係数の候補数が異なる複数の符号帳が記憶されており、
     上記符号帳選択部は、上記η1が大きいほど、上記符号帳記憶部に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補数が多い符号帳を選択する、
     線形予測符号化装置。
  13.  請求項11又は12の線形予測符号化装置において、
     上記符号帳記憶部には、符号帳に記憶された線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列の平坦度合いが異なる複数の符号帳が記憶されており、
     上記符号帳選択部は、上記η1が小さいほど、上記符号帳記憶部に記憶された複数の符号帳の中から、符号帳に記憶された線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列がより平坦である符号帳を選択する、
     線形予測符号化装置。
  14.  請求項11又は12の線形予測符号化装置において、
     上記符号帳記憶部には、線形予測係数に変換可能な係数の候補間の間隔が異なる複数の符号帳が記憶されており、
     上記符号帳選択部は、上記ηが小さいほど、上記符号帳記憶部に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補間の間隔が狭い符号帳を選択する、
     線形予測符号化装置。
  15.  パラメータηを正の数として、時系列信号に対応するパラメータηを、その時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做すことにより推定されたスペクトル包絡で上記周波数領域サンプル列を除算した系列である白色化スペクトル系列のヒストグラムを近似する一般化ガウス分布の形状パラメータとし、η1はパラメータηの所定の値であるとして、
     時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析部と、
     符号帳が記憶された符号帳記憶部と、
     入力されたη1に基づいて、上記符号帳記憶部に記憶された符号帳と上記線形予測係数に変換可能な係数との少なくとも一方を適合させる適合部と、
     上記符号帳又は上記適合された符号帳を用いて、上記線形予測係数に変換可能な係数又は上記適合された線形予測係数に変換可能な係数を符号化する符号化部と、
     を含む線形予測符号化装置。
  16.  符号帳が記憶された符号帳記憶部と、
     η1を正の数として、入力されたη1に基づいて、上記符号帳記憶部に記憶された符号帳と、上記符号帳に格納された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補との少なくとも一方を適合させる適合部を含み、
     上記線形予測係数に変換可能な係数は、上記線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列を得るために用いられる、
     線形予測復号装置。
  17.  請求項16の線形予測復号装置であって、
     上記符号帳に格納された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補を線形予測係数に変換可能な係数として得る復号部を更に含み、
     上記適合部は、上記復号部で得られた線形予測係数に変換可能な係数に対して、所定の正の数であるη1に応じた線形変換をして線形予測係数に変換可能な係数を得る線形変換部である、
     線形予測復号装置。
  18.  請求項16の線形予測復号装置であって、
     上記符号帳には、複数の符号帳が記憶されており、
     η2を正の数として、上記適合部は、上記符号帳記憶部に記憶された複数の符号帳の中から上記ηに応じて符号帳を選択する符号帳選択部と、復号部で得られた線形予測係数に変換可能な係数に対して、所定の正の数であるη1に応じた線形変換をして線形予測係数に変換可能な係数を得る線形変換部とであり、
     上記選択された符号帳に格納された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補を線形予測係数に変換可能な係数として得る上記復号部を更に含む、
     線形予測復号装置。
  19.  請求項17の線形予測復号装置において、
     上記線形変換部は、上記η1が小さいほど上記線形変換部で得られた線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列が平坦になるように上記線形変換を行う、
     線形予測復号装置。
  20.  請求項17から19の何れかの線形予測復号装置において、
     pを線形予測係数に変換可能な係数の次数とし、上記復号部で得られた線形予測係数に変換可能な係数を^ω[k][k=1,2,…,p]とし、上記線形変換後の線形予測係数に変換可能な係数を~ω[k][k=1,2,…,p]とし、x1,x2,…xp,y1,y2,…yp-1,z2,z3,…zpを所定の非負の数とし、y1,y2,…yp-1,z2,z3,…zpの少なくとも1つは所定の正の数であるとし、Kをx1,x2,…xp,y1,y2,…yp-1,z2,z3,…zp以外の要素が0である行列として、
     上記線形変換部は、下記式により線形変換を行う、

     線形予測復号装置。
  21.  請求項17の線形予測復号装置において、
     上記線形変換部は、上記η1が小さいほど上記線形変換後の線形予測係数に変換可能な係数の次数が小さくなるように上記線形変換を行う、
     線形予測復号装置。
  22.  請求項16の線形予測復号装置であって、
     上記符号帳には、複数の符号帳が記憶されており、
     上記適合部は、上記符号帳記憶部に記憶された複数の符号帳の中から上記η1に応じて符号帳を選択する符号帳選択部であり、上記選択された符号帳を用いて、入力された線形予測係数符号を復号して線形予測係数に変換可能な係数を得る復号部を更に含む
     線形予測復号装置。
  23.  請求項22の線形予測復号装置において、
     上記符号帳記憶部には、線形予測係数に変換可能な係数の候補数が異なる複数の符号帳が記憶されており、
     上記符号帳選択部は、上記η1が大きいほど、上記符号帳記憶部に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補数が多い符号帳を選択する、
     線形予測復号装置。
  24.  請求項22又は23の線形予測復号装置において、
     上記符号帳記憶部には、符号帳に記憶された線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列の平坦度合いが異なる複数の符号帳が記憶されており、
     上記符号帳選択部は、上記η1が小さいほど、上記符号帳記憶部に記憶された複数の符号帳の中から、符号帳に記憶された線形予測係数に変換可能な係数の候補に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列が平坦である符号帳を選択する、
     線形予測復号装置。
  25.  請求項22又は23の線形予測復号装置において、
     上記符号帳記憶部には、線形予測係数に変換可能な係数の候補間の間隔が異なる複数の符号帳が記憶されており、
     上記符号帳選択部は、上記η1が小さいほど、上記符号帳記憶部に記憶された複数の符号帳の中から、線形予測係数に変換可能な係数の候補間の間隔が狭い符号帳を選択する、
     線形予測復号装置。
  26.  パラメータηを正の数として、時系列信号に対応するパラメータηを、その時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做すことにより推定されたスペクトル包絡で上記周波数領域サンプル列を除算した系列である白色化スペクトル系列のヒストグラムを近似する一般化ガウス分布の形状パラメータとし、η1はパラメータηの所定の値であるとして、
     線形予測分析部が、時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析ステップと、
     適合部が、N種類(Nは1以上の整数)のパラメータηのそれぞれに対応するN個の符号帳が記憶され、各符号帳にはそれぞれのパラメータηに対応する線形予測係数に変換可能な係数の候補が複数個格納された符号帳記憶部に記憶された符号帳に格納された線形予測係数に変換可能な係数の複数個の候補と、上記線形予測分析ステップが得た線形予測係数に変換可能な係数と、のηの値を適合させる適合ステップと、
     符号化部が、上記ηの値が適合された線形予測係数に変換可能な係数の複数個の候補と線形予測係数に変換可能な係数とを用いて、上記線形予測分析部が得た線形予測係数に変換可能な係数に対応する線形予測係数符号を得る符号化ステップと、
     を含む線形予測符号化方法。
  27.  パラメータηを正の数として、時系列信号に対応するパラメータηを、その時系列信号に対応する周波数領域サンプル列の絶対値のη乗をパワースペクトルと見做すことにより推定されたスペクトル包絡で上記周波数領域サンプル列を除算した系列である白色化スペクトル系列のヒストグラムを近似する一般化ガウス分布の形状パラメータとし、η1はパラメータηの所定の値であるとして、
     時系列信号に対応する周波数領域サンプル列の絶対値のη1乗をパワースペクトルと見做した逆フーリエ変換を行うことにより得られる疑似相関関数信号列を用いて線形予測分析を行い線形予測係数に変換可能な係数を得る線形予測分析ステップ、
     入力されたη1に基づいて、符号帳記憶部に記憶された符号帳と上記線形予測係数に変換可能な係数との少なくとも一方を適合させる適合ステップと、
     上記符号帳又は上記適合された符号帳を用いて、上記線形予測係数に変換可能な係数又は上記適合された線形予測係数に変換可能な係数を符号化する符号化ステップと、
     を含む線形予測符号化方法。
  28.  η1を正の数として、入力されたη1に基づいて、符号帳記憶部に記憶された符号帳と、上記符号帳に格納された複数個の線形予測係数に変換可能な係数の候補のうち、入力された線形予測係数符号に対応する線形予測係数に変換可能な係数の候補との少なくとも一方を適合させる適合ステップを含み、
     上記線形予測係数に変換可能な係数は、上記線形予測係数に変換可能な係数に対応する振幅スペクトル包絡の系列を1/η1乗した系列である非平滑化スペクトル包絡系列を得るために用いられる、
     線形予測復号方法。
  29.  請求項1から15の何れかの線形予測符号化装置又は請求項16から25の何れかの線形予測復号装置の各部としてコンピュータを機能させるためのプログラム。
  30.  請求項1から15の何れかの線形予測符号化装置又は請求項16から25の何れかの線形予測復号装置の各部としてコンピュータを機能させるためのプログラムが記録されたコンピュータ読み取り可能な記録媒体。
PCT/JP2016/061682 2015-04-13 2016-04-11 線形予測符号化装置、線形予測復号装置、これらの方法、プログラム及び記録媒体 Ceased WO2016167215A1 (ja)

Priority Applications (5)

Application Number Priority Date Filing Date Title
US15/562,689 US10325609B2 (en) 2015-04-13 2016-04-11 Coding and decoding a sound signal by adapting coefficients transformable to linear predictive coefficients and/or adapting a code book
KR1020177028710A KR102061300B1 (ko) 2015-04-13 2016-04-11 선형 예측 부호화 장치, 선형 예측 복호 장치, 이들의 방법, 프로그램 및 기록 매체
EP16780006.9A EP3270376B1 (en) 2015-04-13 2016-04-11 Sound signal linear predictive coding
JP2017512523A JP6517924B2 (ja) 2015-04-13 2016-04-11 線形予測符号化装置、方法、プログラム及び記録媒体
CN201680021332.5A CN107408390B (zh) 2015-04-13 2016-04-11 线性预测编码装置、线性预测解码装置、它们的方法以及记录介质

Applications Claiming Priority (4)

Application Number Priority Date Filing Date Title
JP2015-081747 2015-04-13
JP2015081747 2015-04-13
JP2015081746 2015-04-13
JP2015-081746 2015-04-13

Publications (1)

Publication Number Publication Date
WO2016167215A1 true WO2016167215A1 (ja) 2016-10-20

Family

ID=57126589

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/JP2016/061682 Ceased WO2016167215A1 (ja) 2015-04-13 2016-04-11 線形予測符号化装置、線形予測復号装置、これらの方法、プログラム及び記録媒体

Country Status (6)

Country Link
US (1) US10325609B2 (ja)
EP (1) EP3270376B1 (ja)
JP (2) JP6517924B2 (ja)
KR (1) KR102061300B1 (ja)
CN (1) CN107408390B (ja)
WO (1) WO2016167215A1 (ja)

Families Citing this family (6)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP6387117B2 (ja) * 2015-01-30 2018-09-05 日本電信電話株式会社 符号化装置、復号装置、これらの方法、プログラム及び記録媒体
JP6499206B2 (ja) * 2015-01-30 2019-04-10 日本電信電話株式会社 パラメータ決定装置、方法、プログラム及び記録媒体
CN112350760B (zh) * 2019-08-09 2021-07-23 大唐移动通信设备有限公司 一种预编码码本选择的方法及装置
KR20210133554A (ko) * 2020-04-29 2021-11-08 한국전자통신연구원 선형 예측 코딩을 이용한 오디오 신호의 부호화 및 복호화 방법과 이를 수행하는 부호화기 및 복호화기
CN111901004B (zh) * 2020-08-04 2022-04-12 三维通信股份有限公司 平坦度的补偿方法和装置、存储介质和电子设备
US12525248B2 (en) * 2022-08-11 2026-01-13 Electronics And Telecommunications Research Institute Apparatus for encoding and decoding audio signal and method of operation thereof

Citations (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
WO2013176177A1 (ja) * 2012-05-23 2013-11-28 日本電信電話株式会社 符号化方法、復号方法、符号化装置、復号装置、プログラム、および記録媒体

Family Cites Families (25)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPS6253028A (ja) 1985-09-02 1987-03-07 Nec Corp 適応形符号化復号化方式とその装置
JP3186013B2 (ja) * 1995-01-13 2001-07-11 日本電信電話株式会社 音響信号変換符号化方法及びその復号化方法
GB2326572A (en) * 1997-06-19 1998-12-23 Softsound Limited Low bit rate audio coder and decoder
US6453289B1 (en) * 1998-07-24 2002-09-17 Hughes Electronics Corporation Method of noise reduction for speech codecs
CA2733453C (en) * 2000-11-30 2014-10-14 Panasonic Corporation Lpc vector quantization apparatus
JP4365610B2 (ja) * 2003-03-31 2009-11-18 パナソニック株式会社 音声復号化装置および音声復号化方法
CN101556800B (zh) * 2003-10-23 2012-05-23 松下电器产业株式会社 音频频谱编解码方法和装置、声音信号发送和接收装置
JP4493030B2 (ja) * 2005-10-12 2010-06-30 月島機械株式会社 ろ過装置
EP1988544B1 (en) 2006-03-10 2014-12-24 Panasonic Intellectual Property Corporation of America Coding device and coding method
ATE518224T1 (de) * 2008-01-04 2011-08-15 Dolby Int Ab Audiokodierer und -dekodierer
CN102449689B (zh) * 2009-06-03 2014-08-06 日本电信电话株式会社 编码方法、编码装置、编码程序、以及它们的记录介质
RU2510974C2 (ru) * 2010-01-08 2014-04-10 Ниппон Телеграф Энд Телефон Корпорейшн Способ кодирования, способ декодирования, устройство кодера, устройство декодера, программа и носитель записи
JP5602769B2 (ja) * 2010-01-14 2014-10-08 パナソニック インテレクチュアル プロパティ コーポレーション オブ アメリカ 符号化装置、復号装置、符号化方法及び復号方法
FR2961937A1 (fr) * 2010-06-29 2011-12-30 France Telecom Codage/decodage predictif lineaire adaptatif
JP2012163919A (ja) * 2011-02-09 2012-08-30 Sony Corp 音声信号処理装置、および音声信号処理方法、並びにプログラム
CN103620675B (zh) * 2011-04-21 2015-12-23 三星电子株式会社 对线性预测编码系数进行量化的设备、声音编码设备、对线性预测编码系数进行反量化的设备、声音解码设备及其电子装置
JP5958866B2 (ja) * 2012-08-01 2016-08-02 国立研究開発法人産業技術総合研究所 音声分析合成のためのスペクトル包絡及び群遅延の推定システム及び音声信号の合成システム
PL3525208T3 (pl) * 2012-10-01 2021-12-13 Nippon Telegraph And Telephone Corporation Sposób kodowania, koder, program i nośnik zapisu
FR3011408A1 (fr) * 2013-09-30 2015-04-03 Orange Re-echantillonnage d'un signal audio pour un codage/decodage a bas retard
CN103824561B (zh) * 2014-02-18 2015-03-11 北京邮电大学 一种语音线性预测编码模型的缺失值非线性估算方法
WO2015162979A1 (ja) * 2014-04-24 2015-10-29 日本電信電話株式会社 周波数領域パラメータ列生成方法、符号化方法、復号方法、周波数領域パラメータ列生成装置、符号化装置、復号装置、プログラム及び記録媒体
EP3226243B1 (en) * 2014-11-27 2022-01-05 Nippon Telegraph and Telephone Corporation Encoding apparatus, decoding apparatus, and method and program for the same
JP6499206B2 (ja) * 2015-01-30 2019-04-10 日本電信電話株式会社 パラメータ決定装置、方法、プログラム及び記録媒体
JP6387117B2 (ja) * 2015-01-30 2018-09-05 日本電信電話株式会社 符号化装置、復号装置、これらの方法、プログラム及び記録媒体
CN107851442B (zh) * 2015-04-13 2021-07-20 日本电信电话株式会社 匹配装置、判定装置、它们的方法、程序及记录介质

Patent Citations (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
WO2013176177A1 (ja) * 2012-05-23 2013-11-28 日本電信電話株式会社 符号化方法、復号方法、符号化装置、復号装置、プログラム、および記録媒体

Non-Patent Citations (2)

* Cited by examiner, † Cited by third party
Title
JUHA OJANPERA ET AL.: "Long Term Predictor for Transform Domain Perceptual Audio Coding", PROC. 107TH CONVENTION OF AES, 24 September 1999 (1999-09-24), pages 1 - 25 *
See also references of EP3270376A4 *

Also Published As

Publication number Publication date
JP2019079069A (ja) 2019-05-23
CN107408390B (zh) 2021-08-06
JPWO2016167215A1 (ja) 2018-02-01
KR102061300B1 (ko) 2020-02-11
EP3270376A1 (en) 2018-01-17
US10325609B2 (en) 2019-06-18
US20180096694A1 (en) 2018-04-05
CN107408390A (zh) 2017-11-28
EP3270376A4 (en) 2018-08-29
JP6517924B2 (ja) 2019-05-22
KR20170127533A (ko) 2017-11-21
EP3270376B1 (en) 2020-03-18
JP6633787B2 (ja) 2020-01-22

Similar Documents

Publication Publication Date Title
JP6633787B2 (ja) 線形予測復号装置、方法、プログラム及び記録媒体
JP6422813B2 (ja) 符号化装置、復号装置、これらの方法及びプログラム
JP6484358B2 (ja) 符号化装置、及びその方法、プログラム、記録媒体
JP6457552B2 (ja) 符号化装置、復号装置、これらの方法及びプログラム
JP6499206B2 (ja) パラメータ決定装置、方法、プログラム及び記録媒体
JP6387117B2 (ja) 符号化装置、復号装置、これらの方法、プログラム及び記録媒体
CN106233383B (zh) 频域参数串生成方法、频域参数串生成装置以及记录介质
CN110534122B (zh) 解码装置、及其方法、记录介质

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 16780006

Country of ref document: EP

Kind code of ref document: A1

WWE Wipo information: entry into national phase

Ref document number: 15562689

Country of ref document: US

ENP Entry into the national phase

Ref document number: 2017512523

Country of ref document: JP

Kind code of ref document: A

ENP Entry into the national phase

Ref document number: 20177028710

Country of ref document: KR

Kind code of ref document: A

REEP Request for entry into the european phase

Ref document number: 2016780006

Country of ref document: EP

NENP Non-entry into the national phase

Ref country code: DE