PL337717A1 - Method of and apparatus for improving speech in a voice communication system - Google Patents

Method of and apparatus for improving speech in a voice communication system

Info

Publication number
PL337717A1
PL337717A1 PL98337717A PL33771798A PL337717A1 PL 337717 A1 PL337717 A1 PL 337717A1 PL 98337717 A PL98337717 A PL 98337717A PL 33771798 A PL33771798 A PL 33771798A PL 337717 A1 PL337717 A1 PL 337717A1
Authority
PL
Poland
Prior art keywords
speech
unit
determines
intelligible
listener
Prior art date
Application number
PL98337717A
Inventor
Robert James Chance
Ian Vince Mcloughlin
Original Assignee
Simoco International Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Simoco International Ltd filed Critical Simoco International Ltd
Publication of PL337717A1 publication Critical patent/PL337717A1/en

Links

Classifications

    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0208Noise filtering
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/003Changing voice quality, e.g. pitch or formants
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02Speech enhancement, e.g. noise reduction or echo cancellation
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0316Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude
    • G10L21/0364Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude for improving intelligibility
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/003Changing voice quality, e.g. pitch or formants
    • G10L21/007Changing voice quality, e.g. pitch or formants characterised by the process used
    • G10L21/013Adapting to target pitch
    • G10L2021/0135Voice conversion or morphing
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/03Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters
    • G10L25/15Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters the extracted parameters being formant information
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/03Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters
    • G10L25/24Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters the extracted parameters being the cepstrum
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04RLOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R2225/00Details of deaf aids covered by H04R25/00, not provided for in any of its subgroups
    • H04R2225/43Signal processing in hearing aids to enhance the speech intelligibility

Landscapes

  • Engineering & Computer Science (AREA)
  • Quality & Reliability (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Physics & Mathematics (AREA)
  • Computational Linguistics (AREA)
  • Signal Processing (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
  • Telephonic Communication Services (AREA)
  • Reduction Or Emphasis Of Bandwidth Of Signals (AREA)
  • Interconnected Communication Systems, Intercoms, And Interphones (AREA)
  • Machine Translation (AREA)
  • Document Processing Apparatus (AREA)
  • Telephone Function (AREA)

Abstract

The characteristics of the speech received by the decoding unit are altered by a processing unit 10 based upon an analysis of the listener's current background noise before the speech is output to enhance its intelligibility to a listener. An analysis unit 12 determines the type and level of the background noise by use of a microphone 13. A decision unit 11 then determines whether the speech currently being received and replayed would be intelligible to an average listener in the current background noise. If unit 11 determines that the speech is readily intelligible then no processing is necessary and the processing unit 10 does not alter the speech which has been passed to it. However, if unit 11 determines that the speech would be unintelligible, then unit 10 alters the speech before passing it to the output to make the speech more intelligible. In a particularly preferred embodiment, the speech characteristics are altered by altering line spectral pair/formant data representing the speech.
PL98337717A 1997-07-02 1998-07-01 Method of and apparatus for improving speech in a voice communication system PL337717A1 (en)

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
GBGB9714001.6A GB9714001D0 (en) 1997-07-02 1997-07-02 Method and apparatus for speech enhancement in a speech communication system

Publications (1)

Publication Number Publication Date
PL337717A1 true PL337717A1 (en) 2000-08-28

Family

ID=10815285

Family Applications (1)

Application Number Title Priority Date Filing Date
PL98337717A PL337717A1 (en) 1997-07-02 1998-07-01 Method of and apparatus for improving speech in a voice communication system

Country Status (12)

Country Link
EP (1) EP0993670B1 (en)
JP (1) JP2002507291A (en)
KR (1) KR20010014352A (en)
CN (1) CN1265217A (en)
AT (1) ATE214832T1 (en)
AU (1) AU8227798A (en)
CA (1) CA2235455A1 (en)
DE (1) DE69804310D1 (en)
GB (2) GB9714001D0 (en)
PL (1) PL337717A1 (en)
WO (1) WO1999001863A1 (en)
ZA (1) ZA985607B (en)

Families Citing this family (36)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
SE9903553D0 (en) * 1999-01-27 1999-10-01 Lars Liljeryd Enhancing conceptual performance of SBR and related coding methods by adaptive noise addition (ANA) and noise substitution limiting (NSL)
FR2794322B1 (en) * 1999-05-27 2001-06-22 Sagem NOISE SUPPRESSION PROCESS
EP1210765B1 (en) 1999-07-28 2007-03-07 Clear Audio Ltd. Filter banked gain control of audio in a noisy environment
US6876968B2 (en) * 2001-03-08 2005-04-05 Matsushita Electric Industrial Co., Ltd. Run time synthesizer adaptation to improve intelligibility of synthesized speech
DE10124189A1 (en) * 2001-05-17 2002-11-21 Siemens Ag Signal reception procedure
JP2003255993A (en) * 2002-03-04 2003-09-10 Ntt Docomo Inc Speech recognition system, speech recognition method, speech recognition program, speech synthesis system, speech synthesis method, speech synthesis program
AU2003263380A1 (en) * 2002-06-19 2004-01-06 Koninklijke Philips Electronics N.V. Audio signal processing apparatus and method
US20060126859A1 (en) * 2003-01-31 2006-06-15 Claus Elberling Sound system improving speech intelligibility
KR20050049103A (en) * 2003-11-21 2005-05-25 삼성전자주식회사 Method and apparatus for enhancing dialog using formant
BRPI0515643A (en) * 2004-09-07 2008-07-29 Sensear Pty Ltd sound improvement equipment and method
US8280730B2 (en) 2005-05-25 2012-10-02 Motorola Mobility Llc Method and apparatus of increasing speech intelligibility in noisy environments
GB2433849B (en) 2005-12-29 2008-05-21 Motorola Inc Telecommunications terminal and method of operation of the terminal
DE102006001730A1 (en) 2006-01-13 2007-07-19 Robert Bosch Gmbh Sound system, method for improving the voice quality and / or intelligibility of voice announcements and computer program
EP1814109A1 (en) * 2006-01-27 2007-08-01 Texas Instruments Incorporated Voice amplification apparatus for modelling the Lombard effect
JP2007295347A (en) * 2006-04-26 2007-11-08 Mitsubishi Electric Corp Audio processing device
WO2018127263A2 (en) * 2017-01-03 2018-07-12 Lizn Aps Speech intelligibility enhancing system
KR101414233B1 (en) 2007-01-05 2014-07-02 삼성전자 주식회사 Apparatus and method for improving intelligibility of speech signal
JP4926005B2 (en) * 2007-11-13 2012-05-09 ソニー・エリクソン・モバイルコミュニケーションズ株式会社 Audio signal processing apparatus, audio signal processing method, and communication terminal
KR101597375B1 (en) 2007-12-21 2016-02-24 디티에스 엘엘씨 System for adjusting perceived loudness of audio signals
JP5453740B2 (en) * 2008-07-02 2014-03-26 富士通株式会社 Speech enhancement device
US8538042B2 (en) 2009-08-11 2013-09-17 Dts Llc System for increasing perceived loudness of speakers
EP2372700A1 (en) * 2010-03-11 2011-10-05 Oticon A/S A speech intelligibility predictor and applications thereof
US9117455B2 (en) 2011-07-29 2015-08-25 Dts Llc Adaptive voice intelligibility processor
CN103002105A (en) * 2011-09-16 2013-03-27 宏碁股份有限公司 Mobile Communication Method That Increases the Clarity of Communication Content
CN103297896B (en) * 2012-02-27 2016-07-06 联想(北京)有限公司 A kind of audio-frequency inputting method and electronic equipment
US9015044B2 (en) * 2012-03-05 2015-04-21 Malaspina Labs (Barbados) Inc. Formant based speech reconstruction from noisy signals
US9312829B2 (en) 2012-04-12 2016-04-12 Dts Llc System for adjusting loudness of audio signals in real time
EP3010017A1 (en) * 2014-10-14 2016-04-20 Thomson Licensing Method and apparatus for separating speech data from background data in audio communication
JP6565206B2 (en) * 2015-02-20 2019-08-28 ヤマハ株式会社 Audio processing apparatus and audio processing method
EP3107097B1 (en) 2015-06-17 2017-11-15 Nxp B.V. Improved speech intelligilibility
US9847093B2 (en) 2015-06-19 2017-12-19 Samsung Electronics Co., Ltd. Method and apparatus for processing speech signal
JP6790732B2 (en) * 2016-11-02 2020-11-25 ヤマハ株式会社 Signal processing method and signal processing device
WO2019127112A1 (en) * 2017-12-27 2019-07-04 深圳前海达闼云端智能科技有限公司 Voice interaction method and device and intelligent terminal
CN109346058B (en) * 2018-11-29 2024-06-28 西安交通大学 A system for expanding speech acoustic features
KR102845224B1 (en) * 2019-12-09 2025-08-12 삼성전자주식회사 Electronic apparatus and controlling method thereof
US11817114B2 (en) 2019-12-09 2023-11-14 Dolby Laboratories Licensing Corporation Content and environmentally aware environmental noise compensation

Family Cites Families (8)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPS5870292A (en) * 1981-10-22 1983-04-26 日産自動車株式会社 Voice recognition equipment for vehicle
US4538295A (en) * 1982-08-16 1985-08-27 Nissan Motor Company, Limited Speech recognition system for an automotive vehicle
EP0226613B1 (en) * 1985-07-01 1993-09-15 Motorola, Inc. Noise supression system
GB8801014D0 (en) * 1988-01-18 1988-02-17 British Telecomm Noise reduction
US5235669A (en) * 1990-06-29 1993-08-10 At&T Laboratories Low-delay code-excited linear-predictive coding of wideband speech at 32 kbits/sec
CA2056110C (en) * 1991-03-27 1997-02-04 Arnold I. Klayman Public address intelligibility system
FI102337B (en) * 1995-09-13 1998-11-13 Nokia Mobile Phones Ltd Procedure and circuit arrangement for processing audio signal
GB2306086A (en) * 1995-10-06 1997-04-23 Richard Morris Trim Improved adaptive audio systems

Also Published As

Publication number Publication date
ATE214832T1 (en) 2002-04-15
GB9714001D0 (en) 1997-09-10
JP2002507291A (en) 2002-03-05
GB9814279D0 (en) 1998-09-02
CN1265217A (en) 2000-08-30
ZA985607B (en) 2000-06-01
CA2235455A1 (en) 1999-01-02
DE69804310D1 (en) 2002-04-25
EP0993670A1 (en) 2000-04-19
KR20010014352A (en) 2001-02-26
WO1999001863A1 (en) 1999-01-14
GB2327835B (en) 2000-04-19
GB2327835A (en) 1999-02-03
AU8227798A (en) 1999-01-25
EP0993670B1 (en) 2002-03-20

Similar Documents

Publication Publication Date Title
GB2327835B (en) Method and apparatus for speech enhancement in a speech communication system
RU2146394C1 (en) Method and device for alternating rate voice coding using reduced encoding rate
DE69620585D1 (en) METHOD AND DEVICE FOR DETECTING AND Bypassing TANDEM SPEECH CODING
SE9500321L (en) Procedure for noise suppression by spectral subtraction
MX9602391A (en) Method and apparatus for reproducing speech signals and method for transmitting same.
JPH1097296A (en) Speech encoding method and apparatus, speech decoding method and apparatus
US6424942B1 (en) Methods and arrangements in a telecommunications system
EP1312075A1 (en) Method for noise robust classification in speech coding
JPH0644195B2 (en) Speech analysis and synthesis system having energy normalization and unvoiced frame suppression function and method thereof
JP3131249B2 (en) Mixed audio signal receiver
AU1324592A (en) Method and apparatus for the teaching of languages
Espy-Wilson et al. Enhancement of alaryngeal speech by adaptive filtering
US7643991B2 (en) Speech enhancement for electronic voiced messages
GB2343822A (en) Using LSP to alter frequency characteristics of speech
JPH10111699A (en) Audio playback device
JP3166797B2 (en) Voice coding method, voice decoding method, and voice codec
SU1674226A1 (en) Method and apparatus for detecting speech signals and their boundaries
Brandenburg et al. Fast signal processor encodes 48 kHz/16-bit audio into 3-bit in real time
KR100624694B1 (en) Sound quality improvement device for call connection sound and its method
Bertrand Secure narrowband digital conferencing
Riedhammer et al. A software kit for automatic voice descrambling
Gan et al. Implementation of silence compression scheme for G. 723.1 speech coder using TI TMS320C51 DSP chip
JPS5853349B2 (en) Speech analysis and synthesis method
Canas et al. Stability of Adaptive Delta Modulation systems with memory
Pfeifer Speaker identification from a minimal set of training data