PL337717A1 - Method of and apparatus for improving speech in a voice communication system - Google Patents

Method of and apparatus for improving speech in a voice communication system

Info

Publication number
PL337717A1
PL337717A1 PL98337717A PL33771798A PL337717A1 PL 337717 A1 PL337717 A1 PL 337717A1 PL 98337717 A PL98337717 A PL 98337717A PL 33771798 A PL33771798 A PL 33771798A PL 337717 A1 PL337717 A1 PL 337717A1
Authority
PL
Poland
Prior art keywords
speech
unit
determines
intelligible
listener
Prior art date
Application number
PL98337717A
Other languages
English (en)
Inventor
Robert James Chance
Ian Vince Mcloughlin
Original Assignee
Simoco International Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Simoco International Ltd filed Critical Simoco International Ltd
Publication of PL337717A1 publication Critical patent/PL337717A1/xx

Links

Classifications

    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0208—Noise filtering
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/003—Changing voice quality, e.g. pitch or formants
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0316—Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude
    • G10L21/0364—Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude for improving intelligibility
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/003—Changing voice quality, e.g. pitch or formants
    • G10L21/007—Changing voice quality, e.g. pitch or formants characterised by the process used
    • G10L21/013—Adapting to target pitch
    • G10L2021/0135—Voice conversion or morphing
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/03—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters
    • G10L25/15—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters the extracted parameters being formant information
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/03—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters
    • G10L25/24—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters the extracted parameters being the cepstrum
    • H—ELECTRICITY
    • H04—ELECTRIC COMMUNICATION TECHNIQUE
    • H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R2225/00—Details of deaf aids covered by H04R25/00, not provided for in any of its subgroups
    • H04R2225/43—Signal processing in hearing aids to enhance the speech intelligibility

Landscapes

  • Engineering & Computer Science (AREA)
  • Quality & Reliability (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Physics & Mathematics (AREA)
  • Computational Linguistics (AREA)
  • Signal Processing (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
  • Telephonic Communication Services (AREA)
  • Reduction Or Emphasis Of Bandwidth Of Signals (AREA)
  • Machine Translation (AREA)
  • Document Processing Apparatus (AREA)
  • Interconnected Communication Systems, Intercoms, And Interphones (AREA)
  • Telephone Function (AREA)
PL98337717A 1997-07-02 1998-07-01 Method of and apparatus for improving speech in a voice communication system PL337717A1 (en)

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
GBGB9714001.6A GB9714001D0 (en) 1997-07-02 1997-07-02 Method and apparatus for speech enhancement in a speech communication system

Publications (1)

Publication Number Publication Date
PL337717A1 true PL337717A1 (en) 2000-08-28

Family

ID=10815285

Family Applications (1)

Application Number Title Priority Date Filing Date
PL98337717A PL337717A1 (en) 1997-07-02 1998-07-01 Method of and apparatus for improving speech in a voice communication system

Country Status (12)

Country Link
EP (1) EP0993670B1 (de)
JP (1) JP2002507291A (de)
KR (1) KR20010014352A (de)
CN (1) CN1265217A (de)
AT (1) ATE214832T1 (de)
AU (1) AU8227798A (de)
CA (1) CA2235455A1 (de)
DE (1) DE69804310D1 (de)
GB (2) GB9714001D0 (de)
PL (1) PL337717A1 (de)
WO (1) WO1999001863A1 (de)
ZA (1) ZA985607B (de)

Families Citing this family (36)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
SE9903553D0 (sv) * 1999-01-27 1999-10-01 Lars Liljeryd Enhancing percepptual performance of SBR and related coding methods by adaptive noise addition (ANA) and noise substitution limiting (NSL)
FR2794322B1 (fr) * 1999-05-27 2001-06-22 Sagem Procede de suppression de bruit
ATE356469T1 (de) 1999-07-28 2007-03-15 Clear Audio Ltd Verstärkungsregelung von audiosignalen in lärmender umgebung mit hilfe einer filterbank
US6876968B2 (en) * 2001-03-08 2005-04-05 Matsushita Electric Industrial Co., Ltd. Run time synthesizer adaptation to improve intelligibility of synthesized speech
DE10124189A1 (de) * 2001-05-17 2002-11-21 Siemens Ag Verfahren zum Signalempfang
JP2003255993A (ja) * 2002-03-04 2003-09-10 Ntt Docomo Inc 音声認識システム、音声認識方法、音声認識プログラム、音声合成システム、音声合成方法、音声合成プログラム
US20050246170A1 (en) * 2002-06-19 2005-11-03 Koninklijke Phillips Electronics N.V. Audio signal processing apparatus and method
EP1609134A1 (de) * 2003-01-31 2005-12-28 Oticon A/S Schallsystem mit verbessertersprachverständlichkeit
KR20050049103A (ko) * 2003-11-21 2005-05-25 삼성전자주식회사 포만트 대역을 이용한 다이얼로그 인핸싱 방법 및 장치
WO2006026812A2 (en) * 2004-09-07 2006-03-16 Sensear Pty Ltd Apparatus and method for sound enhancement
US8280730B2 (en) 2005-05-25 2012-10-02 Motorola Mobility Llc Method and apparatus of increasing speech intelligibility in noisy environments
GB2433849B (en) 2005-12-29 2008-05-21 Motorola Inc Telecommunications terminal and method of operation of the terminal
DE102006001730A1 (de) 2006-01-13 2007-07-19 Robert Bosch Gmbh Beschallungsanlage, Verfahren zur Verbesserung der Sprachqualität und/oder Verständlichkeit von Sprachdurchsagen sowie Computerprogramm
EP1814109A1 (de) * 2006-01-27 2007-08-01 Texas Instruments Incorporated Sprachsignalverstärker zur Modellierung des Lombard-Effekts
JP2007295347A (ja) * 2006-04-26 2007-11-08 Mitsubishi Electric Corp 音声処理装置
KR101414233B1 (ko) 2007-01-05 2014-07-02 삼성전자 주식회사 음성 신호의 명료도를 향상시키는 장치 및 방법
JP4926005B2 (ja) * 2007-11-13 2012-05-09 ソニー・エリクソン・モバイルコミュニケーションズ株式会社 音声信号処理装置及び音声信号処理方法、通信端末
WO2009086174A1 (en) 2007-12-21 2009-07-09 Srs Labs, Inc. System for adjusting perceived loudness of audio signals
JP5453740B2 (ja) * 2008-07-02 2014-03-26 富士通株式会社 音声強調装置
US8538042B2 (en) 2009-08-11 2013-09-17 Dts Llc System for increasing perceived loudness of speakers
EP2372700A1 (de) * 2010-03-11 2011-10-05 Oticon A/S Sprachverständlichkeitsprädikator und Anwendungen dafür
KR102060208B1 (ko) * 2011-07-29 2019-12-27 디티에스 엘엘씨 적응적 음성 명료도 처리기
CN103002105A (zh) * 2011-09-16 2013-03-27 宏碁股份有限公司 可增加通讯内容清晰度的移动通讯方法
CN103297896B (zh) * 2012-02-27 2016-07-06 联想(北京)有限公司 一种音频输出方法及电子设备
US9015044B2 (en) * 2012-03-05 2015-04-21 Malaspina Labs (Barbados) Inc. Formant based speech reconstruction from noisy signals
US9312829B2 (en) 2012-04-12 2016-04-12 Dts Llc System for adjusting loudness of audio signals in real time
EP3010017A1 (de) * 2014-10-14 2016-04-20 Thomson Licensing Verfahren und Vorrichtung zur Trennung von Sprachdaten von Hintergrunddaten in der Audiokommunikation
JP6565206B2 (ja) * 2015-02-20 2019-08-28 ヤマハ株式会社 音声処理装置および音声処理方法
EP3107097B1 (de) 2015-06-17 2017-11-15 Nxp B.V. Verbesserte sprachverständlichkeit
US9847093B2 (en) 2015-06-19 2017-12-19 Samsung Electronics Co., Ltd. Method and apparatus for processing speech signal
JP6790732B2 (ja) * 2016-11-02 2020-11-25 ヤマハ株式会社 信号処理方法、および信号処理装置
EP3566469B1 (de) 2017-01-03 2020-04-01 Lizn APS System zur sprachverständlichkeitsverbesserung
WO2019127112A1 (zh) * 2017-12-27 2019-07-04 深圳前海达闼云端智能科技有限公司 一种语音交互方法、装置和智能终端
CN109346058B (zh) * 2018-11-29 2024-06-28 西安交通大学 一种语音声学特征扩大系统
US11817114B2 (en) 2019-12-09 2023-11-14 Dolby Laboratories Licensing Corporation Content and environmentally aware environmental noise compensation
KR102845224B1 (ko) * 2019-12-09 2025-08-12 삼성전자주식회사 전자 장치 및 이의 제어 방법

Family Cites Families (8)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPS5870292A (ja) * 1981-10-22 1983-04-26 日産自動車株式会社 車両用音声認識装置
US4538295A (en) * 1982-08-16 1985-08-27 Nissan Motor Company, Limited Speech recognition system for an automotive vehicle
EP0226613B1 (de) * 1985-07-01 1993-09-15 Motorola, Inc. Rauschminderungssystem
GB8801014D0 (en) * 1988-01-18 1988-02-17 British Telecomm Noise reduction
US5235669A (en) * 1990-06-29 1993-08-10 At&T Laboratories Low-delay code-excited linear-predictive coding of wideband speech at 32 kbits/sec
CA2056110C (en) * 1991-03-27 1997-02-04 Arnold I. Klayman Public address intelligibility system
FI102337B (fi) * 1995-09-13 1998-11-13 Nokia Mobile Phones Ltd Menetelmä ja piirijärjestely audiosignaalin käsittelemiseksi
GB2306086A (en) * 1995-10-06 1997-04-23 Richard Morris Trim Improved adaptive audio systems

Also Published As

Publication number Publication date
EP0993670A1 (de) 2000-04-19
ATE214832T1 (de) 2002-04-15
ZA985607B (en) 2000-06-01
CN1265217A (zh) 2000-08-30
KR20010014352A (ko) 2001-02-26
GB2327835A (en) 1999-02-03
CA2235455A1 (en) 1999-01-02
GB9814279D0 (en) 1998-09-02
EP0993670B1 (de) 2002-03-20
GB2327835B (en) 2000-04-19
GB9714001D0 (en) 1997-09-10
DE69804310D1 (de) 2002-04-25
AU8227798A (en) 1999-01-25
WO1999001863A1 (en) 1999-01-14
JP2002507291A (ja) 2002-03-05

Similar Documents

Publication Publication Date Title
GB2327835B (en) Method and apparatus for speech enhancement in a speech communication system
RU2146394C1 (ru) Способ и устройство вокодирования переменной скорости при пониженной скорости кодирования
DE69620585D1 (de) Verfahren und vorrichtung zur detektion und umgehung von tandem-sprachkodierung
CA2362584A1 (en) Speech enhancement with gain limitations based on speech activity
MX9602391A (es) Metodo y aparato para reproducir señales de conversacion y metodo para transmitirlas.
JPH1097296A (ja) 音声符号化方法および装置、音声復号化方法および装置
US6424942B1 (en) Methods and arrangements in a telecommunications system
EP1312075A1 (de) Verfahren zur rauschrobusten klassifikation in der sprachkodierung
JP3131249B2 (ja) 混合音声信号受信装置
AU1324592A (en) Method and apparatus for the teaching of languages
Espy-Wilson et al. Enhancement of alaryngeal speech by adaptive filtering
US7643991B2 (en) Speech enhancement for electronic voiced messages
GB2343822A (en) Using LSP to alter frequency characteristics of speech
JPH10111699A (ja) 音声再生装置
JP3166797B2 (ja) 音声符号化法及び音声復号化法並びに音声符復号化装置
SU1674226A1 (ru) Способ обнаружени речевых сигналов и их границ и устройство дл его осуществлени
Brandenburg et al. Fast signal processor encodes 48 kHz/16-bit audio into 3-bit in real time
KR100624694B1 (ko) 통화 연결음 음질개선장치 및 그 방법
Bertrand Secure narrowband digital conferencing
Riedhammer et al. A software kit for automatic voice descrambling
Suzuki et al. 8 kbps voice transmission by SPAC
Gan et al. Implementation of silence compression scheme for G. 723.1 speech coder using TI TMS320C51 DSP chip
JPS5853349B2 (ja) 音声分析合成方法
Canas et al. Stability of Adaptive Delta Modulation systems with memory
Pfeifer Speaker identification from a minimal set of training data