ATE354850T1 - Kodierung von audiosignalen - Google Patents

Kodierung von audiosignalen

Info

Publication number
ATE354850T1
ATE354850T1 AT01980541T AT01980541T ATE354850T1 AT E354850 T1 ATE354850 T1 AT E354850T1 AT 01980541 T AT01980541 T AT 01980541T AT 01980541 T AT01980541 T AT 01980541T AT E354850 T1 ATE354850 T1 AT E354850T1
Authority
AT
Austria
Prior art keywords
input signal
psychoacoustic
norm
modeled
frame
Prior art date
Application number
AT01980541T
Other languages
English (en)
Inventor
Richard Heusdens
Renat Vafin
Willem B Kleijn
Original Assignee
Koninkl Philips Electronics Nv
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Koninkl Philips Electronics Nv filed Critical Koninkl Philips Electronics Nv
Application granted granted Critical
Publication of ATE354850T1 publication Critical patent/ATE354850T1/de

Links

Classifications

    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/02—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0316—Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude
    • G10L21/0364—Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude for improving intelligibility
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L2019/0001—Codebooks
    • G10L2019/0013—Codebook search algorithms
    • G10L2019/0014—Selection criteria for distances

Landscapes

  • Engineering & Computer Science (AREA)
  • Physics & Mathematics (AREA)
  • Computational Linguistics (AREA)
  • Signal Processing (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Spectroscopy & Molecular Physics (AREA)
  • Quality & Reliability (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
AT01980541T 2000-11-03 2001-10-31 Kodierung von audiosignalen ATE354850T1 (de)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
EP00203856 2000-11-03
EP01201685 2001-05-08

Publications (1)

Publication Number Publication Date
ATE354850T1 true ATE354850T1 (de) 2007-03-15

Family

ID=26072835

Family Applications (1)

Application Number Title Priority Date Filing Date
AT01980541T ATE354850T1 (de) 2000-11-03 2001-10-31 Kodierung von audiosignalen

Country Status (8)

Country Link
US (1) US7120587B2 (de)
EP (1) EP1338001B1 (de)
JP (1) JP2004513392A (de)
KR (1) KR20020070373A (de)
CN (1) CN1216366C (de)
AT (1) ATE354850T1 (de)
DE (1) DE60126811T2 (de)
WO (1) WO2002037476A1 (de)

Families Citing this family (16)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US8271200B2 (en) * 2003-12-31 2012-09-18 Sieracki Jeffrey M System and method for acoustic signature extraction, detection, discrimination, and localization
US8478539B2 (en) 2003-12-31 2013-07-02 Jeffrey M. Sieracki System and method for neurological activity signature determination, discrimination, and detection
US7079986B2 (en) * 2003-12-31 2006-07-18 Sieracki Jeffrey M Greedy adaptive signature discrimination system and method
WO2005091275A1 (en) * 2004-03-17 2005-09-29 Koninklijke Philips Electronics N.V. Audio coding
US7751572B2 (en) 2005-04-15 2010-07-06 Dolby International Ab Adaptive residual audio coding
KR100788706B1 (ko) * 2006-11-28 2007-12-26 삼성전자주식회사 광대역 음성 신호의 부호화/복호화 방법
KR101299155B1 (ko) 2006-12-29 2013-08-22 삼성전자주식회사 오디오 부호화 및 복호화 장치와 그 방법
KR101149448B1 (ko) * 2007-02-12 2012-05-25 삼성전자주식회사 오디오 부호화 및 복호화 장치와 그 방법
KR101346771B1 (ko) * 2007-08-16 2013-12-31 삼성전자주식회사 심리 음향 모델에 따른 마스킹 값보다 작은 정현파 신호를효율적으로 인코딩하는 방법 및 장치, 그리고 인코딩된오디오 신호를 디코딩하는 방법 및 장치
KR101441898B1 (ko) * 2008-02-01 2014-09-23 삼성전자주식회사 주파수 부호화 방법 및 장치와 주파수 복호화 방법 및 장치
US8805083B1 (en) 2010-03-21 2014-08-12 Jeffrey M. Sieracki System and method for discriminating constituents of image by complex spectral signature extraction
US9558762B1 (en) 2011-07-03 2017-01-31 Reality Analytics, Inc. System and method for distinguishing source from unconstrained acoustic signals emitted thereby in context agnostic manner
US9886945B1 (en) 2011-07-03 2018-02-06 Reality Analytics, Inc. System and method for taxonomically distinguishing sample data captured from biota sources
US9691395B1 (en) 2011-12-31 2017-06-27 Reality Analytics, Inc. System and method for taxonomically distinguishing unconstrained signal data segments
JP5799707B2 (ja) * 2011-09-26 2015-10-28 ソニー株式会社 オーディオ符号化装置およびオーディオ符号化方法、オーディオ復号装置およびオーディオ復号方法、並びにプログラム
US11030524B2 (en) * 2017-04-28 2021-06-08 Sony Corporation Information processing device and information processing method

Family Cites Families (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN1062963C (zh) * 1990-04-12 2001-03-07 多尔拜实验特许公司 用于产生高质量声音信号的解码器和编码器
JP3446216B2 (ja) * 1992-03-06 2003-09-16 ソニー株式会社 音声信号処理方法
US5651090A (en) * 1994-05-06 1997-07-22 Nippon Telegraph And Telephone Corporation Coding method and coder for coding input signals of plural channels using vector quantization, and decoding method and decoder therefor
JP3707153B2 (ja) * 1996-09-24 2005-10-19 ソニー株式会社 ベクトル量子化方法、音声符号化方法及び装置
FI973873A7 (fi) * 1997-10-02 1999-04-03 Nokia Mobile Phones Ltd Puhekoodaus

Also Published As

Publication number Publication date
EP1338001B1 (de) 2007-02-21
US20030009332A1 (en) 2003-01-09
CN1408110A (zh) 2003-04-02
DE60126811T2 (de) 2007-12-06
DE60126811D1 (de) 2007-04-05
KR20020070373A (ko) 2002-09-06
CN1216366C (zh) 2005-08-24
JP2004513392A (ja) 2004-04-30
EP1338001A1 (de) 2003-08-27
US7120587B2 (en) 2006-10-10
WO2002037476A1 (en) 2002-05-10

Similar Documents

Publication Publication Date Title
DE60126811D1 (de) Kodierung von audiosignalen
CN110085251B (zh) 人声提取方法、人声提取装置及相关产品
CA1216673A (en) Text to speech system
GB2159997A (en) Speech recognition
ATE386989T1 (de) Verfahren und vorrichtung zum dekodieren handschriftlicher zeichen
TW357313B (en) Methods and apparatus for handwriting recognition
AU2003250669A1 (en) Systems and methods of building and using custom word lists
CN111696580B (zh) 一种语音检测方法、装置、电子设备及存储介质
CN107403619A (zh) 一种应用于自行车环境的语音控制方法及系统
DE50001467D1 (de) Verfahren und vorrichtung zum einbringen von informationen in einen datenstrom sowie verfahren und vorrichtung zum codieren eines audiosignals
DE3275779D1 (en) Recognition of speech or speech-like sounds
EP1569199A4 (de) Datenerzeugungseinrichtung und verfahren für musikkompositionen
Mundodu Krishna et al. Single channel speech separation based on empirical mode decomposition and Hilbert transform
TW200612392A (en) Multi-channel encoder
WO2003014961A3 (en) Methods for efficient filtering of data
Leong et al. An FPGA-based electronic cochlea
CN113129926A (zh) 语音情绪识别模型训练方法、语音情绪识别方法及装置
CN106228976A (zh) 语音识别方法和装置
CN120356460A (zh) 语音识别方法、装置、电子设备及计算机可读存储介质
CN112242134B (zh) 语音合成方法及装置
DE3570784D1 (en) Improved phonemic classification in speech recognition system
Takamichi et al. Do learned speech symbols follow Zipf’s law?
Manjunath et al. Multilingual digits speech-to-text conversion for permanent deaf patients
O'Shaughnessy Design of a real-time French text-to-speech system
Lashkari et al. NMF-based cepstral features for speech emotion recognition

Legal Events

Date Code Title Description
RER Ceased as to paragraph 5 lit. 3 law introducing patent treaties