ATE362165T1 - Anpassung einer umgebungsfehlanpassung für spracherkennungssysteme - Google Patents
Anpassung einer umgebungsfehlanpassung für spracherkennungssystemeInfo
- Publication number
- ATE362165T1 ATE362165T1 AT04770165T AT04770165T ATE362165T1 AT E362165 T1 ATE362165 T1 AT E362165T1 AT 04770165 T AT04770165 T AT 04770165T AT 04770165 T AT04770165 T AT 04770165T AT E362165 T1 ATE362165 T1 AT E362165T1
- Authority
- AT
- Austria
- Prior art keywords
- speech
- environmental
- misadaptation
- adjusting
- voice recognition
- Prior art date
Links
- 230000007613 environmental effect Effects 0.000 title abstract 4
- 238000000034 method Methods 0.000 abstract 2
- 239000013598 vector Substances 0.000 abstract 2
- 230000006978 adaptation Effects 0.000 abstract 1
- 238000004590 computer program Methods 0.000 abstract 1
- 238000001228 spectrum Methods 0.000 abstract 1
- 230000009466 transformation Effects 0.000 abstract 1
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/06—Creation of reference templates; Training of speech recognition systems, e.g. adaptation to the characteristics of the speaker's voice
- G10L15/065—Adaptation
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/20—Speech recognition techniques specially adapted for robustness in adverse environments, e.g. in noise, of stress induced speech
Landscapes
- Engineering & Computer Science (AREA)
- Physics & Mathematics (AREA)
- Computational Linguistics (AREA)
- Health & Medical Sciences (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Human Computer Interaction (AREA)
- Acoustics & Sound (AREA)
- Multimedia (AREA)
- Artificial Intelligence (AREA)
- Compression, Expansion, Code Conversion, And Decoders (AREA)
- Telephonic Communication Services (AREA)
- Two-Way Televisions, Distribution Of Moving Picture Or The Like (AREA)
- Complex Calculations (AREA)
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP03103727 | 2003-10-08 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| ATE362165T1 true ATE362165T1 (de) | 2007-06-15 |
Family
ID=34429460
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| AT04770165T ATE362165T1 (de) | 2003-10-08 | 2004-10-05 | Anpassung einer umgebungsfehlanpassung für spracherkennungssysteme |
Country Status (7)
| Country | Link |
|---|---|
| US (1) | US20070124143A1 (de) |
| EP (1) | EP1673761B1 (de) |
| JP (1) | JP2007508577A (de) |
| CN (1) | CN1864202A (de) |
| AT (1) | ATE362165T1 (de) |
| DE (1) | DE602004006429D1 (de) |
| WO (1) | WO2005036525A1 (de) |
Families Citing this family (6)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US7725316B2 (en) * | 2006-07-05 | 2010-05-25 | General Motors Llc | Applying speech recognition adaptation in an automated speech recognition system of a telematics-equipped vehicle |
| EP2317730B1 (de) | 2009-10-29 | 2015-08-12 | Unify GmbH & Co. KG | Verfahren und System zur automatischen Änderung oder Aktualisierung der Konfiguration oder Einstellung eines Kommunikationssystems |
| GB2482874B (en) * | 2010-08-16 | 2013-06-12 | Toshiba Res Europ Ltd | A speech processing system and method |
| EP2685889B1 (de) * | 2011-03-16 | 2019-08-14 | Koninklijke Philips N.V. | Beurteilung von atemlosigkeits- und ödem-symptomen |
| US8972256B2 (en) * | 2011-10-17 | 2015-03-03 | Nuance Communications, Inc. | System and method for dynamic noise adaptation for robust automatic speech recognition |
| US9338580B2 (en) * | 2011-10-21 | 2016-05-10 | Qualcomm Incorporated | Method and apparatus for packet loss rate-based codec adaptation |
Family Cites Families (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US5604839A (en) * | 1994-07-29 | 1997-02-18 | Microsoft Corporation | Method and system for improving speech recognition through front-end normalization of feature vectors |
| JP2768274B2 (ja) * | 1994-09-08 | 1998-06-25 | 日本電気株式会社 | 音声認識装置 |
| JPH10161692A (ja) * | 1996-12-03 | 1998-06-19 | Canon Inc | 音声認識装置及び音声認識方法 |
| KR100304666B1 (ko) * | 1999-08-28 | 2001-11-01 | 윤종용 | 음성 향상 방법 |
| US7072833B2 (en) * | 2000-06-02 | 2006-07-04 | Canon Kabushiki Kaisha | Speech processing system |
-
2004
- 2004-10-05 AT AT04770165T patent/ATE362165T1/de not_active IP Right Cessation
- 2004-10-05 WO PCT/IB2004/051969 patent/WO2005036525A1/en not_active Ceased
- 2004-10-05 JP JP2006530972A patent/JP2007508577A/ja not_active Withdrawn
- 2004-10-05 DE DE602004006429T patent/DE602004006429D1/de not_active Expired - Lifetime
- 2004-10-05 EP EP04770165A patent/EP1673761B1/de not_active Expired - Lifetime
- 2004-10-05 US US10/574,447 patent/US20070124143A1/en not_active Abandoned
- 2004-10-05 CN CN200480029513.XA patent/CN1864202A/zh active Pending
Also Published As
| Publication number | Publication date |
|---|---|
| JP2007508577A (ja) | 2007-04-05 |
| EP1673761B1 (de) | 2007-05-09 |
| US20070124143A1 (en) | 2007-05-31 |
| EP1673761A1 (de) | 2006-06-28 |
| CN1864202A (zh) | 2006-11-15 |
| WO2005036525A1 (en) | 2005-04-21 |
| DE602004006429D1 (de) | 2007-06-21 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US8731936B2 (en) | Energy-efficient unobtrusive identification of a speaker | |
| Kalinli et al. | Noise adaptive training for robust automatic speech recognition | |
| US7895038B2 (en) | Signal enhancement via noise reduction for speech recognition | |
| Zhan et al. | Speaker normalization based on frequency warping | |
| US6721699B2 (en) | Method and system of Chinese speech pitch extraction | |
| KR20060044629A (ko) | 신경 회로망을 이용한 음성 신호 분리 시스템 및 방법과음성 신호 강화 시스템 | |
| Kim et al. | Cepstrum-domain acoustic feature compensation based on decomposition of speech and noise for ASR in noisy environments | |
| US10522135B2 (en) | System and method for segmenting audio files for transcription | |
| US20050143997A1 (en) | Method and apparatus using spectral addition for speaker recognition | |
| Espinoza-Cuadros et al. | Speaker de-identification system using autoencoders and adversarial training | |
| CN101409073A (zh) | 一种基于基频包络的汉语普通话孤立词识别方法 | |
| Raj et al. | Voice controlled door lock system using Matlab and Arduino | |
| DE602004006429D1 (de) | Anpassung einer umgebungsfehlanpassung für spracherkennungssysteme | |
| JP2003532162A (ja) | 雑音に影響された音声の認識のためのロバストなパラメータ | |
| Zhang et al. | Piecewise-linear transformation-based HMM adaptation for noisy speech | |
| CN119626229A (zh) | 一种基于语音交互中声音状态检测系统及方法 | |
| Sharanyaa et al. | Audio event recognition involving animals and bird species using machine learning | |
| Nishimura et al. | Speaker adaptation method for HMM-based speech recognition | |
| Singh et al. | A critical review on automatic speaker recognition | |
| Gales | Acoustic modelling for speech recognition: Hidden Markov Models and beyond? | |
| Zarazaga et al. | Recovering implicit pitch contours from formants in whispered speech | |
| Alrouqi | Additive noise subtraction for environmental noise in speech recognition | |
| CN106981287A (zh) | 一种提高声纹识别速度的方法及系统 | |
| Pardede et al. | On the Effect of the Implementation of Human Auditory Systems on Q-Log-Based Features for Robustness of Speech Recognition Against Noise. | |
| Kumar | Speech Enhancement Algorithm Analysis for a Reliable Speech Recognition System using Artificial Intelligence Methods |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| RER | Ceased as to paragraph 5 lit. 3 law introducing patent treaties |