ATE391985T1 - Verfahren und vorrichtung zur modellierung eines spracherkennungssystems und zur schätzung einer wort-fehlerrate basierend auf einem text - Google Patents

Verfahren und vorrichtung zur modellierung eines spracherkennungssystems und zur schätzung einer wort-fehlerrate basierend auf einem text

Info

Publication number
ATE391985T1
ATE391985T1 AT04003316T AT04003316T ATE391985T1 AT E391985 T1 ATE391985 T1 AT E391985T1 AT 04003316 T AT04003316 T AT 04003316T AT 04003316 T AT04003316 T AT 04003316T AT E391985 T1 ATE391985 T1 AT E391985T1
Authority
AT
Austria
Prior art keywords
text
recognition system
speech recognition
modeling
error rate
Prior art date
Application number
AT04003316T
Other languages
English (en)
Inventor
Milind Mahajan
Yonggang Deng
Alejandro Acero
Asela J R Gunawardana
Ciprian Chelba
Original Assignee
Microsoft Corp
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Microsoft Corp filed Critical Microsoft Corp
Application granted granted Critical
Publication of ATE391985T1 publication Critical patent/ATE391985T1/de

Links

Classifications

    • BPERFORMING OPERATIONS; TRANSPORTING
    • B29WORKING OF PLASTICS; WORKING OF SUBSTANCES IN A PLASTIC STATE IN GENERAL
    • B29CSHAPING OR JOINING OF PLASTICS; SHAPING OF MATERIAL IN A PLASTIC STATE, NOT OTHERWISE PROVIDED FOR; AFTER-TREATMENT OF THE SHAPED PRODUCTS, e.g. REPAIRING
    • B29C45/00Injection moulding, i.e. forcing the required volume of moulding material through a nozzle into a closed mould; Apparatus therefor
    • B29C45/17Component parts, details or accessories; Auxiliary operations
    • B29C45/38Cutting-off equipment for sprues or ingates
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L15/00Speech recognition
    • G10L15/08Speech classification or search
    • G10L15/18Speech classification or search using natural language modelling
    • G10L15/183Speech classification or search using natural language modelling using context dependencies, e.g. language models
    • G10L15/19Grammatical context, e.g. disambiguation of the recognition hypotheses based on word sequence rules
    • G10L15/197Probabilistic grammars, e.g. word n-grams
    • BPERFORMING OPERATIONS; TRANSPORTING
    • B29WORKING OF PLASTICS; WORKING OF SUBSTANCES IN A PLASTIC STATE IN GENERAL
    • B29CSHAPING OR JOINING OF PLASTICS; SHAPING OF MATERIAL IN A PLASTIC STATE, NOT OTHERWISE PROVIDED FOR; AFTER-TREATMENT OF THE SHAPED PRODUCTS, e.g. REPAIRING
    • B29C2945/00Indexing scheme relating to injection moulding, i.e. forcing the required volume of moulding material through a nozzle into a closed mould
    • B29C2945/76Measuring, controlling or regulating
    • B29C2945/76344Phase or stage of measurement
    • B29C2945/76418Ejection
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L15/00Speech recognition
    • G10L15/08Speech classification or search
    • G10L15/18Speech classification or search using natural language modelling
    • G10L15/183Speech classification or search using natural language modelling using context dependencies, e.g. language models

Landscapes

  • Engineering & Computer Science (AREA)
  • Physics & Mathematics (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Probability & Statistics with Applications (AREA)
  • Artificial Intelligence (AREA)
  • Computational Linguistics (AREA)
  • Health & Medical Sciences (AREA)
  • Multimedia (AREA)
  • Acoustics & Sound (AREA)
  • Mechanical Engineering (AREA)
  • Manufacturing & Machinery (AREA)
  • Machine Translation (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
  • Character Discrimination (AREA)
  • Telephonic Communication Services (AREA)
AT04003316T 2003-02-13 2004-02-13 Verfahren und vorrichtung zur modellierung eines spracherkennungssystems und zur schätzung einer wort-fehlerrate basierend auf einem text ATE391985T1 (de)

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
US10/365,850 US7117153B2 (en) 2003-02-13 2003-02-13 Method and apparatus for predicting word error rates from text

Publications (1)

Publication Number Publication Date
ATE391985T1 true ATE391985T1 (de) 2008-04-15

Family

ID=32681724

Family Applications (1)

Application Number Title Priority Date Filing Date
AT04003316T ATE391985T1 (de) 2003-02-13 2004-02-13 Verfahren und vorrichtung zur modellierung eines spracherkennungssystems und zur schätzung einer wort-fehlerrate basierend auf einem text

Country Status (7)

Country Link
US (2) US7117153B2 (de)
EP (1) EP1447792B1 (de)
JP (1) JP4528535B2 (de)
KR (1) KR101004560B1 (de)
CN (1) CN100589179C (de)
AT (1) ATE391985T1 (de)
DE (1) DE602004012909T2 (de)

Families Citing this family (32)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US8959019B2 (en) 2002-10-31 2015-02-17 Promptu Systems Corporation Efficient empirical determination, computation, and use of acoustic confusability measures
US8050918B2 (en) * 2003-12-11 2011-11-01 Nuance Communications, Inc. Quality evaluation tool for dynamic voice portals
US7505906B2 (en) * 2004-02-26 2009-03-17 At&T Intellectual Property, Ii System and method for augmenting spoken language understanding by correcting common errors in linguistic performance
US7509259B2 (en) * 2004-12-21 2009-03-24 Motorola, Inc. Method of refining statistical pattern recognition models and statistical pattern recognizers
US7877387B2 (en) * 2005-09-30 2011-01-25 Strands, Inc. Systems and methods for promotional media item selection and promotional program unit generation
US7809568B2 (en) * 2005-11-08 2010-10-05 Microsoft Corporation Indexing and searching speech with text meta-data
US7831428B2 (en) * 2005-11-09 2010-11-09 Microsoft Corporation Speech index pruning
US7831425B2 (en) * 2005-12-15 2010-11-09 Microsoft Corporation Time-anchored posterior indexing of speech
US8234120B2 (en) * 2006-07-26 2012-07-31 Nuance Communications, Inc. Performing a safety analysis for user-defined voice commands to ensure that the voice commands do not cause speech recognition ambiguities
EP1887562B1 (de) * 2006-08-11 2010-04-28 Harman/Becker Automotive Systems GmbH Spracherkennung mittels eines statistischen Sprachmodells unter Verwendung von Quadratwurzelglättung
JP2008134475A (ja) * 2006-11-28 2008-06-12 Internatl Business Mach Corp <Ibm> 入力された音声のアクセントを認識する技術
US7925505B2 (en) * 2007-04-10 2011-04-12 Microsoft Corporation Adaptation of language models and context free grammar in speech recognition
JP5327054B2 (ja) * 2007-12-18 2013-10-30 日本電気株式会社 発音変動規則抽出装置、発音変動規則抽出方法、および発音変動規則抽出用プログラム
US9659559B2 (en) * 2009-06-25 2017-05-23 Adacel Systems, Inc. Phonetic distance measurement system and related methods
US8782556B2 (en) * 2010-02-12 2014-07-15 Microsoft Corporation User-centric soft keyboard predictive technologies
US9224386B1 (en) * 2012-06-22 2015-12-29 Amazon Technologies, Inc. Discriminative language model training using a confusion matrix
CN103578463B (zh) * 2012-07-27 2017-12-01 腾讯科技(深圳)有限公司 自动化测试方法及测试装置
US9292487B1 (en) 2012-08-16 2016-03-22 Amazon Technologies, Inc. Discriminative language model pruning
US10019983B2 (en) * 2012-08-30 2018-07-10 Aravind Ganapathiraju Method and system for predicting speech recognition performance using accuracy scores
EP2891147B1 (de) * 2012-08-30 2020-08-12 Interactive Intelligence, INC. Verfahren und system zur vorhersage der spracherkennungsleistung durch genauigkeitsscores
US9613619B2 (en) 2013-10-30 2017-04-04 Genesys Telecommunications Laboratories, Inc. Predicting recognition quality of a phrase in automatic speech recognition systems
US9953646B2 (en) 2014-09-02 2018-04-24 Belleau Technologies Method and system for dynamic speech recognition and tracking of prewritten script
US10540957B2 (en) * 2014-12-15 2020-01-21 Baidu Usa Llc Systems and methods for speech transcription
US10229672B1 (en) * 2015-12-31 2019-03-12 Google Llc Training acoustic models using connectionist temporal classification
KR102801724B1 (ko) * 2016-06-28 2025-04-30 삼성전자주식회사 언어 처리 방법 및 장치
US9977729B1 (en) * 2016-11-23 2018-05-22 Google Llc Testing applications with a defined input format
CN112528611B (zh) * 2019-08-27 2024-07-02 珠海金山办公软件有限公司 一种编辑表格的方法、装置、计算机存储介质及终端
CN111402865B (zh) * 2020-03-20 2023-08-08 北京达佳互联信息技术有限公司 语音识别训练数据的生成方法、语音识别模型的训练方法
CN112151014B (zh) * 2020-11-04 2023-07-21 平安科技(深圳)有限公司 语音识别结果的测评方法、装置、设备及存储介质
WO2022155842A1 (en) * 2021-01-21 2022-07-28 Alibaba Group Holding Limited Quality estimation for automatic speech recognition
CN113192497B (zh) * 2021-04-28 2024-03-01 平安科技(深圳)有限公司 基于自然语言处理的语音识别方法、装置、设备及介质
CN113936643B (zh) * 2021-12-16 2022-05-17 阿里巴巴达摩院(杭州)科技有限公司 语音识别方法、语音识别模型、电子设备和存储介质

Family Cites Families (29)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US4817156A (en) * 1987-08-10 1989-03-28 International Business Machines Corporation Rapidly training a speech recognizer to a subsequent speaker given training data of a reference speaker
JPH01102599A (ja) * 1987-10-12 1989-04-20 Internatl Business Mach Corp <Ibm> 音声認識方法
ES2128390T3 (es) * 1992-03-02 1999-05-16 At & T Corp Metodo de adiestramiento y dispositivo para reconocimiento de voz.
US5467425A (en) * 1993-02-26 1995-11-14 International Business Machines Corporation Building scalable N-gram language models using maximum likelihood maximum entropy N-gram models
US5515475A (en) * 1993-06-24 1996-05-07 Northern Telecom Limited Speech recognition method using a two-pass search
CA2126380C (en) * 1993-07-22 1998-07-07 Wu Chou Minimum error rate training of combined string models
US5737723A (en) * 1994-08-29 1998-04-07 Lucent Technologies Inc. Confusable word detection in speech recognition
US5802057A (en) * 1995-12-01 1998-09-01 Apple Computer, Inc. Fly-by serial bus arbitration
JPH1097271A (ja) * 1996-09-24 1998-04-14 Nippon Telegr & Teleph Corp <Ntt> 言語モデル構成法、音声認識用モデル及び音声認識方法
US6137863A (en) * 1996-12-13 2000-10-24 At&T Corp. Statistical database correction of alphanumeric account numbers for speech recognition and touch-tone recognition
US6101241A (en) * 1997-07-16 2000-08-08 At&T Corp. Telephone-based speech recognition for data collection
JP2001511267A (ja) * 1997-12-12 2001-08-07 コーニンクレッカ フィリップス エレクトロニクス エヌ ヴィ 音声パターン認識用のモデル特殊因子の決定方法
US6182039B1 (en) * 1998-03-24 2001-01-30 Matsushita Electric Industrial Co., Ltd. Method and apparatus using probabilistic language model based on confusable sets for speech recognition
US6185530B1 (en) * 1998-08-14 2001-02-06 International Business Machines Corporation Apparatus and methods for identifying potential acoustic confusibility among words in a speech recognition system
US6684185B1 (en) * 1998-09-04 2004-01-27 Matsushita Electric Industrial Co., Ltd. Small footprint language and vocabulary independent word recognizer using registration by word spelling
US6385579B1 (en) * 1999-04-29 2002-05-07 International Business Machines Corporation Methods and apparatus for forming compound words for use in a continuous speech recognition system
US6611802B2 (en) * 1999-06-11 2003-08-26 International Business Machines Corporation Method and system for proofreading and correcting dictated text
CN1207664C (zh) * 1999-07-27 2005-06-22 国际商业机器公司 对语音识别结果中的错误进行校正的方法和语音识别系统
US6622121B1 (en) * 1999-08-20 2003-09-16 International Business Machines Corporation Testing speech recognition systems using test data generated by text-to-speech conversion
US6711541B1 (en) * 1999-09-07 2004-03-23 Matsushita Electric Industrial Co., Ltd. Technique for developing discriminative sound units for speech recognition and allophone modeling
US20060074664A1 (en) * 2000-01-10 2006-04-06 Lam Kwok L System and method for utterance verification of chinese long and short keywords
US7219056B2 (en) * 2000-04-20 2007-05-15 International Business Machines Corporation Determining and using acoustic confusability, acoustic perplexity and synthetic acoustic word error rate
US20010056345A1 (en) * 2000-04-25 2001-12-27 David Guedalia Method and system for speech recognition of the alphabet
US6912498B2 (en) * 2000-05-02 2005-06-28 Scansoft, Inc. Error correction in speech recognition by correcting text around selected area
US7031908B1 (en) * 2000-06-01 2006-04-18 Microsoft Corporation Creating a language model for a language processing system
WO2002049004A2 (de) * 2000-12-14 2002-06-20 Siemens Aktiengesellschaft Verfahren und anordnung zur spracherkennung für ein kleingerät
US6859774B2 (en) * 2001-05-02 2005-02-22 International Business Machines Corporation Error corrective mechanisms for consensus decoding of speech
US7013276B2 (en) * 2001-10-05 2006-03-14 Comverse, Inc. Method of assessing degree of acoustic confusability, and system therefor
JP4024614B2 (ja) * 2002-08-02 2007-12-19 日本電信電話株式会社 言語モデル生成方法、装置およびプログラム、テキスト分析装置およびプログラム

Also Published As

Publication number Publication date
DE602004012909D1 (de) 2008-05-21
EP1447792B1 (de) 2008-04-09
US7117153B2 (en) 2006-10-03
KR101004560B1 (ko) 2011-01-03
JP4528535B2 (ja) 2010-08-18
US20040162730A1 (en) 2004-08-19
US7103544B2 (en) 2006-09-05
EP1447792A2 (de) 2004-08-18
KR20040073398A (ko) 2004-08-19
US20050228670A1 (en) 2005-10-13
JP2004246368A (ja) 2004-09-02
DE602004012909T2 (de) 2009-06-10
CN1571013A (zh) 2005-01-26
CN100589179C (zh) 2010-02-10
EP1447792A3 (de) 2005-01-19

Similar Documents

Publication Publication Date Title
ATE391985T1 (de) Verfahren und vorrichtung zur modellierung eines spracherkennungssystems und zur schätzung einer wort-fehlerrate basierend auf einem text
EP1575029B1 (de) Generierung von grossen Graphonem-Einheiten mit Kriterium gegenseitiger Information für die Sprachsynthese
ATE417346T1 (de) Spracherkennungs- und korrektursystem, korrekturvorrichtung und verfahren zur erstellung eines lexikons von alternativen
TWI441163B (zh) 中文語音辨識裝置及其辨識方法
DE602005001125D1 (de) Erlernen der Aussprache neuer Worte unter Verwendung eines Aussprachegraphen
DE60111329D1 (de) Anpassung des phonetischen Kontextes zur Verbesserung der Spracherkennung
DE59010131D1 (de) Verfahren zur sprecheradaptiven Erkennung von Sprache
DE59705581D1 (de) Verfahren zur anpassung eines hidden-markov-lautmodelles in einem spracherkennungssystem
DE602004026258D1 (de) Verfahren und Anordnung zur Erkennung semantischer Strukturen aus einem Text
ATE531031T1 (de) Segmentbasierte tonale modellierung für tonale sprachen
GB2617729A (en) Alternative soft label generation
ATE261171T1 (de) Vorrichtung und verfahren zur erzeugung und bewertung von mehrfachen ausprachevarianten eines buchstabierten worts unter verwendung von entscheidungsbäumen
ATE457510T1 (de) Spracherkennungssystem mit riesigem vokabular
CN106710585A (zh) 语音交互过程中的多音字播报方法及系统
DE602006001764D1 (de) Verfahren zur Spracherkennung
ATE282882T1 (de) Vorrichtung zur sprecherunabhängigen spracherkennung , basierend auf einem client- server-system
WO2004049305A3 (en) Discriminative training of hidden markov models for continuous speech recognition
ATE487212T1 (de) Verstekte bedingte zufallfeldermodelle für phonetische klassifizierung und spracherkennung
DE60233238D1 (de) Verfahren und vorrichtung zur codierung aufeinanderfolgender grundperioden in einem sprachsignal
ATE454676T1 (de) Vorrichtung und verfahren zur handschrifterkennung
EP4503021A4 (de) Verfahren und vorrichtung zur sprachkodierung, verfahren und vorrichtung zur sprachdekodierung, computervorrichtung und speichermedium
CN110176251B (zh) 一种声学数据自动标注方法及装置
ATE440359T1 (de) Verfahren und system zur automatischen textunabhängigen bewertung der aussprache für den sprachunterricht
Sung et al. A new dataset for tonal and segmental dialectometry from the Yue-and pinghua-speaking area
CN110858268A (zh) 一种检测语音翻译系统中不流畅现象的方法及系统

Legal Events

Date Code Title Description
RER Ceased as to paragraph 5 lit. 3 law introducing patent treaties