ATE502380T1 - METHOD, APPARATUS AND PROGRAM CODE FOR CONVERTING VOICES - Google Patents

METHOD, APPARATUS AND PROGRAM CODE FOR CONVERTING VOICES

Info

Publication number
ATE502380T1
ATE502380T1 AT08804436T AT08804436T ATE502380T1 AT E502380 T1 ATE502380 T1 AT E502380T1 AT 08804436 T AT08804436 T AT 08804436T AT 08804436 T AT08804436 T AT 08804436T AT E502380 T1 ATE502380 T1 AT E502380T1
Authority
AT
Austria
Prior art keywords
modelling
stage
glottal
vocal tract
converted
Prior art date
Application number
AT08804436T
Other languages
German (de)
Inventor
Pozo Echezarreta Maria Del
Original Assignee
Ct De Tecnologias De Interaccion Visual Y Comunicaciones Vicomtech Asoc
Del Pozo Echezarreta Maria Arantzazu
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Ct De Tecnologias De Interaccion Visual Y Comunicaciones Vicomtech Asoc, Del Pozo Echezarreta Maria Arantzazu filed Critical Ct De Tecnologias De Interaccion Visual Y Comunicaciones Vicomtech Asoc
Application granted granted Critical
Publication of ATE502380T1 publication Critical patent/ATE502380T1/en

Links

Classifications

    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0316—Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude
    • G10L21/0364—Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude for improving intelligibility
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/003—Changing voice quality, e.g. pitch or formants
    • G10L21/007—Changing voice quality, e.g. pitch or formants characterised by the process used
    • G10L21/013—Adapting to target pitch
    • G10L2021/0135—Voice conversion or morphing
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/04—Time compression or expansion
    • G10L21/057—Time compression or expansion for improving intelligibility
    • G10L2021/0575—Aids for the handicapped in speaking

Landscapes

  • Engineering & Computer Science (AREA)
  • Computational Linguistics (AREA)
  • Quality & Reliability (AREA)
  • Signal Processing (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Physics & Mathematics (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
  • Auxiliary Devices For Music (AREA)
  • Numerical Control (AREA)
  • Measurement Of Mechanical Vibrations Or Ultrasonic Waves (AREA)
  • Circuit For Audible Band Transducer (AREA)

Abstract

A method of converting a source speakers speech signal into a converted speech signal, which comprises a stage of training using a given database of parallel source and target data. For each pitch period modelling a glottal waveform and a vocal tract filter to obtain a set of parameters comprising an excitation strength, parameters modelling a glottal waveform, and all-pole vocal tract filter coefficients. Defining a glottal vector to be converted; defining a vocal tract vector to be converted, obtaining an estimate of a glottal aspiration noise and estimating a vocal tract transformation function. The stage of modelling comprises: modelling said aspiration noise estimate by- modulating Gaussian noise with the said modelled glottal waveform and adjusting its energy to match that of the said aspiration noise estimate. The method further comprises a stage of conversion and a stage of synthesis.
AT08804436T 2008-09-19 2008-09-19 METHOD, APPARATUS AND PROGRAM CODE FOR CONVERTING VOICES ATE502380T1 (en)

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
PCT/EP2008/062502 WO2010031437A1 (en) 2008-09-19 2008-09-19 Method and system of voice conversion

Publications (1)

Publication Number Publication Date
ATE502380T1 true ATE502380T1 (en) 2011-04-15

Family

ID=40277465

Family Applications (1)

Application Number Title Priority Date Filing Date
AT08804436T ATE502380T1 (en) 2008-09-19 2008-09-19 METHOD, APPARATUS AND PROGRAM CODE FOR CONVERTING VOICES

Country Status (5)

Country Link
EP (1) EP2215632B1 (en)
AT (1) ATE502380T1 (en)
DE (1) DE602008005641D1 (en)
ES (1) ES2364005T3 (en)
WO (1) WO2010031437A1 (en)

Families Citing this family (11)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
RU2427044C1 (en) * 2010-05-14 2011-08-20 Закрытое акционерное общество "Ай-Ти Мобайл" Text-dependent voice conversion method
CN101901598A (en) * 2010-06-30 2010-12-01 北京捷通华声语音技术有限公司 Humming synthesis method and system
ES2364401B2 (en) * 2011-06-27 2011-12-23 Universidad Politécnica de Madrid METHOD AND SYSTEM FOR ESTIMATING PHYSIOLOGICAL PARAMETERS OF THE FONATION.
RU2510954C2 (en) * 2012-05-18 2014-04-10 Александр Юрьевич Бредихин Method of re-sounding audio materials and apparatus for realising said method
US9607610B2 (en) 2014-07-03 2017-03-28 Google Inc. Devices and methods for noise modulation in a universal vocoder synthesizer
CN111602194B (en) 2018-09-30 2023-07-04 微软技术许可有限责任公司 Speech waveform generation
US20220148570A1 (en) * 2019-02-25 2022-05-12 Technologies Of Voice Interface Ltd. Speech interpretation device and system
EP3839947A1 (en) 2019-12-20 2021-06-23 SoundHound, Inc. Training a voice morphing apparatus
US11600284B2 (en) 2020-01-11 2023-03-07 Soundhound, Inc. Voice morphing apparatus having adjustable parameters
CN113780107B (en) * 2021-08-24 2024-03-01 电信科学技术第五研究所有限公司 Radio signal detection method based on deep learning dual-input network model
CN115641858B (en) * 2022-10-27 2026-04-21 深圳市中科蓝讯科技股份有限公司 Voice changing processing methods, storage media, chips and electronic devices

Family Cites Families (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100809368B1 (en) * 2006-08-09 2008-03-05 한국과학기술원 Voice conversion system using vocal cords

Also Published As

Publication number Publication date
ES2364005T3 (en) 2011-08-22
EP2215632B1 (en) 2011-03-16
WO2010031437A1 (en) 2010-03-25
DE602008005641D1 (en) 2011-04-28
EP2215632A1 (en) 2010-08-11

Similar Documents

Publication Publication Date Title
DE602008005641D1 (en) METHOD, DEVICE AND PROGRAM CODE FOR CONVERTING VOTES
DK2242045T3 (en) Speech synthesis and coding methods
DE602006017673D1 (en) METHOD AND DEVICE FOR VECTOR-QUANTIZING A SPEKTRALENVELOP REPRESENTATION
WO2010148141A3 (en) Apparatus and method for speech analysis
ATE542191T1 (en) METHOD FOR IDENTIFYING A PERSON BY HIS IRIS
ATE403928T1 (en) VOICE DIALOGUE CONTROL BASED ON SIGNAL PREPROCESSING
ATE492873T1 (en) APPARATUS AND PROGRAM FOR SOUND ANALYSIS
ATE539434T1 (en) APPARATUS AND METHOD FOR MULTI-CHANNEL PARAMETER CONVERSION
JP2019101093A5 (en) Speech synthesis method, speech synthesis system and program
WO2010123483A3 (en) Analyzing the prosody of speech
CN105023574B (en) A kind of method and system for realizing synthesis speech enhan-cement
DE602004023134D1 (en) LANGUAGE RECOGNITION AND SYSTEM ADAPTED TO THE CHARACTERISTICS OF NON-NUT SPEAKERS
ATE533146T1 (en) METHOD AND DEVICE FOR SEARCHING A BASE FREQUENCY
ATE456845T1 (en) LANGUAGE DIFFERENTIATION
JP2015161774A (en) Sound synthesis method and sound synthesizer
EP1495465A4 (en) METHOD FOR MODELING VOICE HARMONIC AMPLITUDES
DE502007002566D1 (en) METHOD AND DEVICE FOR PRODUCING A GAS GENERATOR AND GAS GENERATOR PRODUCED BY THE METHOD
DE502007001672D1 (en) Hearing aid adaptation method
ATE407621T1 (en) METHOD AND DEVICE FOR USING A MULTI-CHANNEL MEASUREMENT SIGNAL IN DETERMINING THE POWER DISTRIBUTION OF AN OBJECT
DE602006010395D1 (en) Use of child-oriented language to automatically generate speech segmentation and a model-based speech recognition system
CN112820266A (en) A Parallel End-to-End Speech Synthesis Method Based on Skip Encoders
JP2017520016A5 (en) Excitation signal formation method of glottal pulse model based on parametric speech synthesis system
ATE528748T1 (en) METHOD AND CORRESPONDING SYSTEM FOR EXPANDING THE SPECTRAL BANDWIDTH OF A VOICE SIGNAL
KR100809368B1 (en) Voice conversion system using vocal cords
ATE515019T1 (en) METHOD AND DEVICE FOR EXECUTING OPTIMALIZED AUDIO CODING BETWEEN TWO LONG-TERM PREDICTION MODELS

Legal Events

Date Code Title Description
RER Ceased as to paragraph 5 lit. 3 law introducing patent treaties