JP2014519082A - 文字に基づく映像生成 - Google Patents
文字に基づく映像生成 Download PDFInfo
- Publication number
- JP2014519082A JP2014519082A JP2014509502A JP2014509502A JP2014519082A JP 2014519082 A JP2014519082 A JP 2014519082A JP 2014509502 A JP2014509502 A JP 2014509502A JP 2014509502 A JP2014509502 A JP 2014509502A JP 2014519082 A JP2014519082 A JP 2014519082A
- Authority
- JP
- Japan
- Prior art keywords
- person
- video sequence
- speech
- voice
- character string
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
Images
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
- G10L13/08—Text analysis or generation of parameters for speech synthesis out of text, e.g. grapheme to phoneme translation, prosody generation or stress or intonation determination
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T13/00—Animation
- G06T13/20—Three-dimensional [3D] animation
- G06T13/40—Three-dimensional [3D] animation of characters, e.g. humans, animals or virtual beings
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
- G10L13/08—Text analysis or generation of parameters for speech synthesis out of text, e.g. grapheme to phoneme translation, prosody generation or stress or intonation determination
- G10L13/10—Prosody rules derived from text; Stress or intonation
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04M—TELEPHONIC COMMUNICATION
- H04M1/00—Substation equipment, e.g. for use by subscribers
- H04M1/72—Mobile telephones; Cordless telephones, i.e. devices for establishing wireless links to base stations without route selection
- H04M1/724—User interfaces specially adapted for cordless or mobile telephones
- H04M1/72403—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality
- H04M1/7243—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages
- H04M1/72436—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages for text messaging, e.g. short messaging services [SMS] or e-mails
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04M—TELEPHONIC COMMUNICATION
- H04M1/00—Substation equipment, e.g. for use by subscribers
- H04M1/72—Mobile telephones; Cordless telephones, i.e. devices for establishing wireless links to base stations without route selection
- H04M1/724—User interfaces specially adapted for cordless or mobile telephones
- H04M1/72403—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality
- H04M1/7243—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages
- H04M1/72439—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages for image or video messaging
Landscapes
- Engineering & Computer Science (AREA)
- Human Computer Interaction (AREA)
- Physics & Mathematics (AREA)
- Multimedia (AREA)
- Computer Networks & Wireless Communication (AREA)
- Business, Economics & Management (AREA)
- General Business, Economics & Management (AREA)
- Signal Processing (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Acoustics & Sound (AREA)
- Health & Medical Sciences (AREA)
- Computational Linguistics (AREA)
- General Physics & Mathematics (AREA)
- Theoretical Computer Science (AREA)
- Processing Or Creating Images (AREA)
- Two-Way Televisions, Distribution Of Moving Picture Or The Like (AREA)
Abstract
【選択図】図10
Description
「雨」 : f(k,t)、k = 1 〜K、 t = 1 〜T
「雨(驚きの感情)」 : f1(k,t) 、k = 1 〜K、 t = 1 〜T
Claims (24)
- 処理装置に文字列を入力するステップと、
視覚的かつ可聴的な、人の感情表現をモデル化するために、前記処理装置により、前記文字列に基づいて前記人の映像列を生成するステップであって、前記映像列の音声部分を生成するために、前記人の音声の音声モデルを用いることを有するステップとを含む方法。 - 前記処理装置は、携帯装置であり、前記文字列は、ショートメッセージサービス(SMS)を介して第二携帯装置から入力され、
人の映像列を生成する前記ステップは、前記携帯装置と前記第二携帯装置に記録された共有情報に基づいて人の映像列を前記携帯装置により生成することを含む請求項1に記載の方法。 - 前記文字列は、少なくとも一つの言葉を含む言葉群を有し、前記映像列は、前記人が映像列の中で言葉を発声して見えるように生成される請求項1に記載の方法。
- 前記文字列は、発声を表現する文字を有し、前記映像列は、前記人が映像列の中で言葉を発声して見えるように生成される請求項1に記載の方法。
- 前記文字列は、言葉と該言葉の指標とを有しており、前記指標は、前記映像列の中で前記人が前記言葉を発声して見えるとき、前記映像列の中で前記人の感情表現を同時に表示し、前記指標は、規定の指標群であり、前記規定の指標群の各指標は、異なる感情表現に関連する請求項1に記載の方法。
- 映像列を生成する前記ステップは、前記文字列と前記人の事前知識に基づいて、視覚的かつ可聴的な、前記人の感情表現をモデル化するために、前記処理装置により人の映像列を生成することを含む請求項1に記載の方法。
- 映像列を生成する前記ステップは、前記文字列の言葉を前記人の顔の特徴にマッピングするステップと、前記人の顔の特徴を背景上にレンダリングするステップとを含む請求項1に記載の方法。
- 前記言葉は、前記言葉のための一又は複数の指標に基づいて前記顔の特徴にマッピングされ、前記指標は、前記映像列の中で前記人が前記言葉を発声して見えるとき、前記映像列の中で前記人の感情表現を同時に表示する請求項7に記載の方法。
- 前記顔の特徴は、特定の前記人に適用される特定の顔の特徴を含む請求項7に記載の方法。
- 映像列を生成する前記ステップは、前記人の顔の特徴に適合する前記人の体のジェスチャーを生成することを含む請求項7に記載の方法。
- 映像列を生成する前記ステップは、前記人の音声に基づく音声モデルを用いて、前記文字列内の言葉に基づいて前記人の音声を表現する音声列を生成することを含む請求項1に記載の方法。
- 文字列の受信は、リアルタイムで文字列を受信することを含み、
映像列を生成する前記ステップは、視覚的かつ可聴的な、前記人の感情表現をモデル化するために、前記文字列に基づいて人の映像列をリアルタイムで生成するステップを含み、該ステップは、前記映像列の音声部分を生成するために前記人の音声の音声モデルを用いることを含む請求項1に記載の方法。 - 処理装置に文字列を入力するステップと、
視覚的な、人の感情表現をモデル化するために、前記処理装置により、前記文字列に基づいて前記人の映像列を生成するステップであって、前記映像列の各フレームの顔部が前記人の複数の推測画像の結合により表現されるステップと、
可聴的な、人の感情表現をモデル化するために、前記人の音声の音声モデルを用いて、前記文字列に基づいて前記人の音声列を前記処理装置により生成するステップと、
前記処理装置を用いて、前記映像列と音声列とを結合することにより前記人の映像列を生成するステップであって、前記映像列と音声列が前記文字列に基づいて同期されるステップとを含む方法。 - 前記映像列の各フレームの顔部は、前記人の複数の推測画像の線形結合により表現され、前記人の複数の推測画像における各推測画像は、前記人の平均画像からの偏差に対応する請求項13に記載の方法。
- 前記文字列に基づいて前記人の映像列を生成する前記ステップは、前記映像列の各フレームを2以上の領域に分割することを含み、少なくとも一つの前記領域は、前記人の推測画像の結合により表現されている請求項13に記載の方法。
- 前記人の音声の前記音声モデルは、前記人の音声サンプルから作られる複数の音声の特徴を含み、前記複数の音声の特徴の各音声の特徴は、文字に対応する請求項13に記載の方法。
- 前記複数の音声の特徴における各音声の特徴は、言葉、又は、音素、発声に対応する請求項16に記載の方法。
- 前記人の音声の音声モデルは、前記人の音声サンプルから作られる複数の音声の特徴と、前記文字列に従う第二の人の音声と、前記人の音声の波形と第二の人の音声の波形との一致とを含み、前記人の音声特徴は、前記人の音声の波形と第二の人の音声の波形との前記一致に基づいて、前記第二の人の音声にマッピングされる請求項13に記載の方法。
- 前記人の音声の音声モデルは、前記人の音声サンプルから作られる複数の音声の特徴と、前記文字列に従って、文字を音声に変換するモデルにより生成される音声と、前記人の音声の波形と文字を音声に変換するモデルの音声の波形との一致とを含み、前記人の音声特徴は、前記人の音声の波形と文字を音声に変換するモデルの音声の波形との前記一致に基づいて、前記モデルの音声にマッピングされる請求項13に記載の方法。
- 文字列を生成するステップであって、前記文字列は、人の感情の範囲を視覚的かつ可聴的に表現するために、前記人の音声に基づく音声モデルを用いて生成される映像列の中で人が発声する一又は複数の言葉を表現するように構成されるステップと、
前記文字列内の言葉に関連する指標を特定するステップであって、前記指標は、規定の指標群の一つであり、各指標が前記人の異なる感情表現を示すように構成されるステップと、
前記指標を前記文字列に組み込むステップと、
前記映像列を生成するように構成される装置に前記文字列を送信するステップとを含む方法。 - 指標を特定する前記ステップは、前記文字列内の言葉に関連する項目群の一覧から一項目を選択するステップを含み、前記一覧の各項目は、前記人の感情表現を示す指標である請求項20に記載の方法。
- 指標を特定する前記ステップは、自動音声認識(ASR)装置を用いて、前記文字列内の言葉を話す話者の音声列に基づいて、前記文字列内の言葉に関連する指標を特定するステップを含む請求項20に記載の方法。
- 非個人の複数の項目の情報を処理装置に記憶するステップと、
前記処理装置の前記非個人の複数の項目の事前情報に基づいて、前記非個人の複数の項目のための映像列を生成するステップであって、前記非個人の各項目が独立して管理可能に構成されるステップとを含む方法。 - 前記非個人の複数の項目は、前記映像列の中で他の要素に関連して制約される請求項23に記載の方法。
Applications Claiming Priority (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US201161483571P | 2011-05-06 | 2011-05-06 | |
| US61/483,571 | 2011-05-06 | ||
| PCT/US2012/036679 WO2012154618A2 (en) | 2011-05-06 | 2012-05-04 | Video generation based on text |
Publications (3)
| Publication Number | Publication Date |
|---|---|
| JP2014519082A true JP2014519082A (ja) | 2014-08-07 |
| JP2014519082A5 JP2014519082A5 (ja) | 2016-09-08 |
| JP6019108B2 JP6019108B2 (ja) | 2016-11-02 |
Family
ID=47139917
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| JP2014509502A Active JP6019108B2 (ja) | 2011-05-06 | 2012-05-04 | 文字に基づく映像生成 |
Country Status (5)
| Country | Link |
|---|---|
| US (1) | US9082400B2 (ja) |
| EP (1) | EP2705515A4 (ja) |
| JP (1) | JP6019108B2 (ja) |
| CN (2) | CN103650002B (ja) |
| WO (1) | WO2012154618A2 (ja) |
Cited By (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2014146340A (ja) * | 2013-01-29 | 2014-08-14 | Toshiba Corp | コンピュータ生成ヘッド |
| JP2014146339A (ja) * | 2013-01-29 | 2014-08-14 | Toshiba Corp | コンピュータ生成ヘッド |
| JP2024532728A (ja) * | 2021-08-07 | 2024-09-10 | グーグル エルエルシー | 自動ボイスオーバ生成 |
Families Citing this family (50)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2012088403A2 (en) | 2010-12-22 | 2012-06-28 | Seyyer, Inc. | Video transmission and sharing over ultra-low bitrate wireless communication channel |
| US8682144B1 (en) * | 2012-09-17 | 2014-03-25 | Google Inc. | Method for synchronizing multiple audio signals |
| KR102091003B1 (ko) * | 2012-12-10 | 2020-03-19 | 삼성전자 주식회사 | 음성인식 기술을 이용한 상황 인식 서비스 제공 방법 및 장치 |
| EP2929460A4 (en) * | 2012-12-10 | 2016-06-22 | Wibbitz Ltd | METHOD FOR AUTOMATICALLY TRANSFORMING TEXT IN VIDEO |
| US9639743B2 (en) * | 2013-05-02 | 2017-05-02 | Emotient, Inc. | Anonymization of facial images |
| US10438631B2 (en) | 2014-02-05 | 2019-10-08 | Snap Inc. | Method for real-time video processing involving retouching of an object in the video |
| CN105282621A (zh) * | 2014-07-22 | 2016-01-27 | 中兴通讯股份有限公司 | 一种语音消息可视化服务的实现方法及装置 |
| US9607609B2 (en) * | 2014-09-25 | 2017-03-28 | Intel Corporation | Method and apparatus to synthesize voice based on facial structures |
| US10116901B2 (en) * | 2015-03-18 | 2018-10-30 | Avatar Merger Sub II, LLC | Background modification in video conferencing |
| US10664741B2 (en) | 2016-01-14 | 2020-05-26 | Samsung Electronics Co., Ltd. | Selecting a behavior of a virtual agent |
| WO2017137948A1 (en) * | 2016-02-10 | 2017-08-17 | Vats Nitin | Producing realistic body movement using body images |
| US10595039B2 (en) | 2017-03-31 | 2020-03-17 | Nvidia Corporation | System and method for content and motion controlled action video generation |
| CN107172449A (zh) * | 2017-06-19 | 2017-09-15 | 微鲸科技有限公司 | 多媒体播放方法、装置及多媒体存储方法 |
| KR102421745B1 (ko) * | 2017-08-22 | 2022-07-19 | 삼성전자주식회사 | Tts 모델을 생성하는 시스템 및 전자 장치 |
| CN109992754B (zh) * | 2017-12-29 | 2023-06-16 | 阿里巴巴(中国)有限公司 | 文档处理方法及装置 |
| CN112334973B (zh) * | 2018-07-19 | 2024-04-26 | 杜比国际公司 | 用于创建基于对象的音频内容的方法和系统 |
| KR102136464B1 (ko) * | 2018-07-31 | 2020-07-21 | 전자부품연구원 | 어텐션 메커니즘 기반의 오디오 분할 방법 |
| KR102079453B1 (ko) * | 2018-07-31 | 2020-02-19 | 전자부품연구원 | 비디오 특성에 부합하는 오디오 합성 방법 |
| CN110853614A (zh) * | 2018-08-03 | 2020-02-28 | Tcl集团股份有限公司 | 虚拟对象口型驱动方法、装置及终端设备 |
| CN108986186B (zh) * | 2018-08-14 | 2023-05-05 | 山东师范大学 | 文字转化视频的方法和系统 |
| CN109218629B (zh) * | 2018-09-14 | 2021-02-05 | 三星电子(中国)研发中心 | 视频生成方法、存储介质和装置 |
| TW202014992A (zh) * | 2018-10-08 | 2020-04-16 | 財團法人資訊工業策進會 | 虛擬臉部模型之表情擬真系統及方法 |
| CN109195007B (zh) * | 2018-10-19 | 2021-09-07 | 深圳市轱辘车联数据技术有限公司 | 视频生成方法、装置、服务器及计算机可读存储介质 |
| CN109614537A (zh) * | 2018-12-06 | 2019-04-12 | 北京百度网讯科技有限公司 | 用于生成视频的方法、装置、设备和存储介质 |
| KR102116315B1 (ko) * | 2018-12-17 | 2020-05-28 | 주식회사 인공지능연구원 | 캐릭터의 음성과 모션 동기화 시스템 |
| CN113383384A (zh) * | 2019-01-25 | 2021-09-10 | 索美智能有限公司 | 语音动画的实时生成 |
| CN109978021B (zh) * | 2019-03-07 | 2022-09-16 | 北京大学深圳研究生院 | 一种基于文本不同特征空间的双流式视频生成方法 |
| CN110148406B (zh) * | 2019-04-12 | 2022-03-04 | 北京搜狗科技发展有限公司 | 一种数据处理方法和装置、一种用于数据处理的装置 |
| CN110162598B (zh) * | 2019-04-12 | 2022-07-12 | 北京搜狗科技发展有限公司 | 一种数据处理方法和装置、一种用于数据处理的装置 |
| CN110166844B (zh) * | 2019-04-12 | 2022-05-31 | 北京搜狗科技发展有限公司 | 一种数据处理方法和装置、一种用于数据处理的装置 |
| CN110263203B (zh) * | 2019-04-26 | 2021-09-24 | 桂林电子科技大学 | 一种结合皮尔逊重构的文本到图像生成方法 |
| US11151979B2 (en) * | 2019-08-23 | 2021-10-19 | Tencent America LLC | Duration informed attention network (DURIAN) for audio-visual synthesis |
| CN110728971B (zh) * | 2019-09-25 | 2022-02-18 | 云知声智能科技股份有限公司 | 一种音视频合成方法 |
| CN110933330A (zh) * | 2019-12-09 | 2020-03-27 | 广州酷狗计算机科技有限公司 | 视频配音方法、装置、计算机设备及计算机可读存储介质 |
| CN111061915B (zh) * | 2019-12-17 | 2023-04-18 | 中国科学技术大学 | 视频人物关系识别方法 |
| CN111259148B (zh) * | 2020-01-19 | 2024-03-26 | 北京小米松果电子有限公司 | 信息处理方法、装置及存储介质 |
| CN114242037B (zh) * | 2020-09-08 | 2025-09-05 | 华为技术有限公司 | 一种虚拟人物生成方法及其装置 |
| US11682153B2 (en) | 2020-09-12 | 2023-06-20 | Jingdong Digits Technology Holding Co., Ltd. | System and method for synthesizing photo-realistic video of a speech |
| CN113115104B (zh) * | 2021-03-19 | 2023-04-07 | 北京达佳互联信息技术有限公司 | 视频处理方法、装置、电子设备及存储介质 |
| CN113891150B (zh) * | 2021-09-24 | 2024-10-11 | 北京搜狗科技发展有限公司 | 一种视频处理方法、装置和介质 |
| CN114513706B (zh) * | 2022-03-22 | 2023-07-25 | 中国平安人寿保险股份有限公司 | 视频生成方法和装置、计算机设备、存储介质 |
| CN114579806B (zh) * | 2022-04-27 | 2022-08-09 | 阿里巴巴(中国)有限公司 | 视频检测方法、存储介质和处理器 |
| CN114898735B (zh) * | 2022-06-09 | 2025-04-22 | 上海幻电信息科技有限公司 | 音视频素材的生成方法及装置 |
| CN116582726B (zh) * | 2023-07-12 | 2023-12-01 | 北京红棉小冰科技有限公司 | 视频生成方法、装置、电子设备及存储介质 |
| CN117201706B (zh) * | 2023-09-12 | 2024-10-11 | 深圳市木愚科技有限公司 | 基于控制策略的数字人合成方法、系统、设备及介质 |
| CN117544832B (zh) | 2023-11-17 | 2025-05-02 | 北京有竹居网络技术有限公司 | 用于生成视频的方法、装置、设备和介质 |
| CN119893238A (zh) | 2023-11-17 | 2025-04-25 | 北京有竹居网络技术有限公司 | 用于生成视频的方法、装置、设备和介质 |
| CN120416680A (zh) | 2024-01-30 | 2025-08-01 | 北京有竹居网络技术有限公司 | 用于生成视频的方法、装置、电子设备和计算机程序产品 |
| CN119295618B (zh) * | 2024-12-13 | 2025-02-28 | 成都开心音符科技有限公司 | 音频和视频生成方法、电子设备和计算机可读存储介质 |
| CN119364133B (zh) * | 2024-12-19 | 2025-03-18 | 苏州元脑智能科技有限公司 | 一种视频数据生成方法、电子设备、存储介质及程序产品 |
Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2003216173A (ja) * | 2002-01-28 | 2003-07-30 | Toshiba Corp | 合成音声及び映像の同期制御方法、装置及びプログラム |
| JP2007279776A (ja) * | 2004-07-23 | 2007-10-25 | Matsushita Electric Ind Co Ltd | Cgキャラクタエージェント装置 |
| JP2010277588A (ja) * | 2009-05-28 | 2010-12-09 | Samsung Electronics Co Ltd | アニメーションスクリプト生成装置、アニメーション出力装置、受信端末装置、送信端末装置、携帯用端末装置及び方法 |
Family Cites Families (23)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP4236815B2 (ja) * | 1998-03-11 | 2009-03-11 | マイクロソフト コーポレーション | 顔合成装置および顔合成方法 |
| US6250928B1 (en) * | 1998-06-22 | 2001-06-26 | Massachusetts Institute Of Technology | Talking facial display method and apparatus |
| US6735566B1 (en) * | 1998-10-09 | 2004-05-11 | Mitsubishi Electric Research Laboratories, Inc. | Generating realistic facial animation from speech |
| JP2001034282A (ja) * | 1999-07-21 | 2001-02-09 | Konami Co Ltd | 音声合成方法、音声合成のための辞書構築方法、音声合成装置、並びに音声合成プログラムを記録したコンピュータ読み取り可能な媒体 |
| US6594629B1 (en) * | 1999-08-06 | 2003-07-15 | International Business Machines Corporation | Methods and apparatus for audio-visual speech detection and recognition |
| US6366885B1 (en) * | 1999-08-27 | 2002-04-02 | International Business Machines Corporation | Speech driven lip synthesis using viseme based hidden markov models |
| US6813607B1 (en) * | 2000-01-31 | 2004-11-02 | International Business Machines Corporation | Translingual visual speech synthesis |
| US6539354B1 (en) * | 2000-03-24 | 2003-03-25 | Fluent Speech Technologies, Inc. | Methods and devices for producing and using synthetic visual speech based on natural coarticulation |
| US20070254684A1 (en) * | 2001-08-16 | 2007-11-01 | Roamware, Inc. | Method and system for sending and creating expressive messages |
| DE60132096T2 (de) * | 2000-08-22 | 2008-06-26 | Symbian Ltd. | Verfahren und vorrichtung zur kommunikation von nutzerbezogenen informationen unter verwendung einer drahtlosen informationsvorrichtung |
| US6970820B2 (en) * | 2001-02-26 | 2005-11-29 | Matsushita Electric Industrial Co., Ltd. | Voice personalization of speech synthesizer |
| KR20060090687A (ko) * | 2003-09-30 | 2006-08-14 | 코닌클리케 필립스 일렉트로닉스 엔.브이. | 시청각 콘텐츠 합성을 위한 시스템 및 방법 |
| DE102004012208A1 (de) * | 2004-03-12 | 2005-09-29 | Siemens Ag | Individualisierung von Sprachausgabe durch Anpassen einer Synthesestimme an eine Zielstimme |
| JP4627152B2 (ja) * | 2004-06-01 | 2011-02-09 | 三星電子株式会社 | 危機監視システム |
| KR20070117195A (ko) * | 2006-06-07 | 2007-12-12 | 삼성전자주식회사 | 휴대용 단말기에서 사용자의 감정이 이입된 문자메시지를송수신하는 방법 및 장치 |
| GB0702150D0 (en) * | 2007-02-05 | 2007-03-14 | Amegoworld Ltd | A Communication Network and Devices |
| CN101925952B (zh) | 2008-01-21 | 2012-06-06 | 松下电器产业株式会社 | 音响再现装置 |
| US20090252481A1 (en) | 2008-04-07 | 2009-10-08 | Sony Ericsson Mobile Communications Ab | Methods, apparatus, system and computer program product for audio input at video recording |
| US8224652B2 (en) * | 2008-09-26 | 2012-07-17 | Microsoft Corporation | Speech and text driven HMM-based body animation synthesis |
| EP2491536B1 (en) | 2009-10-20 | 2019-06-12 | Oath Inc. | Method and system for assembling animated media based on keyword and string input |
| CN101751809B (zh) * | 2010-02-10 | 2011-11-09 | 长春大学 | 基于三维头像的聋儿语言康复方法及系统 |
| US8558903B2 (en) | 2010-03-25 | 2013-10-15 | Apple Inc. | Accelerometer / gyro-facilitated video stabilization |
| WO2012088403A2 (en) | 2010-12-22 | 2012-06-28 | Seyyer, Inc. | Video transmission and sharing over ultra-low bitrate wireless communication channel |
-
2012
- 2012-05-04 WO PCT/US2012/036679 patent/WO2012154618A2/en not_active Ceased
- 2012-05-04 CN CN201280033415.8A patent/CN103650002B/zh active Active
- 2012-05-04 JP JP2014509502A patent/JP6019108B2/ja active Active
- 2012-05-04 EP EP12782015.7A patent/EP2705515A4/en not_active Withdrawn
- 2012-05-04 CN CN201810052644.3A patent/CN108090940A/zh active Pending
- 2012-05-04 US US13/464,915 patent/US9082400B2/en active Active
Patent Citations (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2003216173A (ja) * | 2002-01-28 | 2003-07-30 | Toshiba Corp | 合成音声及び映像の同期制御方法、装置及びプログラム |
| JP2007279776A (ja) * | 2004-07-23 | 2007-10-25 | Matsushita Electric Ind Co Ltd | Cgキャラクタエージェント装置 |
| JP2010277588A (ja) * | 2009-05-28 | 2010-12-09 | Samsung Electronics Co Ltd | アニメーションスクリプト生成装置、アニメーション出力装置、受信端末装置、送信端末装置、携帯用端末装置及び方法 |
Cited By (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2014146340A (ja) * | 2013-01-29 | 2014-08-14 | Toshiba Corp | コンピュータ生成ヘッド |
| JP2014146339A (ja) * | 2013-01-29 | 2014-08-14 | Toshiba Corp | コンピュータ生成ヘッド |
| US9959657B2 (en) | 2013-01-29 | 2018-05-01 | Kabushiki Kaisha Toshiba | Computer generated head |
| JP2024532728A (ja) * | 2021-08-07 | 2024-09-10 | グーグル エルエルシー | 自動ボイスオーバ生成 |
| JP7763327B2 (ja) | 2021-08-07 | 2025-10-31 | グーグル エルエルシー | 自動ボイスオーバ生成 |
Also Published As
| Publication number | Publication date |
|---|---|
| US20130124206A1 (en) | 2013-05-16 |
| WO2012154618A2 (en) | 2012-11-15 |
| CN103650002A (zh) | 2014-03-19 |
| JP6019108B2 (ja) | 2016-11-02 |
| WO2012154618A3 (en) | 2013-01-17 |
| EP2705515A2 (en) | 2014-03-12 |
| US9082400B2 (en) | 2015-07-14 |
| CN108090940A (zh) | 2018-05-29 |
| EP2705515A4 (en) | 2015-04-29 |
| CN103650002B (zh) | 2018-02-23 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| JP6019108B2 (ja) | 文字に基づく映像生成 | |
| JP2014519082A5 (ja) | ||
| JP7408048B2 (ja) | 人工知能に基づくアニメキャラクター駆動方法及び関連装置 | |
| CN110688911B (zh) | 视频处理方法、装置、系统、终端设备及存储介质 | |
| JP3664474B2 (ja) | 視覚的スピーチの言語透過的合成 | |
| CN113554737A (zh) | 目标对象的动作驱动方法、装置、设备及存储介质 | |
| CN113592985B (zh) | 混合变形值的输出方法及装置、存储介质、电子装置 | |
| JP2023552854A (ja) | ヒューマンコンピュータインタラクション方法、装置、システム、電子機器、コンピュータ可読媒体及びプログラム | |
| US20120130717A1 (en) | Real-time Animation for an Expressive Avatar | |
| CN112668407B (zh) | 人脸关键点生成方法、装置、存储介质及电子设备 | |
| CN119440254A (zh) | 一种数字人实时交互系统及数字人实时交互方法 | |
| US20250006212A1 (en) | Method and apparatus for training speech conversion model, device, and medium | |
| CN117523088A (zh) | 一种个性化的三维数字人全息互动形成系统及方法 | |
| KR20230095432A (ko) | 텍스트 서술 기반 캐릭터 애니메이션 합성 시스템 | |
| WO2024122284A1 (ja) | 情報処理装置、情報処理方法及び情報処理プログラム | |
| CN117115310A (zh) | 一种基于音频和图像的数字人脸生成方法及系统 | |
| CN114155321B (zh) | 一种基于自监督和混合密度网络的人脸动画生成方法 | |
| CN114972589A (zh) | 虚拟数字形象的驱动方法及其装置 | |
| Verma et al. | Animating expressive faces across languages | |
| Kolivand et al. | Realistic lip syncing for virtual character using common viseme set | |
| TWM652806U (zh) | 互動虛擬人像系統 | |
| CN118945420A (zh) | 数字人口播视频生成方法、装置、设备、存储介质和程序产品 | |
| Mahavidyalaya | Phoneme and viseme based approach for lip synchronization | |
| CN117557692A (zh) | 口型动画生成方法、装置、设备和介质 | |
| KR20220097673A (ko) | 수어 인식 방법 |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| A621 | Written request for application examination |
Free format text: JAPANESE INTERMEDIATE CODE: A621 Effective date: 20150422 |
|
| A977 | Report on retrieval |
Free format text: JAPANESE INTERMEDIATE CODE: A971007 Effective date: 20160113 |
|
| A131 | Notification of reasons for refusal |
Free format text: JAPANESE INTERMEDIATE CODE: A131 Effective date: 20160122 |
|
| A601 | Written request for extension of time |
Free format text: JAPANESE INTERMEDIATE CODE: A601 Effective date: 20160421 |
|
| A601 | Written request for extension of time |
Free format text: JAPANESE INTERMEDIATE CODE: A601 Effective date: 20160622 |
|
| A524 | Written submission of copy of amendment under article 19 pct |
Free format text: JAPANESE INTERMEDIATE CODE: A524 Effective date: 20160722 |
|
| TRDD | Decision of grant or rejection written | ||
| A01 | Written decision to grant a patent or to grant a registration (utility model) |
Free format text: JAPANESE INTERMEDIATE CODE: A01 Effective date: 20160902 |
|
| A61 | First payment of annual fees (during grant procedure) |
Free format text: JAPANESE INTERMEDIATE CODE: A61 Effective date: 20161003 |
|
| R150 | Certificate of patent or registration of utility model |
Ref document number: 6019108 Country of ref document: JP Free format text: JAPANESE INTERMEDIATE CODE: R150 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
| R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |