ATE293274T1 - Umschreibung und anzeige eines eingegebenen sprachsignals - Google Patents

Umschreibung und anzeige eines eingegebenen sprachsignals

Info

Publication number
ATE293274T1
ATE293274T1 AT02716176T AT02716176T ATE293274T1 AT E293274 T1 ATE293274 T1 AT E293274T1 AT 02716176 T AT02716176 T AT 02716176T AT 02716176 T AT02716176 T AT 02716176T AT E293274 T1 ATE293274 T1 AT E293274T1
Authority
AT
Austria
Prior art keywords
present
display
transcription
syllable
displayed
Prior art date
Application number
AT02716176T
Other languages
English (en)
Inventor
Sara Helene Basson
Dimitri Kanevsky
Benoit Emmanuel Maison
Original Assignee
Ibm
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Ibm filed Critical Ibm
Application granted granted Critical
Publication of ATE293274T1 publication Critical patent/ATE293274T1/de

Links

Classifications

    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L15/00Speech recognition
    • G10L15/22Procedures used during a speech recognition process, e.g. man-machine dialogue

Landscapes

  • Engineering & Computer Science (AREA)
  • Computational Linguistics (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Physics & Mathematics (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Electrically Operated Instructional Devices (AREA)
  • Control Of El Displays (AREA)
  • Document Processing Apparatus (AREA)
AT02716176T 2001-03-16 2002-01-28 Umschreibung und anzeige eines eingegebenen sprachsignals ATE293274T1 (de)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
US09/811,053 US6785650B2 (en) 2001-03-16 2001-03-16 Hierarchical transcription and display of input speech
PCT/GB2002/000359 WO2002075723A1 (en) 2001-03-16 2002-01-28 Transcription and display of input speech

Publications (1)

Publication Number Publication Date
ATE293274T1 true ATE293274T1 (de) 2005-04-15

Family

ID=25205414

Family Applications (1)

Application Number Title Priority Date Filing Date
AT02716176T ATE293274T1 (de) 2001-03-16 2002-01-28 Umschreibung und anzeige eines eingegebenen sprachsignals

Country Status (7)

Country Link
US (1) US6785650B2 (de)
EP (1) EP1368808B1 (de)
JP (1) JP3935844B2 (de)
CN (1) CN1206620C (de)
AT (1) ATE293274T1 (de)
DE (1) DE60203705T2 (de)
WO (1) WO2002075723A1 (de)

Families Citing this family (52)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2002132287A (ja) * 2000-10-20 2002-05-09 Canon Inc 音声収録方法および音声収録装置および記憶媒体
US6915258B2 (en) * 2001-04-02 2005-07-05 Thanassis Vasilios Kontonassios Method and apparatus for displaying and manipulating account information using the human voice
US7076429B2 (en) * 2001-04-27 2006-07-11 International Business Machines Corporation Method and apparatus for presenting images representative of an utterance with corresponding decoded speech
DE10138408A1 (de) * 2001-08-04 2003-02-20 Philips Corp Intellectual Pty Verfahren zur Unterstützung des Korrekturlesens eines spracherkannten Textes mit an die Erkennungszuverlässigkeit angepasstem Wiedergabegeschwindigkeitsverlauf
US20030220788A1 (en) * 2001-12-17 2003-11-27 Xl8 Systems, Inc. System and method for speech recognition and transcription
US6990445B2 (en) * 2001-12-17 2006-01-24 Xl8 Systems, Inc. System and method for speech recognition and transcription
US20030115169A1 (en) * 2001-12-17 2003-06-19 Hongzhuan Ye System and method for management of transcribed documents
US7386454B2 (en) * 2002-07-31 2008-06-10 International Business Machines Corporation Natural error handling in speech recognition
US8321427B2 (en) 2002-10-31 2012-11-27 Promptu Systems Corporation Method and apparatus for generation and augmentation of search terms from external and internal sources
US6993482B2 (en) * 2002-12-18 2006-01-31 Motorola, Inc. Method and apparatus for displaying speech recognition results
EP1524650A1 (de) * 2003-10-06 2005-04-20 Sony International (Europe) GmbH Zuverlässigkeitsmass in einem Spracherkennungssystem
JP4604178B2 (ja) * 2004-11-22 2010-12-22 独立行政法人産業技術総合研究所 音声認識装置及び方法ならびにプログラム
WO2006070373A2 (en) * 2004-12-29 2006-07-06 Avraham Shpigel A system and a method for representing unrecognized words in speech to text conversions as syllables
KR100631786B1 (ko) * 2005-02-18 2006-10-12 삼성전자주식회사 프레임의 신뢰도를 측정하여 음성을 인식하는 방법 및 장치
US20070078806A1 (en) * 2005-10-05 2007-04-05 Hinickle Judith A Method and apparatus for evaluating the accuracy of transcribed documents and other documents
JP4757599B2 (ja) * 2005-10-13 2011-08-24 日本電気株式会社 音声認識システムと音声認識方法およびプログラム
JP2007133033A (ja) * 2005-11-08 2007-05-31 Nec Corp 音声テキスト化システム、音声テキスト化方法および音声テキスト化用プログラム
JPWO2007097390A1 (ja) * 2006-02-23 2009-07-16 日本電気株式会社 音声認識システム、音声認識結果出力方法、及び音声認識結果出力プログラム
US8204748B2 (en) * 2006-05-02 2012-06-19 Xerox Corporation System and method for providing a textual representation of an audio message to a mobile device
US8521510B2 (en) * 2006-08-31 2013-08-27 At&T Intellectual Property Ii, L.P. Method and system for providing an automated web transcription service
US8321197B2 (en) * 2006-10-18 2012-11-27 Teresa Ruth Gaudet Method and process for performing category-based analysis, evaluation, and prescriptive practice creation upon stenographically written and voice-written text files
WO2008084476A2 (en) * 2007-01-09 2008-07-17 Avraham Shpigel Vowel recognition system and method in speech to text applications
US8433576B2 (en) * 2007-01-19 2013-04-30 Microsoft Corporation Automatic reading tutoring with parallel polarized language modeling
US8306822B2 (en) * 2007-09-11 2012-11-06 Microsoft Corporation Automatic reading tutoring using dynamically built language model
US8271281B2 (en) * 2007-12-28 2012-09-18 Nuance Communications, Inc. Method for assessing pronunciation abilities
US8326631B1 (en) * 2008-04-02 2012-12-04 Verint Americas, Inc. Systems and methods for speech indexing
JP5451982B2 (ja) * 2008-04-23 2014-03-26 ニュアンス コミュニケーションズ,インコーポレイテッド 支援装置、プログラムおよび支援方法
WO2010024052A1 (ja) * 2008-08-27 2010-03-04 日本電気株式会社 音声認識仮説検証装置、音声認識装置、それに用いられる方法およびプログラム
US8805686B2 (en) * 2008-10-31 2014-08-12 Soundbound, Inc. Melodis crystal decoder method and device for searching an utterance by accessing a dictionary divided among multiple parallel processors
TWI377560B (en) * 2008-12-12 2012-11-21 Inst Information Industry Adjustable hierarchical scoring method and system
KR101634247B1 (ko) * 2009-12-04 2016-07-08 삼성전자주식회사 피사체 인식을 알리는 디지털 촬영 장치, 상기 디지털 촬영 장치의 제어 방법
US9070360B2 (en) * 2009-12-10 2015-06-30 Microsoft Technology Licensing, Llc Confidence calibration in automatic speech recognition systems
KR20130005160A (ko) * 2011-07-05 2013-01-15 한국전자통신연구원 음성인식기능을 이용한 메세지 서비스 방법
US8914277B1 (en) * 2011-09-20 2014-12-16 Nuance Communications, Inc. Speech and language translation of an utterance
US9390085B2 (en) * 2012-03-23 2016-07-12 Tata Consultancy Sevices Limited Speech processing system and method for recognizing speech samples from a speaker with an oriyan accent when speaking english
US9020803B2 (en) * 2012-09-20 2015-04-28 International Business Machines Corporation Confidence-rated transcription and translation
US9697821B2 (en) * 2013-01-29 2017-07-04 Tencent Technology (Shenzhen) Company Limited Method and system for building a topic specific language model for use in automatic speech recognition
GB2511078A (en) * 2013-02-22 2014-08-27 Cereproc Ltd System for recording speech prompts
CN103106900B (zh) * 2013-02-28 2016-05-04 用友网络科技股份有限公司 语音识别装置和语音识别方法
KR20150092996A (ko) * 2014-02-06 2015-08-17 삼성전자주식회사 디스플레이 장치 및 이를 이용한 전자 장치의 제어 방법
US9633657B2 (en) * 2014-04-02 2017-04-25 Speakread A/S Systems and methods for supporting hearing impaired users
US9741342B2 (en) * 2014-11-26 2017-08-22 Panasonic Intellectual Property Corporation Of America Method and apparatus for recognizing speech by lip reading
US10152298B1 (en) * 2015-06-29 2018-12-11 Amazon Technologies, Inc. Confidence estimation based on frequency
US20190221213A1 (en) * 2018-01-18 2019-07-18 Ezdi Inc. Method for reducing turn around time in transcription
CN108615526B (zh) * 2018-05-08 2020-07-07 腾讯科技(深圳)有限公司 语音信号中关键词的检测方法、装置、终端及存储介质
US11138334B1 (en) 2018-10-17 2021-10-05 Medallia, Inc. Use of ASR confidence to improve reliability of automatic audio redaction
US11398239B1 (en) 2019-03-31 2022-07-26 Medallia, Inc. ASR-enhanced speech compression
US12170082B1 (en) * 2019-03-31 2024-12-17 Medallia, Inc. On-the-fly transcription/redaction of voice-over-IP calls
US11410642B2 (en) * 2019-08-16 2022-08-09 Soundhound, Inc. Method and system using phoneme embedding
US12387720B2 (en) 2020-11-20 2025-08-12 SoundHound AI IP, LLC. Neural sentence generator for virtual assistants
US12223948B2 (en) * 2022-02-03 2025-02-11 Soundhound, Inc. Token confidence scores for automatic speech recognition
US12394411B2 (en) 2022-10-27 2025-08-19 SoundHound AI IP, LLC. Domain specific neural sentence generator for multi-domain virtual assistants

Family Cites Families (9)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US4882757A (en) * 1986-04-25 1989-11-21 Texas Instruments Incorporated Speech recognition system
TW323364B (de) * 1993-11-24 1997-12-21 At & T Corp
US5842163A (en) * 1995-06-21 1998-11-24 Sri International Method and apparatus for computing likelihood and hypothesizing keyword appearance in speech
US6567778B1 (en) * 1995-12-21 2003-05-20 Nuance Communications Natural language speech recognition using slot semantic confidence scores related to their word recognition confidence scores
US6006183A (en) 1997-12-16 1999-12-21 International Business Machines Corp. Speech recognition confidence level display
DE19821422A1 (de) 1998-05-13 1999-11-18 Philips Patentverwaltung Verfahren zum Darstellen von aus einem Sprachsignal ermittelten Wörtern
US6502073B1 (en) * 1999-03-25 2002-12-31 Kent Ridge Digital Labs Low data transmission rate and intelligible speech communication
WO2000058942A2 (en) * 1999-03-26 2000-10-05 Koninklijke Philips Electronics N.V. Client-server speech recognition
US6526380B1 (en) * 1999-03-26 2003-02-25 Koninklijke Philips Electronics N.V. Speech recognition system having parallel large vocabulary recognition engines

Also Published As

Publication number Publication date
DE60203705D1 (de) 2005-05-19
US20020133340A1 (en) 2002-09-19
DE60203705T2 (de) 2006-03-02
JP2004526197A (ja) 2004-08-26
WO2002075723A1 (en) 2002-09-26
CN1206620C (zh) 2005-06-15
US6785650B2 (en) 2004-08-31
JP3935844B2 (ja) 2007-06-27
CN1509467A (zh) 2004-06-30
EP1368808B1 (de) 2005-04-13
EP1368808A1 (de) 2003-12-10

Similar Documents

Publication Publication Date Title
ATE293274T1 (de) Umschreibung und anzeige eines eingegebenen sprachsignals
US8489400B2 (en) System and method for audibly presenting selected text
SG135951A1 (en) Presentation of data based on user input
EP1557821A3 (de) Segmentbasierte tonale Modellierung für tonale Sprachen
EP4447041A3 (de) Synthetisierte sprachaudiodaten, die im auftrag eines menschlichen gesprächsteilnehmers erzeugt werden
CN100521708C (zh) 移动信息终端的语音识别与语音标签记录和调用方法
Warner Reduction
JP4859642B2 (ja) 音声情報管理装置
KR100771374B1 (ko) 언어의 상징적 선율 변환기
CN101840703A (zh) 一种语音变调方法及装置
Tihelka ARTIC for assistive technologies: transformation to resource-limited hardware
KR20120041051A (ko) 초성 기반의 음성검색 기능을 갖는 단말장치 및 그 동작 방법
CHAUDHARY et al. Intuitive data representation techniques for representing para-linguistic speech data
Bradley COPIUS Transcription & orthography toolset
Ahmad et al. Towards designing a high intelligibility rule based standard malay text-to-speech synthesis system
Setiyawan Slips of the Ears on Phonetic Knowledge toward Surabaya Indie Rock Music Singer
KR20020054568A (ko) 이동통신 단말기를 이용한 외국어 학습장치 및 방법
TW200643748A (en) Mixed language inquiry method
KR20090002422U (ko) 영어회화를 습득할 수 있는 비디오 책
KR910008648A (ko) 음성합성기의 복합코딩방법
CN201111045Y (zh) 方便快捷语言翻译器
KR940020247A (ko) 맹인용 정보검색 단말장치
JPH01119822A (ja) 文章読み上げ装置
KR20090000304U (ko) 영어회화를 습득할 수 있는 비디오책
Chopra GAYATRI–A FAST HINDI TEXT TO SPEECH SYSTEM WITH INPUT SUPPORT FOR ENGLISH LANGUAGE

Legal Events

Date Code Title Description
RER Ceased as to paragraph 5 lit. 3 law introducing patent treaties