WO2022243778A1 - System and method for smart broadcast management - Google Patents
System and method for smart broadcast management Download PDFInfo
- Publication number
- WO2022243778A1 WO2022243778A1 PCT/IB2022/054124 IB2022054124W WO2022243778A1 WO 2022243778 A1 WO2022243778 A1 WO 2022243778A1 IB 2022054124 W IB2022054124 W IB 2022054124W WO 2022243778 A1 WO2022243778 A1 WO 2022243778A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- speech
- keyword
- segment
- information
- segments
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
- H04R25/00—Electric hearing aids
- H04R25/60—Mounting or interconnection of hearing aid parts, e.g. inside tips, housings or to ossicles
- H04R25/604—Mounting or interconnection of hearing aid parts, e.g. inside tips, housings or to ossicles of acoustic or vibrational transducers
- H04R25/606—Mounting or interconnection of hearing aid parts, e.g. inside tips, housings or to ossicles of acoustic or vibrational transducers acting directly on the eardrum, the ossicles or the skull, e.g. mastoid, tooth, maxillary or mandibular bone, or mechanically stimulating the cochlea, e.g. at the oval window
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/08—Speech classification or search
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
- G10L25/78—Detection of presence or absence of voice signals
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
- H04R25/00—Electric hearing aids
- H04R25/55—Electric hearing aids using an external connection, either wireless or wired
- H04R25/554—Electric hearing aids using an external connection, either wireless or wired using a wireless connection, e.g. between microphone and amplifier or using Tcoils
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/08—Speech classification or search
- G10L2015/088—Word spotting
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
- H04R2225/00—Details of deaf aids covered by H04R25/00, not provided for in any of its subgroups
- H04R2225/43—Signal processing in hearing aids to enhance the speech intelligibility
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
- H04R2225/00—Details of deaf aids covered by H04R25/00, not provided for in any of its subgroups
- H04R2225/55—Communication between hearing aids and external devices via a network for data exchange
Definitions
- the present application relates generally to systems and methods for receiving broadcasted information by a device worn or held by a user and managing (e.g., filtering; annotating; storing) the information prior to being presented to the user.
- Medical devices have provided a wide range of therapeutic benefits to recipients over recent decades.
- Medical devices can include internal or implantable components/devices, external or wearable components/devices, or combinations thereof (e.g., a device having an external component communicating with an implantable component).
- Medical devices such as traditional hearing aids, partially or fully-implantable hearing prostheses (e.g., bone conduction devices, mechanical stimulators, cochlear implants, etc.), pacemakers, defibrillators, functional electrical stimulation devices, and other medical devices, have been successful in performing lifesaving and/or lifestyle enhancement functions and/or recipient monitoring for a number of years.
- implantable medical devices now often include one or more instruments, apparatus, sensors, processors, controllers or other functional mechanical or electrical components that are permanently or temporarily implanted in a recipient. These functional devices are typically used to diagnose, prevent, monitor, treat, or manage a disease/injury or symptom thereof, or to investigate, replace or modify the anatomy or a physiological process. Many of these functional devices utilize power and/or data received from external devices that are part of, or operate in conjunction with, implantable components.
- an apparatus comprises voice activity detection (VAD) circuitry configured to analyze one or more broadcast streams comprising audio data, to identify first segments of the one or more broadcast streams in which the audio data includes speech data, and to identify second segments of the one or more broadcast streams in which the audio data does not include speech data.
- VAD voice activity detection
- the apparatus further comprises derivation circuitry configured to receive the first segments and, for each first segment, to derive one or more words from the speech data of the first segment.
- the apparatus further comprises keyword detection circuitry configured to, for each first segment, receive the one or more words and to generate keyword information indicative of whether at least one word of the one or more words is among a set of stored keywords.
- the apparatus further comprises decision circuitry configured to receive the first segments, the one or more words of each of the first segments, and the keyword information for each of the first segments and, for each first segment, to select, based at least in part on the keyword information, among a plurality of options regarding communication of information indicative of the first segment to a recipient.
- a method comprises receiving one or more electromagnetic wireless broadcast streams comprising audio data.
- the method further comprises dividing the one or more electromagnetic wireless broadcast streams into a plurality of segments comprising speech-including segments and speech-excluding segments.
- the method further comprises evaluating the audio data of each speech-including segment for inclusion of at least one keyword.
- the method further comprises based on said evaluating, communicating information regarding the speech-including segment to a user.
- a non-transitory computer readable storage medium has stored thereon a computer program that instructs a computer system to segment real-time audio information into distinct sections of information by at least: receiving one or more electromagnetic wireless broadcast streams comprising audio information; segmenting the one or more electromagnetic wireless broadcast streams into a plurality of sections comprising speech-including sections and speech-excluding sections; evaluating the audio information of each speech-including section for inclusion of at least one keyword; and based on said evaluating, communicating information regarding the speech-including section to a user.
- FIG. 1 is a perspective view of an example cochlear implant auditory prosthesis implanted in a recipient in accordance with certain implementations described herein;
- FIG. 2 is a perspective view of an example fully implantable middle ear implant auditory prosthesis implanted in a recipient in accordance with certain implementations described herein;
- FIG. 3A schematically illustrates an example system comprising a device worn by a recipient or implanted on and/or within the recipient’s body in accordance with certain implementations described herein;
- FIG. 3B schematically illustrates an example system comprising an external device worn, held, and/or carried by a recipient in accordance with certain implementations described herein;
- FIG. 3C schematically illustrates an example system comprising a device worn by a recipient or implanted on and/or within the recipient’s body and an external device worn, held, and/or carried by the recipient in accordance with certain implementations described herein;
- FIG. 4A schematically illustrates an example apparatus in accordance with certain implementations described herein;
- FIG. 4B schematically illustrates the example apparatus, in accordance with certain implementations described herein, as a component of the device, the external device, or divided among the device and the external device;
- FIGs. 5A-5C are flow diagrams of example methods in accordance with certain implementations described herein.
- a device configured to receive wireless broadcasts (e.g., Bluetooth 5.2 broadcasts; location- based Bluetooth broadcasts) that stream many audio announcements, at least some of which are of interest to the user of the device.
- the received wireless broadcasts can include a large number of audio announcements that are not of interest to the user which can cause various problems (e.g., interfering with the user listening to ambient sounds, conversations, or other audio streams; user missing the small number of announcements of interest within the many audio announcements, thereby creating uncertainty, confusion, and/or stress and potentially impacting the user’s safety).
- a transportation hub e.g., an airport; train station; bus station
- the small fraction of relevant announcements pertaining to the user’s trip e.g., flight number and gate number at an airport.
- Certain implementations described herein utilize a keyword detection based mechanism to analyze the broadcast stream, to segment the broadcast stream into distinct sections of information (e.g., announcements), and to intelligently manage the broadcast streams in the background without the user actively listening to the streams and to notify the user of relevant announcements in an appropriate fashion.
- relevant announcements can be stored and replayed to ensure that none are missed by the user (e.g., by the user listening to them at a more convenient time); preceded by a warning tone (e.g., beep) and played back in response to a user-initiated signal.
- relevant announcements can be converted to text or other visually displayed information relayed to the user (e.g., via a smart phone or smart watch display).
- the keyword detection based mechanism can be tailored directly by the user (e.g., to present only certain categories of announcements selected by the user; on a general basis for all broadcasts; on a per-broadcast basis) and/or can receive information from other integrated services (e.g., calendars; personalized profiling module providing user-specific parameters for keyword detection/notification), thereby ensuring that relevant information is conveyed to the user while streamlining the user’s listening experience.
- integrated services e.g., calendars; personalized profiling module providing user-specific parameters for keyword detection/notification
- inventions detailed herein are applicable, in at least some implementations, to any type of implantable or non-implantable stimulation system or device (e.g., implantable or non-implantable auditory prosthesis device or system). Implementations can include any type of medical device that can utilize the teachings detailed herein and/or variations thereof. Furthermore, while certain implementations are described herein in the context of auditory prosthesis devices, certain other implementations are compatible in the context of other types of devices or systems (e.g., smart phones; smart speakers).
- an implantable transducer assembly including but not limited to: electro-acoustic electrical/acoustic systems, cochlear implant devices, implantable hearing aid devices, middle ear implant devices, bone conduction devices (e.g., active bone conduction devices; passive bone conduction devices, percutaneous bone conduction devices; transcutaneous bone conduction devices), Direct Acoustic Cochlear Implant (DACI), middle ear transducer (MET), electro-acoustic implant devices, other types of auditory prosthesis devices, and/or combinations or variations thereof, or any other suitable hearing prosthesis system with or without one or more external components.
- DACI Direct Acoustic Cochlear Implant
- MET middle ear transducer
- electro-acoustic implant devices other types of auditory prosthesis devices, and/or combinations or variations thereof, or any other suitable hearing prosthesis system with or without one or more external components.
- Implementations can include any type of auditory prosthesis that can utilize the teachings detailed herein and/or variations thereof. Certain such implementations can be referred to as “partially implantable,” “semi-implantable,” “mostly implantable,” “fully implantable,” or “totally implantable” auditory prostheses. In some implementations, the teachings detailed herein and/or variations thereof can be utilized in other types of prostheses beyond auditory prostheses.
- FIG. 1 is a perspective view of an example cochlear implant auditory prosthesis 100 implanted in a recipient in accordance with certain implementations described herein.
- the example auditory prosthesis 100 is shown in FIG. 1 as comprising an implanted stimulator unit 120 and a microphone assembly 124 that is external to the recipient (e.g., a partially implantable cochlear implant).
- An example auditory prosthesis 100 e.g., a totally implantable cochlear implant; a mostly implantable cochlear implant
- the example cochlear implant auditory prosthesis 100 of FIG. 1 can be in conjunction with a reservoir of liquid medicament as described herein.
- the recipient has an outer ear 101, a middle ear 105, and an inner ear 107.
- the outer ear 101 comprises an auricle 110 and an ear canal 102.
- An acoustic pressure or sound wave 103 is collected by the auricle 110 and is channeled into and through the ear canal 102.
- a tympanic membrane 104 Disposed across the distal end of the ear canal 102 is a tympanic membrane 104 which vibrates in response to the sound wave 103.
- This vibration is coupled to oval window or fenestra ovalis 112 through three bones of middle ear 105, collectively referred to as the ossicles 106 and comprising the malleus 108, the incus 109, and the stapes 111.
- the bones 108, 109, and 111 of the middle ear 105 serve to filter and amplify the sound wave 103, causing the oval window 112 to articulate, or vibrate in response to vibration of the tympanic membrane 104.
- This vibration sets up waves of fluid motion of the perilymph within cochlea 140.
- Such fluid motion activates tiny hair cells (not shown) inside the cochlea 140. Activation of the hair cells causes appropriate nerve impulses to be generated and transferred through the spiral ganglion cells (not shown) and auditory nerve 114 to the brain (also not shown) where they are perceived as sound.
- the example auditory prosthesis 100 comprises one or more components which are temporarily or permanently implanted in the recipient.
- the example auditory prosthesis 100 is shown in FIG. 1 with an external component 142 which is directly or indirectly attached to the recipient’s body, and an internal component 144 which is temporarily or permanently implanted in the recipient (e.g., positioned in a recess of the temporal bone adjacent auricle 110 of the recipient).
- the external component 142 typically comprises one or more sound input elements (e.g., an external microphone 124) for detecting sound, a sound processing unit 126 (e.g., disposed in a Behind-The-Ear unit), a power source (not shown), and an external transmitter unit 128.
- the external transmitter unit 128 comprises an external coil 130 (e.g., a wire antenna coil comprising multiple turns of electrically insulated single-strand or multi-strand platinum or gold wire) and, preferably, a magnet (not shown) secured directly or indirectly to the external coil 130.
- the external coil 130 of the external transmitter unit 128 is part of an inductive radio frequency (RF) communication link with the internal component 144.
- the sound processing unit 126 processes the output of the microphone 124 that is positioned externally to the recipient’s body, in the depicted implementation, by the recipient’s auricle 110.
- the sound processing unit 126 processes the output of the microphone 124 and generates encoded signals, sometimes referred to herein as encoded data signals, which are provided to the external transmitter unit 128 (e.g., via a cable).
- the sound processing unit 126 can utilize digital processing techniques to provide frequency shaping, amplification, compression, and other signal conditioning, including conditioning based on recipient-specific fitting parameters.
- the power source of the external component 142 is configured to provide power to the auditory prosthesis 100, where the auditory prosthesis 100 includes a battery (e.g., located in the internal component 144, or disposed in a separate implanted location) that is recharged by the power provided from the external component 142 (e.g., via a transcutaneous energy transfer link).
- the transcutaneous energy transfer link is used to transfer power and/or data to the internal component 144 of the auditory prosthesis 100.
- Various types of energy transfer such as infrared (IR), electromagnetic, capacitive, and inductive transfer, may be used to transfer the power and/or data from the external component 142 to the internal component 144.
- the internal component 144 comprises an internal receiver unit 132, a stimulator unit 120, and an elongate electrode assembly 118.
- the internal receiver unit 132 and the stimulator unit 120 are hermetically sealed within a biocompatible housing.
- the internal receiver unit 132 comprises an internal coil 136 (e.g., a wire antenna coil comprising multiple turns of electrically insulated single-strand or multi strand platinum or gold wire), and preferably, a magnet (also not shown) fixed relative to the internal coil 136.
- the internal receiver unit 132 and the stimulator unit 120 are hermetically sealed within a biocompatible housing, sometimes collectively referred to as a stimulator/receiver unit.
- the internal coil 136 receives power and/or data signals from the external coil 130 via a transcutaneous energy transfer link (e.g., an inductive RF link).
- the stimulator unit 120 generates electrical stimulation signals based on the data signals, and the stimulation signals are delivered to the recipient via the elongate electrode assembly 118.
- the elongate electrode assembly 118 has a proximal end connected to the stimulator unit 120, and a distal end implanted in the cochlea 140.
- the electrode assembly 118 extends from the stimulator unit 120 to the cochlea 140 through the mastoid bone 119.
- the electrode assembly 118 may be implanted at least in the basal region 116, and sometimes further.
- the electrode assembly 118 may extend towards apical end of cochlea 140, referred to as cochlea apex 134.
- the electrode assembly 118 may be inserted into the cochlea 140 via a cochleostomy 122.
- a cochleostomy may be formed through the round window 121, the oval window 112, the promontory 123, or through an apical turn 147 of the cochlea 140.
- the elongate electrode assembly 118 comprises a longitudinally aligned and distally extending array 146 of electrodes or contacts 148, sometimes referred to as electrode or contact array 146 herein, disposed along a length thereof.
- electrode or contact array 146 can be disposed on the electrode assembly 118, in most practical applications, the electrode array 146 is integrated into the electrode assembly 118 (e.g., the electrode array 146 is disposed in the electrode assembly 118).
- the stimulator unit 120 generates stimulation signals which are applied by the electrodes 148 to the cochlea 140, thereby stimulating the auditory nerve 114.
- FIG. 1 schematically illustrates an auditory prosthesis 100 utilizing an external component 142 comprising an external microphone 124, an external sound processing unit 126, and an external power source
- one or more of the microphone 124, sound processing unit 126, and power source are implantable on or within the recipient (e.g., within the internal component 144).
- the auditory prosthesis 100 can have each of the microphone 124, sound processing unit 126, and power source implantable on or within the recipient (e.g., encapsulated within a biocompatible assembly located subcutaneously), and can be referred to as a totally implantable cochlear implant (“TICI”).
- TICI totally implantable cochlear implant
- the auditory prosthesis 100 can have most components of the cochlear implant (e.g., excluding the microphone, which can be an in-the-ear-canal microphone) implantable on or within the recipient, and can be referred to as a mostly implantable cochlear implant (“MICI”).
- MICI implantable cochlear implant
- FIG. 2 schematically illustrates a perspective view of an example fully implantable auditory prosthesis 200 (e.g., fully implantable middle ear implant or totally implantable acoustic system), implanted in a recipient, utilizing an acoustic actuator in accordance with certain implementations described herein.
- the example auditory prosthesis 200 of FIG. 2 comprises a biocompatible implantable assembly 202 (e.g., comprising an implantable capsule) located subcutaneously (e.g., beneath the recipient’s skin and on a recipient's skull). While FIG.
- the implantable assembly 202 includes a signal receiver 204 (e.g., comprising a coil element) and an acoustic transducer 206 (e.g., a microphone comprising a diaphragm and an electret or piezoelectric transducer) that is positioned to receive acoustic signals through the recipient’s overlying tissue.
- the implantable assembly 202 may further be utilized to house a number of components of the fully implantable auditory prosthesis 200.
- the implantable assembly 202 can include an energy storage device and a signal processor (e.g., a sound processing unit).
- Various additional processing logic and/or circuitry components can also be included in the implantable assembly 202 as a matter of design choice.
- the signal processor of the implantable assembly 202 is in operative communication (e.g., electrically interconnected via a wire 208) with an actuator 210 (e.g., comprising a transducer configured to generate mechanical vibrations in response to electrical signals from the signal processor).
- the example auditory prosthesis 100, 200 shown in FIGs. 1 and 2 can comprise an implantable microphone assembly, such as the microphone assembly 206 shown in FIG. 2.
- the signal processor of the implantable assembly 202 can be in operative communication (e.g., electrically interconnected via a wire) with the microphone assembly 206 and the stimulator unit of the main implantable component 120.
- at least one of the microphone assembly 206 and the signal processor e.g., a sound processing unit is implanted on or within the recipient.
- the actuator 210 of the example auditory prosthesis 200 shown in FIG. 2 is supportably connected to a positioning system 212, which in turn, is connected to a bone anchor 214 mounted within the recipient's mastoid process (e.g., via a hole drilled through the skull).
- the actuator 210 includes a connection apparatus 216 for connecting the actuator 210 to the ossicles 106 of the recipient. In a connected state, the connection apparatus 216 provides a communication path for acoustic stimulation of the ossicles 106 (e.g., through transmission of vibrations from the actuator 210 to the incus 109).
- ambient acoustic signals e.g., ambient sound
- a signal processor within the implantable assembly 202 processes the signals to provide a processed audio drive signal via wire 208 to the actuator 210.
- the signal processor may utilize digital processing techniques to provide frequency shaping, amplification, compression, and other signal conditioning, including conditioning based on recipient-specific fitting parameters.
- the audio drive signal causes the actuator 210 to transmit vibrations at acoustic frequencies to the connection apparatus 216 to affect the desired sound sensation via mechanical stimulation of the incus 109 of the recipient.
- the subcutaneously implantable microphone assembly 202 is configured to respond to auditory signals (e.g., sound; pressure variations in an audible frequency range) by generating output signals (e.g., electrical signals; optical signals; electromagnetic signals) indicative of the auditory signals received by the microphone assembly 202, and these output signals are used by the auditory prosthesis 100, 200 to generate stimulation signals which are provided to the recipient’s auditory system.
- auditory signals e.g., sound; pressure variations in an audible frequency range
- output signals e.g., electrical signals; optical signals; electromagnetic signals
- the diaphragm of an implantable microphone assembly 202 can be configured to provide higher sensitivity than are external non-implantable microphone assemblies.
- the diaphragm of an implantable microphone assembly 202 can be configured to be more robust and/or larger than diaphragms for external non-implantable microphone assemblies.
- the example auditory prostheses 100 shown in FIG. 1 utilizes an external microphone 124 and the auditory prosthesis 200 shown in FIG. 2 utilizes an implantable microphone assembly 206 comprising a subcutaneously implantable acoustic transducer.
- the auditory prosthesis 100 utilizes one or more implanted microphone assemblies on or within the recipient.
- the auditory prosthesis 200 utilizes one or more microphone assemblies that are positioned external to the recipient and/or that are implanted on or within the recipient, and utilizes one or more acoustic transducers (e.g., actuator 210) that are implanted on or within the recipient.
- an external microphone assembly can be used to supplement an implantable microphone assembly of the auditory prosthesis 100, 200.
- teachings detailed herein and/or variations thereof can be utilized with any type of external or implantable microphone arrangement, and the acoustic transducers shown in FIGs. 1 and 2 are merely illustrative.
- FIG. 3A schematically illustrates an example system 300 comprising a device 310 worn by a recipient or implanted on and/or within the recipient’s body in accordance with certain implementations described herein.
- FIG. 3B schematically illustrates an example system 300 comprising an external device 320 worn, held, and/or carried by a recipient in accordance with certain implementations described herein.
- FIG. 3C schematically illustrates an example system 300 comprising a device 310 worn by a recipient or implanted on and/or within the recipient’s body and an external device 320 worn, held, and/or carried by the recipient in accordance with certain implementations described herein.
- the audio data can include announcements relevant to one or more users within the range (e.g., spatial extent) of the at least one remote broadcast system 330 (e.g., announcements at an airport, train station, boat dock, or other transportation facility; announcements at a conference, sporting event, or other public or private event).
- announcements relevant to one or more users within the range (e.g., spatial extent) of the at least one remote broadcast system 330 e.g., announcements at an airport, train station, boat dock, or other transportation facility; announcements at a conference, sporting event, or other public or private event.
- the device 310 is configured to receive the electromagnetic signals 332 directly from the at least one remote broadcast system 330 via a wireless communication link 334 (e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection).
- the device 310 can be configured to receive the one or more broadcast streams (e.g., audio broadcast streams) directly from at least one remote broadcast system 330 and to provide information from the audio data (e.g., via stimulation signals; via sound) to the recipient.
- a wireless communication link 334 e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection.
- the device 310 can be configured to receive the one or more broadcast streams (e.g., audio broadcast streams) directly from at least one remote broadcast system 330 and to provide information from the audio data (e.g., via stimulation signals; via sound) to the recipient.
- the one or more broadcast streams e.g., audio broadcast streams
- the external device 320 is configured to receive the electromagnetic signals 332 directly from the at least one broadcast system 330 via a wireless communication link 334 (e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection) and to provide information (e.g., via stimulation signals; via sound) from the audio data to the recipient.
- a wireless communication link 334 e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection
- the external device 320 can be configured to receive the one or more broadcast streams from at least one remote broadcast system 330 and to transmit information (e.g., via sound; via text) from the audio data to the recipient.
- the external device 320 is configured to receive the electromagnetic signals 332 directly from the at least one broadcast system 330 via a first wireless communication link 334 (e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection) and is configured to transmit at least a portion of the one or more broadcast streams to the device 310 via a second wireless communication link 336 (e.g., WiFi; Bluetooth; radio-frequency (RF); magnetic induction).
- a first wireless communication link 334 e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection
- a second wireless communication link 336 e.g., WiFi; Bluetooth; radio-frequency (RF); magnetic induction
- the external device 320 can be configured to receive the one or more broadcast streams from at least one remote broadcast system 330 and to transmit information (e.g., via the second wireless communication link 336) from the audio data to the device 310 which is configured to provide the information (e.g., via stimulation signals; via sound) to the recipient.
- the device 310 of FIGs. 3 A and 3C comprises multiple devices 310 implanted or worn on the recipient’s body.
- the device 310 can comprise two hearing prostheses, one for each of the recipient’s ears (e.g., a bilateral cochlear implant pair; a sound processor and a hearing aid).
- the multiple devices 310 can be operated in sync with one another (e.g., a pair of cochlear implant devices which both receive information from the audio data from the external device 320 or directly from the at least one broadcast system 330).
- the multiple devices 310 operate independently from one another, while in certain other implementations, the multiple devices 310 operate as a “parent” device which controls operation of a “child” device.
- the device 310 and/or the external device 320 are in operative communication with one or more geographically remote computing devices (e.g., remote servers and/or processors; “the cloud”) which are configured to perform one or more functionalities as described herein.
- the device 310 and/or the external device 320 can be configured to transmit signals to the one or more geographically remote computing devices via the at least one broadcast system 330 (e.g., via one or both of the wireless communication links 334, 336) as schematically illustrated by FIGs. 3A-3C.
- the device 310 and/or the external device 320 can be configured to transmit signals to the one or more geographically remote computing devices via other wireless communication links (e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection) that are not coupled to the at least one broadcast system 330.
- other wireless communication links e.g., WiFi; Bluetooth; cellphone connection, telephony, or other Internet connection
- the device 310 comprises a transducer assembly, examples of which include but are not limited to: an implantable and/or wearable sensory prosthesis (e.g., cochlear implant auditory prosthesis 100; fully implantable auditory prosthesis 200; implantable hearing aid; wearable hearing aid, an example of which is a hearing aid that is partially or wholly within the ear canal); at least one wearable speaker (e.g., in-the- ear; over-the-ear; ear bud; headphone).
- an implantable and/or wearable sensory prosthesis e.g., cochlear implant auditory prosthesis 100; fully implantable auditory prosthesis 200; implantable hearing aid; wearable hearing aid, an example of which is a hearing aid that is partially or wholly within the ear canal
- at least one wearable speaker e.g., in-the- ear; over-the-ear; ear bud; headphone.
- the device 310 is configured to receive auditory information from the ambient environment (e.g., sound detected by one or more microphones of the device 310) and/or to receive audio input from at least one remote system (e.g., mobile phone, television, computer), and to receive user input from the recipient (e.g., for controlling the device 310).
- the external device 320 comprises at least one portable device worn, held, and/or carried by the recipient.
- the external device 320 can comprise an externally-worn sound processor (e.g., sound processing unit 126) that is configured to be in wired communication or in wireless communication (e.g., via RF communication link; via magnetic induction link) with the device 310 and is dedicated to operation in conjunction with the device 310.
- the external device 320 can comprise a device remote to the device 310 (e.g., smart phone, smart tablet, smart watch, laptop computer, other mobile computing device configured to be transported away from a stationary location during normal use).
- the external device 320 can comprise multiple devices (e.g., a handheld computing device in communication with an externally-worn sound processor that is in communication with the device 310).
- the external device 320 comprises an input device (e.g., keyboard; touchscreen; buttons; switches; voice recognition system) configured to receive user input from the recipient and an output device (e.g., display; speaker) configured to provide information to the recipient.
- an input device e.g., keyboard; touchscreen; buttons; switches; voice recognition system
- an output device e.g., display; speaker
- the external device 320 can comprise a touchscreen 322 configured to be operated as both the input device and the output device.
- the external device 320 is configured to transmit control signals to the device 310 and/or to receive data signals from the device 310 indicative of operation or performance of the device 310.
- the external device 320 can further be configured to receive user input from the recipient (e.g., for controlling the device 310) and/or to provide the recipient with information regarding the operation or performance of the device 310 (e.g., via a graphical user interface displayed on the touchscreen 322).
- the external device 320 is configured to transmit information (e.g., audio information) to the device 310 and the device 310 is configured to provide the information (e.g., via stimulation signals; via sound) to the recipient.
- information e.g., audio information
- the device 310 is configured to provide the information (e.g., via stimulation signals; via sound) to the recipient.
- FIG. 4A schematically illustrates an example apparatus 400 in accordance with certain implementations described herein.
- the apparatus 400 comprises voice activity detection (VAD) circuitry 410 configured to analyze one or more broadcast streams 412 comprising audio data, to identify first segments 414 of the one or more broadcast streams 412 in which the audio data includes speech data, and to identify second segments of the one or more broadcast streams 412 in which the audio data does not include speech data.
- VAD voice activity detection
- the apparatus 400 further comprises derivation circuitry 420 configured to receive the first segments 414 and, for each first segment 414, to derive one or more words 422 from the speech data of the first segment 414.
- the apparatus 400 further comprises keyword detection circuitry 430 configured to, for each first segment 414, receive the one or more words 422 and to generate keyword information indicative of whether at least one word of the one or more words 422 is among a set of stored keywords 434.
- the apparatus 400 further comprises decision circuitry 440 configured to receive the first segments 414, the one or more words 422 of each of the first segments 414, and the keyword information 432 for each of the first segments 414 and, for each first segment 414, to select, based at least in part on the keyword information 432, among a plurality of options regarding communication of information 442 indicative of the first segment 414 to a recipient.
- FIG. 4B schematically illustrates the example apparatus 400, in accordance with certain implementations described herein, as a component of the device 310, the external device 320, or divided among the device 310 and the external device 320.
- at least a portion of the apparatus 400 resides in one or more one or more geographically remote computing devices that are remote from both the device 310 and the external device 320.
- the apparatus 400 comprises one or more microprocessor (e.g., application-specific integrated circuits; generalized integrated circuits programmed by software with computer executable instructions; microelectronic circuitry; microcontrollers) of which the VAD circuitry 410, the derivation circuitry 420, the keyword detection circuitry 430, and/or the decision circuitry 440 are components.
- the one or more microprocessors comprise control circuitry configured to control the VAD circuitry 410, the derivation circuitry 420, the keyword detection circuitry 430, and/or the decision circuitry 440, as well as other components of the apparatus 400.
- the external device 320 can comprise at least one microprocessor of the one or more microprocessors.
- the device 310 e.g., a sensory prosthesis configured to be worn by a recipient or implanted on and/or within a recipient’s body
- the one or more microprocessors comprise and/or are in operative communication with at least one storage device configured to store information (e.g., data; commands) accessed by the one or more microprocessors during operation (e.g., while providing the functionality of certain implementations described herein).
- the at least one storage device can comprise at least one tangible (e.g., non-transitory) computer readable storage medium, examples of which include but are not limited to: read only memory (ROM); random access memory (RAM); magnetic disk storage media; optical storage media; flash memory.
- the at least one storage device can be encoded with software (e.g., a computer program downloaded as an application) comprising computer executable instructions for instructing the one or more microprocessors (e.g., executable data access logic, evaluation logic, and/or information outputting logic).
- software e.g., a computer program downloaded as an application
- microprocessors e.g., executable data access logic, evaluation logic, and/or information outputting logic.
- the one or more microprocessors execute the instructions of the software to provide functionality as described herein.
- the apparatus 400 can be in operative communication with at least one data input interface 450 (e.g., a component of the device 310 and/or the external device 320) configured to receive the one or more broadcast streams 412.
- the at least one data input interface 450 include but are not limited to ports and/or antennas configured for receiving at least one of: WiFi signals; Bluetooth signals; cellphone connection signals, telephony signals, or other Internet signals.
- the at least one data input interface 450 is configured to detect the electromagnetic signals 332 from the at least one remote broadcast system 330 and to receive the broadcast stream 412 comprising the electromagnetic signals 332 in response to user input (e.g., the user responding to a prompt indicating that the broadcast remote broadcast system 330 has been detected) and/or automatically (e.g., based on learned behavior, such as detection of electromagnetic signals 332 from a remote broadcast system 330 that was connected to during a previous visit within the range of the remote broadcast system 330).
- user input e.g., the user responding to a prompt indicating that the broadcast remote broadcast system 330 has been detected
- automatically e.g., based on learned behavior, such as detection of electromagnetic signals 332 from a remote broadcast system 330 that was connected to during a previous visit within the range of the remote broadcast system 330.
- the apparatus 400 can be configured to operate in at least two modes: a first (e.g., “normal”) operation mode in which the functionalities described herein are disabled and a second (e.g., “smart”) operation mode in which the functionalities described herein are enabled.
- the apparatus 400 can switch between the first and second modes in response to user input (e.g., the user responding to a prompt indicating that the broadcast remote broadcast system 330 has been detected) and/or automatically (e.g., based on connection to and/or disconnection from a remote broadcast system 330).
- the one or more broadcast streams 412 are encoded (e.g., encrypted)
- the at least one data input interface 450 and/or other portions of the apparatus 400 are configured to decode (e.g., decrypt) the broadcast stream 412.
- the apparatus 400 can be in operative communication with at least one data output interface 460 (e.g., a component of the device 310 and/or the external device 320) configured to be operatively coupled to a communication component (e.g., another component of the device 310 and/or the external device 320; a component separate from the device 310 and the external device 320) configured to communicate the information 442 indicative of the first segment 414 to the recipient.
- the at least one data output interface 460 can comprise any combination of wired and/or wireless ports, including but not limited to: Universal Serial Bus (USB) ports; Institute of Electrical and Electronics Engineers (IEEE) 1394 ports; PS/2 ports; network ports; Ethernet ports; Bluetooth ports; wireless network interfaces.
- the first segments 414 e.g., segments including speech data
- the first segments 414 of the one or more broadcast streams 412 contain messages (e.g., sentences) with specific information of possible interest to the recipient (e.g., announcements regarding updates to scheduling or gates at an airport or train station; announcements regarding event schedules or locations at a conference, cultural event, or sporting event).
- the first segments 414 of a broadcast stream 412 can be separated from one another by one or more second segments (e.g., segments not including speech data) of the broadcast stream 412 that contain either no audio data or only non-speech audio data (e.g., music; background noise).
- the VAD circuitry 410 is configured to identify the first segments 414 and to identify the second segments by analyzing one or more characteristics of the audio data of the one or more broadcast streams 412. For example, based on the one or more characteristics (e.g., modulation depth; signal-to-noise ratio; zero crossing rate; cross correlations; sub-band/full-band energy measures; spectral structure in frequency range corresponding to speech (e.g., 80 Hz to 400 Hz); long term time-domain behavior characteristics), the VAD circuitry 410 can identify time intervals of the audio data of the one or more broadcast streams 412 that contain speech activity and time intervals of the audio data of the one or more broadcast streams 412 that do not contain speech activity.
- the one or more characteristics e.g., modulation depth; signal-to-noise ratio; zero crossing rate; cross correlations; sub-band/full-band energy measures; spectral structure in frequency range corresponding to speech (e.g., 80 Hz to 400 Hz); long term time-domain behavior characteristics
- voice activity detection processes that can be performed by the VAD circuitry 410 in accordance with certain implementations described herein are described by S. Graf et al., “Features for voice activity detection: a comparative analysis,” EURASIP J. Adv. in Signal Processing, 2015:91 (2015); International Telecommunications Union, “ITU-T
- the VAD circuitry 410 is local (e.g., a component of the device 410 and/or the external device 420), while in certain other implementations, the VAD circuitry 410 is part of a remote server (e.g., “in the cloud”).
- the VAD circuitry 410 can identify the first segments 414 as being the segments broadcasted between time intervals without broadcasted segments.
- the VAD circuitry 410 is configured to append information to at least some of the segments, the appended information indicative of whether the segment is a first segment 414 (e.g., speech-including segment) or a second segment (e.g., speech-excluding segment).
- a first segment 414 e.g., speech-including segment
- a second segment e.g., speech-excluding segment
- the appended information can be in the form of a value (e.g., zero or one) appended to (e.g., overlaid on) the segment based on whether the one or more characteristics (e.g., modulation depth; signal-to-noise ratio; zero crossing rate; cross correlations; sub-band/full-band energy measures; spectral structure in frequency range corresponding to speech (e.g., 80 Hz to 400 Hz); long term time-domain behavior characteristics) of the audio data of the segment is indicative of either the segment being a first segment 414 or a second segment.
- the VAD circuitry 410 is configured to parse (e.g., divide) the first segments 414 from the second segments.
- the VAD circuitry 410 can transmit the first segments 414 to circuitry for further processing (e.g., to memory circuitry for storage and further processing by other circuitry) and can discard the second segments.
- the VAD circuitry 410 can exclude the second segments from further processing (e.g., by transmitting the first segments 414 to the derivation circuitry 420 and to the decision circuitry 440 while not transmitting the second segments to either the derivation circuitry 420 or the decision circuitry 440).
- the derivation circuitry 420 is configured to analyze the speech data from the first segments 414 (e.g., received from the VAD circuitry 410) for the one or more words 422 contained within the speech data.
- the derivation circuitry 420 can be configured to perform speech-to-text conversion (e.g., using a speech-to-text engine or application programming interface, examples of which are available from Google and Amazon) and/or other speech recognition processes (e.g., translation from one language into another).
- the derivation circuitry 420 can be configured to extract the one or more words 422 from the speech data in a form (e.g., text) compatible with further processing and/or with communication to the recipient as described herein.
- the derivation circuitry 420 is configured to transmit the one or more words 422 to the keyword detection circuitry 430 and the decision circuitry 440. In certain other implementations, the derivation circuitry 420 is configured to transmit the one or more words 422 to the keyword detection circuitry 430 and the keyword detection circuitry 430 is configured to transmit the one or more words 422 to the decision circuitry 440. In certain implementations, the derivation circuitry 420 is part of the VAD circuitry 410 or vice versa.
- the derivation circuitry 420 is local (e.g., a component of the device 410 and/or the external device 420), while in certain other implementations, the derivation circuitry 420 is part of a remote server (e.g., “in the cloud”).
- the keyword detection circuitry 430 is configured to receive the one or more words 422 (e.g., from the derivation circuitry 420), to retrieve the set of stored keywords 434 from memory circuitry and to compare the one or more words 422 to keywords of the set of stored keywords 434 (e.g., to determine the relevance of the first segment 414 to the user or recipient).
- the set of stored keywords 434 e.g., a keyword list
- the apparatus 400 can access a plurality of sets of stored keywords 434 (e.g., different sets of stored keywords 434 for different broadcast streams 412, different broadcast systems 330, and/or different times of day) and one or more of the sets of stored keywords 434 can change (e.g., edited automatically or by the recipient) over time.
- the set of stored keywords 434 to be accessed for the comparison with the one or more words 422 can be selected based, at least in part, on the identity of the currently-received broadcast stream 412 and/or the identity of the broadcast system 330 broadcasting the currently-received broadcast stream 412.
- the keyword detection circuitry 430 can access a set of stored keywords 434 that is compatible for comparison with keywords expected to be within the broadcast stream 412 (e.g., gate changes; schedule changes).
- the keyword detection circuitry 430 is in operative communication with keyword generation circuitry 470 configured to generate at least some keywords of the set of stored keywords 434 to be accessed by the keyword detection circuitry 430.
- the keyword generation circuitry 470 is a component of the keyword detection circuitry 430 or is another component of the apparatus 400.
- the keyword generation circuitry 470 of certain implementations is in operative communication with at least one input interface 480 configured to receive input information 482, and the keyword generation circuitry 470 is configured to generate the set of stored keywords 434 at least partially based on the input information 482.
- the input information 482 can comprise information provided by the recipient (e.g., user input; manually entered via a keyboard or touchscreen; verbally entered via a microphone) indicative of keywords of interest to the recipient, information from a clock, calendar, or other software application of the device 310 and/or external device 320 (e.g., a clock/calendar app providing information regarding scheduled events and/or time of day; ticketing app providing information regarding information regarding stored tickets; geolocating app providing information regarding the recipient’s location, such as work or transportation station) from which the keyword generation circuitry 470 can extract keywords, and/or other information from which the keyword generation circuitry 470 can extract (e.g., pluck; scrape) keywords or keyword-relevant information.
- a clock/calendar app providing information regarding scheduled events and/or time of day
- ticketing app providing information regarding information regarding stored tickets
- geolocating app providing information regarding the recipient’s location, such as work or transportation station
- the keyword generation circuitry 470 is configured to generate the set of stored keywords 434 automatically (e.g., based on learned behavior, such as using a set of stored keywords 434 that was previously used when the apparatus 400 was previously receiving a broadcast stream 412 from the same broadcast system 330 that is providing the currently-received broadcast stream 412) and/or based on predetermined rules (e.g., words such as “evacuate” and “emergency” automatically being included in the set of stored keywords 434).
- predetermined rules e.g., words such as “evacuate” and “emergency” automatically being included in the set of stored keywords 434.
- the set of stored keywords 434 comprises, for each stored keyword 434, information indicative of an importance of the stored keyword 434.
- the keyword generation circuitry 470 of certain implementations is in operative communication with at least one input interface 490 configured to receive input information 492, and the keyword generation circuitry 470 is configured to generate the set of stored keywords 434 at least partially based on the input information 492.
- the at least one input interface 490 and the at least one input interface 480 can be the same as one another or can be separate from one another.
- the importance of a keyword is indicative of its relative importance as compared to other keywords.
- the input information 492 can comprise information provided by the recipient (e.g., user input; manually entered via a keyboard or touchscreen; verbally entered via a microphone) indicative of the importance of one or more keywords of interest to the recipient, information from a clock, calendar, or other software application of the device 310 and/or external device 320 (e.g., a clock/calendar app providing information regarding scheduled events and/or time of day; ticketing app providing information regarding stored tickets; geolocating app providing information regarding the recipient’s location) from which the keyword generation circuitry 470 can extract the importance of one or more keywords, and/or other information from which the keyword generation circuitry 470 can extract (e.g., pluck; scrape) the importance of one or more keywords.
- the recipient e.g., user input; manually entered via a keyboard or touchscreen; verbally entered via a microphone
- information from a clock, calendar, or other software application of the device 310 and/or external device 320 e.g., a clock/calendar app providing information regarding scheduled events and
- the keyword generation circuitry 470 is configured to assign an importance to one or more keywords of the set of stored keywords 434 automatically (e.g., based on learned or past behavior, such as an importance of a keyword 434 that was previously used when the apparatus 400 was previously receiving a broadcast stream 412 from the same broadcast system 330 that is providing the currently-received broadcast stream 412) and/or based on predetermined rules (e.g., keywords such as “evacuate” and “emergency” automatically having the highest level of importance).
- predetermined rules e.g., keywords such as “evacuate” and “emergency” automatically having the highest level of importance.
- the decision circuitry 440 is configured to, in response at least in part to the keyword information 432 (e.g., received from the keyword detection circuitry 430) corresponding to the first segment 414, select whether any information 442 indicative of the first segment 414 is to be communicated to the recipient.
- the decision circuitry 440 is configured to compare the keyword information 432 for a first segment 414 to a predetermined set of rules to determine whether the first segment 414 is of sufficient interest (e.g., importance) to the recipient to warrant communication to the recipient. If the keyword information 432 indicates that the first segment 414 is not of sufficient interest, the decision circuitry 440 does not generate any information 442 regarding the first segment 414. If the keyword information 432 indicates that the first segment 414 is of sufficient interest, the decision circuitry 440 generates the information 442 regarding the first segment 414.
- the decision circuitry 440 in response at least in part to the keyword information 432 corresponding to the first segment 414, can select among the data output interfaces 460 and can select the form and/or content of the information 442 indicative of the first segment 414 to be communicated to the recipient.
- the first segments 414 and/or the one or more words 422 comprise at least part of the content of the information 422 to be communicated to the recipient via the data output interfaces 460.
- the decision circuitry 440 can transmit the information 442 in the form of at least one text message indicative of the one or more words 422 of the first segment 414 to a data output interface 460a configured to receive the information 442 and to communicate the information 442 a screen configured to display the at least one text message to the recipient.
- the decision circuitry 440 can transmit the information 442 in the form of at least one signal indicative of a notification (e.g., alert; alarm) regarding the information 442 (e.g., indicative of whether the one or more words 422 of the first segment 414 comprises a stored keyword 434, indicative of an identification of the stored keyword 434, and/or indicative of an importance of the stored keyword 434) to a data output interface 460b configured to receive the at least one signal and to communicate the notification to the recipient as at least one visual signal (e.g., outputted by an indicator light or display screen), at least one audio signal (e.g., outputted as a tone or other sound from a speaker), and/or at least one tactile or haptic signal (e.g., outputted as a vibration from a motor).
- a notification e.g., alert; alarm
- a data output interface 460b configured to receive the at least one signal and to communicate the notification to the recipient as at least one visual signal (e.g., outputted by an indicator light or
- the decision circuitry 440 can transmit the information 442 in the form of at least one signal indicative of the audio data of the first segment 414 to a data output interface 460c configured to receive the at least one signal and to communicate the audio data to the recipient (e.g., outputted as sound from a speaker, such as a hearing aid or headphone; outputted as stimulation signals from a hearing prosthesis).
- a data output interface 460c configured to receive the at least one signal and to communicate the audio data to the recipient (e.g., outputted as sound from a speaker, such as a hearing aid or headphone; outputted as stimulation signals from a hearing prosthesis).
- the decision circuitry 440 can transmit the information 442 in the form of at least one signal compatible for storage to a data output interface 460d configured to receive the at least one signal and to communicate the information 442 to memory circuitry (e.g., at least one storage device, such as flash memory) to be stored and subsequently retrieved and communicated to the recipient (e.g., via one or more of the other data output interfaces 460a-c).
- memory circuitry e.g., at least one storage device, such as flash memory
- the decision circuitry 440 can be further configured to track the intent of the first segment 414 over time and can correspondingly manage the queue of information 442 in the memory circuitry (e.g., deleting older information 442 upon receiving newer information 442 about the same topic; learning the intent and/or interests of the user over time and stopping notifications to the user for certain types of information 442 not of interest).
- One or more of the data output interfaces 460 can be configured to receive the information 442 in multiple forms and/or can be configured to be in operative communication with multiple communication components. Other types of data output interfaces 460 (e.g., interfaces to other communication components) are also compatible with certain implementations described herein.
- FIG. 5A is a flow diagram of an example method 500 in accordance with certain implementations described herein. While the method 500 is described by referring to some of the structures of the example apparatus 400 of FIGs. 4A-4B, other apparatus and systems with other configurations of components can also be used to perform the method 500 in accordance with certain implementations described herein.
- a non-transitory computer readable storage medium has stored thereon a computer program that instructs a computer system to perform the method 500.
- the method 500 comprises receiving one or more electromagnetic wireless broadcast streams 412 (e.g., at least one Bluetooth broadcast stream from at least one remote broadcast system 330) comprising audio data.
- the one or more electromagnetic wireless broadcast streams 412 can be received by a personal electronic device (e.g., external device 320) worn, held, and/or carried by the user or implanted on or within the user’s body (e.g., device 310).
- the method 500 further comprises dividing the one or more broadcast streams 412 into a plurality of segments comprising speech-including segments (e.g., first segments 414) and speech-excluding segments.
- FIG. 5B is a flow diagram of an example of the operational block 520 in accordance with certain implementations described herein.
- dividing the one or more broadcast streams 412 can comprise detecting at least one characteristic (e.g., modulation depth; signal-to-noise ratio; zero crossing rate; cross correlations; sub-band/full-band energy measures; spectral structure in frequency range corresponding to speech (e.g., 80 Hz to 400 Hz); long term time-domain behavior characteristics) for each segment of the plurality of segments.
- at least one characteristic e.g., modulation depth; signal-to-noise ratio; zero crossing rate; cross correlations; sub-band/full-band energy measures; spectral structure in frequency range corresponding to speech (e.g., 80 Hz to 400 Hz); long term time-domain behavior characteristics
- dividing the one or more broadcast streams 412 can further comprise determining, for each segment of the plurality of segments, whether the at least one characteristic is indicative of either the segment being a speech-including segment or a speech-excluding segment.
- dividing the one or more broadcast streams 412 can further comprise appending information to at least some of the segments, the information indicative of whether the segment is a speech-including segment or a speech-excluding segment.
- dividing the one or more broadcast streams 412 can further comprise excluding the speech-excluding segments from further processing in an operational block 528.
- the method 500 further comprises evaluating the audio data of each speech-including segment for inclusion of at least one keyword 434.
- FIG. 5C is a flow diagram of an example of the operational block 530 in accordance with certain implementations described herein.
- evaluating the audio data can comprise extracting one or more words 422 from the audio data of the speech- including segment.
- evaluating the audio data can further comprise comparing the one or more words 422 to a set of keywords 434 to detect the at least one keyword 434 within the one or more words 422.
- the set of keywords 434 can be compiled from at least one of: user input, time of day, user’s geographic location when the speech- including segment is received, history of previous user input, and/or information from computer memory or one or more computing applications.
- evaluating the audio data can further comprise appending information to at least some of the speech-including segments, the information indicative of existence and/or identity of the detected at least one keyword 434 within the one or more words 422 of the speech-including segment.
- evaluating the audio data can further comprise assigning an importance level to the speech-including segment in an operational block 538.
- the importance level can be based at least in part on existence and/or identity of the at least one keyword, user input, time of day, user’s geographic location when the speech-including segment is received, history of previous user input, and/or information from computer memory or one or more computing applications.
- the method 500 further comprises, based on said evaluating, communicating information regarding the speech-including segment to a user. For example, based on whether the one or more words 422 includes at least one keyword 434, the identity of the included at least one keyword 434, and/or the importance level of the speech- including segment, the information regarding the speech-including segment can be selected to be communicated to the user or to not be communicated to the user.
- said communicating information can be selected from the group consisting of: displaying at least one text message to the user, the at least one text message indicative of the one or more words of the speech-including segment; providing at least one visual, audio, and/or tactile signal to the user, the at least one visual, audio, and/or tactile signal indicative of whether the speech-including segment comprises a keyword, an identification of the keyword, and/or an importance of the keyword; providing at least one signal indicative of the audio data of the speech-including segment to the user; and storing at least one signal indicative of the audio data of the speech-including segment in memory circuitry, and subsequently retrieving the stored at least one signal from the memory circuitry and providing the stored at least one signal to the user.
- a recipient with a hearing prosthesis (e.g., device 310) with an external sound processor (e.g., external device 320) and a mobile device (e.g., smart phone; smart watch; another external device 320) in communication with the sound processor in accordance with certain implementations described herein can enter an airport where a location-based Bluetooth wireless broadcast (e.g., broadcast stream 412) is being used to mirror the normal announcements made over the speaker system.
- the mobile device can connect to the wireless broadcast (e.g., received via the data input interface 450) and can be toggled into a mode of operation (e.g., “smart mode”) enabling the functionality of certain implementations described herein.
- the recipient can enter keywords corresponding to the flight information (e.g., airline, flight number, gate number) and/or other relevant information into a dialog box of key terms via an input interface 480.
- the mobile device can receive announcements from the wireless broadcast, split them into segments, and check for one or more of the keywords.
- a gate change for the recipient’s flight number can be announced, and the mobile device can store this announcement in audio form and can notify the recipient via a tone (e.g., triple ascending beep) via the hearing prosthesis.
- the recipient can select to hear the announcement when the recipient chooses (e.g., once the recipient is done ordering a coffee; by pressing a button on the mobile device), and the mobile device can stream the stored audio of the announcement to the sound processor of the recipient’s hearing prosthesis.
- the recipient can also select to replay the announcement when the recipient chooses (e.g., by pressing the button again within five seconds of completion of the streaming of the stored audio the previous time).
- the recipient can also select to receive a text version of the announcement (e.g., if text is more convenient for the recipient; if the streaming of the stored audio is unclear to the recipient).
- a recipient with a hearing prosthesis (e.g., device 310) with an external sound processor (e.g., external device 320) and a mobile device (e.g., smart phone; smart watch; another external device 320) in communication with the sound processor in accordance with certain implementations described herein can enter a mass transit train station where a location-based Bluetooth wireless broadcast (e.g., broadcast stream 412) is being used to mirror the normal announcements made over the speaker system.
- a location-based Bluetooth wireless broadcast e.g., broadcast stream 412
- the station can be one that the recipient is at every workday morning to ride the same commuter train, and the mobile device can present a notification pop-up text message offering to connect to the station’s wireless broadcast (e.g., to receive the wireless broadcast via the data input interface 450) and to enable the functionality of certain implementations described herein.
- the mobile device can access keywords relevant to the recipient’s normal commuter train (e.g., name; time; track; platform). These keywords can be received from input from the recipient, automatically from information obtained from a calendar application on the mobile device, and/or automatically from previously-stored keywords corresponding to previous commutes by the recipient.
- the announcement can be presented to the recipient via a warning buzz by the mobile device followed by a text message informing the recipient of the platform change. The recipient can then go to the new platform without interruption of the music that the recipient had been listening to.
- a recipient with a hearing prosthesis with an external sound processor (e.g., external device 320) and a mobile device (e.g., smart phone; smart watch; another external device 320) in communication with the sound processor in accordance with certain implementations described herein can attend an event with their family where a location-based Bluetooth wireless broadcast (e.g., broadcast stream 412) is being used to mirror the normal announcements made over the speaker system.
- the announcements can be about the location of certain keynote talks, and the recipient can scroll through a list of these announcements, with the most recent announcements appearing at the top of the list in real-time.
- the recipient can configure the mobile device to not play audible notifications for this category of announcements, but to play audible notifications for one or more second categories of announcements having higher importance to the recipient (e.g., announcements including one or more keywords having higher priority or importance over others). If an announcement is broadcast referring to the recipient’s automobile by its license plate number (e.g., an automobile with the license plate number is about to be towed), because the recipient had previously entered the license plate number in a list of high-priority keywords, the announcement can trigger an audible notification for the recipient so the recipient can immediately check it and respond.
- an announcement is broadcast referring to the recipient’s automobile by its license plate number (e.g., an automobile with the license plate number is about to be towed)
- the announcement can trigger an audible notification for the recipient so the recipient can immediately check it and respond.
- the terms “generally parallel” and “substantially parallel” refer to a value, amount, or characteristic that departs from exactly parallel by ⁇ 10 degrees, by ⁇ 5 degrees, by ⁇ 2 degrees, by ⁇ 1 degree, or by ⁇ 0.1 degree
- the terms “generally perpendicular” and “substantially perpendicular” refer to a value, amount, or characteristic that departs from exactly perpendicular by ⁇ 10 degrees, by ⁇ 5 degrees, by ⁇ 2 degrees, by ⁇ 1 degree, or by ⁇ 0.1 degree.
- the ranges disclosed herein also encompass any and all overlap, sub-ranges, and combinations thereof. Language such as “up to,” “at least,” “greater than,” less than,” “between,” and the like includes the number recited.
- ordinal adjectives e.g., first, second, etc.
- the ordinal adjective are used merely as labels to distinguish one element from another (e.g., one signal from another or one circuit from one another), and the ordinal adjective is not used to denote an order of these elements or of their use.
Landscapes
- Engineering & Computer Science (AREA)
- Health & Medical Sciences (AREA)
- Physics & Mathematics (AREA)
- Acoustics & Sound (AREA)
- Signal Processing (AREA)
- Computer Networks & Wireless Communication (AREA)
- General Health & Medical Sciences (AREA)
- Otolaryngology (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Computational Linguistics (AREA)
- Human Computer Interaction (AREA)
- Multimedia (AREA)
- Neurosurgery (AREA)
- Prostheses (AREA)
Abstract
Description
Claims
Priority Applications (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US18/556,177 US20240185881A1 (en) | 2021-05-18 | 2022-05-04 | System and method for smart broadcast management |
| CN202280032496.3A CN117242518A (en) | 2021-05-18 | 2022-05-04 | Systems and methods for intelligent broadcast management |
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US202163190112P | 2021-05-18 | 2021-05-18 | |
| US63/190,112 | 2021-05-18 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2022243778A1 true WO2022243778A1 (en) | 2022-11-24 |
Family
ID=84141144
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/IB2022/054124 Ceased WO2022243778A1 (en) | 2021-05-18 | 2022-05-04 | System and method for smart broadcast management |
Country Status (3)
| Country | Link |
|---|---|
| US (1) | US20240185881A1 (en) |
| CN (1) | CN117242518A (en) |
| WO (1) | WO2022243778A1 (en) |
Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20170359661A1 (en) * | 2016-06-08 | 2017-12-14 | Michael Goorevich | Electro-acoustic adaption in a hearing prosthesis |
| US20190009090A1 (en) * | 2003-12-22 | 2019-01-10 | Gunther Van der Borght | Hearing prosthesis system having interchangeable housings |
| US20190325862A1 (en) * | 2018-04-23 | 2019-10-24 | Eta Compute, Inc. | Neural network for continuous speech segmentation and recognition |
| US20200222697A1 (en) * | 2014-11-21 | 2020-07-16 | Cochlear Limited | Systems and methods for non-obtrusive adjustment of auditory prostheses |
| US20210058720A1 (en) * | 2018-01-16 | 2021-02-25 | Cochlear Limited | Individualized own voice detection in a hearing prosthesis |
Family Cites Families (11)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| DK200400191A (en) * | 2004-02-09 | 2005-08-10 | Gn Netcom As | Amplifier coupling connected between a landline telephone and a headset |
| FR2932920A1 (en) * | 2008-06-19 | 2009-12-25 | Archean Technologies | METHOD AND APPARATUS FOR MEASURING THE INTELLIGIBILITY OF A SOUND DIFFUSION DEVICE |
| KR101759190B1 (en) * | 2011-01-04 | 2017-07-19 | 삼성전자주식회사 | Method for reporting emergency stateduring call service in portable wireless terminal and apparatus thereof |
| US10242664B2 (en) * | 2014-03-31 | 2019-03-26 | NetTalk.com, Inc. | System and method for processing flagged words or phrases in audible communications |
| US10212263B2 (en) * | 2017-01-16 | 2019-02-19 | Lenovo (Singapore) Pte. Ltd. | Notifying a user of external audio |
| US10817252B2 (en) * | 2018-03-10 | 2020-10-27 | Staton Techiya, Llc | Earphone software and hardware |
| US10791404B1 (en) * | 2018-08-13 | 2020-09-29 | Michael B. Lasky | Assisted hearing aid with synthetic substitution |
| US11195518B2 (en) * | 2019-03-27 | 2021-12-07 | Sonova Ag | Hearing device user communicating with a wireless communication device |
| US10997970B1 (en) * | 2019-07-30 | 2021-05-04 | Abbas Rafii | Methods and systems implementing language-trainable computer-assisted hearing aids |
| WO2021159369A1 (en) * | 2020-02-13 | 2021-08-19 | 深圳市汇顶科技股份有限公司 | Hearing aid method and apparatus for noise reduction, chip, earphone and storage medium |
| WO2022054978A1 (en) * | 2020-09-09 | 2022-03-17 | 올리브유니온(주) | Smart hearing device and artificial intelligence hearing system for distinguishing natural language or non-natural language, and method therefor |
-
2022
- 2022-05-04 US US18/556,177 patent/US20240185881A1/en active Pending
- 2022-05-04 CN CN202280032496.3A patent/CN117242518A/en active Pending
- 2022-05-04 WO PCT/IB2022/054124 patent/WO2022243778A1/en not_active Ceased
Patent Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20190009090A1 (en) * | 2003-12-22 | 2019-01-10 | Gunther Van der Borght | Hearing prosthesis system having interchangeable housings |
| US20200222697A1 (en) * | 2014-11-21 | 2020-07-16 | Cochlear Limited | Systems and methods for non-obtrusive adjustment of auditory prostheses |
| US20170359661A1 (en) * | 2016-06-08 | 2017-12-14 | Michael Goorevich | Electro-acoustic adaption in a hearing prosthesis |
| US20210058720A1 (en) * | 2018-01-16 | 2021-02-25 | Cochlear Limited | Individualized own voice detection in a hearing prosthesis |
| US20190325862A1 (en) * | 2018-04-23 | 2019-10-24 | Eta Compute, Inc. | Neural network for continuous speech segmentation and recognition |
Also Published As
| Publication number | Publication date |
|---|---|
| CN117242518A (en) | 2023-12-15 |
| US20240185881A1 (en) | 2024-06-06 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| CN110072434B (en) | Use of acoustic biomarkers to assist hearing device use | |
| US8170677B2 (en) | Recording and retrieval of sound data in a hearing prosthesis | |
| US11189265B2 (en) | Systems and methods for assisting the hearing-impaired using machine learning for ambient sound analysis and alerts | |
| US20110093039A1 (en) | Scheduling information delivery to a recipient in a hearing prosthesis | |
| US12347554B2 (en) | Dynamic virtual hearing modelling | |
| WO2020261148A1 (en) | Prediction and identification techniques used with a hearing prosthesis | |
| EP3930346A1 (en) | A hearing aid comprising an own voice conversation tracker | |
| US20190149928A1 (en) | Hearing aid configured to be operating in a communication system | |
| EP2876899A1 (en) | Adjustable hearing aid device | |
| CN113195043A (en) | Evaluating responses to sensory events and performing processing actions based thereon | |
| CN108141681A (en) | functional migration | |
| EP4210646A1 (en) | New tinnitus management techniques | |
| US20170359661A1 (en) | Electro-acoustic adaption in a hearing prosthesis | |
| US12348933B2 (en) | Audio training | |
| EP2876902A1 (en) | Adjustable hearing aid device | |
| WO2020084342A1 (en) | Systems and methods for customizing auditory devices | |
| US20240185881A1 (en) | System and method for smart broadcast management | |
| CN111133774B (en) | Acoustic point recognition | |
| US9901736B2 (en) | Cochlea hearing aid fixed on eardrum | |
| US12375196B2 (en) | Broadcast selection | |
| US20240430626A1 (en) | Method for operating a hearing device, and hearing device | |
| WO2025041006A1 (en) | Audio processing device operable as remote sensor | |
| EP2835983A1 (en) | Hearing instrument presenting environmental sounds | |
| Satpute et al. | 15 MaximizingPeoplewith ParticipationHearing for | |
| WO2025078920A1 (en) | System and method to facilitate finding a misplaced device |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 22804132 Country of ref document: EP Kind code of ref document: A1 |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 18556177 Country of ref document: US |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 202280032496.3 Country of ref document: CN |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 22804132 Country of ref document: EP Kind code of ref document: A1 |