MX2022008071A - Sistemas y metodos para mezclar automaticamente audio para escenas acusticas. - Google Patents

Sistemas y metodos para mezclar automaticamente audio para escenas acusticas.

Info

Publication number
MX2022008071A
MX2022008071A MX2022008071A MX2022008071A MX2022008071A MX 2022008071 A MX2022008071 A MX 2022008071A MX 2022008071 A MX2022008071 A MX 2022008071A MX 2022008071 A MX2022008071 A MX 2022008071A MX 2022008071 A MX2022008071 A MX 2022008071A
Authority
MX
Mexico
Prior art keywords
audio sample
systems
methods
obtaining
impulse response
Prior art date
Application number
MX2022008071A
Other languages
English (en)
Inventor
Shilpa Jois Rao
Yadong Wang
Murthy Parthasarathi
Kyle Tacke
Original Assignee
Netflix Inc
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Netflix Inc filed Critical Netflix Inc
Publication of MX2022008071A publication Critical patent/MX2022008071A/es

Links

Classifications

    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/48Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use
    • G10L25/51Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use for comparison or discrimination
    • GPHYSICS
    • G11INFORMATION STORAGE
    • G11BINFORMATION STORAGE BASED ON RELATIVE MOVEMENT BETWEEN RECORD CARRIER AND TRANSDUCER
    • G11B27/00Editing; Indexing; Addressing; Timing or synchronising; Monitoring; Measuring tape travel
    • G11B27/10Indexing; Addressing; Timing or synchronising; Measuring tape travel
    • G11B27/19Indexing; Addressing; Timing or synchronising; Measuring tape travel by using information detectable on the record carrier
    • G11B27/28Indexing; Addressing; Timing or synchronising; Measuring tape travel by using information detectable on the record carrier by using information signals recorded by the same method as the main recording
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06NCOMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
    • G06N3/00Computing arrangements based on biological models
    • G06N3/02Neural networks
    • G06N3/08Learning methods
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06NCOMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
    • G06N3/00Computing arrangements based on biological models
    • G06N3/02Neural networks
    • G06N3/08Learning methods
    • G06N3/09Supervised learning
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10HELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
    • G10H1/00Details of electrophonic musical instruments
    • G10H1/0091Means for obtaining special acoustic effects
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10KSOUND-PRODUCING DEVICES; METHODS OR DEVICES FOR PROTECTING AGAINST, OR FOR DAMPING, NOISE OR OTHER ACOUSTIC WAVES IN GENERAL; ACOUSTICS NOT OTHERWISE PROVIDED FOR
    • G10K15/00Acoustics not otherwise provided for
    • G10K15/08Arrangements for producing a reverberation or echo sound
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L15/00Speech recognition
    • G10L15/005Language recognition
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L15/00Speech recognition
    • G10L15/22Procedures used during a speech recognition process, e.g. man-machine dialogue
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02Speech enhancement, e.g. noise reduction or echo cancellation
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/78Detection of presence or absence of voice signals
    • G10L25/81Detection of presence or absence of voice signals for discriminating voice from music
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/78Detection of presence or absence of voice signals
    • G10L25/84Detection of presence or absence of voice signals for discriminating voice from noise
    • GPHYSICS
    • G11INFORMATION STORAGE
    • G11BINFORMATION STORAGE BASED ON RELATIVE MOVEMENT BETWEEN RECORD CARRIER AND TRANSDUCER
    • G11B27/00Editing; Indexing; Addressing; Timing or synchronising; Monitoring; Measuring tape travel
    • G11B27/02Editing, e.g. varying the order of information signals recorded on, or reproduced from, record carriers
    • G11B27/031Electronic editing of digitised analogue information signals, e.g. audio or video signals
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10HELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
    • G10H2210/00Aspects or methods of musical processing having intrinsic musical character, i.e. involving musical theory or musical parameters or relying on musical knowledge, as applied in electrophonic musical tools or instruments
    • G10H2210/155Musical effects
    • G10H2210/265Acoustic effect simulation, i.e. volume, spatial, resonance or reverberation effects added to a musical sound, usually by appropriate filtering or delays
    • G10H2210/281Reverberation or echo
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10HELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
    • G10H2250/00Aspects of algorithms or signal processing methods without intrinsic musical character, yet specifically adapted for or used in electrophonic musical processing
    • G10H2250/311Neural networks for electrophonic musical instruments or musical processing, e.g. for musical recognition or control, automatic composition or improvisation
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04SSTEREOPHONIC SYSTEMS 
    • H04S2400/00Details of stereophonic systems covered by H04S but not provided for in its groups
    • H04S2400/15Aspects of sound capture and related signal processing for recording or reproduction
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04SSTEREOPHONIC SYSTEMS 
    • H04S7/00Indicating arrangements; Control arrangements, e.g. balance control
    • H04S7/30Control circuits for electronic adaptation of the sound field
    • H04S7/305Electronic adaptation of stereophonic audio signals to reverberation of the listening space
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04SSTEREOPHONIC SYSTEMS 
    • H04S7/00Indicating arrangements; Control arrangements, e.g. balance control
    • H04S7/30Control circuits for electronic adaptation of the sound field
    • H04S7/307Frequency adjustment, e.g. tone control

Landscapes

  • Engineering & Computer Science (AREA)
  • Physics & Mathematics (AREA)
  • Multimedia (AREA)
  • Health & Medical Sciences (AREA)
  • Acoustics & Sound (AREA)
  • Computational Linguistics (AREA)
  • Human Computer Interaction (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Signal Processing (AREA)
  • Theoretical Computer Science (AREA)
  • Biophysics (AREA)
  • Data Mining & Analysis (AREA)
  • Computing Systems (AREA)
  • General Engineering & Computer Science (AREA)
  • General Physics & Mathematics (AREA)
  • Mathematical Physics (AREA)
  • Software Systems (AREA)
  • General Health & Medical Sciences (AREA)
  • Evolutionary Computation (AREA)
  • Molecular Biology (AREA)
  • Biomedical Technology (AREA)
  • Artificial Intelligence (AREA)
  • Life Sciences & Earth Sciences (AREA)
  • Quality & Reliability (AREA)
  • Management Or Editing Of Information On Record Carriers (AREA)
  • Circuit For Audible Band Transducer (AREA)
  • Circuits Of Receivers In General (AREA)

Abstract

El método implementado por computadora divulgado puede incluir obtener una muestra de audio a partir de una fuente de contenido, ingresar la muestra de audio obtenida en un modelo de aprendizaje por máquina entrenado, obtener la salida del modelo de aprendizaje por máquina entrenado, en donde la salida es un perfil de un ambiente en el cual se grabó la muestra de audio de entrada, obtener una respuesta de impulso acústico correspondiente al perfil del ambiente en el cual se grabó la muestra de audio de entrada, obtener una segunda muestra de audio, procesar la respuesta de impulso acústico obtenida con la segunda muestra de audio, e insertar un resultado de procesar la respuesta de impulso acústico obtenida y la segunda muestra de audio en una pista de audio. También se divulgan diversos otros métodos, sistemas, y medios legibles por computadora.
MX2022008071A 2019-12-31 2020-12-31 Sistemas y metodos para mezclar automaticamente audio para escenas acusticas. MX2022008071A (es)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
US16/732,142 US11238888B2 (en) 2019-12-31 2019-12-31 System and methods for automatically mixing audio for acoustic scenes
PCT/US2020/067661 WO2021138557A1 (en) 2019-12-31 2020-12-31 System and methods for automatically mixing audio for acoustic scenes

Publications (1)

Publication Number Publication Date
MX2022008071A true MX2022008071A (es) 2022-07-27

Family

ID=74285591

Family Applications (1)

Application Number Title Priority Date Filing Date
MX2022008071A MX2022008071A (es) 2019-12-31 2020-12-31 Sistemas y metodos para mezclar automaticamente audio para escenas acusticas.

Country Status (7)

Country Link
US (2) US11238888B2 (es)
EP (1) EP4085456A1 (es)
AU (1) AU2020417822B2 (es)
BR (1) BR112022012975A2 (es)
CA (1) CA3160724A1 (es)
MX (1) MX2022008071A (es)
WO (1) WO2021138557A1 (es)

Families Citing this family (9)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US12051438B1 (en) * 2021-03-26 2024-07-30 T-Mobile Usa, Inc. Using machine learning to locate mobile device
US11705148B2 (en) * 2021-06-11 2023-07-18 Microsoft Technology Licensing, Llc Adaptive coefficients and samples elimination for circular convolution
US12579984B2 (en) 2022-01-20 2026-03-17 Microsoft Technology Licensing, Llc. Data augmentation system and method for multi-microphone systems
US12456456B2 (en) * 2022-01-20 2025-10-28 Microsoft Technology Licensing, Llc Data augmentation system and method for multi-microphone systems
US12469513B2 (en) * 2022-12-06 2025-11-11 Microsoft Technology Licensing, Llc System and method for replicating background acoustic properties using neural networks
CN116861182A (zh) * 2023-06-14 2023-10-10 钉钉(中国)信息技术有限公司 房间声学冲激响应的估计方法、训练方法及装置
US12554943B2 (en) * 2023-07-14 2026-02-17 Robert Bosch Gmbh System and method for anomaly detection in unlabeled collections of audio recording
CN118737172B (zh) * 2024-07-25 2025-09-02 武汉大学 基于拾音环境因素采集的音频数据增强方法、装置及介质
CN119479679B (zh) * 2024-11-19 2025-07-04 嘉兴市科讯电子有限公司 舰船通讯设备中音频失真调节装置及方法

Family Cites Families (12)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20070213987A1 (en) 2006-03-08 2007-09-13 Voxonic, Inc. Codebook-less speech conversion method and system
US20130151251A1 (en) 2011-12-12 2013-06-13 Advanced Micro Devices, Inc. Automatic dialog replacement by real-time analytic processing
LV14747B (lv) * 2012-04-04 2014-03-20 Sonarworks, Sia Elektroakustisko izstarotāju akustisko parametru korekcijas paņēmiens un iekārta tā realizēšanai
US9449613B2 (en) 2012-12-06 2016-09-20 Audeme Llc Room identification using acoustic features in a recording
US9185199B2 (en) * 2013-03-12 2015-11-10 Google Technology Holdings LLC Method and apparatus for acoustically characterizing an environment in which an electronic device resides
US10063965B2 (en) * 2016-06-01 2018-08-28 Google Llc Sound source estimation using neural networks
CN108780643B (zh) 2016-11-21 2023-08-25 微软技术许可有限责任公司 自动配音方法和装置
US10991379B2 (en) 2018-06-22 2021-04-27 Babblelabs Llc Data driven audio enhancement
CN109119063B (zh) 2018-08-31 2019-11-22 腾讯科技(深圳)有限公司 视频配音生成方法、装置、设备及存储介质
US11112389B1 (en) * 2019-01-30 2021-09-07 Facebook Technologies, Llc Room acoustic characterization using sensors
US11074925B2 (en) * 2019-11-13 2021-07-27 Adobe Inc. Generating synthetic acoustic impulse responses from an acoustic impulse response
US11545134B1 (en) * 2019-12-10 2023-01-03 Amazon Technologies, Inc. Multilingual speech translation with adaptive speech synthesis and adaptive physiognomy

Also Published As

Publication number Publication date
BR112022012975A2 (pt) 2022-09-13
US20210201931A1 (en) 2021-07-01
CA3160724A1 (en) 2021-07-08
AU2020417822A1 (en) 2022-06-23
WO2021138557A1 (en) 2021-07-08
US11238888B2 (en) 2022-02-01
US20220115030A1 (en) 2022-04-14
EP4085456A1 (en) 2022-11-09
AU2020417822B2 (en) 2023-07-06

Similar Documents

Publication Publication Date Title
BR112022012975A2 (pt) Sistema e métodos para mixagem automática de áudio para cenas acústicas
AU2016331881A8 (en) Q-compensated full wavefield inversion
EP4488992A3 (en) Adaptive anc based on enironmental triggers
MX2016013015A (es) Métodos y sistemas de administrar un dialogo con un robot.
GB2590555B (en) Methods for characterizing and evaluating well integrity using unsupervised machine learning of acoustic data
MX2021014721A (es) Sistemas y metodos para aprendizaje de maquina de atributos de voz.
WO2021127660A3 (en) Machine and deep learning process modeling of performance and behavioral data
SG10201707702YA (en) Collaborative Voice Controlled Devices
MX2016014193A (es) Caracterizacion de entorno del interior del pozo mediante el uso de coeficientes de rigidez.
NZ725145A (en) Methods and systems for managing dialogs of a robot
EP4679426A3 (en) Text-to-speech synthesis system and method
GB2554601A (en) Method for analyzing cement integrity in casing strings using machine learning
GB2542054A (en) Virtual simulation of spatial audio characteristics
MX385727B (es) Método y aparato para generar una señal de audio filtrada realizando representación de elevación.
RU2016106913A (ru) Обработка пространственно дифузных или больших звуковых объектов
EP4425488A3 (en) Acoustic model training using corrected terms
AU2015364405A8 (en) Methods for simultaneous source separation
DE502008003378D1 (de) Vorrichtung und verfahren zum erzeugen eines multikanalsignals mit einer sprachsignalverarbeitung
GB2565965A (en) Method and system for determining equity index for a brand
WO2017106610A8 (en) Method and system for providing automated localized feedback for an extracted component of an electronic document file
MX2021012019A (es) Deteccion rapida de fusiones genicas.
JP2017511896A5 (es)
EP3185133A3 (en) Computing device and corresponding method for generating data representing text
KR102203048B9 (ko) 유사도를 이용한 교육 프로그램 추천방법 및 그 시스템
MX383427B (es) Aparato y método para mejorar una transición desde una porción de señal de audio oculta hasta una porción de señal de audio subsiguiente de una señal de audio.