Patent search ap:("Adobe Inc.") AND inv:"Zeyu Jin" Page 1

1.

发明申请
Searching for Music 有权

公开(公告)号：US20210294840A1

公开(公告)日：2021-09-23

申请号：US16823538

申请日：2020-03-19

Applicant: Adobe Inc.

Inventor： Jongpil Lee , Nicholas J. Bryan , Justin J. Salamon , Zeyu Jin

IPC: G06F16/635 , G10L25/51 , G10H1/00 , G06F16/683 , G06F16/2457 , G06N3/08

Abstract: In implementations of searching for music, a music search system can receive a music search request that includes a music file including music content. The music search system can also receive a selected musical attribute from a plurality of musical attributes. The music search system includes a music search application that can generate musical features of the music content, where a respective one or more of the musical features correspond to a respective one of the musical attributes. The music search application can then compare the musical features that correspond to the selected musical attribute to audio features of audio files, and determine similar audio files to the music file based on the comparison of the musical features to the audio features of the audio files.

2.

发明申请
USING A PREDICTIVE MODEL TO AUTOMATICALLY ENHANCE AUDIO HAVING VARIOUS AUDIO QUALITY ISSUES 有权

公开(公告)号：US20210343305A1

公开(公告)日：2021-11-04

申请号：US16863591

申请日：2020-04-30

Applicant: Adobe Inc. , THE TRUSTEES OF PRINCETON UNIVERSITY

Inventor： Zeyu Jin , Jiaqi Su , Adam Finkelstein

IPC: G10L21/02 , G10L25/30 , G10L25/18 , G06N3/04 , G06N3/08

Abstract: Operations of a method include receiving a request to enhance a new source audio. Responsive to the request, the new source audio is input into a prediction model that was previously trained. Training the prediction model includes providing a generative adversarial network including the prediction model and a discriminator. Training data is obtained including tuples of source audios and target audios, each tuple including a source audio and a corresponding target audio. During training, the prediction model generates predicted audios based on the source audios. Training further includes applying a loss function to the predicted audios and the target audios, where the loss function incorporates a combination of a spectrogram loss and an adversarial loss. The prediction model is updated to optimize that loss function. After training, based on the new source audio, the prediction model generates a new predicted audio as an enhanced version of the new source audio.

3.

发明授权
Real-time speaker-dependent neural vocoder 有权

公开(公告)号：US10770063B2

公开(公告)日：2020-09-08

申请号：US16108996

申请日：2018-08-22

Applicant: Adobe Inc. , The Trustees of Princeton University

Inventor： Zeyu Jin , Gautham J. Mysore , Jingwan Lu , Adam Finkelstein

IPC: G10L15/16 , G06F17/14 , G10L15/22 , G06N3/08 , G06N3/04

Abstract: Techniques for a recursive deep-learning approach for performing speech synthesis using a repeatable structure that splits an input tensor into a left half and right half similar to the operation of the Fast Fourier Transform, performs a 1-D convolution on each respective half, performs a summation and then applies a post-processing function. The repeatable structure may be utilized in a series configuration to operate as a vocoder or perform other speech processing functions.

4.

发明授权
Neural pitch-shifting and time-stretching 有权

公开(公告)号：US11915714B2

公开(公告)日：2024-02-27

申请号：US17558580

申请日：2021-12-21

Applicant: Adobe Inc. , Northwestern University

Inventor： Maxwell Morrison , Juan Pablo Caceres Chomali , Zeyu Jin , Nicholas Bryan , Bryan A. Pardo

IPC: G10L21/013 , G10L15/02 , G10L15/18 , G10L25/90 , G10L25/30 , G10L19/032 , G10L21/04 , G10L25/24 , G10L15/06 , G10L19/028

CPC classification number: G10L21/013 , G10L15/02 , G10L15/063 , G10L15/1807 , G10L19/028 , G10L19/032 , G10L21/04 , G10L25/24 , G10L25/30 , G10L25/90 , G10L2021/0135

Abstract: Methods for modifying audio data include operations for accessing audio data having a first prosody, receiving a target prosody differing from the first prosody, and computing acoustic features representing samples. Computing respective acoustic features for a sample includes computing a pitch feature as a quantized pitch value of the sample by assigning a pitch value, of the target prosody or the audio data, to at least one of a set of pitch bins having equal widths in cents. Computing the respective acoustic features further includes computing a periodicity feature from the audio data. The respective acoustic features for the sample include the pitch feature, the periodicity feature, and other acoustic features. A neural vocoder is applied to the acoustic features to pitch-shift and time-stretch the audio data from the first prosody toward the target prosody.

5.

发明公开
Music Enhancement Systems 审中-公开

公开(公告)号：US20230343312A1

公开(公告)日：2023-10-26

申请号：US17726289

申请日：2022-04-21

Applicant: Adobe Inc.

Inventor： Nikhil Kandpal , Oriol Nieto-Caballero , Zeyu Jin

IPC: G10H1/00 , G10H1/06

CPC classification number: G10H1/0008 , G10H1/06 , G10H2210/066 , G10H2250/005

Abstract: In implementations of music enhancement systems, a computing device implements an enhancement system to receive input data describing a recorded acoustic waveform of a musical instrument. The recorded acoustic waveform is represented as an input mel spectrogram. The enhancement system generates an enhanced mel spectrogram by processing the input mel spectrogram using a first machine learning model trained on a first type of training data to generate enhanced mel spectrograms based on input mel spectrograms. An acoustic waveform of the musical instrument is generated by processing the enhanced mel spectrogram using a second machine learning model trained on a second type of training data to generate acoustic waveforms based on mel spectrograms. The acoustic waveform of the musical instrument does not include an acoustic artifact that is included in the recorded waveform of the musical instrument.

6.

发明授权
Secure audio watermarking based on neural networks 有权

公开(公告)号：US11170793B2

公开(公告)日：2021-11-09

申请号：US16790301

申请日：2020-02-13

Applicant: ADOBE INC.

Inventor： Zeyu Jin , Oona Shigeno Risse-Adams

IPC: G10L19/018 , G10L17/18 , G10L17/04 , G06N3/08 , G10L15/08 , G10L15/06

Abstract: Embodiments provide systems, methods, and computer storage media for secure audio watermarking and audio authenticity verification. An audio watermark detector may include a neural network trained to detect a particular audio watermark and embedding technique, which may indicate source software used in a workflow that generated an audio file under test. For example, the watermark may indicate an audio file was generated using voice manipulation software, so detecting the watermark can indicate manipulated audio such as deepfake audio and other attacked audio signals. In some embodiments, the audio watermark detector may be trained as part of a generative adversarial network in order to make the underlying audio watermark more robust to neural network-based attacks. Generally, the audio watermark detector may evaluate time domain samples from chunks of an audio clip under test to detect the presence of the audio watermark and generate a classification for the audio clip.

7.

发明授权
Text-based insertion and replacement in audio narration 有权

公开(公告)号：US10347238B2

公开(公告)日：2019-07-09

申请号：US15796292

申请日：2017-10-27

Applicant: Adobe Inc. , The Trustees of Princeton University

Inventor： Zeyu Jin , Gautham J. Mysore , Stephen DiVerdi , Jingwan Lu , Adam Finkelstein

IPC: G10L13/08 , G10L15/02 , G10L13/04 , G10L13/07

Abstract: Systems and techniques are disclosed for synthesizing a new word or short phrase such that it blends seamlessly in the context of insertion or replacement in an existing narration. In one such embodiment, a text-to-speech synthesizer is utilized to say the word or phrase in a generic voice. Voice conversion is then performed on the generic voice to convert it into a voice that matches the narration. An editor and interface are described that support fully automatic synthesis, selection among a candidate set of alternative pronunciations, fine control over edit placements and pitch profiles, and guidance by the editors own voice.

8.

发明授权
High fidelity audio super resolution 有权

公开(公告)号：US12217742B2

公开(公告)日：2025-02-04

申请号：US17534221

申请日：2021-11-23

Applicant: Adobe Inc. , The Trustees of Princeton University

Inventor： Zeyu Jin , Jiaqi Su , Adam Finkelstein

IPC: G10L15/16 , G06N3/045 , G10L15/06

Abstract: Embodiments are disclosed for generating full-band audio from narrowband audio using a GAN-based audio super resolution model. A method of generating full-band audio may include receiving narrow-band input audio data, upsampling the narrow-band input audio data to generate upsampled audio data, providing the upsampled audio data to an audio super resolution model, the audio super resolution model trained to perform bandwidth expansion from narrow-band to wide-band, and returning wide-band output audio data corresponding to the narrow-band input audio data.

9.

发明公开
NEURAL PITCH-SHIFTING AND TIME-STRETCHING 审中-公开

公开(公告)号：US20230197093A1

公开(公告)日：2023-06-22

申请号：US17558580

申请日：2021-12-21

Applicant: Adobe Inc. , Northwestern University

Inventor： Maxwell Morrison , Juan Pablo Caceres Chomali , Zeyu Jin , Nicholas Bryan , Bryan A. Pardo

IPC: G10L21/013 , G10L15/02 , G10L15/18 , G10L25/90 , G10L25/30 , G10L19/028 , G10L19/032 , G10L21/04 , G10L25/24 , G10L15/06

CPC classification number: G10L21/013 , G10L15/02 , G10L15/1807 , G10L25/90 , G10L25/30 , G10L19/028 , G10L19/032 , G10L21/04 , G10L25/24 , G10L15/063 , G10L2021/0135

Abstract: Methods for modifying audio data include operations for accessing audio data having a first prosody, receiving a target prosody differing from the first prosody, and computing acoustic features representing samples. Computing respective acoustic features for a sample includes computing a pitch feature as a quantized pitch value of the sample by assigning a pitch value, of the target prosody or the audio data, to at least one of a set of pitch bins having equal widths in cents. Computing the respective acoustic features further includes computing a periodicity feature from the audio data. The respective acoustic features for the sample include the pitch feature, the periodicity feature, and other acoustic features. A neural vocoder is applied to the acoustic features to pitch-shift and time-stretch the audio data from the first prosody toward the target prosody.

10.

发明公开
CONTEXT-AWARE PROSODY CORRECTION OF EDITED SPEECH 审中-公开

公开(公告)号：US20230169961A1

公开(公告)日：2023-06-01

申请号：US17538683

申请日：2021-11-30

Applicant: Adobe Inc.

Inventor： Maxwell Morrison , Zeyu Jin , Nicholas Bryan , Juan Pablo Caceres Chomali , Lucas Rencker

IPC: G10L15/18 , G10L25/90 , G10L15/187 , G10L15/02 , G10L15/04 , G10L21/0208 , G10L15/16 , G06N3/08

CPC classification number: G10L15/1807 , G10L25/90 , G10L15/187 , G10L15/02 , G10L15/04 , G10L21/0208 , G10L15/16 , G06N3/088 , G10L2015/025 , G10L2021/02082 , G06N3/0454

Abstract: Methods are performed by one or more processing devices for correcting prosody in audio data. A method includes operations for accessing subject audio data in an audio edit region of the audio data. The subject audio data in the audio edit region potentially lacks prosodic continuity with unedited audio data in an unedited audio portion of the audio data. The operations further include predicting, based on a context of the unedited audio data, phoneme durations including a respective phoneme duration of each phoneme in the unedited audio data. The operations further include predicting, based on the context of the unedited audio data, a pitch contour comprising at least one respective pitch value of each phoneme in the unedited audio data. Additionally, the operations include correcting prosody of the subject audio data in the audio edit region by applying the phoneme durations and the pitch contour to the subject audio data.

Search Results

Country/Region

Patent validity

Application date

Publication (announcement) day

applicant

The country/region where the applicant is located

Inventor

IPC

IPC Department

IPC class

IPC subclass

IPC group

IPC team

Appearance classification