Descript, Inc.

United States of America

Back to Profile

1-57 of 57 for Descript, Inc. Sort by
Query
Aggregations
IP Type
        Patent 34
        Trademark 23
Jurisdiction
        United States 47
        Canada 5
        World 4
        Europe 1
Date
2026 March 3
2026 (YTD) 8
2025 14
2024 8
2023 7
See more
IPC Class
G06F 3/16 - Sound inputSound output 16
G10L 15/26 - Speech to text systems 14
G06F 16/61 - IndexingData structures thereforStorage structures 9
G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks 8
G06F 17/00 - Digital computing or data processing equipment or methods, specially adapted for specific functions 6
See more
NICE Class
42 - Scientific, technological and industrial services, research and design 21
09 - Scientific and electric apparatus and instruments 20
35 - Advertising and business services 4
41 - Education, entertainment, sporting and cultural services 4
Status
Pending 24
Registered / In Force 33

1.

DESCRIPT

      
Application Number 1908388
Status Registered
Filing Date 2026-02-03
Registration Date 2026-02-03
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable and recorded computer software platforms using artificial intelligence for making podcasts and other multimedia content; downloadable and recorded computer software platforms using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable and recorded audio word processing computer software platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; downloadable digital image files of avatars; downloadable computer software using artificial intelligence for creating avatars. Providing on-line non-downloadable computer software using artificial intelligence for making podcasts and other multimedia content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence for making podcasts and other multimedia content; providing on-line non-downloadable computer software using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence (AI) for recording, transcribing, editing, and mixing audio, video, text, and other media content; technical support services relating to recording, transcribing, editing, and mixing audio, video, text, and other media content, namely, troubleshooting in the nature of diagnosing computer software problems using artificial intelligence; providing on-line non-downloadable computer software using artificial intelligence for enabling editing of sound files and lyrics in text form; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; providing on-line non-downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; providing on-line non-downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; providing on-line non-downloadable computer software using artificial intelligence (AI) for creating avatars; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for creating avatars.

2.

D

      
Application Number 1907524
Status Registered
Filing Date 2026-02-03
Registration Date 2026-02-03
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable and recorded computer software platforms using artificial intelligence for making podcasts and other multimedia content; downloadable and recorded computer software platforms using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable and recorded audio word processing computer software platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; downloadable digital image files of avatars; downloadable computer software using artificial intelligence for creating avatars. Providing on-line non-downloadable computer software using artificial intelligence for making podcasts and other multimedia content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence for making podcasts and other multimedia content; providing on-line non-downloadable computer software using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence (AI) for recording, transcribing, editing, and mixing audio, video, text, and other media content; technical support services relating to recording, transcribing, editing, and mixing audio, video, text, and other media content, namely, troubleshooting in the nature of diagnosing computer software problems using artificial intelligence; providing on-line non-downloadable computer software using artificial intelligence for enabling editing of sound files and lyrics in text form; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; providing on-line non-downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; providing on-line non-downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; providing on-line non-downloadable computer software using artificial intelligence (AI) for creating avatars; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for creating avatars.

3.

APPROACHES TO MULTIMEDIA EDITING USING AN ARTIFICIAL INTELLIGENCE MODEL AND SYSTEMS FOR ACCOMPLISHING THE SAME

      
Application Number 19316409
Status Pending
Filing Date 2025-09-02
First Publication Date 2026-03-12
Owner Descript, Inc. (USA)
Inventor
  • Dodero, David
  • Lui, Katrina
  • Yuan, Raymond
  • Arasanipalai, Ajay
  • Lam, Cora
  • Ramabhadran, Pranav

Abstract

The disclosed technology uses a media production platform to edit multimedia files with an AI model (e.g., a neural network). The technology can remove retakes, identify highlight clips, and/or generate layouts for multimedia files. The technology can process audio transcripts to exclude retakes by generating a refined transcript and highlighting removed segments. Additionally, the technology can edit audiovisual files by generating scenes based on content and mapping the scenes to relevant layouts, dynamically adjusting based on user input. The technology can generate highlights by applying AI models to create clips and identify topics within the audiovisual file, producing an edited file indicative of the topics. The results, such as the edited files, are presented on the client device.

IPC Classes  ?

  • H04N 21/81 - Monomedia components thereof
  • H04N 21/431 - Generation of visual interfacesContent or additional data rendering
  • H04N 21/466 - Learning process for intelligent management, e.g. learning user preferences for recommending movies

4.

DESCRIPT

      
Application Number 246349600
Status Pending
Filing Date 2026-02-03
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Downloadable and recorded computer software platforms using artificial intelligence for making podcasts and other multimedia content; downloadable and recorded computer software platforms using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable and recorded audio word processing computer software platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; downloadable digital image files of avatars; downloadable computer software using artificial intelligence for creating avatars. (1) Providing on-line non-downloadable computer software using artificial intelligence for making podcasts and other multimedia content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence for making podcasts and other multimedia content; providing on-line non-downloadable computer software using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence (AI) for recording, transcribing, editing, and mixing audio, video, text, and other media content; technical support services relating to recording, transcribing, editing, and mixing audio, video, text, and other media content, namely, troubleshooting in the nature of diagnosing computer software problems using artificial intelligence; providing on-line non-downloadable computer software using artificial intelligence for enabling editing of sound files and lyrics in text form; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; providing on-line non-downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; providing on-line non-downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; providing on-line non-downloadable computer software using artificial intelligence (AI) for creating avatars; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for creating avatars.

5.

D

      
Application Number 246200100
Status Pending
Filing Date 2026-02-03
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Downloadable and recorded computer software platforms using artificial intelligence for making podcasts and other multimedia content; downloadable and recorded computer software platforms using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable and recorded audio word processing computer software platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; downloadable digital image files of avatars; downloadable computer software using artificial intelligence for creating avatars. (1) Providing on-line non-downloadable computer software using artificial intelligence for making podcasts and other multimedia content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence for making podcasts and other multimedia content; providing on-line non-downloadable computer software using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; platform as a service (PaaS) featuring computer software platforms using artificial intelligence (AI) for recording, transcribing, editing, and mixing audio, video, text, and other media content; technical support services relating to recording, transcribing, editing, and mixing audio, video, text, and other media content, namely, troubleshooting in the nature of diagnosing computer software problems using artificial intelligence; providing on-line non-downloadable computer software using artificial intelligence for enabling editing of sound files and lyrics in text form; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; providing on-line non-downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; providing on-line non-downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; providing on-line non-downloadable computer software using artificial intelligence (AI) for creating avatars; platform as a service (PaaS) featuring computer software and audio and word processing platforms using artificial intelligence for creating avatars.

6.

APPROACHES TO TRAINING AND IMPLEMENTING A UNIVERSAL VARIABLE MODEL FOR DYNAMIC VOICE SYNTHESIS AND SYSTEMS FOR ACCOMPLISHING THE SAME

      
Application Number 19266486
Status Pending
Filing Date 2025-07-11
First Publication Date 2026-01-15
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Kundan
  • Kumar, Rithesh
  • Kumar, Ishaan
  • Luebs, Alejandro

Abstract

Introduced here are approaches to training and then employing computer-implemented models designed to generate synthesized speech using a Universal Variable Model (UVM). The UVM is pre-trained using reference audio samples and associated text prompts to comprehend and replicate various aspects of human speech, including intonation, rhythm, and pronunciation. In the training process, the UVM learns general patterns and relationships between the acoustic properties of speech and the linguistic features of text from a dataset covering different linguistic contexts, accents, and speakers. This enables the UVM to generate natural-sounding speech without the need for personalized training on the user's voice. Users of the media production platform can submit text inputs along with a reference audio sample, and the UVM will produce corresponding audio output in the same voice as the reference sample.

IPC Classes  ?

  • G10L 13/02 - Methods for producing synthetic speechSpeech synthesisers
  • G10L 13/10 - Prosody rules derived from textStress or intonation
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

7.

APPROACHES TO EDITING AUDIO CONTENT USING DYNAMIC VOICE SYNTHESIS AND SYSTEMS FOR ACCOMPLISHING THE SAME

      
Application Number 19266381
Status Pending
Filing Date 2025-07-11
First Publication Date 2026-01-15
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Kundan
  • Kumar, Rithesh
  • Kumar, Ishaan
  • Luebs, Alejandro

Abstract

Introduced here are approaches to editing audio content using dynamic voice synthesis and systems for accomplishing the same. The system uses a transcript associated with an audio file and received input that indicates a location to add or remove text to identify preceding and succeeding segments around the indicated location. The system constructs a modified transcript, and applies a model (e.g., a Universal Variable Model (UVM)) to generate new audio content in accordance with the modified transcript. The model aligns the acoustic properties of the original audio file with linguistic features of the transcript, therefore enabling the new audio content to emulate the original audio file's properties. The system produces a final audio file by inserting the new audio content into the original audio file. This approach allows for dynamic editing of audio content, maintaining coherence and acoustic consistency while accommodating textual modifications. Prior to generating the final audio file, the system can perform one or more authentication operations using a dynamically generated consent statement.

IPC Classes  ?

  • G11B 27/031 - Electronic editing of digitised analogue information signals, e.g. audio or video signals
  • G06F 40/166 - Editing, e.g. inserting or deleting

8.

TRAINING MACHINE LEARNING FRAMEWORKS TO GENERATE STUDIO-QUALITY RECORDINGS THROUGH MANIPULATION OF NOISY AUDIO SIGNALS

      
Application Number 19323569
Status Pending
Filing Date 2025-09-09
First Publication Date 2026-01-08
Owner Descript, Inc. (USA)
Inventor
  • Seetharaman, Prem S.
  • Kumar, Kundan

Abstract

Introduced here are computer programs and associated computer-implemented techniques for manipulating noisy audio signals to produce clean audio signals that are sufficiently high quality so as to be largely, if not entirely, indistinguishable from “rich” recordings generated by recording studios. When a noisy audio signal is obtained by a media production platform, the noisy audio signal can be manipulated to sound as if recording occurred with sophisticated equipment in a soundproof environment. Manipulation can be performed by a model that, when applied to the noisy audio signal, can manipulate its characteristics so as to emulate the characteristics of clean audio signals that are learned through training.

IPC Classes  ?

  • G10L 21/0232 - Processing in the frequency domain
  • G06F 3/16 - Sound inputSound output
  • G10L 15/06 - Creation of reference templatesTraining of speech recognition systems, e.g. adaptation to the characteristics of the speaker's voice
  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G10L 25/21 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being power information
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

9.

LYREBIRD

      
Serial Number 99483679
Status Pending
Filing Date 2025-11-06
Owner Descript, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Research and development in the field of artificial intelligence, and machine learning; Research and development in the field of media, namely, development of software for recording, transcribing, editing, and mixing audio, video, text, and other media content

10.

REGENERATE

      
Serial Number 99483682
Status Registered
Filing Date 2025-11-06
Registration Date 2026-06-09
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable software for creating, recording, editing, mixing, processing, enhancing, and repairing video, audio, text, and other media content; Downloadable computer software using artificial intelligence (AI) for for creating, recording, editing, mixing, processing, enhancing, and repairing video, audio, text, and other media content; Downloadable software for editing audio and dialogue to improve tone; Downloadable computer software using artificial intelligence (AI) for editing audio and dialogue to improve tone; Downloadable software for reducing background noise in video, audio, and other media content; Downloadable computer software using artificial intelligence (AI) for reducing background noise in video, audio, and other media content; Downloadable software for facilitating production of multimedia compilations; Downloadable computer software using artificial intelligence (AI) for facilitating production of multimedia compilations Providing on-line non-downloadable software for creating, recording, editing, mixing, processing, enhancing, and transcribing video and audio; Providing on-line non-downloadable software using artificial intelligence (AI) for creating, recording, editing, mixing, processing, enhancing, and transcribing video and audio; Providing on-line non-downloadable software for for facilitating production of multimedia compilations; Providing on-line non-downloadable software using artificial intelligence (AI) for facilitating production of multimedia compilations; Providing on-line non-downloadable software for reducing background noise in videos and audio; Providing on-line non-downloadable software using artificial intelligence (AI) for reducing background noise in videos and audio

11.

STUDIO SOUND

      
Serial Number 99483680
Status Pending
Filing Date 2025-11-06
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable software for creating, recording, editing, mixing, processing, enhancing, and transcribing video and audio; Downloadable computer software using artificial intelligence (AI) for creating, recording, editing, mixing, processing, enhancing, and transcribing video and audio; Downloadable software for reducing background noise in videos and audio; Downloadable computer software using artificial intelligence (AI) for reducing background noise in videos and audio; Downloadable software for facilitating production of multimedia compilations; Downloadable computer software using artificial intelligence (AI) for facilitating production of multimedia compilations Providing on-line non-downloadable software for creating, recording, editing, mixing, processing, enhancing, and transcribing video and audio; Providing on-line non-downloadable software using artificial intelligence (AI) for creating, recording, editing, mixing, processing, enhancing, and transcribing video and audio; Providing on-line non-downloadable software for reducing background noise in video, audio, and other media content; Providing on-line non-downloadable software using artificial intelligence (AI) for reducing background noise in video, audio, and other media content; Providing on-line non-downloadable software for for facilitating production of multimedia compilations; Providing on-line non-downloadable software using artificial intelligence (AI) for facilitating production of multimedia compilations

12.

EYE CONTACT

      
Serial Number 99483699
Status Pending
Filing Date 2025-11-06
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable software for facilitating production of multimedia compilations; Downloadable computer software using artificial intelligence (AI) for facilitating production of multimedia compilations; Downloadable software for creating, recording, editing, mixing, and processing video; Downloadable computer software using artificial intelligence (AI) for creating, recording, editing, mixing, and processing video Providing on-line non-downloadable software for creating, recording, editing, mixing, and processing video; Providing on-line non-downloadable software for facilitating production of multimedia compilations; Providing on-line non-downloadable software using artificial intelligence (AI) for facilitating production of multimedia compilations; Providing on-line non-downloadable software using artificial intelligence (AI) for creating, recording, editing, mixing, and processing video

13.

DESCRIPT

      
Serial Number 99401028
Status Registered
Filing Date 2025-09-18
Registration Date 2026-05-05
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable and recorded computer software platforms using artificial intelligence for making podcasts and other multimedia content; downloadable and recorded computer software platforms using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable and recorded audio word processing computer software platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; Downloadable digital image files of avatars; downloadable computer software using artificial intelligence for creating avatars Providing on-line non-downloadable computer software using artificial intelligence for making podcasts and other multimedia content; platform as a service (PAAS) featuring computer software platforms using artificial intelligence for making podcasts and other multimedia content; providing on-line non-downloadable computer software using artificial intelligence for recording, transcribing, editing, and mixing audio, video, text, and other media content; platform as a service (PAAS) featuring computer software platforms using artificial intelligence (AI) for recording, transcribing, editing, and mixing audio, video, text, and other media content; technical support services relating to recording, transcribing, editing, and mixing audio, video, text, and other media content, namely, troubleshooting in the nature of diagnosing computer software problems using artificial intelligence; providing on-line non-downloadable computer software using artificial intelligence for enabling editing of sound files and lyrics in text form; platform as a service (PAAS) featuring computer software and audio and word processing platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; providing on-line non-downloadable computer software using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; platform as a service (PAAS) featuring computer software and audio and word processing platforms using artificial intelligence for use in prompting questions, offering observations, and providing challenges to users in the field of audio, video, text, and other media content; providing on-line non-downloadable computer software using artificial intelligence for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; providing on-line non-downloadable computer software using artificial intelligence (AI) for creating avatars; platform as a service (PAAS) featuring computer software and audio and word processing platforms using artificial intelligence for creating avatars

14.

D

      
Serial Number 99401021
Status Pending
Filing Date 2025-09-18
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable and recorded computer software platforms using artificial intelligence for making podcasts and other audio and video content; downloadable and recorded computer software platforms using artificial intelligence for recording, transcribing, and editing audio, video, and text and for mixing audio and video; downloadable and recorded audio word processing computer software platforms using artificial intelligence for editing of sound files and lyrics in text form; downloadable computer software using artificial intelligence for use in providing analysis and recommendations for improving and editing audio, video, and text; downloadable computer software using artificial intelligence for recording audio and using that audio to create text, audio, images, video, and combinations thereof; downloadable digital image files of avatars; downloadable computer software using artificial intelligence for creating avatars Providing on-line non-downloadable computer software using artificial intelligence for making podcasts and other audio and video content; platform as a service (PAAS) featuring computer software platforms using artificial intelligence for making podcasts and other audio and video content; providing on-line non-downloadable computer software using artificial intelligence for recording, transcribing, and editing audio, video, and text and for mixing audio and video; platform as a service (PAAS) featuring computer software platforms using artificial intelligence (AI) for recording, transcribing, and editing audio, video, and text and for mixing audio and video; technical support services relating to recording, transcribing, editing, and mixing audio, video, text, and other media content, namely, troubleshooting in the nature of diagnosing computer software problems using artificial intelligence; providing on-line non-downloadable computer software using artificial intelligence for editing of sound files and lyrics in text form; platform as a service (PAAS) featuring computer software and audio and word processing platforms using artificial intelligence for enabling editing of sound files and lyrics in text form; providing on-line non-downloadable computer software using artificial intelligence for use in providing analysis and recommendations for improving and editing audio, video, and text; platform as a service (PAAS) featuring computer software and audio and word processing platforms using artificial intelligence for use in providing analysis and recommendations for improving and editing audio, video, and text; providing online non-downloadable computer software using artificial intelligence for recording audio and using that audio to create text, audio, images, video, and combinations thereof; providing on-line non-downloadable computer software using artificial intelligence (AI) for creating avatars; platform as a service (PAAS) featuring computer software and audio and word processing platforms using artificial intelligence for creating avatars

15.

SIMULTANEOUS RECORDING AND UPLOADING OF MULTIPLE AUDIO FILES OF THE SAME CONVERSATION AND AUDIO DRIFT NORMALIZATION SYSTEMS AND METHODS

      
Application Number 19203010
Status Pending
Filing Date 2025-05-08
First Publication Date 2025-08-21
Owner Descript, Inc. (USA)
Inventor Moreno, Zachariah Steven

Abstract

As advances in Internet communications including video and audio mediums continue to increase, the need to effectively and efficiently communicate media has also increased. Many of the platforms that are designed for the creation and communication of media content cannot ensure that the media content is accessible in the event of a problem, for example, with the recording process. Moreover, these platforms generally do not eliminate the need to record a complete session in its entirety prior to uploading or using that content. Introduced here is an approach to content recording that involves two computer programs. A first computer program executing on a first device can segment media into chunks as the media is being recorded and then transmit those chunks, one by one, to a second device. A second computer program executing on the second device can then reconstruct the chunks in numbered sequence.

IPC Classes  ?

  • H04N 21/854 - Content authoring
  • G06F 16/11 - File system administration, e.g. details of archiving or snapshots
  • G06F 16/16 - File or folder operations, e.g. details of user interfaces specifically adapted to file systems
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G11B 27/034 - Electronic editing of digitised analogue information signals, e.g. audio or video signals on discs
  • H04L 9/40 - Network security protocols
  • H04L 65/70 - Media network packetisation
  • H04L 67/02 - Protocols based on web technology, e.g. hypertext transfer protocol [HTTP]
  • H04L 67/06 - Protocols specially adapted for file transfer, e.g. file transfer protocol [FTP]
  • H04L 67/10 - Protocols in which an application is distributed across nodes in the network
  • H04N 21/845 - Structuring of content, e.g. decomposing content into time segments

16.

TECHNOLOGIES FOR CREATING, ALTERING, AND PRESENTING MEDIA CONTENT

      
Application Number 19047889
Status Pending
Filing Date 2025-02-07
First Publication Date 2025-06-05
Owner Descript, Inc. (USA)
Inventor
  • Holmes, Ryan Terrill
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Different types of media experiences can be developed based on characteristics of the consumer. “Linear” experiences may require execution of a pre-built script, although the script could be dynamically modified by a media production platform. Linear experiences can include guided audio tours that are modified or updated based on the location of the consumer. “Enhanced” experiences include conventional media content that is supplemented with intelligent media content. For example, turn-by-turn directions could be supplemented with audio descriptions about the surrounding area. “Freeform” experiences, meanwhile, are those that can continually morph based on information gleaned from a consumer. For example, a radio station may modify what content is being presented based on the geographical metadata uploaded by a computing device associated with the consumer.

IPC Classes  ?

  • G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
  • G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
  • G06F 16/68 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
  • G06F 16/683 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content
  • G10L 15/187 - Phonemic context, e.g. pronunciation rules, phonotactical constraints or phoneme n-grams
  • G10L 15/26 - Speech to text systems

17.

UPSAMPLING OF AUDIO USING GENERATIVE ADVERSARIAL NETWORKS

      
Application Number 18934075
Status Pending
Filing Date 2024-10-31
First Publication Date 2025-02-20
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Rithesh
  • Kumar, Kundan

Abstract

Introduced here are approaches to training and then employing computer-implemented models designed to upsample discrete audio signals to higher sampling rates. Assume, for example, that a media production platform obtains a first discrete signal at a relatively low sampling rate. The relatively low sampling frequency may make the first discrete audio signal unsuitable for inclusion in media compilations, so the media production platform may attempt to improve its quality through upsampling. To accomplish this, the media production platform can apply a transform to the first discrete signal to produce a first magnitude spectrogram. Then, the media production platform can apply a computer-implemented model to the first magnitude spectrogram to produce a second magnitude spectrogram. Thereafter, the media production platform can apply an inverse transform to the second magnitude spectrogram to create a second discrete signal that has a higher sampling rate than the first discrete audio signal.

IPC Classes  ?

  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G06F 3/16 - Sound inputSound output
  • G06N 3/045 - Combinations of networks
  • G06N 3/088 - Non-supervised learning, e.g. competitive learning
  • G10L 19/02 - Speech or audio signal analysis-synthesis techniques for redundancy reduction, e.g. in vocodersCoding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

18.

TRAINING GENERATIVE ADVERSARIAL NETWORKS TO UPSAMPLE AUDIO

      
Application Number 18932072
Status Pending
Filing Date 2024-10-30
First Publication Date 2025-02-13
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Rithesh
  • Kumar, Kundan

Abstract

Introduced here are approaches to training and then employing computer-implemented models designed to upsample discrete audio signals to higher sampling rates. Assume, for example, that a media production platform obtains a first discrete signal at a relatively low sampling rate. The relatively low sampling frequency may make the first discrete audio signal unsuitable for inclusion in media compilations, so the media production platform may attempt to improve its quality through upsampling. To accomplish this, the media production platform can apply a transform to the first discrete signal to produce a first magnitude spectrogram. Then, the media production platform can apply a computer-implemented model to the first magnitude spectrogram to produce a second magnitude spectrogram. Thereafter, the media production platform can apply an inverse transform to the second magnitude spectrogram to create a second discrete signal that has a higher sampling rate than the first discrete audio signal.

IPC Classes  ?

  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G06F 3/16 - Sound inputSound output
  • G06N 3/045 - Combinations of networks
  • G06N 3/088 - Non-supervised learning, e.g. competitive learning
  • G10L 19/02 - Speech or audio signal analysis-synthesis techniques for redundancy reduction, e.g. in vocodersCoding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

19.

FILLER WORD DETECTION THROUGH TOKENIZING AND LABELING OF TRANSCRIPTS

      
Application Number 18934035
Status Pending
Filing Date 2024-10-31
First Publication Date 2025-02-13
Owner Descript, Inc. (USA)
Inventor
  • De Brébisson, Alexandre
  • D’andigné, Antoine

Abstract

Introduced here are computer programs and associated computer-implemented techniques for discovering the presence of filler words through tokenization of a transcript derived from audio content. When audio content is obtained by a media production platform, the audio content can be converted into text content as part of a speech-to-text operation. The text content can then be tokenized and labeled using a Natural Language Processing (NLP) library. Tokenizing/labeling may be performed in accordance with a series of rules associated with filler words. At a high level, these rules may examine the text content (and associated tokens/labels) to determine whether patterns, relationships, verbatim, and context indicate that a term is a filler word. Any filler words that are discovered in the text content can be identified as such so that appropriate action(s) can be taken.

IPC Classes  ?

20.

TRANSCRIPTION CORRECTION THROUGH PROGRAMMATIC COMPARISON OF INDEPENDENTLY GENERATED TRANSCRIPTS

      
Application Number 18903746
Status Pending
Filing Date 2024-10-01
First Publication Date 2025-01-16
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Kundan
  • Anand, Vicki

Abstract

Introduced here are computer programs and associated computer-implemented techniques for facilitating the creation of a master transcription (or simply “transcript”) that more accurately reflects underlying audio by comparing multiple independently generated transcripts. The master transcript may be used to record and/or produce various forms of media content, as further discussed below. Thus, the technology described herein may be used to facilitate editing of text content, audio content, or video content. These computer programs may be supported by a media production platform that is able to generate the interfaces through which individuals (also referred to as “users”) can create, edit, or view media content. For example, a computer program may be embodied as a word processor that allows individuals to edit voice-based audio content by editing a master transcript, and vice versa.

IPC Classes  ?

  • G10L 15/26 - Speech to text systems
  • G06F 3/16 - Sound inputSound output
  • G10L 15/01 - Assessment or evaluation of speech recognition systems
  • G10L 15/08 - Speech classification or search
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G10L 15/30 - Distributed recognition, e.g. in client-server systems, for mobile phones or network applications
  • G10L 15/32 - Multiple recognisers used in sequence or in parallelScore combination systems therefor, e.g. voting systems

21.

DYNAMIC MODIFICATION OF CONTENT-BASED EXPERIENCES BASED ON REAL-TIME ANALYSIS OF LOCATION

      
Application Number 18885381
Status Pending
Filing Date 2024-09-13
First Publication Date 2025-01-02
Owner DESCRIPT, INC. (USA)
Inventor
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Media content can be created and/or modified using a network-accessible platform. Scripts for content-based experiences could be readily created using one or more interfaces generated by the network-accessible platform. For example, a script for a content-based experience could be created using an interface that permits triggers to be inserted directly into the script. Interface(s) may also allow different media formats to be easily aligned for post-processing. For example, a transcript and an audio file may be dynamically aligned so that the network-accessible platform can globally reflect changes made to either item. User feedback may also be presented directly on the interface(s) so that modifications can be made based on actual user experiences.

IPC Classes  ?

  • G06F 3/16 - Sound inputSound output
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G06Q 10/101 - Collaborative creation, e.g. joint development of products or services
  • G10L 19/008 - Multichannel audio signal coding or decoding using interchannel correlation to reduce redundancy, e.g. joint-stereo, intensity-coding or matrixing
  • G11B 27/031 - Electronic editing of digitised analogue information signals, e.g. audio or video signals
  • G11B 27/10 - IndexingAddressingTiming or synchronisingMeasuring tape travel
  • G11B 27/34 - Indicating arrangements

22.

APPROACHES TO SYNTHESIZING AND BLENDING AUDIO IN RESPONSE TO MODIFICATION OF A TRANSCRIPT

      
Application Number 18886937
Status Pending
Filing Date 2024-09-16
First Publication Date 2025-01-02
Owner DESCRIPT, INC. (USA)
Inventor
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Media content can be created and/or modified using a network-accessible platform. Scripts for content-based experiences could be readily created using one or more interfaces generated by the network-accessible platform. For example, a script for a content-based experience could be created using an interface that permits triggers to be inserted directly into the script. Interface(s) may also allow different media formats to be easily aligned for post-processing. For example, a transcript and an audio file may be dynamically aligned so that the network-accessible platform can globally reflect changes made to either item. User feedback may also be presented directly on the interface(s) so that modifications can be made based on actual user experiences.

IPC Classes  ?

  • G06F 3/16 - Sound inputSound output
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G06Q 10/101 - Collaborative creation, e.g. joint development of products or services
  • G10L 19/008 - Multichannel audio signal coding or decoding using interchannel correlation to reduce redundancy, e.g. joint-stereo, intensity-coding or matrixing
  • G11B 27/031 - Electronic editing of digitised analogue information signals, e.g. audio or video signals
  • G11B 27/10 - IndexingAddressingTiming or synchronisingMeasuring tape travel
  • G11B 27/34 - Indicating arrangements

23.

UNDERLORD

      
Application Number 1829060
Status Registered
Filing Date 2024-07-29
Registration Date 2024-07-29
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content through a software platform for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI assisting in content creation; downloadable computer software that utilizes artificial intelligence to engage users of a software platform through prompting questions, offering observations, and providing challenges, the software platform allowing the users to record, transcribe, edit, and mix audio, video, text, and other media content; downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively; downloadable files of avatars for use in virtual environments and for use in software platforms for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable computer software for making audio, video and text content, in a semi- or fully-autonomous manner; downloadable computer software for facilitating the recording, transcribing, editing, and mixing of audio, video, text, and other media content; downloadable computer software for transforming ideas specified, either audibly or textually, by an individual into usable outputs in an automated manner, while also allowing the individual to edit those outputs for the purpose of producing content; downloadable word processor computer programs for enabling individuals to edit audio through the manipulation of corresponding text, and vice versa; downloadable computer programs that enable individuals to edit audio and lyrics in text form; downloadable computer software for artificial intelligence; downloadable computer software for avatars; downloadable computer software through which an individual is able to audibly or textually record thoughts and receive text, audio, images, video, or combinations thereof that are produced as output; downloadable computer software through which inputs are identified, provided, or generated in audible, visual, or textual form and those inputs are used to guide identification or generation of outputs in audible, visual, or textual form. Providing non-downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content through a software platform for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI assisting in content creation; providing non-downloadable computer software that utilizes artificial intelligence to engage users of a software platform through prompting questions, offering observations, and providing challenges, the software platform allowing the users to record, transcribe, edit, and mix audio, video, text, and other media content; providing non-downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively; providing non-downloadable computer software for creating avatars for use in virtual environments and for use in software platforms for recording, transcribing, editing, and mixing audio, video, text, and other media content; providing non-downloadable computer software for making audio, video and text content, in a semi- or fully-autonomous manner; providing non-downloadable computer software for facilitating the recording, transcribing, editing, and mixing of audio, video, text, and other media content; providing non-downloadable computer software for transforming ideas specified, either audibly or textually, by an individual into usable outputs in an automated manner, while also allowing the individual to edit those outputs for the purpose of producing content; providing non-downloadable word processor computer programs for enabling individuals to edit audio through the manipulation of corresponding text, and vice versa; providing non-downloadable computer programs that enable individuals to edit audio and lyrics in text form; providing non-downloadable computer software for artificial intelligence; providing non-downloadable computer software for avatars; providing non-downloadable computer software through which an individual is able to audibly or textually record thoughts and receive text, audio, images, video, or combinations thereof that are produced as output; providing non-downloadable computer software through which inputs are identified, provided, or generated in audible, visual, or textual form and those inputs are used to guide identification or generation of outputs in audible, visual, or textual form.

24.

AUTOMATED GENERATION OF TRANSCRIPTS THROUGH INDEPENDENT TRANSCRIPTION

      
Application Number 18793589
Status Pending
Filing Date 2024-08-02
First Publication Date 2024-11-28
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Kundan
  • Anand, Vicki

Abstract

Introduced here are computer programs and associated computer-implemented techniques for facilitating the creation of a master transcription (or simply “transcript”) that more accurately reflects underlying audio by comparing multiple independently generated transcripts. The master transcript may be used to record and/or produce various forms of media content, as further discussed below. Thus, the technology described herein may be used to facilitate editing of text content, audio content, or video content. These computer programs may be supported by a media production platform that is able to generate the interfaces through which individuals (also referred to as “users”) can create, edit, or view media content. For example, a computer program may be embodied as a word processor that allows individuals to edit voice-based audio content by editing a master transcript, and vice versa.

IPC Classes  ?

  • G10L 15/26 - Speech to text systems
  • G06F 3/16 - Sound inputSound output
  • G10L 15/01 - Assessment or evaluation of speech recognition systems
  • G10L 15/08 - Speech classification or search
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G10L 15/30 - Distributed recognition, e.g. in client-server systems, for mobile phones or network applications
  • G10L 15/32 - Multiple recognisers used in sequence or in parallelScore combination systems therefor, e.g. voting systems

25.

BRAIN BUDDY

      
Application Number 1804233
Status Registered
Filing Date 2024-05-02
Registration Date 2024-05-02
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable computer software for providing artificial intelligence (AI) integrated into software platforms for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI assisting in content creation; downloadable computer software for providing artificial intelligence (AI) integrated into software platforms for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI engaging users through prompting questions, offering observations, and providing constructive challenges; downloadable software providing artificial intelligence which simulates stimulus provided by a great conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure their ideas effectively; downloadable image files of avatars for use in virtual environments and for use in platforms for recording, transcribing, editing, and mixing audio, video, text and other media content; downloadable computer software platforms for making audio, video and text content; downloadable computer software platforms for recording, transcribing, editing, and mixing audio, video, text and other media content; downloadable audio word processing computer software platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form; downloadable computer software for artificial intelligence; downloadable computer software for avatars. Application service provider for providing artificial intelligence (AI) integrated into platform as a service for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI assisting in content creation; application service provider for providing artificial intelligence (AI) integrated into platform as a service for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI engaging users through prompting questions, offering observations, and providing constructive challenges; application service provider providing artificial intelligence (AI) integrated into platform as a service providing artificial intelligence which simulates stimulus provided by a great conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure their ideas effectively; software as a service providing avatars for use in virtual environments and for use in platforms for recording, transcribing, editing, and mixing audio, video, text and other media content; software as a service for providing artificial intelligence; software as a service for providing avatars.

26.

UNDERLORD

      
Application Number 236943600
Status Pending
Filing Date 2024-07-29
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content through a software platform for recording, transcribing, editing, and mixing audio, video, text, images and graphics, the AI assisting in content creation; downloadable computer software that utilizes artificial intelligence to engage users of a software platform through prompting questions, offering observations, and providing challenges, the software platform allowing the users to record, transcribe, edit, and mix audio, video, text, images and graphics; downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively; downloadable files of avatars for use in virtual environments and for use in software platforms for recording, transcribing, editing, and mixing audio, video, text, images and graphics; downloadable computer software for making audio, video and text content, in a semi- or fully-autonomous manner; downloadable computer software for facilitating the recording, transcribing, editing, and mixing of audio, video, text, images and graphics; downloadable computer software for transforming ideas specified, either audibly or textually, by an individual into usable outputs in an automated manner, while also allowing the individual to edit those outputs for the purpose of producing content; downloadable word processor computer programs for enabling individuals to edit audio through the manipulation of corresponding text, and vice versa; downloadable computer programs that enable individuals to edit audio and lyrics in text form; downloadable computer software that utilizes artificial intelligence for providing feedback in the field of writing; downloadable computer software for creating avatars in virtual worlds; downloadable computer software through which an individual is able to audibly or textually record thoughts and receive text, audio, images, video, or combinations thereof that are produced as output; downloadable computer software through which inputs are identified, provided, or generated in audible, visual, or textual form and those inputs are used to guide identification or generation of outputs in audible, visual, or textual form. (1) Providing non-downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content through a software platform for recording, transcribing, editing, and mixing audio, video, text, images and graphics, the AI assisting in content creation; providing non-downloadable computer software that utilizes artificial intelligence to engage users of a software platform through prompting questions, offering observations, and providing challenges, the software platform allowing the users to record, transcribe, edit, and mix audio, video, text, images and graphics; providing non-downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively; providing non-downloadable computer software for creating avatars for use in virtual environments and for use in software platforms for recording, transcribing, editing, and mixing audio, video, text, images and graphics; providing non-downloadable computer software for making audio, video and text content, in a semi- or fully-autonomous manner; providing non-downloadable computer software for facilitating the recording, transcribing, editing, and mixing of audio, video, text, images and graphics; providing non-downloadable computer software for transforming ideas specified, either audibly or textually, by an individual into usable outputs in an automated manner, while also allowing the individual to edit those outputs for the purpose of producing content; providing non-downloadable word processor computer programs for enabling individuals to edit audio through the manipulation of corresponding text, and vice versa; providing non-downloadable computer programs that enable individuals to edit audio and lyrics in text form; providing non-downloadable computer software that utilizes artificial intelligence for providing feedback in the field of writing; providing non-downloadable computer software for creating avatars in virtual worlds; providing non-downloadable computer software through which an individual is able to audibly or textually record thoughts and receive text, audio, images, video, or combinations thereof that are produced as output; providing non-downloadable computer software through which inputs are identified, provided, or generated in audible, visual, or textual form and those inputs are used to guide identification or generation of outputs in audible, visual, or textual form.

27.

UNDERLORD

      
Serial Number 98586641
Status Registered
Filing Date 2024-06-05
Registration Date 2026-05-12
Owner Descript, Inc. (USA)
NICE Classes  ? 09 - Scientific and electric apparatus and instruments

Goods & Services

Downloadable computer software that utilizes artificial intelligence (AI) for use in creating content, namely, for recording, transcribing, editing, and mixing audio, video, text and other media content; downloadable computer software that utilizes artificial intelligence (AI) for use in prompting questions, offering observations, and providing challenges to users in the field of recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable computer software that utilizes artificial intelligence (AI) for use in providing feedback in the field of writing; downloadable image files of avatars for use in virtual worlds; downloadable computer software using artificial intelligence (AI) for use in creating and editing media content and producing media compilations containing audio, video and text content; downloadable computer software for recording, transcribing, editing, and mixing audio, video, text, and other media content; downloadable computer software using artificial intelligence (AI) for use in transforming and editing audio and text into multimedia compilations; downloadable computer programs for word processing, namely, for editing corresponding text and audio; downloadable computer programs for editing audio and lyrics in text form; downloadable computer software using artificial intelligence (AI) for creating and editing media content and producing media compilations; downloadable computer software for creating avatars in virtual worlds; downloadable computer software for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; downloadable computer software for recording, processing, and editing audio, images, video, and text

28.

COMMUNICATOR

      
Serial Number 98586647
Status Pending
Filing Date 2024-06-05
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content through a software platform for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI assisting in content creation in the field of media editing; downloadable computer software that utilizes artificial intelligence to engage users of a software platform through prompting questions, offering observations, and providing challenges, the software platform allowing the users to record, transcribe, edit, and mix audio, video, text, and other media content in the field of media editing; downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively in the field of media editing; downloadable digital image files of avatars for use in virtual environments and for use in software platforms in connection with downloadable software for recording, transcribing, editing, and mixing audio, video, text, and other media content in the field of media editing; downloadable computer software using artificial intelligence for creating, editing, or compiling audio, video and text content, in a semi- or fully-autonomous manner; downloadable computer software for facilitating the recording, transcribing, editing, and mixing of audio, video, text, and other media content in the field of media editing; downloadable computer software for transforming ideas specified, either audibly or textually, by an individual into usable outputs in an automated manner, while also allowing the individual to edit those outputs for the purpose of producing content, namely, for use in facilitating production of multimedia compilations; downloadable word processor computer programs for enabling individuals to edit audio through the manipulation of corresponding text, and vice versa; downloadable computer programs that enable individuals to edit audio and lyrics in text form; downloadable computer software that uses artificial intelligence for facilitating production of multimedia compilations; downloadable computer software for creating avatars in virtual worlds; downloadable computer software using artificial intelligence through which an individual is able to audibly or textually record thoughts and receive text, audio, images, video, or combinations thereof that are produced as output for creating multimedia compilations; downloadable computer software using artificial intelligence through which inputs are identified, provided, or generated in audible, visual, or textual form and those inputs are used to guide identification or generation of outputs in audible, visual, or textual form for facilitating production of multimedia compilations Providing temporary use of online non-downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content through a software platform for recording, transcribing, editing, and mixing audio, video, text and other media content, the AI assisting in content creation in the field of media editing; providing temporary use of online non-downloadable computer software that utilizes artificial intelligence to engage users of a software platform through prompting questions, offering observations, and providing challenges, the software platform allowing the users to record, transcribe, edit, and mix audio, video, text, and other media content in the field of media editing; providing temporary use of online non-downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively in the field of media editing; providing temporary use of online non-downloadable files of avatars for use in virtual environments and for use in software platforms in connection with downloadable software for recording, transcribing, editing, and mixing audio, video, text, and other media content in the field of media editing; providing temporary use of online non-downloadable computer software for creating, editing, or compiling audio, video and text content, in a semi- or fully-autonomous manner; providing temporary use of online non-downloadable computer software for facilitating the recording, transcribing, editing, and mixing of audio, video, text, and other media content in the field of media editing; providing temporary use of online non-downloadable computer software for transforming ideas specified, either audibly or textually, by an individual into usable outputs in an automated manner, while also allowing the individual to edit those outputs for the purpose of producing content, namely, for use in facilitating production of multimedia compilations; providing temporary use of online non-downloadable word processor computer programs for enabling individuals to edit audio through the manipulation of corresponding text, and vice versa; providing temporary use of online non-downloadable computer programs that enable individuals to edit audio and lyrics in text form; providing temporary use of online non-downloadable computer software for creating avatars in virtual worlds; providing temporary use of online non-downloadable computer software using artificial intelligence through which an individual is able to audibly or textually record thoughts and receive text, audio, images, video, or combinations thereof that are produced as output for creating multimedia compilations; providing temporary use of online non-downloadable computer software using artificial intelligence through which inputs are identified, provided, or generated in audible, visual, or textual form and those inputs are used to guide identification or generation of outputs in audible, visual, or textual form for facilitating production of multimedia compilations

29.

UNDERLORD

      
Serial Number 98586645
Status Registered
Filing Date 2024-06-05
Registration Date 2025-03-18
Owner Descript, Inc. ()
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Providing temporary use of online non-downloadable computer software that utilizes artificial intelligence (AI) for use in creating content, namely, for recording, transcribing, editing, and mixing audio, video, text and other media content; providing temporary use of online non-downloadable computer software that utilizes artificial intelligence (AI) for use in prompting questions, offering observations, and providing challenges to users in the field of recording, transcribing, editing, and mixing audio, video, text, and other media content; providing temporary use of online non-downloadable computer software that utilizes artificial intelligence (AI) for use in providing feedback in the field of writing; providing temporary use of online non-downloadable image files of avatars for use in virtual worlds; providing temporary use of online non-downloadable computer software using artificial intelligence (AI) for use in creating and editing media content and producing media compilations containing audio, video and text content; providing temporary use of online non-downloadable computer software for recording, transcribing, editing, and mixing audio, video, text, and other media content; providing temporary use of online non-downloadable computer software using artificial intelligence (AI) for use in transforming and editing audio and text into multimedia compilations; providing temporary use of online non-downloadable computer programs for word processing, namely, for editing corresponding text and audio; providing temporary use of online non-downloadable computer programs for editing audio and lyrics in text form; providing temporary use of online non-downloadable computer software using  artificial intelligence (AI) for creating and editing media content and producing media compilations; providing temporary use of online non-downloadable computer software for creating avatars in virtual worlds; providing temporary use of online non-downloadable computer software for recording audio and transforming that audio into text, audio, images, video, or combinations thereof; providing temporary use of online non-downloadable computer software for recording, processing, and editing audio, images, video, and text

30.

BRAIN BUDDY

      
Application Number 234266300
Status Pending
Filing Date 2024-05-02
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Downloadable computer software that utilizes artificial intelligence (AI) integrated into software platforms for recording, transcribing, editing, and mixing audio, video, text, images, graphics, the AI assisting in creating audio-visual media content; downloadable computer software that utilizes artificial intelligence (AI) integrated into software platforms in the field of audio and video editing allowing the users to record, transcribe, edit, and mix audio, video, text, with the AI engaging users through prompting questions, offering observations, and providing constructive challenges in the field of audio and video editing; downloadable software providing artificial intelligence which simulates stimulus provided by a great conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure their ideas effectively; downloadable image files of avatars for use in virtual environments and for use in platforms for recording, transcribing, editing, and mixing audio, video, text and other media content; downloadable computer software platforms for recording, transcribing, editing, and mixing of audio, video, and text for virtual computer games; downloadable computer software platforms for recording, transcribing, editing, and mixing audio, video, text, images and graphics; downloadable audio word processing computer software platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form; downloadable computer software using artificial intelligence for audio and video editing; downloadable computer software for use in creating avatars in virtual worlds. (1) Application service provider (ASP) providing computer software applications of others, such computer software applications utilizing artificial intelligence (AI) integrated into software platforms for recording, transcribing, editing, and mixing audio, video, text, images, graphics, the AI assisting in creating audio-visual media content; Application service provider (ASP), providing computer software applications of others featuring computer software that utilizes artificial intelligence (AI) integrated into software platforms in the field of audio and video editing, allowing the users to record, transcribe, edit, and mix audio, video, and text, with the AI engaging users through prompting questions, offering observations, and providing constructive challenges in the field of audio and video editing; application service provider providing artificial intelligence (AI) integrated into platform as a service providing artificial intelligence which simulates stimulus provided by a great conversationalist, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure their ideas effectively; software as a service providing avatars for use in virtual environments and for use in platforms for recording, transcribing, editing, and mixing audio, video, text and other media content; Software as a service (SAAS) services offering computer software using artificial intelligence for audio and video editing; Software as a service (SAAS) services offering software for use in creating avatars in virtual worlds.

31.

Tokenization of text data to facilitate automated discovery of speech disfluencies

      
Application Number 18352145
Grant Number 12651119
Status In Force
Filing Date 2023-07-13
First Publication Date 2023-11-09
Grant Date 2026-06-09
Owner Descript, Inc. (USA)
Inventor
  • De Brébisson, Alexandre
  • D'Andigné, Antoine

Abstract

Introduced here are computer programs and associated computer-implemented techniques for discovering the presence of filler words through tokenization of a transcript derived from audio content. When audio content is obtained by a media production platform, the audio content can be converted into text content as part of a speech-to-text operation. The text content can then be tokenized and labeled using a Natural Language Processing (NLP) library. Tokenizing/labeling may be performed in accordance with a series of rules associated with filler words. At a high level, these rules may examine the text content (and associated tokens/labels) to determine whether patterns, relationships, verbatim, and context indicate that a term is a filler word. Any filler words that are discovered in the text content can be identified as such so that appropriate action(s) can be taken.

IPC Classes  ?

32.

BRAIN BUDDY

      
Serial Number 98258793
Status Pending
Filing Date 2023-11-07
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable computer software that utilizes artificial intelligence (AI) to assist in creating content, namely, content for social media platforms and other distribution channels through a software platform for recording, transcribing, editing, and mixing audio, video, and text, with the AI assisting in content creation; downloadable computer software that utilizes artificial intelligence (AI) integrated into software platforms in the field of audio and video editing allowing the users to record, transcribe, edit, and mix audio, video, and text, with the AI engaging users through prompting questions, offering observations, and providing constructive challenges in the field of audio and video editing; downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist in the field of writing, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively; downloadable image files of avatars for use in virtual environments and for use in software platforms for recording, transcribing, editing, and mixing audio, video, text, and other media content for customizing avatars; downloadable computer software platforms for recording, transcribing, editing, and mixing of audio, video, text, and other media content for virtual computer games; downloadable audio word processing computer software platforms for enabling editors and producers to edit sound files and writers to edit lyrics through the manipulation of corresponding text format; downloadable computer software using artificial intelligence for audio and video editing; downloadable computer software for generating avatars for virtual computer environments Application service provider (ASP), namely, hosting computer software applications of others featuring computer software that utilizes artificial intelligence (AI) to assist in creating content, namely, content for social media platforms and other distribution channels through a software platform for recording, transcribing, editing, and mixing audio, video, and text, with the AI assisting in content creation; Application service provider (ASP), namely, hosting computer software applications of others featuring computer software that utilizes artificial intelligence (AI) integrated into software platforms in the field of audio and video editing, allowing the users to record, transcribe, edit, and mix audio, video, and text, with the AI engaging users through prompting questions, offering observations, and providing constructive challenges in the field of audio and video editing; Application service provider (ASP), namely, hosting computer software applications of others featuring non-downloadable computer software that utilizes artificial intelligence to simulate stimulus provided by a conversationalist in the field of writing, guiding a writer's creative flow and nurturing the writer's ability to flesh out and structure ideas effectively; Software as a service (SAAS) services featuring non-downloadable image files of avatars for use in virtual environments and for use in software platforms for recording, transcribing, editing, and mixing audio, video, text, and other media content for customizing avatars; Software as a service (SAAS) services featuring computer software using artificial intelligence for audio and video editing; Software as a service (SAAS) services featuring software for generating avatars for virtual computer environments

33.

Filler word detection through tokenizing and labeling of transcripts

      
Application Number 18295684
Grant Number 12169691
Status In Force
Filing Date 2023-04-04
First Publication Date 2023-08-03
Grant Date 2024-12-17
Owner Descript, Inc. (USA)
Inventor
  • De Brébisson, Alexandre
  • D'Andigné, Antoine

Abstract

Introduced here are computer programs and associated computer-implemented techniques for discovering the presence of filler words through tokenization of a transcript derived from audio content. When audio content is obtained by a media production platform, the audio content can be converted into text content as part of a speech-to-text operation. The text content can then be tokenized and labeled using a Natural Language Processing (NLP) library. Tokenizing/labeling may be performed in accordance with a series of rules associated with filler words. At a high level, these rules may examine the text content (and associated tokens/labels) to determine whether patterns, relationships, verbatim, and context indicate that a term is a filler word. Any filler words that are discovered in the text content can be identified as such so that appropriate action(s) can be taken.

IPC Classes  ?

34.

Training machine learning frameworks to generate studio-quality recordings through manipulation of noisy audio signals

      
Application Number 18154707
Grant Number 12431154
Status In Force
Filing Date 2023-01-13
First Publication Date 2023-07-20
Grant Date 2025-09-30
Owner Descript, Inc. (USA)
Inventor
  • Seetharaman, Prem S.
  • Kumar, Kundan

Abstract

Introduced here are computer programs and associated computer-implemented techniques for manipulating noisy audio signals to produce clean audio signals that are sufficiently high quality so as to be largely, if not entirely, indistinguishable from “rich” recordings generated by recording studios. When a noisy audio signal is obtained by a media production platform, the noisy audio signal can be manipulated to sound as if recording occurred with sophisticated equipment in a soundproof environment. Manipulation can be performed by a model that, when applied to the noisy audio signal, can manipulate its characteristics so as to emulate the characteristics of clean audio signals that are learned through training.

IPC Classes  ?

  • G10L 21/0232 - Processing in the frequency domain
  • G06F 3/16 - Sound inputSound output
  • G10L 15/06 - Creation of reference templatesTraining of speech recognition systems, e.g. adaptation to the characteristics of the speaker's voice
  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G10L 25/21 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being power information
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

35.

APPROACHES TO GENERATING STUDIO-QUALITY RECORDINGS THROUGH MANIPULATION OF NOISY AUDIO

      
Application Number 18154718
Status Pending
Filing Date 2023-01-13
First Publication Date 2023-07-20
Owner DESCRIPT INC. (USA)
Inventor
  • Seetharaman, Prem S.
  • Kumar, Kundan

Abstract

Introduced here are computer programs and associated computer-implemented techniques for manipulating noisy audio signals to produce clean audio signals that are sufficiently high quality so as to be largely, if not entirely, indistinguishable from “rich” recordings generated by recording studios. When a noisy audio signal is obtained by a media production platform, the noisy audio signal can be manipulated to sound as if recording occurred with sophisticated equipment in a soundproof environment. Manipulation can be performed by a model that, when applied to the noisy audio signal, can manipulate its characteristics so as to emulate the characteristics of clean audio signals that are learned through training.

IPC Classes  ?

  • G10L 21/0232 - Processing in the frequency domain
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks
  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G10L 25/21 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being power information
  • G06F 3/16 - Sound inputSound output

36.

Technologies for creating, altering, and presenting media content

      
Application Number 18179987
Grant Number 12277303
Status In Force
Filing Date 2023-03-07
First Publication Date 2023-07-13
Grant Date 2025-04-15
Owner DESCRIPT, INC. (USA)
Inventor
  • Holmes, Ryan Terrill
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Different types of media experiences can be developed based on characteristics of the consumer. “Linear” experiences may require execution of a pre-built script, although the script could be dynamically modified by a media production platform. Linear experiences can include guided audio tours that are modified or updated based on the location of the consumer. “Enhanced” experiences include conventional media content that is supplemented with intelligent media content. For example, turn-by-turn directions could be supplemented with audio descriptions about the surrounding area. “Freeform” experiences, meanwhile, are those that can continually morph based on information gleaned from a consumer. For example, a radio station may modify what content is being presented based on the geographical metadata uploaded by a computing device associated with the consumer.

IPC Classes  ?

  • G06F 17/00 - Digital computing or data processing equipment or methods, specially adapted for specific functions
  • G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
  • G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
  • G06F 16/68 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
  • G06F 16/683 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content
  • G10L 15/187 - Phonemic context, e.g. pronunciation rules, phonotactical constraints or phoneme n-grams
  • G10L 15/26 - Speech to text systems

37.

Simultaneous recording and uploading of multiple audio files of the same conversation and audio drift normalization systems and methods

      
Application Number 18061609
Grant Number 12342055
Status In Force
Filing Date 2022-12-05
First Publication Date 2023-05-18
Grant Date 2025-06-24
Owner Descript, Inc. (USA)
Inventor Moreno, Zachariah Steven

Abstract

The invention relates to simultaneous recording and uploading systems and methods, and, more particularly to a simultaneous recording and uploading multiple files from the same conversation.

IPC Classes  ?

  • H04N 21/854 - Content authoring
  • G06F 16/11 - File system administration, e.g. details of archiving or snapshots
  • G06F 16/16 - File or folder operations, e.g. details of user interfaces specifically adapted to file systems
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G11B 27/034 - Electronic editing of digitised analogue information signals, e.g. audio or video signals on discs
  • H04L 65/70 - Media network packetisation
  • H04L 67/02 - Protocols based on web technology, e.g. hypertext transfer protocol [HTTP]
  • H04L 67/06 - Protocols specially adapted for file transfer, e.g. file transfer protocol [FTP]
  • H04L 67/10 - Protocols in which an application is distributed across nodes in the network
  • H04N 21/845 - Structuring of content, e.g. decomposing content into time segments
  • H04L 9/40 - Network security protocols

38.

Platform for producing and delivering media content

      
Application Number 17652610
Grant Number 12118266
Status In Force
Filing Date 2022-02-25
First Publication Date 2022-11-24
Grant Date 2024-10-15
Owner Descript, Inc. (USA)
Inventor
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Media content can be created and/or modified using a network-accessible platform. Scripts for content-based experiences could be readily created using one or more interfaces generated by the network-accessible platform. For example, a script for a content-based experience could be created using an interface that permits triggers to be inserted directly into the script. Interface(s) may also allow different media formats to be easily aligned for post-processing. For example, a transcript and an audio file may be dynamically aligned so that the network-accessible platform can globally reflect changes made to either item. User feedback may also be presented directly on the interface(s) so that modifications can be made based on actual user experiences.

IPC Classes  ?

  • G06F 3/16 - Sound inputSound output
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G06Q 10/101 - Collaborative creation, e.g. joint development of products or services
  • G10L 19/008 - Multichannel audio signal coding or decoding using interchannel correlation to reduce redundancy, e.g. joint-stereo, intensity-coding or matrixing
  • G11B 27/031 - Electronic editing of digitised analogue information signals, e.g. audio or video signals
  • G11B 27/10 - IndexingAddressingTiming or synchronisingMeasuring tape travel
  • G11B 27/34 - Indicating arrangements

39.

Simultaneous recording and uploading of multiple audio files of the same conversation and audio drift normalization systems and methods

      
Application Number 17858363
Grant Number 11876850
Status In Force
Filing Date 2022-07-06
First Publication Date 2022-10-27
Grant Date 2024-01-16
Owner DESCRIPT, INC. (USA)
Inventor Moreno, Zachariah Steven

Abstract

The invention relates to audio drift normalization, and more particularly to audio drift normalization systems and methods that can normalize audio drift of a plurality of recordings from a source.

IPC Classes  ?

  • H04L 65/70 - Media network packetisation
  • H04L 67/10 - Protocols in which an application is distributed across nodes in the network

40.

Techniques for creating and presenting media content

      
Application Number 17657931
Grant Number 11747967
Status In Force
Filing Date 2022-04-04
First Publication Date 2022-09-08
Grant Date 2023-09-05
Owner Descript, Inc. (USA)
Inventor
  • Holmes, Ryan Terrill
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Different types of media experiences can be developed based on characteristics of the consumer. “Linear” experiences may require execution of a pre-built script, although the script could be dynamically modified by a media production platform. Linear experiences can include guided audio tours that are modified or updated based on the location of the consumer. “Enhanced” experiences include conventional media content that is supplemented with intelligent media content. For example, turn-by-turn directions could be supplemented with audio descriptions about the surrounding area. “Freeform” experiences, meanwhile, are those that can continually morph based on information gleaned from a consumer. For example, a radio station may modify what content is being presented based on the geographical metadata uploaded by a computing device associated with the consumer.

IPC Classes  ?

  • G06F 17/00 - Digital computing or data processing equipment or methods, specially adapted for specific functions
  • G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
  • G10L 15/187 - Phonemic context, e.g. pronunciation rules, phonotactical constraints or phoneme n-grams
  • G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
  • G06F 16/68 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
  • G10L 15/26 - Speech to text systems
  • G06F 16/683 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content

41.

DESCRIPT

      
Application Number 018705944
Status Registered
Filing Date 2022-05-19
Registration Date 2022-10-05
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 35 - Advertising and business services
  • 41 - Education, entertainment, sporting and cultural services
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable computer software platforms for making podcasts and other audio content; downloadable computer software platforms for recording, transcribing, editing, and mixing podcasts and other media content; downloadable audio word processing computer software platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form; downloadable computer software platforms for multimedia production. Providing business support services in the nature of start-up support for businesses of others related to podcasting and podcasting creation; providing business support services in the nature of multimedia production. Audio recording and production services, namely, making podcasts and other audio content; providing a website featuring blogs in the field of podcasting, podcasting creation, and audio production. Providing temporary use of on-line non-downloadable computer software for making podcasts and other audio content; platform as a service (PAAS) featuring computer software platforms for making podcasts and other audio content; providing temporary use of on-line non-downloadable computer software for recording, transcribing, editing, and mixing podcasts and other media content; platform as a service (PAAS) featuring computer software platforms for recording, transcribing, editing, and mixing podcasts and other media content; technical support services related to podcasting and podcasting creation, namely, troubleshooting in the nature of diagnosing computer software problems in the recording, creating, and editing of media; providing temporary use of on-line non-downloadable computer software enabling editors and producers to edit sound files and writers to edit lyrics in text form; platform as a service (PAAS) featuring computer software and audio and word processing platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form; platform as a service (PAAS) featuring computer software platforms for multimedia production.

42.

DESCRIPT

      
Application Number 218653700
Status Registered
Filing Date 2022-05-18
Registration Date 2026-03-04
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 35 - Advertising and business services
  • 41 - Education, entertainment, sporting and cultural services
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Downloadable computer software platforms for making podcasts and other audio content; downloadable computer software platforms for recording, transcribing, editing, and mixing podcasts and audio-visual media content; downloadable audio word processing computer software platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form; downloadable computer software platforms for podcasts, film, video, television shows, music and radio production. (1) Providing business support services in the nature of start-up support for businesses of others related to podcasting and podcasting creation; providing business support services in the nature of multimedia production. (2) Audio recording and production services, namely, making podcasts and other audio content; providing blogs in the field of podcasting, podcasting creation, and audio production via a website. (3) Providing temporary use of on-line non-downloadable computer software for making podcasts and other audio content; platform as a service (PAAS) featuring computer software platforms for making podcasts and other audio content; providing temporary use of on-line non-downloadable computer software for recording, transcribing, editing, and mixing podcasts and audio-visual media content; platform as a service (PAAS) featuring computer software platforms for recording, transcribing, editing, and mixing podcasts and audio-visual media content; technical support services related to podcasting and podcasting creation, namely, troubleshooting in the nature of diagnosing computer software problems in the recording, creating, and editing of media; providing temporary use of on-line non-downloadable computer software enabling editors and producers to edit sound files and writers to edit lyrics in text form; platform as a service (PAAS) featuring computer software and audio and word processing platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form; platform as a service (PAAS) featuring computer software platforms for podcasts, film, video, television shows, music and radio production.

43.

Upsampling of audio using generative adversarial networks

      
Application Number 17478722
Grant Number 12170096
Status In Force
Filing Date 2021-09-17
First Publication Date 2022-03-31
Grant Date 2024-12-17
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Rithesh
  • Kumar, Kundan

Abstract

Introduced here are approaches to training and then employing computer-implemented models designed to upsample discrete audio signals to higher sampling rates. Assume, for example, that a media production platform obtains a first discrete signal at a relatively low sampling rate. The relatively low sampling frequency may make the first discrete audio signal unsuitable for inclusion in media compilations, so the media production platform may attempt to improve its quality through upsampling. To accomplish this, the media production platform can apply a transform to the first discrete signal to produce a first magnitude spectrogram. Then, the media production platform can apply a computer-implemented model to the first magnitude spectrogram to produce a second magnitude spectrogram. Thereafter, the media production platform can apply an inverse transform to the second magnitude spectrogram to create a second discrete signal that has a higher sampling rate than the first discrete audio signal.

IPC Classes  ?

  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G06F 3/16 - Sound inputSound output
  • G06N 3/045 - Combinations of networks
  • G06N 3/088 - Non-supervised learning, e.g. competitive learning
  • G10L 19/02 - Speech or audio signal analysis-synthesis techniques for redundancy reduction, e.g. in vocodersCoding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

44.

Training generative adversarial networks to upsample audio

      
Application Number 17478734
Grant Number 12159645
Status In Force
Filing Date 2021-09-17
First Publication Date 2022-03-31
Grant Date 2024-12-03
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Rithesh
  • Kumar, Kundan

Abstract

Introduced here are approaches to training and then employing computer-implemented models designed to upsample discrete audio signals to higher sampling rates. Assume, for example, that a media production platform obtains a first discrete signal at a relatively low sampling rate. The relatively low sampling frequency may make the first discrete audio signal unsuitable for inclusion in media compilations, so the media production platform may attempt to improve its quality through upsampling. To accomplish this, the media production platform can apply a transform to the first discrete signal to produce a first magnitude spectrogram. Then, the media production platform can apply a computer-implemented model to the first magnitude spectrogram to produce a second magnitude spectrogram. Thereafter, the media production platform can apply an inverse transform to the second magnitude spectrogram to create a second discrete signal that has a higher sampling rate than the first discrete audio signal.

IPC Classes  ?

  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G06F 3/16 - Sound inputSound output
  • G06N 3/045 - Combinations of networks
  • G06N 3/088 - Non-supervised learning, e.g. competitive learning
  • G10L 19/02 - Speech or audio signal analysis-synthesis techniques for redundancy reduction, e.g. in vocodersCoding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

45.

Tokenization of text data to facilitate automated discovery of speech disfluencies

      
Application Number 17094554
Grant Number 11741303
Status In Force
Filing Date 2020-11-10
First Publication Date 2022-02-03
Grant Date 2023-08-29
Owner Descript, Inc. (USA)
Inventor
  • De Brébisson, Alexandre
  • D'Andigné, Antoine

Abstract

Introduced here are computer programs and associated computer-implemented techniques for discovering the presence of filler words through tokenization of a transcript derived from audio content. When audio content is obtained by a media production platform, the audio content can be converted into text content as part of a speech-to-text operation. The text content can then be tokenized and labeled using a Natural Language Processing (NLP) library. Tokenizing/labeling may be performed in accordance with a series of rules associated with filler words. At a high level, these rules may examine the text content (and associated tokens/labels) to determine whether patterns, relationships, verbatim, and context indicate that a term is a filler word. Any filler words that are discovered in the text content can be identified as such so that appropriate action(s) can be taken.

IPC Classes  ?

46.

Filler word detection through tokenizing and labeling of transcripts

      
Application Number 17094533
Grant Number 11651157
Status In Force
Filing Date 2020-11-10
First Publication Date 2022-02-03
Grant Date 2023-05-16
Owner Descript, Inc. (USA)
Inventor
  • De Brébisson, Alexandre
  • D'Andigné, Antoine

Abstract

Introduced here are computer programs and associated computer-implemented techniques for discovering the presence of filler words through tokenization of a transcript derived from audio content. When audio content is obtained by a media production platform, the audio content can be converted into text content as part of a speech-to-text operation. The text content can then be tokenized and labeled using a Natural Language Processing (NLP) library. Tokenizing/labeling may be performed in accordance with a series of rules associated with filler words. At a high level, these rules may examine the text content (and associated tokens/labels) to determine whether patterns, relationships, verbatim, and context indicate that a term is a filler word. Any filler words that are discovered in the text content can be identified as such so that appropriate action(s) can be taken.

IPC Classes  ?

47.

Transcript correction through programmatic comparison of independently generated transcripts

      
Application Number 17127166
Grant Number 12136423
Status In Force
Filing Date 2020-12-18
First Publication Date 2021-06-24
Grant Date 2024-11-05
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Kundan
  • Anand, Vicki

Abstract

Introduced here are computer programs and associated computer-implemented techniques for facilitating the creation of a master transcription (or simply “transcript”) that more accurately reflects underlying audio by comparing multiple independently generated transcripts. The master transcript may be used to record and/or produce various forms of media content, as further discussed below. Thus, the technology described herein may be used to facilitate editing of text content, audio content, or video content. These computer programs may be supported by a media production platform that is able to generate the interfaces through which individuals (also referred to as “users”) can create, edit, or view media content. For example, a computer program may be embodied as a word processor that allows individuals to edit voice-based audio content by editing a master transcript, and vice versa.

IPC Classes  ?

  • G10L 15/26 - Speech to text systems
  • G06F 3/16 - Sound inputSound output
  • G10L 15/01 - Assessment or evaluation of speech recognition systems
  • G10L 15/08 - Speech classification or search
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G10L 15/30 - Distributed recognition, e.g. in client-server systems, for mobile phones or network applications
  • G10L 15/32 - Multiple recognisers used in sequence or in parallelScore combination systems therefor, e.g. voting systems

48.

Automated generation of transcripts through independent transcription

      
Application Number 17127235
Grant Number 12062373
Status In Force
Filing Date 2020-12-18
First Publication Date 2021-06-24
Grant Date 2024-08-13
Owner Descript, Inc. (USA)
Inventor
  • Kumar, Kundan
  • Anand, Vicki

Abstract

Introduced here are computer programs and associated computer-implemented techniques for facilitating the creation of a master transcription (or simply “transcript”) that more accurately reflects underlying audio by comparing multiple independently generated transcripts. The master transcript may be used to record and/or produce various forms of media content, as further discussed below. Thus, the technology described herein may be used to facilitate editing of text content, audio content, or video content. These computer programs may be supported by a media production platform that is able to generate the interfaces through which individuals (also referred to as “users”) can create, edit, or view media content. For example, a computer program may be embodied as a word processor that allows individuals to edit voice-based audio content by editing a master transcript, and vice versa.

IPC Classes  ?

  • G10L 15/26 - Speech to text systems
  • G06F 3/16 - Sound inputSound output
  • G10L 15/01 - Assessment or evaluation of speech recognition systems
  • G10L 15/08 - Speech classification or search
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G10L 15/30 - Distributed recognition, e.g. in client-server systems, for mobile phones or network applications
  • G10L 15/32 - Multiple recognisers used in sequence or in parallelScore combination systems therefor, e.g. voting systems

49.

Simultaneous recording and uploading of multiple audio files of the same conversation and audio drift normalization systems and methods

      
Application Number 17119764
Grant Number 11540030
Status In Force
Filing Date 2020-12-11
First Publication Date 2021-06-17
Grant Date 2022-12-27
Owner DESCRIPT, INC. (USA)
Inventor Moreno, Zachariah Steven

Abstract

The invention relates to simultaneous recording and uploading systems and methods, and, more particularly to a simultaneous recording and uploading of multiple files from the same conversation.

IPC Classes  ?

  • H04L 29/08 - Transmission control procedure, e.g. data link level control procedure
  • H04N 21/854 - Content authoring
  • G11B 27/034 - Electronic editing of digitised analogue information signals, e.g. audio or video signals on discs
  • H04N 21/845 - Structuring of content, e.g. decomposing content into time segments
  • G06F 16/16 - File or folder operations, e.g. details of user interfaces specifically adapted to file systems
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G06F 16/11 - File system administration, e.g. details of archiving or snapshots
  • H04L 67/02 - Protocols based on web technology, e.g. hypertext transfer protocol [HTTP]
  • H04L 67/06 - Protocols specially adapted for file transfer, e.g. file transfer protocol [FTP]
  • H04L 67/10 - Protocols in which an application is distributed across nodes in the network
  • H04L 65/70 - Media network packetisation
  • H04L 9/40 - Network security protocols

50.

Simultaneous recording and uploading of multiple audio files of the same conversation and audio drift normalization systems and methods

      
Application Number 17119784
Grant Number 11388489
Status In Force
Filing Date 2020-12-11
First Publication Date 2021-06-17
Grant Date 2022-07-12
Owner DESCRIPT, INC. (USA)
Inventor Moreno, Zachariah Steven

Abstract

The invention relates to audio drift normalization, and more particularly to audio drift normalization systems and methods that can normalize audio drift of a plurality of recordings from a source.

IPC Classes  ?

  • H04N 21/854 - Content authoring
  • H04N 21/845 - Structuring of content, e.g. decomposing content into time segments
  • G06F 16/11 - File system administration, e.g. details of archiving or snapshots
  • H04L 67/06 - Protocols specially adapted for file transfer, e.g. file transfer protocol [FTP]
  • G11B 27/034 - Electronic editing of digitised analogue information signals, e.g. audio or video signals on discs
  • H04L 65/60 - Network streaming of media packets
  • G06F 16/16 - File or folder operations, e.g. details of user interfaces specifically adapted to file systems
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • H04L 67/02 - Protocols based on web technology, e.g. hypertext transfer protocol [HTTP]
  • H04L 67/10 - Protocols in which an application is distributed across nodes in the network
  • H04L 9/40 - Network security protocols

51.

Techniques for creating and presenting media content

      
Application Number 16736156
Grant Number 11294542
Status In Force
Filing Date 2020-01-07
First Publication Date 2020-05-07
Grant Date 2022-04-05
Owner Descript, Inc. (USA)
Inventor
  • Holmes, Ryan Terrill
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Different types of media experiences can be developed based on characteristics of the consumer. “Linear” experiences may require execution of a pre-built script, although the script could be dynamically modified by a media production platform. Linear experiences can include guided audio tours that are modified or updated based on the location of the consumer. “Enhanced” experiences include conventional media content that is supplemented with intelligent media content. For example, turn-by-turn directions could be supplemented with audio descriptions about the surrounding area. “Freeform” experiences, meanwhile, are those that can continually morph based on information gleaned from a consumer. For example, a radio station may modify what content is being presented based on the geographical metadata uploaded by a computing device associated with the consumer.

IPC Classes  ?

  • G06F 17/00 - Digital computing or data processing equipment or methods, specially adapted for specific functions
  • G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
  • G10L 15/187 - Phonemic context, e.g. pronunciation rules, phonotactical constraints or phoneme n-grams
  • G06F 3/04817 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
  • G06F 16/68 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
  • G10L 15/26 - Speech to text systems
  • G06F 16/683 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content

52.

DESCRIPT

      
Serial Number 88981249
Status Registered
Filing Date 2020-03-16
Registration Date 2021-12-14
Owner Descript, Inc. ()
NICE Classes  ?
  • 35 - Advertising and business services
  • 41 - Education, entertainment, sporting and cultural services

Goods & Services

Providing business support services in the nature of start-up support for businesses of others related to podcasting and podcasting creation Audio recording and production services, namely, making podcasts and other audio content; providing a website featuring blogs in the field of podcasting, podcasting creation, and audio production

53.

D

      
Serial Number 88836546
Status Registered
Filing Date 2020-03-16
Registration Date 2021-03-16
Owner Descript, Inc. (USA)
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 35 - Advertising and business services
  • 41 - Education, entertainment, sporting and cultural services
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable and recorded computer software platforms for making podcasts and other audio content; downloadable and recorded computer software platforms for recording, transcribing, editing, mixing podcasts and other media content; downloadable and recorded audio word processing computer software platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form Providing business support services related to podcasting and podcasting creation Audio recording and production services, namely, making podcasts and other audio content; providing a website featuring blogs in the field of podcasting, podcasting creation, and audio production Providing temporary use of on-line non-downloadable computer software for making podcasts and other audio content; platform as a service (PAAS) featuring computer software platforms for making podcasts and other audio content; providing temporary use of on-line non-downloadable computer software for recording, transcribing, editing, mixing podcasts and other media content; platform as a service (PAAS) featuring computer software platforms for recording, transcribing, editing, mixing podcasts and other media content; technical support services related to podcasting and podcasting creation, namely, troubleshooting in the nature of diagnosing computer software problems; providing temporary use of on-line non-downloadable computer software enabling editors and producers to edit sound files and writers to edit lyrics in text form; platform as a service (PAAS) featuring computer software and audio and word processing platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form

54.

DESCRIPT

      
Serial Number 88836547
Status Registered
Filing Date 2020-03-16
Registration Date 2022-09-06
Owner Descript, Inc. ()
NICE Classes  ?
  • 09 - Scientific and electric apparatus and instruments
  • 42 - Scientific, technological and industrial services, research and design

Goods & Services

Downloadable computer software platforms for making podcasts and other audio content; downloadable computer software platforms for recording, transcribing, editing, and mixing podcasts and other media content; downloadable audio word processing computer software platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form Providing temporary use of on-line non-downloadable computer software for making podcasts and other audio content; platform as a service (PAAS) featuring computer software platforms for making podcasts and other audio content; providing temporary use of on-line non-downloadable computer software for recording, transcribing, editing, and mixing podcasts and other media content; platform as a service (PAAS) featuring computer software platforms for recording, transcribing, editing, and mixing podcasts and other media content; technical support services related to podcasting and podcasting creation, namely, troubleshooting in the nature of diagnosing computer software problems in the recording, creating, and editing of media; providing temporary use of on-line non-downloadable computer software enabling editors and producers to edit sound files and writers to edit lyrics in text form; platform as a service (PAAS) featuring computer software and audio and word processing platforms enabling editors and producers to edit sound files and writers to edit lyrics in text form

55.

Platform for producing and delivering media content

      
Application Number 16600095
Grant Number 11262970
Status In Force
Filing Date 2019-10-11
First Publication Date 2020-02-06
Grant Date 2022-03-01
Owner Descript, Inc. (USA)
Inventor
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Media content can be created and/or modified using a network-accessible platform. Scripts for content-based experiences could be readily created using one or more interfaces generated by the network-accessible platform. For example, a script for a content-based experience could be created using an interface that permits triggers to be inserted directly into the script. Interface(s) may also allow different media formats to be easily aligned for post-processing. For example, a transcript and an audio file may be dynamically aligned so that the network-accessible platform can globally reflect changes made to either item. User feedback may also be presented directly on the interface(s) so that modifications can be made based on actual user experiences.

IPC Classes  ?

  • G06F 3/16 - Sound inputSound output
  • G11B 27/031 - Electronic editing of digitised analogue information signals, e.g. audio or video signals
  • G10L 19/008 - Multichannel audio signal coding or decoding using interchannel correlation to reduce redundancy, e.g. joint-stereo, intensity-coding or matrixing
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G11B 27/34 - Indicating arrangements
  • G11B 27/10 - IndexingAddressingTiming or synchronisingMeasuring tape travel
  • G06Q 10/10 - Office automationTime management

56.

Techniques for creating and presenting media content

      
Application Number 15835266
Grant Number 10564817
Status In Force
Filing Date 2017-12-07
First Publication Date 2018-06-21
Grant Date 2020-02-18
Owner DESCRIPT, INC. (USA)
Inventor
  • Holmes, Ryan Terrill
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Different types of media experiences can be developed based on characteristics of the consumer. “Linear” experiences may require execution of a pre-built script, although the script could be dynamically modified by a media production platform. Linear experiences can include guided audio tours that are modified or updated based on the location of the consumer. “Enhanced” experiences include conventional media content that is supplemented with intelligent media content. For example, turn-by-turn directions could be supplemented with audio descriptions about the surrounding area. “Freeform” experiences, meanwhile, are those that can continually morph based on information gleaned from a consumer. For example, a radio station may modify what content is being presented based on the geographical metadata uploaded by a computing device associated with the consumer.

IPC Classes  ?

  • G06F 17/20 - Handling natural language data
  • G06F 3/0484 - Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
  • G10L 15/26 - Speech to text systems
  • G10L 15/187 - Phonemic context, e.g. pronunciation rules, phonotactical constraints or phoneme n-grams
  • G06F 3/0481 - Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance
  • G06F 16/68 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually
  • G06F 16/683 - Retrieval characterised by using metadata, e.g. metadata not derived from the content or metadata generated manually using metadata automatically derived from the content

57.

Platform for producing and delivering media content

      
Application Number 15716957
Grant Number 10445052
Status In Force
Filing Date 2017-09-27
First Publication Date 2018-04-05
Grant Date 2019-10-15
Owner DESCRIPT, INC. (USA)
Inventor
  • Rubin, Steven Surmacz
  • Schwekendiek, Ulf
  • Williams, David John

Abstract

Media content can be created and/or modified using a network-accessible platform. Scripts for content-based experiences could be readily created using one or more interfaces generated by the network-accessible platform. For example, a script for a content-based experience could be created using an interface that permits triggers to be inserted directly into the script. Interface(s) may also allow different media formats to be easily aligned for post-processing. For example, a transcript and an audio file may be dynamically aligned so that the network-accessible platform can globally reflect changes made to either item. User feedback may also be presented directly on the interface(s) so that modifications can be made based on actual user experiences.

IPC Classes  ?

  • G06F 17/00 - Digital computing or data processing equipment or methods, specially adapted for specific functions
  • G06F 3/16 - Sound inputSound output
  • G11B 27/031 - Electronic editing of digitised analogue information signals, e.g. audio or video signals
  • G10L 21/02 - Speech enhancement, e.g. noise reduction or echo cancellation
  • G10L 19/008 - Multichannel audio signal coding or decoding using interchannel correlation to reduce redundancy, e.g. joint-stereo, intensity-coding or matrixing
  • G06F 16/61 - IndexingData structures thereforStorage structures
  • G06Q 10/10 - Office automationTime management