Pindrop Security, Inc.

United States of America

Back to Profile

1-100 of 282 for Pindrop Security, Inc. Sort by
Query
Aggregations
IP Type
        Patent 244
        Trademark 38
Jurisdiction
        United States 163
        World 63
        Canada 56
Date
New (last 4 weeks) 1
2026 August 1
2026 July 1
2026 (YTD) 16
2025 35
See more
IPC Class
G10L 17/04 - Training, enrolment or model building 69
G10L 17/18 - Artificial neural networksConnectionist approaches 52
H04M 3/42 - Systems providing special services or facilities to subscribers 46
H04M 3/22 - Arrangements for supervision, monitoring or testing 44
G06N 20/00 - Machine learning 38
See more
NICE Class
42 - Scientific, technological and industrial services, research and design 36
09 - Scientific and electric apparatus and instruments 22
37 - Construction and mining; installation and repair services 13
Status
Pending 69
Registered / In Force 213
  1     2     3        Next Page

1.

ACTIVE VOICE LIVENESS DETECTION SYSTEM

      
Application Number 19633427
Status Pending
Filing Date 2026-03-30
First Publication Date 2026-08-13
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Klein, Nicholas
  • Stankus, Anthony

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination
  • G10L 25/60 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination for measuring the quality of voice signals
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

2.

REAL HUMAN + RIGHT HUMAN

      
Application Number 1926944
Status Registered
Filing Date 2026-04-15
Registration Date 2026-04-15
Owner Pindrop Security, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Providing temporary use of non-downloadable cloud-based software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention in connection with software applications and digital communications, namely, text, data, audio, and video communications, telephony and voice communications, contact center interactions, interactive voice response (IVR) systems, conferencing, messaging, remote access, screen sharing, data transfer, and document collaboration; application service provider (ASP) services featuring application programming interface (API) software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention in connection with software applications and digital communications, including telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; software as a service (SaaS) services featuring software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, identity verification, and fraud prevention in connection with software applications and digital communications, including telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; cloud computing featuring software for detection and analysis of synthetic or manipulated media, including audio, voice, music, and video content generated or altered by artificial intelligence, including deepfakes and replay attacks; providing temporary use of online non-downloadable software for real-time analysis of audio and video communications, including telephony and voice communications and contact center interactions, for detection of synthetic or manipulated media and fraud; providing temporary use of online non-downloadable software for user authentication, biometric authentication, liveness detection, and identity and geolocation verification in connection with telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; technical support services, namely, remote and on-site installation, deployment, implementation, management, and maintenance of computer software systems for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention.

3.

DEEPFAKE DETECTION

      
Application Number 19414714
Status Pending
Filing Date 2025-12-10
First Publication Date 2026-04-23
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 15/08 - Speech classification or search
  • G06N 20/00 - Machine learning
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 15/26 - Speech to text systems
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

4.

REAL HUMAN + RIGHT HUMAN

      
Application Number 248720700
Status Pending
Filing Date 2026-04-15
Owner Pindrop Security, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Providing temporary use of non-downloadable cloud-based software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention in connection with software applications and digital communications, namely, text, data, audio, and video communications, telephony and voice communications, contact center interactions, interactive voice response (IVR) systems, conferencing, messaging, remote access, screen sharing, data transfer, and document collaboration; application service provider (ASP) services featuring application programming interface (API) software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention in connection with software applications and digital communications, including telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; software as a service (SaaS) services featuring software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, identity verification, and fraud prevention in connection with software applications and digital communications, including telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; cloud computing featuring software for detection and analysis of synthetic or manipulated media, including audio, voice, music, and video content generated or altered by artificial intelligence, including deepfakes and replay attacks; providing temporary use of online non-downloadable software for real-time analysis of audio and video communications, including telephony and voice communications and contact center interactions, for detection of synthetic or manipulated media and fraud; providing temporary use of online non-downloadable software for user authentication, biometric authentication, liveness detection, and identity and geolocation verification in connection with telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; technical support services, namely, remote and on-site installation, deployment, implementation, management, and maintenance of computer software systems for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention.

5.

REAL HUMAN + RIGHT HUMAN

      
Serial Number 99729663
Status Pending
Filing Date 2026-03-27
Owner Pindrop Security, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Providing temporary use of non-downloadable cloud-based software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention in connection with software applications and digital communications, namely, text, data, audio, and video communications, telephony and voice communications, contact center interactions, interactive voice response (IVR) systems, conferencing, messaging, remote access, screen sharing, data transfer, and document collaboration; Application service provider (ASP) services featuring application programming interface (API) software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention in connection with software applications and digital communications, including telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; Software as a service (SaaS) services featuring software for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, identity verification, and fraud prevention in connection with software applications and digital communications, including telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; Cloud computing featuring software for detection and analysis of synthetic or manipulated media, including audio, voice, music, and video content generated or altered by artificial intelligence, including deepfakes and replay attacks; Providing temporary use of online non-downloadable software for real-time analysis of audio and video communications, including telephony and voice communications and contact center interactions, for detection of synthetic or manipulated media and fraud; Providing temporary use of online non-downloadable software for user authentication, biometric authentication, liveness detection, and identity and geolocation verification in connection with telephony and voice communications, contact center interactions, and interactive voice response (IVR) systems; Technical support services, namely, remote and on-site installation, deployment, implementation, management, and maintenance of computer software systems for authentication, biometric authentication, liveness detection, behavioral analysis, risk scoring, and fraud prevention.

6.

SYSTEMS AND METHODS TO PREVENT DENIAL OF SERVICE ATTACKS FROM GENERATIVE AI VOICE BOTS

      
Application Number US2025045724
Publication Number 2026/059982
Status In Force
Filing Date 2025-09-10
Publication Date 2026-03-19
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Balasubramaniyan, Vijay
  • Gupta, Amit

Abstract

Embodiments disclosed herein include software processes and of machine-learning architectures for detecting and mitigating against synthetic speech instances. A computer analyzes audio speech data and metadata received with contact events associated with source identifiers. The computer executes machine-learning architecture(s) that determine whether the contact events likely include human-generated speech or machine-generated synthetic speech. The computer may determine the likelihood that contact events represent a DoS attack launched by a source device, by analyzing behavior features in metadata associated with the source identifier. The computer determines whether the contact events originated from the source user device having the source identifier launched a DoS attack and, if so, may update a blocklist. The blocklist may be stored in a database and includes one or more source identifiers that should be rejected or blocked at the current or inbound contact event or at future contact events for the particular source identifiers.

IPC Classes  ?

7.

SYSTEMS AND METHODS TO PREVENT DENIAL OF SERVICE ATTACKS FROM GENERATIVE AI VOICE BOTS

      
Application Number 19324843
Status Pending
Filing Date 2025-09-10
First Publication Date 2026-03-12
Owner Pindrop Security, Inc. (USA)
Inventor
  • Balasubramaniyan, Vijay
  • Gupta, Amit

Abstract

Embodiments disclosed herein include software processes and of machine-learning architectures for detecting and mitigating against synthetic speech instances. A computer analyzes audio speech data and metadata received with contact events associated with source identifiers. The computer executes machine-learning architecture(s) that determine whether the contact events likely include human-generated speech or machine-generated synthetic speech. The computer may determine the likelihood that contact events represent a DoS attack launched by a source device, by analyzing behavior features in metadata associated with the source identifier. The computer determines whether the contact events originated from the source user device having the source identifier launched a DoS attack and, if so, may update a blocklist. The blocklist may be stored in a database and includes one or more source identifiers that should be rejected or blocked at the current or inbound contact event or at future contact events for the particular source identifiers.

IPC Classes  ?

  • H04L 9/40 - Network security protocols
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices

8.

SINGING VOICE DEEPFAKE DETECTION

      
Application Number 19309516
Status Pending
Filing Date 2025-08-25
First Publication Date 2026-03-05
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Tak, Hemlata
  • Casal, Ricardo
  • Sivaraman, Ganesh
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that detect machine-generated synthetic singing vocals in a vocal audio signal of an audio signal using a multi-stage machine-learning architecture. A singing detector identifies vocal segments containing singing. A singing liveness detector includes a fakeprint embedding extractor that extracts fakeprint feature vector embeddings representing artifacts of machine-generated vocal signals, scoring layers or classifier layers to generate a singing liveness score for identifying the likelihood a vocal signal is human-generated or synthetic. An optional singer detector includes a vocalprint embedding extractor that extracts vocalprint feature vector embeddings representing singer-specific vocal identity characteristics and generates a singer identification score or attribution score for identifying a particular singer in the vocal signal.

IPC Classes  ?

  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/18 - Artificial neural networksConnectionist approaches

9.

SINGING VOICE DEEPFAKE DETECTION

      
Application Number US2025043455
Publication Number 2026/050204
Status In Force
Filing Date 2025-08-26
Publication Date 2026-03-05
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Tak, Hemlata
  • Casal, Ricardo
  • Sivaraman, Ganesh
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that detect machine-generated synthetic singing vocals in a vocal audio signal of an audio signal using a multi-stage machine-learning architecture. A singing detector identifies vocal segments containing singing. A singing liveness detector includes a fakeprint embedding extractor that extracts fakeprint feature vector embeddings representing artifacts of machine-generated vocal signals, scoring layers or classifier layers to generate a singing liveness score for identifying the likelihood a vocal signal is human-generated or synthetic. An optional singer detector includes a vocalprint embedding extractor that extracts vocalprint feature vector embeddings representing singer-specific vocal identity characteristics and generates a singer identification score or attribution score for identifying a particular singer in the vocal signal.

IPC Classes  ?

  • G10L 25/81 - Detection of presence or absence of voice signals for discriminating voice from music
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/12 - Score normalisation
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices

10.

CALL AUTHENTICATION AT THE CALL CENTER USING A MOBILE DEVICE

      
Document Number 03274834
Status Pending
Filing Date 2020-08-27
Open to Public Date 2026-03-02
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gupta, Payas
  • Nelms Ii, Terry

Abstract

Embodiments described herein provide for automatically authenticating telephone calls to an enterprise call center. The system disclosed herein builds on the trust of a data channel for the telephony channel. Certain types of authentication information can be received through the telephony channel, as well. But the mobile application associated with the call center system may provide additional or alternative forms of data through the data channel. The system may send requests to a mobile application of a device to provide information that can reliably be assumed to be coming from that particular device, such as a state of the device and/or a user's response to push notifications. In some cases, the authentication processes may be based on quantity and quality of matches between certain metadata or attributes expected to be received from a given device as compared to the metadata or attributes received.

IPC Classes  ?

  • H04M 3/436 - Arrangements for screening incoming calls
  • H04M 3/487 - Arrangements for providing information services, e.g. recorded voice services or time announcements
  • H04W 12/72 - Subscriber identity

11.

PASSIVE AND CONTINUOUS MULTI-SPEAKER VOICE BIOMETRICS

      
Application Number 19372835
Status Pending
Filing Date 2025-10-29
First Publication Date 2026-02-26
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Kumar, Avrosh
  • Antolic-Soban, Ivan

Abstract

Embodiments described herein provide for a voice biometrics system execute machine-learning architectures capable of passive, active, continuous, or static operations, or a combination thereof. Systems passively and/or continuously, in some cases in addition to actively and/or statically, enrolling speakers as the speakers speak into or around an edge device (e.g., car, television, radio, phone). The system identifies users on the fly without requiring a new speaker to mirror prompted utterances for reconfiguring operations. The system manages speaker profiles as speakers provide utterances to the system. Machine-learning architectures implement a passive and continuous voice biometrics system, possibly without knowledge of speaker identities. The system creates identities in an unsupervised manner, sometimes passively enrolling and recognizing known or unknown speakers. The system offers personalization and security across a wide range of applications, including media content for over-the-top services and IoT devices (e.g., personal assistants, vehicles), and call centers.

IPC Classes  ?

  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G06N 20/00 - Machine learning
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

12.

OMNI CHANNEL AUTHENTICATION

      
Application Number 19370564
Status Pending
Filing Date 2025-10-27
First Publication Date 2026-02-19
Owner Pindrop Security, Inc. (USA)
Inventor
  • Merchant, Mohammedali
  • Gupta, Payas

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures providing improved omni-channel authentication solutions. Embodiments include one or more computing devices that provide an authentication interface by which various communication channels may deposit contact or session data received via a first-channel session into a non-transitory storage medium of an authentication database for another channel to obtain and employ (e.g., verify users). This allows the customer to access an online data channel and enter the contact center through a telephony communication channel, but further allows the enterprise contact center systems to passively maintain access to various types of information about the user's identity captured from each contact channel, allowing the call center to request or capture authenticating information (e.g., voice biometrics) from both channels to employ authentication processes for one or both channels, such as voice biometrics authentication processes or other types of authentication functions.

IPC Classes  ?

13.

SILENT CALLER ID VERIFICATION USING CALLBACK REQUEST

      
Application Number 19370212
Status Pending
Filing Date 2025-10-27
First Publication Date 2026-02-19
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gupta, Payas
  • Nelms, Ii, Terry

Abstract

Disclosed herein are embodiments of systems, methods, and products comprises an authentication server for caller ID verification. When a caller makes a phone call, the server receives the phone call and verifies whether the phone call is from a registered device associated with the phone number. The server queries the registered device to retrieve one or more current call states via an authentication function on the registered device. The server compares the states and/or state transitions to the observed states and/or state transitions of the phone call. If the registered device states and/or state transitions match the observed phone call states and/or state transitions, the server verifies that the phone call is from the registered device and not some imposter's device. If there is no such match, the server rejects the phone call before the call phone is connected or terminates the phone call after the phone call is connected.

IPC Classes  ?

  • H04M 3/436 - Arrangements for screening incoming calls
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 19/04 - Current supply arrangements for telephone systems providing ringing current or supervisory tones, e.g. dialling tone or busy tone the ringing-current being generated at the substations

14.

CROSS-LINGUAL SPEAKER RECOGNITION

      
Application Number 19359565
Status Pending
Filing Date 2025-10-15
First Publication Date 2026-02-12
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Chen, Tianxiang
  • Kumar, Avrosh
  • Sivaraman, Ganesh
  • Phatak, Kedar

Abstract

Disclosed are systems and methods including computing-processes executing machine-learning architectures for voice biometrics, in which the machine-learning architecture implements one or more language compensation functions. Embodiments include an embedding extraction engine (sometimes referred to as an “embedding extractor”) that extracts speaker embeddings and determines a speaker similarity score for determine or verifying the likelihood that speakers in different audio signals are the same speaker. The machine-learning architecture further includes a multi-class language classifier that determines a language likelihood score that indicates the likelihood that a particular audio signal includes a spoken language. The features and functions of the machine-learning architecture described herein may implement the various language compensation techniques to provide more accurate speaker recognition results, regardless of the language spoken by the speaker.

IPC Classes  ?

  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/10 - Multimodal systems, i.e. based on the integration of multiple recognition engines or fusion of expert systems

15.

CARRIER SIGNALING BASED AUTHENTICATION AND FRAUD DETECTION

      
Application Number 19343941
Status Pending
Filing Date 2025-09-29
First Publication Date 2026-01-22
Owner Pindrop Security, Inc. (USA)
Inventor
  • Casal, Ricky
  • Maddali, Vinay
  • Gupta, Payas
  • Patil, Kailash

Abstract

Disclosed are systems and methods including computing-processes, which may include layers of machine-learning architectures, for assessing risk for calls directed to call center systems using carrier signaling metadata. A computer evaluates carrier signaling metadata to perform various new risk-scoring techniques to determine riskiness of calls and authenticate calls. When determining a risk score for an incoming call is received at a call center system, the computer may obtain certain metadata values from inbound metadata, prior call metadata, or from third-party telecommunications services and executes processes for determining the risk score for the call. The risk score operations include several scoring components, including appliance print scoring, carrier detection scoring, ANI location detection scoring, location similarity scoring, and JIP-ANI location similarity scoring, among others.

IPC Classes  ?

  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

16.

ROBUST SPOOFING DETECTION SYSTEM USING DEEP RESIDUAL NEURAL NETWORKS

      
Application Number 19328724
Status Pending
Filing Date 2025-09-15
First Publication Date 2026-01-15
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Embodiments described herein provide for systems and methods for implementing a neural network architecture for spoof detection in audio signals. The neural network architecture contains a layers defining embedding extractors that extract embeddings from input audio signals. Spoofprint embeddings are generated for particular system enrollees to detect attempts to spoof the enrollee's voice. Optionally, voiceprint embeddings are generated for the system enrollees to recognize the enrollee's voice. The voiceprints are extracted using features related to the enrollee's voice. The spoofprints are extracted using features related to features of how the enrollee speaks and other artifacts. The spoofprints facilitate detection of efforts to fool voice biometrics using synthesized speech (e.g., deepfakes) that spoof and emulate the enrollee's voice.

IPC Classes  ?

  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/22 - Interactive proceduresMan-machine interfaces

17.

METHOD AND APPARATUS FOR DETECTING SPOOF CONDITIONS

      
Document Number 03289758
Status Pending
Filing Date 2018-03-02
Open to Public Date 2025-11-29
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Nagarsheth, Parav
  • Patil, Kailash
  • Garland, Matthew

IPC Classes  ?

  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination

18.

ONE TIME VOICE PASSPHRASE TO PROTECT AGAINST MAN-IN-THE-MIDDLE ATTACK

      
Application Number US2025030590
Publication Number 2025/245352
Status In Force
Filing Date 2025-05-22
Publication Date 2025-11-27
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gupta, Amit
  • Merchant, Mohammedali
  • Balasubramaniyan, Vijay

Abstract

Embodiments described herein provide for automatically authenticating operation requests and end-users who submit operation requests during contact events. A server obtains an operation request for an operation originated at an end-user device. The server generates a voice-based one-time password (OTP) using contextual information associated with the requested operation. The server generates and transmits an OTP prompt having text representing the OTP for display at a user interface of the user device. The server receives a response including an audio signal that contains the recording of the user speaking the OTP text aloud. The server uses the audio signal to authenticate the user and the operation request based on the speaker's voice, the accuracy of the user speaking the OTP, and liveness or fraud detection features extracted from the audio signal or metadata from the user device.

IPC Classes  ?

  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase
  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G10L 15/20 - Speech recognition techniques specially adapted for robustness in adverse environments, e.g. in noise or of stress induced speech
  • H04W 12/06 - Authentication
  • G06Q 20/40 - Authorisation, e.g. identification of payer or payee, verification of customer or shop credentialsReview and approval of payers, e.g. check of credit lines or negative lists
  • G10L 17/04 - Training, enrolment or model building

19.

ONE TIME VOICE PASSPHRASE TO PROTECT AGAINST MAN-IN-THE-MIDDLE ATTACK

      
Application Number 19216344
Status Pending
Filing Date 2025-05-22
First Publication Date 2025-11-27
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gupta, Amit
  • Merchant, Mohammedali
  • Balasubramaniyan, Vijay

Abstract

Embodiments described herein provide for automatically authenticating operation requests and end-users who submit operation requests during contact events. A server obtains an operation request for an operation originated at an end-user device. The server generates a voice-based one-time password (OTP) using contextual information associated with the requested operation. The server generates and transmits an OTP prompt having text representing the OTP for display at a user interface of the user device. The server receives a response including an audio signal that contains the recording of the user speaking the OTP text aloud. The server uses the audio signal to authenticate the user and the operation request based on the speaker's voice, the accuracy of the user speaking the OTP, and liveness or fraud detection features extracted from the audio signal or metadata from the user device.

IPC Classes  ?

20.

ONE TIME VOICE PASSPHRASE TO PROTECT AGAINST MAN-IN-THE-MIDDLE ATTACK

      
Application Number 19216408
Status Pending
Filing Date 2025-05-22
First Publication Date 2025-11-27
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gupta, Amit
  • Merchant, Mohammedali
  • Balasubramaniyan, Vijay

Abstract

Embodiments described herein provide for automatically authenticating operation requests and end-users who submit operation requests during contact events. A server obtains an operation request for an operation originated at an end-user device. The server generates a voice-based one-time password (OTP) using contextual information associated with the requested operation. The server generates and transmits an OTP prompt having text representing the OTP for display at a user interface of the user device. The server receives a response including an audio signal that contains the recording of the user speaking the OTP text aloud. The server uses the audio signal to authenticate the user and the operation request based on the speaker's voice, the accuracy of the user speaking the OTP, and liveness or fraud detection features extracted from the audio signal or metadata from the user device.

IPC Classes  ?

  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction

21.

ONE TIME VOICE PASSPHRASE TO PROTECT AGAINST MAN-IN-THE-MIDDLE ATTACK

      
Application Number 19216469
Status Pending
Filing Date 2025-05-22
First Publication Date 2025-11-27
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gupta, Amit
  • Merchant, Mohammedali
  • Balasubramaniyan, Vijay

Abstract

Embodiments described herein provide for automatically authenticating operation requests and end-users who submit operation requests during contact events. A server obtains an operation request for an operation originated at an end-user device. The server generates a voice-based one-time password (OTP) using contextual information associated with the requested operation. The server generates and transmits an OTP prompt having text representing the OTP for display at a user interface of the user device. The server receives a response including an audio signal that contains the recording of the user speaking the OTP text aloud. The server uses the audio signal to authenticate the user and the operation request based on the speaker's voice, the accuracy of the user speaking the OTP, and liveness or fraud detection features extracted from the audio signal or metadata from the user device.

IPC Classes  ?

  • H04L 9/40 - Network security protocols
  • G10L 17/00 - Speaker identification or verification techniques

22.

SYSTEMS AND METHODS OF SPEAKER-INDEPENDENT EMBEDDING FOR IDENTIFICATION AND VERIFICATION FROM AUDIO

      
Application Number 19281043
Status Pending
Filing Date 2025-07-25
First Publication Date 2025-11-20
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Phatak, Kedar
  • Khoury, Elie

Abstract

Embodiments described herein provide for audio processing operations that evaluate characteristics of audio signals that are independent of the speaker's voice. A neural network architecture trains and applies discriminatory neural networks tasked with modeling and classifying speaker-independent characteristics. The task-specific models generate or extract feature vectors from input audio data based on the trained embedding extraction models. The embeddings from the task-specific models are concatenated to form a deep-phoneprint vector for the input audio signal. The DP vector is a low dimensional representation of the each of the speaker-independent characteristics of the audio signal and applied in various downstream operations.

IPC Classes  ?

  • G06F 8/65 - Updates
  • H04L 9/14 - Arrangements for secret or secure communicationsNetwork security protocols using a plurality of keys or algorithms
  • H04L 9/32 - Arrangements for secret or secure communicationsNetwork security protocols including means for verifying the identity or authority of a user of the system

23.

ROBUST SPOOFING DETECTION SYSTEM USING DEEP RESIDUAL NEURAL NETWORKS

      
Document Number 03256373
Status Pending
Filing Date 2021-01-22
Open to Public Date 2025-10-31
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Chen, Tianxiang
  • Khoury, Elie

IPC Classes  ?

  • G10L 17/18 - Artificial neural networksConnectionist approaches

24.

SYSTEMS AND METHODS OF SPEAKER-INDEPENDENT EMBEDDING FOR IDENTIFICATION AND VERIFICATION FROM AUDIO

      
Document Number 03274583
Status Pending
Filing Date 2021-03-04
Open to Public Date 2025-10-31
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Phatak, Kedar
  • Khoury, Elie

IPC Classes  ?

  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 25/27 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique

25.

CHANNEL-COMPENSATED LOW-LEVEL FEATURES FOR SPEAKER RECOGNITION

      
Application Number 19245159
Status Pending
Filing Date 2025-06-20
First Publication Date 2025-10-09
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Garland, Matthew

Abstract

A system for generating channel-compensated features of a speech signal includes a channel noise simulator that degrades the speech signal, a feed forward convolutional neural network (CNN) that generates channel-compensated features of the degraded speech signal, and a loss function that computes a difference between the channel-compensated features and handcrafted features for the same raw speech signal. Each loss result may be used to update connection weights of the CNN until a predetermined threshold loss is satisfied, and the CNN may be used as a front-end for a deep neural network (DNN) for speaker recognition/verification. The DNN may include convolutional layers, a bottleneck features layer, multiple fully-connected layers, and an output layer. The bottleneck features may be used to update connection weights of the convolutional layers, and dropout may be applied to the convolutional layers.

IPC Classes  ?

  • G10L 17/20 - Pattern transformations or operations aimed at increasing system robustness, e.g. against channel noise or different working conditions
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 19/028 - Noise substitution, e.g. substituting non-tonal spectral components by noisy source

26.

DEEPFAKE DETECTION SYSTEM COMBINING DISCRIMINATIVE FEATURES FROM VOICED AND UNVOICED SEGMENTS OF SPEECH

      
Application Number US2025019476
Publication Number 2025/193772
Status In Force
Filing Date 2025-03-12
Publication Date 2025-09-18
Owner PINDROP SECURITY, INC. (USA)
Inventor Sivaraman, Ganesh

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech ("deepfakes"). The server applies a machine-learning architecture that includes a segmentation engine trained to parse an audio signal into voiced- speech segments and unvoiced-speech segments. Each segment type is analyzed by respective deepfake detectors. A first deepfake detector generates a first risk score for the voiced-speech segment, while a second deepfake detector generates a second risk score for the unvoiced- speech segment. The machine-learning architecture includes fusion layers to algorithmically combine the risk scores to determine and overall risk score. In training, the server uses loss functions to calculate losses indicating distances or discrepancies between the generated risk scores and expected risk scores provided by training labels. Based on the loss, the server updates the parameters of the respective deepfake detectors or segmentation engine.

IPC Classes  ?

  • G06F 21/31 - User authentication
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 15/06 - Creation of reference templatesTraining of speech recognition systems, e.g. adaptation to the characteristics of the speaker's voice
  • G06N 20/00 - Machine learning
  • G06F 21/00 - Security arrangements for protecting computers, components thereof, programs or data against unauthorised activity

27.

DIFFUSION-BASED AUDIO PURIFICATION FOR DEFENDING AGAINST ADVERSARIAL DEEPFAKE ATTACKS

      
Application Number US2025019477
Publication Number 2025/193773
Status In Force
Filing Date 2025-03-12
Publication Date 2025-09-18
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Kassis, Andre
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech ("deepfakes"). Embodiments implement a machine-learning architecture having a diffusion model that generates purified features that are fed to a deepfake detection model. The machine-learning architecture includes input layers that convert an audio signal into a Gaussian or frequency space representation (e.g., log spectrogram) to extract a set of initial features indicative of spoofing or deepfake attacks. The diffusion model identifies adversarial noise on the audio signal in the initial features and generates purified features or clean version of the input audio signal. A deepfake detector includes a neural network architecture and classifier programmed and trained to generate a deepfake detection score and classify the audio signal as genuine or fraudulent using the purified features.

IPC Classes  ?

  • G10L 19/018 - Audio watermarking, i.e. embedding inaudible data in the audio signal
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination
  • G06F 21/82 - Protecting input, output or interconnection devices
  • G06N 20/00 - Machine learning
  • G06F 21/00 - Security arrangements for protecting computers, components thereof, programs or data against unauthorised activity

28.

SOURCE TRACING OF AUDIO DEEPFAKE SYSTEMS

      
Document Number 03322682
Status Pending
Filing Date 2025-03-12
Open to Public Date 2025-09-18
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Klein, Nicholas
  • Tak, Hemlata
  • Casal, Ricardo
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that implement a machine-learning architecture for audio source tracing for deepfake detection. The computer extracts a feature vector representing features of the input audio signal. The machine-learning architecture includes one or more embedding extractors for extracting one or more feature vectors from the input audio signal. An attribute detector ingests an embedding and scoring layers generate a source-indicating attribute score. A source tracer includes a multi-class classifier to generate a signal source score using the attribute scores and generates a signal source class.

IPC Classes  ?

  • G06F 21/82 - Protecting input, output or interconnection devices
  • G06N 20/00 - Machine learning
  • G10L 19/018 - Audio watermarking, i.e. embedding inaudible data in the audio signal
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination

29.

DIFFUSION-BASED AUDIO PURIFICATION FOR DEFENDING AGAINST ADVERSARIAL DEEPFAKE ATTACKS

      
Application Number 19076895
Status Pending
Filing Date 2025-03-11
First Publication Date 2025-09-18
Owner Pindrop Security, Inc. (USA)
Inventor
  • Kassis, Andre
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”). Embodiments implement a machine-learning architecture having a diffusion model that generates purified features that are fed to a deepfake detection model. The machine-learning architecture includes input layers that convert an audio signal into a Gaussian or frequency space representation (e.g., log spectrogram) to extract a set of initial features indicative of spoofing or deepfake attacks. The diffusion model identifies adversarial noise on the audio signal in the initial features and generates purified features or clean version of the input audio signal. A deepfake detector includes a neural network architecture and classifier programmed and trained to generate a deepfake detection score and classify the audio signal as genuine or fraudulent using the purified features.

IPC Classes  ?

  • G06F 21/55 - Detecting local intrusion or implementing counter-measures
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches

30.

DEEPFAKE DETECTION SYSTEM COMBINING DISCRIMINATIVE FEATURES FROM VOICED AND UNVOICED SEGMENTS OF SPEECH

      
Application Number 19076935
Status Pending
Filing Date 2025-03-11
First Publication Date 2025-09-18
Owner Pindrop Security, Inc. (USA)
Inventor Sivaraman, Ganesh

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”). The server applies a machine-learning architecture that includes a segmentation engine trained to parse an audio signal into voiced-speech segments and unvoiced-speech segments. Each segment type is analyzed by respective deepfake detectors. A first deepfake detector generates a first risk score for the voiced-speech segment, while a second deepfake detector generates a second risk score for the unvoiced-speech segment. The machine-learning architecture includes fusion layers to algorithmically combine the risk scores to determine and overall risk score. In training, the server uses loss functions to calculate losses indicating distances or discrepancies between the generated risk scores and expected risk scores provided by training labels. Based on the loss, the server updates the parameters of the respective deepfake detectors or segmentation engine.

IPC Classes  ?

  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/06 - Decision making techniquesPattern matching strategies

31.

SOURCE TRACING OF AUDIO DEEPFAKE SYSTEMS

      
Application Number 19076960
Status Pending
Filing Date 2025-03-11
First Publication Date 2025-09-18
Owner Pindrop Security, Inc. (USA)
Inventor
  • Klein, Nicholas
  • Tak, Hemlata
  • Casal, Ricardo
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that implement a machine-learning architecture for audio source tracing for deepfake detection. The computer extracts a feature vector representing features of the input audio signal. The machine-learning architecture includes one or more embedding extractors for extracting one or more feature vectors from the input audio signal. An attribute detector ingests an embedding and scoring layers generate a source-indicating attribute score. A source tracer includes a multi-class classifier to generate a signal source score using the attribute scores and generates a signal source class.

IPC Classes  ?

  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G06F 9/451 - Execution arrangements for user interfaces
  • G10L 17/04 - Training, enrolment or model building

32.

SOURCE TRACING OF AUDIO DEEPFAKE SYSTEMS

      
Application Number US2025019475
Publication Number 2025/193771
Status In Force
Filing Date 2025-03-12
Publication Date 2025-09-18
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Klein, Nicholas
  • Tak, Hemlata
  • Casal, Ricardo
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server that implement a machine-learning architecture for audio source tracing for deepfake detection. The computer extracts a feature vector representing features of the input audio signal. The machine-learning architecture includes one or more embedding extractors for extracting one or more feature vectors from the input audio signal. An attribute detector ingests an embedding and scoring layers generate a source-indicating attribute score. A source tracer includes a multi-class classifier to generate a signal source score using the attribute scores and generates a signal source class.

IPC Classes  ?

  • G10L 19/018 - Audio watermarking, i.e. embedding inaudible data in the audio signal
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination
  • G06F 21/82 - Protecting input, output or interconnection devices
  • G06N 20/00 - Machine learning
  • G06F 21/00 - Security arrangements for protecting computers, components thereof, programs or data against unauthorised activity

33.

FRAUD IMPORTANCE SYSTEM

      
Application Number 19209628
Status Pending
Filing Date 2025-05-15
First Publication Date 2025-09-04
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Phatak, Kedar
  • Raghuram, Jayaram

Abstract

Embodiments described herein provide for a fraud detection engine for detecting various types of fraud at a call center and a fraud importance engine for tailoring the fraud detection operations to relative importance of fraud events. Fraud importance engine determines which fraud events are comparative more important than others. The fraud detection engine comprises machine-learning models that consume contact data and fraud importance information for various anti-fraud processes. The fraud importance engine calculates importance scores for fraud events based on user-customized attributes, such as fraud-type or fraud activity. The fraud importance scores are used in various processes, such as model training, model selection, and selecting weights or hyper-parameters for the ML models, among others. The fraud detection engine uses the importance scores to prioritize fraud alerts for review. The fraud importance engine receives detection feedback, which contacts involved false negatives, where fraud events were undetected but should have been detected.

IPC Classes  ?

34.

PRESENTATION ATTACKS IN REVERBERANT CONDITIONS

      
Document Number 03231427
Status Pending
Filing Date 2024-03-08
Open to Public Date 2025-06-25
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gaubitch, Nikolay
  • Looney, David

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures including obtaining training audio signals having corresponding training impulse responses associated with reverberation degradation, training a machine-learning model of a presentation attack detection engine to generate one or more acoustic parameters by executing the presentation attack detection engine using the training impulse responses of the training audio signals and a loss function, obtaining an audio signal having an acoustic impulse response associated with reverberation degradation caused by one or more rooms, generating the one or more acoustic parameters for the audio signal by executing the machine-learning model using the audio signal as input, and generating an attack score for the audio signal based upon the one or more parameters generated by the machine learning model.

IPC Classes  ?

  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/06 - Decision making techniquesPattern matching strategies

35.

PINDROP

      
Serial Number 99233705
Status Registered
Filing Date 2025-06-13
Registration Date 2026-01-20
Owner Pindrop Security, Inc. ()
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Providing temporary use of non-downloadable cloud-based software for use in authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; application service provider featuring application programming interface (API) software for authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; software as a service (SaaS) services featuring software for authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; remote and on-site installation, deployment, implementation, management and maintenance of computer software systems for use in authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration

36.

END-TO-END SPEAKER RECOGNITION USING DEEP NEURAL NETWORK

      
Document Number 03244290
Status Pending
Filing Date 2017-09-11
Open to Public Date 2025-06-13
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Garland, Matthew

IPC Classes  ?

  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/18 - Artificial neural networksConnectionist approaches

37.

P

      
Serial Number 99233708
Status Registered
Filing Date 2025-06-13
Registration Date 2026-01-20
Owner Pindrop Security, Inc. ()
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Providing temporary use of non-downloadable cloud-based software for use in authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; application service provider featuring application programming interface (API) software for authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; software as a service (SaaS) services featuring software for authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; remote and on-site installation, deployment, implementation, management and maintenance of computer software systems for use in authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration

38.

P PINDROP

      
Serial Number 99234033
Status Registered
Filing Date 2025-06-13
Registration Date 2026-01-20
Owner Pindrop Security, Inc. ()
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Providing temporary use of non-downloadable cloud-based software for use in authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; application service provider featuring application programming interface (API) software for authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; software as a service (SaaS) services featuring software for authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration; remote and on-site installation, deployment, implementation, management and maintenance of computer software systems for use in authentication and fraud prevention for software applications, text, data, audio, and video communications, interactions and transactions, video, audio, and data conferencing, instant messaging and discussion forums, remote computer access and screen sharing, data transfer, and document and data collaboration

39.

METHOD AND APPARATUS FOR THREAT IDENTIFICATION THROUGH ANALYSIS OF COMMUNICATIONS SIGNALING, EVENTS, AND PARTICIPANTS

      
Document Number 03244288
Status Pending
Filing Date 2017-08-02
Open to Public Date 2025-06-13
Owner PINDROP SECURITY, INC. (USA)
Inventor Douglas, Lance

IPC Classes  ?

  • H04M 3/436 - Arrangements for screening incoming calls

40.

CENTRALIZED SYNTHETIC SPEECH DETECTION SYSTEM USING WATERMARKING

      
Document Number 03249059
Status Pending
Filing Date 2024-07-19
Open to Public Date 2025-06-05
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Looney, David
  • Gaubitch, Nikolay
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server for obtaining, by a computer, an audio signal including synthetic speech, extracting, by the computer, metadata from a watermark of the audio signal by applying a set of keys associated with a plurality of text-to-speech (TTS) services to the audio signal, the metadata indicating an origin of the synthetic speech in the audio signal, and generating, by the computer, based on the extracted metadata, a notification indicating that the audio signal includes the synthetic speech.

IPC Classes  ?

  • G10L 17/00 - Speaker identification or verification techniques

41.

ROBUST SPREAD-SPECTRUM SPEECH WATERMARKING USING LINEAR PREDICTION AND DEEP SPECTRAL SHAPING

      
Document Number 03254055
Status Pending
Filing Date 2024-09-11
Open to Public Date 2025-06-03
Owner Pindrop Security, Inc. (USA)
Inventor
  • Looney, David
  • Gaubitch, Nikolay

Abstract

Embodiments disclosed herein include software processes executed by a computer for encoding and decoding watermarks for a speech signal in a call signal communicated via telephony channels. An encoder uses Linear Predictive Coding (LPC) to analyzes the call signal’s spectral envelope and embeds the watermark into the LPC log-spectrum of the speech signal of the call signal. The encoder may reduce the watermark’s strength at a formant peak of the speech signal, balancing the watermark’s robustness and detectability. A deep decoder includes a neural network architecture trained on watermarked and watermark-free speech signals having various types of degradation to extract a feature vector of a call signal and compute a watermark detection score for one or more frames or for the call signal. At inference time, the deep decoder detects the watermark when the watermark detection score satisfies a detection threshold.

IPC Classes  ?

  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 19/018 - Audio watermarking, i.e. embedding inaudible data in the audio signal

42.

CALLER VERIFICATION VIA CARRIER METADATA

      
Application Number 19023199
Status Pending
Filing Date 2025-01-15
First Publication Date 2025-05-15
Owner Pindrop Security, Inc. (USA)
Inventor
  • Cornwell, John
  • Nelms, Ii, Terry

Abstract

Embodiments described herein provide for passive caller verification and/or passive fraud risk assessments for calls to customer call centers. Systems and methods may be used in real time as a call is coming into a call center. An analytics server of an analytics service looks at the purported Caller ID of the call, as well as the unaltered carrier metadata, which the analytics server then uses to generate or retrieve one or more probability scores using one or more lookup tables and/or a machine-learning model. A probability score indicates the likelihood that information derived using the Caller ID information has occurred or should occur given the carrier metadata received with the inbound call. The one or more probability scores be used to generate a risk score for the current call that indicates the probability of the call being valid (e.g., originated from a verified caller or calling device, non-fraudulent).

IPC Classes  ?

  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention
  • G06F 18/214 - Generating training patternsBootstrap methods, e.g. bagging or boosting
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers

43.

SPEAKER RECOGNITION WITH QUALITY INDICATORS

      
Application Number 18989690
Status Pending
Filing Date 2024-12-20
First Publication Date 2025-04-17
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Rao, Hrishikesh
  • Phatak, Kedar
  • Khoury, Elie

Abstract

Embodiments described herein provide for a machine-learning architecture for modeling quality measures for enrollment signals. Modeling these enrollment signals enables the machine-learning architecture to identify deviations from expected or ideal enrollment signal in future test phase calls. These differences can be used to generate quality measures for the various audio descriptors or characteristics of audio signals. The quality measures can then be fused at the score-level with the speaker recognition's embedding comparisons for verifying the speaker. Fusing the quality measures with the similarity scoring essentially calibrates the speaker recognition's outputs based on the realities of what is actually expected for the enrolled caller and what was actually observed for the current inbound caller.

IPC Classes  ?

  • G10L 25/60 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination for measuring the quality of voice signals
  • G06N 20/20 - Ensemble learning
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit

44.

PINDROP PULSE

      
Application Number 1845118
Status Registered
Filing Date 2025-02-08
Registration Date 2025-02-08
Owner Pindrop Security, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Cloud computing featuring software for use in the detection of various types of synthetic content such as voices, audio, music, or video including deepfakes, machine-based replay, and other media manipulated or generated by artificial intelligence.

45.

ROBUST SPREAD-SPECTRUM SPEECH WATERMARKING USING LINEAR PREDICTION AND DEEP SPECTRAL SHAPING

      
Application Number 18883681
Status Pending
Filing Date 2024-09-12
First Publication Date 2025-03-20
Owner Pindrop Security, Inc. (USA)
Inventor
  • Looney, David
  • Gaubitch, Nikolay

Abstract

Embodiments disclosed herein include software processes executed by a computer for encoding and decoding watermarks for a speech signal in a call signal communicated via telephony channels. An encoder uses Linear Predictive Coding (LPC) to analyzes the call signal's spectral envelope and embeds the watermark into the LPC log-spectrum of the speech signal of the call signal. The encoder may reduce the watermark's strength at a formant peak of the speech signal, balancing the watermark's robustness and detectability. A deep decoder includes a neural network architecture trained on watermarked and watermark-free speech signals having various types of degradation to extract a feature vector of a call signal and compute a watermark detection score for one or more frames or for the call signal. At inference time, the deep decoder detects the watermark when the watermark detection score satisfies a detection threshold.

IPC Classes  ?

  • G10L 19/018 - Audio watermarking, i.e. embedding inaudible data in the audio signal
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

46.

METHOD AND APPARATUS FOR THREAT IDENTIFICATION THROUGH ANALYSIS OF COMMUNICATIONS SIGNALING, EVENTS, AND PARTICIPANTS

      
Application Number 18943686
Status Pending
Filing Date 2024-11-11
First Publication Date 2025-02-27
Owner Pindrop Security, Inc. (USA)
Inventor Douglas, Lance

Abstract

Aspects of the invention determining a threat score of a call traversing a telecommunications network by leveraging the signaling used to originate, propagate and terminate the call. Outer-edge data utilized to originate the call may be analyzed against historical, or third party real-time data to determine the propensity of calls originating from those facilities to be categorized as a threat. Storing the outer edge data before the call is sent over the communications network permits such data to be preserved and not subjected to manipulations during traversal of the communications network. This allows identification of threat attempts based on the outer edge data from origination facilities, thereby allowing isolation of a compromised network facility that may or may not be known to be compromised by its respective network owner. Other aspects utilize inner edge data from an intermediate node of the communications network which may be analyzed against other inner edge data from other intermediate nodes and/or outer edge data.

IPC Classes  ?

  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04L 9/40 - Network security protocols
  • H04M 3/436 - Arrangements for screening incoming calls
  • H04M 7/00 - Arrangements for interconnection between switching centres

47.

PINDROP PULSE

      
Application Number 238718900
Status Pending
Filing Date 2025-02-08
Owner Pindrop Security, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

(1) Cloud computing featuring software for use in the detection of various types of synthetic content such as voices, audio, music, or video including deepfakes, machine-based replay, and other media manipulated or generated by artificial intelligence.

48.

TELECOMMUNICATIONS VALIDATION SYSTEM AND METHOD

      
Application Number 18913429
Status Pending
Filing Date 2024-10-11
First Publication Date 2025-01-30
Owner Pindrop Security, Inc. (USA)
Inventor
  • Merchant, Mohammedali
  • Williams, Matthew
  • Prugar, Tim

Abstract

According to an embodiment of the disclosure, a toll-free telecommunications validation system determines a confidence value that an incoming phone call to an enterprises' toll-free number is originating from the station it purports to be by incorporating one or more layers of signals and data in determining said confidence value. The data and signals can include one or more call identifiers and/or toll-free call routing logs, service control point (SCP) signals and data, service data point (SDP) signals and data, dialed number information service (DNIS) signals and data, session initiation protocol (SIP) signals and data, carrier identification code (CIC) signals and data, location routing number (LRN) signals and data, jurisdiction information parameter (JIP) signals and data, charge number (CN) signals and data, billing number (BN) signals and data, and originating carrier information (such as information derived from the ANI).

IPC Classes  ?

  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers

49.

AUDIOVISUAL DEEPFAKE DETECTION

      
Application Number 18918928
Status Pending
Filing Date 2024-10-17
First Publication Date 2025-01-30
Owner Pindrop Security, Inc. (USA)
Inventor
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

The embodiments execute machine-learning architectures for biometric-based identity recognition (e.g., speaker recognition, facial recognition) and deepfake detection (e.g., speaker deepfake detection, facial deepfake detection). The machine-learning architecture includes layers defining multiple scoring components, including sub-architectures for speaker deepfake detection, speaker recognition, facial deepfake detection, facial recognition, and lip-sync estimation engine. The machine-learning architecture extracts and analyzes various types of low-level features from both audio data and visual data, combines the various scores, and uses the scores to determine the likelihood that the audiovisual data contains deepfake content and the likelihood that a claimed identity of a person in the video matches to the identity of an expected or enrolled person. This enables the machine-learning architecture to perform identity recognition and verification, and deepfake detection, in an integrated fashion, for both audio data and visual data.

IPC Classes  ?

  • G06V 40/40 - Spoof detection, e.g. liveness detection
  • G06F 18/21 - Design or setup of recognition systems or techniquesExtraction of features in feature spaceBlind source separation
  • G06F 18/22 - Matching criteria, e.g. proximity measures
  • G06V 20/40 - ScenesScene-specific elements in video content
  • G06V 40/16 - Human faces, e.g. facial parts, sketches or expressions
  • G06V 40/70 - Multimodal biometrics, e.g. combining information from different biometric modalities
  • G10L 17/22 - Interactive proceduresMan-machine interfaces

50.

AUDIOVISUAL DEEPFAKE DETECTION

      
Application Number 18919049
Status Pending
Filing Date 2024-10-17
First Publication Date 2025-01-30
Owner Pindrop Security, Inc. (USA)
Inventor
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

The embodiments execute machine-learning architectures for biometric-based identity recognition (e.g., speaker recognition, facial recognition) and deepfake detection (e.g., speaker deepfake detection, facial deepfake detection). The machine-learning architecture includes layers defining multiple scoring components, including sub-architectures for speaker deepfake detection, speaker recognition, facial deepfake detection, facial recognition, and lip-sync estimation engine. The machine-learning architecture extracts and analyzes various types of low-level features from both audio data and visual data, combines the various scores, and uses the scores to determine the likelihood that the audiovisual data contains deepfake content and the likelihood that a claimed identity of a person in the video matches to the identity of an expected or enrolled person. This enables the machine-learning architecture to perform identity recognition and verification, and deepfake detection, in an integrated fashion, for both audio data and visual data.

IPC Classes  ?

  • G06V 40/40 - Spoof detection, e.g. liveness detection
  • G06F 18/21 - Design or setup of recognition systems or techniquesExtraction of features in feature spaceBlind source separation
  • G06F 18/22 - Matching criteria, e.g. proximity measures
  • G06V 20/40 - ScenesScene-specific elements in video content
  • G06V 40/16 - Human faces, e.g. facial parts, sketches or expressions
  • G06V 40/70 - Multimodal biometrics, e.g. combining information from different biometric modalities
  • G10L 17/22 - Interactive proceduresMan-machine interfaces

51.

CENTRALIZED SYNTHETIC SPEECH DETECTION SYSTEM USING WATERMARKING

      
Application Number 18777278
Status Pending
Filing Date 2024-07-18
First Publication Date 2025-01-23
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Looney, David
  • Gaubitch, Nikolay
  • Khoury, Elie

Abstract

Disclosed are systems and methods including software processes executed by a server for obtaining, by a computer, an audio signal including synthetic speech, extracting, by the computer, metadata from a watermark of the audio signal by applying a set of keys associated with a plurality of text-to-speech (TTS) services to the audio signal, the metadata indicating an origin of the synthetic speech in the audio signal, and generating, by the computer, based on the extracted metadata, a notification indicating that the audio signal includes the synthetic speech.

IPC Classes  ?

  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction

52.

ACTIVE VOICE LIVENESS DETECTION SYSTEM

      
Document Number 03289953
Status Pending
Filing Date 2024-04-25
Open to Public Date 2024-10-31
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Khoury, Elie
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Stankus, Anthony
  • Klein, Nicholas

IPC Classes  ?

  • G06F 21/31 - User authentication
  • G06N 3/04 - Architecture, e.g. interconnection topology
  • G10L 13/10 - Prosody rules derived from textStress or intonation
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches

53.

DEEPFAKE DETECTION

      
Application Number 18388412
Status Pending
Filing Date 2023-11-09
First Publication Date 2024-10-31
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

54.

Deepfake detection

      
Application Number 18388466
Grant Number 12562150
Status In Force
Filing Date 2023-11-09
First Publication Date 2024-10-31
Grant Date 2026-02-24
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/16 - Speech classification or search using artificial neural networks

55.

ACTIVE VOICE LIVENESS DETECTION SYSTEM

      
Application Number 18646375
Status Pending
Filing Date 2024-04-25
First Publication Date 2024-10-31
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Klein, Nicholas
  • Stankus, Anthony

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/00 - Speaker identification or verification techniques

56.

ACTIVE VOICE LIVENESS DETECTION SYSTEM

      
Application Number 18646493
Status Pending
Filing Date 2024-04-25
First Publication Date 2024-10-31
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Klein, Nicholas
  • Stankus, Anthony

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 25/60 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination for measuring the quality of voice signals
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

57.

Active voice liveness detection system

      
Application Number 18646228
Grant Number 12706098
Status In Force
Filing Date 2024-04-25
First Publication Date 2024-10-31
Grant Date 2026-08-11
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Klein, Nicholas
  • Stankus, Anthony

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/00 - Speaker identification or verification techniques
  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination
  • G10L 25/60 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination for measuring the quality of voice signals
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks

58.

Active voice liveness detection system

      
Application Number 18646310
Grant Number 12592239
Status In Force
Filing Date 2024-04-25
First Publication Date 2024-10-31
Grant Date 2026-03-31
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Klein, Nicholas
  • Stankus, Anthony

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination
  • G10L 25/60 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination for measuring the quality of voice signals
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

59.

ACTIVE VOICE LIVENESS DETECTION SYSTEM

      
Application Number 18646431
Status Pending
Filing Date 2024-04-25
First Publication Date 2024-10-31
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Klein, Nicholas
  • Stankus, Anthony

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates

60.

ACTIVE VOICE LIVENESS DETECTION SYSTEM

      
Application Number US2024026212
Publication Number 2024/226757
Status In Force
Filing Date 2024-04-25
Publication Date 2024-10-31
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Sivaraman, Ganesh
  • Chen, Tianxiang
  • Gaubitch, Nikolay
  • Looney, David
  • Khoury, Elie
  • Gupta, Amit
  • Balasubramaniyan, Vijay
  • Stankus, Anthony
  • Klein, Nicholas

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech ("deepfakes") in a call conversation. Embodiments include systems and methods for detecting fraudulent presentation attacks using multiple functional engines that implement various fraud-detection techniques, to produce calibrated scores and/or fused scores. A computer may, for example, evaluate the audio quality of speech signals within audio signals, where speech signals contain the speech portions having speaker utterances.

IPC Classes  ?

  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/04 - Training, enrolment or model building
  • G10L 13/10 - Prosody rules derived from textStress or intonation
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G06N 3/04 - Architecture, e.g. interconnection topology
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog
  • G06F 21/31 - User authentication

61.

DEEPFAKE DETECTION

      
Application Number 18388364
Status Pending
Filing Date 2023-11-09
First Publication Date 2024-10-24
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

62.

DEEPFAKE DETECTION

      
Application Number 18388428
Status Pending
Filing Date 2023-11-09
First Publication Date 2024-10-24
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

63.

Deepfake detection

      
Application Number 18388447
Grant Number 12592220
Status In Force
Filing Date 2023-11-09
First Publication Date 2024-10-24
Grant Date 2026-03-31
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 15/08 - Speech classification or search
  • G06N 20/00 - Machine learning
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 15/26 - Speech to text systems
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

64.

Deepfake detection

      
Application Number 18388457
Grant Number 12711951
Status In Force
Filing Date 2023-11-09
First Publication Date 2024-10-24
Grant Date 2026-08-18
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 15/08 - Speech classification or search
  • G06N 20/00 - Machine learning
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 15/26 - Speech to text systems
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

65.

DEEPFAKE DETECTION

      
Application Number US2024023576
Publication Number 2024/220274
Status In Force
Filing Date 2024-04-08
Publication Date 2024-10-24
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech ("deepfakes") in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice "liveness" detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • H04M 3/436 - Arrangements for screening incoming calls
  • H04W 12/12 - Detection or prevention of fraud
  • H04M 3/50 - Centralised arrangements for answering callsCentralised arrangements for recording messages for absent or busy subscribers
  • G06V 40/40 - Spoof detection, e.g. liveness detection
  • G10L 15/04 - SegmentationWord boundary detection
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • G10L 21/0272 - Voice signal separating

66.

DEEPFAKE DETECTION

      
Document Number 03288456
Status Pending
Filing Date 2024-04-08
Open to Public Date 2024-10-24
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

IPC Classes  ?

  • G06V 40/40 - Spoof detection, e.g. liveness detection
  • G10L 15/04 - SegmentationWord boundary detection
  • G10L 21/0272 - Voice signal separating
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/436 - Arrangements for screening incoming calls
  • H04M 3/50 - Centralised arrangements for answering callsCentralised arrangements for recording messages for absent or busy subscribers
  • H04W 12/12 - Detection or prevention of fraud

67.

Deepfake detection

      
Application Number 18388385
Grant Number 12525224
Status In Force
Filing Date 2023-11-09
First Publication Date 2024-10-24
Grant Date 2026-01-13
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 15/00 - Speech recognition
  • G06N 20/00 - Machine learning
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/08 - Speech classification or search
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 15/26 - Speech to text systems
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

68.

Deepfake detection

      
Application Number 18439049
Grant Number 12676144
Status In Force
Filing Date 2024-02-12
First Publication Date 2024-10-24
Grant Date 2026-07-07
Owner Pindrop Security, Inc. (USA)
Inventor
  • Altaf, Umair
  • Peri, Sai Pradeep
  • Phatela, Lakshay
  • Gupta, Payas
  • Sun, Yitao
  • Afanaseva, Svetlana
  • Patil, Kailash
  • Khoury, Elie
  • Magnetta, Bradley
  • Balasubramaniyan, Vijay
  • Chen, Tianxiang

Abstract

Disclosed are systems and methods including software processes executed by a server that detect audio-based synthetic speech (“deepfakes”) in a call conversation. The server applies an NLP engine to transcribe call audio and analyze the text for anomalous patterns to detect synthetic speech. Additionally or alternatively, the server executes a voice “liveness” detection system for detecting machine speech, such as synthetic speech or replayed speech. The system performs phrase repetition detection, background change detection, and passive voice liveness detection in call audio signals to detect liveness of a speech utterance. An automated model update module allows the liveness detection model to adapt to new types of presentation attacks, based on the human provided feedback.

IPC Classes  ?

  • G10L 15/26 - Speech to text systems
  • G06N 20/00 - Machine learning
  • G10L 15/02 - Feature extraction for speech recognitionSelection of recognition unit
  • G10L 15/08 - Speech classification or search
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 17/06 - Decision making techniquesPattern matching strategies
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase

69.

Call classification through analysis of DTMF events

      
Application Number 18680327
Grant Number 12621382
Status In Force
Filing Date 2024-05-31
First Publication Date 2024-09-26
Grant Date 2026-05-05
Owner Pindrop Security, Inc. (USA)
Inventor
  • Gaubitch, Nick
  • Strong, Scott
  • Cornwell, John
  • Kingravi, Hassan
  • Dewey, David

Abstract

Systems, methods, and computer-readable media for call classification and for training a model for call classification, an example method comprising: receiving DTMF information from a plurality of calls; determining, for each of the calls, a feature vector including statistics based on DTMF information such as DTMF residual signal comprising channel noise and additive noise; training a model for classification; comparing a new call feature vector to the model; predicting a device type and geographic location based on the comparison of the new call feature vector to the model; classifying the call as spoofed or genuine; and authenticating a call or altering an IVR call flow.

IPC Classes  ?

  • H04M 1/56 - Arrangements for indicating or recording the called number at the calling subscriber's set
  • H04L 25/02 - Baseband systems Details
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/493 - Interactive information services, e.g. directory enquiries
  • H04M 7/12 - Arrangements for interconnection between switching centres for working between exchanges having different types of switching equipment, e.g. power-driven and step by step or decimal and non-decimal
  • H04M 15/06 - Recording class or number of calling party or called party
  • H04Q 1/45 - Signalling arrangementsManipulation of signalling currents using AC with voice-band signalling frequencies using multi-frequency signalling
  • H04Q 3/70 - Identification of class of calling subscriber
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination

70.

PRESENTATION ATTACKS IN REVERBERANT CONDITIONS

      
Application Number 18598595
Status Pending
Filing Date 2024-03-07
First Publication Date 2024-09-19
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Gaubitch, Nikolay
  • Looney, David

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures including obtaining training audio signals having corresponding training impulse responses associated with reverberation degradation, training a machine-learning model of a presentation attack detection engine to generate one or more acoustic parameters by executing the presentation attack detection engine using the training impulse responses of the training audio signals and a loss function, obtaining an audio signal having an acoustic impulse response associated with reverberation degradation caused by one or more rooms, generating the one or more acoustic parameters for the audio signal by executing the machine-learning model using the audio signal as input, and generating an attack score for the audio signal based upon the one or more parameters generated by the machine-learning model.

IPC Classes  ?

  • G06F 21/55 - Detecting local intrusion or implementing counter-measures
  • G06N 20/00 - Machine learning
  • G10L 25/18 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
  • G10L 25/51 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination
  • G10L 25/69 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for evaluating synthetic or decoded voice signals

71.

Joint estimation of acoustic parameters from single-microphone speech

      
Application Number 17079082
Grant Number 12087319
Status In Force
Filing Date 2020-10-23
First Publication Date 2024-09-10
Grant Date 2024-09-10
Owner Pindrop Security, Inc. (USA)
Inventor
  • Looney, David
  • Gaubitch, Nikolay

Abstract

Embodiments described herein provide for end-to-end joint determination of degradation parameter scores for certain types of degradation. Degradation parameters include degradation describing additive noise and multiplicative noise such as Signal-to-Noise Ratio (SNR), reverberation time (T60), and Direct-to-Reverberant Ratio (DRR). Various neural network architectures are described such that the inherent interplay between the degradation parameters is considered in both the degradation parameter score and degradation score determination. The neural network architectures are trained according to computer generated audio datasets.

IPC Classes  ?

  • G10L 25/30 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique using neural networks
  • G06N 3/048 - Activation functions
  • G06N 3/08 - Learning methods

72.

PINDROP PULSE

      
Serial Number 98691836
Status Registered
Filing Date 2024-08-09
Registration Date 2026-05-05
Owner Pindrop Security, Inc. (USA)
NICE Classes  ? 42 - Scientific, technological and industrial services, research and design

Goods & Services

Cloud computing featuring software for use in the detection of various types of synthetic content such as voices, audio, music, or video including deepfakes, machine-based replay, and other media manipulated or generated by artificial intelligence.

73.

SYSTEMS AND METHODS FOR CALL FRAUD ANALYSIS USING A MACHINE-LEARNING ARCHITECTURE AND MAINTAINING CALLER ANI PRIVACY

      
Application Number 18413524
Status Pending
Filing Date 2024-01-16
First Publication Date 2024-08-08
Owner Pindrop Security, Inc. (USA)
Inventor Merchant, Mohammedali

Abstract

Disclosed are systems and methods including processes executed by a server that executes software routines for machine-learning architectures that receive call-invite messages containing data from a terminating carrier. The server a caller ANI and types of call data. The server further requests data from a telephony database. The server applies and executes the software programming of the machine-learning architecture on the call data (from the terminating carrier) and the portability data (from the telephony database) to generate risk scores. The server stores the data and the risk scores into a request database, until a provider server requests the risk scores in a threat assessment request. The server returns a threat assessment message to the provider server in response to the threat assessment request. The threat assessment message includes information about the caller or caller device, and the risk scores, but not the caller ANI.

IPC Classes  ?

  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/436 - Arrangements for screening incoming calls

74.

End-to-end speaker recognition using deep neural network

      
Application Number 18422523
Grant Number 12525244
Status In Force
Filing Date 2024-01-25
First Publication Date 2024-07-25
Grant Date 2026-01-13
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Garland, Matthew

Abstract

The present invention is directed to a deep neural network (DNN) having a triplet network architecture, which is suitable to perform speaker recognition. In particular, the DNN includes three feed-forward neural networks, which are trained according to a batch process utilizing a cohort set of negative training samples. After each batch of training samples is processed, the DNN may be trained according to a loss function, e.g., utilizing a cosine measure of similarity between respective samples, along with positive and negative margins, to provide a robust representation of voiceprints.

IPC Classes  ?

  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G06N 3/04 - Architecture, e.g. interconnection topology
  • G06N 3/08 - Learning methods
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/22 - Interactive proceduresMan-machine interfaces

75.

Systems and methods of speaker-independent embedding for identification and verification from audio

      
Application Number 18585366
Grant Number 12437751
Status In Force
Filing Date 2024-02-23
First Publication Date 2024-07-11
Grant Date 2025-10-07
Owner Pindrop Security, Inc. (USA)
Inventor
  • Phatak, Kedar
  • Khoury, Elie

Abstract

Embodiments described herein provide for audio processing operations that evaluate characteristics of audio signals that are independent of the speaker's voice. A neural network architecture trains and applies discriminatory neural networks tasked with modeling and classifying speaker-independent characteristics. The task-specific models generate or extract feature vectors from input audio data based on the trained embedding extraction models. The embeddings from the task-specific models are concatenated to form a deep-phoneprint vector for the input audio signal. The DP vector is a low dimensional representation of the each of the speaker-independent characteristics of the audio signal and applied in various downstream operations.

IPC Classes  ?

  • G10L 15/06 - Creation of reference templatesTraining of speech recognition systems, e.g. adaptation to the characteristics of the speaker's voice
  • G06N 3/045 - Combinations of networks
  • G06N 20/00 - Machine learning
  • G10L 15/16 - Speech classification or search using artificial neural networks
  • G10L 25/27 - Speech or voice analysis techniques not restricted to a single one of groups characterised by the analysis technique

76.

Fraud importance system

      
Application Number 18432316
Grant Number 12309316
Status In Force
Filing Date 2024-02-05
First Publication Date 2024-06-27
Grant Date 2025-05-20
Owner Pindrop Security, Inc. (USA)
Inventor
  • Phatak, Kedar
  • Raghuram, Jayaram

Abstract

Embodiments described herein provide for a fraud detection engine for detecting various types of fraud at a call center and a fraud importance engine for tailoring the fraud detection operations to relative importance of fraud events. Fraud importance engine determines which fraud events are comparative more important than others. The fraud detection engine comprises machine-learning models that consume contact data and fraud importance information for various anti-fraud processes. The fraud importance engine calculates importance scores for fraud events based on user-customized attributes, such as fraud-type or fraud activity. The fraud importance scores are used in various processes, such as model training, model selection, and selecting weights or hyper-parameters for the ML models, among others. The fraud detection engine uses the importance scores to prioritize fraud alerts for review. The fraud importance engine receives detection feedback, which contacts involved false negatives, where fraud events were undetected but should have been detected.

IPC Classes  ?

77.

Speaker recognition in the call center

      
Application Number 18436911
Grant Number 12175983
Status In Force
Filing Date 2024-02-08
First Publication Date 2024-06-27
Grant Date 2024-12-24
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Ellie
  • Garland, Matthew

Abstract

Utterances of at least two speakers in a speech signal may be distinguished and the associated speaker identified by use of diarization together with automatic speech recognition of identifying words and phrases commonly in the speech signal. The diarization process clusters turns of the conversation while recognized special form phrases and entity names identify the speakers. A trained probabilistic model deduces which entity name(s) correspond to the clusters.

IPC Classes  ?

  • G10L 17/00 - Speaker identification or verification techniques
  • G06N 7/01 - Probabilistic graphical models, e.g. probabilistic networks
  • G10L 15/07 - Adaptation to the speaker
  • G10L 15/26 - Speech to text systems
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase
  • H04M 1/27 - Devices whereby a plurality of signals may be stored simultaneously

78.

BEHAVIORAL BIOMETRICS USING KEYPRESS TEMPORAL INFORMATION

      
Application Number US2023080542
Publication Number 2024/112672
Status In Force
Filing Date 2023-11-20
Publication Date 2024-05-30
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Rao, Hrishikesh
  • Casal, Ricardo
  • Khoury, Elie
  • Lorimer, Eric
  • Cornwell, John
  • Patil, Kailash

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures including a neural network-based embedding extraction system that to produce an embedding vector representing a user's behavior's keypresses, where the system extracts the behaviorprint embedding vector using the keypress features that the system references later for authenticating users. Embodiments may extract and evaluate keypress features, such as keypress sequences, keypress pressure or volume, and temporal keypress features, such as the duration of keypresses and the interval between keypresses, among others. Some embodiments employ a deep neural network architecture that generates a behaviorprint embedding vector representation of the keypress duration and interval features that is used for enrollment and at inference time to authenticate users.

IPC Classes  ?

  • G06F 21/31 - User authentication
  • H04L 69/28 - Timers or timing mechanisms used in protocols
  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G06N 5/022 - Knowledge engineeringKnowledge acquisition

79.

BEHAVIORAL BIOMETRICS USING KEYPRESS TEMPORAL INFORMATION

      
Application Number 18515128
Status Pending
Filing Date 2023-11-20
First Publication Date 2024-05-23
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Rao, Hrishikesh
  • Casal, Ricky
  • Khoury, Elie
  • Lorimer, Eric
  • Cornwell, John
  • Patil, Kailash

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures including a neural network-based embedding extraction system that to produce an embedding vector representing a user's behavior's keypresses, where the system extracts the behaviorprint embedding vector using the keypress features that the system references later for authenticating users. Embodiments may extract and evaluate keypress features, such as keypress sequences, keypress pressure or volume, and temporal keypress features, such as the duration of keypresses and the interval between keypresses, among others. Some embodiments employ a deep neural network architecture that generates a behaviorprint embedding vector representation of the keypress duration and interval features that is used for enrollment and at inference time to authenticate users.

IPC Classes  ?

80.

Caller verification via carrier metadata

      
Application Number 18423858
Grant Number 12250344
Status In Force
Filing Date 2024-01-26
First Publication Date 2024-05-23
Grant Date 2025-03-11
Owner Pindrop Security, Inc. (USA)
Inventor
  • Cornwell, John
  • Nelms, Ii, Terry

Abstract

Embodiments described herein provide for passive caller verification and/or passive fraud risk assessments for calls to customer call centers. Systems and methods may be used in real time as a call is coming into a call center. An analytics server of an analytics service looks at the purported Caller ID of the call, as well as the unaltered carrier metadata, which the analytics server then uses to generate or retrieve one or more probability scores using one or more lookup tables and/or a machine-learning model. A probability score indicates the likelihood that information derived using the Caller ID information has occurred or should occur given the carrier metadata received with the inbound call. The one or more probability scores be used to generate a risk score for the current call that indicates the probability of the call being valid (e.g., originated from a verified caller or calling device, non-fraudulent).

IPC Classes  ?

  • H04M 3/00 - Automatic or semi-automatic exchanges
  • G06F 18/214 - Generating training patternsBootstrap methods, e.g. bagging or boosting
  • H04L 12/66 - Arrangements for connecting between networks having differing types of switching systems, e.g. gateways
  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention
  • H04M 5/00 - Manual exchanges

81.

Robust spoofing detection system using deep residual neural networks

      
Application Number 18394300
Grant Number 12417772
Status In Force
Filing Date 2023-12-22
First Publication Date 2024-05-09
Grant Date 2025-09-16
Owner Pindrop Security, Inc. (USA)
Inventor
  • Chen, Tianxiang
  • Khoury, Elie

Abstract

Embodiments described herein provide for systems and methods for implementing a neural network architecture for spoof detection in audio signals. The neural network architecture contains a layers defining embedding extractors that extract embeddings from input audio signals. Spoofprint embeddings are generated for particular system enrollees to detect attempts to spoof the enrollee's voice. Optionally, voiceprint embeddings are generated for the system enrollees to recognize the enrollee's voice. The voiceprints are extracted using features related to the enrollee's voice. The spoofprints are extracted using features related to features of how the enrollee speaks and other artifacts. The spoofprints facilitate detection of efforts to fool voice biometrics using synthesized speech (e.g., deepfakes) that spoof and emulate the enrollee's voice.

IPC Classes  ?

  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/22 - Interactive proceduresMan-machine interfaces

82.

UNSUPERVISED KEYWORD SPOTTING AND WORD DISCOVERY FOR FRAUD ANALYTICS

      
Application Number 18385632
Status Pending
Filing Date 2023-10-31
First Publication Date 2024-02-22
Owner PINDROP SECURITY, INC. (USA)
Inventor Rao, Hrishikesh

Abstract

Embodiments described herein provide for a computer that detects one or more keywords of interest using acoustic features, to detect or query commonalities across multiple fraud calls. Embodiments described herein may implement unsupervised keyword spotting (UKWS) or unsupervised word discovery (UWD) in order to identify commonalities across a set of calls, where both UKWS and UWD employ Gaussian Mixture Models (GMM) and one or more dynamic time-warping algorithms. A user may indicate a training exemplar or occurrence of call-specific information, referred to herein as “a named entity,” such as a person's name, an account number, account balance, or order number. The computer may perform a redaction process that computationally nullifies the import of the named entity in the modeling processes described herein.

IPC Classes  ?

  • G10L 15/197 - Probabilistic grammars, e.g. word n-grams
  • G10L 15/04 - SegmentationWord boundary detection
  • G10L 15/30 - Distributed recognition, e.g. in client-server systems, for mobile phones or network applications
  • G10L 15/22 - Procedures used during a speech recognition process, e.g. man-machine dialog

83.

Omni channel authentication

      
Application Number 18235321
Grant Number 12489760
Status In Force
Filing Date 2023-08-17
First Publication Date 2024-02-22
Grant Date 2025-12-02
Owner Pindrop Security, Inc. (USA)
Inventor
  • Merchant, Mohammedali
  • Gupta, Payas

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures providing improved omni-channel authentication solutions. Embodiments include one or more computing devices that provide an authentication interface by which various communication channels may deposit contact or session data received via a first-channel session into a non-transitory storage medium of an authentication database for another channel to obtain and employ (e.g., verify users). This allows the customer to access an online data channel and enter the contact center through a telephony communication channel, but further allows the enterprise contact center systems to passively maintain access to various types of information about the user's identity captured from each contact channel, allowing the call center to request or capture authenticating information (e.g., voice biometrics) from both channels to employ authentication processes for one or both channels, such as voice biometrics authentication processes or other types of authentication functions.

IPC Classes  ?

84.

OMNI CHANNEL AUTHENTICATION

      
Document Number 03209649
Status Pending
Filing Date 2023-08-17
Open to Public Date 2024-02-18
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Merchant, Mohammedali
  • Gupta, Payas

Abstract

Embodiments include a computing device that executes software routines and/or one or more machine-learning architectures providing improved omni-channel authentication solutions. Embodiments include one or more computing devices that provide an authentication interface by which various communication channels may deposit contact or session data received via a first-channel session into a non-transitory storage medium of an authentication database for another channel to obtain and employ (e.g., verify users). This allows the customer to access an online data channel and enter the contact center through a telephony communication channel, but further allows the enterprise contact center systems to passively maintain access to various types of information about the user's identity captured from each contact channel, allowing the call center to request or capture authenticating information (e.g., voice biometrics) from both channels to employ authentication processes for one or both channels, such as voice biometrics authentication processes or other types of authentication functions.

IPC Classes  ?

  • G06F 21/32 - User authentication using biometric data, e.g. fingerprints, iris scans or voiceprints
  • G06N 20/00 - Machine learning
  • G06Q 30/015 - Providing customer assistance, e.g. assisting a customer within a business location or via helpdesk
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • H04L 9/32 - Arrangements for secret or secure communicationsNetwork security protocols including means for verifying the identity or authority of a user of the system

85.

Carrier signaling based authentication and fraud detection

      
Application Number 18221802
Grant Number 12432299
Status In Force
Filing Date 2023-07-13
First Publication Date 2024-01-18
Grant Date 2025-09-30
Owner Pindrop Security, Inc. (USA)
Inventor
  • Casal, Ricky
  • Maddali, Vinay
  • Gupta, Payas
  • Patil, Kailash

Abstract

Disclosed are systems and methods including computing-processes, which may include layers of machine-learning architectures, for assessing risk for calls directed to call center systems using carrier signaling metadata. A computer evaluates carrier signaling metadata to perform various new risk-scoring techniques to determine riskiness of calls and authenticate calls. When determining a risk score for an incoming call is received at a call center system, the computer may obtain certain metadata values from inbound metadata, prior call metadata, or from third-party telecommunications services and executes processes for determining the risk score for the call. The risk score operations include several scoring components, including appliance print scoring, carrier detection scoring, ANI location detection scoring, location similarity scoring, and JIP-ANI location similarity scoring, among others.

IPC Classes  ?

  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

86.

CARRIER SIGNALING BASED AUTHENTICATION AND FRAUD DETECTION

      
Document Number 03206586
Status Pending
Filing Date 2023-07-13
Open to Public Date 2024-01-14
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Casal, Ricky
  • Maddali, Vinay
  • Gupta, Payas
  • Patil, Kallash

Abstract

Disclosed are systems and methods including computing-processes, which may include layers of machine-learning architectures, for assessing risk for calls directed to call center systems using carrier signaling metadata. A computer evaluates carrier signaling metadata to perform various new risk-scoring techniques to determine riskiness of calls and authenticate calls. When determining a risk score for an incoming call is received at a call center system, the computer may obtain certain metadata values from inbound metadata, prior call metadata, or from third-party telecommunications services and executes processes for determining the risk score for the call. The risk score operations include several scoring components, including appliance print scoring, carrier detection scoring, ANI location detection scoring, location similarity scoring, and JIP-ANI location similarity scoring, among others.

IPC Classes  ?

  • H04M 3/436 - Arrangements for screening incoming calls
  • H04M 3/493 - Interactive information services, e.g. directory enquiries

87.

Speaker recognition in the call center

      
Application Number 18329138
Grant Number 12711960
Status In Force
Filing Date 2023-06-05
First Publication Date 2023-10-12
Grant Date 2026-08-18
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Garland, Matthew

Abstract

Utterances of at least two speakers in a speech signal may be distinguished and the associated speaker identified by use of diarization together with automatic speech recognition of identifying words and phrases commonly in the speech signal. The diarization process clusters turns of the conversation while recognized special form phrases and entity names identify the speakers. A trained probabilistic model deduces which entity name(s) correspond to the clusters.

IPC Classes  ?

  • G10L 17/00 - Speaker identification or verification techniques
  • G06N 7/01 - Probabilistic graphical models, e.g. probabilistic networks
  • G10L 15/07 - Adaptation to the speaker
  • G10L 15/19 - Grammatical context, e.g. disambiguation of recognition hypotheses based on word sequence rules
  • G10L 15/26 - Speech to text systems
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/08 - Use of distortion metrics or a particular distance between probe pattern and reference templates
  • G10L 17/24 - the user being prompted to utter a password or a predefined phrase
  • H04M 1/27 - Devices whereby a plurality of signals may be stored simultaneously

88.

Channel-compensated low-level features for speaker recognition

      
Application Number 18321353
Grant Number 12354608
Status In Force
Filing Date 2023-05-22
First Publication Date 2023-09-14
Grant Date 2025-07-08
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Garland, Matthew

Abstract

A system for generating channel-compensated features of a speech signal includes a channel noise simulator that degrades the speech signal, a feed forward convolutional neural network (CNN) that generates channel-compensated features of the degraded speech signal, and a loss function that computes a difference between the channel-compensated features and handcrafted features for the same raw speech signal. Each loss result may be used to update connection weights of the CNN until a predetermined threshold loss is satisfied, and the CNN may be used as a front-end for a deep neural network (DNN) for speaker recognition/verification. The DNN may include convolutional layers, a bottleneck features layer, multiple fully-connected layers, and an output layer. The bottleneck features may be used to update connection weights of the convolutional layers, and dropout may be applied to the convolutional layers.

IPC Classes  ?

  • G10L 17/20 - Pattern transformations or operations aimed at increasing system robustness, e.g. against channel noise or different working conditions
  • G10L 17/02 - Preprocessing operations, e.g. segment selectionPattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal componentsFeature selection or extraction
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/18 - Artificial neural networksConnectionist approaches
  • G10L 19/028 - Noise substitution, e.g. substituting non-tonal spectral components by noisy source

89.

Authentication using DTMF tones

      
Application Number 18317799
Grant Number 12256040
Status In Force
Filing Date 2023-05-15
First Publication Date 2023-09-07
Grant Date 2025-03-18
Owner Pindrop Security, Inc. (USA)
Inventor Gupta, Payas

Abstract

A method of obtaining and automatically providing secure authentication information includes registering a client device over a data line, storing information and a changeable value for authentication in subsequent telephone-only transactions. In the subsequent transactions, a telephone call placed from the client device to an interactive voice response server is intercepted and modified to include dialing of a delay and at least a passcode, the passcode being based on the unique information and the changeable value, where the changeable value is updated for every call session. The interactive voice response server forwards the passcode and a client device identifier to an authentication function, which compares the received passcode to plural passcodes generated based on information and iterations of a value stored in correspondence with the client device identifier. Authentication is confirmed when a generated passcode matches the passcode from the client device.

IPC Classes  ?

  • H04M 3/38 - Graded-service arrangements, i.e. some subscribers prevented from establishing certain connections
  • H04L 9/40 - Network security protocols
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/493 - Interactive information services, e.g. directory enquiries
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

90.

Systems and methods for stir-shaken attestation using SPID

      
Application Number 18117134
Grant Number 12413975
Status In Force
Filing Date 2023-03-03
First Publication Date 2023-09-07
Grant Date 2025-09-09
Owner Pindrop Security, Inc. (USA)
Inventor
  • Merchant, Mohammedali
  • Sun, Yitao

Abstract

Embodiments described herein provide for evaluating call metadata and certificates of inbound calls for authentication. The computer identifies a service provider indicated by the SPID and/or the ANI (or other identifier) of the metadata and identifies a service provider indicated by the SPID and/or ANI (or other identifier) of the certificate, then compares identities of the service providers and/or compares the data values associated with the service providers (e.g., SPIDs, ANIs). Based on this comparison, the computer determines whether the service provider that signed the certificate is first-party signer (e.g., carrier) for the ANI or a third-party signer that is signing certificates as the first-party signer for the ANI.

IPC Classes  ?

91.

SYSTEMS AND METHODS FOR STIR-SHAKEN ATTESTATION USING SPID

      
Document Number 03191780
Status Pending
Filing Date 2023-03-03
Open to Public Date 2023-09-04
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Merchant, Mohammedali
  • Sun, Yitao

Abstract


Embodiments described herein provide for evaluating call metadata and
certificates of inbound
calls for authentication. The computer identifies a service provider indicated
by the SPID
and/or the ANI (or other identifier) of the metadata and identifies a service
provider indicated
by the SPID and/or ANI (or other identifier) of the certificate, then compares
identities of the
service providers and/or compares the data values associated with the service
providers
(e.g., SPIDs, ANIs). Based on this comparison, the computer determines whether
the service
provider that signed the certificate is first-party signer (e.g., carrier) for
the ANI or a third-party
signer that is signing certificates as the first-party signer for the ANI.

IPC Classes  ?

  • H04M 3/436 - Arrangements for screening incoming calls
  • H04W 12/069 - Authentication using certificates or pre-shared keys
  • H04W 12/40 - Security arrangements using identity modules
  • H04W 12/72 - Subscriber identity

92.

Systems and methods for JIP and CLLI matching in telecommunications metadata and machine-learning

      
Application Number 18109095
Grant Number 12701184
Status In Force
Filing Date 2023-02-13
First Publication Date 2023-08-17
Grant Date 2026-08-04
Owner Pindrop Security, Inc. (USA)
Inventor
  • Merchant, Mohammed Ali
  • Sun, Yitao

Abstract

Embodiments described herein provide for systems and methods for verifying authentic JIPs associated with ANIs using CLLIs known to be associated with the ANIs, allowing a computer to authenticate calls using the verified JIPs, among various factors. The computer builds a trust model for JIPs by correlating unique CLLIs to JIPs. A malicious actor might spoof numerous ANIs mapped to a single CLLI, but the malicious actor is unlikely to spoof multiple CLLIs due to the complexity of spoofing the volumes of ANIs associated with multiple CLLIs, so the CLLIs can be trusted when determining whether a JIP is authentic. The computer identifies an authentic JIP when the trust model indicates that a number of CLLIs associated with the JIP satisfies one or more thresholds. A machine-learning architecture references the fact that the JIP is authentic as an authentication factor for downstream call authentication functions.

IPC Classes  ?

  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/22 - Arrangements for supervision, monitoring or testing

93.

SYSTEMS AND METHODS FOR JIP AND CLLI MATCHING IN TELECOMMUNICATIONS METADATA AND MACHINE-LEARNING

      
Document Number 03189600
Status Pending
Filing Date 2023-02-14
Open to Public Date 2023-08-15
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Merchant, Mohammedali
  • Sun, Yitao

Abstract

Embodiments described herein provide for systems and methods for verifying authentic JIPs associated with ANIs using CLLIs known to be associated with the ANIs, allowing a computer to authenticate calls using the verified JIPs, among various factors. The computer builds a trust model for JIPs by correlating unique CLLIs to JIPs. A malicious actor might spoof numerous ANIs mapped to a single CLLI, but the malicious actor is unlikely to spoof multiple CLLIs due to the complexity of spoofing the volumes of ANIs associated with multiple CLLIs, so the CLLIs can be trusted when determining whether a JIP is authentic. The computer identifies an authentic JIP when the trust model indicates that a number of CLLIs associated with the JIP satisfies one or more thresholds. A machine-learning architecture references the fact that the JIP is authentic as an authentication factor for downstream call authentication functions.

IPC Classes  ?

  • H04M 1/66 - Substation equipment, e.g. for use by subscribers with means for preventing unauthorised or fraudulent calling
  • H04M 3/436 - Arrangements for screening incoming calls

94.

Systems and methods employing graph-derived features for fraud detection

      
Application Number 18301897
Grant Number 12022024
Status In Force
Filing Date 2023-04-17
First Publication Date 2023-08-10
Grant Date 2024-06-25
Owner Pindrop Security, Inc. (USA)
Inventor
  • Casal, Ricardo
  • Walker, Theo
  • Patil, Kailash
  • Cornwell, John

Abstract

Embodiments described herein provide for performing a risk assessment using graph-derived features of a user interaction. A computer receives interaction information and infers information from the interaction based on information provided to the computer by a communication channel used in transmitting the interaction information. The computer may determine a claimed identity of the user associated with the user interaction. The computer may extract features from the inferred identity and claimed identity. The computer generates a graph representing the structural relationship between the communication channels and claimed identities associated with the inferred identity and claimed identity. The computer may extract additional features from the inferred identity and claimed identity using the graph. The computer may apply the features to a machine learning model to generate a risk score indicating the probability of a fraudulent interaction associated with the user interaction.

IPC Classes  ?

  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • G06F 18/214 - Generating training patternsBootstrap methods, e.g. bagging or boosting
  • G06N 20/00 - Machine learning
  • G06N 3/08 - Learning methods
  • G06N 20/10 - Machine learning using kernel methods, e.g. support vector machines [SVM]
  • G06N 20/20 - Ensemble learning
  • H04M 3/42 - Systems providing special services or facilities to subscribers
  • H04M 3/51 - Centralised call answering arrangements requiring operator intervention

95.

Telecommunications validation system and method

      
Application Number 18123463
Grant Number 12120263
Status In Force
Filing Date 2023-03-20
First Publication Date 2023-08-03
Grant Date 2024-10-15
Owner Pindrop Security, Inc. (USA)
Inventor
  • Merchant, Mohammedali
  • Williams, Matthew
  • Prugar, Tim

Abstract

According to an embodiment of the disclosure, a toll-free telecommunications validation system determines a confidence value that an incoming phone call to an enterprises' toll-free number is originating from the station it purports to be, i.e., is not a spoofed call by incorporating one or more layers of signals and data in determining said confidence value, the data and signals including, but not limited to, toll-free call routing logs, service control point (SCP) signals and data, service data point (SDP) signals and data, dialed number information service (DNIS) signals and data, automatic number identification (ANI) signals and data, session initiation protocol (SIP) signals and data, carrier identification code (CIC) signals and data, location routing number (LRN) signals and data, jurisdiction information parameter (JIP) signals and data, charge number (CN) signals and data, billing number (BN) signals and data, and originating carrier information (such as information derived from the ANI, including, but not limited to, alternative service provider ID (ALTSPID), service provider ID (SPID), or operating company number (OCN)). In certain configurations said enterprise provides an ANI and DNIS associated with said incoming toll-free call, which is used to query a commercial toll-free telecommunications routing platform for any corresponding log entries. The existence of any such log entries, along with the originating carrier information in the event log entries do exist, is used to determine a confidence value that said incoming toll-free call is originating from the station it purports to be. As a result, said entities or enterprises operating a toll-free number may be provided a confidence value regarding an incoming telephone call, and using that confidence value, further determine whether or not to accept the authenticity of the incoming telephone call and/or based on said confidence value, service the incoming call differently.

IPC Classes  ?

  • H04M 3/22 - Arrangements for supervision, monitoring or testing
  • H04M 3/42 - Systems providing special services or facilities to subscribers

96.

CROSS-LINGUAL SPEAKER RECOGNITION

      
Document Number 03236335
Status Pending
Filing Date 2022-10-31
Open to Public Date 2023-05-04
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Chen, Tianxiang
  • Kumar, Avrosh
  • Sivaraman, Ganesh
  • Phatak, Kedar

Abstract

Disclosed are systems and methods including computing-processes executing machine-learning architectures for voice biometrics, in which the machine-learning architecture implements one or more language compensation functions. Embodiments include an embedding extraction engine (sometimes referred to as an "embedding extractor") that extracts speaker embeddings and determines a speaker similarity score for determine or verifying the likelihood that speakers in different audio signals are the same speaker. The machine-learning architecture further includes a multi-class language classifier that determines a language likelihood score that indicates the likelihood that a particular audio signal includes a spoken language. The features and functions of the machine-learning architecture described herein may implement the various language compensation techniques to provide more accurate speaker recognition results, regardless of the language spoken by the speaker.

IPC Classes  ?

  • G06F 40/263 - Language identification
  • G06N 3/02 - Neural networks
  • G06N 3/045 - Combinations of networks
  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/10 - Multimodal systems, i.e. based on the integration of multiple recognition engines or fusion of expert systems
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices

97.

CROSS-LINGUAL SPEAKER RECOGNITION

      
Application Number US2022048365
Publication Number 2023/076653
Status In Force
Filing Date 2022-10-31
Publication Date 2023-05-04
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Khoury, Elie
  • Chen, Tianxiang
  • Kumar, Avrosh
  • Sivaraman, Ganesh
  • Phatak, Kedar

Abstract

Disclosed are systems and methods including computing-processes executing machine-learning architectures for voice biometrics, in which the machine-learning architecture implements one or more language compensation functions. Embodiments include an embedding extraction engine (sometimes referred to as an "embedding extractor") that extracts speaker embeddings and determines a speaker similarity score for determine or verifying the likelihood that speakers in different audio signals are the same speaker. The machine-learning architecture further includes a multi-class language classifier that determines a language likelihood score that indicates the likelihood that a particular audio signal includes a spoken language. The features and functions of the machine-learning architecture described herein may implement the various language compensation techniques to provide more accurate speaker recognition results, regardless of the language spoken by the speaker.

IPC Classes  ?

  • G10L 25/60 - Speech or voice analysis techniques not restricted to a single one of groups specially adapted for particular use for comparison or discrimination for measuring the quality of voice signals
  • G06F 40/263 - Language identification
  • G06N 3/02 - Neural networks
  • G06N 3/045 - Combinations of networks
  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 17/10 - Multimodal systems, i.e. based on the integration of multiple recognition engines or fusion of expert systems
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices

98.

Cross-lingual speaker recognition

      
Application Number 17977521
Grant Number 12451138
Status In Force
Filing Date 2022-10-31
First Publication Date 2023-05-04
Grant Date 2025-10-21
Owner Pindrop Security, Inc. (USA)
Inventor
  • Khoury, Elie
  • Chen, Tianxiang
  • Kumar, Avrosh
  • Sivaraman, Ganesh
  • Phatak, Kedar

Abstract

Disclosed are systems and methods including computing-processes executing machine-learning architectures for voice biometrics, in which the machine-learning architecture implements one or more language compensation functions. Embodiments include an embedding extraction engine (sometimes referred to as an “embedding extractor”) that extracts speaker embeddings and determines a speaker similarity score for determine or verifying the likelihood that speakers in different audio signals are the same speaker. The machine-learning architecture further includes a multi-class language classifier that determines a language likelihood score that indicates the likelihood that a particular audio signal includes a spoken language. The features and functions of the machine-learning architecture described herein may implement the various language compensation techniques to provide more accurate speaker recognition results, regardless of the language spoken by the speaker.

IPC Classes  ?

  • G10L 15/00 - Speech recognition
  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 17/04 - Training, enrolment or model building
  • G10L 17/10 - Multimodal systems, i.e. based on the integration of multiple recognition engines or fusion of expert systems

99.

AGE ESTIMATION FROM SPEECH

      
Document Number 03232220
Status Pending
Filing Date 2022-10-05
Open to Public Date 2023-04-13
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Saraf, Amruta
  • Khoury, Elie
  • Sivaraman, Ganesh

Abstract

Disclosed are systems and methods including computing-processes executing machine- learning architectures implementing label distribution loss functions to improve age estimation performance and generalization. The machine-learning architecture includes a front-end neural network architecture defining a speaker embedding extraction engine of the machine-learning architecture, and a backend neural network architecture defining an age estimation engine of the machine-learning architecture. The embedding extractor is trained to extract low-level acoustic features of a speaker's speech, such as mel-frequency cepstral coefficients (MFCCs), from audio signals, and then extract a feature vector or speaker embedding vector that mathematically represents the low-level features of the speaker. The age estimator is trained to generate an estimated age for the speaker and a Gaussian probability distribution around the estimated age, by applying the various types of layers of the age estimator on the speaker embedding.

IPC Classes  ?

  • G06N 20/00 - Machine learning
  • G10L 15/26 - Speech to text systems
  • G10L 17/00 - Speaker identification or verification techniques
  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices

100.

AGE ESTIMATION FROM SPEECH

      
Application Number US2022045777
Publication Number 2023/059717
Status In Force
Filing Date 2022-10-05
Publication Date 2023-04-13
Owner PINDROP SECURITY, INC. (USA)
Inventor
  • Saraf, Amruta
  • Khoury, Elie
  • Sivaraman, Ganesh

Abstract

Disclosed are systems and methods including computing-processes executing machine- learning architectures implementing label distribution loss functions to improve age estimation performance and generalization. The machine-learning architecture includes a front-end neural network architecture defining a speaker embedding extraction engine of the machine-learning architecture, and a backend neural network architecture defining an age estimation engine of the machine-learning architecture. The embedding extractor is trained to extract low-level acoustic features of a speaker's speech, such as mel-frequency cepstral coefficients (MFCCs), from audio signals, and then extract a feature vector or speaker embedding vector that mathematically represents the low-level features of the speaker. The age estimator is trained to generate an estimated age for the speaker and a Gaussian probability distribution around the estimated age, by applying the various types of layers of the age estimator on the speaker embedding.

IPC Classes  ?

  • G10L 17/26 - Recognition of special voice characteristics, e.g. for use in lie detectorsRecognition of animal voices
  • G10L 15/26 - Speech to text systems
  • G10L 17/00 - Speaker identification or verification techniques
  • G06N 20/00 - Machine learning
  1     2     3        Next Page