San Francisco Startup Builds AI to Detect Deepfake Voices

San Francisco Startup Builds AI to Detect Deepfake Voices
Photo Credit: Unsplash.com

San Francisco AI startup DetectifAI is developing technology designed to identify deepfake and cloned voices during calls and other audio interactions. Founder Tarini Padmanabhuni started the company after a voice-cloning scam targeted her grandfather, leading the startup to focus on detecting synthetic speech directly on smartphones.

Key Takeaways

  • DetectifAI is developing technology to identify AI-generated and cloned voices
  • Founder Tarini Padmanabhuni launched the company after a voice-cloning scam involving her grandfather
  • The startup’s technology is designed to operate directly on smartphones
  • DetectifAI is targeting phone calls, voice messages and other audio interactions
  • Its approach is designed to detect synthetic speech without requiring audio to be processed entirely through the cloud

DetectifAI Develops On-Device AI for Deepfake Voice Detection

DetectifAI is developing artificial intelligence technology intended to distinguish synthetic voices from human speech during audio interactions. The San Francisco startup was founded by Tarini Padmanabhuni, who began developing the company after a voice-cloning scam involving her grandfather.

The company’s focus is deepfake voice detection. Its technology is designed to analyze speech and identify signs that an audio sample was generated or altered using artificial intelligence.

DetectifAI is also developing its system for on-device use. Rather than depending entirely on cloud-based processing, the company’s approach is designed to allow detection technology to operate directly on a smartphone.

That approach places the detection process on the device handling the audio interaction. The technology is intended for situations involving phone calls, voice messages and other forms of spoken communication.

The company’s development addresses a specific problem created by synthetic voices. A person can use voice-cloning technology to produce speech that resembles another individual, creating a challenge for people attempting to determine whether they are communicating with the person they expect.

DetectifAI’s product focus is therefore centered on identifying the audio itself rather than relying only on the identity presented by a caller or message sender.

Founder Launches Startup After Family Voice-Cloning Scam

Padmanabhuni’s decision to start DetectifAI followed a voice-cloning scam involving her grandfather. The incident became the basis for the company’s focus on identifying synthetic voices.

In the scam, a cloned voice was used to make Padmanabhuni’s grandfather believe he was speaking with his brother. The incident showed the practical difficulty of distinguishing a familiar person’s voice from an AI-generated imitation.

The scam involved a false claim that the grandfather’s brother had been kidnapped. The deception led to a ransom payment before the situation was identified as fraudulent. Similar phone impersonation scams can rely on urgency and familiar details to pressure recipients before they have time to verify a caller’s claims.

Padmanabhuni subsequently founded DetectifAI to work on technology intended to identify this type of synthetic speech. The company’s stated focus connects the founding experience directly to the technical problem it is attempting to address.

The family incident provides the immediate context for the startup’s product development. DetectifAI is working on technology intended for use during ordinary audio communications.

The startup’s work centers on detection rather than voice generation. Its technology is intended to identify whether speech is synthetic, providing a technical layer for evaluating audio received through calls and messages.

Technology Targets Synthetic Speech During Audio Interactions

DetectifAI is designing its technology for several types of spoken communication, including phone calls and voice messages. The system is intended to analyze speech during those interactions and identify whether the voice may have been generated through AI.

The focus on audio interactions gives the technology a specific use case. A deepfake voice can be presented through a familiar communication channel, meaning the recipient may initially treat the voice as evidence of the speaker’s identity.

Voice-cloning scams can exploit that assumption by reproducing characteristics associated with a real person’s speech. DetectifAI’s technology is being developed to provide a separate method for assessing the authenticity of the voice.

The challenge extends beyond voice cloning as anti-deepfake detection tools are also being developed to assess synthetic media and manipulated identity signals.

The startup’s approach centers on the technical detection of synthetic speech rather than on identifying a particular scam scenario. The same detection capability can be applied to audio in different communication settings when the system is able to analyze the voice.

Phone calls are one of the situations targeted by the technology. Voice messages are another. Both involve audio that can potentially be separated from other identifying information and evaluated for characteristics associated with synthetic generation.

The company’s product development therefore focuses on a defined technical task: determining whether spoken audio is genuine or AI-generated.

That task differs from conventional caller identification. Caller identification can provide information about a telephone number or sender, while deepfake voice detection examines the characteristics of the speech itself.

DetectifAI’s development combines artificial intelligence for analyzing speech with smartphone-based processing for conducting that analysis on the device.

On-Device Processing Forms Part of DetectifAI’s Approach

DetectifAI’s technology is designed to operate directly on smartphones rather than requiring all audio to be sent to a remote cloud system for analysis. On-device AI places the processing capability on the device where the communication occurs.

The approach is relevant to the startup’s voice-detection objective because audio can be analyzed within the same device used for a call or message. The company’s technology is being developed around that local-processing model.

Cloud-based systems and on-device systems handle AI processing differently. A cloud-dependent system sends information to remote computing infrastructure, while an on-device system performs at least part of the computation locally.

DetectifAI is applying the latter approach to voice authentication and deepfake detection. The company is seeking to determine whether speech is synthetic while keeping the detection process tied to the user’s smartphone.

The startup’s focus on local processing is part of its technical design rather than a separate product category. The central purpose remains the identification of AI-generated voices.

The technology is intended for audio interactions in which a recipient needs to determine whether the person speaking is genuine. This includes communications where a cloned voice is presented as someone the recipient already knows.

The technical distinction remains important for DetectifAI because its proposed system is designed around local voice analysis rather than a general-purpose cloud AI platform.

San Francisco Startup Focuses on Voice Authentication Technology

DetectifAI’s work places voice authentication at the center of its San Francisco startup operation. The company is developing technology intended to help distinguish genuine speech from cloned or AI-generated voices.

Padmanabhuni’s founding experience provides the direct connection between the company’s creation and the problem it is addressing. Her grandfather’s experience with a cloned voice led to the development of a startup focused specifically on synthetic-speech detection.

The company is focused on detecting rather than producing cloned voices. Its stated product direction includes identifying AI-generated speech used during phone calls and voice messages.

Its on-device approach also distinguishes the technology’s intended operation. DetectifAI is developing a system that can perform voice analysis directly through a smartphone instead of relying entirely on remote processing.

The company’s work addresses a specific challenge for voice communications: a familiar-sounding voice is no longer necessarily sufficient evidence of a speaker’s identity when AI tools can reproduce a person’s speech.

DetectifAI’s technology is intended to provide an additional method for assessing the authenticity of spoken communication. The system’s role is to analyze the voice and identify whether it shows characteristics associated with synthetic generation.

For DetectifAI, the product objective remains focused on identifying whether spoken audio is synthetic. The startup’s development is rooted in the family incident that prompted Padmanabhuni to create the company and is centered on detecting deepfake voices across common audio communication formats.

Frequently Asked Questions

What is DetectifAI?

DetectifAI is a San Francisco startup developing artificial intelligence technology for detecting deepfake and cloned voices. Its technology is designed for audio interactions including phone calls and voice messages.

What does DetectifAI’s deepfake voice technology do?

The technology is designed to analyze spoken audio and identify whether a voice may have been generated or altered using artificial intelligence. The company’s focus is detecting synthetic speech rather than generating cloned voices.

Who founded the San Francisco AI startup?

Tarini Padmanabhuni founded DetectifAI after a voice-cloning scam involving her grandfather. The incident led her to focus the company’s work on detecting synthetic voices.

How does on-device deepfake voice detection work?

DetectifAI is developing technology designed to perform voice analysis directly on smartphones. Its approach does not require all audio processing to take place through a remote cloud system.

Why are deepfake voice scams difficult to identify?

Deepfake voice scams can use AI-generated speech that resembles a familiar person’s voice. DetectifAI is developing technology intended to provide a technical method for assessing whether spoken audio is synthetic.

San Francisco Post

Chronicles of the Bay Area’s heartbeat.