AI

Voice.ai raises $6M as its real-time voice changer approaches 500K users

Comment

grapic depiction of white soundwaves on pinkish background
Image Credits: Bryce Durbin/TechCrunch

Services like Midjourney and ChatGPT have pushed the boundaries of how AI can create images and text out of basic text prompts. Now, audio appears to be the inevitable next frontier. Music generation based on word prompts, AI tutors for language learning and voice simulators have all seen developments in recent months. Voice.ai hopes to be a part of that conversation (heh) with technology that lets users change (and disguise) their voices in real time, and now it has raised its first outside funding on the heels of early growth.

With more than 480,000 users and a library of more than 50,000 voice filters, Voice.ai has picked up $6 million, funding that it plans to use to take its voice changing tech into new places.

Mucker Capital and M13 are leading the round. Before now, Voice.ai has grown by word of mouth — the startup has a Discord channel with more than 120,000 people — on the back of $3 million in self-funding.

Currently the company’s tools — available as apps for Mac, PC, Android and iOS — are getting adopted by gamers, content creators, Vtubers and others on TikTok, Zoom, Discord, Minecraft, GTA5, Fortnite, Valorant, League of Legends, Among Us, Skype, WhatsApp and other platforms. The Voice.ai interface lets them create a new voice, or select from some 50,000 different pre-created voices (created and shared by users like themselves), which can be used as-is or modified, to use live in supported platforms, or for recordings.

The plan is to use the funding to hire more technical talent and to build new SDKs and APIs to work with further platforms like Meta, Unreal and Unity; bring on multi-language support; and add in new applications like singing where voice is center stage.

The startup doesn’t single it out, but it will be interesting to see if it uses some of the funding also to increase server capacity.

That is no small burden. Anecdotally, we’ve heard that GPU pain is one of the biggest gating factors in how a lot of AI apps are able to scale at the moment. (It’s partly why you’re seeing big deals being made that include strategics providing processing and server capacity.)

For Voice.ai specifically, your voice is processed locally and channeled into wherever it will be used through what founder and CEO Heath Ahrens described to me as a “virtual audio cable.” But when you look at reviews of its apps, a common lament is that when you sign up you are put on a waitlist because “overwhelming demand has our servers at max capacity” with a promise that you’ll be informed when the service increases that capacity.

There are dozens of speech-to-voice and voice-to-speech services in the market today, and already a lot of activity among them: Last year Spotify acquired Sonantic and Snap bought an AI voice assistant even earlier than that; another startup, Sanas, is working on changing your accent and there are the voice simulators Murf and Acapela, among many others. Voice.ai counts itself in the same general category as Respeecher and ElevenLabs, two voice-to-voice AI startups, letting users apply masks to tweak or completely transform their voices — in some cases creating completely synthetic voices in place of the real thing.

Respeecher, founded and based in Ukraine, made a name for itself by helping build a new Darth Vader voice for new Star Wars installments, based on how James Earl Jones sounded 45 years ago when he originated the role. (In keeping with a character hell-bent on destroying worlds, Darth’s voice was delivered to the Hollywood client from its offices in Ukraine as Russia marched into the country.)

ElevenLabs — famously (or infamously as the case may be) — has built a platform that is frighteningly good at cloning voices, and earlier this month it picked up its most recent funding round of $19 million from a group of big-name investors.

Voice.ai is trying, in that mix, to position itself as the AI voice modifying app for Everyman.

“There are plenty of companies that are trying to provide a different flavor of voice tech to businesses,” Ahrens told TechCrunch in an email (ironically, it wasn’t possible to arrange a live interview with him). Ahrens has some experience with the building of B2B AI tech: his two previous companies — iSpeech for text-to-speech and Haystack for face recognition — are built around API offerings.

“What sets Voice.ai apart is that we are focused on bringing tech that was previously reserved for enterprise companies directly into the hands of consumers in an affordable fashion.” Many users, he noted, “come to us from classical DSP voice changers and voice modulators which they had been using in the past and which are still popular among many gamers and streamers.”

“Affordable” comes in two tiers, with most users now on a free service that requires them to opt in to providing computational power to train Voice.ai’s models, with its service built on its own private data set comprised of “millions of unique users.” No pricing is provided on the site: we’re asking for those details.

“We believe in making technology accessible and plan on working together with the open source community to democratize Voice AI technology,” added Ahrens.

Voice.ai also claims it takes what is a fundamentally different approach to the challenge of changing a voice, tapping into some of the ethos that has built up around the use of avatars by Vtubers, gamers and others online.

“Most voice AI companies that are coming into the space try to build scalable enterprise focused text-to-speech solutions or expensive voice-to-voice services for production studios,” Ahrens said. “We start from the opposite spectrum and try to deliver value to individuals who are looking to expand how they sound online. The core value proposition of our speech-to-speech AI isn’t that it can perfectly replicate any given person. It’s that it retains the core elements of a user’s speech: their emotion, pacing and emphasis while replacing the sound of the voice, in order to create a completely unique new end result, in real-time.”

It might be because of how the demographics in interactive platforms like gaming skew, but for now Voice.ai’s audience is 70% male versus 30% female with new categories opening not just around who is using the tech, but why.

That includes not just those using avatars and building voices to match them, or those looking for more privacy protection, but also, he said, “transgender users who can represent themselves with voices that match their identity, as well as users exploring completely new online personas for themselves.”

There is already a base of users tapping into Voice.ai’s direct-to-consumer offerings, but one of the reasons why Mucker is investing in the startup is because it believes that there is an opportunity to build out a network of developers using and integrating its tech.

“Voice.ai is poised to revolutionize the AI developer community in a manner akin to AdMob’s impact on the mobile app developer community,” said Omar Hamoui, a partner at lead investor Mucker Capital. (Hamoui previously founded the mobile ad startup AdMob, eventually acquired by Google, so he has some direct experience building mobile developer tools.) “By offering user-friendly solutions that were once exclusive to large enterprises, Voice.ai aims to democratize access for developers worldwide.”

Karl Alomar, the former COO of Digital Ocean, who led the investment for M13, said investors will be taking an active role in the next stage of development. “At Digital Ocean too we saw the value of building a community of builders by builders,” he said. “We’re excited for creators and developers to build on the Voice.ai platform.”

More TechCrunch

Reddit announced on Wednesday that it is reintroducing its awards system after shutting down the program last year. The company said that most of the mechanisms related to awards will…

Reddit reintroduces its awards system

Sigma Computing, a startup building a range of data analytics and business intelligence tools, has raised $200 million in a fresh VC round.

Sigma is building a suite of collaborative data analytics tools

European Union enforcers of the bloc’s online governance regime, the Digital Services Act (DSA), said Thursday they’re closely monitoring disinformation campaigns on the Elon Musk-owned social network X (formerly Twitter)…

EU ‘closely’ monitoring X in wake of Fico shooting as DSA disinfo probe rumbles on

Wind is the largest source of renewable energy in the U.S., according to the U.S. Energy Information Administration, but wind farms come with an environmental cost as wind turbines can…

Spoor uses AI to save birds from wind turbines

The key to taking on legacy players in the financial technology industry may be to go where they have not gone before. That’s what Chicago-based Aeropay is doing. The provider…

Cannabis and gaming payments startup Aeropay is now offering an alternative to Mastercard and Visa

Facebook and Instagram are under formal investigation in the European Union over child protection concerns, the Commission announced Thursday. The proceedings follow a raft of requests for information to parent…

EU opens child safety probes of Facebook and Instagram, citing addictive design concerns

Bedrock Materials is developing a new type of sodium-ion battery, which promises to be dramatically cheaper than lithium-ion.

Forget EVs: Why Bedrock Materials is targeting gas-powered cars for its first sodium-ion batteries

Private equity giant Thoma Bravo has announced that its security information and event management (SIEM) company LogRhythm will be merging with Exabeam, a rival cybersecurity company backed by the likes…

Thoma Bravo’s LogRhythm merges with Exabeam in more cybersecurity consolidation

Consumer protection groups around the European Union have filed coordinated complaints against Temu, accusing the Chinese-owned ultra low-cost e-commerce platform of a raft of breaches related to the bloc’s Digital…

Temu accused of breaching EU’s DSA in bundle of consumer complaints

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

The AI industry moves faster than the rest of the technology sector, which means it outpaces the federal government by several orders of magnitude.

Senate study proposes ‘at least’ $32B yearly for AI programs

The FBI along with a coalition of international law enforcement agencies seized the notorious cybercrime forum BreachForums on Wednesday.  For years, BreachForums has been a popular English-language forum for hackers…

FBI seizes hacking forum BreachForums — again

The announcement signifies a significant shake-up in the streaming giant’s advertising approach.

Netflix to take on Google and Amazon by building its own ad server

It’s tough to say that a $100 billion business finds itself at a critical juncture, but that’s the case with Amazon Web Services, the cloud arm of Amazon, and the…

Matt Garman taking over as CEO with AWS at crossroads

Back in February, Google paused its AI-powered chatbot Gemini’s ability to generate images of people after users complained of historical inaccuracies. Told to depict “a Roman legion,” for example, Gemini would show…

Google still hasn’t fixed Gemini’s biased image generator

A feature Google demoed at its I/O confab yesterday, using its generative AI technology to scan voice calls in real time for conversational patterns associated with financial scams, has sent…

Google’s call-scanning AI could dial up censorship by default, privacy experts warn

Google’s going all in on AI — and it wants you to know it. During the company’s keynote at its I/O developer conference on Tuesday, Google mentioned “AI” more than…

The top AI announcements from Google I/O

Uber is taking a shuttle product it developed for commuters in India and Egypt and converting it for an American audience. The ride-hail and delivery giant announced Wednesday at its…

Uber has a new way to solve the concert traffic problem

Google is preparing to launch a new system to help address the problem of malware on Android. Its new live threat detection service leverages Google Play Protect’s on-device AI to…

Google takes aim at Android malware with an AI-powered live threat detection service

Users will be able to access the AR content by first searching for a location in Google Maps.

Google Maps is getting geospatial AR content later this year

The heat pump startup unveiled its first products and revealed details about performance, pricing and availability.

Quilt heat pump sports sleek design from veterans of Apple, Tesla and Nest

The space is available from the launcher and can be locked as a second layer of authentication.

Google’s new Private Space feature is like Incognito Mode for Android

Gemini, the company’s family of generative AI models, will enhance the smart TV operating system so it can generate descriptions for movies and TV shows.

Google TV to launch AI-generated movie descriptions

When triggered, the AI-powered feature will automatically lock the device down.

Android’s new Theft Detection Lock helps deter smartphone snatch and grabs

The company said it is increasing the on-device capability of its Google Play Protect system to detect fraudulent apps trying to breach sensitive permissions.

Google adds live threat detection and screen-sharing protection to Android

This latest release, one of many announcements from the Google I/O 2024 developer conference, focuses on improved battery life and other performance improvements, like more efficient workout tracking.

Wear OS 5 hits developer preview, offering better battery life

For years, Sammy Faycurry has been hearing from his registered dietitian (RD) mom and sister about how poorly many Americans eat and their struggles with delivering nutritional counseling. Although nearly…

Dietitian startup Fay has been booming from Ozempic patients and emerges from stealth with $25M from General Catalyst, Forerunner

Apple is bringing new accessibility features to iPads and iPhones, designed to cater to a diverse range of user needs.

Apple announces new accessibility features for iPhone and iPad users

TechCrunch Disrupt, our flagship startup event held annually in San Francisco, is back on October 28-30 — and you can expect a bustling crowd of thousands of startup enthusiasts. Exciting…

Startup Blueprint: TC Disrupt 2024 Builders Stage agenda sneak peek!

Mike Krieger, one of the co-founders of Instagram and, more recently, the co-founder of personalized news app Artifact (which TechCrunch corporate parent Yahoo recently acquired), is joining Anthropic as the…

Anthropic hires Instagram co-founder as head of product