Latka logo

Top 2 Text to Speech Software Companies $100M+ Revenue (September 2026)

As of September 2026, Latka tracks 2 text to speech software companies with $100M+ in annual revenue. They have combined revenues of $1B and employ 400 people. They have raised $881M.

Every company below sells text to speech software to other businesses and is ranked by its most recent annual revenue. Revenue, funding, headcount and customer figures come from CEO interviews on the Latka podcast, public company filings, and Latka estimates where a company has not disclosed a number.

What Text to Speech Software Companies do

Text to Speech (TTS) software converts written text into spoken words, enabling users to synthesize speech from digital content. This technology is primarily utilized in applications such as accessibility for individuals with visual impairments, voiceovers for videos, language learning tools, and customer service automation. TTS software often includes features like voice selection, speed control, and the ability to handle various languages and accents, providing flexibility for different user needs. The software is commonly used by a diverse range of professionals, including educators looking to enhance learning experiences, content creators producing multimedia presentations, and businesses implementing automated customer service solutions. With the growing emphasis on accessibility and user engagement, TTS has become an essential tool across educational, corporate, and creative sectors, facilitating seamless communication and broadening the reach of digital content.

Companies
2
Revenue
$1B
Funding
$881M
Employees
400

Filters

Sorting: Highest -> Lowest

Filters

Top Text to Speech Software Companies $100M+ Revenue

Showing 2 of 2 companies ranked by annual revenue.

1Eleven Labs logo
Eleven Labs

New York City, New York, United Kingdom

Research Lab: Exploring new frontiers of voice generation. We are dedicated to researching and implementing innovative techniques in voice artificial intelligence (AI) to enhance the appeal of content across different languages and voices. Our goal is to reach new audiences and viewers by ensuring a more immersive and engaging experience.

Revenue
$500M
Year founded
2022
Funding
$881M
Team size
400
Growth
51.52%
2edream,Inc logo
edream,Inc

China

Immerse yourself in the world of spoken languages with our AI-powered language app. Learn multiple languages, practice real-life scenarios, receive grammar corrections, and choose from a variety of voices. Start your language learning journey today!

Revenue
$500M

Frequently asked questions about Text to Speech Software Companies $100M+ Revenue

How many text to speech software companies with $100M+ in annual revenue are there?

Latka tracks 2 text to speech software companies with $100M+ in annual revenue. Together they generate $1B in annual revenue and employ 400 people.

Which text to speech software company with $100M+ in annual revenue is the largest?

Eleven Labs is the largest, with $500M in annual revenue, founded in 2022.

How much revenue does a typical text to speech software company with $100M+ in annual revenue make?

The average text to speech software company in this list makes $500M a year, across 2 companies with reported revenue.

Who are the leading Text to Speech software vendors with $100M+ in annual revenue?

Ranked by annual revenue, the leaders are Eleven Labs and edream,Inc.

How much funding have text to speech software companies with $100M+ in annual revenue raised?

The 2 text to speech software companies with $100M+ in annual revenue tracked here have raised $881M in disclosed funding between them.

Related Artificial Intelligence Software categories

Inclusion Criteria

- Must enable conversion of written text to spoken words - Should support multiple languages and accents - Must include voice selection options for varied user preferences - Should offer features for customization, such as speech speed control - Must be applicable for accessibility purposes, helping individuals with visual impairments - Not limited to basic text reading; must also support enhanced functionalities like emotion in speech synthesis