NEWSLETTER

By clicking submit, you agree to share your email address with TFN to receive marketing, updates, and other emails from the site owner. Use the unsubscribe link in the emails to opt out at any time.

Wispr raises $280M at $2B valuation from Menlo Ventures to build voice layer beneath every app

Wispr founders
Image credits: Wispr
  • Wispr raises $280M Series B at a $2B valuation, led by Menlo Ventures.
  • Total funding hits $361M, less than 10 months after its last raise.
  • New Canto speech model cuts dictation word error rates from 30% to under 10%.

Millions of people now talk to their laptops instead of typing into them, and Wispr wants to own the layer that makes that possible. The AI dictation startup Wispr has raised $280 million in Series B funding at a $2 billion valuation, led by Menlo Ventures, in one of the firm’s largest bets on a single AI company to date. 

“Dictation was always the starting point for something bigger. The real ambition is to build voice into the foundational layer beneath every other piece of software and hardware,” said Tanay Kothari, co-founder and CEO of Wispr.

A $361 million war chest, less than a year after the last one

The round takes Wispr’s total funding to $361 million, arriving less than 10 months after its previous raise. 

Existing backers Notable Capital, NEA, Neo Ventures, 8VC, and MVP Ventures doubled down, while Acrew, Forerunner, Goodwater, Peak XV, Together Fund, and PLUS Capital joined as new investors.  It follow Tech Funding News‘ earlier reporting that Menlo Ventures was in talks to back a $260 million round at a valuation close to the same.

Founded in San Francisco in 2021 by Kothari and Sahaj Garg, the company built Wispr Flow, an AI dictation tool that turns speech into formatted text across desktop and mobile apps. Flow has since spread across more than 125,000 businesses and most of the Fortune 500, largely through employees adopting it on their own rather than top-down enterprise sales.

Backing the model

Alongside the raise, Wispr unveiled Canto, its first proprietary speech model, addressing user complaints about a recent dip in Flow’s accuracy. In noisy, real-world conditions, Wispr says Canto cuts word error rates from more than 30% to between 5% and 10%, which it expects will mean 30% to 35% fewer dictations need editing. 

Until now, Wispr has been built on top of other companies’ speech models. Canto is the first one it owns. The work is housed in the newly launched Wispr Advanced Interfaces Lab, led by chief scientist Ariya Rastrow, a founding member of Amazon’s Alexa team who later led multimodal foundation model work at Meta. 

“The bottleneck in AI has moved from the model to the human interface to the model, and Wispr is the interface. We watched Flow spread through most of the Fortune 500 before there was a sales team to speak of, pulled in by employees one desk at a time. That is why this is one of the largest AI investments in our firm’s history: voice is becoming part of every department at every company, and Wispr is building what comes after the text box,” said Matt Kraning, partner at Menlo Ventures. 

Wispr competes with dictation rivals Willow, Monologue, Aqua, and Superwhisper on one side, and meeting-productivity platforms on the other: Granola, which raised $125 million at a $1.5 billion valuation in March; Fireflies, which crossed a $1 billion valuation through an employee tender offer; and Read AI, which raised a $50 million Series B. 

Where those tools structure meeting conversations, Wispr is saying that voice will become the default way people talk to software.

Capital is chasing the wider category 

ElevenLabs raised $500 million at an $11 billion valuation in February and was later reported to be in talks for a secondary sale near $22 billion, a tender offer letting existing shareholders cash out, not fresh capital. 

The voice recognition market is projected to grow from roughly $22 billion in 2026 to $62 billion by 2031, a 22% annual growth rate, according to Mordor Intelligence. 

The question this round doesn’t answer is whether voice actually replaces the keyboard, or just becomes a faster way to fill the same text box. Wispr is wagering hundreds of millions of dollars on the former.

Total
0
Shares
Related Posts
Total
0
Share

Get daily funding news briefings in the tech world delivered right to your inbox.

Enter Your Email
join our newsletter. thank you
TFN Banner