AIToday
Audio & SpeechAI Business & IndustrySiliconANGLE AIPublished: Aug 18, 2026, 01:01 JST3 min read

Wispr raises $280M for AI speech-to-text at $2B valuation

Wispr raises $280M for AI speech-to-text at $2B valuation

Key takeaway

  • Wispr AI, developer of the Flow dictation platform, raised $280 million in Series B funding at a $2 billion valuation to expand its AI-powered speech-to-text technology.

  • The company announced Canto, its proprietary AI model designed to detect human speech in noisy environments, reducing error rates from 30% to nearly 5%–10% in loud settings.

  • This advancement makes dictation viable for real-world use cases beyond quiet offices—such as in cars or crowded venues—and has potential value for both mobile workers and the deaf and hard-of-hearing community who depend on transcription services.

3 Key Points

  1. What happened

    Wispr AI, which makes Flow—a dictation tool that converts spoken words into clean text—raised $280 million in Series B funding at a $2 billion valuation, led by Menlo Ventures. The company also previewed Canto, its first proprietary AI model trained to isolate human speech in noisy environments.

  2. Why it matters

    Canto reduces transcription error rates in loud settings from 30% to nearly 5%–10%, making dictation practical beyond quiet rooms—in cars, offices, or near traffic. This matters for mobile workers, but also for deaf and hard-of-hearing users who rely on AI transcription in real-world, multi-speaker settings. Wispr says people have written over 60 billion words on Flow and the tool is used by almost all Fortune 500 companies and over 10,000 enterprises.

  3. What to watch

    Canto is in preview. The company's total raised now stands at $361 million after this round, which included new investors Acrew, Forerunner, Goodwater, Peak XV, Together Fund, and PLUS Capital alongside existing backers.

Ask the AI about this article →

Context & Analysis

Wispr's Series B round reflects investor confidence in a narrow but high-value wedge: AI-powered dictation that works in the real world, not just in controlled settings. The company's existing reach—Fortune 500 adoption and 60 billion words authored on Flow—suggests the product has already proven its utility across industries. What Canto represents is a step from "occasionally useful" to "reliable in messy reality," a shift that unlocks use cases previously out of reach: taking notes while driving, capturing ideas in noisy offices, or enabling multi-speaker conversations for users who depend on live transcription.

The technical challenge Wispr is addressing is real. Most modern AI transcription models train on clean audio because that is the easiest and cheapest path; Canto's different training regimen—learning from noisy, overlapping, interrupted speech—is more expensive upfront but solves a genuine pain point. The 30%-to-5–10% error-rate reduction in loud environments is a concrete claim that, if borne out in practice, changes the utility profile of dictation for mobile and remote workers. The company's emphasis on the deaf and hard-of-hearing community also signals an understanding that transcription improvements matter beyond convenience—they enable access. The Series B's scale ($280 million) and the roster of tier-one investors (Menlo, NEA, Notable Capital, and new backers including Peak XV and Forerunner) suggest that this particular AI application—voice-to-text in noisy real-world conditions—is viewed as a durable, expanding market.

FAQ

What is Canto and how does it work differently from other speech models?
Canto is Wispr's proprietary AI model trained to detect speech across a wide variety of speech variances, loud environments, noise, interruptions, background voices, and environmental sounds like dogs barking or traffic. Unlike most speech models trained on clean voice assets, Canto isolates human voices from environmental noise, enabling people to use dictation apps where they live—in cars, offices, or near traffic.
How much does the error rate improve with Canto in noisy environments?
When Canto is built into Flow, Wispr's dictation tool, the error rate in noisy environments falls from 30% to nearly 5%–10%.
Who is using Wispr's Flow product?
Wispr says almost all Fortune 500 companies and over 10,000 enterprises use Flow. People have written more than 60 billion words on the platform. The company also counts three-time NBA All-Star Domantas Sabonis among its users.
SiliconANGLE AIRead Original Article

Get the latest Audio & Speech news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleVertiv, Supermicro report divergent AI infrastructure paths—margin vs. growth

The AI news that matters, in one minute each morning.

Sign up free