❯ Voice-Input Company Wispr Closes $280M Series B at $2B Valuation
VOICE REPLACES TYPINGVoice-input company Wispr has closed a $280 million Series B at a $2 billion valuation, led by Menlo Ventures. Its product, Flow, is a cross-app dictation layer: on iOS, Android, and Windows, users speak into any input box and Flow converts the speech into clean, formatted written text on the spot. The company has raised $361 million in total, and more than 60 billion words have already been written on the platform.
10-MONTH RE-RAISEThe previous round was a $25 million Series A extension led by Notable Capital in November 2025, when cumulative funding stood at $81 million. In less than ten months, that figure has jumped to $361 million. In between, Wispr moved itself up a notch from a dictation tool. Enterprise penetration has been fast: the company says people at nearly every Fortune 500 company and more than 10,000 enterprises are using it. These tools typically get in through employees installing them on their own, with IT procurement taking over once they catch on. All existing investors — Notable Capital, NEA, Neo Ventures, 8VC, and MVP Ventures — added to the round, joined by new investors Acrew, Forerunner, Goodwater, and Peak XV, with a batch of athletes and celebrities on the list as well.
IN-HOUSE MODELAlongside the round, Wispr also previewed Canto, its first self-developed speech recognition model — the dividing line between calling someone else’s model and building in-house. The company’s figures: in noisy real-world environments, word error rate drops from above 30% to 5%–10%, roughly a fourfold improvement. The model supports multiple languages and mid-sentence language switching, and draws on the user’s own vocabulary and contact list. Its effectiveness metric is zero-edit rate, the share of a dictation session that comes out needing no edits at all. Wispr expects Canto to reduce what needs fixing in daily use by about another 30%. The product line is expanding, too: meeting transcription tool Flow Notetaker is already live, and it has set up the Wispr Advanced Interfaces Lab, led by former Amazon Alexa researcher Ariya Rastrow, with a mission to make systems understand intent and deliver results directly. Proceeds from this round are mainly directed toward model R&D and expanding coverage.
THE INPUT LAYERA $2 billion valuation for a dictation tool looks absurd on the surface, but what the capital is buying is the position at the input layer. The keyboard is the gateway to all software — whoever stands between the input box and the application sees what users want to do in every scenario. That is also what emboldens Wispr to push toward understanding intent and delivering results directly. The risks are just as obvious: OS vendors already ship their own dictation, and Apple and Google could build this capability into the system at any time. Wispr’s room to maneuver is cross-platform reach and enterprise-side manageability. Voice’s positioning needs to be reassessed — it is moving from accessibility features and in-car scenarios into everyday input at the office desk.
▪ SIGNALThis round isn’t betting on transcription accuracy — it’s betting on the position at the input box. Whoever catches the spoken word sees the user’s intent first.