21 points sleepypandas 3 hours ago 9 comments

hs86 1 hour ago | parent

They also have a video announcement: https://x.com/WisprFlow/status/2100640514186072347

ks2048 1 hour ago | parent

They need to show some examples. You beat all the top models on your private data set? Show at least a couple examples - audio and transcripts - from examples that your model got right and others got wrong.

ymaws 44 minutes ago | parent

after reading the article I still have no idea how their thing performs, or if I should care how it performs since a majority of voice benchmarks still don't map to human evals

IOT_Apprentice 1 hour ago | parent

What languages are supported?

simonjgreen 1 hour ago | parent

Congrats Wisprflow, love to see this. Especially learning the users nuances and corrections, i think that may be novel in the dictation app space. Things like Handy have the ability to specify common typos but some method of automatic learning is new.

This is a pretty hot space right now. For me, it's all incredible for two things:

- I get my unabridged thoughts down on the page substantially quicker and cleaner using dictation. I believe dictation is the perfect first draft tool, and an amazing way to interact with AI too as you can dump tonnes of personal opinion in to every prompt without the overhead of keyboard interface.

- Accessibility! I know a couple of folks whos ability to use a computer has been tremendously elevated by the recent improvements in dictation. Due to mobility issues, they feel largely locked out of interacting online and tools like Handy and Wisprflow have been a game changer for them.

Very excited about all this, keep it coming!

qprofyeh 1 hour ago | parent

Wonder if they are aware of Cantonese, the language that is often abbreviated as Canto.

monknomo 43 minutes ago | parent

Probably this is after the italian word for singing

hbn 26 minutes ago | parent

Or Spanish, specifically it's the first person conjugation, meaning "I sing" or "I'm singing"

nr378 15 minutes ago | parent

Is it actually better than Microsoft's MAI-Transcribe-2? That generally seems like the best model right now and it's not included in their benchmarks.

I switched from Superwhisper->WisprFlow->Spokenly->Fieldwork and found WisprFlow the least accurate of the 4.