Benchmark run
2026 09 v2 flow
Wispr Flow 1.6.774 external-product benchmark; benchmark-v2 publication batch 2026-09-v2; wispr-flow wispr-flow; clips [0, 400) of the consumable range; hotkey option+z
Sep 5, 2026
- Models
- 1
- Samples per dataset
- 400
- Warmup runs
- 3
- Hardware
- Apple M4 Max / 36 GB / macOS 26.6.2
Full results
Accuracy is the share of words transcribed correctly per language condition, so higher is better; speed is milliseconds per second of audio, so lower is better. The best value in each column is emphasised, and a cell reading a dash says on hover why there is no figure in it.
Accuracy (%) per language, and speed in ms per second of audio, so lower is better. A cell reading a dash was never measured, and hovering any cell says what is behind it.
| Wispr Flow | 400 | No disk file: it is a cloud product. | Memory of a cloud product cannot be observed. | 132 ms | 91.8%Every word from every language counted together, not an average of per-language scores. | 94.5%Every word from every language counted together, not an average of per-language scores. | 90.1%Every word from every language counted together, not an average of per-language scores. | 95.5% | 93.5% | 96.1% | 94.6% | 76.7% |
|---|
At a glance
Ratings computed from benchmark data, scaled 1 to 10, and written with their scale in every cell. Accuracy is based on Word Error Rate (WER) and does not include punctuation yet.
Ratings run 1 to 10, higher is better. A cell reading a dash was never measured, and hovering any cell says what is behind it.
| Wispr Flow | The languages this row covers were never recorded. | Whether this row translates was never recorded. | 7/10 | 9/10Every word from every language counted together, not an average of per-language scores. |
|---|