EnviousWispr
AI Polished Dictation

Tuned for better dictation.
The results are here.

Better writing takes careful work. Here’s how our local polish models performed, what we graded, and the decisions behind the results.

EG-1 v2 overall90.3%Pass + minor, across 1,462 cases
EG-1 self-corrections77.6%Across 219 correction cases
S1-mini median polish87 msSingle-request text cleanup
How we stack up

The details make
the difference.

Explore the strengths and tradeoffs of each model. This is our archived English text-polish comparison, not a test of complete dictation apps.

August 26, 2026

Complete English results by category

Pass rate includes pass + minor. Serious errors are in parentheses. Bold marks the highest pass rate, including ties.

CategoryCasesEG-1Pass (serious)S1-miniPass (serious)Fluid-1Pass (serious)
Overall1,46290.3% (66)86.5% (64)82.9% (81)
Clean speech30996.1% (8)93.9% (14)93.5% (14)
Self-correction21977.6% (22)59.8% (24)51.1% (21)
Fillers only20091.0% (9)95.0% (3)94.0% (8)
Voice at risk13491.8% (6)92.5% (6)86.6% (9)
Unfinished thought12091.7% (3)91.7% (3)93.3% (1)
Spoken list11485.1% (7)82.5% (9)54.4% (4)
Topic shift10294.1% (5)93.1% (2)94.1% (2)
Inline enumeration8792.0% (0)72.4% (1)97.7% (1)
Connected prose7494.6% (3)98.6% (0)95.9% (2)
Numbers and dates7397.3% (2)100.0% (0)93.2% (5)
Quoted instruction3080.0% (1)70.0% (2)43.3% (14)
Overall1,462 cases
EG-190.3%66 serious
S1-mini86.5%64 serious
Fluid-182.9%81 serious
Clean speech309 cases
EG-196.1%8 serious
S1-mini93.9%14 serious
Fluid-193.5%14 serious
Self-correction219 cases
EG-177.6%22 serious
S1-mini59.8%24 serious
Fluid-151.1%21 serious
Fillers only200 cases
EG-191.0%9 serious
S1-mini95.0%3 serious
Fluid-194.0%8 serious
Voice at risk134 cases
EG-191.8%6 serious
S1-mini92.5%6 serious
Fluid-186.6%9 serious
Unfinished thought120 cases
EG-191.7%3 serious
S1-mini91.7%3 serious
Fluid-193.3%1 serious
Spoken list114 cases
EG-185.1%7 serious
S1-mini82.5%9 serious
Fluid-154.4%4 serious
Topic shift102 cases
EG-194.1%5 serious
S1-mini93.1%2 serious
Fluid-194.1%2 serious
Inline enumeration87 cases
EG-192.0%0 serious
S1-mini72.4%1 serious
Fluid-197.7%1 serious
Connected prose74 cases
EG-194.6%3 serious
S1-mini98.6%0 serious
Fluid-195.9%2 serious
Numbers and dates73 cases
EG-197.3%2 serious
S1-mini100.0%0 serious
Fluid-193.2%5 serious
Quoted instruction30 cases
EG-180.0%1 serious
S1-mini70.0%2 serious
Fluid-143.3%14 serious

Serious errors changed meaning or dropped content. Lower is better. All models were scored on the same cases.

Our archived test · M5 Max · 64 GB · macOS 26
EG-1 v2 · S1-mini in lists mode · Fluid-1 with reasoning
What went into the test

Built around the way
people actually speak.

Corrections. Unfinished thoughts. Spoken lists. The test checks more than whether a sentence looks tidy.

1

Define good writing.

Write the policy for meaning, voice and structure before creating answer keys.

2

Bring speech into it.

Speak realistic scenarios with TTS, then transcribe them with a real recognizer.

3

Check the answer keys.

Two authors work independently. Resolve disagreements and exclude undecidable cases.

4

Grade consistently.

Use the same judge and rubric for every model, with an adjudication pass.

Explore the benchmark methodology +

The exam was sealed August 15, 2026. It contains 1,462 English cases across 11 categories; 11 undecidable cases were excluded. Initial answer-author agreement was 77.94%.

The August 26 run used an Apple M5 Max with 64 GB memory and macOS 26. GPT-5.6-luna graded the model outputs with the same rubric. Pass rates include pass and minor; serious errors are reported separately.

Envious Labs builds EG-1 and selected this test. Model grading has uncertainty. Fluid-1’s production instructions and custom decoder were not replicated, so this is not its official app performance.

Read the benchmark details ↗
Why it performs

More care for
what you meant.

We tune for the hard parts of spoken writing: keeping your voice, resolving corrections, and creating structure only where it belongs. We test those changes against the same written standards.

The thought you meant.

With AI polish
You say

Let’s meet Thursday, actually Friday.

Keep the corrected detail
Your finished text

Let’s meet Friday.

Illustrative example
A little more room for your ideas.

Try it with your own words.

Free, private dictation for Mac. No account. No subscription.

Download for MacmacOS 14+ · Apple Silicon