diff --git a/README.md b/README.md index bd9b431..58e78f4 100644 --- a/README.md +++ b/README.md @@ -61,21 +61,48 @@ Stream: hotkey → mic PCM chunks → Riva/NIM WebSocket → live words typed ## Screenshots -Everything is configured in the GTK Settings window — presets, engines, input, -hands-free wakeword, and a built-in benchmark, all with tooltips and -screen-reader (ATK) support. Click any thumbnail to view full size. +Everything is configured in the GTK **Settings** window — every tab has tooltips +and screen-reader (ATK) support. Click any image to open it full size. -| Presets | Engines | Input | General | -|:---:|:---:|:---:|:---:| -| [](Screenshots/settings-presets.png) | [](Screenshots/settings-engines.png) | [](Screenshots/settings-input.png) | [](Screenshots/settings-general.png) | -| **Benchmark** | **Log** | **About** | **Wakeword (hands-free)** | -| [](Screenshots/settings-benchmark.png) | [](Screenshots/settings-log.png) | [](Screenshots/settings-about.png) | [](Screenshots/wakeword.png) | +
+ 
+ Presets — your dictation actions. Each preset is either a plain transcription or an LLM rewrite, and carries its own spoken keyword(s) for voice routing, an optional global hotkey, and a custom rewrite prompt.
+
+ 
+ Engines — your speech-to-text and language-model back-ends, local or remote. Add and rename engines, watch live online/offline status, and pick models from a searchable list fetched straight from the endpoint.
+
+ 
+ Input — how you start and stop dictation: the modifier-key scheme (Ctrl+Win / Ctrl / Alt / Esc) or custom hotkeys, plus the silence-based auto-stop (VAD), the quality gate, and audio cues.
+
+ 
+ Wakeword (hands-free) — point Blitztext at a Wyoming/openWakeWord server, choose a wake model, and test the connection live so a spoken keyword starts dictation with no keys at all.
+
+ 
+ General — core preferences: microphone with a live level meter, output mode (type vs. paste), language hint, type delay, and autostart on login.
+
+ 
+ Benchmark — compare every configured STT engine against a reference WAV + transcript to find the fastest and most accurate, with a Device column (CPU / GPU / remote).
+
+ 
+ Log — the in-app log buffer: a live view of recording, transcription, routing, and wakeword events for quick troubleshooting.
+
+ 
+ About — version, source link, changelog, and licence.
+
+ 
+ Presets — your dictation actions. Each preset is either a plain transcription or an LLM rewrite, and carries its own spoken keyword(s) for voice routing, an optional global hotkey, and a custom rewrite prompt.
+
+ 
+ Engines — your speech-to-text and language-model back-ends, local or remote. Add and rename engines, watch live online/offline status, and pick models from a searchable list fetched straight from the endpoint.
+
+ 
+ Input — how you start and stop dictation: the modifier-key scheme (Ctrl+Win / Ctrl / Alt / Esc) or custom hotkeys, plus the silence-based auto-stop (VAD), the quality gate, and audio cues.
+
+ 
+ Wakeword (hands-free) — point Blitztext at a Wyoming/openWakeWord server, choose a wake model, and test the connection live so a spoken keyword starts dictation with no keys at all.
+
+ 
+ General — core preferences: microphone with a live level meter, output mode (type vs. paste), language hint, type delay, and autostart on login.
+
+ 
+ Benchmark — compare every configured STT engine against a reference WAV + transcript to find the fastest and most accurate, with a Device column (CPU / GPU / remote).
+
+ 
+ Log — the in-app log buffer: a live view of recording, transcription, routing, and wakeword events for quick troubleshooting.
+