Changelog
Shipped, fixed, and improved
Every 1AudioTool release, straight from the release notes that ship with the product.
New
- Speaker separation in meetings — a Listen recording now shows who said what. Voices are told apart on this Mac after the meeting stops, each turn is attributed, and clicking a name renames that speaker everywhere in the transcript. Teach my voice records a few seconds of you and labels you Me instead of a number. Nothing is ever labelled Me until you enrol, and the voiceprint never leaves this Mac. Two engines, chosen per profile in Settings - Speech. Built-in needs nothing extra and separates an ordinary two- or three-person meeting. Speaker model is a download that is sharper on noisy calls and on voices that sound alike. Markers you drop during a meeting and the chapters built from them appear beside the speaker list, so a long recording can be read before it is replayed.
- Pick a microphone and see it working — the input picker in Dictate, Settings and the recording overlay lists every microphone by name and type, Built-in, Bluetooth, USB, Virtual or External, and draws a live level meter under the list. A microphone that another app is holding, or that has quietly stopped responding, is now visible before a recording is spent finding out. The picker is on the overlay while a dictation runs, so you can move to another microphone without losing what you have already said. When a capture fails, the same picker is on the failure card, and choosing an input there is what retries it.
- Create a profile from the sidebar — choosing "New profile" in the profile menu opens a name field in place. The new profile starts on the default settings and becomes active straight away, so a second voice or a different set of languages is a few keystrokes away. The free tier keeps one profile, and the menu explains what Pro adds rather than failing silently.
Improved
- The global gesture on macOS is Option twice, not Control twice. macOS binds double-Control to its own Dictation and keeps that binding even when Dictation is switched off, so reclaiming it meant swallowing every Control press system-wide and posting a synthetic one back a fraction of a second later. Nothing binds a bare double tap of Option, so the watcher now only listens: no modifier is ever swallowed or replayed, and Option-click, Option-drag and every Option chord are untouched. Existing profiles move across on their own. Windows keeps Control twice, where a lone Alt press would open the menu bar.
- Dictation is faster to finish — the model now loads while you are still speaking rather than after you stop, and it stays loaded between dictations, so the wait between finishing a sentence and seeing it appear went from about a second to about a fifth of one on the machine this was measured on. The timing, and how long the model stays loaded, are in Settings - Performance.
- The transcript reaches your cursor sooner — pasting used to wait for the history entry, the recording and the profile to be written first; now it goes first and the bookkeeping follows. Bringing your app back to the front no longer waits out a fixed pause, and no longer asks macOS for permission to control other applications.
Fixed
- Fixed dictation from the shortcut always failing when the app was not on screen — minimised, hidden behind another app, or on another Space, the shortcut would show the overlay with no timer, no microphone name and no input picker, then fail with "did not respond, on two attempts". The system will not open a microphone for a window that is not visible, and it does not refuse the request either; it simply holds it until the window comes back, which for a dictation into another app never happens. The recording overlay, which is on screen for exactly as long as a capture runs, now opens the microphone. Dictating from another app works the same way it does inside the app.
- Fixed the end of the last word being cut off — the final fraction of a second of every recording was dropped as the microphone was released. It is now waited for and included in the clip.
- Fixed the overlay going blank after the main window was reloaded — it stopped accepting updates for the rest of the session and showed a stale card.
SHIPPED · v0.1.0
Local dictation, nine spaces and correction memory
New
- Fully local macOS dictation powered by whisper.cpp, with Verbatim and Clean output modes.
- Speaker profiles with per-profile languages, model selection, vocabulary, microphone choice, correction memory, and transcript history.
- A global Control-twice shortcut, floating recording overlay, clipboard delivery, and optional insertion into the previously active application.
Improved
- Accuracy defaults that match how people actually dictate — new profiles start on explicit English rather than auto-detect, because a short clip in accented English is easily classified as the speaker's first language. The app warns when a short recording still uses Auto-detect.
- Cleaner audio before transcription — recordings have DC offset removed, outside silence trimmed and conservative gain applied, and the app reports quiet, clipped or mostly-silent input instead of leaving it to look like a model failure.
- Speech enhancement is now a choice — browser echo cancellation, noise suppression and automatic gain can smear consonants on some microphones, so each profile can A/B test it against the raw microphone.