SHIPPED · v0.3.0
Live translation on this machine, a speech model shelf and updates you read first
New
Live translation, on this machine — talk in one language and the captions arrive in another, about a second behind your voice. Each sentence is translated during the pause that follows it, so nobody waits for you to stop. Settings - Speech picks the caption language, and the pack for that pair is a download with its size and its progress. Nothing ships in the bundle and nothing downloads on its own: until you take a pack, the feature says it needs one rather than fetching it behind you. Captions window opens the same captions at reading-across-a-room size — three type sizes, a high-contrast switch, and it stays above other windows. Captions on this network hands anyone on the same Wi-Fi a read-only page, as a link to copy or a QR code to scan on a phone. It carries the pairing token and dies with the pairing session after fifteen minutes; the page then stops updating and says so. Keep the sentence you spoke above its translation, or show only the language your audience reads. The transcript stays what was heard. A translation is never pasted, never stored as the transcript, and never trusted to carry names or filenames.
A shelf of speech models, with a Settings tab of its own — Settings - Models lists every pack the app can use, with its size, a line on what that size buys, a Download button and progress. A fresh install shows the same shelf's door on Dictate instead of a picker that only says "No model installed", and the first pack to land becomes the profile's model, so the next dictation just works.
An update is now something you read before you take it. When 1AudioTool finds a newer signed release it opens a window with the version, when it was published, and the full release notes — headings, bullets, emphasis and any screenshots the release carried — then offers one button that downloads it and restarts. The download reports its size as it goes, and if anything fails it says so rather than leaving the button looking pressed. Later asks again next time you open the app; Skip this version stops that one release being offered at all. Either way Settings - General still has it, and asking for it there is what un-skips it. The check runs shortly after launch instead of only when you go looking, so the status line can say an update is waiting. On a first run it waits its turn behind the question about anonymous counts.
Open a meeting — a Listen recording now has a page of its own. The recording plays back and every speaker turn seeks it, the transcript can be made again from that recording with today's languages and a bigger model, and correcting it there teaches the same phrase rules a corrected dictation does. A second pass rebuilds everything the first one produced: speakers, chapters and action items. The language picker is in Listen as well as in Settings, because the languages a meeting is decoded with are a decision made where the meeting starts. A note says when the languages and the model disagree — an English-only pack cannot mix languages — and a transcript that came back mostly silence says so plainly rather than leaving you to infer it.
Play back the clip you just dictated — the result card has a player with a scrubber you can seek with the arrow keys. A clip whose recording was not kept says so instead of showing dead controls.
Read your corrections with a coding CLI you already have — Review's new Analysis tab groups the review queue by cause: the vocabulary the model keeps missing, the homophones, the places a second language starts mid-sentence. It names the near-duplicate rules that should be one rule, and says what to do about each group. Settings - Agents lists the coding CLIs found on this machine and which one to use; the default is to ask each time, and nothing runs until you press the button. This is the one feature that sends anything off this machine — see Privacy.
The room says when it is working against you — a dictation now notices music or a TV playing behind you, a room your voice comes back off, and a background sitting nearly as loud as you are. It names the worst one, once, quietly, and never stops the recording. Settings - Speech turns it off.
Every paste mode explains itself — Verbatim, Clean, Prompt, Message and Command each show one line on what lands in the clipboard and the same spoken sentence run through them, so a mode can be understood without spending a recording to find out.
Improved
The Inbox is now Review — the same clips waiting to be looked at and the same rules waiting to be learned, with the Analysis tab beside them.
Closing the main window no longer stops dictation working. It hides, the tray brings it back, and in between the shortcut still starts and stops recordings. Quit still quits.
The app stops drawing what nobody is looking at — a hidden window stops rendering instead of painting to a screen that is not there, a level meter releases the microphone when its panel is put away, and the timers in a window you cannot see stop counting. This was background processor use, on both platforms, for doing nothing.
The shortcut is written as an instruction — "Double-tap ⌥ to stop" rather than a pair of symbols you have to decode.
Actions is hidden while the action engine is unfinished — it was a space you could open and a mode you could choose that did not yet do anything for you. Command mode is still in the mode list, now marked as not yet wired up, and delivers Verbatim until it is.
Fixed
Fixed every recording in a signed release failing with the microphone blocked, while System Settings showed 1AudioTool switched on — the shipped build did not declare the microphone the way macOS demands of a notarized app, so the toggle in Privacy & Security granted nothing and no prompt ever appeared. Release builds now declare it, and a release is checked for that declaration before it ships.
Fixed the double tap not stopping a recording when permission arrived after launch — watching for the gesture needs Accessibility or Input Monitoring, and granting it used to take a restart before the gesture came alive. The watcher now starts within seconds of the grant, and a recording begun from inside the app can be stopped the way the overlay says it can.
Speech models can now be downloaded from inside the app, so a first recording on a new install no longer fails with "whisper-cli was not found". The transcription path looked for an engine before it looked for a model, and reported a missing model as a missing developer tool. It now says that no model is installed and where to get one — and when the bundled engine has not finished loading the model, it waits for it rather than falling back to a command-line tool that packaged builds do not carry.
Fixed Vietnamese input methods (UniKey, EVKey) stopping while 1AudioTool runs on Windows, and Control shortcuts arriving out of order. The double-Control watcher swallowed every Control press for a quarter of a second and re-injected it later, which starved other keyboard hooks and reordered chords. It now only listens, like the Option watcher on macOS: every key passes straight through, nothing synthetic is posted, and an input method's own replacement text is left alone.
Fixed Esc doing nothing on the overlay once it had been clicked on Windows. The panel now answers Escape itself — cancel while listening or transcribing, dismiss on a failure — in addition to the global watcher.
Fixed a black console window flashing on Windows whenever the app started its speech or translation engine, downloaded a model, checked the licence, or read the machine id. Every helper the app runs is now started without a console.
The Windows helpers are built against the static C runtime, so they start on a machine that has no Visual C++ redistributable installed.
Fixed pressing Install doing nothing for several seconds, and sometimes failing after the app had already offered the update. Installing re-ran the whole check over the network before it began, so the press could not be answered until a second round trip had finished, and that round trip had its own way to fail. The release found by the check is now the one that installs.
Fixed a recording cancelled while the microphone was still opening leaving the microphone open, and stopping one recording muting the next when the two overlapped. A capture that is no longer wanted now releases its input the moment it arrives, and only the capture that owns the audio is allowed to tear it down.
Fixed a correction you saved looking as though it had been thrown away — the result card went back to showing what the model heard rather than what you corrected it to.
Fixed Prompt mode cutting your sentence at the dot in a filename — "fix the login bug in auth.ts before the demo" arrived as "fix the login bug in auth." A full stop now has to end a sentence before it is treated as one.

Settings, Speech tab with live translation on, German chosen as the caption language and the English to German pack installed 
Caption layout set to Source above target, with the Captions on this network switch that opens a read-only page for the same Wi-Fi 
Translation packs list showing each language pair with its download size, one pack installed and the rest waiting to download 
Speech models shelf listing tiny through large-v3 with a one-line trade-off and size for each, two packs installed