Writher: The Windows Voice Tool That Solves the Last Mile

A local-first dictation and assistant app for Windows, built around Whisper, Ollama, and a clipboard-safe way to paste speech into any focused app.

7 min read · benmaster82/writher

A wide Windows desktop scene with a floating voice widget, a focused text field, and a local processing path that turns speech into pasted text. It explains that Writher's real job is not just transcription, but reliable delivery into the active app while keeping the workflow on-device.
Writher's promise is not voice input alone. It is voice that lands safely in the right Windows app.
Key Takeaways

The hard problem is not speech, it is delivery

Most voice tools stop at transcription. Writher starts where they usually stop: getting the words into the right Windows app without breaking whatever was already in the clipboard. That last mile is why it feels more like desktop plumbing than a chatbot.

The repo's cleverest move is to treat paste as a managed operation. `injector.py` saves the user's clipboard, injects the dictated text, then restores the original state. If the handoff fails, it writes a recovery copy to disk, which turns a fragile UX into a survivable one.

The last mile is a safety system, not a single paste command.

A close-up of a clipboard mechanism inside a desktop workflow, where original clipboard contents are held in a side compartment while dictated text slides into the focused app. It explains the repo's safety trick: Writher can paste into the current app without destroying what was already on the clipboard.
The clipboard swap is the part most voice tools gloss over.

Why benmaster82 built Writher

Writher comes from a simple complaint: cloud voice tools make you trade privacy, subscriptions, or both for a basic productivity win. The project answers that complaint with a local stack, so the microphone, transcription, and assistant logic stay on the machine after the models are installed.

I built Writher because I was tired of voice tools that require subscriptions or send audio to remote servers. It runs fully local using faster-whisper and Ollama.

bcorp, Author · Show HN post

That explains the shape of the product. It is not trying to be a universal assistant, and it is not trying to be a cloud service in disguise. It is trying to be a reliable Windows utility that happens to understand speech.

How the pipeline works

`main.py` coordinates the whole system. It listens for hotkeys, starts and stops `recorder.py`, pushes audio through `transcriber.py`, and keeps the UI alive with queues so transcription and LLM calls do not freeze the widget or tray icon.

The pipeline then splits. In dictation mode, speech becomes text and heads straight to injection. In assistant mode, `assistant.py` turns speech into structured actions through Ollama, then writes notes, lists, reminders, or appointments into `database.py`.

Two modes: hold AltGr to dictate text anywhere on Windows, or Ctrl+R for the AI assistant (notes, reminders, appointments). Works in any app – editor, browser, chat.

bcorp, Author · Show HN post

The date logic matters more than it sounds. Turning phrases like "next Monday at 3" into absolute timestamps means the model can do useful work without inventing a new interaction pattern, and the SQLite WAL setup lets the notes window read while the assistant writes.

The UI makes the invisible system feel safe

Writher's floating widget is not ornamental. The glassy pill, the expressive eyes, the tray icon, and the notes window all do state communication, which matters when the system is listening, thinking, or about to paste into whatever app happens to be focused.

That is a design pattern worth stealing. When a tool acts at the OS boundary, uncertainty is the enemy. Clear feedback lowers the cognitive tax of trusting it, especially when the app is living in the tray instead of on the taskbar.

What Writher gets right, and what it is really competing with

Writher is not winning by being the biggest model or the flashiest brand. It is winning by combining offline transcription, local assistant actions, clipboard-safe insertion, and global hotkeys into one small workflow that behaves like a native part of Windows.

Tool typeOfflineDictation anywhereAssistant actionsClipboard-safe pasteBest for
Single-purpose dictation appOften yesUsually yesNoSometimesFast transcription only
Cloud voice assistantNoSometimesYesRarelyConvenience over control
WritherYesYesYesYesWindows users who want one local workflow
A split scene showing two voice workflows. On the left, audio leaves a laptop for a cloud tower and comes back as text. On the right, the same audio stays inside the Windows machine and flows through local transcription, local assistant logic, and direct insertion into an app.
Writher's differentiator is architectural. The speech never has to leave the machine.

The contrast is not just privacy versus convenience. It is also about ownership. Writher gives the user a local path from voice to action, which means the core interaction does not depend on another company's uptime, policy changes, or account model.

The bigger bet

Writher points to a useful future for voice tools. Not a chat window with a microphone, but a local input primitive that behaves like part of the operating system. If the last mile stays reliable, voice stops feeling like a novelty and starts feeling like a normal way to work.