They pay to answer calls on the number clients already know, no bot and no second phone — and for every transcript to arrive already filed to the right project and counted toward the right invoice, instead of a folder of audio they triage themselves.
REPLACEMENT BRIEF
dictation
Can AI replace Superscribe?
The dictation half is genuinely solved — a hotkey, a realtime STT stream, one cleanup pass, done in a sitting, with open-source clones to fork. The product's spine is the other half: capturing live calls on your existing iPhone number takes carrier forwarding, a pooled Twilio line, and a CallKit softphone debugged against real ringing phones. And a transcript is not the product either — around it sits a workspace that files every call to the right client using your GitHub activity, then turns it into billable time, reports, and an API.
Bind a global hotkey, stream mic audio to a realtime STT API, run one LLM cleanup pass, and paste the result into whatever app has focus.
Build the personal version →AT A GLANCE
- price
- $38/mo
- listed annual price
- $456/yr
- replaceable scope
- Bind a global hotkey, stream mic audio to a realtime STT API, run one LLM cleanup pass, and paste the result into whatever app has focus.
- build time
- one sitting (dictation only), multi-week with call capture
Build promptreviewed prompt · not run
the prompt
Curated promptBuild me a push-to-talk dictation tool for macOS to replace Superscribe's desktop app. Requirements: - Swift menu bar app, SPM only, no Xcode project. A global hotkey (default: hold Option+Space) records while held, stops on release. - Capture the mic with AVAudioEngine, downsample to 16kHz mono PCM, and stream it over WebSocket to ElevenLabs Scribe realtime (key in .env). Show partial transcripts in a small floating panel while I speak. - On release, run one LLM cleanup pass over the final transcript (fix punctuation, drop filler words; key in the same .env), then paste it into the focused app via NSPasteboard + CGEvent Cmd-V and restore my previous clipboard afterwards. - Fallback: with no ElevenLabs key, record to a temp wav and transcribe locally with whisper.cpp instead. Slower is fine. - Menu bar icon shows idle/recording/transcribing states; a history window lists the last 20 transcripts with copy buttons, persisted to ~/Dictation/history.jsonl. - Ad-hoc codesign for my own machine only. README documents the Microphone and Accessibility permission prompts and the whisper.cpp model download. - Out of scope: live word-by-word typing into the field (paste on release only), Windows support, accounts and billing, and the phone-call capture product. If I want call notes I will upload recordings by hand.
The prompt stays readable first. Choose a launch option when you are ready.
$ open in your agent (prompt prefilled, you press enter) or copy it raw
what AI can build
Bind a global hotkey, stream mic audio to a realtime STT API, run one LLM cleanup pass, and paste the result into whatever app has focus.
The prompt was editor-reviewed; no completed build is recorded.
The honest tradeoff
why people still pay
what you lose
xcall capture on your existing number (carrier forwarding into a CallKit softphone)
xglitch-free live word-by-word insertion tuned per app: terminals, Electron editors, browser contenteditables
xauto-filing: dictations and calls alike are matched to the right project via embeddings (enriched with GitHub repo context) plus live commit activity during the work block, so billable time lands on the right client without tagging
xthe workspace around the transcripts: semantic search, invoice-ready PDF reports, CRM note drafts, a public API, and an MCP server
xWindows parity and signed, notarized auto-updating installers
Start with existing software
prior art · use these instead of building, if you'd rather
EVIDENCE LEDGER
What this page can prove
The verdict judges replaceability. The evidence level records what DeepFeather actually checked.
The prompt was editor-reviewed; no completed build is recorded.
Business Voice · monthly
known limits · call capture on your existing number (carrier forwarding into a CallKit softphone); glitch-free live word-by-word insertion tuned per app: terminals, Electron editors, browser contenteditables
BUILD FEEDBACK
Did you try this build?
Report the outcome. Submissions enter a manual evidence queue and never auto-upgrade the verdict.
Want next week’s replacements?
New verdicts + most-wanted, weekly. Free. One-click out.
Share this verdict
questions
Can AI replace Superscribe?
Possibly for a narrower core workflow, but this catalog judgment is not a verified build. Expected gaps include: call capture on your existing number (carrier forwarding into a CallKit softphone), glitch-free live word-by-word insertion tuned per app: terminals, Electron editors, browser contenteditables. Validate the prompt against your own acceptance criteria before committing.
How much does Superscribe cost?
Superscribe is listed at about $38/month (Business Voice, checked 2026-07-30), or $456 per year. This is a pricing reference, not evidence of a completed replacement or realized savings.
What do I lose by replacing Superscribe?
Honestly: call capture on your existing number (carrier forwarding into a CallKit softphone); glitch-free live word-by-word insertion tuned per app: terminals, Electron editors, browser contenteditables; auto-filing: dictations and calls alike are matched to the right project via embeddings (enriched with GitHub repo context) plus live commit activity during the work block, so billable time lands on the right client without tagging; the workspace around the transcripts: semantic search, invoice-ready PDF reports, CRM note drafts, a public API, and an MCP server; Windows parity and signed, notarized auto-updating installers. If any of those are load-bearing for you, keep paying.
Is there an open-source alternative to Superscribe?
Yes — Handy (MIT cross-platform push-to-talk dictation app in Tauri, built to be forked; covers the whole desktop core loop.), VoiceInk (GPL native Swift macOS dictation app with local whisper.cpp and per-app modes; closest open clone of the mac side.), FreeFlow (Solo-built open Wispr Flow clone; working example of cloud STT plus LLM cleanup with active-window context.), Twilio Media Streams (Official docs and tutorials for streaming live call audio to STT; the happy path of the call-capture half.). Using prior art is also a valid exit; the prompt is for when you want it exactly your way.