Notes on speech,
voice typing and AI notes.
Explainers on how on-device transcription works, which models to pick, and how to get more out of your recordings.
Cloud transcription API pricing, per hour of audio
Cloud transcription API pricing per hour of audio: verified batch and streaming list prices, worked annual costs, and the charges not in the rate.
12 min readHow much RAM a Mac needs for Whisper and a local LLM
How much RAM a Mac needs for Whisper, Parakeet and a local LLM: unified memory, model sizes by quantization, KV cache growth, and 8, 16, 32 and 64 GB budgets.
12 min readHow Speaker Diarization and Speaker Recognition Work
Speaker diarization explained in plain words: how software splits audio by speaker, matches voices across recordings, where it fails, and how to help it.
8 min readLive captions and transcripts as an accessibility tool
How live captions and recorded transcripts solve different problems for deaf and hard of hearing people, where each one helps, and what this kind of app is not.
9 min readOffline speech to text on iPhone: what the platform allows
Offline speech recognition on iPhone: Apple's on-device dictation, the iOS 26 SpeechAnalyzer API, Whisper models, and why keyboard extensions cannot use the mic.
9 min readOn-device transcription vs cloud: privacy, cost and speed
What on-device transcription means technically, what cloud services do with your audio, the per-minute billing math, and the cases where cloud still wins.
8 min readPrivate Medical and Legal Dictation on Mac and iPhone
Why clinicians and lawyers need a private medical dictation app or legal dictation software that keeps audio on the device, and what to check with compliance.
9 min readWhisper vs Parakeet: Which On-Device Speech Model Is Better?
Whisper vs Parakeet TDT compared in plain words: architecture, speed, languages, accuracy, hallucinations and when to pick which on-device speech to text model.
9 min readWhy Whisper hallucinates and repeats itself, and how to stop it
Whisper hallucination and repetition loops explained: why silence becomes 'Thank you for watching' and the whisper.cpp, openai-whisper and faster-whisper fixes.
10 min read