github TypeWhisper/typewhisper-mac v1.5.0

one month ago

TypeWhisper 1.5.0 is the stable 1.5 macOS release. It brings app-aware dictation insertion, virtual audio input support, broader number normalization, expanded bundled speech providers, local-model memory controls, dictionary-learning improvements, and a focused reliability pass across hotkeys, workflows, plugins, model handling, and release metadata.

Highlights

  • Added app-aware smart insertion so dictated text better respects surrounding text, sentence position, trailing spaces, terminal paste behavior, and target-app context.
  • Added virtual audio input device support and hardened media pause/resume behavior, recording start feedback, mouse-button shortcuts, and fullscreen indicator handling.
  • Expanded number normalization across generic transcription output, multilingual number words, Dutch number words, English ordinals, English digit sequences, and French decimal phrases.
  • Added and refreshed bundled speech and AI providers, including Gemini speech transcription, Cartesia speech transcription, Sber SaluteSpeech, OpenRouter speech-to-text, Reson8, Mistral AI, OpenAI-compatible profiles, Soniox regions and TTS, and updated cloud ASR providers.
  • Added Japanese localization and dictation support, ordered language hints, recent transcriptions in the workflow palette, source progress for file transcription, recorder-specific engine overrides, and a short-dictation workflow AI-skip setting.
  • Added local MLX memory-footprint controls, idle local-model auto-unload behavior, and recovery for stalled Gemma 4 model downloads.
  • Added pro transcription fallback, auto-learned dictionary corrections, per-term dictionary CTC tuning plumbing, and tighter target-app correction learning.

Reliability and Fixes

  • Fixed delayed Hybrid hotkey behavior, restored non-Control hybrid modifier taps, improved global push-to-talk capture, aligned Pages workflow hotkeys with palette behavior, added menu-based dictation pause, and required confirmation before Esc cancels dictation.
  • Fixed fullscreen indicator menu-bar strip issues and the top overlay menu-bar strip.
  • Fixed Groq and cloud ASR compressed M4A upload handling, including a WAV fallback for incomplete M4A finalization.
  • Fixed Soniox async file transcription routing, region persistence, and realtime model selection.
  • Fixed OpenAI Compatible GPT-5 token parameter handling, host SDK compatibility, and configurable request-timeout behavior.
  • Fixed plugin host compatibility metadata, raised cloud ASR plugin host floors to TypeWhisper 1.5 where required, and prevented plugin releases from taking over the repository-level GitHub Latest marker.
  • Fixed TaskForge preserve-clipboard insertion fallback, dictionary scroll crashes, event tap teardown, FaceTime built-in microphone capture, plugin uninstall keychain cleanup, and development build cleanup scope.
  • Fixed Parakeet v2/v3 model setup by repairing missing vocabulary downloads before model loading.
  • Improved Japanese dictation post-processing and settings localization, Latin word replacement boundaries, terminal clipboard restore verification, and recorder final transcription failure surfacing.

Developer and Distribution Notes

  • Stable 1.5.0 builds use the default Sparkle channel and require macOS 14.0 or later.
  • Release candidates and daily builds remain on their dedicated Sparkle channels and do not update Homebrew.
  • The final stable tag publishes DMG and ZIP assets, updates the stable Sparkle appcast entry, triggers the website release dispatch, and updates the Homebrew cask.
  • The 1.5 plugin line uses the community-capable registry feed and keeps host compatibility metadata aligned with the bundled plugin SDK.

Don't miss a new typewhisper-mac release

NewReleases is sending notifications on new releases.