github FluidInference/FluidAudio v0.13.5

latest releases: v0.17.7, v0.17.6, v0.17.5...
6 months ago

What's New in v0.13.5

Features

  • Add experimental CTC zh-CN Mandarin ASR (8.23% CER on THCHS-30) (#476)
  • Add PocketTTS sessions with persistent KV-cache (#471)
  • Add PunctuationCommitLayer for punctuation-aware streaming ASR (#466)

Improvements

  • Refactor TDT decoder: Extract reusable components (#474)
  • ASR architecture cleanup: naming, dead code, file organization (#468)
  • Clarify custom vocabulary model compatibility and approach selection (#469)

Bug Fixes

  • Fix Swift 6 concurrency errors in SlidingWindowAsrManager (#472, #476)
  • Fix use-after-free when mic and system transcription run concurrently (#473)
  • Fix fatal error in levenshteinDistance with empty arrays (#476)

Documentation

  • Fix stale references in ASR documentation (#462)
  • Update Documentation index, remove espeak-ng licenses (#461)
  • Clean up CI workflows (#463, #464)

Full Changelog: v0.13.4...v0.13.5

Note: CTC zh-CN is experimental. API may change in future releases.

Don't miss a new FluidAudio release

NewReleases is sending notifications on new releases.