github FluidInference/FluidAudio v0.14.6

latest releases: v0.17.5, v0.17.4, v0.17.3...
4 months ago

What's Changed

  • feat(tts): Supertonic-3 multilingual CoreML TTS in #617
  • refactor(tts): async StyleTTS2 predict + drop non-native Magpie synthesizeStream in #589
  • Make SpeakerManager a struct and de-async DiarizerManager by @panv-kw in #591
  • feat(tts/magpie): warmup API for cold-start mitigation (#60 Track 2) in #595* Fixed LS-EEND Memory Leak + Updated Docs by @SGD2718 in #605
  • fix(tts/pocket-tts): repair v1 voice cloning for pocket-tts 2.0.0 (#592) in #601
  • Timestamping RTTN decoder by @SGD2718 in #608
  • fix(asr/nemotron): seed cache_len=1 to avoid ios17.slice_by_index zero-shape warning (#607) in #609
  • fix(tts/pockettts): normalize French text and preserve mid-sentence chunks (#584) in #606
  • feat(asr): expose tdtConfig in SlidingWindowAsrConfig by @execsumo in #611
  • deprecate: remove CosyVoice3 and mono Kokoro (#571) in #602
  • Diarization progress by @nburns in #615

Full Changelog: v0.14.5...v0.14.6

Don't miss a new FluidAudio release

NewReleases is sending notifications on new releases.