Skip to content

v0.0.23

Latest

Choose a tag to compare

@ivan-digital ivan-digital released this 18 Jul 06:44
c1aa219

Highlights

  • Adds an OpenAI-compatible POST /v1/audio/speech endpoint to speech-server, supporting WAV and raw PCM output for drop-in local TTS integrations.
  • Adds incremental Sortformer streaming, Nemotron emission-aligned word timestamps, and the ReDimNet2-B6 speaker identity encoder.
  • Improves cached Whisper model startup.

What's Changed

Full Changelog: v0.0.22...v0.0.23