Skip to content

v0.5.0 — Parakeet backend, faster toggles, overlay motion

Choose a tag to compare

@acailic acailic released this 04 Sep 03:13
· 80 commits to linux since this release

SayItErmano — community Linux port of FluidVoice

Local voice dictation, rebranded and its own app: 23 commits since v0.4.0, headlined by a second speech-recognition engine and a snappier hotkey.

New: NVIDIA Parakeet TDT backend

  • parakeet-tdt-0.6b-v2 runs locally via ONNX Runtime alongside the faster-whisper backends — pick it in Settings → Model (backend = parakeet), auto-selection still prefers whisper on CUDA unless you choose Parakeet.
  • Real-audio integration fixtures and a factory-run test suite came with it.

Faster hotkeys

  • Toggle-on dropped from a flat ~350 ms to ~100 ms: the recorder now polls for flowing PCM instead of sleeping a fixed fail-fast window.
  • First dictation after daemon start no longer pays a ~390 ms CUDA warm-up tax — warmup runs one throwaway inference at startup (measured ~170 ms, same as steady state).

Overlay & history polish

  • Pill motion science: fade-in, done-beat, elapsed cue, and reduced-motion support.
  • History gains confidence bands (solid/mixed/shaky), date grouping, and inline repair.

Also

  • Environment overrides renamed FLUIDVOICE_*SAYITERMANO_* (no aliases — update scripts/custom units).
  • Original app icon across all sizes, symbolic icons refreshed.
  • Test suite: 620+ tests green; CI stays manual-trigger only.

Install

curl -fsSL https://raw.githubusercontent.com/acailic/SayItErmano/linux/scripts/install-one-shot.sh | bash

Or grab sayit-ermano_0.5.0-1_amd64.deb below and sudo apt install ./….

sha256  b1f46555a468ee7dd9cb4f35cd54e3b66fbcf4ec30d7d4f8e3107032f2cfb13f