Skip to content

v7.7.0

Latest

Choose a tag to compare

@github-actions github-actions released this 12 Aug 13:25
2815b70

7.7.0 (2026-08-12)

Flux TTS streaming controls and Listen v2 redaction.

Features

  • Speak v2 (Flux TTS streaming): barge-in via send_interrupt() (optional playback_offset, {type: "time_ms", value: N}), answered by a SpeechInterrupted server message whose metadata carries the new controls_applied.breaks_applied counter; mid-stream send_configure() to change speed, acknowledged by ConfigureSuccess / ConfigureFailure; new speed and expressivity connect query parameters. Inline pause and pronunciation controls are not applied at launch — they are stripped before synthesis and support is coming soon. (#758) (aab1eae)
  • Listen v2: redact connect parameter (ListenV2Redact: numbers, aggressive_numbers); send_configure() is now properly typed (ListenV2Configure + ListenV2ConfigureSuccess in the response union), replacing the previous typing.Any shim. (#758) (aab1eae)
  • Other: GoogleThinkProviderVersion adds ai-studio-v1beta and gemini-enterprise-agent-v1; AgentV1UpdateListenListenProvider discriminated union (_V1 / _V2, discriminant version); client_wrapper now derives its version from importlib.metadata rather than a hardcoded string. (#758) (aab1eae)

Compatibility

  • No breaking changes against v7.6.0: 0 removed public exports, 0 deleted modules, baseline socket-client signatures intact, and enum changes are widenings only.
  • The deepgram speak provider version widens from Literal["v1"] to str.
  • AgentV1UpdateListenListen.provider moves from a bare DeepgramListenProviderV2 to a required discriminated union; a compatibility validator coerces a legacy provider instance or bare dict into the new shape (both serialize to version: "v2"), so existing callers are unaffected.