v2.2.0: 听写兜底更稳、新增 OpenRouter 与 DeepSeek、词典全面生效 / Sturdier dictation, OpenRouter & DeepSeek, dictionary everywhere
·
0 commits
to b2265d8c5bd7c8b9e67df8f8f183acb714b69912
since this release
听写失败有了多层兜底,个人词典对所有云服务生效,还多了两个新服务商可选。
✨ 改进
- 听写不再轻易丢语音:百炼实时识别中断或失败时,自动改用一次性转写重试,仍失败再落到本地 Whisper——弱网或服务抖动时,说过的话不会无声消失。
- 个人词典全服务生效:词库里的人名、术语现在对 Soniox、AssemblyAI、Mistral 的听写同样生效(此前仅百炼),换服务商不用重新攒词。
- 文字整理会纠"听岔"的词:整理时结合整段上下文,把发音相近但明显不合语境的词改回正确写法;拿不准的词保持原样,不会过度改写。
- 百炼模型升级:默认识别模型升级至 Qwen Audio 3.1 消息版,识别结果与此前一致、费用约降 62%;另可在设置中选择 3.1 流式版。已有设置自动升级,无需操作。
- 切换服务商更省心:文字处理的模型名按服务商分别记忆——切走清空、切回自动恢复,不会再把一家的模型名带到另一家导致报错。
- 默认识别模型更新:OpenAI、Soniox、AssemblyAI 的新装用户默认使用各家当前推荐的识别模型;已自定义的模型不受影响。
🎉 新功能
- 新增 OpenRouter 听写服务:一个账号即可在 20 多家语音识别模型之间随意切换(默认为微软 MAI-Transcribe 2,中文表现优秀),无需逐家注册。
- 新增 DeepSeek 文字处理服务:整理与翻译可选用 DeepSeek(默认使用其快速模型),价格低、中文表现出色;提供与百炼同款的深度思考开关,默认关闭以保持整理速度。
Dictation now falls back instead of failing, the personal dictionary works with every cloud provider, and two new providers join the lineup.
✨ Improvements
- Your speech no longer vanishes on a hiccup: When Bailian's realtime recognition drops mid-dictation, the recording is automatically re-transcribed in one shot, falling back to local Whisper only if that fails too — network hiccups no longer swallow what you said.
- Personal dictionary works everywhere: Names and terms from your dictionary now also boost Soniox, AssemblyAI, and Mistral dictations (previously Bailian only); switching providers no longer means starting your vocabulary over.
- Cleanup fixes misheard words: Text cleanup now reads the whole transcript and corrects words that sound similar but clearly don't fit the context; uncertain words are left untouched, so nothing gets over-edited.
- Bailian model upgrade: The default recognition model moves to Qwen Audio 3.1 Message — identical transcripts at roughly 62% lower cost — with the 3.1 streaming model also selectable. Existing settings upgrade automatically.
- Smoother provider switching: The text-processing model is remembered per provider — switching away clears the field and switching back restores your model, so one vendor's model name never leaks into another's request.
- Fresh defaults: New OpenAI, Soniox, and AssemblyAI installs default to each vendor's current recommended recognition model; existing custom models are untouched.
🎉 New features
- OpenRouter dictation: One account, 20+ speech-recognition models to switch between at will (Microsoft MAI-Transcribe 2 by default, strong in Mandarin) — no per-vendor signups.
- DeepSeek text processing: Cleanup and translation can now run on DeepSeek (its fast model by default) — inexpensive and strong in Chinese, with a deep-thinking toggle mirroring Bailian's that stays off by default to keep cleanup fast.