Repository navigation
VRCS 0.1.2
VRCS 0.1.2
This release expands multilingual translation workflows and subtitle interaction, while improving recognition, OSC output, audio capture, SteamVR integration, and release reliability.
Highlights
- Added Ask AI for selected subtitle text. Ask a custom question or use quick prompts for meaning, grammar, and phrasing.
- Added up to three ordered translation routes for both microphone and speaker audio, with independent target languages, API profiles, and models.
- Added reusable language presets for recognition, translation routes, and OSC output preferences.
- Added a dedicated Glossaries settings section shared by LLM translation and supported ASR services.
- Added a glossary table editor with search, filtering, inline editing, multi-selection, and bulk actions.
- Added three OSC multilingual output strategies:
- Preferred language only
- Round robin
- All languages in one message
- Added an option to send translated text without including the original subtitle.
Improvements
- Recognition model lists can now load compatible versioned Qwen Realtime and FunASR models dynamically.
- OpenAI Realtime transcription now uses the model selected in settings.
- Speaker and microphone capture pipelines now start in parallel for faster initialization.
- VR overlay reconnection now waits until SteamVR is running.
- Explicit manual Chatbox messages can be sent while mute synchronization is blocking automatic output.
- Existing translation routes and glossary settings are migrated to the new configuration format automatically.
- Improved release validation, updater configuration, and CUDA architecture checks.
Fixes
- Suppressed repeated VRCX world-name echoes in partial and final recognition results.
- Prevented duplicate final ASR events for the same utterance.
- Improved OSC revision handling to discard stale automatic output after configuration or mute-state changes.
- Fixed unbalanced WASAPI COM initialization during audio device discovery and capture.
- Improved rollback behavior when one of the audio capture pipelines fails to start.
Compatibility note
The CUDA edition now requires the CUDA 13.x Runtime, cuBLAS, and a compatible NVIDIA driver. The standard edition does not require CUDA.
本次更新扩展了多语言翻译流程与字幕交互能力,并改进了语音识别、OSC 输出、音频采集、SteamVR 集成和发布可靠性。
主要更新
- 新增字幕划词 “问 AI” 功能,可自由提问,也可快速查询词义、语法和表达方式。
- 麦克风与扬声器翻译现在分别支持最多三条有序翻译路线,每条路线可独立设置目标语言、API 配置和模型。
- 新增语言预设,可保存识别语言、翻译路线和 OSC 输出偏好。
- 新增独立的 “术语表” 设置页面,术语可同时供 LLM 翻译和支持上下文的 ASR 服务使用。
- 新增术语表格编辑器,支持搜索、筛选、行内编辑、多选和批量操作。
- 新增三种 OSC 多语言输出策略:
- 仅输出首选语言
- 按语言轮换输出
- 在同一条消息中输出所有语言
- 新增仅发送译文、不附带原文的选项。
改进
- Qwen Realtime 与 FunASR 现在可以动态加载兼容的版本化识别模型。
- OpenAI Realtime 转写现在会使用设置中选择的模型。
- 扬声器与麦克风采集管线改为并行启动,缩短初始化时间。
- VR 覆盖层会等待 SteamVR 启动后再尝试连接。
- 静音同步阻止自动输出时,用户主动发送的 Chatbox 消息仍可正常发送。
- 旧版翻译和术语表配置会自动迁移到新的配置格式。
- 改进发布验证、自动更新配置和 CUDA 架构检查。
修复
- 过滤 VRCX 上下文造成的世界名称重复识别结果。
- 防止同一段语音产生重复的最终识别事件。
- 改进 OSC 修订状态处理,配置或静音状态变化后不再发送过期的自动输出。
- 修复音频设备查询和采集过程中的 WASAPI COM 初始化不平衡问题。
- 改进音频管线启动失败时的回滚处理。