Skip to content

VRCS 0.1.2

Choose a tag to compare

@Dreaminko Dreaminko released this 20 Aug 12:55
· 90 commits to main since this release

VRCS 0.1.2

This release expands multilingual translation workflows and subtitle interaction, while improving recognition, OSC output, audio capture, SteamVR integration, and release reliability.

Highlights

  • Added Ask AI for selected subtitle text. Ask a custom question or use quick prompts for meaning, grammar, and phrasing.
  • Added up to three ordered translation routes for both microphone and speaker audio, with independent target languages, API profiles, and models.
  • Added reusable language presets for recognition, translation routes, and OSC output preferences.
  • Added a dedicated Glossaries settings section shared by LLM translation and supported ASR services.
  • Added a glossary table editor with search, filtering, inline editing, multi-selection, and bulk actions.
  • Added three OSC multilingual output strategies:
    • Preferred language only
    • Round robin
    • All languages in one message
  • Added an option to send translated text without including the original subtitle.

Improvements

  • Recognition model lists can now load compatible versioned Qwen Realtime and FunASR models dynamically.
  • OpenAI Realtime transcription now uses the model selected in settings.
  • Speaker and microphone capture pipelines now start in parallel for faster initialization.
  • VR overlay reconnection now waits until SteamVR is running.
  • Explicit manual Chatbox messages can be sent while mute synchronization is blocking automatic output.
  • Existing translation routes and glossary settings are migrated to the new configuration format automatically.
  • Improved release validation, updater configuration, and CUDA architecture checks.

Fixes

  • Suppressed repeated VRCX world-name echoes in partial and final recognition results.
  • Prevented duplicate final ASR events for the same utterance.
  • Improved OSC revision handling to discard stale automatic output after configuration or mute-state changes.
  • Fixed unbalanced WASAPI COM initialization during audio device discovery and capture.
  • Improved rollback behavior when one of the audio capture pipelines fails to start.

Compatibility note

The CUDA edition now requires the CUDA 13.x Runtime, cuBLAS, and a compatible NVIDIA driver. The standard edition does not require CUDA.


本次更新扩展了多语言翻译流程与字幕交互能力,并改进了语音识别、OSC 输出、音频采集、SteamVR 集成和发布可靠性。

主要更新

  • 新增字幕划词 “问 AI” 功能,可自由提问,也可快速查询词义、语法和表达方式。
  • 麦克风与扬声器翻译现在分别支持最多三条有序翻译路线,每条路线可独立设置目标语言、API 配置和模型。
  • 新增语言预设,可保存识别语言、翻译路线和 OSC 输出偏好。
  • 新增独立的 “术语表” 设置页面,术语可同时供 LLM 翻译和支持上下文的 ASR 服务使用。
  • 新增术语表格编辑器,支持搜索、筛选、行内编辑、多选和批量操作。
  • 新增三种 OSC 多语言输出策略:
    • 仅输出首选语言
    • 按语言轮换输出
    • 在同一条消息中输出所有语言
  • 新增仅发送译文、不附带原文的选项。

改进

  • Qwen Realtime 与 FunASR 现在可以动态加载兼容的版本化识别模型。
  • OpenAI Realtime 转写现在会使用设置中选择的模型。
  • 扬声器与麦克风采集管线改为并行启动,缩短初始化时间。
  • VR 覆盖层会等待 SteamVR 启动后再尝试连接。
  • 静音同步阻止自动输出时,用户主动发送的 Chatbox 消息仍可正常发送。
  • 旧版翻译和术语表配置会自动迁移到新的配置格式。
  • 改进发布验证、自动更新配置和 CUDA 架构检查。

修复

  • 过滤 VRCX 上下文造成的世界名称重复识别结果。
  • 防止同一段语音产生重复的最终识别事件。
  • 改进 OSC 修订状态处理,配置或静音状态变化后不再发送过期的自动输出。
  • 修复音频设备查询和采集过程中的 WASAPI COM 初始化不平衡问题。
  • 改进音频管线启动失败时的回滚处理。

Full Changelog / 完整修改记录