Skip to content

v0.6.2

Choose a tag to compare

@github-actions github-actions released this 24 Jul 05:19
· 33 commits to main since this release

0.6.2

  • Updated the available Gemini models, with Gemini 3.6 Flash as the new default and Gemini 3.5 Flash-Lite as a low-cost option.

0.6.1

  • Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
  • Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.

0.6.0

  • Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
  • Added Markdown rendering for supported recognition and image description results.
  • Added follow-up question support for conversational image description engines.
  • Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
  • Improved engine settings handling and result presentation reliability.

SHA256:
fd025f5a18c921703ff789b9c45e0d27c57c70b0fdc756709ae92062f44fff8d visAware-0.6.2.nvda-addon