Releases: cary-rowen/visAware
Releases · cary-rowen/visAware
Release list
v0.8.1
0.8.1
- Added Kimi image description, follow-up questions, OCR with structured coordinates, and AI Agent support through the OpenAI-compatible Kimi API.
- Improved image-description prompts across supported engines with localized defaults.
- Improved mathematical image descriptions by converting visible formulas to LaTeX and returning formula-only images as LaTeX formulas.
- Improved automatic image recognition with concise 20 to 30 word descriptions without Markdown.
- Improved browsable image-description output by using standard Markdown tables for tabular content and avoiding code fences.
- Removed model and model-provider names from the follow-up dialog.
- Added Copy and Close buttons to browsable recognition and follow-up result dialogs.
0.7.0
- Added the Apple Vision (OCR Server) engine for local-network OCR through the open-source OCR Server iOS app.
0.6.5
- Improved settings and follow-up dialog layouts, including high-DPI scaling.
- Improved editing of long custom and automatic-recognition prompts.
- Improved follow-up questions so initial descriptions and latest answers can be viewed as formatted content.
0.6.4
- Improved formula rendering in HTML and PaddleOCR results.
- Fixed AI Agent text input and automatic recognition on English interfaces.
- Improved English and Simplified Chinese UI text and documentation.
0.6.2
- Updated the available Gemini models, with Gemini 3.6 Flash as the new default and Gemini 3.5 Flash-Lite as a low-cost option.
0.6.1
- Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
- Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.
0.6.0
- Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
- Added Markdown rendering for supported recognition and image description results.
- Added follow-up question support for conversational image description engines.
- Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
- Improved engine settings handling and result presentation reliability.
SHA256:
98b571c3e815a67be79faa97be2a3d38fc8c2eccbfee6ffa5417a99a9be83d2c visAware-0.8.1.nvda-addon
v0.7.0
0.7.0
- Added the Apple Vision (OCR Server) engine for local-network OCR through the open-source OCR Server iOS app.
0.6.5
- Improved settings and follow-up dialog layouts, including high-DPI scaling.
- Improved editing of long custom and automatic-recognition prompts.
- Improved follow-up questions so initial descriptions and latest answers can be viewed as formatted content.
0.6.4
- Improved formula rendering in HTML and PaddleOCR results.
- Fixed AI Agent text input and automatic recognition on English interfaces.
- Improved English and Simplified Chinese UI text and documentation.
0.6.2
- Updated the available Gemini models, with Gemini 3.6 Flash as the new default and Gemini 3.5 Flash-Lite as a low-cost option.
0.6.1
- Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
- Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.
0.6.0
- Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
- Added Markdown rendering for supported recognition and image description results.
- Added follow-up question support for conversational image description engines.
- Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
- Improved engine settings handling and result presentation reliability.
SHA256:
b8537ad41abe882dda8f0c5b65b336b58bf7774b8af7fa072013a4aa43b2d9a0 visAware-0.7.0.nvda-addon
v0.6.5
0.6.5
- Improved settings and follow-up dialog layouts, including high-DPI scaling.
- Improved editing of long custom and automatic-recognition prompts.
- Improved follow-up questions so initial descriptions and latest answers can be viewed as formatted content.
0.6.4
- Improved formula rendering in HTML and PaddleOCR results.
- Fixed AI Agent text input and automatic recognition on English interfaces.
- Improved English and Simplified Chinese UI text and documentation.
0.6.2
- Updated the available Gemini models, with Gemini 3.6 Flash as the new default and Gemini 3.5 Flash-Lite as a low-cost option.
0.6.1
- Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
- Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.
0.6.0
- Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
- Added Markdown rendering for supported recognition and image description results.
- Added follow-up question support for conversational image description engines.
- Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
- Improved engine settings handling and result presentation reliability.
SHA256:
20d8d77ecba99ca7fdc7569f6db93b41a0a677f80979f75d31916f26b9a24fc4 visAware-0.6.5.nvda-addon
v0.6.4
0.6.4
- Improved formula rendering in HTML and PaddleOCR results.
- Fixed AI Agent text input and automatic recognition on English interfaces.
- Improved English and Simplified Chinese UI text and documentation.
0.6.2
- Updated the available Gemini models, with Gemini 3.6 Flash as the new default and Gemini 3.5 Flash-Lite as a low-cost option.
0.6.1
- Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
- Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.
0.6.0
- Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
- Added Markdown rendering for supported recognition and image description results.
- Added follow-up question support for conversational image description engines.
- Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
- Improved engine settings handling and result presentation reliability.
SHA256:
079e8174ab3d413ecd7076f4fc729d11266b5eab477cb80ddd2e3e50c5862ac9 visAware-0.6.4.nvda-addon
v0.6.2
0.6.2
- Updated the available Gemini models, with Gemini 3.6 Flash as the new default and Gemini 3.5 Flash-Lite as a low-cost option.
0.6.1
- Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
- Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.
0.6.0
- Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
- Added Markdown rendering for supported recognition and image description results.
- Added follow-up question support for conversational image description engines.
- Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
- Improved engine settings handling and result presentation reliability.
SHA256:
fd025f5a18c921703ff789b9c45e0d27c57c70b0fdc756709ae92062f44fff8d visAware-0.6.2.nvda-addon
v0.6.1
0.6.1
- Removed the separate AI Agent start/stop command from Input Gestures; use the main Vis Aware command in Agent mode.
- Updated the English and Simplified Chinese documentation for commands, automatic recognition, settings, data handling, and the recommended Gemma 4 setup through Ollama.
0.6.0
- Added PaddleOCR / PaddleOCR-VL and Google Gemma engines.
- Added Markdown rendering for supported recognition and image description results.
- Added follow-up question support for conversational image description engines.
- Added per-engine enable / disable controls so unused engines can be hidden from normal use and engine cycling.
- Improved engine settings handling and result presentation reliability.
SHA256:
e095bcc7686926cb18e2f51b99f4ce618271db1e852366cc1869fac4f8e73da7 visAware-0.6.1.nvda-addon