Skip to content

anymd 8.5.0

Choose a tag to compare

@github-actions github-actions released this 02 Oct 22:24
d80a930

Minor Changes

  • The default anymd binary is smaller: about 25 MB instead of 35 MB on Linux x64, with no change to anything it converts. The local VLM OCR engine moved out into its own small program, anymd-ocr-vlm (about 10 MB). anymd setup ocr downloads it in the same one-time step as the model weights; anymd checks a signature from a key built into anymd itself (bound to this exact version), downloads over HTTPS only, and checks the program's own version before installing it. --ocr vlm needs the same setup as before, and its output is unchanged. After an upgrade, anymd refreshes the 10 MB engine once by itself for anyone who already ran anymd setup ocr; the 2 GB model weights are never downloaded without asking. If you ask for VLM OCR before running anymd setup ocr, anymd says VLM OCR needs a one-time setup: run anymd setup ocr``; over MCP that is a normal reply rather than an error. Tesseract OCR and every other feature need no setup. Builds from source with --features ocr-vlm keep the engine inside the binary.

  • read now marks scanned PDF pages it could not read. A page with no text layer that images cover gets <!-- page N: scanned image, no text layer; enable OCR to read it: anymd setup ocr / ocr: true --> after its page marker, the front matter gains scanned_pages: [N, …], and a result made only of scans opens with a one-line hint, so agents can tell an empty page from an unread scan. Pages with text, and pages OCR reads, are unchanged.