Repository navigation
anymd 8.5.0
Minor Changes
-
The default
anymdbinary is smaller: about 25 MB instead of 35 MB on Linux x64, with no change to anything it converts. The local VLM OCR engine moved out into its own small program,anymd-ocr-vlm(about 10 MB).anymd setup ocrdownloads it in the same one-time step as the model weights; anymd checks a signature from a key built into anymd itself (bound to this exact version), downloads over HTTPS only, and checks the program's own version before installing it.--ocr vlmneeds the same setup as before, and its output is unchanged. After an upgrade, anymd refreshes the 10 MB engine once by itself for anyone who already rananymd setup ocr; the 2 GB model weights are never downloaded without asking. If you ask for VLM OCR before runninganymd setup ocr, anymd saysVLM OCR needs a one-time setup: runanymd setup ocr``; over MCP that is a normal reply rather than an error. Tesseract OCR and every other feature need no setup. Builds from source with--features ocr-vlmkeep the engine inside the binary. -
readnow marks scanned PDF pages it could not read. A page with no text layer that images cover gets<!-- page N: scanned image, no text layer; enable OCR to read it: anymd setup ocr / ocr: true -->after its page marker, the front matter gainsscanned_pages: [N, …], and a result made only of scans opens with a one-line hint, so agents can tell an empty page from an unread scan. Pages with text, and pages OCR reads, are unchanged.