Skip to content

LLMario 0.1.0

Choose a tag to compare

@mannysinghx mannysinghx released this 29 Sep 18:50

First release of LLMario: a free, open-source Mac app for chatting with open models (Qwen, Gemma, Llama, gpt-oss, Mistral and more) on your own computer. Nothing you type leaves your Mac.

Download

LLMario-0.1.0-macos-universal.dmg: open it and drag LLMario into Applications. The app is signed with a Developer ID and notarized by Apple, so it opens without warnings.

To verify the download, compare its SHA-256 with the .sha256 file:

shasum -a 256 -c LLMario-0.1.0-macos-universal.dmg.sha256

SHA-256: 572860c9400d11eb432228563f78ed9b45d9f4a0a2dfa9605d3c93e1ef51e183

Before you start: install an engine

LLMario runs models with the open-source engines llama.cpp and MLX-LM. The app does not include them yet, so install at least one:

  • llama.cpp (any Mac): brew install llama.cpp
  • MLX-LM (Apple Silicon, usually fastest): pip3 install mlx-lm

The app's Models → Engines tab shows which engines it found.

What's in 0.1.0

  • Chat with streaming replies, Markdown and code blocks, a collapsible Thinking section for reasoning models, and a Stop button that actually cancels generation.
  • Hardware-aware: picks the right engine for your Mac and checks that a model fits in memory before loading it.
  • Model library: 37 open-model families and 69 verified downloads, with search, task filters and "fits this computer". Every file is checked against Hugging Face checksums and pinned to an exact version.
  • Bring your own models: drag a .gguf file or an MLX model folder onto the window.
  • Private by default: works offline once a model is downloaded, has no accounts and no telemetry, and never writes messages to logs.
  • CLI and local API: build from source for the llmario command and an OpenAI-compatible API on 127.0.0.1 (see the README).

Requirements and status

  • macOS 13 or later.
  • Tested end to end on Apple Silicon (M4 Max) with both engines.
  • Intel Macs: included in the universal build (llama.cpp only). Not yet tested on real Intel hardware.

Website: https://llmario.com · Source: https://github.com/mannysinghx/LLMario · Report issues: https://github.com/mannysinghx/LLMario/issues