EchoSub 1.0.2
EchoSub shows live, translated captions for anything your PC plays — videos, streams, calls, games. It listens to your speakers, recognizes speech in almost 100 languages, translates it on your own GPU, and shows it in a customizable always-on-top caption box, with a line and a color for every speaker.
What's new in 1.0.2
- Zoom the caption box: hold Shift and turn the mouse wheel over the box to make the box and its text bigger or smaller (50–300 %). The value is also in Settings → Position & Alignment → Box style → Size (zoom).
- Google Translate limits handled: when Google answers "too many requests", EchoSub spaces out and retries requests, pauses Google for a while after repeated refusals, and meanwhile translates offline with an NLLB model you have already downloaded. It switches back to Google automatically.
- Translation errors now appear as a short message that disappears after a few seconds, instead of a long message that stayed in the caption box.
Download
EchoSub-Setup-1.0.2.exe (973 MB) — Windows 10/11, 64-bit
SHA-256: 554420654BA134784D35A49B1B9C774BED8CCDA2E1BE4C0F652FF34EB27FC3D6
- Close EchoSub if it is running, then run the installer (no admin rights needed; you can also install for all users). Installing over an earlier version keeps your settings and models.
- Start EchoSub. The first start downloads the AI models (~2.3 GB) in a progress window — you can cancel and it resumes later.
- Play anything with speech. Right-click the caption box or the tray icon for the menu.
Windows SmartScreen: the installer isn't code-signed yet, so Windows may show "Windows protected your PC". Click More info → Run anyway. You can verify the file with the SHA-256 above:
Get-FileHash EchoSub-Setup-1.0.2.exe.
System requirements
| Minimum | Recommended | |
|---|---|---|
| Operating system | Windows 10, 64-bit | Windows 11, 64-bit |
| Processor | 64-bit, 4 cores | 6+ cores (e.g. Intel Core i5-9400F / AMD Ryzen 5 3600) |
| Memory (RAM) | 8 GB | 16 GB |
| Graphics | None required — CPU with the Small or Base speech model (captions lag several seconds) | NVIDIA GeForce GTX 1060 6 GB or better |
| Graphics memory | 4 GB (default models on an NVIDIA GPU) | 6 GB or more |
| NVIDIA driver | 527.41 or newer | Latest |
| Free disk space | 6 GB | 10 GB on an SSD |
| Internet | Once, to download the models (~2.3 GB) | Also for the optional Google Translate engine |
Highlights
- Any app, any language — WASAPI loopback capture, Whisper large-v3-turbo recognition with automatic language detection (~100 languages).
- On-device translation — NLLB-200 (600M / 1.3B), or Google Translate. The original text appears instantly and the translation follows a moment later.
- Speakers — each voice gets its own line and color.
- Arabic — Moroccan Arabic (Darija) support and optional diacritics (تشكيل).
- Caption box your way — fonts, colors and spacing, language labels with flags, 9 screen positions or drag anywhere, autosize, corner radius, padding, smooth slide / fade animations.
- Handy — global hotkeys (
Ctrl+Alt+H/P/L), caption history and transcripts, click-through lock, auto-hide when nobody talks, cancelable model downloads, self-recovery when an audio device disappears.
| Any language in, your language out | Arabic diacritics (تشكيل) |
|---|---|
![]() |
![]() |
| Settings | About |
|---|---|
![]() |
![]() |
More screenshots in the README · full changelog · user guide.
License
Copyright © 2026 Mohammad Al-Safadi — GPL-3.0.
The default offline translation model (NLLB-200) is licensed for non-commercial use only; see third-party notices.




