Skip to content

Releases: parseh-io/Parseh

Parseh a0.5.0

Choose a tag to compare

@github-actions github-actions released this 07 Oct 09:42

Added

  • ⚙ page: one panel of settings on a book, a video, a document and its editor and an exercise deck, on a computer and on a phone
  • Page zoom, 70 to 200 %, for all of Parseh, also in the installed app
  • Levels have names, per language, which you can change and which follow you
  • Diacritics: put the vowel marks of Persian and Arabic away from a book's first two levels and a video's lines
  • Review later: flag a chunk or a phrase, then go to it, test yourself, copy the list or make its card; a hub door and a page for every flag

Changed

  • Plainer names on the page's own buttons: keep going (now remembered, and following you), between repeats, hide bars, highlight, keep in view, keep video, beside the text, kana only, pinyin only
  • One colour control on every page: the studio's sheet has no theme menu of its own
  • The foot of the hub, Settings and the guide: GitHub leads to Parseh's repository, and parseh.io is a link of its own beside it
  • The name Parseh is set in TeX Gyre Chorus in the bars and on the hub's title

Fixed

  • Two writes at once to the settings file no longer fail
  • Colours set on a phone's document page repaint its sheet

Parseh a0.4.4

Choose a tag to compare

@github-actions github-actions released this 06 Oct 05:25

Added

  • Parseh lives in the parseh-io organisation on GitHub; Settings → Updating Parseh looks in the new repository
  • The guide is at parseh.io/guide, under a bar that leads to Parseh's site; pages exported from Parseh link to it for good
  • The phone app's icons are published with the guide again
  • Every address Parseh writes for itself is kept in one place, lib/project.py
  • release.py links asks every address Parseh names and says what each answers
  • Settings → About shows a few friendly numbers: cards answered, videos and books on the shelves, since when
  • From a0.4.2: prompts of your own, the studio's prompt in parts, IPA and short-vowel options, skills for your chatbot, a book made by an agent from any device, pictograms (ARASAAC)
  • From a0.4.2: a video's vocabulary in the books' entries, notes on a video's phrases, a sound in place of a video, flashcards that turn
  • From a0.4.3: Whisper models for each language, optional exact word times kept through every edit of a transcript
  • From a0.4.3: a transcript workspace with LM likelihood, an LLM connection, a reasoning workspace and chatbot copy-and-paste review
  • From a0.4.3: partial-word colours in the studio, mobile arrow keys that skip, a card maker for books and videos

Changed

  • Speech to text lists the Whisper program first, above the second pass and the models
  • a0.4.2 and a0.4.3 were finished but not released on their own: their full entries follow

Fixed

  • LLM Integration, LM likelihood and About are drawn with their own stylesheet and end in a foot like the other doors
  • The README and the settings pages link to guide pages that exist

[a0.4.3] - 2026-10-05

Added

  • A standalone reasoning-workspace CSV checker, shipped with the helper and skill, for draft checks, exact source spans and required-entry coverage
  • Standard and language-specific Whisper models in one pinned catalogue, with independent resumable installs, per-language model choices and model provenance in transcription jobs
  • Optional Hugging Face downloads for Persian fast/accuracy, Arabic, Italian, Hindi fast and Spanish, including merged LoRA models, matching offline assets and original licence notices
  • Settings → About with version, author, launch-script locations and installation information using the host's native paths
  • Saved pending transcription reviews, with pause/resume from Videos across server and browser restarts
  • Checked-word locks, selected-section review, single-word LM likelihood and bulk selection of the latest method's best choices
  • Optional similar-sound and candidate search-probability filters for additional model-derived words; every genuine Whisper alternative remains eligible
  • LM likelihood review with installed-model selection, raw-token candidate search, numerical rankings and an isolated CPU/GPU scoring worker
  • External chatbot copy/paste review for suspect words, whole text and reasoning workspaces, with bounded prompts, validated answer import and downloadable text/CSV workspace files
  • Additional reasoning workspace review with numbered transcript blanks, CSV evidence, isolated Python tools, separate model selection and two-pass word progress; existing review modes and thresholds stay unchanged
  • Browser transcription workspace with video beside the transcript, word/caption replay, ±1/2/5-second playback controls, caption filters and a persistent word editor
  • Brief word playback with automatic pause, installed-dictionary suspect flags and original/replacement meanings in browser transcript review
  • Dictionary meanings for native Whisper alternatives alongside their recognition scores in the word inspector
  • Native faster-whisper beam alternatives mapped to source words, with sequence log scores kept separate from word probabilities
  • Browser Settings → LLM Integration: a reusable OpenAI-compatible endpoint, host-local credentials, model discovery and Unsloth run-settings link import
  • Review-first Whisper results with short sentence LLM replies, word-count progress, local failure recovery, retry of unresolved words, keyboard edits and inspectable responses
  • Remote installed-model profiles from Studio links, with explicit API loading of quantization, KV cache, context and vision options
  • Separate models, prompts and downloadable skills for suspect-word and whole-text review, including confident words and contiguous ASR spans
  • Individual pending draft decisions and explicit Use this transcript preserve the existing box and timing guards
  • Optional, local exact word times for every speech-to-text language, with hash-pinned int8 CTC networks and a remembered add-page switch
  • A Parseh-made transcript carries its word timing through the add-page editor and stores it with the matching video
  • Two transcript tidy cues: the existing text cues and recorded pauses
  • Mobile-mode keyboard arrows skip narration or video by the selected amount
  • Partial-word foreground colours in Studio, with selection-first editing, Unicode-safe boundaries, one-word linguistic clouds, PDF and standalone-HTML output

Changed

  • Transcript word editors can include adjacent words for a short-span correction; reasoning workspaces can join split words across suspect/confident neighbors
  • Transcription opens directly into a shared editing workspace; correction methods are reusable tools, with optional automatic and later per-word/section Whisper rechecks
  • All four transcript-correction methods are labeled experimental; LM likelihood offers no-cutoff, suggested 0.1 and custom cutoff choices
  • LLM settings and transcript review show simpler controls, with technical setup and numerical diagnostics folded away
  • Markdown and exercise-deck cards made from books or videos mark target and context fields as the target language automatically
  • Studio and exported word clouds are linguistic only; foreground colour is applied from the source editor

Fixed

  • Speech program and model downloads no longer fail before their progress starts
  • Exported guides include the segmented-colour parser needed for independent builds
  • Speech model notices list every catalogue source and converted package with its own licence
  • Persian dictionary lookup recognizes colloquial copulas and combined nominal endings, with direct entries preferred and dictionary-backed stem guards
  • LM likelihood candidate search preserves source whitespace with SentencePiece tokenizers, including PersianMind
  • Partial-word colours compile and remain visible in PDFs, including headings, target-language blocks and exercise prompts

[a0.4.2] - 2026-09-30

Added

  • Glosses an LLM writes are aligned to their chunks: a meaning says what its own words say, in the text's order
  • Every prompt is assembled by one kit from its instructions, its answer contract and its data, and opens with a version line
  • The LLM row: one control for every prompt, saying its size before the copy, on the studio, the exercise dialog, the add page, the tidy, the player, the reader and Ask LLM
  • A video's vocabulary in the books' entries, asked of the chatbots and drawn as in the reader, with macro buttons that say what they write
  • A book made by an agent in place, from any device, its text given in parts, watched from the library, steered through asks, finished with one button
  • Prompts of your own, from a menu at every copy button, exported and imported under Settings → Your prompts
  • The studio's prompt in parts you tick, with presets, a level and a length, and the exercise dialog pre-ticking what the page uses
  • Options of a prompt: the transliteration in IPA, and the short vowels of Persian and Arabic
  • Every language's conventions carry a worked example and one vocabulary convention
  • Skills for your chatbot: gloss, markdown and book, built on request and downloaded under Settings → Skills
  • A sound in place of a video, with its waveform
  • A video phrase's note can be edited, and a ✱ notes button lights every phrase that has one
  • A vocabulary or opposites flashcard's example, notes and source go on the side the card turns to
  • Settings → Pictograms (ARASAAC): the symbols and their words fetched on request, with their licence

Changed

  • The LLM prompt page's Edit prompt, Save custom prompt and Reset to default give way to the prompt menu

Parseh a0.4.1

Choose a tag to compare

@github-actions github-actions released this 29 Sep 03:31

Added

  • Speech to text, optional and local: a transcript made on this computer while adding a video, from a film on this machine or a YouTube video recorded through the tab
  • Settings → Speech to text: the program, two Whisper models and the processor, open to any device let in
  • The graphics card is used only once Parseh has proved it works; the CPU always does
  • A YouTube video recorded for its transcript keeps that recording's waveform when it is added
  • Hover ⏸ in a book's reader, in both interfaces: the narration waits while a gloss is open
  • Edit times by ear in a book's header, and a video's timings, open where the listener is

Changed

  • An update, and going back, leave the speech program and its models as they are (stt/)
  • Typing in the transcript, the address or the film's path makes a prepared LLM prompt stale
  • The offline page's Try again says that it tried, and Go back is a button of its own

Fixed

  • A page not kept on the device opens from a computer that is only slow, instead of "Parseh cannot be reached"
  • Stopping the tab share after a waveform was drawn no longer pauses the video or resets its speed
  • The transcript editor saves the draft as it hands the transcript back

Parseh a0.4.0

Choose a tag to compare

@github-actions github-actions released this 28 Sep 10:21

Added

  • LaTeX drawings: a ::::latex block, or [...]{latex} inline, compiled by LaTeX itself, beside the formulas MathJax draws
  • The same drawing on screen, in the PDF, in an HTML export and on a phone; a live preview kept only once it is saved
  • A drawing that cannot be made shows its source and one line saying why, with the button that mends it
  • The sheet that draws a block: its preview filling its pane beside the source, a caption, size and position folded away
  • Settings → LaTeX drawings, styled like the rest of Settings: named themes of packages, languages, a font, a preamble and a compiler
  • Themes exported and imported as files; a rename rewrites every block that names the theme
  • A theme's packages beyond the base kept in Parseh's own folder, listed with the base, asked about and gotten one at a time
  • More packages to pick from: syntax trees and dependency arcs, the IPA, pinyin over Chinese characters
  • A deck's exercises filtered by where they came from, beside its text and tag filters
  • Flashcards that are asked both ways: both (random) and both (repeat), the second two linked cards in a deck
  • The transliteration cloud works on an exported page and on the guide's own words
  • On a phone's video, the subtitles have a text size of their own, reached in full-screen too

Changed

  • Every Settings page starts with the bar of its doors, Network included
  • Deleting a document asks in the studio's own window, and the card leaves the list at once
  • Documents, decks, bundles and shelves are stored in a new shape: going back to a0.3.3 says so first
  • The size sliders of books and videos reach further, at both ends

Fixed

  • Exporting a page to HTML from another device could land on "Parseh cannot be reached": it shows a bar now

Parseh a0.3.3

Choose a tag to compare

@github-actions github-actions released this 25 Sep 18:33

Added

  • Estimate the next N seconds, beside estimate the rest

Parseh a0.3.2

Choose a tag to compare

@github-actions github-actions released this 25 Sep 11:18

Added

  • Estimate the rest by the sound: boundaries placed in the pauses of the waveform, in books and videos
  • Update Parseh from Settings, from GitHub or a zip, to any version, keeping your content and settings
  • Releases on GitHub, one zip per version; Parseh shows its version on the hub and in Settings
  • Downloads show their size, progress and time left, can be stopped, and resume where they stopped
  • On a phone, the dictionary opens in a sheet, and sparsely glossed books and videos mark their glossed chunks

Changed

  • Dictionaries and the other downloads moved to Settings → Reading help, laid out by language
  • "Offline" only when the computer is really gone, with ↻ to check again
  • Who may change a setting is decided per setting: Network and updates on the computer only
  • Languages you add are kept in config/
  • Exported pages open on sepia

Fixed

  • Pages of other websites could make Parseh change things
  • A phone could read the Wi-Fi pairing code
  • The test suites wrote into config/