Skip to content

2.0.0 Performance Update

Choose a tag to compare

@CoffeeMethod CoffeeMethod released this 28 Jan 06:08
· 112 commits to main since this release

[v2.0.0] - Latest

2.0.0 is basically a entirely different program from the 1.0.0 release almost everything is better in this update. Finally got around to making it better, after I had to use it for another project. Hopefully it will be useful for you too.

Added

Document Support: Added ability to load and parse .pdf (via PyPDF) and .epub (via EbookLib) files directly.

Subtitle Generation: New feature to export .srt subtitle files synchronized with the generated audio.

Performance: Implemented parallel processing to utilize multiple CPU threads for faster generation.

Speed Control: Added audio speed adjustment slider (0.5x to 2.0x).

Text Cleaning: Integrated BeautifulSoup4 to strip HTML tags and formatting from EPUB imports.

Smart Splitting: Added configurable text splitting options (Newlines, Paragraphs, Sentences).

Changed

UI Overhaul: Updated graphical interface to support new configuration options and file loading.

Output Management: Enhanced file merging logic to handle larger batches of audio segments.

[v1.0.0] - Initial Release

Features

Text Input: Basic text field for direct input.

Voice Selection: Dropdown menu for standard Kokoro voices (af_alloy, af_bella, etc.).

Basic Output: Option to save as separate segments or a combined .wav file.

Timecodes: Customizable timestamp formatting for filenames.

Status Updates: Real-time feedback in the GUI during conversion.

Error Handling: Basic alerts for missing text or initialization failures.