Skip to content

v3.0.0 The UI & Polish Update

Choose a tag to compare

@CoffeeMethod CoffeeMethod released this 30 Jan 06:20
· 187 commits to main since this release
3b50740

[v3.0.0] - The UI & Polish Update

Version 3.0.0 represents a massive leap in user experience and audio control. I have completely replaced the old tkinter interface with CustomTkinter, giving the app a modern, sleek look with theme support. Under the hood, I've added granular control over the audio output and workflow improvements like Presets and Previews. And yes, it has been two days with two major overhauls in code, but I honestly barely know what I'm doing and I'm kind of just having fun with it, soo. Another New Program structure!!! Yay… 

Added

  • Modern UI Framework: Migrated the entire GUI to Customtkinter.
    • Added support for Dark, Light, and System themes.
    • Added UI Scaling options for high-DPI displays.
  • Audio Control: Added fine-tuning sliders for Volume and Pitch.
  • Post-Processing:
    • Added Normalize Audio option to ensure consistent volume levels.
    • Added Trim Silence option to remove dead air from the start/end of generated clips.
  • Presets System: You can now Save and Load your favorite voice, speed, and pitch configurations.
  • Audio Preview: Added a "Preview" button to generate a short sample before committing to a full render.
  • Enhanced Feedback: The status bar now includes estimated time remaining for large batch jobs.

Changed

  • Dependencies: Updated requirements to include Customtkinter.
  • Progress Tracking: Improved the accuracy of the progress bar during multi-threaded operations.

Installation

  1. Clone the repository:

    git https://github.com/CoffeeMethod/KokoroGUI.git
    cd KokoroGUI
  2. Create a virtual environment (recommended):

    python -m venv .venv
    # On Windows:
    .venv\Scripts\activate
    # On macOS/Linux:
    source .venv/bin/activate
  3. Install dependencies:

    pip install -r requirements.txt

    Note: If you have issues with torch, visit pytorch.org for specific installation instructions tailored to your OS and hardware.

Usage

  1. Run the application:

    python main.py
  2. Configure your conversion:

    • Choose your input method (Direct Text or Load File).
    • Select a voice from the dropdown menu.
    • (Optional) Adjust speed, volume, and pitch.
    • (Optional) Use Presets to save or load configurations.
    • Set your desired output directory and filename.
    • Choose whether to keep separate chunks or merge them into one file.
  3. Preview & Convert:

    • Click "Preview Audio" to hear a short sample of the current settings.
    • Click "Start Conversion" to begin the full process. You can monitor progress via the status label and progress bar.