Skip to content

v1.2.7

Choose a tag to compare

@mann1x mann1x released this 12 Jan 23:00
· 20 commits to master since this release

v1.2.7

  • Separate Best Answer Judge Model (--judgebest) - New command-line argument for best answer determination
    • Use a different model for best answer judgment vs similarity scoring (--judge)
    • --judgebest can be used alone or combined with --judge for different models
    • Same configuration options as --judge: local model name or http://host:port/model for remote
    • New system prompt focused purely on qualitative best answer determination
    • Supports both serial and parallel execution modes
    • Works with --rejudge to re-run only best answer judgment with new model
  • Version Tracking in QC Results - Record osync and Ollama versions in test results
    • OsyncVersion - Version of osync used for testing
    • OllamaVersion - Ollama server version for test quantizations
    • OllamaJudgeVersion - Ollama version for judge server (similarity scoring)
    • OllamaJudgeBestAnswerVersion - Ollama version for best answer judge server
    • Versions captured automatically from Ollama /api/version endpoint
  • QCView Output Updates - All output formats updated with new information
    • Table output shows Best Answer Judge model (when different from Judge) and versions
    • JSON output includes all version fields and JudgeModelBestAnswer
    • Markdown output includes Best Answer Judge and versions in header
    • HTML output shows Best Answer Judge in info grid and versions row
    • PDF output includes Best Answer Judge and versions in header tables
  • Manage TUI Multi-Select Delete Fix - Fixed batch delete for multiple selected models
    • Multi-selection delete now works correctly (previously only deleted single model)
    • Added batch confirmation dialog showing count and list of models to be deleted
  • Judge Retry Output Improvements - Better visibility into retry attempts during judgment
    • Both judge and judgebest operations now show retry warnings with error codes at each attempt
    • Displays retry delay countdown before each retry attempt
  • Fixed Copy to Remote Server - Resolved HTTP 500 errors when loading copied models
    • Fixed stop parameter serialization (now correctly sent as array instead of string)
    • Fixed numeric/boolean parameter type conversion (top_k, temperature, seed, etc.)
    • New ConvertParameterValue helper ensures correct JSON types for all Ollama model parameters
  • Fixed HuggingFace Model Copy - Correct path resolution for hf.co/... models
    • HuggingFace models now use correct manifest path (not under registry.ollama.ai)
    • Fixed cross-platform path separator handling for model paths with forward slashes
  • Load/Unload URL Format Support - Both commands now accept URL format with embedded model name
    • Supports osync load http://host:port/modelname in addition to osync load modelname -d host
    • Same URL parsing for unload command
  • Fixed --rejudge Model Pulling - Rejudge mode no longer attempts to download test models
    • When using --rejudge with existing results, only the judge model is needed
    • Wildcard expansion now filters to only tags present in the results file
    • Skips model verification for all existing results in rejudge mode
    • Properly queues partial results for re-judgment without resuming tests