Skip to content

v0.2.6

Choose a tag to compare

@holgern holgern released this 14 Aug 14:27
· 3 commits to main since this release

Full Changelog: v0.2.5...v0.2.6

[0.2.6] - 2026-08-13

Added

  • Added conservative acronym mode for context-aware initialism handling
  • Added contextual long-number mode for intelligent verbalization of 4+ digit numbers
  • Added registered acronym mode for controlling registered initialism expansion
  • Added bibliographic citation and reference sequence recognizers
  • Added decade rendering with natural English ordinal decade forms
  • Added duration recognition for HH:MM:SS timestamp formats
  • Added score recognition for sports results with contextual plausibility checks
  • Added chained score recognition for multi-set or multi-game results
  • Added software version recognition for named software packages
  • Added spaced ISBN label recognition for OCR or speech-recognized text
  • Added German Roman numeral title rendering for monarchs and popes
  • Added parenthesized initialism and ticker symbol recognition
  • Added superscript and Greek letter symbol rendering
  • Added additional Unicode fraction characters for fifths and sixths
  • Added compact license plate and vehicle code recognition
  • Added Google TN benchmark for English text normalization evaluation
  • Added benchmark ownership table with safety, extended, and quarantine gates
  • Added benchmark comparison identity validation and fresh report enforcement

Changed

  • Improved year detection with expanded temporal keyword context and bibliographic patterns
  • Improved formula recognition with balanced parenthesis context checking
  • Improved product code recognition with stronger label evidence requirements
  • Improved Roman numeral recognition with expanded context keywords for articles, acts, and scenes
  • Changed structured semantic precedence from regex iteration order to centralized precedence module

Fixed

  • Fixed fraction rendering to use natural word forms for fifths and sixths instead of Unicode fallback
  • Fixed formula recognition to reject fragments from unmatched parenthesized spans
  • Fixed slash fraction rendering to skip incomplete numerator or denominator matches
  • Fixed identifier separator handling to accept empty marker values
  • Fixed version literal detection to skip bare v-prefixed versions without contextual evidence
  • Fixed biology recognition to exclude temporal prepositions as false species markers
  • Fixed code recognition to classify vehicle-shaped codes with appropriate digit policies

Quality

  • Added Proteno diagnostic aggregates by rule, phase, ownership, and ambiguity family
  • Added PolyNorm ownership table documentation for comparison compatibility and quarantine policy