v0.2.6
Full Changelog: v0.2.5...v0.2.6
[0.2.6] - 2026-08-13
Added
- Added conservative acronym mode for context-aware initialism handling
- Added contextual long-number mode for intelligent verbalization of 4+ digit numbers
- Added registered acronym mode for controlling registered initialism expansion
- Added bibliographic citation and reference sequence recognizers
- Added decade rendering with natural English ordinal decade forms
- Added duration recognition for HH:MM:SS timestamp formats
- Added score recognition for sports results with contextual plausibility checks
- Added chained score recognition for multi-set or multi-game results
- Added software version recognition for named software packages
- Added spaced ISBN label recognition for OCR or speech-recognized text
- Added German Roman numeral title rendering for monarchs and popes
- Added parenthesized initialism and ticker symbol recognition
- Added superscript and Greek letter symbol rendering
- Added additional Unicode fraction characters for fifths and sixths
- Added compact license plate and vehicle code recognition
- Added Google TN benchmark for English text normalization evaluation
- Added benchmark ownership table with safety, extended, and quarantine gates
- Added benchmark comparison identity validation and fresh report enforcement
Changed
- Improved year detection with expanded temporal keyword context and bibliographic patterns
- Improved formula recognition with balanced parenthesis context checking
- Improved product code recognition with stronger label evidence requirements
- Improved Roman numeral recognition with expanded context keywords for articles, acts, and scenes
- Changed structured semantic precedence from regex iteration order to centralized precedence module
Fixed
- Fixed fraction rendering to use natural word forms for fifths and sixths instead of Unicode fallback
- Fixed formula recognition to reject fragments from unmatched parenthesized spans
- Fixed slash fraction rendering to skip incomplete numerator or denominator matches
- Fixed identifier separator handling to accept empty marker values
- Fixed version literal detection to skip bare v-prefixed versions without contextual evidence
- Fixed biology recognition to exclude temporal prepositions as false species markers
- Fixed code recognition to classify vehicle-shaped codes with appropriate digit policies
Quality
- Added Proteno diagnostic aggregates by rule, phase, ownership, and ambiguity family
- Added PolyNorm ownership table documentation for comparison compatibility and quarantine policy