v0.2.3: Major Extraction & Performance Improvements
·
110 commits
to master
since this release
What's Changed
🔧 Core Architecture Improvements
- Removed code formatting feature to significantly improve performance
- Eliminated markdown extraction component to reduce complexity and resource usage
- Simplified crawl configuration for better reliability
🤖 Enhanced Code Extraction
- Improved non-LLM code extraction with better HTML processing
- Added markdown content storage capabilities
- Enhanced language detection without external dependencies
🐳 Docker & Environment
- Consolidated Docker environment configuration to single .env file
- Streamlined setup process with unified configuration management
📊 UI Improvements
- Fixed bulk selection and deletion for filtered sources
- Enhanced table structure in CrawlJobs component
- Added domain prefix search to library search functionality
- Unified search interface for better user experience
🖼️ Performance Optimization
- Converted all screenshots from PNG to WebP format (68% space reduction)
- Updated README.md with improved documentation
Technical Highlights
- Performance: Code formatting removed for faster extraction times
- Reliability: Simplified architecture reduces point of failure
- User Experience: Enhanced search and bulk operations
- Documentation: WebP screenshots improve loading times
- Configuration: Single .env file simplifies deployment
Breaking Changes
- Code formatting feature has been removed - extraction is now raw HTML based
- Markdown extraction component eliminated - code extraction is simplified
Installation
This release maintains full backward compatibility for existing installations. Simply update to the latest version:
# For Docker users
cd codedox
git pull origin master
./docker-setup.sh # Rebuild with new configuration
# For manual installations
cd codedox
git pull origin master
source .venv/bin/activate
pip install -r requirements.txt
python cli.py init # Required schema migration
python cli.py serveMigration Notes
- Mandatory Migration Required: Users upgrading from v0.2.2 or earlier must run
python cli.py initto apply database schema changes - No API changes - existing setups should work after migration
- Code snippets are extracted directly from HTML without formatting overhead
- Search functionality remains fully compatible after migration
Database schema changes include:
- Updated tables for simplified crawl configuration
- Removal of formatting-related columns
- Optimization for new extraction workflow
Run python cli.py init to automatically apply required migrations.