Skip to content

2.6.5

Choose a tag to compare

@KennethEnevoldsen KennethEnevoldsen released this 05 Jan 09:12
· 1136 commits to main since this release

2.6.5 (2026-01-05)

Documentation

  • docs: Fix docs build strict mode errors (#3809)

  • fix: resolve mkdocs strict mode errors

  • fix: remove duplicate line in installation.md

  • build: add --strict flag to mkdocs build

  • fix: resolve invalid BibTeX keys in task citations

  • feat: filter BibTeX warnings in strict docs build

  • Update docs/usage/defining_the_model.md

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • fix: dynamic mkdocs path discovery for CI

  • fix: improve docs build script with clear warning counts

  • fix: resolve 6 real docs build warnings

  • fix: remove broken PylateSearchEncoder reference

  • Remove unused build scripts

  • docs: wrap multimodal example as code snippet

  • fix: export SklearnModelProtocol for docs API

  • docs: add API reference for SklearnModelProtocol

  • fix: remove SklearnModelProtocol export to avoid circular import

  • feat: add SklearnModelProtocol docs with lazy import

  • fix: convert Sphinx cross-references to MkDocs syntax for proper linking

Convert :class: Sphinx syntax to [Text][module.path] MkDocs syntax to ensure
cross-references are properly clickable in the generated documentation.

🤖 Generated with Claude Code

Co-Authored-By: Claude Sonnet 4 <noreply@anthropic.com>

  • chore: remove SklearnModelProtocol docs to avoid circular import

  • chore: remove SklearnModelProtocol export from _evaluators

  • style: fix indentation in _evaluators/init.py

  • fix: enable mkdocs build --strict without warnings

🤖 Generated with Claude Code

Co-Authored-By: Claude Sonnet 4 <noreply@anthropic.com>

  • fix: rename evaluator to multilabel_classifier to avoid type conflict

  • fix: use evaluator_model instead of evaluator to avoid type conflict

  • Change parameter name from multilabel_classifier to evaluator_model
  • Maintains SklearnModelProtocol type hint as requested in PR review
  • Resolves mypy type error by using parent class's evaluator_model field
  • Keeps explicit protocol reference in docstring for clarity

🤖 Generated with Claude Code

Co-Authored-By: Claude Sonnet 4 <noreply@anthropic.com>


Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>
Co-authored-by: Claude Sonnet 4 <noreply@anthropic.com> (3723d27)

Fix

  • fix: Extend framework annotations for ModelMeta (#3819)

  • Update framework and filter based on them

  • update ModelMeta of models

  • update ModelMeta of models

  • update ModelMeta of models

  • update ModelMeta of models

  • update ModelMeta of models

  • update ModelMeta of models

  • update ModelMeta

  • add csv

  • update ModelMeta

  • added framework to ModelMeta

  • update ModelMeta of models

  • update ModelMeta of models

  • update framework in ModelMeta of models

  • update framework

  • update framework

  • update framework in ModelMeta

  • fix tests

  • Add models

  • fix tests

  • add tags extraction in from_hub()

  • fix typecheck

  • apply suggestions

  • apply suggestions

  • keep only static method

  • delete csv and script (d033c24)

Unknown

  • fix dataset generation tags (#3835) (bf2627a)

  • model: Add SauerkrautLM-ColPali visual document retrieval models (#3804)

  • model: Add SauerkrautLM-ColPali visual document retrieval models

Add inference code and requirements for SauerkrautLM-ColPali visual document retrieval models.

These are multi-vector embedding models based on the ColPali architecture:

  • ColQwen3 (Qwen3-VL backbone): 1.7B Turbo, 2B, 4B, 8B variants
  • ColLFM2 (LFM2-VL backbone): 450M variant
  • ColMinistral3 (Ministral3 backbone): 3B variant

All models produce 128-dimensional embeddings per text/image token and use MaxSim (late interaction) for retrieval scoring.

Model checkpoints:

  • fix: Address review comments
  • Remove loader functions, use classes directly in ModelMeta
  • Remove unused get_fused_embeddings method
  • Move model.to(device) and model.eval() to base class init
  • Pass torch_dtype directly to ColMinistral3.from_pretrained
  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • Update pyproject.toml

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • fix: Update release_date to 2025-12-20

  • fix: address review comments - remove partial, add adapted_from and training_datasets

  • Update mteb/models/model_implementations/slm_models.py

Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>

  • fix: import COLPALI_CITATION from colpali_models and add model_type

  • add training datasets

  • fix: remove section headers and use PyPI package instead of Git URL

  • fix: resolve merge conflicts and remove section headers

  • fix: use COLPALI_TRAINING_DATA for training_datasets

  • fix: use exact n_parameters and memory_usage_mb values from HuggingFace

  • don't build 3.14

  • lint


Co-authored-by: David Golchinfar <d.golchin@web.de>
Co-authored-by: Roman Solomatin <samoed.roman@gmail.com>
Co-authored-by: Roman Solomatin <36135455+Samoed@users.noreply.github.com> (44e9b20)