FastEmbed-go

Go implementation of @Qdrant/fastembed

🍕 Features

Supports batch embeddings with parallelism using go-routines.
Uses @sugarme/tokenizer for fast tokenization.
Optimized embedding models.

The default embedding supports "query" and "passage" prefixes for the input text. The default model is Flag Embedding, which is top of the MTEB leaderboard.

🔍 Not looking for Go?

Python 🐍: fastembed
Rust 🦀: fastembed-rs
JavaScript 🌐: fastembed-js

🤖 Models

🚀 Installation

Run the following Go CLI command in your project directory:

go get -u github.com/anush008/fastembed-go

📖 Usage

import "github.com/anush008/fastembed-go"

// With default options
model, err := fastembed.NewFlagEmbedding(nil)
if err != nil {
 panic(err)
}
defer model.Destroy()

// With custom options
options := fastembed.InitOptions{
 Model:     fastembed.BGEBaseEN,
 CacheDir:  "model_cache",
 MaxLength: 200,
}

model, err = fastembed.NewFlagEmbedding(&options)
if err != nil {
 panic(err)
}
defer model.Destroy()

documents := []string{
 "passage: Hello, World!",
 "query: Hello, World!",
 "passage: This is an example passage.",
 // You can leave out the prefix but it's recommended
 "fastembed-go is licensed under MIT",
}

// Generate embeddings with a batch-size of 25, defaults to 256
embeddings, err := model.Embed(documents, 25)  //  -> Embeddings length: 4
if err != nil {
 panic(err)
}

Supports passage and query embeddings for more accurate results

// Generate embeddings for the passages
// The texts are prefixed with "passage" for better results
// The batch size is set to 1 for demonstration purposes
passages := []string{
 "This is the first passage. It contains provides more context for retrieval.",
 "Here's the second passage, which is longer than the first one. It includes additional information.",
 "And this is the third passage, the longest of all. It contains several sentences and is meant for more extensive testing.",
}

embeddings, err := model.PassageEmbed(passages, 1)  //  -> Embeddings length: 3
if err != nil {
 panic(err)
}

// Generate embeddings for the query
// The text is prefixed with "query" for better retrieval
query := "What is the answer to this generic question?";

embeddings, err := model.QueryEmbed(query)
if err != nil {
 panic(err)
}

ℹ︎ Notice:

The Onnx runtime path is automatically loaded on most environments. However, if you encounter

panic: Platform-specific initialization failed: Error loading ONNX shared library

Set the ONNX_PATH env to your Onnx installation. For eg, on MacOS:

export ONNX_PATH="/path/to/onnx/lib/libonnxruntime.dylib"

On Linux:

export ONNX_PATH="/path/to/onnx/lib/libonnxruntime.so"

You can find the Onnx runtime releases here.

🚒 Under the hood

Why fast?

It's important we justify the "fast" in FastEmbed. FastEmbed is fast because:

Quantized model weights
ONNX Runtime which allows for inference on CPU, GPU, and other dedicated runtimes

Name		Name	Last commit message	Last commit date
Latest commit History 31 Commits
.github/workflows		.github/workflows
vendor		vendor
.gitignore		.gitignore
.golangci.yml		.golangci.yml
.releaserc		.releaserc
LICENSE		LICENSE
README.md		README.md
fastembed.go		fastembed.go
fastembed_test.go		fastembed_test.go
go.mod		go.mod
go.sum		go.sum

License

Anush008/fastembed-go

Folders and files

Latest commit

History

Repository files navigation

Go implementation of @Qdrant/fastembed

🍕 Features

🔍 Not looking for Go?

🤖 Models

🚀 Installation

📖 Usage

Supports passage and query embeddings for more accurate results

ℹ︎ Notice:

🚒 Under the hood

Why fast?

Why light?

Why accurate?

📄 LICENSE

About

Topics

Resources

License

Stars

Watchers

Forks

Languages