GitHub - devflowinc/trieve: All-in-one infrastructure for search, recommendations, RAG, and analytics offered via API

All-in-one solution for search, recommendations, and RAG

Quick Links

Features

🔒 Self-Hosting in your VPC or on-prem: We have full self-hosting guides for AWS, GCP, Kubernetes generally, and docker compose available on our documentation page here.
🧠 Semantic Dense Vector Search: Integrates with OpenAI or Jina embedding models and Qdrant to provide semantic vector search.
🔍 Typo Tolerant Full-Text/Neural Search: Every uploaded chunk is vector'ized with naver/efficient-splade-VI-BT-large-query for typo tolerant, quality neural sparse-vector search.
🖊️ Sub-Sentence Highlighting: Highlight the matching words or sentences within a chunk and bold them on search to enhance UX for your users. Shout out to the simsearch crate!
🌟 Recommendations: Find similar chunks (or files if using grouping) with the recommendation API. Very helpful if you have a platform where users' favorite, bookmark, or upvote content.
🤖 Convenient RAG API Routes: We integrate with OpenRouter to provide you with access to any LLM you would like for RAG. Try our routes for fully-managed RAG with topic-based memory management or select your own context RAG.
💼 Bring Your Own Models: If you'd like, you can bring your own text-embedding, SPLADE, cross-encoder re-ranking, and/or large-language model (LLM) and plug it into our infrastructure.
🔄 Hybrid Search with cross-encoder re-ranking: For the best results, use hybrid search with BAAI/bge-reranker-large re-rank optimization.
📆 Recency Biasing: Easily bias search results for what was most recent to prevent staleness
🛠️ Tunable Merchandizing: Adjust relevance using signals like clicks, add-to-carts, or citations
🕳️ Filtering: Date-range, substring match, tag, numeric, and other filter types are supported.
👥 Grouping: Mark multiple chunks as being part of the same file and search on the file-level such that the same top-level result never appears twice

Are we missing a feature that your use case would need? - call us at 628-222-4090, make a Github issue, or join the Matrix community and tell us! We are a small company who is still very hands-on and eager to build what you need; professional services are available.

Local development with Linux

Installing via Smithery

To install Trieve for Claude Desktop automatically via Smithery:

npx -y @smithery/cli install trieve-mcp-server --client claude

Debian/Ubuntu Packages needed packages

sudo apt install curl \
gcc \
g++ \
make \
pkg-config \
python3 \
python3-pip \
libpq-dev \
libssl-dev \
openssl

Arch Packages needed

sudo pacman -S base-devel postgresql-libs

Install NodeJS and Yarn

You can install NVM using its install script.

curl -o- https://raw.githubusercontent.com/nvm-sh/nvm/v0.39.5/install.sh | bash

You should restart the terminal to update bash profile with NVM. Then, you can install NodeJS LTS release and Yarn.

nvm install --lts
npm install -g yarn

Make server tmp dir

mkdir server/tmp

Install rust

curl --proto '=https' --tlsv1.2 -sSf https://sh.rustup.rs | sh

Install cargo-watch

cargo install cargo-watch

Setup env's

cp .env.analytics ./frontends/analytics/.env
cp .env.chat ./frontends/chat/.env
cp .env.search ./frontends/search/.env
cp .env.example ./server/.env
cp .env.dashboard ./frontends/dashboard/.env

Add your `LLM_API_KEY` to `./server/.env`

Here is a guide for acquiring that.

Steps once you have the key

Open the ./server/.env file
Replace the value for LLM_API_KEY to be your own OpenAI API key.
Replace the value for OPENAI_API_KEY to be your own OpenAI API key.

Export the following keys in your terminal for local dev

The PAGEFIND_CDN_BASE_URL and S3_SECRET_KEY_CSVJSONL could be set to a random list of strings.

export OPENAI_API_KEY="your_OpenAI_api_key" \
LLM_API_KEY="your_OpenAI_api_key" \
PAGEFIND_CDN_BASE_URL="lZP8X4h0Q5Sj2ZmV,aAmu1W92T6DbFUkJ,DZ5pMvz8P1kKNH0r,QAqwvKh8rI5sPmuW,YMwgsBz7jLfN0oX8" \
S3_SECRET_KEY_CSVJSONL="Gq6wzS3mjC5kL7i4KwexnL3gP8Z1a5Xv,V2c4ZnL0uHqBzFvR2NcN8Pb1g6CjmX9J,TfA1h8LgI5zYkH9A9p7NvWlL0sZzF9p8N,pKr81pLq5n6MkNzT1X09R7Qb0Vn5cFr0d,DzYwz82FQiW6T3u9A4z9h7HLOlJb7L2V1"

Start docker container services needed for local dev

cat .env.chat .env.search .env.server .env.docker-compose > .env

./convenience.sh -l

Start services for local dev

We recommend managing this through tmuxp, see the guide here or terminal tabs.

cd clients/ts-sdk
yarn build

cd frontends
yarn
yarn dev

cd server
cargo watch -x run

cd server
cargo run --bin ingestion-worker

cd server
cargo run --bin file-worker

cd server
cargo run --bin delete-worker

cd search
yarn
yarn dev

Verify Working Setup

After the cargo build has finished (after the tmuxp load trieve):

check that you can see redoc with the OpenAPI reference at localhost:8090/redoc
make an account create a dataset with test data at localhost:5173
search that dataset with test data at localhost:5174

Additional Instructions for testing cross encoder reranking models

To test the Cross Encoder rerankers in local dev,

click on the dataset, go to the Dataset Settings -> Dataset Options -> Additional Options and uncheck the Fulltext Enabled option.
in the Embedding Settings, select your reranker model and enter the respective key in the adjacent textbox, and hit save.
in the search playground, set Type -> Semantic and select Rerank By -> Cross Encoder
if AIMon Reranker is selected in the Embedding Settings, you can enter an optional Task Definition in the search playground to specify the domain of context documents to the AIMon reranker.

Debugging issues with local dev

Reach out to us on discord for assistance. We are available and more than happy to assist.

Debug diesel by getting the exact generated SQL

diesel::debug_query(&query).to_string();

Local Setup for Testing Stripe Features

Install Stripe CLI.

stripe login
stripe listen --forward-to localhost:8090/api/stripe/webhook
set the STRIPE_WEBHOOK_SECRET in the server/.env to the resulting webhook signing secret
stripe products create --name trieve --default-price-data.unit-amount 1200 --default-price-data.currency usd
stripe plans create --amount=1200 --currency=usd --interval=month --product={id from response of step 3}

Name		Name	Last commit message	Last commit date
Latest commit History 4,984 Commits
.github		.github
.vscode		.vscode
batch-etl		batch-etl
charts/trieve		charts/trieve
clients		clients
docker		docker
frontends		frontends
glasskube		glasskube
hallucination-detection		hallucination-detection
pdf2md		pdf2md
scripts		scripts
server		server
terraform		terraform
.dockerignore		.dockerignore
.env.analytics		.env.analytics
.env.chat		.env.chat
.env.dashboard		.env.dashboard
.env.docker-compose		.env.docker-compose
.env.example		.env.example
.env.search		.env.search
.env.server		.env.server
.gitattributes		.gitattributes
.gitignore		.gitignore
.release-please-manifest.json		.release-please-manifest.json
CHANGELOG.md		CHANGELOG.md
CHANGE_LOG.md		CHANGE_LOG.md
CODE_OF_CONDUCT.md		CODE_OF_CONDUCT.md
CONTRIBUTING.md		CONTRIBUTING.md
Caddyfile		Caddyfile
LICENSE.txt		LICENSE.txt
Makefile		Makefile
README.md		README.md
SECURITY.md		SECURITY.md
convenience.bat		convenience.bat
convenience.sh		convenience.sh
docker-compose-cpu-embeddings.yml		docker-compose-cpu-embeddings.yml
docker-compose-firecrawl.yml		docker-compose-firecrawl.yml
docker-compose-gpu-embeddings.yml		docker-compose-gpu-embeddings.yml
docker-compose.yml		docker-compose.yml
package.json		package.json
release-please-config.json		release-please-config.json
turbo.json		turbo.json
version.txt		version.txt
yarn.lock		yarn.lock

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

All-in-one solution for search, recommendations, and RAG

Quick Links

Features

Local development with Linux

Installing via Smithery

Debian/Ubuntu Packages needed packages

Arch Packages needed

Install NodeJS and Yarn

Make server tmp dir

Install rust

Install cargo-watch

Setup env's

Add your `LLM_API_KEY` to `./server/.env`

Steps once you have the key

Export the following keys in your terminal for local dev

Start docker container services needed for local dev

Start services for local dev

Verify Working Setup

Additional Instructions for testing cross encoder reranking models

Debugging issues with local dev

Debug diesel by getting the exact generated SQL

Local Setup for Testing Stripe Features

Contributors

About

Releases 7

Contributors 51

Languages

License

devflowinc/trieve

Folders and files

Latest commit

History

Repository files navigation

All-in-one solution for search, recommendations, and RAG

Quick Links

Features

Local development with Linux

Installing via Smithery

Debian/Ubuntu Packages needed packages

Arch Packages needed

Install NodeJS and Yarn

Make server tmp dir

Install rust

Install cargo-watch

Setup env's

Add your LLM_API_KEY to ./server/.env

Steps once you have the key

Export the following keys in your terminal for local dev

Start docker container services needed for local dev

Start services for local dev

Verify Working Setup

Additional Instructions for testing cross encoder reranking models

Debugging issues with local dev

Debug diesel by getting the exact generated SQL

Local Setup for Testing Stripe Features

Contributors

About

Topics

Resources

License

Code of conduct

Security policy

Stars

Watchers

Forks

Releases 7

Contributors 51

Languages

Add your `LLM_API_KEY` to `./server/.env`