---
sidebar_label: Qualcomm Inference Suite
---

# QISEmbeddings

This will help you get started with QIS embedding models using LangChain. For detailed documentation on `QISEmbeddings` features and configuration options, please refer to the [API reference](https://python.langchain.com/v0.2/api_reference/langchain_qualcomm_inference_suite/embeddings/langchain_qualcomm_inference_suite.embeddingsQISEmbeddings.html).

## Overview
### Integration details

| Provider | Package |
|:--------:|:-------:|
| [QIS](/docs/integrations/providers/qualcomm-inference-suite/) | [langchain-qualcomm-inference-suite](https://python.langchain.com/v0.2/api_reference/langchain_qualcomm_inference_suite/embeddings/langchain_qualcomm_inference_suite.embeddingsQISEmbeddings.html) |

## Setup

To access Qualcomm Inference Suite embedding models please reach out to your Qualcomm Inference Suite service provider for support. They will provide you an API key and API endpoint. Then you can and install and make use of the `langchain-qualcomm-inference-suite` integration package.

### Credentials

Please reach out to your Qualcomm Inference Suite service provider for support to generate an API key and obtain the API endpoint. Once you've done this set the `IMAGINE_API_KEY` and `IMAGINE_API_ENDPOINT` environment variables:

In [1]:
import getpass
import os

if not os.getenv("IMAGINE_API_KEY"):
    os.environ["IMAGINE_API_KEY"] = getpass.getpass("Enter your Qualcomm Inference Suite API key: ")
if not os.getenv("IMAGINE_API_ENDPOINT"):
    os.environ["IMAGINE_API_ENDPOINT"] = input("Enter your Qualcomm Inference Suite API endpoint: ")

### Installation

The LangChain Qualcomm Inference Suite integration lives in the `langchain-qualcomm-inference-suite` package:

In [None]:
%pip install -qU langchain-qualcomm-inference-suite

## Instantiation

Now we can instantiate our model object and generate chat completions:


In [2]:
from langchain_qualcomm_inference_suite import QISEmbeddings

embeddings = QISEmbeddings(
    model="BAAI/bge-large-en-v1.5",
)

## Indexing and Retrieval

Embedding models are often used in retrieval-augmented generation (RAG) flows, both as part of indexing data as well as later retrieving it. For more detailed instructions, please see our [RAG tutorials](/docs/tutorials/).

Below, see how to index and retrieve data using the `embeddings` object we initialized above. In this example, we will index and retrieve a sample document in the `InMemoryVectorStore`.

In [4]:
# Create a vector store with a sample text
from langchain_core.vectorstores import InMemoryVectorStore

text = "LangChain is the framework for building context-aware reasoning applications"

vectorstore = InMemoryVectorStore.from_texts(
    [text],
    embedding=embeddings,
)

# Use the vectorstore as a retriever
retriever = vectorstore.as_retriever()

# Retrieve the most similar text
retrieved_documents = retriever.invoke("What is LangChain?")

# show the retrieved document's content
retrieved_documents[0].page_content

'LangChain is the framework for building context-aware reasoning applications'

## Direct Usage

Under the hood, the vectorstore and retriever implementations are calling `embeddings.embed_documents(...)` and `embeddings.embed_query(...)` to create embeddings for the text(s) used in `from_texts` and retrieval `invoke` operations, respectively.

You can directly call these methods to get embeddings for your own use cases.

### Embed single texts

You can embed single texts or documents with `embed_query`:

In [5]:
single_vector = embeddings.embed_query(text)
print(str(single_vector)[:100]) # Show the first 100 characters of the vector

[0.029815673828125, 0.037322998046875, -0.036651611328125, -0.02178955078125, -0.0212249755859375, 0


### Embed multiple texts

You can embed multiple texts with `embed_documents`:

In [6]:
text2 = (
    "LangGraph is a library for building stateful, multi-actor applications with LLMs"
)
two_vectors = embeddings.embed_documents([text, text2])
for vector in two_vectors:
    print(str(vector)[:100]) # Show the first 100 characters of the vector

[0.029815673828125, 0.037322998046875, -0.036651611328125, -0.02178955078125, -0.0212249755859375, 0
[0.059814453125, 0.0281829833984375, -0.035552978515625, 0.003520965576171875, -0.023681640625, -0.0


## API Reference

For detailed documentation on `QISEmbeddings` features and configuration options, please refer to the [API reference](https://api.python.langchain.com/en/latest/embeddings/langchain_qualcomm_inference_suite.embeddings.QISEmbeddings.html).
