Skip to main content
Weaviate Docs (migrated from docs.weaviate.io) Docs

Search documentation

Type to search this documentation.

On this pageOverview

Reranker

Weaviate's integration with NVIDIA's APIs allows you to access their models' capabilities directly from Weaviate.

Configure a Weaviate collection to use an NVIDIA reranker model, and Weaviate will use the specified model and your NVIDIA NIM API key to rerank search results.

This two-step process involves Weaviate first performing a search and then reranking the results using the specified model.

Reranker integration illustration

Your Weaviate instance must be configured with the NVIDIA reranker integration (reranker-nvidia) module.

For Weaviate Cloud (WCD) users

This integration is enabled by default on Weaviate Cloud (WCD) instances.

For self-hosted users

You must provide a valid NVIDIA NIM API key to Weaviate for this integration. Go to NVIDIA to sign up and obtain an API key.

Provide the API key to Weaviate using one of the following methods:

  • Set the NVIDIA_APIKEY environment variable that is available to Weaviate.
  • Provide the API key at runtime, as shown in the examples below.
Python
# Recommended: save sensitive data as environment variables
nvidia_key = os.getenv("NVIDIA_API_KEY")
JavaScript/TypeScript
const nvidiaApiKey = process.env.NVIDIA_API_KEY || '';  // Replace with your inference API key

Configure a Weaviate collection to use an NVIDIA reranker model as follows:

Python
from weaviate.classes.config import Configureclient.collections.create(    "DemoCollection",    reranker_config=Configure.Reranker.nvidia()    # Additional parameters not shown)
JavaScript/TypeScript
await client.collections.create({  name: 'DemoCollection',  reranker: weaviate.configure.reranker.nvidia(),});

You can specify one of the available models for Weaviate to use, as shown in the following configuration example:

Python
from weaviate.classes.config import Configureclient.collections.create(    "DemoCollection",    reranker_config=Configure.Reranker.nvidia(        model="nvidia/llama-3.2-nv-rerankqa-1b-v2",        base_url="https://integrate.api.nvidia.com/v1",    )    # Additional parameters not shown)
JavaScript/TypeScript
await client.collections.create({  name: 'DemoCollection',  reranker: weaviate.configure.reranker.nvidia({    model: "nvidia/llama-3.2-nv-rerankqa-1b-v2",    baseURL: "https://integrate.api.nvidia.com/v1"  }),});

The default model is used if no model is specified.

Once the reranker is configured, Weaviate performs reranking operations using the specified NVIDIA model.

More specifically, Weaviate performs an initial search, then reranks the results using the specified model.

Any search in Weaviate can be combined with a reranker to perform reranking operations.

Reranker integration illustration

Python
from weaviate.classes.query import Rerankcollection = client.collections.use("DemoCollection")response = collection.query.near_text(    query="A holiday film",  # The model provider integration will automatically vectorize the query    limit=2,    rerank=Rerank(        prop="title",                   # The property to rerank on        query="A melodic holiday film"  # If not provided, the original query will be used    ))for obj in response.objects:    print(obj.properties["title"])
JavaScript/TypeScript
let myCollection = client.collections.use('DemoCollection');const results = await myCollection.query.nearText(  ['A holiday film'],  {    limit: 2,    rerank: {      property: 'title',                // The property to rerank on      query: 'A melodic holiday film'   // If not provided, the original query will be used    }  });for (const obj of results.objects) {  console.log(obj.properties['title']);}

You can use any reranker model on NVIDIA NIM APIs with Weaviate.

The default model is nvidia/rerank-qa-mistral-4b.

Once the integrations are configured at the collection, the data management and search operations in Weaviate work identically to any other collection. See the following model-agnostic examples:

Have a question or feedback? Here's how to reach us.

Suggest an edit

Propose a replacement for this page. The site team reviews it before applying any changes.

Export
Documentation menu