Skip to content

Embeddings and vector stores

The embedding model and the LanceDB or Qdrant backend behind retrieval.

Embeddings

EMBEDDING_PROVIDER

Default: huggingface · Type: one of openai, huggingface, ollama

Where vector-store embeddings are computed: huggingface runs a local sentence-transformers model, openai and ollama call a service.

EMBEDDING_MODEL_NAME

Default: sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2 · Type: str

Embedding model identifier used by the selected provider. Spelled with the org prefix so it shares one SharedEncoder slot with AGG_EMBEDDING_MODEL and CHUNK_EMBEDDING_MODEL when aligned.

EMBEDDING_API_KEY

Default: unset · Type: str

Provider API key for hosted embedding services.

EMBEDDING_BASE_URL

Default: unset · Type: str

Provider base URL (for Ollama-compatible endpoints).

EMBEDDING_DIMENSION

Default: 384 · Type: int

Expected dense embedding vector size for core and neighborhood vectors.

EMBEDDING_BM25_MODEL_NAME

Default: Qdrant/bm25 · Type: str

fastembed SparseTextEmbedding model id for the BM25 sparse lane.

EMBEDDING_QUERY_PREFIX

Default: empty · Type: str

Prefix prepended to text embedded as a query. Asymmetric retrieval models underperform their spec without it — BGE wants 'Represent this sentence for searching relevant passages: ', E5 wants 'query: '. Empty (default) suits the symmetric paraphrase model. Part of the stored embedding contract: changing it requires a reindex.

EMBEDDING_DOCUMENT_PREFIX

Default: empty · Type: str

Prefix prepended to text embedded as a document during indexing (E5 wants 'passage: '; BGE wants nothing). Part of the stored embedding contract: changing it requires a reindex.

LanceDB

LANCEDB_ENABLED

Default: false · Type: bool

Enable embedded LanceDB when QDRANT_URI is unset. Uses a local directory via lancedb.connect(data_dir).

LANCEDB_DATA_DIR

Default: ~/.lancedb_data · Type: Path or str

Local filesystem directory passed to lancedb.connect(...) (supports ~ expansion).

LANCEDB_ONTOLOGY_TABLE

Default: unset · Type: str

Lance table for ontology atom vectors; derived like FUSEKI_DATASET, and replaced the same way by ontocast serve and ontocast process.

LANCEDB_FACTS_TABLE

Default: unset · Type: str

Lance table reserved for future fact vectors; created on init.

Qdrant

QDRANT_URI

Default: unset · Type: str

Qdrant server URL, such as http://localhost:6333. Setting it selects Qdrant as the vector store.

QDRANT_API_KEY

Default: unset · Type: str

API key for a Qdrant server that requires one.

QDRANT_ONTOLOGY_COLLECTION

Default: unset · Type: str

Qdrant collection for ontology atom vectors; derived like FUSEKI_DATASET, and replaced the same way by ontocast serve and ontocast process.

QDRANT_FACTS_COLLECTION

Default: unset · Type: str

Qdrant collection reserved for future fact vectors; created on init.

QDRANT_GRPC_PORT

Default: 6334 · Type: int

Qdrant gRPC port, used when QDRANT_USE_GRPC is on.

QDRANT_USE_GRPC

Default: false · Type: bool

Talk to Qdrant over gRPC instead of HTTP.

QDRANT_VECTOR_SIZE

Default: unset · Type: int

Vector size override. When set, must equal EmbeddingConfig.dimension; when unset, the embedding dimension is used.

QDRANT_DISTANCE

Default: Cosine · Type: one of Cosine, Dot, Euclid, Manhattan

Qdrant vector distance when creating collections (Cosine, Dot, Euclid, Manhattan; same as qdrant_client Distance).

QDRANT_UPSERT_BATCH_SIZE

Default: 256 · Type: int

Batch size used for Qdrant upsert operations.

QDRANT_TIMEOUT_SECONDS

Default: 30 · Type: int

Per-request timeout for Qdrant calls, in whole seconds (the client accepts nothing finer). Without one, an unreachable or hung Qdrant blocks a pipeline worker indefinitely.