Skip to main content

AI service

Semantic Search / Embeddings

We provide Semantic Search / Embeddings at any scale, on a private cluster dedicated to you — from a single GPU to a hundred, on premise or in any cloud.

If you need such a service please contact us.

Retrieval works by meaning rather than keyword, so a query finds the passage that answers it even when it shares no words with it. Embeddings are generated and stored inside your deployment, and access is scoped per user group. Embeddings can be regenerated when the model changes, without touching the source documents.

You can use the models we operate or bring your own model and we run it on your cluster. Work arrives as an API call, a batch over a folder or bucket, or a resident on-premise service, and results return as structured output with a per-job audit record.

More services