Levelrail
Skip to content

Self-host Text Embeddings Inference with Docker ​

A lean server that turns text into embeddings with popular open models, on CPU with no extra setup.

This page describes the Levelrail template for Text Embeddings Inference (AI). It deploys the services below as one Docker Compose app on your own server.

Project site: https://huggingface.co/docs/text-embeddings-inference.

Recommended memory: about 2048 MiB.

Services, ports and volumes ​

ServiceImageContainer portsVolumes
teighcr.io/huggingface/text-embeddings-inference:cpu-1.880
tei_data -> /data

Environment variables ​

This template sets no environment variables.

Passwords and keys marked as generated are created for you when the app is deployed and stored as secrets. Values are not shown here.

Deploy Text Embeddings Inference with Levelrail ​

In the dashboard, open Apps, choose New app, then Browse templates, and select Text Embeddings Inference. Review the Compose body and deploy.

With the CLI:

levelrail-cli templates deploy text-embeddings-inference --name my-text-embeddings-inference

See Service template catalog for how templates work, and Getting started if you have not installed Levelrail yet.

More AI templates ​

All templates are listed in the self-host gallery.

Released under the Apache 2.0 License.