multilingual-e5-large-instruct

Live
intfloat

The Multilingual-E5 models, initialized from XLM-RoBERTa, support up to 512 tokens per input — any longer text will be silently truncated. To ensure optimal performance, always prefix inputs with “query:” or “passage:”, as the model was explicitly trained with this format.

Modality

Embedding

Context window

512 tokens

Region

US

Published

Jul 21, 2026

Pricing

Input

0.56

EGP per 1M tokens

Output

Free

EGP per 1M tokens

Use it via the API

multilingual-e5-large-instruct works with any OpenAI-compatible SDK — point the base URL at https://preprod-backend.sovereigneg.com/v1 and use your SovereignEG API key.

Run inference

OpenAI-compatible — POST /v1/embeddings — drop-in for any OpenAI SDK.

from openai import OpenAI

client = OpenAI(
    base_url="https://preprod-backend.sovereigneg.com/v1",
    api_key="YOUR_API_KEY",
)

response = client.embeddings.create(
    model="multilingual-e5-large-instruct",
    input="The food was delicious and the waiter...",
    encoding_format="float",
)

vector = response.data[0].embedding
print(f"dim={len(vector)}, first 8 dims: {vector[:8]}")

Frequently asked questions

How much does multilingual-e5-large-instruct cost?

0.56 EGP per 1M input tokens; Output tokens are free. Billing is metered per token in Egyptian pounds (EGP), with no minimum commitment.

What is the context window of multilingual-e5-large-instruct?

multilingual-e5-large-instruct supports a context window of 512 tokens (512).

Where is multilingual-e5-large-instruct hosted?

multilingual-e5-large-instruct is served from US (US-hosted inference).

How do I use multilingual-e5-large-instruct via the API?

multilingual-e5-large-instruct is available through the OpenAI-compatible SovereignEG API: point your SDK's base URL at https://preprod-backend.sovereigneg.com/v1 and call /v1/embeddings with model "multilingual-e5-large-instruct" and your API key.