all-minilm-l6-v2
LiveWe present a sentence transformation model that achieves state-of-the-art results on various NLP tasks without requiring task-specific architectures or fine-tuning. Our approach leverages contrastive learning and utilizes a variety of datasets to learn robust sentence representations. We evaluate our model on several benchmarks and demonstrate its effectiveness in various applications such as text classification, sentiment analysis, named entity recognition, and question answering.
Modality
Embedding
Context window
512 tokens
Region
US
Published
Jul 21, 2026
Pricing
Input
0.28
EGP per 1M tokens
Output
Free
EGP per 1M tokens
Use it via the API
all-minilm-l6-v2 works with any OpenAI-compatible SDK — point the base URL at https://preprod-backend.sovereigneg.com/v1 and use your SovereignEG API key.
Run inference
OpenAI-compatible — POST /v1/embeddings — drop-in for any OpenAI SDK.
from openai import OpenAI
client = OpenAI(
base_url="https://preprod-backend.sovereigneg.com/v1",
api_key="YOUR_API_KEY",
)
response = client.embeddings.create(
model="all-minilm-l6-v2",
input="The food was delicious and the waiter...",
encoding_format="float",
)
vector = response.data[0].embedding
print(f"dim={len(vector)}, first 8 dims: {vector[:8]}")import OpenAI from "openai"
const client = new OpenAI({
baseURL: "https://preprod-backend.sovereigneg.com/v1",
apiKey: "YOUR_API_KEY",
})
const response = await client.embeddings.create({
model: "all-minilm-l6-v2",
input: "The food was delicious and the waiter...",
encoding_format: "float",
})
const vector = response.data[0].embedding
console.log(`dim=${vector.length}, first 8 dims:`, vector.slice(0, 8))curl https://preprod-backend.sovereigneg.com/v1/embeddings \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "all-minilm-l6-v2",
"input": "The food was delicious and the waiter...",
"encoding_format": "float"
}'Frequently asked questions
How much does all-minilm-l6-v2 cost?
0.28 EGP per 1M input tokens; Output tokens are free. Billing is metered per token in Egyptian pounds (EGP), with no minimum commitment.
What is the context window of all-minilm-l6-v2?
all-minilm-l6-v2 supports a context window of 512 tokens (512).
Where is all-minilm-l6-v2 hosted?
all-minilm-l6-v2 is served from US (US-hosted inference).
How do I use all-minilm-l6-v2 via the API?
all-minilm-l6-v2 is available through the OpenAI-compatible SovereignEG API: point your SDK's base URL at https://preprod-backend.sovereigneg.com/v1 and call /v1/embeddings with model "all-minilm-l6-v2" and your API key.
Related models
View all models →all-minilm-l12-v2
SBERT · Embedding · 512 context
0.28 EGP / 1M input tokens
all-mpnet-base-v2
SBERT · Embedding · 512 context
0.28 EGP / 1M input tokens
clip-vit-b-32
SBERT · Embedding · 77 context
0.28 EGP / 1M input tokens
clip-vit-b-32-multilingual-v1
SBERT · Embedding · 512 context
0.28 EGP / 1M input tokens