Models Overview
SovereignEG serves models through an OpenAI-compatible API — switch models by changing one string. See the Model Library for what is live today.
Available models
| Model | Context | Input EGP/1M | Output EGP/1M | Status |
|---|---|---|---|---|
all-minilm-l12-v2all-minilm-l12-v2 | 512 | 0.28 | Free | Live |
all-minilm-l6-v2all-minilm-l6-v2 | 512 | 0.28 | Free | Live |
all-mpnet-base-v2all-mpnet-base-v2 | 512 | 0.28 | Free | Live |
bge-base-en-v1.5bge-base-en-v1.5 | 512 | 0.28 | Free | Live |
bge-en-iclbge-en-icl | 8K | 0.56 | Free | Live |
bge-large-en-v1.5bge-large-en-v1.5 | 512 | 0.56 | Free | Live |
bge-m3bge-m3 | 8K | 0.56 | Free | Live |
bge-m3-multibge-m3-multi | 512 | 0.56 | Free | Live |
bge-m3-multi-8kbge-m3-multi-8k | 8K | 0.56 | Free | Live |
blur-backgroundblur-background | — | — | — | Live |
bria-3.2bria-3.2 | — | — | — | Live |
bria-3.2-vectorbria-3.2-vector | — | — | — | Live |
claude-fable-5claude-fable-5 | 1M | 564.94 | 2824.68 | Live |
claude-haiku-4.5claude-haiku-4.5 | 200K | 56.49 | 282.47 | Live |
claude-opus-4.7claude-opus-4.7 | 1M | 282.47 | 1412.34 | Live |
claude-opus-4.8claude-opus-4.8 | 1M | 282.47 | 1412.34 | Live |
claude-opus-5claude-opus-5 | 1M | 282.47 | 1412.34 | Live |
claude-sonnet-4.6claude-sonnet-4.6 | 1M | 169.48 | 847.40 | Live |
claude-sonnet-5claude-sonnet-5 | 1M | 169.48 | 847.40 | Live |
clip-vit-b-32clip-vit-b-32 | 77 | 0.28 | Free | Live |
clip-vit-b-32-multilingual-v1clip-vit-b-32-multilingual-v1 | 512 | 0.28 | Free | Live |
deepseek-r1-0528deepseek-r1-0528 | 164K | 28.25 | 121.46 | Live |
deepseek-v3deepseek-v3 | 164K | 18.08 | 50.28 | Live |
deepseek-v3-0324deepseek-v3-0324 | 164K | 13.56 | 50.84 | Live |
Showing 24 of 152 models. Browse the full filterable catalog →
Rates are per million tokens against your prepaid balance. Coming soon models are listed but not yet callable.
Choosing a model
Every model above is callable through the same OpenAI-compatible API —
switch between them by changing the model string. When picking one:
- Latency + cost — smaller models return tokens faster and cost less per token. A good default for chat, drafting, and high-volume classification.
- Quality — larger models handle complex reasoning, long instructions, and code generation better. Reach for them when a smaller model falls short.
- Arabic — Arabic-native models produce stronger Arabic than translated output. See the Arabic Guide for recommendations.
- Long context — sort by context window in the Model Library when you need to fit large documents into a single request.
Live status, context windows, and per-model EGP pricing always reflect the current catalog on the Model Library.