Models Overview
SovereignEG serves models through an OpenAI-compatible API — switch models by changing one string. See the Model Library for what is live today.
Available models
| Model | Context | Input EGP/1M | Output EGP/1M | Status |
|---|---|---|---|---|
all-minilm-l12-v2all-minilm-l12-v2 | 512 | 0.33 | Free | Live |
all-minilm-l6-v2all-minilm-l6-v2 | 512 | 0.33 | Free | Live |
all-mpnet-base-v2all-mpnet-base-v2 | 512 | 0.33 | Free | Live |
Arcee AI: Trinity Large Thinkingtrinity-large-thinking | 262K | 14.30 | 45.76 | Live |
Arize AI Qwen 2 1.5B Instructqwen-2-1.5b-instruct | 33K | 5.72 | 5.72 | Live |
bge-base-en-v1.5bge-base-en-v1.5 | 512 | 0.33 | Free | Live |
bge-en-iclbge-en-icl | 8K | 0.65 | Free | Live |
bge-large-en-v1.5bge-large-en-v1.5 | 512 | 0.65 | Free | Live |
bge-m3bge-m3 | 8K | 0.65 | Free | Live |
bge-m3-multibge-m3-multi | 8K | 0.65 | Free | Live |
Claude Opus 4.5claude-opus-4.5 | 200K | 277.97 | 1389.85 | Live |
Claude Opus 4.6claude-opus-4.6 | 1M | 277.97 | 1389.85 | Live |
Claude Sonnet 4.5claude-sonnet-4.5 | 1M | 166.78 | 833.91 | Live |
claude-fable-5claude-fable-5 | 1M | 555.94 | 2779.70 | Live |
claude-haiku-4.5claude-haiku-4.5 | 200K | 55.59 | 277.97 | Live |
claude-opus-4.7claude-opus-4.7 | 1M | 277.97 | 1389.85 | Live |
claude-opus-4.8claude-opus-4.8 | 1M | 277.97 | 1389.85 | Live |
claude-sonnet-4.6claude-sonnet-4.6 | 1M | 166.78 | 833.91 | Live |
claude-sonnet-5claude-sonnet-5 | 1M | 111.19 | 555.94 | Live |
clip-vit-b-32clip-vit-b-32 | 77 | 0.33 | Free | Live |
clip-vit-b-32-multilingual-v1clip-vit-b-32-multilingual-v1 | 512 | 0.33 | Free | Live |
deepseek-r1-0528deepseek-r1-0528 | 164K | 28.60 | 122.98 | Live |
deepseek-v3deepseek-v3 | 164K | 18.30 | 50.91 | Live |
deepseek-v3.1deepseek-v3.1 | 164K | 14.30 | 54.34 | Live |
Showing 24 of 144 models. Browse the full filterable catalog →
Rates are per million tokens against your prepaid balance. Coming soon models are listed but not yet callable.
Choosing a model
Every model above is callable through the same OpenAI-compatible API —
switch between them by changing the model string. When picking one:
- Latency + cost — smaller models return tokens faster and cost less per token. A good default for chat, drafting, and high-volume classification.
- Quality — larger models handle complex reasoning, long instructions, and code generation better. Reach for them when a smaller model falls short.
- Arabic — Arabic-native models produce stronger Arabic than translated output. See the Arabic Guide for recommendations.
- Long context — sort by context window in the Model Library when you need to fit large documents into a single request.
Live status, context windows, and per-model EGP pricing always reflect the current catalog on the Model Library.