Models Overview
SovereignEG serves models through an OpenAI-compatible API — switch models by changing one string. See the Model Library for what is live today.
Available models
| Model | Context | Input EGP/1M | Output EGP/1M | Status |
|---|---|---|---|---|
AionLabs: Aion 3.5aion-3.5 | 262K | 171.90 | 343.81 | Live |
AionLabs: Aion 3.5 Miniaion-3.5-mini | 262K | 40.11 | 80.22 | Live |
AionLabs: Aion-2.0aion-2.0 | 131K | 45.84 | 91.68 | Live |
AionLabs: Aion-3.0aion-3.0 | 131K | 171.90 | 343.81 | Live |
AionLabs: Aion-3.0-Miniaion-3.0-mini | 131K | 40.11 | 80.22 | Live |
AionLabs: Aion-RP 1.0 (8B)aion-rp-llama-3.1-8b | 33K | 45.84 | 91.68 | Live |
all-minilm-l12-v2all-minilm-l12-v2 | 512 | 0.29 | — | Live |
all-minilm-l6-v2all-minilm-l6-v2 | 512 | 0.29 | — | Live |
all-mpnet-base-v2all-mpnet-base-v2 | 512 | 0.29 | — | Live |
Amazon: Nova 2 Litenova-2-lite-v1 | 1M | 17.19 | 143.25 | Live |
Amazon: Nova Lite 1.0nova-lite-v1 | 300K | 3.44 | 13.75 | Live |
Amazon: Nova Micro 1.0nova-micro-v1 | 128K | 2.01 | 8.02 | Live |
Amazon: Nova Pro 1.0nova-pro-v1 | 300K | 45.84 | 183.36 | Live |
Anthropic: Claude Fable 5claude-fable-5 | 1M | 573.02 | 2865.08 | Live |
Anthropic: Claude Fable 5.1claude-fable-5.1 | 1M | 573.02 | 2865.08 | Live |
Anthropic: Claude Fable Latestclaude-fable-latest | 1M | 573.02 | 2865.08 | Live |
Anthropic: Claude Haiku Latestclaude-haiku-latest | 200K | 57.30 | 286.51 | Live |
Anthropic: Claude Opus 4.1claude-opus-4.1 | 200K | 859.52 | 4297.61 | Live |
Anthropic: Claude Opus 4.5claude-opus-4.5 | 200K | 286.51 | 1432.54 | Live |
Anthropic: Claude Opus 4.6claude-opus-4.6 | 1M | 286.51 | 1432.54 | Live |
Anthropic: Claude Opus 4.7claude-opus-4.7 | 1M | 286.51 | 1432.54 | Live |
Anthropic: Claude Opus 4.8claude-opus-4.8 | 1M | 286.51 | 1432.54 | Live |
Anthropic: Claude Opus 5claude-opus-5 | 1M | 286.51 | 1432.54 | Live |
Anthropic: Claude Opus 5.5claude-opus-5.5 | 1M | 229.21 | 1146.03 | Live |
Showing 24 of 402 models. Browse the full filterable catalog →
Rates are per million tokens against your prepaid balance. Coming soon models are listed but not yet callable.
Choosing a model
Every model above is callable through the same OpenAI-compatible API —
switch between them by changing the model string. When picking one:
- Latency + cost — smaller models return tokens faster and cost less per token. A good default for chat, drafting, and high-volume classification.
- Quality — larger models handle complex reasoning, long instructions, and code generation better. Reach for them when a smaller model falls short.
- Arabic — Arabic-native models produce stronger Arabic than translated output. See the Arabic Guide for recommendations.
- Long context — sort by context window in the Model Library when you need to fit large documents into a single request.
Live status, context windows, and per-model EGP pricing always reflect the current catalog on the Model Library.