bge-m3-multi-8k - API in Egypt, billed in EGP
LiveBGE-M3 is a multilingual text embedding model developed by BAAI, distinguished by its Multi-Linguality (supporting 100+ languages), Multi-Functionality (unified dense, multi-vector, and sparse retrieval), and Multi-Granularity (handling inputs from short queries to long documents). It achieves state-of-the-art retrieval performance across diverse benchmarks while maintaining a single model for multiple retrieval modes. This endpoint serves the model's full 8192-token context.
Modality
Embedding
Context window
8K tokens
Region
US
Weights
Open weights
Published
29 Sept 2026
Pricing
Input
0.57
EGP per 1M tokens
Output
—
EGP per 1M tokens
Monthly cost example
28.65 EGP
Illustrative estimate: 50M input + 15M output tokens per month at this model's catalog rates.
Use it via the API
bge-m3-multi-8k works with any OpenAI-compatible SDK — point the base URL at https://dev-backend.sovereigneg.com/v1 and use your SovereignEG API key.
Run inference
OpenAI-compatible — POST /v1/embeddings — drop-in for any OpenAI SDK.
from openai import OpenAI
client = OpenAI(
base_url="https://dev-backend.sovereigneg.com/v1",
api_key="YOUR_API_KEY",
)
response = client.embeddings.create(
model="bge-m3-multi-8k",
input="The food was delicious and the waiter...",
encoding_format="float",
)
vector = response.data[0].embedding
print(f"dim={len(vector)}, first 8 dims: {vector[:8]}")import OpenAI from "openai"
const client = new OpenAI({
baseURL: "https://dev-backend.sovereigneg.com/v1",
apiKey: "YOUR_API_KEY",
})
const response = await client.embeddings.create({
model: "bge-m3-multi-8k",
input: "The food was delicious and the waiter...",
encoding_format: "float",
})
const vector = response.data[0].embedding
console.log(`dim=${vector.length}, first 8 dims:`, vector.slice(0, 8))curl https://dev-backend.sovereigneg.com/v1/embeddings \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "bge-m3-multi-8k",
"input": "The food was delicious and the waiter...",
"encoding_format": "float"
}'Frequently asked questions
How much does bge-m3-multi-8k cost?
0.57 EGP per 1M input tokens; Output tokens are free. Billing is metered per token in Egyptian pounds (EGP), with no minimum commitment.
What is the context window of bge-m3-multi-8k?
bge-m3-multi-8k supports a context window of 8,192 tokens (8K).
Where is bge-m3-multi-8k hosted?
bge-m3-multi-8k is served from US (US-hosted inference).
Is bge-m3-multi-8k an open-weights model?
Yes — bge-m3-multi-8k is an open-weights model. You call it through the SovereignEG API like any other catalog model, with per-token EGP billing.
How do I use bge-m3-multi-8k via the API?
bge-m3-multi-8k is available through the OpenAI-compatible SovereignEG API: point your SDK's base URL at https://dev-backend.sovereigneg.com/v1 and call /v1/embeddings with model "bge-m3-multi-8k" and your API key.