Pricing

AI API pricing, billed in EGP

Every model on one OpenAI-compatible API, priced in Egyptian pounds. You pay per token used — no FX conversion on your bill.

244 live models · input from 0.11 to 1707.84 EGP per million tokens

Chat models

191 models

Billed per million tokens, input and output priced separately. Output is usually the larger share of a chat bill, so compare both columns.

ModelContextInput EGP / 1MOutput EGP / 1M
mistral-nemo-instruct-2407Mistral AI131K1.081.71
meta-llama-3.1-8b-instruct-turboMeta131K1.142.28
Meta Llama 3.2 1B InstructMeta60K1.5411.44
gpt-oss-20bOpenAI131K1.717.97
granite-4.2-3bibm-granite131K1.716.83
Qwen: Qwen3.7 FlashAlibaba1M1.717.40
gpt-oss-120bOpenAI131K2.119.68
gemma-3-12b-itGoogle131K2.858.54
gemma-3-4b-itGoogle131K2.855.69
gpt-5-nanoOpenAI400K2.8522.77
Meta Llama 3.2 3B InstructMeta131K2.8518.79
mistral-small-24b-instruct-2501Mistral AI33K2.854.55
nemotron-3-nano-30b-a3bNVIDIA262K2.8511.39
nim/meta/llama-3.1-8b-instructnim131K2.854.55
DeepSeek: DeepSeek V4 Flash 0731DeepSeek1M3.4210.25
granite-4.2-8bibm-granite131K3.4214.23
Ling-3.0-flashinclusionai131K3.4210.25
ling-3.0-flash-fininclusionAI262K3.4210.25
ling-3.0-flash-vlinclusionAI131K3.4210.25
glm-4.7-flashzai-org131K3.4422.77
Qwen: Qwen3.5-FlashAlibaba1M3.7014.80
gemma-4-26b-a4b-itGoogle262K3.9819.36
phi-4Microsoft16K3.987.97
Qwen3 Coder 30B A3b InstructAlibaba262K3.9815.94
mistral-small-3.2-24b-instruct-2506Mistral AI128K4.2711.39
gemma-3-27b-itGoogle131K4.559.11
nvidia-nemotron-3.5-lightningNVIDIA262K4.5511.39
qwen3-32bAlibaba41K4.5515.94
nvidia-nemotron-3-super-120b-a12bNVIDIA262K4.8422.77
Qwen3-0.6BAlibaba4K5.0010.00
deepseek-v4-flashDeepSeek1M5.1210.25
gemma-4-31b-it-turboGoogle262K5.1219.36
Qwen: Qwen3.8 FlashAlibaba1M5.1216.05
qwen3-235b-a22b-instruct-2507Alibaba262K5.1231.31
qwen3-next-80b-a3b-instructAlibaba262K5.1262.62
Arize AI Qwen 2 1.5B Instructarize-ai33K5.695.69
gpt-4.1-nanoOpenAI1M5.6922.77
llama-3.3-70b-instruct-turboMeta131K5.6918.22
llama-4-scout-17b-16e-instructMeta328K5.6917.08
nim/meta/llama-3.3-70b-instructnim131K5.6918.22
qwen3.5-9bAlibaba262K5.698.54
qwen3.6-35b-a3bAlibaba262K5.6954.08
seed-2.0-miniByteDance256K5.6922.77
voxtral-small-24b-2507Mistral AI33K5.6917.08
Qwen3-VL-32B-InstructAlibaba131K5.9223.68
Qwen3 8BAlibaba131K6.6625.90
Qwen3-VL-8B-InstructAlibaba131K6.6625.90
Qwen: Qwen3 Coder NextAlibaba262K6.8345.54
qwen3-14bAlibaba41K6.8313.66
qwen3-30b-a3bAlibaba41K6.8328.46
gemma-4-31b-itGoogle262K7.4021.63
hy3tencent262K7.4030.17
mimo-v2.5XiaomiMiMo1M7.9715.94
qwen3.5-35b-a3bAlibaba262K7.9756.93
Xiaomi: MiMo-V2.6-Flashxiaomi1M7.9715.94
glm-5.3-flashzai-org1M8.5428.46
gpt-4o-miniOpenAI128K8.5434.16
gpt-4o-mini-2024-07-18OpenAI128K8.5434.16
gpt-oss-120b-turboOpenAI131K8.5434.16
Qwen3 Next 80B A3b ThinkingAlibaba262K8.5468.31
qwen3-vl-30b-a3b-instructAlibaba262K8.5434.16
gpt-4o-mini-search-previewOpenAI128K8.5834.32
granite-4.2-30bibm-granite131K9.1137.00
llama-guard-4-12bMeta164K10.2510.25
Qwen: Qwen3.6 FlashAlibaba1M10.6764.04
Qwen: Qwen3 Coder FlashAlibaba1M11.1055.50
deepseek-v4.1-flashDeepSeek1M11.3934.16
gpt-5.4-nanoOpenAI400K11.3971.16
gpt-6-lunaOpenAI1M11.3956.93
nemotron-content-safety-3.5NVIDIA131K11.3911.39
Qwen: Qwen3.8 27BAlibaba262K11.39142.32
qwen3-vl-235b-a22b-instructAlibaba262K11.3950.10
step-3.7-flashstepfun-ai256K11.3965.47
qwen3-235b-a22b-thinking-2507Alibaba131K13.09130.93
Qwen3.8-27B-FP8Alibaba131K14.00105.00
deepseek-v3.1DeepSeek164K14.2354.08
DeepSeek: DeepSeek V3.1DeepSeek164K14.2354.08
gemini-3.1-flash-liteGoogle1M14.2385.39
gpt-5-miniOpenAI400K14.23113.86
gpt-5.1-codex-miniOpenAI400K14.23113.86
seed-1.8ByteDance256K14.23113.86
MiniMax M2MiniMaxAI205K14.5258.07
deepseek-v3.2DeepSeek164K14.8021.63
Qwen: Qwen3.5 Plus 2026-02-15Alibaba1M14.8088.81
qwen3.5-122b-a10bAlibaba262K14.80118.41
qwen3.5-27bAlibaba262K14.80148.01
deepseek-v3.1-terminusDeepSeek131K15.3756.93
gemma-4-31b-it-ultraGoogle131K15.3743.27
minimax-m3MiniMaxAI524K15.9462.62
DeepSeek: DeepSeek V3 0324DeepSeek164K16.5164.90
gemini-2.5-flashGoogle1M17.08142.32
minimax-m2.7MiniMaxAI205K17.0868.31
Qwen: Qwen3.5 Plus 2026-04-20Alibaba1M17.08102.47
deepseek-v3DeepSeek164K18.2250.67
qwen3.6-27bAlibaba262K18.22182.17
Qwen3.7 PlusAlibaba1M18.2272.87
Muse Glimmer 30Bmeta-models131K19.9285.39
minimax-m2.7-turboMiniMaxAI197K21.6396.78
glm-4.7zai-org203K22.7799.62
gpt-4.1-miniOpenAI1M22.7791.08
gpt-5.6-lunaOpenAI1M22.77136.63
meta-llama-3.1-70b-instruct-turboMeta131K22.7722.77
nim/meta/llama-3.1-70b-instructnim131K22.7722.77
Xiaomi: MiMo-V2.6-Proxiaomi1M24.4849.53
mimo-v2.5-proXiaomiMiMo1M24.7649.53
DeepSeek: DeepSeek V4 Flash Vision ExpDeepSeek1M25.0575.15
kimi-k2.5moonshotai262K25.62128.09
qwen3.5-397b-a17bAlibaba262K25.62170.78
Thinking Machines: Inkling Smallthinkingmachines524K25.6268.31
deepseek-r1-0528DeepSeek164K28.46122.40
glm-4.6zai-org203K28.46113.86
gpt-3.5-turboOpenAI16K28.4685.39
nvidia-nemotron-3-ultra-550b-a55bNVIDIA262K28.46125.24
Qwen3.6 PlusAlibaba1M28.46170.78
seed-2.0-codeByteDance256K28.46170.78
seed-2.0-proByteDance256K28.46170.78
MoonshotAI: Kimi K2 0711moonshotai131K32.45130.93
GLM 4.5Vzai-org66K34.16102.47
glm-5zai-org198K34.16109.30
MoonshotAI: Kimi K2 0905moonshotai262K34.16142.32
NVIDIA Nemotron 3 Ultra 550B A55B NVFP4NVIDIA203K34.16136.63
Gemma-2 Instruct (27B)Google8K37.0037.00
kimi-k2.7-codemoonshotai262K37.36187.86
DeepSeek: R1DeepSeek64K39.85142.32
hermes-3-llama-3.1-70bNousResearch131K39.8539.85
glm-5.2zai-org1M42.70136.63
Google: Gemini 3.6 FlashGoogle1M42.70213.48
Google: Gemini 3.7 FlashGoogle1M42.70213.48
Google: Gemini 3.8 FlashGoogle1M42.70213.48
gpt-5.4-miniOpenAI400K42.70256.18
kimi-k2.6moonshotai262K42.70199.25
DeepSeek R1 Distill Llama 70BDeepSeek8K45.5445.54
Qwen2.5-VL (72B) InstructAlibaba128K45.5456.93
Tencent: Hy4 previewtencent1M47.48142.38
Z.ai: GLM 5.3z-ai1M51.24227.71
Inkling FP4thinkingmachines524K54.08230.56
claude-haiku-4.5Anthropic200K56.93284.64
Qwen: Qwen3.6 Max PreviewAlibaba262K58.47350.79
o3-miniOpenAI200K62.62250.48
o4-miniOpenAI200K62.62250.48
qwen3-maxAlibaba256K68.31341.57
qwen3-max-thinkingAlibaba256K68.31341.57
gemini-2.5-proGoogle1M71.16569.28
gpt-5OpenAI400K71.16569.28
gpt-5.1OpenAI400K71.16569.28
gpt-5.1-codexOpenAI400K71.16569.28
gpt-5.1-codex-maxOpenAI400K71.16569.28
deepseek-v4-proDeepSeek1M74.01148.01
DeepSeek: DeepSeek V4 Pro 0813DeepSeek1M74.01148.01
glm-5.1zai-org203K79.70250.48
gemini-3.5-flashGoogle1M85.39512.35
gpt-3.5-turbo-instructOpenAI4K85.39113.86
Qwen3.7 MaxAlibaba1M85.39256.18
Qwen: Qwen3.8 MaxAlibaba256K93.93281.85
gpt-5.2OpenAI400K99.62796.99
gpt-5.2-codexOpenAI400K99.62796.99
gpt-5.3-codexOpenAI400K99.62796.99
claude-sonnet-5Anthropic1M113.86569.28
gemini-3.1-proGoogle1M113.86683.14
gpt-4.1OpenAI1M113.86455.42
gpt-5.6-solOpenAI1M113.86569.28
gpt-5.6-terraOpenAI1M113.86683.14
gpt-6-solOpenAI1M113.86569.28
o3OpenAI200K113.86455.42
qwen3.8-2.4t-a95bAlibaba262K113.86341.57
gpt-4oOpenAI128K142.32569.28
gpt-4o-2024-08-06OpenAI128K142.32569.28
gpt-4o-2024-11-20OpenAI128K142.32569.28
gpt-5.4OpenAI1M142.32853.92
gpt-4o-search-previewOpenAI128K143.00572.00
MoonshotAI: Kimi K3moonshotai1M162.25811.23
Claude Sonnet 4.5Anthropic1M170.78853.92
claude-sonnet-4.6Anthropic1M170.78853.92
gpt-3.5-turbo-16kOpenAI16K170.78227.71
Claude Opus 5.5Anthropic1M227.711138.56
Claude Opus 4.5Anthropic200K284.641423.20
Claude Opus 4.6Anthropic1M284.641423.20
Claude Opus 5Anthropic1M284.641423.20
claude-opus-4.7Anthropic1M284.641423.20
claude-opus-4.8Anthropic1M284.641423.20
gpt-4o-2024-05-13OpenAI128K284.64853.92
gpt-5.5OpenAI1M284.641707.84
gpt-image-2Provider400K455.42455.42
Claude Fable 5.1Anthropic1M569.282846.40
claude-fable-5Anthropic1M569.282846.40
gpt-6-astraOpenAI1M569.282846.40
o1OpenAI200K853.923415.68
o3-proOpenAI200K1138.564554.25
gpt-5.2-proOpenAI400K1195.499563.92
gpt-4OpenAI8K1707.843415.68
gpt-4-turboOpenAI8K1707.843415.68

Embedding models

28 models

Billed on input tokens only. Embeddings return vectors, not generated text, so there is no output charge.

ModelContextInput EGP / 1M
embeddinggemma-300mGoogle2K0.11
all-minilm-l12-v2SBERT5120.28
all-minilm-l6-v2SBERT5120.28
all-mpnet-base-v2SBERT5120.28
bge-base-en-v1.5BAAI5120.28
clip-vit-b-32SBERT770.28
clip-vit-b-32-multilingual-v1SBERT5120.28
e5-base-v2intfloat5120.28
gte-basethenlper5120.28
multi-qa-mpnet-base-dot-v1SBERT5120.28
paraphrase-minilm-l6-v2SBERT5120.28
text2vec-base-chineseshibing6245120.28
bge-en-iclBAAI8K0.57
bge-large-en-v1.5BAAI5120.57
bge-m3BAAI8K0.57
bge-m3-multiBAAI5120.57
bge-m3-multi-8kBAAI8K0.57
e5-large-v2intfloat5120.57
gte-largethenlper5120.57
llama-nemotron-embed-vl-1b-v2NVIDIA10K0.57
multilingual-e5-largeintfloat5120.57
multilingual-e5-large-instructintfloat5120.57
qwen3-embedding-0.6bAlibaba33K0.57
qwen3-embedding-8bAlibaba33K0.57
qwen3-embedding-4bAlibaba33K1.14
text-embedding-3-smallOpenAI—1.14
text-embedding-ada-002OpenAI—5.69
text-embedding-3-largeOpenAI—7.40

Image generation models

25 models

Billed per generated image. Where a figure is marked ≈, the exact charge depends on the request and is reported on every response.

ModelPrice
sdxl-turbostabilityai0.0114 EGP per 1024p image
flux-1-schnellblack-forest-labs0.0285 EGP per 1024p image
p-imagePrunaAI0.2846 EGP per image
flux-1-devblack-forest-labs0.5124 EGP per 1024p image
ming-image-0.1-designinclusionAI0.5693 EGP per 1024p image
ming-image-0.1-design-layerinclusionAI0.5693 EGP per 1024p image
flux-2-klein-4bblack-forest-labs0.797 EGP per 1024p image
flux-2-klein-9bblack-forest-labs0.8539 EGP per 1024p image
flux-2-problack-forest-labs≈ 0.8539 EGP per image
flux-2-devblack-forest-labs1.02 EGP per 1024p image
wan2.6-t2iWan-AI1.71 EGP per image
nano-banana-2-liteGoogle1.91 EGP per image
bria-3.2Bria2.28 EGP per image
bria-3.2-vectorBria2.28 EGP per image
fiboBria2.28 EGP per image
fibo-1.5Bria2.28 EGP per image
flux-1.1-problack-forest-labs2.28 EGP per image
seedream-4ByteDance2.28 EGP per image
seedream-4.5ByteDance2.28 EGP per image
nano-banana-2Google3.83 EGP per image
flux-2-maxblack-forest-labs≈ 3.98 EGP per image
qwen-image-maxAlibaba4.27 EGP per image
seedream-5.0-proByteDance5.64 EGP per image
gemini-3-pro-imageGoogle7.65 EGP per image
nano-banana-proGoogle7.65 EGP per image

How billing works

Rates are per million tokens, and you are charged only for the tokens a request actually uses. Each model page shows its context window, modalities and full rate details. For request formats and limits, see the documentation; to filter and compare models, use the model library.