UAE Sovereign Cloud
EN ع
Sign in
← Catalogue
Tool Open ○ Public · National
vllm-project /

vLLM

High-throughput, memory-efficient inference and serving engine for LLMs.

Serve

Readme

High-throughput, memory-efficient inference and serving engine for LLMs. This asset is published to AI Majlis under the namespace shown above and runs entirely within the UAE Sovereign Cloud perimeter. All invocations are authenticated, classification-checked, rate-limited, and written to an append-only audit trail.

# Connect over the governed gateway
POST https://aimajlis.gov.ae/vllm-project/serve
Authorization: Bearer <scoped-token>
X-Classification: Open