How Core42's sovereign AI cloud gave AIREV the elastic, resilient infrastructure to scale OnDemand's 300+ production agents.
THE CHALLENGE
OnDemand coordinates more than 300 AI agents across 16 microservices. As the platform expanded, AIREV needed GPU infrastructure that could grow rapidly, maintain enterprise-grade uptime and place different inference workloads on the most suitable hardware. Enterprise deployments also required UAE data sovereignty and regional governance alignment. Fixed capacity, provisioning delays or dependence on one accelerator family would constrain performance, economics and speed to market.
THE CORE42 SOLUTION
Core42 delivered a fully managed, sovereign AI cloud that combines elastic capacity, multi-accelerator choice and production-grade operations:
→ Elastic GPU scaling — capacity expands on demand without provisioning delays.
→ Mixed inferencing — NVIDIA, AMD and Cerebras match workload price-performance needs.
→ Managed AI operations — Kubernetes, Slurm and built-in observability unify the environment.
→ Multi-zone sovereignty — 99.95% availability with UAE data residency and governance.
PLATFORM CAPABILITIES
| CAPABILITIES | WHAT IT DOES | |
| Elastic GPU Infrastructure | Scales on demand without provisioning delays, giving a growing agentic platform capacity when production workloads increase. | |
| Multi-Accelerator Inference | Offers NVIDIA, AMD and Cerebras options so workloads can be aligned with different performance-cost profiles. | |
| Unified AI Operations | Integrates Kubernetes and Slurm orchestration with built-in observability in a fully managed operating environment. | |
| Multi-Zone Sovereign Architecture |
|
OnDemand is AIREV's agentic AI operating system for building, deploying and managing AI applications. Behind the user experience, more than 300 AI agents operate across 16 microservices, creating a fast-changing infrastructure profile. Agent workloads rise unevenly, inference economics vary by model and accelerator, and enterprise customers expect high availability alongside regional control of data. AIREV therefore needed a production environment that could scale quickly, absorb changing demand and support new deployment models without locking the platform to one hardware architecture.
To achieve this, Core42 delivered a fully managed, sovereign AI cloud rather than a fixed GPU cluster. Elastic capacity expands without provisioning delays, while accelerator choice across NVIDIA, AMD and Cerebras allows each inference workload to be aligned with an appropriate performance-cost profile. Kubernetes and Slurm provide common orchestration, and built-in observability gives AIREV one operating view across the environment. A multi-zone architecture adds resilience, while UAE-based deployment supports data residency and governance requirements for regional enterprise customers.
With this foundation, AIREV scaled compute 3.5x, with the environment carrying a 99.95% availability SLA and supportingmixed inferencing across all three accelerator families. OnDemand Enterprise is also expanding through software-only, self-hosted and cloud-hosted distribution models across the GCC and Africa. Core42's infrastructure gives AIREV a flexible production base for that next phase: capacity can grow with adoption, workloads can move to the right hardware, and enterprise deployments can maintain the resilience, governance and sovereign readiness required in each market.
OUTCOMES AT A GLANCE
3.5xCompute scale achieved |
US $1.225MReported 2025 compute spend |
99.95%Availability SLA |
|---|---|---|
300+AI agents in production |
16Microservices supporting OnDemand |
3Accelerator families used for inference |