Our Thesis
Energy becomes intelligence here.
Every industrial economy now runs on two inputs: energy and intelligence. Saudi Arabia has the first in abundance. Navid builds the machine that converts it into the second — anchored in the Kingdom, engineered to serve the region and the world.
That machine is an AI Factory. We are building it in Dammam, and opening it to everyone who needs what it makes.
The Facility
Engineered for the physics of frontier compute.
15 MW across 6,000 m² of industrial land in MODON's Dammam 2nd Industrial City. Liquid-cooled from day one, engineered in partnership with ProteCooling — Mulhim & Partners Engineering Company.
Power train, cooling loop and rack density are designed as one system around the accelerators they carry. That is what lets frontier silicon run at its ceiling, every hour of the year.
Key specifications
NVIDIA DGX B300 Reference Architecture · NEO Cloud AI Factory · NVIDIA Preferred Partner
| IT Load Capacity | 15MW |
|---|---|
| Power Infrastructure | 33kV Dual Feed | N+1 Redundancy |
| Cooling System | Direct Liquid Cooling (DLC), NVIDIA Reference Design |
| PUE Target | ≤ 1.15 |
| Rack Density | Up to 140kW per Rack |
| Total POD Capacity | 1,000+ DGX B300 Nodes (Phased) |
| Network | NVIDIA Quantum-2 InfiniBand |
| Storage | High Performance Parallel File System |
| Security | Physical + Cybersecurity by Design |
The Compute
Three tiers of NVIDIA silicon, each matched to its work.
| Tier | Role | What it runs |
|---|---|---|
| NVIDIA GB300 | Flagship | Blackwell Ultra, rack-scale and NVLink-connected. Domain-adaptive training and high-throughput inference for the largest open-weight models. |
| NVIDIA HGX B200 | Workhorse | Eight Blackwell SXM GPUs per node, 1.4 TB of HBM3e and 14.4 TB/s of NVLink bandwidth. The volume tier for industrial inference, computer vision and fine-tuning — where cost per inference decides whether a use case ships. |
| NVIDIA Rubin | Roadmap | NVIDIA's next generation is in full production, with partner availability from the second half of 2026. It joins the floor as the facility scales. |
Heterogeneous by design. Every workload lands on the silicon built for it — and you pay for exactly that.
The Economics
We separate the two halves of inference.
Prefill is compute-bound. Decode is memory-bound. We run them on separate GPU pools, each scaled to its own curve, so every request meets hardware matched to the work it is doing.
NVIDIA reports that disaggregated serving can raise throughput by up to 15× without sacrificing latency. We designed the floor for it from the first rack.
It arrives as the one number that decides economics at scale: tokens per megawatt. That is what lets us serve frontier models at the labs' published rates while owning the metal underneath.
The Token Factory
Navid is a token manufacturer.
We self-host open-weight frontier models on our own GPU infrastructure and serve them as production inference — including agentic-class models such as Kimi K3 and GLM-5.2.
Frontier models, published prices, sovereign ground — open to teams in the Kingdom, across the GCC and internationally, on infrastructure you are entitled to inspect.
Arabic-First
Models that think in Arabic.
Fourteen centuries of written heritage, carried forward into models that reason natively in the language.
Our deepest investment is Arabic-native AI: the sovereign compute layer for training, adapting and serving Arabic models in agentic use across government, industry and enterprise. The Kingdom leads the region in Arabic model development. We are building the factory floor beneath it.
Sovereign by Design
Three commitments, written into the architecture.
Jurisdiction
Anchored in Saudi Arabia and built to the Kingdom's regulatory standard — the strictest bar our customers ask us to clear.
Residency
Your prompts, documents and model weights stay where you place them, with in-Kingdom residency for workloads that require it.
Isolation
Multi-tenant NeoCloud with tenant-level separation — or dedicated capacity where regulation requires it.
NVIDIA NEO Cloud Stack
One platform, from facility to application.
- AI Applications Training | Inference | HPC | Digital Twins
- NEO Cloud Services Scheduling | Orchestration | Monitoring | IAM
- AI Infrastructure DGX B300 | Networking | Storage | Security
- Cloud Foundation Compute | Network | Storage | Facility
- Sovereign Cloud Hosted in Kingdom of Saudi Arabia
Sustainable by Design
Efficiency engineered in, not bolted on.
High Efficiency
Low PUE ≤ 1.15
Liquid Cooling
Reduced Water Usage
Renewable Ready
Designed for Clean Energy Integration
Carbon Conscious
Supporting KSA Vision 2030
Reach
Built in Dammam. Serving three continents.
Capacity is open wherever the demand is: national entities and industrial operators in the Kingdom, enterprises across the GCC, and international teams that need frontier inference close to their users.
Dammam sits between Europe, Africa and Asia — one of the shortest network paths to a large share of the world's inference demand, on a grid with the power to feed it.
Built for Industrial AI
Where AI runs the plant.
Built for the operators of refineries, ports, utilities, factories and national platforms — where a model is a control loop with consequences.
What Comes Next
15 MW is phase one.
The site, the cooling architecture and the fabric are designed to scale well beyond it — and demand is already arriving from the Kingdom, the GCC and abroad. Early capacity is being allocated now.
TAKING AI FORWARD
Experience Products in AI
Empower your organization with Navid's AI-driven solutions, merging human expertise with advanced technology for unparalleled growth and security.
Copyright © 2024 Navid.