Eight cards, one chassis
Eight NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs in a single 6.5U rackmount: 768 GB of GDDR7 across PCIe Gen5, enough headroom to serve 70B-class models to 100–500 concurrent users without leaving the rack.
8× NVIDIA RTX PRO 6000 Blackwell Server Edition (96 GB GDDR7 each, 768 GB total) in a 6.5U rackmount: the EMARQUE AI PRO 800. Supplied and supported through EMARQUE in Malaysia.

Manufacturer-defined features from the published datasheet.
Eight NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs in a single 6.5U rackmount: 768 GB of GDDR7 across PCIe Gen5, enough headroom to serve 70B-class models to 100–500 concurrent users without leaving the rack.
In Win IW-RG658-PRO 6.5U platform (a 4.5U GPU chamber plus 2U system chamber, purpose-built for the 8× RTX PRO 6000 SE thermal envelope), supplied through EMARQUE with local commissioning, warranty handling, and Tier-1 support in Malaysia.
MIG-style partitioning per card lets you carve the server into isolated inference workloads: separate teams, separate models, separate quotas on shared hardware. No noisy-neighbour fights, full observability via the BMC.
8× E1.S NVMe plus 4× 2.5-inch hot-swap bays keep the eight GPUs fed; NVIDIA ConnectX-class SuperNIC networking up to 400 / 800 Gb/s with a 10 GbE management port for scale-out clusters. No PCIe contention, no storage starvation under sustained load.
Redundant 3+1 CRPS Titanium power (9,600 W) with 14 hot-swap fans and front-to-back airflow, engineered to run reliably in a Malaysian rack or data-centre environment. EMARQUE confirms site power and cooling readiness during scoping.
GPU count, memory (up to 8 TB ECC DDR5), storage (E1.S NVMe + 2.5-inch bays), networking (ConnectX SuperNIC 400 / 800 Gb/s), and CPU choice (Intel Xeon or AMD EPYC). All selectable at quote. Need six GPUs instead of eight? The EMARQUE AI PRO 700 is the 6-GPU tier.
The four sub-systems that determine real-workload behaviour. We tune each before delivery.
Tell us your workload. EMARQUE sizes the AI Server and sends a quote.
Workload categories documented in the manufacturer's reference materials. Sizing is confirmed with your technical team during scoping.
Run a private RAG stack with a 70B-class model against your document store, code repos, and ticketing system. Eight GPUs split four ways gives four logical inference endpoints sized for 100+ concurrent users each: the right shape for finance, legal, engineering, and ops to share one server.
TGI / vLLM / Triton serving multiple fine-tuned variants of a base model. PCIe Gen5 isolation per card means no NVLink coherence overhead: the right architecture when each request fits on one GPU and you want predictable per-tenant throughput rather than tightly coupled training.
Eight independent cards parallelise across camera or audio streams cleanly: object detection on 50 4K streams, speech-to-text on 200 concurrent calls, or pose estimation across a factory floor. Each stream gets a dedicated GPU slice with consistent latency.
Run production inference on six cards during business hours, reallocate to LoRA / QLoRA fine-tuning runs on all eight overnight. The BMC + Redfish API makes the rebalance scriptable; no separate dev cluster required.
Configurable. Final BOM, GPU mix, RAM and storage, and networking topology are confirmed in writing at quotation.
Server-optimised variant of the RTX PRO 6000 Blackwell: same Blackwell silicon and 96 GB GDDR7 memory, with passive cooling (300 W TDP) designed to be cooled by the server chassis airflow. Targets enterprise rackmount deployments where the active 600 W variant would be impractical.
EMARQUE supplies the AI Server, the EMARQUE AI PRO 800, on the In Win IW-RG658-PRO 6.5U rackmount platform: a 4.5U GPU chamber plus a 2U system chamber, purpose-built for the 8× RTX PRO 6000 SE thermal envelope. Final BOM is documented at quotation.
Yes. GPU count (the 6-GPU tier is the EMARQUE AI PRO 700), memory (up to 8 TB ECC DDR5), storage (E1.S NVMe + 2.5-inch bays), networking (ConnectX SuperNIC 400 / 800 Gb/s + InfiniBand), and OS are all configurable. CPU choice (Intel Xeon vs AMD EPYC) is customer-selectable per workload.
A 3-year EMARQUE warranty with Malaysia-based local support on the AI PRO 800, plus NVIDIA's warranty entitlement on the RTX PRO 6000 SE GPUs. EMARQUE handles warranty and RMA locally.
EMARQUE handles in-country delivery, customs, commissioning, acceptance testing, and Tier-1 support response in Malaysia. Tier-2/3 escalation routes to the manufacturer per the warranty entitlement. Optional service contracts for extended response SLAs and on-site engineering visits can be added.
Step into NVIDIA DGX B200 or NVIDIA DGX B300 (or HGX B200 / B300 OEM platforms from Dell, Giga Computing, or Supermicro). Those configurations are presented on the respective NVIDIA DGX product pages. Each shows both NVIDIA-branded DGX systems and HGX OEM alternatives.
Manufacturer specifications and warranty terms apply. EMARQUE issues a formal quotation through your Key Account Manager.
Model size, concurrency, latency budget, deployment site. EMARQUE returns a quote in MYR within one Malaysian business day, sized to the workload, not the salesperson’s quota.
Enterprise AI servers are supplied for deployment in Malaysia only, and are subject to end-user, application and location verification. Start with the short pre-screening form.