Home / Products / Nova AI Rack
The full system

Nova AI Rack.
AI, delivered whole.

Cooled by Nexalus

One training server, up to ten inference servers, one liquid loop. Delivered racked, plumbed and benchmarked, with one partner to call.

1
Nova 8 Training
10
Nova 8 Inference
1
liquid loop, one partner
Per rack

What one Nova AI Rack delivers.

Cooled by Nexalus
Inference
550ktokens / s
Up to 10 × Nova 8 Inference at up to 55,000 tokens a second each.1
GPU memory
8.8TB
7.7 TB GDDR7 for serving plus 1.1 TB HBM3e for training.4
Heat out
~60kW of hot water
Captured at source as water up to 60 °C, ready to reject or reuse (design value).5
How teams use it

Train once. Serve all day.

Cooled by Nexalus

Teams typically run 10 to 20 inference servers for every training server. The rack is sized for that ratio, so the model you fine-tune at night is the model your users talk to in the morning.

1 · Train or fine-tune

Nova 8 Training adapts open models to your data on site.

2 · Deploy

Push the new model to the inference servers in the same rack, on the same network.

3 · Serve

Ten Nova 8 Inference servers answer users at up to 550,000 tokens a second.

Heat, not waste

Seal the rack. Lose the HVAC.

Cooled by Nexalus

Start in a standard air-cooled rack. Connect the liquid loop and the heat leaves as hot water instead of hot air, ready for a simple dry cooler or for heating buildings.

Today · air-cooled hall

Serverheat into air
→
CRAC/CRAHair handling
→
Chillercompressors
→
Towerevaporates water

Heat is mixed into room air, chilled back out, then evaporated away: three stages of plant after the server.

Sealed rack · cooled by Nexalus

Nova rackheat captured at source
60 °C
→
Dry coolerfans only
or
Heat reusebuildings, district heat

60 °C water is hotter than outside air on the hottest day, so a dry cooler rejects it all year. No compressors, no chilled water.

~100%
of server heat into water

Up to 98% independently measured on a Nexalus sealed platform.5

0
CRACs, CRAHs or chillers

In a sealed rack the room no longer needs air conditioning for the servers.

0 L
water evaporated

Closed loop with dry coolers. Nothing lost to cooling towers.

Delivered whole

One partner from quote to running model.

Cooled by Nexalus

Sized to your workload

We size the training-to-inference ratio, power and cooling with you.

Racked and plumbed

Servers, loop and quick-connects installed and pressure-tested before shipping.

Benchmarked on delivery

Every server is benchmarked and the results are handed over with the rack.

Supported as one

One support line for the servers and the cooling, through the Renda network.

Don't take our numbers

Run your model on a live Nova 8.

Book a two-hour remote session in a fresh, isolated environment. See your own tokens per second before you commit.