Renda Nova · AI infrastructure

More AI from
every watt.
Cooled by Nexalus.

Cooled by Nexalus

The Nova range: liquid-cooled 8-GPU servers for AI inference and training. Built by Renda, cooled end to end by Nexalus, benchmarked before they ship.

Renda Nova 8 top view: eight liquid-cooled GPUs and the CPU on one loop
Replay · 1-hour run54,968tokens / second, one server
Throughput
55,000tok/s
Up to 55,000 tokens a second from one 4U server, Llama 3.3 70B.1
Versus air-cooled
+20%tokens
More tokens than a leading tier-1 air-cooled 8-GPU server, same GPUs.2
Efficiency
+22%per watt
8.6 tokens per watt vs 7.0. Less power for every token served.3
The Nova range

One platform. Inference and training.

Cooled by Nexalus

Most teams don't want a spec sheet, they want AI that works. Pick the server for the job, or take both as a rack.

Performance you can hold

Not a burst. An hour at full load.

Cooled by Nexalus

All eight GPUs flat out for an hour, 175 back-to-back runs. Throughput held at ~54,000 tokens a second and the hottest GPU stayed at 80 °C or below.

Server throughput Replay

54,968 tokens / s
Run 1 / 175Elapsed 0 minHour average 53,984Hottest GPU 65 °C
5 Oct 2026 · 12:28–13:26 UTC

Per GPU, this run

Every card averaged more than 6,600 tokens a second across the hour.

Replay of measured data from the Nova 8 Inference development build.1 See the full benchmark →

Why liquid

Cooling is where the performance is.

Cooled by Nexalus

Every GPU in a Nova server sits on a Nexalus liquid loop. Heat leaves at the chip, so the silicon holds its boost clocks, the server spends less power on fans, and the heat comes out as hot water instead of hot air.

More tokens per box

Cooler GPUs hold higher clocks under sustained load: +20% tokens from the same eight GPUs.

More tokens per watt

Power goes into inference, not into moving air: 8.6 tokens per watt vs 7.0.

Two slots become one

Single-slot liquid blocks put eight full-power GPUs in one 4U chassis.

Heat you can reuse

Up to 60 °C water out. Reject it with a dry cooler or heat buildings with it.

How the Nexalus loop works

Built for

From one server to a full AI floor.

Cooled by Nexalus

Enterprise private AI

Run models on your own premises. Your data never leaves the building.

Learn more →

Data centres & neoclouds

More sellable tokens per kilowatt, less cooling plant per rack.

Learn more →

Research & education

Train and serve on campus, with heat that can warm the building.

Learn more →

Engineering & media

The same liquid-cooled GPU platform for simulation, rendering and VFX.

Learn more →

See it in person

Meet Nova.

Cooled by Nexalus
OCT2026

NVIDIA GTC · Berlin

Liquid-cooled GPU and CPU hardware on hand, plus a live remote session with the Nova 8 test server. Book a meeting.

NOV18–19

Data Centre World · Paris

Nova 8 on show with our cooling partners. Come and see the loop up close.

Don't take our numbers

Run your model on a live Nova 8.

Book a two-hour remote session in a fresh, isolated environment. See your own tokens per second before you commit.