On-demand bare-metal 8× RTX 5090 nodes for training, inference and rendering — P2P-accelerated, in our own Tier III data center, SSH-ready in under 30 seconds.
$0.80 / GPU · hour — no egress fees
$ entrim launch --node 8x-rtx5090
✓ provisioning node-93 (8× RTX 5090, 256GB)…
✓ node ready in 28s — ssh root@node-93.entrim.io
$ nvidia-smi --query-gpu=name,memory.total --format=csv,noheader
8× NVIDIA GeForce RTX 5090, 32768 MiB8×
RTX 5090 / node
256 GB
GDDR7 / node
<30s
Cold-boot to SSH
99.9%
Uptime target
Performance
Consumer RTX cards ship with peer-to-peer transfers disabled, so multi-GPU jobs bounce every tensor through system memory. We run patched open GPU kernel modules that re-enable direct GPU-to-GPU P2P over PCIe — activations, weights and gradients move card-to-card without the detour, speeding up both distributed training and tensor-parallel inference.
Representative multi-GPU scaling — relative tokens/sec on the same hardware and job. Figures illustrative.
Thermals
Eight RTX 5090s in one box throw a lot of heat. Each node lives in a custom-built chassis with a high-static-pressure fan wall and tuned front-to-back airflow, so every card holds its boost clocks through the longest runs — no throttling, no surprises.
The hardware
The fastest consumer GPU ever built, on the Blackwell architecture. Every Entrim node packs eight of them — 256 GB of VRAM in one box — to fit larger models and bigger batches.
The eight GPUs sit on a serious platform — plenty of CPU cores, memory and fast local storage to keep them fed — all on redundant power.
Reliability & compliance
Not a reseller. We run our own bare-metal in our certified EU data center — redundancy at every layer, and engineers who actually answer.
A concurrently maintainable Tier III data center in Slovenia, central EU — owned and operated by us, with redundant power and cooling paths and 24/7 on-site staff. No reselling.
Information security managed to ISO 27001. Data is processed and stored in the EU under a fully GDPR-compliant chain — and wiped on instance release.
Every instance gets its own dedicated public IPv4. Host an endpoint, whitelist by IP, open ports — no shared NAT, no noisy neighbours.
Full root on real bare-metal. Load custom kernel modules (including our P2P drivers), tune the OS and run any container — nothing is locked down.
If a node ever fails, we don't wait on an RMA — we migrate you to a pre-provisioned hot spare from the same rack and get you running again fast.
Dual hot-swap PSUs per node, UPS battery backup and on-site generators — your job rides straight through a utility power event.
Local NVMe runs in redundant RAID, so a single drive failure never takes your data — or your training run — down with it.
Every node has dual uplinks into redundant top-of-rack switches. No single switch, port or cable is a point of failure.
Real engineers monitor the fleet around the clock. When you need help, you talk to the people who actually run the hardware.
Pricing
Rent a full 8× RTX 5090 node, billed per second. No egress fees — stop an instance and billing stops with it.
≈ $6.40 / hour for the full 8-GPU node
Need several nodes or a multi-node cluster? Contact sales for volume & reserved pricing.
FAQ
Create an account, open the console and reserve a node. You only pay when an instance is running.