GPU Lite — RTX 4090
Offshore RTX 4090 for AI inference & rendering.
- NVIDIA RTX 4090 (24 GB VRAM)
- AI inference outside US jurisdiction
- vLLM / Ollama pre-installed
- 10 Gbps uplink
RTX 4090, A6000, and H100 dedicated GPUs for AI inference, ML training, and rendering.
GPU servers are no longer a niche line item. Open-source LLM inference, fine-tuning, computer-vision training, 3D rendering, and scientific compute all benefit from dedicated GPU access — and the largest cloud providers…
Offshore RTX 4090 for AI inference & rendering.
Workstation-class GPU for ML training offshore.
Datacenter-grade H100 for serious LLM workloads.
All gpu plans include Bare-metal GPU access — no virtualization tax · RTX 4090 (24 GB) → A6000 (48 GB) → H100 (80 GB) · Pre-installed: CUDA 12, PyTorch, vLLM, Ollama, Transformers.
Six things that come standard on every gpu plan. No upsells, no surprise fees.
no virtualization tax
Included on every plan in this tier.
PyTorch, vLLM, Ollama, Transformers
Included on every plan in this tier.
Switzerland, and the Netherlands
Included on every plan in this tier.
Jurisdictional notice
We do not process DMCA-format notices — the DMCA has no legal force in any of our operating jurisdictions.
Substantive complaints under the local law of the datacenter jurisdiction (Iceland, Switzerland, Netherlands, Romania, Moldova, Bulgaria, Russia, Panama) are reviewed by counsel. AUP-prohibited content (CSAM, malware C2, fraud, phishing) is acted upon globally regardless of claim format.
Inference of 7B-13B models, single-user rendering, light fine-tuning: RTX 4090 (GPU Lite). Inference up to 30B 4-bit, sustained training, ECC for reliability: A6000 (GPU Pro). Frontier-class LLMs (70B FP16, large fine-tuning runs, production inference): H100 (GPU Beast).
Yes — open a sales ticket for 2× or 4× H100 with NVLink. Provisioning lead time is typically 7-14 days. We don't list multi-GPU configurations on the public catalog because customer requirements vary too much for a fixed SKU to make sense.
Bare-metal — the customer has direct PCIe access to the physical GPU with no hypervisor. This matters for inference batch performance and for CUDA driver flexibility on custom workflows.
Yes — Windows Server 2022 images with CUDA drivers are available on request. Rendering workflows (Octane, Redshift, V-Ray) often prefer the Windows toolchain; AI/ML workflows are typically Linux.
Significantly better at our scale: a single H100 at $2,899/month is equivalent to roughly 200 hours/month on AWS p5.48xlarge ($98/hr × 0.125 GPU share). For sustained workloads, bare-metal rental is 3-5× cheaper than hyperscaler GPU pricing.
Monthly is the default. For short-burst research workloads, GPU Pro and Beast are available with a 7-day minimum; shorter periods need a sales conversation.
We host H100s in Iceland, Switzerland, and the Netherlands under standard commercial agreements; no export-control approvals are required for normal end-user inference or research training. Customers in restricted jurisdictions should consult counsel before signing.
KVM virtualization, full root access, NVMe storage — anonymous signup and DMCA-resilient by default.
Bare-metal Xeon, EPYC, and Ryzen dedicated servers — full hardware, RAID NVMe, IPMI access.
High-capacity HDD-backed VPS for seedboxes, off-site backup, and long-term media archives.
NGINX-RTMP, Wowza, and FFmpeg pre-installed for IPTV, live streaming, and on-demand video.
Anonymous signup. Bitcoin & Monero accepted. Provisioned across 8 jurisdictions.
No credit card required · 7-day money-back guarantee