NVIDIA GPU servers, CUDA-ready in minutes.
Monthly GPU servers on NVIDIA A16 and A40 hardware: from a $90 fractional slice for inference and encoding up to 2x and 4x A16 boxes with 64 GB of GPU memory. Seven cities, full root access, drivers on first boot, payable by PayPal or crypto.
Every GPU plan we sell
Orderable today, with the cities that have capacity for each size. The configurator shows the live price for your city.
A16, 1/8 GPU entry
Inference, encoding, CUDA dev boxes
A40, 1/24 GPU
Ampere card, larger NVMe and transfer
A16, 1/4 GPU
Image generation, mid-size models
A16, 1/2 GPU
Half a card, 8 GB of GPU memory
2x A16 multi-gpu
Two full cards, parallel inference workers
4x A16 multi-gpu
Four cards, 64 GB of GPU memory in one box
Our retail pricing in USD at the lowest-priced location; some metros carry a surcharge and not every plan is offered in every location. The configurator shows live pricing and availability for your chosen city. A single full A16 (16 GB) and the larger A40 tiers appear in the configurator only while a city has capacity for them.
Rent a slice of the card, or several cards
How fractional GPU works: a fixed share of the silicon and its memory, not a time-slice you queue for.
1/8 to 1/2 of a GPU
A dedicated share of the card with its own slice of GPU memory (2 to 8 GB). CUDA works exactly as on a full card; you just size the memory to your model or workload.
2x and 4x A16 boxes
Scale to 32 or 64 GB of total GPU memory across multiple cards for parallel inference workers or batch pipelines.
CUDA-ready on first boot
GPU plans deploy CUDA-ready Linux images with NVIDIA drivers matched to the plan family, so nvidia-smi works on first boot. Full root, reinstall from the panel, your keys.
Seven cities with GPU capacity
Plan counts per city as of today. Capacity moves; the configurator is the source of truth at checkout.
- 01/BangaloreIndia5 plans · up to 4x A16
- 02/Silicon ValleyUnited States4 plans · up to 2x A16
- 03/ChicagoUnited States3 plans · up to 1/2 A16
- 04/FrankfurtGermany3 plans · up to 1/2 A16
- 05/SingaporeSingapore3 plans · up to 1/2 A16
- 06/TokyoJapan3 plans · up to 1/2 A16
- 07/New YorkUnited States2 plans · A16 1/8, A40 1/24
- networkVultr, 10 Gbps ports
- ipv4dedicated, included
- imagesUbuntu, Debian + drivers
- deploy timeminutes after payment
- supportthe engineer who runs it
Honest sizing: A16 and A40, not H100
What these cards are for, and what they are not.
Runs well here
Inference on small and mid-size models, image generation on the 8 to 16 GB tiers, video encoding and transcoding, virtual desktops and cloud gaming, CUDA development and CI runners.
Not the right tool
Training or serving frontier-scale LLMs. An A16 is not an H100, and we would rather say so here than after you have deployed. If your job needs 80 GB of HBM, rent that class of card elsewhere and keep us for everything around it.
Why monthly GPU
A flat monthly price from your prepaid balance instead of a metered hourly bill. Top up by PayPal or crypto, deploy, and know what the month costs before it starts.
01/Which GPU models do you offer?
NVIDIA A16 (fractional slices up to 4x full cards) and an entry NVIDIA A40 slice. We do not sell A100 or H100 instances, and we will not pretend an A16 substitutes for one.
02/Can I run or train large language models on these?
Inference on small and mid-size models fits the 8 to 64 GB tiers well. Fine-tuning small models is possible on the multi-GPU boxes. Large-scale training belongs on data-center cards with more memory bandwidth than this class offers.
03/Do the images come with drivers and CUDA?
Yes. GPU plans deploy CUDA-ready Linux images with NVIDIA drivers matched to the plan family, so nvidia-smi works on first boot.
04/How does billing work?
Plans are monthly and paid from your prepaid account balance. Top up with PayPal or any of 18 cryptocurrencies including Bitcoin and Monero; no card or ID required.
05/Is every plan available in every location?
No. GPU capacity varies by city across the seven locations that carry it, and some metros price slightly higher. The order form shows live availability and the exact price for your chosen location.
Deploy a CUDA-ready GPU server
From a $90 slice to a 4x A16 box. Seven cities, monthly pricing, crypto welcome.