Initech
Cloud GPU · NVIDIA A16 / A40

NVIDIA GPU servers, CUDA-ready in minutes.

Monthly GPU servers on NVIDIA A16 and A40 hardware: from a $90 fractional slice for inference and encoding up to 2x and 4x A16 boxes with 64 GB of GPU memory. Seven cities, full root access, drivers on first boot, payable by PayPal or crypto.

hardwareNVIDIA A16 / A40
from$90/mo
cities with GPU capacity7
billingmonthly, PayPal or crypto
01/

Every GPU plan we sell

Orderable today, with the cities that have capacity for each size. The configurator shows the live price for your city.

Plan
Spec
Cities
Price

A16, 1/8 GPU entry

Inference, encoding, CUDA dev boxes

2 GB VRAM2 vCPU8 GB RAM50 GB NVMe1 TB transfervcg-a16-2c-8g-2vram
7Bangalore, Chicago, Frankfurt, New York, Silicon Valley, Singapore, Tokyo

A40, 1/24 GPU

Ampere card, larger NVMe and transfer

2 GB VRAM1 vCPU5 GB RAM90 GB NVMe3 TB transfervcg-a40-1c-5g-2vram
1New York
$120/moDeploy

A16, 1/4 GPU

Image generation, mid-size models

4 GB VRAM2 vCPU16 GB RAM80 GB NVMe2 TB transfervcg-a16-2c-16g-4vram
6Bangalore, Chicago, Frankfurt, Silicon Valley, Singapore, Tokyo
$180/moDeploy

A16, 1/2 GPU

Half a card, 8 GB of GPU memory

8 GB VRAM3 vCPU32 GB RAM170 GB NVMe3 TB transfervcg-a16-3c-32g-8vram
6Bangalore, Chicago, Frankfurt, Silicon Valley, Singapore, Tokyo
$360/moDeploy

2x A16 multi-gpu

Two full cards, parallel inference workers

32 GB VRAM12 vCPU128 GB RAM700 GB NVMe10 TB transfervcg-a16-12c-128g-32vram
2Bangalore, Silicon Valley
$1,440/moDeploy

4x A16 multi-gpu

Four cards, 64 GB of GPU memory in one box

64 GB VRAM24 vCPU256 GB RAM1.2 TB NVMevcg-a16-24c-256g-64vram
1Bangalore
$2,873.75/moDeploy

Our retail pricing in USD at the lowest-priced location; some metros carry a surcharge and not every plan is offered in every location. The configurator shows live pricing and availability for your chosen city. A single full A16 (16 GB) and the larger A40 tiers appear in the configurator only while a city has capacity for them.

02/

Rent a slice of the card, or several cards

How fractional GPU works: a fixed share of the silicon and its memory, not a time-slice you queue for.

a/

1/8 to 1/2 of a GPU

A dedicated share of the card with its own slice of GPU memory (2 to 8 GB). CUDA works exactly as on a full card; you just size the memory to your model or workload.

2 to 8 GB VRAMfrom $90/mo
b/

2x and 4x A16 boxes

Scale to 32 or 64 GB of total GPU memory across multiple cards for parallel inference workers or batch pipelines.

32 or 64 GB VRAMBangalore, Silicon Valley
c/

CUDA-ready on first boot

GPU plans deploy CUDA-ready Linux images with NVIDIA drivers matched to the plan family, so nvidia-smi works on first boot. Full root, reinstall from the panel, your keys.

nvidia drivercuda toolkitroot access
03/

Seven cities with GPU capacity

Plan counts per city as of today. Capacity moves; the configurator is the source of truth at checkout.

  • 01/BangaloreIndia5 plans · up to 4x A16
  • 02/Silicon ValleyUnited States4 plans · up to 2x A16
  • 03/ChicagoUnited States3 plans · up to 1/2 A16
  • 04/FrankfurtGermany3 plans · up to 1/2 A16
  • 05/SingaporeSingapore3 plans · up to 1/2 A16
  • 06/TokyoJapan3 plans · up to 1/2 A16
  • 07/New YorkUnited States2 plans · A16 1/8, A40 1/24
  • networkVultr, 10 Gbps ports
  • ipv4dedicated, included
  • imagesUbuntu, Debian + drivers
  • deploy timeminutes after payment
  • supportthe engineer who runs it
04/

Honest sizing: A16 and A40, not H100

What these cards are for, and what they are not.

Runs well here

Inference on small and mid-size models, image generation on the 8 to 16 GB tiers, video encoding and transcoding, virtual desktops and cloud gaming, CUDA development and CI runners.

Not the right tool

Training or serving frontier-scale LLMs. An A16 is not an H100, and we would rather say so here than after you have deployed. If your job needs 80 GB of HBM, rent that class of card elsewhere and keep us for everything around it.

Why monthly GPU

A flat monthly price from your prepaid balance instead of a metered hourly bill. Top up by PayPal or crypto, deploy, and know what the month costs before it starts.

05/

GPU VPS questions

Straight answers. Anything else, ask before you order.

Ask us a question first

01/Which GPU models do you offer?

NVIDIA A16 (fractional slices up to 4x full cards) and an entry NVIDIA A40 slice. We do not sell A100 or H100 instances, and we will not pretend an A16 substitutes for one.

02/Can I run or train large language models on these?

Inference on small and mid-size models fits the 8 to 64 GB tiers well. Fine-tuning small models is possible on the multi-GPU boxes. Large-scale training belongs on data-center cards with more memory bandwidth than this class offers.

03/Do the images come with drivers and CUDA?

Yes. GPU plans deploy CUDA-ready Linux images with NVIDIA drivers matched to the plan family, so nvidia-smi works on first boot.

04/How does billing work?

Plans are monthly and paid from your prepaid account balance. Top up with PayPal or any of 18 cryptocurrencies including Bitcoin and Monero; no card or ID required.

05/Is every plan available in every location?

No. GPU capacity varies by city across the seven locations that carry it, and some metros price slightly higher. The order form shows live availability and the exact price for your chosen location.

Deploy a CUDA-ready GPU server

From a $90 slice to a 4x A16 box. Seven cities, monthly pricing, crypto welcome.