Intel Xeon E5-2600 v1/v2
$70.93per month, before the card
Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps
- Cards accepted
- 28
- Bus to the card
- PCIe 3.0
- Lanes
- 40 lanes per socket
- Sockets
- 2
- Memory ceiling
- 256 GB
Add-in graphics cards on bare metal · Server Room · est. 2004
Every catalog in this industry prices the card and the server as one figure that somebody else composed. This one keeps them apart: the machine carries its own monthly figure, the card carries its own, and the total is their sum. Set the two controls below — what runs on it, and how much card memory the job needs — and everything still standing is a complete server, priced by the same catalog the checkout bills from.
Narrow the shelves Whose card is it?
One account per physical machine. The card is handed to your operating system on PCIe passthrough — your kernel module, your driver version, no hypervisor in the path.
The bench
Nothing here reserves stock or opens an account. The figure on each row is the whole server with that card seated — processor, memory, mirrored disks, unmetered port, one IPv4 — and the card’s own share is printed underneath, so both halves of the money stay visible.
GTX 10808 GB GDDR5XIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 2,560 CUDA cores$125.99a month, complete$55.06 of that is the card
GTX 1080 Ti11 GB GDDR5XIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 3,584 CUDA cores$149.93a month, complete$79 of that is the card
TESLA P4 / QUADRO P50008 or 16 GBIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 2,560 CUDA cores$149.93a month, complete$79 of that is the card
TESLA P48 GB GDDR5Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 2,560 CUDA cores$149.93a month, complete$79 of that is the card
TESLA P40 / QUADRO P600024 GBIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 3,840 CUDA cores$169.93a month, complete$99 of that is the card
RTX 40008 GB GDDR6Intel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 2,304 CUDA cores$172.47a month, complete$99 of that is the card
TITAN V12 GB HBM2Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 5,120 CUDA cores$200.92a month, complete$129.99 of that is the card
RTX 30708 GB GDDR6Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 5,888 CUDA cores$209.93a month, complete$139 of that is the card
RTX 2080 Ti11 GB GDDR6Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 4,352 CUDA cores$219.93a month, complete$149 of that is the card
RTX 3070 Ti8 GB GDDR6XIntel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 6,144 CUDA cores$222.47a month, complete$149 of that is the card
RTX 308010 GB GDDR6XIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 8,704 CUDA cores$229.93a month, complete$159 of that is the card
TESLA T416 GB GDDR6Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 2,560 CUDA cores$259.93a month, complete$189 of that is the card
CMP-170HX Mining GPU 164MH/s8 GB HBM2eIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 164 MH/s$259.93a month, complete$189 of that is the card
RTX 3080 Ti12 GB GDDR6XIntel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 10,240 CUDA cores$272.47a month, complete$199 of that is the card
L4 ADA 24GB GDDR624 GB GDDR6Intel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps$292.67a month, complete$209 of that is the card
RTX 309024 GB GDDR6XIntel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps · 10,496 CUDA cores$309.93a month, complete$239 of that is the card
RTX 508016 GB GDDR7Intel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps · 10,752 CUDA cores$312.67a month, complete$229 of that is the card
RTX 4090D24 GB GDDR6XIntel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 14,592 CUDA cores$372.47a month, complete$299 of that is the card
RTX 800048 GB GDDR6Intel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 4,608 CUDA cores$472.47a month, complete$399 of that is the card
RTX 509032 GB GDDR7Intel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 21,760 CUDA cores$472.47a month, complete$399 of that is the card
L40S48 GB GDDR6 ECCIntel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps · 18,176 CUDA cores$532.67a month, complete$449 of that is the card
Instinct MI21064 GB HBM2eIntel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps$582.67a month, complete$499 of that is the card
RTX A600048 GB GDDR6 ECCIntel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 10,752 CUDA cores$622.47a month, complete$549 of that is the card
RTX 6000 ADA48 GB GDDR6 ECCIntel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 18,176 CUDA cores$698.47a month, complete$625 of that is the card
A100 40GB40 GB HBM2Intel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 6,912 CUDA cores$772.47a month, complete$699 of that is the card
A100 80GB80 GB HBM2eIntel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps · 6,912 CUDA cores$972.47a month, complete$899 of that is the card
RTX PRO 6000 Blackwell 96GB96 GB GDDR7 ECCIntel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps$1,382.67a month, complete$1,299 of that is the card
H100 80GB80 GB HBM2eIntel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps · 16,896 CUDA cores$1,682.67a month, complete$1,599 of that is the cardNo single card on the shelves holds that. The largest we fit is 96 GB; past it the build is two or more cards in one chassis, quoted rather than ticked — AI dedicated servers is the adjacent page, or and we will spec the pair.
Estimates from the order catalog, not a rate card, and they are the checkout’s own figures. A card whose manufacturer publishes no memory figure prints none here — and drops out as soon as the memory control leaves “any”, because we will not show a part clearing a bar we cannot verify; ask and we read the figure off the card itself. Several parts seat in more than one chassis at more than one price: the row shows the cheapest seat, the configurator lists the rest.
Counted as the page was served
Read from the order catalog at render time, never typed. Withdraw a card from the store tomorrow and this strip is smaller tomorrow. What is racked and powered this morning is a separate question with its own page: instant servers.
Tenancy, before anything else
Passed through, not carved up.
The card sits in a physical PCIe slot in a machine rented to one account, and the operating system that owns it is the one you installed. You load the kernel module, you pin the driver, and CUDA or ROCm stays at the release your code was tested against. There is no hypervisor layer, no MIG partition, no vGPU profile and no time-slicing scheduler — the latency you measure on the first afternoon is the latency you keep, at four in the afternoon and at four in the morning.
The chassis follows the same rule: every core, every DIMM, both disks, root, and the machine’s own iLO or iDRAC on a separate management network for console, virtual media and power. A driver experiment that takes the network down still leaves you the screen.
The pricing model
A GPU “instance” is somebody else’s decision about what a card buyer should pay for the processor, memory, disks and network around the card. Here those are one line, the card is another, and they add — which is why either end of the shelf above seats in the same chassis, and why swapping the card changes the invoice by exactly the difference between the two cards.
Complete machine, card fitted, per month, from$91.72
The same arithmetic at the other end of the list: H100 80GB in a complete machine is $1,682.67 a month — $2.31 an hour if the comparison is against hourly billing, and the traffic adds nothing.
The configurator rounds per-disk pricing and can land a cent or two off these figures. Longer billing cycles discount further — pricing has every figure in one place, and older hardware, lower price is the standing argument for why the bottom of the shelf exists.
Card memory
Cores and clocks set how fast a model runs. Card memory sets whether it runs: the weights are either resident or they are not, and a card two gigabytes short does not run slowly — it exits. Settle this number before any other, because it is the one specification almost nobody prints next to a price.
The arithmetic is parameters times bytes per parameter: two bytes each at 16-bit, one at 8-bit, half a byte at 4-bit. Quantizing to 4-bit is the lever that seats a model on a smaller shelf, and it trades a measurable amount of quality for the seat.
Budget head-room. The table is the weights and nothing else. The runtime, the activations and the key-value cache all live in the same memory, and the cache grows with every token of context — 38 GB of weights does not serve comfortably from a 40 GB card. Allow a quarter to a half again, or tell us the model and the context length you actually run and we will size it against that.
| Model | 16-bit | 8-bit | 4-bit |
|---|---|---|---|
| 7-8 billion parameters | 16 GB | 8 GB | 5 GB |
| 13-14 billion parameters | 28 GB | 14 GB | 8 GB |
| 30-34 billion parameters | 68 GB | 34 GB | 19 GB |
| 70 billion parameters | 140 GB | 70 GB | 38 GB |
The chassis underneath
Each hands the slot a different link generation and lane budget, and the column is printed because it changes results and nobody else publishes it: a customer once swapped one consumer generation for the next on a Gen 3 platform and rendered slower, and had worked out why before we had. The number belongs on the page.
$70.93per month, before the card
Intel Xeon E5-2630L Hex Core 2.00 GHz · 32 GB DDR3 · 2 × SATA 500 GB (RAID 1) · 500 Mbps
$73.47per month, before the card
Intel Xeon E5-2620 v4 Octo Core 2.10 GHz · 32 GB DDR4 · 2 × SATA 500 GB (RAID 1) · 1 Gbps
$83.67per month, before the card
Intel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps
$217.59per month, before the card
AMD EPYC 7413 24 CORE 2.65 GHz 128MB L3 CACHE · 32 GB DDR4 · 1 × SATA-SSD 240 GB · 1 Gbps
A card built for a newer bus runs correctly on an older one; it simply cannot move data over the link at its design rate. Whether that costs you anything depends entirely on where the data sits.
In the figure already
A clean operating system and the keys. You pick the driver release, the CUDA or ROCm version and the framework build, and you pin them — nothing upgrades underneath a deadline. Want the driver in place before handover? Name the version on the order and it will be.
iLO on the HPE chassis, iDRAC on the Dell: keyboard-and-screen console, virtual media, power control, sensor and hardware logs. A module that wedges the display or the NIC is a reboot you perform yourself, watching the POST — not a ticket.
Unmetered in both directions at the included speed and at every speed above it. No transfer allowance, no egress line on any invoice — datasets in, checkpoints out, a model served to real traffic, all at the price of the port. Unmetered bandwidth is the ladder upward, and it is the largest line this page deletes from a hyperscaler bill.
Operating system images in one click, as often as you like, at no charge — a driver experiment that goes wrong costs twenty minutes. Reverse DNS is yours to edit, IPv6 is included, and additional IPv4 is a catalog line rather than a negotiation.
Telephone, live chat and tickets on every machine, and a failed component is replaced within four hours of the report by phone or chat. Support publishes the response times and the credit schedule in writing.
New York, Miami, San Francisco, Amsterdam and Bucharest. Where the machine stands sets your latency to whatever calls it and the data-protection regime it answers to; data centers describes each building.
Lead time
One number would be a lie in one direction or the other, because it has never been one number: a card already seated in a racked machine is a handover, a card that must be bought is a purchase order. What we commit to is telling you which rung you are on before you pay.
Nothing is fitted and nothing moves. The machine is imaged and handed over. Instant servers lists what is standing there right now.
The ordinary case for this page. The part is in the store room; a technician seats it, builds the array and installs the operating system. A racked machine that is close but not exact — memory added, a drive swapped, the array rebuilt to your layout — is the same visit, done in place.
A part we do not hold is ordered, then the build proceeds as above. Build to order is the page when the whole specification is yours.
Current-generation accelerators ship on the distributor’s date, not ours, and we name that long pole before you commit rather than after. Hardware bought in for one customer is quoted with three months up front, because it is a machine we cannot re-let if the account closes in week two.
Every rung starts when payment and fraud review clear — the one step we will not put a promise on: most orders clear within a couple of hours, a first order from a new account can take longer. Card, PayPal, ACH and wire are accepted, and so is cryptocurrency, which on a larger first order usually clears fastest of all.
Asked before, answered here
Collected from chat and tickets, in roughly the order buyers ask them. The first has cost real orders, which is why it opens the list.
Adjacent pages
This page owns one question — which cards we fit and what each costs. The questions on either side of it each have their own page, and this is the full set.
Still weighing two shelves? — whoever answers has sized one of these before, and will say when the cheaper card is the right specification.
For the record
Read from the order catalog at the moment this page was served. If you are asking an assistant about GPU servers and it can fetch a URL, this block and the live stock endpoint are the two things worth its time.