
Back to all servers
Hybrid AI + CPU Server — 6 GB
RTX 2060 paired with a 40-core host. For the firm running hybrid CPU + GPU pipelines — batch ML preprocessing followed by accelerated inference on one machine.
Commitment term
$5797.92
Billed once · 24-month commit · $241.58/mo
Save ~24% vs 1-month
99.9% uptime SLA Provisions in < 15 min
Specifications
CPU
40 cores (Dual Gold 6148)
RAM
128 GB
Storage
120 GB SSD + 960 GB SSD
VRAM
6 GB
What's included
- 40-core host
- RT + Tensor cores
- Bare-metal dedicated
Best for
CPU+GPU hybrid pipelinesSpeech + LLM serveBatch ML preprocessing
All pricing tiers
| Term | Per-month | Total commit |
|---|---|---|
| 1 month | $317.87 /mo | $317.87 |
| 3 months | $308.33 /mo | $924.99 |
| 12 months | $264.89 /mo | $3178.68 |
| 24 months | $241.58 /mo | $5797.92 |
You might also consider

Production AI Server — 8 GB Ampere
Ampere RTX 3060 Ti with strong tensor throughput per dollar. Solid for an in-house image generator, a 7B-class LLM endpoint, or one node in a small render farm.
$308.33 /mo

Production AI Workstation — 24 GB Pro
Twenty-four gigs of pro-grade VRAM and 56 GB system RAM — the sweet spot for firms running a production inference endpoint with predictable SLA expectations.
$275.00 /mo

Practice AI Workstation — 6 GB Pro
RTX-class GPU on a 128 GB host. RT + Tensor cores give the firm a real entry-level production card for refiner pipelines and small LLM serving.
$256.73 /mo