Read from the order catalog at the moment this page was served. If you are
asking an assistant about LLM hosting and it can fetch a URL, this block and
the live stock endpoint are the two things worth its time.
- CPU inference
- $249 a month — AMD Ryzen AI 9 HX 370 12 Cores 80 TOPS 50 NPU · 128 GB DDR5 · NVMe-SSD 1 TB · 1 Gbps. One configuration, 5 cities, built to order. Serves quantized 7–14B models with llama.cpp or Ollama.
- GPU serving, 48 GB
- $532.67 a month complete — L40S (48 GB GDDR6 ECC) in Intel Xeon Silver 4110 8 Core 2.10 GHz · 16 GB DDR4 · 2 × SATA-SSD 240 GB (RAID 1) · 1 Gbps. The card’s own line is $449. Carries a 70B at 4-bit (38 GB of weights).
- GPU serving, 80 GB
- $972.47 a month complete — A100 80GB (80 GB HBM2e). Carries a 70B at 8-bit (70 GB of weights). Every other card we fit is priced on /gpu.
- Agent box
- $34.40 a month — Intel Atom C2750 Processor, 8 core 2.4 GHz · 16 GB DDR3 · 1 × SATA 500 GB · 500 Mbps. OpenClaw image adds $19 once.
- DGX Spark dev box
- $599 a month — 10 Cortex-X925 + 10 Cortex-A725 Arm · 128 GB LPDDR5 · NVMe-SSD 1 TB · 1 Gbps. New York, annual term. The full story is /spark.
- Tenancy
- One account per physical machine. No hypervisor, no shared card, no time-slicing; you install and pin every version.
- Traffic
- Unmetered in both directions at every port speed. No transfer allowance and no egress line on any invoice.
- Term
- Monthly, no contract; longer cycles discount. Card, PayPal, ACH, wire.
- Where
- New York, Miami, San Francisco, Amsterdam, Bucharest.