GPU Compute Rental and Hosted Inference
Rent GPU infrastructure for model training, fine-tuning, and other demanding workloads. Choose from the systems below and pay by the hour. We manage the physical hardware and hosting so your team can focus on building and running its models.
We are preparing hosted inference endpoints for open models. The table below shows the planned models, token pricing, and expected performance. These endpoints are not active yet, but you can contact us about early access or a custom deployment.
NODES
| 2x Xeon 128c @ 2.3 GHz | 2 TB @ 6400 MT/s | 2.30 TB HBM3e | 8x 3.84 TB @ 14 GB/s | $59.12/HR | RENT | |
| 2x Xeon 112c @ 2.1 GHz | 2 TB @ 5600 MT/s | 1.44 TB HBM3e | 8x 3.84 TB @ 7 GB/s | $47.12/HR | ||
| 2x EPYC 96c @ 2.75 GHz | 1 TB @ 4800 MT/s | 564 GB HBM3e | 4x 3.84 TB @ 7 GB/s | $17.56/HR | ||
| 1x EPYC 96c @ 5.0 GHz | 1 TB @ 8533 MT/s | 1.73 TB HBM4 | 4x 7.68 TB @ 14 GB/s | $39.84/HR |
ENDPOINTS
| $3.00 | $15.00 | $0.30 | 463 MS | 73 TOK/S | ||
| $0.475 | $1.493 | $0.095 | 192 MS | 98 TOK/S |
CONTACT
SALES
- [email protected]
CONTACT
- [email protected]
PHONE
- UNAVAILABLE
Have a question about availability, pricing, or a custom configuration? Send us a message and we’ll get back to you.