GPU ComputeRent dedicated GPUs on demand across a distributed network.0.00 reviewsCloud Infrastructure & Hosting
Pre-Built EnvironmentsPre-configured VMs built for specific workflows such as CI/CD runners and AI agents, ready in minutes with all dependencies pre-installed.0.00 reviewsCloud Infrastructure & Hosting
Dedicated Virtual MachinesSpin up GPU-accelerated VMs with full root SSH access, Docker, NVIDIA drivers, persistent storage across reboots, and custom container image support.0.00 reviewsCloud Infrastructure & Hosting
Inference APIOpenAI-compatible endpoint for hosted open models including GPT-OSS 120B, GPT-OSS 20B, Gemma 4 31B, GLM 5.2, and DeepSeek-OCR 2.0.00 reviewsModel Hosting & Inference
Batch Inference APIRun large offline batch inference jobs at half the per-token rate, with results delivered within 24 hours.0.00 reviewsModel Hosting & Inference
GPU ComputeRent dedicated GPUs on demand across a distributed network.Pricing modelUsage BasedStarting priceStarting at $0.18/hr (RTX 3090); H100 $2.60/hr; A100 from $0.80/hr; RTX 4090 $0.29/hr; RTX 5090 $0.39/hr; RTX 4080 $0.39/hr. No contracts, no minimums.
Pre-Built EnvironmentsPre-configured VMs built for specific workflows such as CI/CD runners and AI agents, ready in minutes with all dependencies pre-installed.Pricing modelUsage BasedStarting pricePriced per dedicated VM per hour; example RTX 3090 runner at $0.18/hr, RTX 4090 at $0.29/hr.
Dedicated Virtual MachinesSpin up GPU-accelerated VMs with full root SSH access, Docker, NVIDIA drivers, persistent storage across reboots, and custom container image support.Pricing modelUsage BasedStarting pricePer-GPU, per-hour; same hourly GPU rates apply. No contracts, no minimums.
Inference APIOpenAI-compatible endpoint for hosted open models including GPT-OSS 120B, GPT-OSS 20B, Gemma 4 31B, GLM 5.2, and DeepSeek-OCR 2.Pricing modelUsage BasedStarting pricePer million tokens; GPT-OSS 120B $0.15 in / $0.60 out; GPT-OSS 20B $0.05 / $0.20; GLM 5.2 $1.82 / $5.72; DeepSeek-OCR 2 $0.039. No minimum spend; streaming included; batch at 50% off.
Batch Inference APIRun large offline batch inference jobs at half the per-token rate, with results delivered within 24 hours.Pricing modelUsage BasedStarting price50% off per-token rates; results within 24 hours.