What is Inference?
AI model inference endpoint with over 70 models; run serverless, dedicated, or batch inference with an Inference Router.
Product typeModel Hosting & Inference
PricingStarting at $0.05/M tokens; Pay only for what you use · Usage Based
Deployment
APINot available
Integrations and access
Platforms, integrations, and language support vary by plan and region. Confirm final requirements with the vendor.
Platforms
IntegrationsAPI and webhooks
Languages
Support
What reviewers say
Loading community reviews…
Enterprise readiness
Compliance claims are normalized from current vendor documentation and independently reviewed by SOTA2.
SOC 2Kata Containers for App Platform isolation Encrypted environment variables Standard Kubernetes network policies Kubernetes Private Distributors List membership Trusted Sources for Managed Databases VPC isolation for worker nodes HTTPS and TLS by default for Functions Benchmarks and comparisons
Independent benchmarksStructured task results are being verified.
Side-by-side comparisonsComparison workspaces will be available in a later release.