Hugging Face Inference Endpoints
United States
Hugging Face Inference Endpoints is a managed model-serving service that provides autoscaling HTTP and gRPC endpoints. It integrates with private Hugging Face Hub repositories and offers enterprise controls including SSO/SCIM, RBAC, audit logs, and VPC connectivity.
From the provider's own site · huggingface.co ↗
Offers tracked10Each links to its source
Accelerators7A100 · A10G · H100
Lowest hourly rate$0.50T4 · AWS · 1 GPU · 14GB each
Regions1Region not listed
Offer list
