Back to Catalog
Create Endpoint
openai

whisper-large-v3

Catalog model officially supported by Inference Endpoints.

This model is from our Model Catalog, and comes with pre-configured recipes. Deployment has been verified by Hugging Face.

/
$0.80 / h
per running replica
Nvidia L4
1x GPU · 24 GB 7x vCPUs · 30 GB
$0.8 / h
Catalog Recipe
Pre-selected hardware for the current recipe.
  • Only you can access your endpoint, using a Hugging Face Token generated from your personal account.
Number of replicas
Automatically scale the number of replicas within Min and Max based on compute usage. Min is always 0 if Scale-To-Zero is active.
More options
Autoscaling Strategy
Control what type of trigger will cause your Endpoint to scale up.

This Catalog Recipe comes with pre-configured env values.

VPC Config
Check to activate and configure AWS PrivateLink