Managed model endpoints
Call supported models without provisioning containers, selecting GPU hardware, or maintaining a private model deployment.
- No customer-managed inference servers
- Documented model endpoints
- One authentication pattern
Use the model without running the deployment