This API calls models hosted by Serverless Inference. To manage Dedicated Inference gateways, deployments, and capacity claims, see the Dedicated Inference API reference. To compare Serverless Inference, Dedicated Inference, and Inference on CKS, see About CoreWeave Inference.
Base URL
Access the Serverless Inference service at:Prerequisites
To call the Serverless Inference API, you need:- A CoreWeave Forge account with Serverless Inference credits.
- A valid CoreWeave Forge API key.
-
Install the required libraries:
[YOUR-TEAM]/[YOUR-PROJECT]. If you don’t specify these, Serverless Inference uses your default entity and the project name inference.
Available methods
The Serverless Inference API provides OpenAI-compatible endpoints for interacting with foundation models. The following methods are available:- Chat Completions: Create chat completions using foundation models.
- List Models: Get all available models and their IDs.
Authentication
All API requests require authentication using your CoreWeave Forge API key. Create an API key at forge.coreweave.com/settings. Include your API key in the request headers:- For the OpenAI SDK, set the
api_keyparameter. - For direct API calls, use
Authorization: Bearer [YOUR-API-KEY].
Error handling
For a complete list of error codes and how to resolve them, see API errors.Next steps
After you have your API key, continue with one of the following:- Try the usage examples to see how the API works.
- Explore models in the Serverless Inference UI.
- Check usage limits for your account.