Skip to main content
This reference describes the Serverless Inference REST API, which lets you call foundation models programmatically from your own applications. Use it to integrate hosted inference into services, scripts, or notebooks without managing model infrastructure.
This API calls models hosted by Serverless Inference. To manage Dedicated Inference gateways, deployments, and capacity claims, see the Dedicated Inference API reference. To compare Serverless Inference, Dedicated Inference, and Inference on CKS, see About CoreWeave Inference.

Base URL

Access the Serverless Inference service at:

Prerequisites

To call the Serverless Inference API, you need:
  • A CoreWeave Forge account with Serverless Inference credits.
  • A valid CoreWeave Forge API key.
  • Install the required libraries:
If you belong to more than one team, or want to attribute your usage to a project, you’ll also need team and project IDs. In code samples, these appear as [YOUR-TEAM]/[YOUR-PROJECT]. If you don’t specify these, Serverless Inference uses your default entity and the project name inference.

Available methods

The Serverless Inference API provides OpenAI-compatible endpoints for interacting with foundation models. The following methods are available:

Authentication

All API requests require authentication using your CoreWeave Forge API key. Create an API key at forge.coreweave.com/settings. Include your API key in the request headers:
  • For the OpenAI SDK, set the api_key parameter.
  • For direct API calls, use Authorization: Bearer [YOUR-API-KEY].

Error handling

For a complete list of error codes and how to resolve them, see API errors.

Next steps

After you have your API key, continue with one of the following:
Last modified on September 30, 2026