Skip to main content
Serverless Inference gives you access to leading open-source foundation models through an OpenAI-compatible API. You can build AI applications and agents without signing up for a hosting provider or self-hosting a model.

Try out inference in the UI

Navigate to https://forge.coreweave.com/inference to explore available models and try them out in the Playground. For more information on the web interface, see the Try models in the Playground UI.

Use Serverless Inference through the API

This Python example uses Serverless Inference to send a chat completion request to an LLM.

Next steps

Last modified on September 29, 2026