Develop and evaluate
Track experiments, evaluate and improve agents and LLM applications, prototype in notebooks, and manage models across their lifecycle.
Weights & BiasesW&B Models for experiment tracking, hyperparameter sweeps, artifacts, reports, and automations, with W&B Weave alongside for LLM and agent evaluation.Explore Weights & BiasesARIAAI research agent that analyzes experiments, explains results, and builds visualizations and reports. Preview.Agent LensIntelligent agent observability and improvement.NotebooksReactive Python notebooks for exploring data, prototyping model workflows, and analyzing runs and artifacts.RegistryVersioned hub for models and datasets with lineage and deployment handoff.
Serverless services
Serve, fine-tune, and run models on CoreWeave infrastructure without managing it.
Post-TrainingPost-train and fine-tune LLMs with Serverless RL and SFT, and distill specialized models from real application traffic.Serverless InferenceOpen-source foundation models through an OpenAI-compatible API, integrated with Weave.Serverless SandboxOn-demand, isolated compute environments you create and discard from Python.
Premiers pas
Premiers pas avec W&B
Comparez les produits, puis suivez le guide de démarrage rapide adapté à votre cas d’usage.
Exemples de code et notebooks
Tutoriels pratiques, guides pas à pas et notebooks pour chaque produit Forge.