> ## Documentation Index
> Fetch the complete documentation index at: https://wb-21fd5541-new-articles-log.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# New articles

> A weekly log of new articles published to the W&B documentation site.

This page lists new articles as they are published to the W\&B documentation site, with the newest articles at the top. It is updated automatically every Friday. Articles published in previous years are collapsed under a heading for that year.

Within each month, articles are grouped by the part of the documentation they belong to, such as **Models**, **Weave**, or **Support**. Auto-generated reference pages and localized content are not included.

{/* new-articles:start */}

## July 2026

### Models

<Card title="Link a Weave prompt to a collection" href="/models/registry/link_prompt/" arrow="true" horizontal>
  Track, version, and manage your AI prompts in the W\&B Registry by linking Weave prompt objects to a collection.
</Card>

<Card title="View Eval Tables" href="/models/evaltables/visualize-evaluation-tables/" arrow="true" horizontal>
  When you create an Eval Table, specify which columns contain inputs, outputs, and scores.
</Card>

<Card title="Create an Eval Table" href="/models/evaltables/create-an-evaluation-table/" arrow="true" horizontal>
  Learn how to create an Eval Table in W\&B.
</Card>

<Card title="Compare runs with Eval Tables" href="/models/evaltables/compare-runs/" arrow="true" horizontal>
  Use Eval Tables to compare inputs, outputs, and scores across multiple runs.
</Card>

<Card title="Overview" href="/models/evaltables/" arrow="true" horizontal>
  Learn how to create, compare, and visualize Eval Tables in W\&B.
</Card>

### Weave

<Card title="Compare evaluations" href="/weave/guides/evaluation/compare_evals/" arrow="true" horizontal>
  Visually compare two or more evaluations to spot regressions, improvements, and scoring differences across runs
</Card>

<Card title="Reference media in your own bucket (BYOB) using Weave Op" href="/weave/guides/tracking/byob-references/" arrow="true" horizontal>
  Render images and video that live in your own cloud bucket in Weave traces by logging their URIs, without importing the bytes into Weave.
</Card>

<Card title="Reference media in your own bucket (BYOB) using agent spans" href="/weave/guides/tracking/agents-byob-references/" arrow="true" horizontal>
  Render images and video that live in your own cloud bucket in the Weave Agents view by returning their URIs from agent tool spans, without importing the bytes into Weave.
</Card>

<Card title="Configure ingest sampling for Weave Self-Managed" href="/weave/guides/platform/ingest-sampling/" arrow="true" horizontal>
  Keep only a share of incoming traces to control storage and LLM scoring costs on a self-managed Weave instance
</Card>

<Card title="Bedrock Agents" href="/weave/guides/integrations/bedrock_agents/" arrow="true" horizontal>
  Trace Amazon Bedrock Agents invocations with Weave, capturing agent inputs, foundation model usage, and completion output.
</Card>

## June 2026

### Models

<Card title="LEET terminal UI" href="/models/app/leet-tui/" arrow="true" horizontal>
  Explore and compare local W\&B runs from the terminal with the LEET (Lightweight Experiment Exploration Tool) TUI.
</Card>

### Weave

<Card title="View agent activity" href="/weave/guides/tracking/view-agent-activity/" arrow="true" horizontal>
  Use W\&B Weave's Agents view to understand what your agent did, how much it cost, and exactly where things went right or wrong.
</Card>

<Card title="Trace sub-agents" href="/weave/guides/tracking/trace-sub-agents/" arrow="true" horizontal>
  Use Weave's sub-agent span to trace sub-agent delegations and view nested agent invocations.
</Card>

<Card title="Trace your agents" href="/weave/guides/tracking/trace-agents/" arrow="true" horizontal>
  Use the Weave SDK to instrument multi-turn agentic applications and view them in the Agents tab.
</Card>

<Card title="Batch logging for your agent" href="/weave/guides/tracking/trace-agents-batch/" arrow="true" horizontal>
  Manually log agent traces for frameworks that have already completed the LLM call and need to record it.
</Card>

<Card title="Pi extension" href="/weave/guides/integrations/agents/pi-dev-harness/" arrow="true" horizontal>
  Trace Pi agentic sessions, LLM calls, and tool executions in Weave.
</Card>

<Card title="OpenClaw plugin" href="/weave/guides/integrations/agents/openclaw-harness/" arrow="true" horizontal>
  Track OpenClaw agent sessions in W\&B Weave for observability and debugging.
</Card>

<Card title="OpenAI Agents SDK" href="/weave/guides/integrations/agents/openai-agents-sdk/" arrow="true" horizontal>
  Trace an agent built with the OpenAI Agents SDK using Weave.
</Card>

<Card title="Google ADK" href="/weave/guides/integrations/agents/google-adk/" arrow="true" horizontal>
  Trace an agent built with Google's Agent Development Kit (ADK) using Weave.
</Card>

<Card title="Codex plugin" href="/weave/guides/integrations/agents/codex-harness/" arrow="true" horizontal>
  Trace Codex agentic sessions, LLM calls, and tool executions in W\&B Weave.
</Card>

<Card title="Claude Code plugin" href="/weave/guides/integrations/agents/claude-code-harness/" arrow="true" horizontal>
  Track Claude Code sessions in W\&B Weave for observability and debugging.
</Card>

<Card title="Claude Agent SDK" href="/weave/guides/integrations/agents/claude-agents-sdk/" arrow="true" horizontal>
  Trace an agent built with the Claude Agent SDK using Weave.
</Card>

### ARIA

<Card title="Overview" href="/aria/overview/" arrow="true" horizontal>
  ARIA, W\&B's AI Research and Iteration Agent, is your personalized research assistant that helps you analyze and run experiments, explain results, identify patterns across runs, recommend next steps, build visualizations and reports, and more in W\&B.
</Card>

<Card title="Governance and security" href="/aria/governance/" arrow="true" horizontal>
  Learn how W\&B Admins can control how your organization's data is used by W\&B's AI features like ARIA.
</Card>

<Card title="Chat with ARIA" href="/aria/chat/" arrow="true" horizontal>
  The following page describes how to start a new chat, view your chat history, delete previous conversations, undock the chat window, reference a specific run in your prompt, add an image to a conversation, and provide feedback on ARIA's responses.
</Card>

<Card title="Use ARIA for autoresearch" href="/aria/autoresearch/" arrow="true" horizontal>
  Learn how to use ARIA, W\&B's AI Research and Iteration Agent, to analyze results and run experiments.
</Card>

### HiveMind

<Card title="W&B HiveMind" href="/hivemind/" arrow="true" horizontal>
  A shared dashboard for AI coding sessions. Track activity, spend, and outcomes across Claude Code, Cursor, Codex, Gemini CLI, and more with leaderboards, team views, and efficiency insights.
</Card>

### Support

<Card title="Why does my Weave cost or token estimate differ from my provider?" href="/support/weave/articles/why-does-my-weave-cost-or-token-estimate-differ-from-my-provider/" arrow="true" horizontal>
  Weave displays cost and token usage estimates based on data captured from your LLM calls, and discrepancies between Weave's numbers and your provider's invoice can be caused by the following issues.
</Card>

<Card title="Why is my W&B run slow to initialize or upload?" href="/support/models/articles/why-is-my-wandb-run-slow-to-initialize-or-upload/" arrow="true" horizontal>
  Slow wandb.init() or sluggish metric uploads are usually caused by network latency, large media payloads, high logging frequency, or slow startup of the W\&B service process.
</Card>

<Card title="Why is my sweep agent not picking up new runs?" href="/support/models/articles/why-is-my-sweep-agent-not-picking-up-new-runs/" arrow="true" horizontal>
  If your sweep agent starts but does not receive new run configurations, or receives one run and then idles, there are several common causes.
</Card>

<Card title="Why is my run showing as crashed?" href="/support/models/articles/why-is-my-run-showing-as-crashed/" arrow="true" horizontal>
  W\&B marks a run as Crashed when it stops receiving heartbeats from the process that called wandb.init(), without the process having called wandb.finish().
</Card>

<Card title="Why is console output not captured for my run?" href="/support/models/articles/why-is-console-output-not-captured-for-my-run/" arrow="true" horizontal>
  W\&B captures your script's stdout and stderr and stores it as output.log on the run's Files tab.
</Card>

<Card title="Why does my API key fail with 'must be 40 characters long'?" href="/support/models/articles/why-does-my-api-key-fail-with-must-be-40-characters/" arrow="true" horizontal>
  W\&B now issues longer API keys (about 86 characters).
</Card>

<Card title="Why do my workspace settings not persist between sessions?" href="/support/models/articles/why-do-my-workspace-settings-not-persist/" arrow="true" horizontal>
  Workspace layout (panels, filters, grouping) persists only when you save a view.
</Card>

<Card title="Why are my metrics missing from wandb.log()?" href="/support/models/articles/why-are-my-metrics-missing-from-wandb-log/" arrow="true" horizontal>
  If metrics logged with wandb.log() are not appearing in the W\&B UI, there are several common causes.
</Card>

<Card title="I deleted my team and now I can't create a new one — what do I do?" href="/support/models/articles/i-deleted-my-team-and-now-i-cant-create-a-new-one/" arrow="true" horizontal>
  When you sign up for W\&B, the platform automatically creates a personal team with the same name as your username.
</Card>

<Card title="How do I use W&B with JAX?" href="/support/models/articles/how-do-i-use-wandb-with-jax/" arrow="true" horizontal>
  W\&B has no JAX-specific integration.
</Card>

<Card title="How do I use the parallel coordinates chart in W&B?" href="/support/models/articles/how-do-i-use-the-parallel-coordinates-chart-in-wandb/" arrow="true" horizontal>
  The parallel coordinates chart shows how hyperparameters relate to metrics across many runs.
</Card>

<Card title="How do I update run config, tags, and notes via the W&B API?" href="/support/models/articles/how-do-i-update-run-config-tags-and-notes-via-the-wandb-api/" arrow="true" horizontal>
  After a run finishes, use the Public API guide to edit config, display name, tags, and notes without re-running the experiment.
</Card>

<Card title="How do I set up W&B alerts and notifications?" href="/support/models/articles/how-do-i-set-up-wandb-alerts-and-notifications/" arrow="true" horizontal>
  You can set up alerts in and notifications using the W\&B Settings page.
</Card>

<Card title="How do I page through large API results in W&B?" href="/support/models/articles/how-do-i-paginate-through-large-api-results-in-wandb/" arrow="true" horizontal>
  You can page through API result using the standard lazy-iterator pattern and perpage parameter.
</Card>

<Card title="How do I log NLP metrics and text outputs in W&B?" href="/support/models/articles/how-do-i-log-nlp-metrics-and-text-outputs-in-wandb/" arrow="true" horizontal>
  You can log corpus-level NLP scores (BLEU, ROUGE, perplexity) with wandb.log() and per-example outputs with wandb.Table.
</Card>

<Card title="How do I log gradients and model weights with wandb.watch()?" href="/support/models/articles/how-do-i-log-gradients-and-model-weights-with-wandb-watch/" arrow="true" horizontal>
  wandb.watch() hooks into a PyTorch model's parameters and gradients and logs histograms of their values at regular intervals.
</Card>

<Card title="How do I invite a user to my W&B team?" href="/support/models/articles/how-do-i-invite-a-user-to-my-wb-team/" arrow="true" horizontal>
  Only team admins can send invitations.
</Card>

<Card title="How do I download the console log file from a run?" href="/support/models/articles/how-do-i-download-the-console-log-file-from-a-run/" arrow="true" horizontal>
  W\&B stores your script's stdout and stderr as output.log (or multipart chunks under logs/).
</Card>

<Card title="How do I create a new team in W&B?" href="/support/models/articles/how-do-i-create-a-new-team-in-wandb/" arrow="true" horizontal>
  Teams are the primary unit of collaboration in W\&B.
</Card>

<Card title="How do I connect to W&B Self-Managed?" href="/support/models/articles/how-do-i-connect-to-wandb-self-managed/" arrow="true" horizontal>
  W\&B Self-Managed is a self-hosted deployment that runs in your infrastructure.
</Card>

<Card title="Can I resume a run inside a sweep?" href="/support/models/articles/can-i-resume-a-run-inside-a-sweep/" arrow="true" horizontal>
  Run resumption is not supported inside a W\&B sweep.
</Card>

<Card title="API error code 422 - Invalid request parameters" href="/support/inference/articles/api-error-code-422-invalid-request-parameters/" arrow="true" horizontal>
  Invalid parameters on chat completion requests often surface as HTTP 422 (Unprocessable Entity) or HTTP 400 (Bad Request) depending on where validation runs.
</Card>

<Card title="API error code 404 - Model not found" href="/support/inference/articles/api-error-code-404-model-not-found/" arrow="true" horizontal>
  A 404 response from the W\&B Inference API means the server could not find the model or resource you asked for.
</Card>

<Card title="How do I log in to W&B Self-Managed?" href="/support/models/articles/how-do-i-log-in-to-wandb-self-managed/" arrow="true" horizontal>
  To authenticate to a W\&B Self-Managed instance, point the CLI and SDK at your instance URL before you log in:
</Card>

<Card title="Which Weave API should I use to trace my agent?" href="/support/weave/articles/which-weave-api-should-i-use-for-agents/" arrow="true" horizontal>
  It depends on what you're building:
</Card>

<Card title="What is the difference between @weave.op and weave.start_session?" href="/support/weave/articles/what-is-the-difference-between-weave-op-and-weave-sdk-for-agents/" arrow="true" horizontal>
  @weave.op traces individual Python functions and surfaces results in the Traces tab.
</Card>

## May 2026

### Models

<Card title="Automation tutorial overview" href="/models/automations/tutorial/" arrow="true" horizontal>
  Learn to build a project run-failure alert or a registry alias automation.
</Card>

<Card title="Tutorial: Registry artifact alias automation" href="/models/automations/registry-automation-tutorial/" arrow="true" horizontal>
  Build an automation that runs a webhook when a Registry artifact gets a specific alias like "production".
</Card>

<Card title="Tutorial: Project run-failure alert automation" href="/models/automations/project-automation-tutorial/" arrow="true" horizontal>
  Build a run-failure alert that sends a Slack notification when a run in your project fails.
</Card>

<Card title="Manage automations with the API" href="/models/automations/api/" arrow="true" horizontal>
  Programmatic automation management with the Python API. Create and update may be affected on some client versions. Prefer the W\&B App until the SDK fix ships.
</Card>

## April 2026

### Weave

<Card title="Haystack" href="/weave/guides/integrations/haystack/" arrow="true" horizontal>
  Trace Deepset Haystack pipelines with W\&B Weave using the WeaveConnector integration.
</Card>

<Card title="Write-ahead log" href="/weave/guides/tracking/write-ahead-log/" arrow="true" horizontal>
  Improve the resilience of W\&B Weave trace data capture with the write-ahead log
</Card>

<Card title="Vercel AI SDK" href="/weave/guides/integrations/vercel_ai_sdk/" arrow="true" horizontal>
  Trace Vercel AI SDK calls in Weave using OpenTelemetry
</Card>

<Card title="Set up automations" href="/weave/guides/evaluation/automations/" arrow="true" horizontal>
  Create event-driven automations that trigger actions based on monitor metrics and trace activity.
</Card>

### Platform

<Card title="Manage bucket storage and costs" href="/platform/hosting/managing-bucket-storage/" arrow="true" horizontal>
  Understand how W\&B uses object storage, how deletion maps to bucket bytes, and how to reduce storage usage and costs.
</Card>

### Sandboxes

<Card title="Secrets" href="/sandboxes/secrets/" arrow="true" horizontal>
  Learn how to manage secrets in Serverless Sandboxes.
</Card>

<Card title="Run commands in a sandbox" href="/sandboxes/run-commands/" arrow="true" horizontal>
  Learn how to run commands in a sandbox environment and read output and exit codes.
</Card>

<Card title="Tutorial: Train a PyTorch model" href="/sandboxes/mltrain-in-sandbox-tutorial/" arrow="true" horizontal>
  Learn how to train a PyTorch model in a Serverless Sandbox environment with this step-by-step tutorial.
</Card>

<Card title="Sandbox lifecycle" href="/sandboxes/lifecycle/" arrow="true" horizontal>
  Learn about the lifecycle of a Serverless Sandbox, including its states, how to wait for state changes, and how to stop a sandbox.
</Card>

<Card title="Tutorial: Invoke an agent in a sandbox" href="/sandboxes/invoke-agent-sandbox-tutorial/" arrow="true" horizontal>
  Learn how to invoke an OpenAI agent within a Serverless Sandbox environment
</Card>

<Card title="File operations" href="/sandboxes/file-access/" arrow="true" horizontal>
  Learn how to read, write, and mount files in Serverless Sandboxes.
</Card>

<Card title="Create sandboxes" href="/sandboxes/create-sandbox/" arrow="true" horizontal>
  Create W\&B Serverless Sandboxes that run code in isolated containers with their own filesystem, network, and process space.
</Card>

<Card title="Serverless Sandboxes" href="/sandboxes/" arrow="true" horizontal>
  On-demand, isolated compute environments that you can create, use, and discard from Python with W\&B Serverless Sandboxes.
</Card>

### Support

<Card title="Why doesn't `.call()` raise exceptions?" href="/support/weave/articles/weave-call-does-not-raise-exceptions/" arrow="true" horizontal>
  By default, Weave's .call() method captures exceptions and stores them in call.exception instead of raising them.
</Card>

<Card title="Why does my workspace load slowly?" href="/support/models/articles/workspace-loads-slowly-with-many-metric/" arrow="true" horizontal>
  Workspaces can load slowly when a project has many metrics, runs, or panels.
</Card>

<Card title="Why does my training hang with distributed training?" href="/support/models/articles/training-hangs-with-distributed-trainin/" arrow="true" horizontal>
  This article helps you resolve training hangs when you use W\&B with distributed training frameworks, so your runs can start and finish without stalling.
</Card>

<Card title="How do I fix service account errors like `Unauthorized` or missing runs?" href="/support/models/articles/service-account-unauthorized-or-runs-no/" arrow="true" horizontal>
  When you use a service account to automate W\&B workflows, you might encounter authorization errors or find that runs don't appear where you expect them.
</Card>

<Card title="How do I fix `Rate limit exceeded` errors when logging metrics?" href="/support/models/articles/rate-limit-exceeded-on-metric-logging/" arrow="true" horizontal>
  If you receive an HTTP 429 Rate limit exceeded error when you call wandb.log(), you're exceeding the rate limit quota for your project.
</Card>

<Card title="Why does my process stop responding when using Hydra with W&B?" href="/support/models/articles/process-hangs-when-using-hydra-with-wan/" arrow="true" horizontal>
  This page explains how to resolve unresponsive processes that occur when you start a process with Hydra alongside W\&B.
</Card>

<Card title="Why is my enterprise license not recognized?" href="/support/models/articles/enterprise-license-not-recognized/" arrow="true" horizontal>
  This page helps administrators of W\&B enterprise deployments diagnose common issues when W\&B doesn't recognize an enterprise license or its features are unavailable after you set the license.
</Card>

<Card title="How do I fix `Cuda out of memory` during a sweep?" href="/support/models/articles/cuda-out-of-memory-during-a-sweep/" arrow="true" horizontal>
  If you see Cuda out of memory during a sweep, refactor your code to use process-based execution.
</Card>

<Card title="How do I fix `CommError, Run does not exist` during a sweep?" href="/support/models/articles/commerror-run-does-not-exist-during-swee/" arrow="true" horizontal>
  If you see both CommError, Run does not exist and ERROR Error uploading during a sweep, the most likely cause is that you're setting a run ID manually in your code:
</Card>

<Card title="Why can't I link my artifact to the Registry?" href="/support/models/articles/cannot-link-artifact-to-registry-from-p/" arrow="true" horizontal>
  If you can't link an artifact to a W\&B Registry, the most common cause is that the artifact was logged with a personal entity instead of a team entity.
</Card>

<Card title="How do I fix an `anaconda 400 error` during a sweep?" href="/support/models/articles/anaconda-400-error-during-a-sweep/" arrow="true" horizontal>
  An anaconda 400 error during a sweep often means you didn't log the metric you're optimizing.
</Card>

## March 2026

### Models

<Card title="Signal handling and sweep runs" href="/models/sweeps/signal-handling-sweep-runs/" arrow="true" horizontal>
  Learn how W\&B Sweeps handle UNIX signals, exit codes, and preemption in sweep runs.
</Card>

### Weave

<Card title="OpenAI Realtime API" href="/weave/guides/integrations/openai-realtime-audio/" arrow="true" horizontal>
  Use Weave to automatically trace your calls to the OpenAI Realtime API.
</Card>

<Card title="Store and track versions of prompts" href="/weave/guides/core-types/prompts-version/" arrow="true" horizontal>
  Retrieve and manage versions of your prompts for LLM applications
</Card>

<Card title="Review items in an annotation queue" href="/weave/guides/tracking/annotation-review/" arrow="true" horizontal>
  Evaluate trace items and submit structured feedback using a simplified review interface.
</Card>

<Card title="Set up annotation queues" href="/weave/guides/tracking/annotation-queues/" arrow="true" horizontal>
  Create annotation queues, route traces to domain experts, and export structured feedback.
</Card>

### Platform

<Card title="W&B Mobile App (iOS)" href="/platform/hosting/monitoring-usage/mobile-app/" arrow="true" horizontal>
  Track training runs, review console logs, view line plots and system metrics, and explore W\&B Models projects from your iPhone or iPad.
</Card>

<Card title="Use W&B Skills" href="/platform/wb-skills/" arrow="true" horizontal>
  Install W\&B Skills to teach your coding agent how to train models, build agents, and analyze experiments using W\&B's AI development platform.
</Card>

<Card title="Rate limits" href="/platform/hosting/self-managed/rate-limits/" arrow="true" horizontal>
  Optional rate limits on Self-Managed instances for stability
</Card>

<Card title="Supported Dedicated Cloud regions" href="/platform/hosting/hosting-options/dedicated-cloud/regions/" arrow="true" horizontal>
  View all supported AWS, Google Cloud, and Azure regions available for W\&B Dedicated Cloud instances.
</Card>

<Card title="Rate limits" href="/platform/hosting/hosting-options/dedicated-cloud/rate-limits/" arrow="true" horizontal>
  Default rate limits on Dedicated Cloud and how to request changes
</Card>

<Card title="Export data from Dedicated Cloud" href="/platform/hosting/hosting-options/dedicated-cloud/export-data/" arrow="true" horizontal>
  Export runs, metrics, artifacts, and reports from a W\&B Dedicated Cloud instance using the Python SDK API.
</Card>

### Inference

<Card title="Cline with Serverless Inference" href="/inference/tutorials/integration-cline/" arrow="true" horizontal>
  >
</Card>

## February 2026

### Weave

<Card title="Manage Weave projects" href="/weave/guides/platform/weave-projects/" arrow="true" horizontal>
  Use Weave projects to organize related assets like traces, prompts, evaluations, models, and dashboards.
</Card>

### Platform

<Card title="Use the W&B MCP Server" href="/platform/mcp-server/" arrow="true" horizontal>
  Connect your IDE or AI agent to the W\&B Model Context Protocol (MCP) server to query and analyze your W\&B data and documentation in natural language.
</Card>

### Inference

<Card title="Model lifecycle" href="/inference/lifecycle/" arrow="true" horizontal>
  >
</Card>

## January 2026

### Models

<Card title="Pin and compare runs" href="/models/runs/compare-runs/" arrow="true" horizontal>
  Learn how to use pinned and baseline runs to keep track of important runs and efficiently evaluate model experiments.
</Card>

### Weave

<Card title="Create dynamic Leaderboards in Evaluations" href="/weave/guides/evaluation/dynamic_leaderboards/" arrow="true" horizontal>
  Dynamic Leaderboards let you configure, customize, save, and update Leaderboard views directly from an evaluation.
</Card>

<Card title="Map columns in datasets" href="/weave/guides/tools/column-mapping/" arrow="true" horizontal>
  Map columns in datasets to different names. This helps you align the column names in your dataset with the column names expected by the scorer.
</Card>

### Inference

<Card title="Creating a fine-tuned LoRA" href="/inference/tutorials/creating-lora/" arrow="true" horizontal>
  >
</Card>

<Accordion title="2025">
  ### December 2025

  #### Models

  <Card title="View an automation's history" href="/models/automations/view-automation-history/" arrow="true" horizontal>
    View the execution history of your W\&B Automations to check status, triggering events, and action results.
  </Card>

  <Card title="Evaluation benchmark catalog" href="/models/launch/evaluations/" arrow="true" horizontal>
    >
  </Card>

  <Card title="Evaluate a model checkpoint" href="/models/launch/evaluate-model-checkpoint/" arrow="true" horizontal>
    Evaluate a VLLM-compatible model checkpoint using infrastructure managed by CoreWeave
  </Card>

  <Card title="Evaluate a hosted API model" href="/models/launch/evaluate-hosted-model/" arrow="true" horizontal>
    Evaluate a hosted API model using infrastructure managed by CoreWeave
  </Card>

  <Card title="LLM Evaluation Jobs" href="/models/launch/" arrow="true" horizontal>
    Evaluate model checkpoints or hosted API models within W\&B and analyze the results using automatically generated leaderboards.
  </Card>
</Accordion>
