Skip to content

Agents > Inference & providers

Bring Your Own API Key

Open in ChatGPT ↗
Ask ChatGPT about this page
Open in Claude ↗
Ask Claude about this page
Copied!

Warp lets you bring your own API keys (BYOK) for OpenAI, Anthropic, and Google AI models.

Warp supports Bring Your Own API Key (BYOK) for users who want to connect agents to their own Anthropic, OpenAI, or Google API accounts.

This lets you use your own API keys for model access, giving you control over model selection, billing, and data routing. See Model Choice for a list of supported models. You can also connect a ChatGPT subscription for eligible OpenAI models or a SuperGrok subscription for Grok models.

Your provider bills inference routed through your keys. Applicable platform charges remain.

How BYOK differs from custom inference endpoints and BYOLLM

Section titled “How BYOK differs from custom inference endpoints and BYOLLM”

Warp offers several ways to bring your own AI infrastructure. Use this table to pick the right one, and follow the links for full details.

NameMeaningPlans
Personal BYOKUse your own API key for OpenAI, Anthropic, or Google models in local Warp Agent runs. Keys are stored on your device.Free and all eligible paid plans
Personal custom inference endpointConnect local Warp Agent runs to an OpenAI-compatible endpoint such as OpenRouter, LiteLLM, z.ai, or an internal gateway.Free and all eligible paid plans
Team-managed API keys and endpointsAn admin configures shared provider keys or endpoints for local and cloud Warp Agent runs.Business and Enterprise
Bring Your Own LLM (BYOLLM)Route inference through your organization’s AWS Bedrock or Gemini Enterprise (Vertex AI) account. See the provider guides for cloud support.Enterprise only
SuperGrok subscriptionConnect your SuperGrok subscription to use Grok models through your xAI account. Tokens are stored locally on your device.Free and all eligible paid plans
ChatGPT subscriptionConnect your ChatGPT plan to use eligible OpenAI models with the Warp Agent in the Warp app.See Warp pricing

See Warp pricing for current plan availability.

Cloud runs incur platform charges for billable agent time. Local runs on Business and Enterprise that use customer-supplied inference also incur platform charges. See platform usage for rates and contract-specific terms.

You can’t connect a Claude consumer subscription to the Warp Agent. To use your own Anthropic account, configure BYOK; Anthropic bills API usage separately from a subscription.

Personal API keys are stored in your device’s secure storage. When you select a provider model with a key icon, your prompt, context, and key pass through Warp’s servers to that provider. Warp does not retain your personal API key.

To configure personal API keys in the Warp app:

  1. Open Settings and search for API keys to find the BYOK configuration.
  2. Add your API key(s) for Anthropic, OpenAI, or Google.
  3. Once added, you’ll see a key icon next to supported models in the model picker.

For CLI runs, configure provider keys separately from the Warp app.

Key icon shown next to supported models in the model picker after BYOK API keys are configured.

When you select a model with a key icon, Warp routes inference through your own API key instead of consuming Warp-provided inference usage.

Warp’s Auto models use Warp-provided inference and consume your available usage, even if you’ve configured your own API keys.

To use your own key, select a specific provider model (for example, Claude Opus 4.7, Claude Sonnet 4.6, GPT-5.5, or Gemini 3.1 Pro) directly from the model picker with a key icon.

Custom routers resolve each task to a model you chose, then apply your API keys to that model. Requests covered by one of your keys bill inference through your provider account. See usage and model availability for details.

When you select a model with the key icon in your model picker, Warp routes the request through your API key. In that case:

  • Inference is billed through your provider account rather than your Warp usage balance.
  • Agent Mode prioritizes BYOK over Warp-provided inference.

Other AI features in Warp

Personal BYOK doesn’t change the inference configuration of these features:

FeatureUses Warp-provided usageDescription
Active AI RecommendationsNoAlways included with Build and higher plans.
Codebase ContextYesUses Warp-provided inference.
Cloud AgentsYesPersonal keys aren’t available. Business and Enterprise teams can use team-managed inference; applicable platform and Warp-hosted compute charges remain.

If Warp detects an issue with your API key, you’ll see a clear error message corresponding with the AI request.

If your key:

  • Is invalid: Warp notifies you and halts the request.
  • Hits usage or rate limits: Warp will not retry using Warp-provided inference unless fallback is enabled.

You can update or replace your keys anytime by opening Settings and searching for API keys.

Warp credit fallback is off by default. Enable it to retry a failed BYOK request with Warp-provided inference. Fallback requests consume your available Warp usage.

Setting to enable Warp credit fallback when a BYOK request fails.

Warp is SOC 2 compliant and has Zero Data Retention (ZDR) policies with all of its contracted LLM providers. No customer AI data is retained, stored, or used for training by the model providers.

BYOK prompts and responses transit Warp’s backend (see How BYOK works). Warp does not use this content for training; retention and analytics handling follow the same account-level privacy and telemetry settings that apply to Warp-billed traffic.

However, when you use your own API key:

  • Data retention policies on the provider side depend on your provider’s account settings.
  • Warp cannot enforce ZDR for requests sent through your API keys.
  • If your Anthropic, OpenAI, or Google account does not have ZDR enabled, your requests may be retained by the provider according to their terms.

Personal BYOK is configured on each member’s device, including on Business and Enterprise. Those keys work for local Warp Agent runs in the app or CLI, not cloud agents.

For shared keys that work in local and cloud runs, configure team-managed API keys and endpoints in the Admin Panel.