LM Warden

Cursor on your own GPUs.

Cursor’s OpenAI base URL override, for chat with a served model.

Protocol
OpenAI Chat Completions
Support
Configured from its docs
Verified
Checked against its docs on 2026-10-04. Config checked against the docs; the streaming Chat Completions request it sends, with a tool call, passed with curl.

Requests are issued from Cursor’s servers, so this warden’s URL must be reachable from the internet; a LAN or localhost warden will not work. Custom keys only work with chat models: Tab completion keeps using Cursor’s own models, and Agent with Cursor’s models rejects custom keys. Cursor’s current docs no longer describe the Override OpenAI Base URL setting; it is known from Cursor staff forum replies.

Requirements

  • Cursor’s servers call this URL, so it must be reachable from the internet.

Setup

In the console, Connect writes these files with your warden’s address, your key and the model you picked. Here they are with placeholders.

Chat with a local model

Cursor Settings > Models, then add the model name as a custom model.

Settings
OpenAI API Key: vw_YOUR_KEY
Override OpenAI Base URL: https://your-warden/v1
Custom model: your-served-model-name

Official documentation: https://cursor.com/docs/settings/api-keys (read 2026-10-04).