Cursor on your own GPUs.
Cursor’s OpenAI base URL override, for chat with a served model.
- Protocol
- OpenAI Chat Completions
- Support
- Configured from its docs
- Verified
- Checked against its docs on 2026-10-04. Config checked against the docs; the streaming Chat Completions request it sends, with a tool call, passed with curl.
Requests are issued from Cursor’s servers, so this warden’s URL must be reachable from the internet; a LAN or localhost warden will not work. Custom keys only work with chat models: Tab completion keeps using Cursor’s own models, and Agent with Cursor’s models rejects custom keys. Cursor’s current docs no longer describe the Override OpenAI Base URL setting; it is known from Cursor staff forum replies.
Requirements
- Cursor’s servers call this URL, so it must be reachable from the internet.
Setup
In the console, Connect writes these files with your warden’s address, your key and the model you picked. Here they are with placeholders.
Chat with a local model
Cursor Settings > Models, then add the model name as a custom model.
OpenAI API Key: vw_YOUR_KEY Override OpenAI Base URL: https://your-warden/v1 Custom model: your-served-model-name
Official documentation: https://cursor.com/docs/settings/api-keys (read 2026-10-04).