Settings
All of Sirius’s own settings live under sirius.ai in File → Preferences → Settings
(search for Sirius), or in settings.json. This table is generated from the
extension manifest that ships in the build, so it cannot drift from the product — the
descriptions below are the ones the Settings editor shows.
API keys are not settings. They live in the system keyring and are set with the Sirius: Set API Key command; see Providers and keys.
Current settings
Section titled “Current settings”| Setting | Default | What it does |
|---|---|---|
sirius.ai.defaultProviderstring: anthropic · gemini · ollama · openai · openrouter · groq · deepseek · mistral · xai · lmstudio · llamacpp · custom | anthropic | Default AI model provider |
sirius.ai.defaultModelstring | claude-opus-5 | Model used for AI chat. Prefer the "Sirius: Select AI Model" command, which lists everything the configured providers offer. |
sirius.ai.ollama.endpointstring | http://localhost:11434 | Ollama server endpoint |
sirius.ai.thinking.enabledboolean | true | Enable thinking/reasoning mode for supported models (Claude Opus, Sonnet, Gemini 3.5) |
sirius.ai.thinking.effortstring: low · medium · high · xhigh · max | high | How deeply the AI should reason before responding |
sirius.ai.maxTokensnumber | 16384 | Maximum tokens in AI response |
sirius.ai.temperaturenumber | 0.7 | AI response creativity (0 = deterministic, 2 = very creative) |
sirius.ai.streamResponsesboolean | true | Stream AI responses token by token |
sirius.ai.enableobject | {"*":false} | Tab completion, per language. "*" is the default for every language; a language id such as "markdown" or "json" overrides it. This is the switch the chat status dashboard toggles. Completions need a local FIM model — Ollama or llama.cpp; see `sirius.ai.completions.model |
sirius.ai.openai.baseUrlstring | https://api.openai.com/v1 | Override the OpenAI endpoint — e.g. an Azure deployment or a corporate proxy. |
sirius.ai.lmstudio.baseUrlstring | http://localhost:1234/v1 | LM Studio local server endpoint. |
sirius.ai.llamacpp.baseUrlstring | http://localhost:8080/v1 | llama.cpp or vLLM OpenAI-compatible endpoint. |
sirius.ai.custom.baseUrlstring | "" | Base URL of any other OpenAI-compatible service, including the /v1 suffix. |
sirius.ai.completions.modelstring | auto | Model for Tab completion: "auto" picks a running Ollama code model automatically; or "ollama/<model>"; or "llamacpp" for the configured llama.cpp server. |
sirius.ai.completions.maxLinesnumber | 12 | Longest suggestion Tab completion will offer, in lines. |
sirius.ai.ollama.largeModelBytesnumber | 12000000000 | Refuse to load Ollama models larger than this many bytes — protection against freezing the machine with a model that far exceeds its memory. 0 disables the check. |
sirius.ai.nextEditSuggestions.enabledboolean | false | Predict your next edit after each change and offer it as a Tab-able diff (experimental; needs a local FIM model). Requires Tab completion to be on — see sirius.ai.enable. |
Deprecated settings
Section titled “Deprecated settings”These still work for one more release cycle but should not be used in new configurations. The replacement is named in each row.
| Setting | Default | What it does |
|---|---|---|
sirius.ai.gemini.apiKeystring | "" | Google Gemini API key (get one from ai.google.dev) Deprecated: keys are now stored in the system keyring. Run "Sirius: Set API Key" instead. |
sirius.ai.anthropic.apiKeystring | "" | Anthropic API key (get one from console.anthropic.com) Deprecated: keys are now stored in the system keyring. Run "Sirius: Set API Key" instead. |
sirius.ai.openai.apiKeystring | "" | OpenAI API key (get one from platform.openai.com) Deprecated: keys are now stored in the system keyring. Run "Sirius: Set API Key" instead. |
sirius.ai.inlineCompletionsboolean | false | Enable AI-powered inline code completions (experimental) Use sirius.ai.enable — it is per language and is what the chat status dashboard toggles. true here still turns completions on everywhere. |