• English
  • How to Configure LLM Providers and Models in Octop

    Octop ships built-in presets for a wide range of LLM backends and lets you add any custom endpoint. Providers are managed by admins under Management → Models; the models they expose then feed a shared pool that every agent can draw from.

    Add a provider

    Management → Models

    Model management

    Model management

    Pick a built-in preset from Cloud or Local, enter your API key, and enable it. The provider's models immediately become available to agents under your account. Voice and search settings live on this same Models page.

    Built-in provider presets

    The Cloud tab covers hosted services:

    PresetProtocol
    Tencent Cloud (including Token Plan enterprise and Hy), Kimi, MiniMax, OpenCode, SiliconFlow, Aliyun, Volcano Engine, Zhipuopenai, multi-site
    OpenAI, OpenAI (ChatGPT), OpenRouter, DeepSeek, Google Gemini, Groq, ModelScope, Xiaomi MiMoopenai
    Anthropicanthropic

    Tencent Cloud includes Token Plan, Token Plan Enterprise (CN), Coding Plan, and Hy Token Plan variants. Hy models such as hy3 expose reasoning status automatically when the provider supports it.

    The Local tab ships an Ollama (Local) preset pointing at http://localhost:11434/v1. If you need something else entirely, use Custom provider to supply your own base URL and key.

    NOTE

    Presets are labelled by the API protocol they speak — openai for the OpenAI chat completions spec and anthropic for the Anthropic Messages API. Any endpoint that follows one of those specs works as a custom provider.

    The available model pool and auto-routing

    Every model from an enabled provider joins the available model pool at the top of the page. Octop can route each message automatically, choosing the model that best fits that message. Models you mark with a star are preferred when nothing about the request calls for something else.

    Each model in the pool shows its modality and context window, so you can see at a glance which ones handle long inputs.

    Use Test connection after entering credentials. Octop maps common authentication, quota, rate-limit, and provider-availability failures to localized guidance so you can fix the provider without reading raw upstream errors first.

    Set a model for an agent

    Agents fall back to the global default unless you assign a model to them specifically. Override it per agent when you want, for example, a cheaper model for everyday chat and a stronger one for complex work.

    Each account can also store a preferred model and per-model reasoning defaults under personal settings. Those preferences apply to new turns for that user and do not change the shared provider catalog.

    TIP

    To configure providers and models from the terminal, use octop models config, octop models list, and octop models active — see Agent commands.

    Use Ollama for local models

    Enable the Ollama (Local) preset under the Local tab. No API key is required; Octop talks to your Ollama instance at http://localhost:11434/v1.

    NOTE

    Ollama must be running before you send the first message to an agent that uses it. Use octop models ollama-list and octop models ollama-pull to inspect and download local models.

    Knowledge-base models

    Knowledge bases can index documents on this machine or through an online embedding model. Administrators pick the model under Knowledge Bases → Knowledge base settings, or download a local model on Models. See Knowledge bases.

    Volcengine image and video

    When a Volcengine (Ark) provider is enabled, you can configure Seedream image and Seedance video generation, test the models, and preview results from the web UI. These generation models are separate from chat completions.