Baltor Get started

Endpoints

Vancine

A hosted service with its own key and prices. This page lists the addresses it answers, the facts its documentation states, and the setup of each harness.

Addresses and facts

  • OpenAI Chat Completions: https://vancine.com/v1 models.dev, read
Authentication
Authorization: Bearer, with the key in VANCINE_API_KEY models.dev, read
Structured output
Unknown
Tool calling
8 of 8 listed models The share of this provider's models that models.dev records with tool calling. models.dev, read
Rate limits
Unknown
Prices
Unknown
Data retention
Unknown
Documentation
Read the page models.dev, read
Output limits Baltor recorded
Unknown

Harness setup

Replace MiniMax-M3 with the model you want.

OpenCode

Put this in opencode.json in your project folder:

{
  "$schema": "https://opencode.ai/config.json",
  "provider": {
    "vancine": {
      "npm": "@ai-sdk/openai-compatible",
      "name": "Vancine",
      "options": {
        "baseURL": "https://vancine.com/v1",
        "apiKey": "{env:VANCINE_API_KEY}"
      },
      "models": {
        "MiniMax-M3": {
          "name": "MiniMax-M3"
        }
      }
    }
  }
}
  • OpenCode reads any OpenAI-compatible address through the @ai-sdk/openai-compatible package, and an address that speaks the Responses API through @ai-sdk/openai.

From OpenCode documentation, read .

Pi

Put this in ~/.pi/agent/models.json:

{
  "providers": {
    "vancine": {
      "baseUrl": "https://vancine.com/v1",
      "api": "openai-completions",
      "apiKey": "$VANCINE_API_KEY",
      "models": [
        {
          "id": "MiniMax-M3"
        }
      ]
    }
  }
}
  • The apiKey field can name an environment variable as $NAME.

From Pi documentation, read .

Codex

Codex speaks only the Responses API, and Vancine documents no Responses address. A gateway that offers one can sit in between.

  • Codex speaks the Responses API only: responses is the one supported wire API of a custom provider. Ollama and LM Studio are built in and start with --oss.

From Codex documentation, read .

Claude Code

Claude Code sends Anthropic Messages requests, and Vancine documents no such address. A gateway that translates to that API can sit in between.

  • Claude Code sends Anthropic Messages requests to ANTHROPIC_BASE_URL. Anthropic says it does not support routing Claude Code to models other than Claude through any gateway, so some features may not work with another model.

From Claude Code documentation, read .

Models it lists

Prices in US dollars per million tokens, input and output, as models.dev, read records them.

ModelInputOutputContextAs of
MiniMax-M3 MiniMax-M30.240 USD0.960 USD1,048,576 older than 30 days
DeepSeek V4.1 Flash deepseek-v4.1-flash0.240 USD0.960 USD1,000,000
GLM-5.3 glm-5.31.12 USD3.52 USD1,000,000 older than 30 days
GLM-5.3-Flash glm-5.3-flash0.120 USD0.400 USD1,000,000
Hy4 preview hy4-preview0.670 USD2.00 USD1,024,000
Kimi K3 kimi-k32.40 USD12.00 USD1,048,576 older than 30 days
Qwen3.8 Flash qwen3.8-flash0.120 USD0.380 USD1,000,000
Qwen3.8 Max qwen3.8-max1.60 USD4.80 USD1,000,000 older than 30 days

Sources of this page