Use NeverMetered with OpenCode

Open-source terminal agent with an OpenAI-compatible provider.

OpenAI Chat
Model
qwen3.8-27b
8,192 max output tokens on the free planReplace YOUR_API_KEY with your key.
  1. 1

    Create an API key

    Create a key in API keys and name it after the tool, for example “OpenCode”. A dedicated key per tool lets you see its usage separately and give it its own parallel-stream cap. The key is shown once; the snippets below use YOUR_API_KEY where it goes.

  2. 2

    Add the provider

    Save this as opencode.json in your project, or merge it into ~/.config/opencode/opencode.json to use it everywhere.

    opencode.json
    {
      "$schema": "https://opencode.ai/config.json",
      "provider": {
        "nevermetered": {
          "npm": "@ai-sdk/openai-compatible",
          "name": "NeverMetered",
          "options": {
            "baseURL": "https://api.nevermetered.com/v1",
            "apiKey": "{env:NEVERMETERED_API_KEY}"
          },
          "models": {
            "qwen3.8-27b": {
              "name": "qwen3.8-27b"
            }
          }
        }
      },
      "model": "nevermetered/qwen3.8-27b"
    }
  3. 3

    Export your key and start OpenCode

    export NEVERMETERED_API_KEY="YOUR_API_KEY"
    opencode
  4. 4

    Verify it works

    Run a single prompt. The request shows up in Requests within a few seconds, with its input, cached and output tokens.

    opencode run "Reply with the word ready"
Tips
  • If you know the model's context window, add "limit": { "context": …, "output": … } to the model entry so OpenCode compacts long sessions in time.
  • Subagents and parallel tool runs each use one of your parallel streams.

Troubleshooting

404 model_not_found
Use a model id exactly as GET /v1/models lists it, for example qwen3.8-27b. Plans can restrict which models a key may call.
429 concurrency_limit_exceeded
All your parallel streams are busy and your queue allowance is full. Lower the tool's parallelism (subagents, batch size), or get more streams on your plan.
Answers stop mid-sentence
The response hit your plan's output cap of 8,192 tokens (finish reason length or max_tokens). Ask for shorter steps, or upgrade for a higher cap.
403 email_verification_required
Free requests need a verified email. Click the link in your verification email (you can resend it from the dashboard), or upgrade.
429 daily_quota_exceeded
You've used today's free requests; agents send one request per step, so they go quickly. The allowance resets at 00:00 UTC, and paid plans have no daily limit.
503 / 504 request_queue_timeout
The service stayed busy longer than the queue timeout and nothing was generated. Retry with backoff, and give the tool a generous request timeout.