Document multi-provider model choices (Tinfoil private + GreenPT green)

This commit is contained in:
inference-bot committed 2026-09-23 20:34:33 -06:00
1 parent 89152f0724
commit 5724b0ca2f
3 files changed
+68 -20

No files matched your search

+37 -6
View File
@@ -9,14 +9,43 @@ client works.
- **Base URL:** `https://gateway.inference.coop/v1`
- **Auth:** a bearer API key (create one in the [Member Dashboard](https://dashboard.inference.coop))
- **Format:** OpenAI chat completions — `POST https://gateway.inference.coop/v1/chat/completions`
- **Models:**
- `deepseek-v4-1-flash` — default, best for agentic tasks (1M context, tool calling)
- `gpt-oss-120b` — lightweight fallback
- `glm-5-3-flash` — fast, efficient MoE model
- **Models:** see the table below.
API usage draws from the **same monthly allowance as chat** — there's no separate
quota.
## Models
Models are named `provider/model-name`, so you can choose not just *which* model
but *where* it runs. Providers carry a value you can route by:
| Model | Provider | Value |
|---|---|---|
| `deepseek-v4-1-flash` | Tinfoil | **private** (TEE) |
| `gpt-oss-120b` | Tinfoil | **private** (TEE) |
| `glm-5-3-flash` | Tinfoil | **private** (TEE) |
| `greenpt/green-r` | GreenPT | **green** (renewable) |
| `greenpt/green-l` | GreenPT | **green** (renewable) |
- **Tinfoil** — *private*. Models run inside hardware enclaves (TEEs) with
end-to-end encryption, so neither we nor the provider can read your prompts
or responses. Best privacy.
- **GreenPT** — *green*. Models served on 100% renewable energy in the EU.
Cleaner energy, standard (non-enclave) hosting.
> **Note:** the three Tinfoil models currently use their bare names (no
> `tinfoil/` prefix). They'll be renamed to `tinfoil/deepseek-v4-1-flash` and so
> on in an upcoming change — we'll email members before that happens. GreenPT
> models already use the `greenpt/` prefix.
**Which to pick:**
- `deepseek-v4-1-flash` — default; best for agentic tasks (1M context, tool calling).
- `gpt-oss-120b` — lightweight fallback.
- `glm-5-3-flash` — fast, efficient MoE model.
- `greenpt/green-r` — reasoning, renewable energy.
- `greenpt/green-l` — lightweight, renewable energy.
## Get a key
Log in to the [Member Dashboard](https://dashboard.inference.coop) and create an
@@ -60,7 +89,7 @@ endpoint configuration:
- **API type / protocol:** "OpenAI compatible" (or `openai-completions` where a
protocol string must be named) — not "Ollama", "Anthropic", or "Responses"
- **API key:** your member key (the `sk-…` string)
- **Model:** one of `deepseek-v4-1-flash`, `gpt-oss-120b`, `glm-5-3-flash`
- **Model:** one from the table above, e.g. `deepseek-v4-1-flash` or `greenpt/green-r`
A typical provider entry looks like:
@@ -74,7 +103,9 @@ A typical provider entry looks like:
"models": [
{ "id": "deepseek-v4-1-flash" },
{ "id": "gpt-oss-120b" },
{ "id": "glm-5-3-flash" }
{ "id": "glm-5-3-flash" },
{ "id": "greenpt/green-r" },
{ "id": "greenpt/green-l" }
]
}
}