diff --git a/api-access.md b/api-access.md index 4f6815c..512515e 100644 --- a/api-access.md +++ b/api-access.md @@ -25,7 +25,7 @@ API key. The key is an `sk-…` string you pass as a bearer token. curl https://gateway.inference.coop/v1/chat/completions \ -H "Authorization: Bearer sk-your-key" \ -H "Content-Type: application/json" \ - -d '{"model": "deepseek-v4-1-flash", "messages": [{"role": "user", "content": "Hello"}]}' + -d '{"model": "tinfoil/deepseek-v4-1-flash", "messages": [{"role": "user", "content": "Hello"}]}' ``` ## OpenAI SDK (Python, Node, etc.) @@ -40,7 +40,7 @@ client = OpenAI( api_key="sk-your-key", ) resp = client.chat.completions.create( - model="deepseek-v4-1-flash", + model="tinfoil/deepseek-v4-1-flash", messages=[{"role": "user", "content": "Hello"}], ) ``` @@ -57,7 +57,7 @@ endpoint configuration: - **API type / protocol:** "OpenAI compatible" (or `openai-completions` where a protocol string must be named) — not "Ollama", "Anthropic", or "Responses" - **API key:** your member key (the `sk-…` string) -- **Model:** one from the [model list](models.md), e.g. `deepseek-v4-1-flash` or `greenpt/green-r` +- **Model:** one from the [model list](models.md), e.g. `tinfoil/deepseek-v4-1-flash` or `greenpt/green-r` A typical provider entry looks like: @@ -69,9 +69,9 @@ A typical provider entry looks like: "api": "openai-completions", "apiKey": "sk-your-key", "models": [ - { "id": "deepseek-v4-1-flash" }, - { "id": "gpt-oss-120b" }, - { "id": "glm-5-3-flash" }, + { "id": "tinfoil/deepseek-v4-1-flash" }, + { "id": "tinfoil/gpt-oss-120b" }, + { "id": "tinfoil/glm-5-3-flash" }, { "id": "greenpt/green-r" }, { "id": "greenpt/green-l" } ] diff --git a/models.md b/models.md index 40d43ce..137f11f 100644 --- a/models.md +++ b/models.md @@ -22,9 +22,9 @@ provider — your own cost is your monthly allowance, not per-token billing). | Model | Provider | Value | Pricing (in / out) | Notes | |---|---|---|---|---| -| `deepseek-v4-1-flash` | Tinfoil | **private** | $0.65 / $1.45 | Default; best for agentic tasks (1M context, tool calling). | -| `gpt-oss-120b` | Tinfoil | **private** | $0.15 / $0.60 | Lightweight fallback. | -| `glm-5-3-flash` | Tinfoil | **private** | $0.40 / $1.25 | Fast, efficient MoE model. | +| `tinfoil/deepseek-v4-1-flash` | Tinfoil | **private** | $0.65 / $1.45 | Default; best for agentic tasks (1M context, tool calling). | +| `tinfoil/gpt-oss-120b` | Tinfoil | **private** | $0.15 / $0.60 | Lightweight fallback. | +| `tinfoil/glm-5-3-flash` | Tinfoil | **private** | $0.40 / $1.25 | Fast, efficient MoE model. | | `greenpt/green-r` | GreenPT | **green** | $0.35 / $0.95 | Reasoning, renewable energy. | | `greenpt/green-l` | GreenPT | **green** | $0.25 / $0.80 | Lightweight, renewable energy. | @@ -34,11 +34,6 @@ numbers. Pricing here mirrors the `model_info` in the LiteLLM config (`code/litellm/config.yaml`), which is the authoritative source — the two are kept in sync. -> **Note:** the three Tinfoil models currently use their bare names (no -> `tinfoil/` prefix). They'll be renamed to `tinfoil/deepseek-v4-1-flash` and so -> on in an upcoming change — we'll email members before that happens. GreenPT -> models already use the `greenpt/` prefix. - ## Limitations on privacy Only **Tinfoil** offers architectural privacy. The other providers are chosen