diff --git a/README.md b/README.md index 87cf875..d82bf31 100644 --- a/README.md +++ b/README.md @@ -41,10 +41,11 @@ The cooperative's software is open source and lives in the [`code`](https://git. ### Models -The gateway currently exposes two models: +The gateway currently exposes three models (all served through Tinfoil's TEE-protected enclaves): -- **DeepSeek V4 Flash** — default model, best for agentic tasks +- **DeepSeek V4 Flash** — default model, best for agentic tasks (1M context, tool calling) - **GPT-OSS 120B** — lightweight fallback +- **GLM-5.3 Flash** — fast, efficient MoE model ### Features diff --git a/open-questions.md b/open-questions.md index 423400f..22437d8 100644 --- a/open-questions.md +++ b/open-questions.md @@ -30,6 +30,12 @@ Tinfoil's catalog includes `nomic-embed-text` (embeddings), `websearch` (private **Question:** Do we standardize on Tinfoil for all model needs (consistency, privacy), or keep the local embedding model (no per-token cost, no dependency)? +### 3b. Which models should we offer? + +We currently expose three chat models: DeepSeek V4 Flash (default), GPT-OSS 120B, and GLM-5.3 Flash. Tinfoil's full catalog also includes GLM-5.3 (full), Kimi K3, Llama 3.3 70B, Gemma 4 31B, and others. + +**Question:** Which models should members have access to? Should we offer a curated few (simpler, cheaper, easier to govern) or the full catalog (more choice, but more cost and governance overhead)? Who decides when to add or remove a model — the General Manager, or members via Loomio? + --- ## Pricing & Fairness