Add GLM-5.3 Flash to model list + 'which models' open question

This commit is contained in:
inference-bot committed 2026-09-09 14:20:35 -06:00
1 parent 94f6fded13
commit 483962a4fa
2 files changed
+9 -2

No files matched your search

+3 -2
View File
@@ -41,10 +41,11 @@ The cooperative's software is open source and lives in the [`code`](https://git.
### Models
The gateway currently exposes two models:
The gateway currently exposes three models (all served through Tinfoil's TEE-protected enclaves):
- **DeepSeek V4 Flash** — default model, best for agentic tasks
- **DeepSeek V4 Flash** — default model, best for agentic tasks (1M context, tool calling)
- **GPT-OSS 120B** — lightweight fallback
- **GLM-5.3 Flash** — fast, efficient MoE model
### Features
+6
View File
@@ -30,6 +30,12 @@ Tinfoil's catalog includes `nomic-embed-text` (embeddings), `websearch` (private
**Question:** Do we standardize on Tinfoil for all model needs (consistency, privacy), or keep the local embedding model (no per-token cost, no dependency)?
### 3b. Which models should we offer?
We currently expose three chat models: DeepSeek V4 Flash (default), GPT-OSS 120B, and GLM-5.3 Flash. Tinfoil's full catalog also includes GLM-5.3 (full), Kimi K3, Llama 3.3 70B, Gemma 4 31B, and others.
**Question:** Which models should members have access to? Should we offer a curated few (simpler, cheaper, easier to govern) or the full catalog (more choice, but more cost and governance overhead)? Who decides when to add or remove a model — the General Manager, or members via Loomio?
---
## Pricing & Fairness