Add GLM-5.3 Flash to model list + 'which models' open question
This commit is contained in:
1 parent
94f6fded13
commit
483962a4fa
2 files changed
+9
-2
No files matched your search
@@ -41,10 +41,11 @@ The cooperative's software is open source and lives in the [`code`](https://git.
|
||||
|
||||
### Models
|
||||
|
||||
The gateway currently exposes two models:
|
||||
The gateway currently exposes three models (all served through Tinfoil's TEE-protected enclaves):
|
||||
|
||||
- **DeepSeek V4 Flash** — default model, best for agentic tasks
|
||||
- **DeepSeek V4 Flash** — default model, best for agentic tasks (1M context, tool calling)
|
||||
- **GPT-OSS 120B** — lightweight fallback
|
||||
- **GLM-5.3 Flash** — fast, efficient MoE model
|
||||
|
||||
### Features
|
||||
|
||||
|
||||
@@ -30,6 +30,12 @@ Tinfoil's catalog includes `nomic-embed-text` (embeddings), `websearch` (private
|
||||
|
||||
**Question:** Do we standardize on Tinfoil for all model needs (consistency, privacy), or keep the local embedding model (no per-token cost, no dependency)?
|
||||
|
||||
### 3b. Which models should we offer?
|
||||
|
||||
We currently expose three chat models: DeepSeek V4 Flash (default), GPT-OSS 120B, and GLM-5.3 Flash. Tinfoil's full catalog also includes GLM-5.3 (full), Kimi K3, Llama 3.3 70B, Gemma 4 31B, and others.
|
||||
|
||||
**Question:** Which models should members have access to? Should we offer a curated few (simpler, cheaper, easier to govern) or the full catalog (more choice, but more cost and governance overhead)? Who decides when to add or remove a model — the General Manager, or members via Loomio?
|
||||
|
||||
---
|
||||
|
||||
## Pricing & Fairness
|
||||
|
||||
Reference in new issue
Block a user