Add GLM-5.3 Flash to model list + 'which models' open question
This commit is contained in:
1 parent
94f6fded13
commit
483962a4fa
2 files changed
+9
-2
No files matched your search
@@ -41,10 +41,11 @@ The cooperative's software is open source and lives in the [`code`](https://git.
|
|||||||
|
|
||||||
### Models
|
### Models
|
||||||
|
|
||||||
The gateway currently exposes two models:
|
The gateway currently exposes three models (all served through Tinfoil's TEE-protected enclaves):
|
||||||
|
|
||||||
- **DeepSeek V4 Flash** — default model, best for agentic tasks
|
- **DeepSeek V4 Flash** — default model, best for agentic tasks (1M context, tool calling)
|
||||||
- **GPT-OSS 120B** — lightweight fallback
|
- **GPT-OSS 120B** — lightweight fallback
|
||||||
|
- **GLM-5.3 Flash** — fast, efficient MoE model
|
||||||
|
|
||||||
### Features
|
### Features
|
||||||
|
|
||||||
|
|||||||
@@ -30,6 +30,12 @@ Tinfoil's catalog includes `nomic-embed-text` (embeddings), `websearch` (private
|
|||||||
|
|
||||||
**Question:** Do we standardize on Tinfoil for all model needs (consistency, privacy), or keep the local embedding model (no per-token cost, no dependency)?
|
**Question:** Do we standardize on Tinfoil for all model needs (consistency, privacy), or keep the local embedding model (no per-token cost, no dependency)?
|
||||||
|
|
||||||
|
### 3b. Which models should we offer?
|
||||||
|
|
||||||
|
We currently expose three chat models: DeepSeek V4 Flash (default), GPT-OSS 120B, and GLM-5.3 Flash. Tinfoil's full catalog also includes GLM-5.3 (full), Kimi K3, Llama 3.3 70B, Gemma 4 31B, and others.
|
||||||
|
|
||||||
|
**Question:** Which models should members have access to? Should we offer a curated few (simpler, cheaper, easier to govern) or the full catalog (more choice, but more cost and governance overhead)? Who decides when to add or remove a model — the General Manager, or members via Loomio?
|
||||||
|
|
||||||
---
|
---
|
||||||
|
|
||||||
## Pricing & Fairness
|
## Pricing & Fairness
|
||||||
|
|||||||
Reference in new issue
Block a user