Commit Graph

  • b6e8542146 config: remove duplicate broken audio-model entries (audio_transcription//audio_speech/ prefixes); keep the correct openai/ routing main Inferencebot 2026-10-02 15:39:07 -06:00
  • b9af62a614 README: current base image v1.103.2 Inferencebot 2026-10-02 15:32:49 -06:00
  • a2864ce34a Upgrade to LiteLLM v1.103.2 (schema pre-migrated live; CVE fix line + streaming-usage bugfix window) Inferencebot 2026-10-02 15:25:27 -06:00
  • 23f73cf6ae README: current architecture (3 providers, portal key injection, hardening), remove MVP model list (point to co-op/docs as source of truth), all provider env vars, member-key verification kept Inferencebot 2026-10-02 01:08:34 -06:00
  • e8c7a55933 Add upgrade runbook + bake memory limit and LITELLM_MIGRATION_DIR into packaging fix/cve-upgrade-v1.84.0-repo-sync inference-co-op-bot 2026-10-02 00:52:50 -06:00
  • fd34e33508 Sync packaging to live v1.84.0 deployment (CVE-2026-35029/59822 fix line) inference-bot 2026-10-01 23:32:43 -06:00
  • 8f42be988e Add Tinfoil audio models (whisper STT, voxtral-tts) for voice mode + API use inference-bot 2026-09-27 11:10:51 -06:00
  • 5ea2256776 Add greenpt/glm-5.3 (full flagship, $1.10/$4.40 per 1M) inference-bot 2026-09-26 16:52:21 -06:00
  • 057507d67c Add PublicAI Apertus models (publicai/apertus-v1.5-8b, -70b) inference-bot 2026-09-25 15:56:08 -06:00
  • 2369c3e848 Add greenpt/glm-5.3-flash (/usr/bin/bash.11//usr/bin/bash.44 per 1M) inference-bot 2026-09-24 10:43:14 -06:00
  • f86e005062 Rename Tinfoil models to tinfoil/* provider prefix inference-bot 2026-09-24 08:16:16 -06:00
  • 67192de232 Document pricing sync with co-op/docs/models.md inference-bot 2026-09-23 21:08:33 -06:00
  • 09ee36014c Add GreenPT models (greenpt/green-r, greenpt/green-l) via provider/model naming convention inference-bot 2026-09-23 20:10:18 -06:00
  • d1f57c45d5 Retry connection errors but never re-bill timeouts inference-bot 2026-09-22 18:18:07 -06:00
  • 576885ed21 Fix Tinfoil spend gap: raise timeout 30->180s, num_retries 2->0 inference-bot 2026-09-22 18:05:27 -06:00
  • 9cbbc33bc7 Fix CORS patch to target site-packages (the copy the running CLI imports) inference-bot 2026-09-14 18:01:13 -06:00
  • 0a797ba50e Patch hardcoded origins=[*] to honour LITELLM_CORS_ALLOWED_ORIGINS (v1.74.0 ignores the env var) inference-bot 2026-09-14 17:55:08 -06:00
  • dae9bf4f31 Lock CORS to chat origin (was wildcard * with allow-credentials) inference-bot 2026-09-14 17:51:38 -06:00
  • fb921b6e09 Disable admin UI (/ui) and API docs (Swagger/ReDoc/openapi) on the public gateway inference-bot 2026-09-14 16:21:34 -06:00
  • 700154c3d4 Add custom per-token pricing (model_info) so spend and budget enforcement work inference-bot 2026-09-14 14:28:13 -06:00
  • da50b0438f Replace deprecated deepseek-v4-flash with deepseek-v4-1-flash (default + fallback) inference-bot 2026-09-14 14:06:23 -06:00
  • 878ab8fa8d Add GLM-5.3 Flash model inference-bot 2026-09-09 14:18:14 -06:00
  • 8a54e75d4e Download tinfoil-proxy at runtime (build sandbox has no GitHub access) inference-bot 2026-09-09 13:23:53 -06:00
  • 26b0edc6a8 Hardcode tinfoil-proxy version in Dockerfile inference-bot 2026-09-09 13:22:21 -06:00
  • e05fec0ca8 Switch to Tinfoil (TEE): add tinfoil-proxy sidecar for EHBP encryption, point LiteLLM at local proxy inference-bot 2026-09-09 13:21:01 -06:00
  • 2ce5c66b9b Switch to Ollama Cloud (OpenAI-compatible /v1 endpoint) for plumbing smoke test inference-bot 2026-09-04 14:15:49 -06:00
  • 41068a6264 Fix Dockerfile for Cloudron: run as root (base image default), override ENTRYPOINT, fix .dockerignore inference-bot 2026-09-04 10:29:15 -06:00
  • b9b4620e52 Initial commit: LiteLLM Cloudron app package for Inference Cooperative inference-bot 2026-09-03 21:55:54 -06:00