chore(env): switch AI_LLM_BASE_URL to omniroute (imrnes:20128)
Build & Deploy (Nix) / build-and-deploy (backend) (push) Successful in 1m38s
Build & Deploy (Nix) / build-and-deploy (discord-gateway) (push) Successful in 2m22s
Build & Deploy (Nix) / build-and-deploy (proxy) (push) Successful in 2m38s

9router → omniroute router. /api/v1 exposes OpenAI-compatible chat
(model 'text' → gemma-4-31b-it, verified) + embeddings (verified:
gemini-embedding-001/-2, nemotron-embed-vl-1b-v2:free, qwen3-embedding).
Runtime env already switched via /etc/gmw/discord-gateway.env +
GATEWAY_ENV secret; gateway restarted 21:10, embedding writes flowing.
This commit is contained in:
Developer
2026-07-31 21:11:39 +07:00
parent 2f298cfe22
commit dc119b5d5a
+1 -1
View File
@@ -85,7 +85,7 @@ BACKLOG_SYNC_BATCH_SIZE=100 # Messages per backlog batch, max 100 (d
# === AI Analysis ===
AI_ANALYSIS_ENABLED=false # Enable AI content moderation (default: false)
# AI_LLM_API_KEY= # REQUIRED if AI_ANALYSIS_ENABLED=true. LLM API key
AI_LLM_BASE_URL=https://9router.asepharyana.my.id/v1 # LLM API base URL (default)
AI_LLM_BASE_URL=http://100.121.180.82:20128/api/v1 # LLM API base URL (omniroute on imrnes; /api/v1 exposes OpenAI-compatible chat+embeddings)
AI_LLM_MODEL=text # LLM text model name (default: text)
# AI_LLM_VISION_MODEL= # Vision model for image analysis (falls back to AI_LLM_MODEL)
# AI_LLM_EMBEDDING_MODEL= # Embedding model for semantic moderation cache (optional; enables near-duplicate text reuse to save LLM calls)