docs: update .env.example Ollama endpoint/model to match the working config
Root-caused a real bug while fixing this: mint_profile was silently falling back to the tiny offline template pool on nearly every call, producing near-identical spirits (same 3-4 persona openings verbatim, compound-noun names from a 5x5 pool). Traced it to the old Ollama endpoint being unreachable/overloaded and, separately, a model-tag mismatch (minicpm-v4.5:latest was requested but only :8b exists on the working box). Verified 6 candidate models against the real mint task on the new endpoint — only minicpm-v4.5:8b reliably returns valid JSON; the others (qwen3.5, ministral-3, granite4.1, lfm2.5, ornith) output conversational prose instead of the structured schema. Confirmed live through the actual guest summon path post-restart: genuinely distinct, well-written personas now, not fallback template text. .env itself is gitignored and was updated locally; this just keeps the example honest. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
This commit is contained in:
@@ -1,8 +1,8 @@
|
|||||||
DATABASE_URL=postgresql+asyncpg://quantumancy:quantumancy@localhost:5432/quantumancy
|
DATABASE_URL=postgresql+asyncpg://quantumancy:quantumancy@localhost:5432/quantumancy
|
||||||
OLLAMA_BASE_URL=http://10.30.20.107:11434
|
OLLAMA_BASE_URL=http://10.30.20.186:11434
|
||||||
PORT=7777
|
PORT=7777
|
||||||
# LLM tiers on the Ollama box (fast = fragments/ambient, chat = direct contact/minting)
|
# LLM tiers on the Ollama box (fast = fragments/ambient, chat = direct contact/minting)
|
||||||
OLLAMA_FAST_MODEL=granite4.1:3b
|
OLLAMA_FAST_MODEL=granite4.1:3b
|
||||||
OLLAMA_CHAT_MODEL=minicpm-v4.5:latest
|
OLLAMA_CHAT_MODEL=minicpm-v4.5:8b
|
||||||
LLM_MAX_CONCURRENCY=2
|
LLM_MAX_CONCURRENCY=2
|
||||||
LLM_MAX_QUEUE_DEPTH=8
|
LLM_MAX_QUEUE_DEPTH=8
|
||||||
|
|||||||
Reference in New Issue
Block a user